---
title: ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning
url: https://www.ml-quant.com/papers/arxiv/2411.05003/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2411.05003
source_url: https://arxiv.org/html/2411.05003v1
featured: 2024-11-13
citations: 76
topic: ML & AI Methods
---


# ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning

ReCapture is a method for creating new videos with unique camera trajectories from a single video, allowing for the regeneration of the video from different angles and cinematic camera motion.

- Source: https://arxiv.org/html/2411.05003v1
- Identifier: arXiv:2411.05003
- Released: 2024-11-07
- First featured: Quant Letter No. 74 (2024-11-13): https://www.ml-quant.com/issues/2024-11-13/
- Citations (Semantic Scholar): 76
- Published in: 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
- Topic: ML & AI Methods

## Related

- [Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal Control](https://www.ml-quant.com/papers/arxiv/2409.08861/): The study presents Adjoint Matching, a new algorithm that enhances dynamical generative models by refining reward fine-tuning, leading to improved consistency, realism, and adaptability to unseen human preference reward models.
- [Scaling Proprioceptive-Visual Learning with Heterogeneous Pre-trained Transformers](https://www.ml-quant.com/papers/arxiv/2409.20537/): The paper presents Heterogeneous Pre-trained Transformers (HPT), a method for training robotic models across various tasks, improving the performance of fine-tuned policies by over 20% on unseen tasks in both simulated and real-world environments.
- [Continual Diffusion: Continual Customization of Text-to-Image Diffusion with C-LoRA](https://www.ml-quant.com/papers/arxiv/2304.06027/): CLoRA, a new method, prevents catastrophic forgetting in text-to-image models when introducing new concepts, achieving top performance in continual learning settings for image classification.
- [Fine-tuning can cripple your foundation model; preserving features may be the solution](https://www.ml-quant.com/papers/arxiv/2308.13320/): Concept forgetting in AI models can be significantly reduced by a new fine-tuning method called LDIFS, which helps retain pre-trained knowledge while working on different tasks.
- [Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data](https://www.ml-quant.com/papers/arxiv/2412.07762/): The article introduces Warm-start RL (WSRL), a new reinforcement learning approach that doesn't require offline data, leading to quicker learning and better performance than previous algorithms.
- [Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate](https://www.ml-quant.com/papers/arxiv/2501.17703/): The article introduces Critique Fine-Tuning (CFT), a new method for training language models that critiques incorrect responses, showing better results than the traditional Supervised Fine-Tuning (SFT) method in math benchmarks.
