---
title: Scaling Laws of Synthetic Images for Model Training … for Now
url: https://www.ml-quant.com/papers/arxiv/2312.04567/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2312.04567
source_url: https://arxiv.org/abs/2312.04567
featured: 2023-12-13
citations: 138
topic: Other
---


# Scaling Laws of Synthetic Images for Model Training … for Now

The study investigates the scaling laws of synthetic images used in training supervised models, identifying factors that influence scaling behavior and situations where scaling synthetic data is most effective.

- Source: https://arxiv.org/abs/2312.04567
- Identifier: arXiv:2312.04567
- Released: 2023-12-07
- First featured: Quant Letter No. 29 (2023-12-13): https://www.ml-quant.com/issues/2023-12-13/
- Citations (Semantic Scholar): 138
- Published in: 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
- Topic: Other

## Related

- [HumanVid: Demystifying Training Data for Camera-controllable Human Image Animation](https://www.ml-quant.com/papers/arxiv/2407.17438/): Camera-controllable Human Image Animation: HumanVid, a new large-scale dataset for human image animation that combines real and synthetic data, has been developed by researchers, setting a new standard in the field.
- [R.I.P.: Better Models by Survival of the Fittest Prompts](https://www.ml-quant.com/papers/arxiv/2501.18578/): The study introduces Rejecting Instruction Preferences (RIP), a method for evaluating data integrity that can filter prompts or create synthetic datasets, enhancing performance across various benchmarks.
- [Depth Anything V2](https://www.ml-quant.com/papers/arxiv/2406.09414/): Depth Anything V2 is a new model for monocular depth estimation, using synthetic and large-scale pseudo-labeled real images for faster, more accurate results and setting a new evaluation benchmark.
- [MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark](https://www.ml-quant.com/papers/arxiv/2406.01574/): MMLU-Pro, an improved dataset, expands the Massive Multitask Language Understanding benchmark by adding tougher questions and more choices, serving as a better benchmark to monitor progress in the field.
- [Qwen2.5-Coder Technical Report](https://www.ml-quant.com/papers/arxiv/2409.12186/): The report unveils the Qwen2.5-Coder series, an improvement from its predecessor, showcasing remarkable code generation abilities and achieving top-tier performance in various code-related tasks.
- [Real-time Photorealistic Dynamic Scene Representation and Rendering with 4D Gaussian Splatting](https://www.ml-quant.com/papers/arxiv/2310.10642/): The 4DGS model is introduced, capable of reconstructing dynamic 3D scenes from 2D images and generating diverse views over time, providing real-time rendering efficiency.
