0
Applied AI·August 19, 2026·1 min read

Using synthetic data for AI training is 'a big mistake,' says AI pioneer Rich Sutton

Share

Sutton’s warning on synthetic training data is a shot across the bow for teams leaning on self-generated corpora to escape data scarcity. If your roadmap assumes heavy synthetic data use, you need a parallel track validating whether performance gains are durable or just overfitting to your own model’s biases.