0
Applied AI·September 14, 2026·1 min read

When LLM judges agree, should we believe them?

Share

The question of when to trust consensus among LLM judges goes straight at the foundation of automated evals and agent oversight. If you're leaning on model-as-judge for QA, ranking, or safety checks, treat this as a research dependency — not a solved problem — and keep a human-in-the-loop for any high-stakes decisions.

Applied AI

Sources: OpenAI bought Glass Imaging, which is developing AI-powered smartphone camera tech, in a deal valuing it at $300M+; it was valued at ~$100M last year

Paying $300M+ for AI-first camera tech is a bet that perception quality is now a core input to model performance and user experience, not a handset-side afterthought. If your product depends on real-world capture — images, video, sensor data — expect model vendors to move closer to that hardware and compress your room to differentiate there.