0
Applied AI·September 1, 2026·1 min read

The efficient frontier of LLM inference

Share

Inference is becoming a capital allocation problem—model size, quantization, and hardware choice define an efficient frontier of cost vs. latency vs. quality. If you’re running LLMs in production, you need someone explicitly accountable for moving workloads along that frontier instead of just “using the latest model” by default.

Applied AI

‘Horrifying’ — Reddit users are creating uncanny valley photos with ChatGPT, so I asked it to make something creepy and now I deeply regret it

The fact that “uncanny” and “horrifying” are now casual prompts for mainstream tools means your brand will eventually be adjacent to content you didn’t intend. If you operate consumer surfaces, tighten your content policies and think through how you’ll handle user-generated AI images that cross the line without being obviously illegal.