0
Applied AI·July 26, 2026·1 min read

Agentic test processes, LLM benchmarks, and other notes on agentic coding

Share

The hard part of agentic coding isn’t generation, it’s building test harnesses and benchmarks that keep autonomous changes from quietly rotting your codebase. If you’re piloting AI coding agents, invest early in automated regression suites and sandboxed evaluation—governance will determine whether these tools compound or create cleanup debt.

Applied AI

Exclusive: 'We want to keep people away from doctors': Why Samsung has gone all-in on AI health, according to its executives — but do users trust it?

Consumer health is shifting from 'track and nudge' to 'triage and substitute'—Samsung talking about keeping people away from doctors puts them closer to care delivery than wellness. If you're building in digital health, assume regulators and incumbents will scrutinize any AI that meaningfully deflects clinical visits, and design your data, liability, and UX posture accordingly.

Applied AI

Monday.com is the latest tech company to blame AI for layoffs — here are 20 others

When 20+ larger tech employers explicitly cite AI in layoff narratives, it's not just cost-cutting language — it's a signal that exec teams now feel board cover to rebase headcount around AI-augmented workflows. If you're an operator, assume 2026 planning cycles will pressure you to show either AI-driven margin expansion or a credible story on why your org structure is exception-worthy.