0
Applied AI·August 7, 2026·1 min read

The next age of LLMs? Dev gets a small LLM running at 10 tokens a second locally on a $10 microcontroller

Share

A 28.9M-parameter model running near 10 tok/s on a sub-$10 microcontroller — with most weights staying in flash — pushes “good enough” language intelligence into the bill-of-materials noise. Hardware teams building devices, sensors, and appliances should be scoping where on-device LLMs can replace cloud calls and unlock offline, low-latency features.

Applied AI

WhatsApp scam costs Hong Kong man $1.27 million after criminals used AI voice notes to impersonate his father — experts say secret codewords are the best way to stay safe

Voice is no longer an authentication factor when $1.27 million can move on the back of a cloned WhatsApp note. If you run any high-value workflows over consumer messaging — family offices, VIP client services, internal approvals — you need out-of-band codewords and hard limits this week, not just awareness training.