01 · Context engineering
Anthropic turns context engineering into a first-class discipline
The new Claude 5 playbook is being treated by practitioners as a reference document rather than launch copy. The implication is practical: agent quality now depends as much on context selection, tool-result shaping, and memory layering as on the base model.
Read the playbook →
02 · Open-weight infrastructure
Open-weight AI is having its Kubernetes moment
Tooling, orchestration, and deployment layers are becoming the strategic center of gravity around open models. The competitive question is shifting from “which weights?” to “which platform makes those weights usable?”
Read the analysis →
03 · Edge inference
A 28.9M-parameter LLM runs on an $8 microcontroller
The ESP32 demonstration makes always-on inference feel like a product constraint rather than a research stunt. Removing the server, GPU bill, and network round-trip opens a new lane for small, private, resilient AI products.
Inspect the project →
04 · Voice models
Full text-to-voice capability fits under 10M parameters
Inflect-Micro-v2 is a compact reminder that model size is no longer a reliable proxy for product surface. Small, complete systems can win where latency, privacy, and unit economics matter more than benchmark prestige.
Explore the model →
05 · Web economics
Cloudflare gives publishers more control over AI traffic
New controls for crawler and inference traffic make content access an explicit infrastructure and business decision. Agent builders will need to design for permission, attribution, and sustainable access instead of assuming the web is frictionless.
Read Cloudflare’s announcement →