llamafile v0.10.5 is out, and it tracks a much more recent llama.cpp, which lets it run two models people have been using lately: Ternary Bonsai 27B and Poolside's Laguna-S-2.1. Both already existed a...
Writing is one of the fundamental tools that enable and enhance human thought, a common refrain I used back in 2023, as ChatGPT exploded on the scene. It felt like everyone I knew, colleagues, friends...
An agent stuck in a reasoning loop doesn't crash. It just quietly burns through your monthly budget until someone notices the bill. A week later, a provider has an outage and your app goes down with i...
Meet Otari, an open-source LLM gateway powered by any-llm, and, the hosted platform built on the same foundation. Run frontier or open-weights models through one API with usage tracking, budget contro...
Cloud AI pricing changed fast in 2026. This post looks at why more teams are moving back to local models, the tradeoffs behind tools like Ollama and LM Studio, and why portability and ownership are be...
The future of AI may not be agents using today’s apps. It may be apps rebuilt around structured representations agents can inspect, modify, and validate directly. The deck, doc, or dashboard becomes t...
Six weeks ago, Daniel Nissani at shared cq, Stack Overflow for agents. One of the top concerns in that thread was security and trust around shared we worked together to build VIBE, a first line of def...
The Octonous open beta is live. Learn what we discovered during closed beta, the workflow patterns users kept returning to, and the biggest improvements shipped since launch.
Sovereign AI shows up across nations, companies, communities, and individuals. This piece, based on a conversation with John Dickerson, CEO at, looks at control over AI systems, avoiding single points...