Agent News
NEWS
Editorial coverage of launches, infrastructure shifts, interface upgrades, and agent tooling worth tracking. Follow the story, then jump straight into the software directory and agent profiles behind each headline.
GPT-5.6 after day one: the benchmarks, the caveats, and the real work
OpenAI’s Sol, Terra, and Luna family is now live. Here is what the official scorecards actually say, where independent and early practitioner evidence agrees or pushes back, and which use cases look ready for production testing.
Grok 4.5 gets a Thursday public launch, but the benchmark story is still missing
Elon Musk says xAI will make Grok 4.5 public Thursday after SpaceX and Tesla beta feedback. The model is framed as Opus-class and lower cost, but xAI has not yet published pricing, benchmarks, API details, or a model card.
Loop Engineering: The Complete Guide to Building Self-Improving AI Agents
Stop prompting your coding agents one shot at a time. Here is how to design loops that prompt them for you—and when the extra complexity is worth it.

