Agents
Long-horizon agent benchmarks are fragmenting: a field guide to what each one actually measures
Originally published on the Arize AI blog: Long-horizon agent benchmarks are fragmenting: a field …
The Phoenix Project still holds up, even if you replaced all the code with agents
I reread The Phoenix Project last month. I do this every year or two — it’s one of those books …
Build a Star Wars Copilot in C# - Lesson 8: Agents and Orchestration
Final lesson, and probably my favorite.
We move from a copilot with tools to a system that also uses …


