Topics

2026-09-21

Can Jev Be a Better Agent Evaluator?

We tested using Jev-as-a-Judge against LLM judges on accuracy, repeatability, latency, and cost to see whether System One models could offer a new approach to agent evaluation.

LangChainAgents and workflows
2026-09-19

v2.1.277

What's changed Added AGENTS.md support: in a project with no CLAUDE.md

Claude Code ReleasesAgents and workflows
2026-09-17

v2.1.274

What's changed Added a visible warning when memory usage is critical

Claude Code ReleasesAgents and workflows
2026-09-16

v2.1.273

What's changed Added x-claude-code-request-class , x-claude-code-agent-type , x-claude-code-prev-tool-durations

Claude Code ReleasesAgents and workflows
2026-09-15

v2.1.271

What's changed Added fast mode in Claude Code Remote sessions (cloud and self-hosted runners): the host's fast-mode setting or /fast typed in the session applies where your

Claude Code ReleasesAgents and workflows
2026-09-12

v2.1.269

What's changed Added claude plugin eval : run a plugin's eval suite against Claude Code and get scored

Claude Code ReleasesAgents and workflows
2026-09-10

Now everyone can put data to work

Meet the Data agent in ChatGPT Work. Connect company data, uncover insights, and build interactive dashboards with AI using natural language.

OpenAI NewsAgents and workflows

The Open Source AI Stack

A deep dive into the open model AI stack — model, inference, gateways and routers, harness

Together AIAgents and workflows