AI Engineer Study Library

local-evals: A Local LLM Eval App Built by Codex Subagents

Melvin Vivas · X post · 2026-09-10 · Open on X

Topics: Evaluation (Evals) & Testing, AI Agents, Tool Use & MCP, AI Dev Tools & Productivity · Level: intermediate

Summary

Melvin Vivas shares local-evals, an open-source eval app that Codex built from scratch with an Astra orchestrator and Luna subagents. The app runs locally and can test local models or any OpenAI-compatible API.

Key points

Resources mentioned

Try this

More in Evaluation (Evals) & Testing

All of Evaluation (Evals) & Testing