AI Engineer Study Library

Local Evals: Open-Source App for Evaluating Local or OpenAI-Compatible Models

Melvin Vivas · X post · 2026-09-08 · Open on X

Topics: Evaluation (Evals) & Testing, LLM Fundamentals, AI Dev Tools & Productivity · Level: intermediate

Summary

Melvin Vivas released local-evals, an open-source evaluation app that runs locally. It works with local models or any OpenAI-compatible API. It evaluates document/image-to-JSON, text-to-JSON and tool calling. He built it entirely with Codex, using GPT-6 Astra as orchestrator and Luna as subagents, and tested it with LM Studio and OpenRouter.

Key points

Resources mentioned

Try this

More in Evaluation (Evals) & Testing

All of Evaluation (Evals) & Testing