AI Engineer Study Library

Cutting Token Costs in Codex Orchestrator/Subagent Setups via config.toml

Melvin Vivas · X post · 2026-09-07 · Open on X

Topics: AI Agents, Tool Use & MCP, AI Dev Tools & Productivity, LLMOps, Deployment & Monitoring · Level: intermediate

Summary

Melvin Vivas explains how he changed the default model settings in his Codex Astra-Luna Subagents skill after users said it used too many tokens. The root orchestrator (gpt-6-astra) now uses low reasoning effort instead of high. Subagents (gpt-5.6-luna) stay at medium reasoning, and he cut the number of concurrent threads from 6 to 4. He recommends setting up skills and subagents per project, because projects differ in complexity.

Key points

Resources mentioned

Try this

More in AI Agents, Tool Use & MCP

All of AI Agents, Tool Use & MCP