Running Local Qwen3.8-27B in Codex on a Long Goal
Melvin Vivas · X post · 2026-08-30 · Open on X
Topics: AI Dev Tools & Productivity, LLM Fundamentals · Level: intermediate
Summary
The creator runs the local open-weight model Qwen3.8-27B inside Codex. It had been working on a single goal for 31 minutes, which shows a local model can handle long autonomous coding tasks.
Key points
- Codex can be pointed at a local model such as Qwen3.8-27B.
- The model worked on one goal autonomously for 31 minutes and was still going.
- Local models are a way to avoid cloud usage limits.
Resources mentioned
- Qwen3.8-27B · tool · huggingface.co · free
A 27B-parameter open-weight model from Alibaba's Qwen family that you can run locally or call through hosted APIs.
Also in: Deploying Qwen3.8 27B on Hugging Face Inference Endpoints (Melvin Vivas on X · notes), Using Hugging Face credits: Jobs, Inference Endpoints and Open Models (Melvin Vivas on X · notes), Running Qwen3.8-27B Locally on an M5 Max MacBook with Inco Splash (Melvin Vivas on X · notes), Ternary Bonsai 2 27B: 9x smaller model keeping 98.2% of benchmark scores (Melvin Vivas on X · notes) and 24 more - OpenAI Codex · tool · openai.com · paid
OpenAI's coding agent. In the diagram it writes code, fixes review findings and drives the build loop. The creator also used it to make this video.
Also in: An agent bot that installs and drives Codex on its own (Melvin Vivas on X · notes), Asking a Coder bot to install Codex (Melvin Vivas on X · notes), Sign in with ChatGPT: Setting Usage Limits for Each App (Melvin Vivas on X · notes), Codex Cloud Environments Must Be Saved & Published Before Use (Melvin Vivas on X · notes) and 240 more
More in AI Dev Tools & Productivity
- Update Pi to Get the Thinking Level Selector
- Codex Workflow: Try GPT-5.6 Luna First, Escalate to Sol
- GPT-5.6 Luna at Highest Reasoning Handled a Release in Codex
- Codex + Local Qwen3.8-27B (Q4_K_M) with /goal
- Plan With a Frontier Model, Code With a Local Quantized Qwen
- Synthetic (synthetic.new): A Clear Way to Show LLM Subscription Usage Limits