Free Personal Assistant: Hermes Agent + Qwen3.6-35B-A3B-MTP in LM Studio
Melvin Vivas · X post · 2026-07-20 · Open on X
Topics: AI Agents, Tool Use & MCP, LLMOps, Deployment & Monitoring, AI Dev Tools & Productivity · Level: intermediate
Summary
Melvin Vivas is testing Hermes Agent with the local model qwen3.6-35b-a3b-mtp served by LM Studio. If it works well over a few days, he will have a personal AI assistant with no API costs. It shows the pattern of pairing an agent framework with a locally served open-weight model.
Key points
- Agent: Hermes Agent. Model: qwen3.6-35b-a3b-mtp (35B total parameters, about 3B active, with multi-token prediction).
- LM Studio serves the model locally, and the agent connects to it.
- Test the setup for a few days before relying on it as your daily assistant.
- A local model means a free AI assistant with no per-token API costs.
Resources mentioned
- LM Studio · tool · x.com · free
Desktop app for downloading and running LLMs locally, with a developer mode that serves models through an API.
Also in: Running LLMs Locally Without an Expensive Rig (Melvin Vivas on X · notes), Adding Vercel AI Gateway as a provider in AIBackends with Devin (Melvin Vivas on X · notes), LoRA Fine-Tune Qwen3.5-2B on Your Tweets with Unsloth Studio (Melvin Vivas on X · notes), AIBackends: An API Layer Between Your App and AI Models (Now with Jev) (Melvin Vivas on X · notes) and 31 more - Hermes Agent · tool · github.com · free
Nous Research's open-source AI agent with CLI, TUI and desktop interfaces. It now supports hands-free activation with a wake word.
Also in: Inspecting Coding-Agent Traces Live with JSONL Viewer (Codex, Claude Code) (Melvin Vivas on X · notes), Hermes Agent: each bot is its own profile (Melvin Vivas on X · notes), An X research bot built with Hermes (Melvin Vivas on X · notes), Run Your Hermes Agent on Free LFM2.5-2.6B via OpenRouter (Melvin Vivas on X · notes) and 55 more - Qwen 3.6 35B (MTP) · tool · huggingface.co · free
Qwen 3.6 mixture-of-experts model (35B total, about 3B active parameters) with multi-token prediction for local inference.
Also in: Use a Local Model for Confidential Data with Your Agent (Melvin Vivas on X · notes), Running Qwen 3.6 35B locally on an RTX 3090 for agent tool calling (Melvin Vivas on X · notes), Models That Work With the Hermes Agent for Personal Productivity (Melvin Vivas on X · notes), Run Hermes Agent locally with Qwen 3.6 35B MTP in LM Studio (Melvin Vivas on X · notes) and 4 more
Try this
- Try running an agent on a local open-weight model in LM Studio for a few days to see if it can replace paid APIs.
- Set up a free local personal assistant: Hermes Agent connected to a Qwen model served by LM Studio.
More in AI Agents, Tool Use & MCP
- Run Hermes Agent locally with Qwen 3.6 35B MTP in LM Studio
- Using Hermes Agent as a personal assistant to write documents
- Local Personal AI Agent to Explain Your Investment Portfolio
- Agent-Made Video in 10 Minutes: Hermes Agent + Grok 4.5 + Hyperframes
- Personal Assistant Agent on Hermes: Morning Briefings and Inbox Triage
- Hermes Agent Automates LinkedIn Posting