Using Hugging Face credits: Jobs, Inference Endpoints and Open Models
Melvin Vivas · X post · 2026-10-01 · Open on X
Topics: LLMOps, Deployment & Monitoring, AI Agents, Tool Use & MCP · Level: intermediate
Summary
The creator thanks Victor Mustar of Hugging Face for a credit giveaway and plans to use the credits to deploy Qwen3.8 27B on an Inference Endpoint. The quoted post lists things you can try with HF credits: having agents launch HF Jobs, running DeepSeek V4.1 Flash, deploying dedicated endpoints, and running agents with Blender on a GPU.
Key points
- Victor Mustar gave away $1,000 in HF credits ($10 each to 100 people).
- Things to try with the credits: have an agent launch hundreds of HF Jobs, run about 33M tokens of DeepSeek V4.1 Flash, and deploy a dedicated Qwen3.8 27B endpoint.
- The creator plans to deploy Qwen3.8 27B on an Inference Endpoint.
Resources mentioned
- Victor M (@victormustar) on X · person · x.com · free
Hugging Face team member who posts about HF products, open models and agent workflows. - Hugging Face · website · x.com · free · recommended by both Bashiri Smith & Melvin Vivas
Platform for hosting and finding ML models, datasets and papers. The quoted post says LocateAnything was trending there.
Also in: Using an ML agent to train an open-source TTS model on your voice (Melvin Vivas on X · notes), Deploy Open-Source Models with Hugging Face Inference Endpoints (Melvin Vivas on X · notes), Deploying Qwen3.8 27B on Hugging Face Inference Endpoints (Melvin Vivas on X · notes), Unsloth passes 500M model downloads on Hugging Face (Melvin Vivas on X · notes) and 38 more - Hugging Face Inference Endpoints · tool · endpoints.huggingface.co · paid
Managed service for deploying models from the Hugging Face Hub on dedicated infrastructure.
Also in: Deploy Open-Source Models with Hugging Face Inference Endpoints (Melvin Vivas on X · notes), Deploying Qwen3.8 27B on Hugging Face Inference Endpoints (Melvin Vivas on X · notes) - Hugging Face Jobs · tool · huggingface.co · paid
Hugging Face service for running compute jobs on HF infrastructure, and agents can launch them. - DeepSeek V4.1 Flash · tool · api-docs.deepseek.com · paid
A fast DeepSeek language model available through DeepSeek's official API.
Also in: DeepSeek 4.1 Flash speed on the official API: about 325 tokens/s (Melvin Vivas on X · notes), One-Shot Agent Management UI with DeepSeek Harness for $0.19 (Melvin Vivas on X · notes), DeepSeek Harness: Open-Source, Browser-Based Agent for DeepSeek V4.1 (Melvin Vivas on X · notes), DeepSeek V4.1 Flash Off-Peak Pricing as a Cheap Fallback Model (Melvin Vivas on X · notes) and 1 more - Qwen3.8-27B · tool · huggingface.co · free
A 27B-parameter open-weight model from Alibaba's Qwen family that you can run locally or call through hosted APIs.
Also in: Deploying Qwen3.8 27B on Hugging Face Inference Endpoints (Melvin Vivas on X · notes), Running Qwen3.8-27B Locally on an M5 Max MacBook with Inco Splash (Melvin Vivas on X · notes), Ternary Bonsai 2 27B: 9x smaller model keeping 98.2% of benchmark scores (Melvin Vivas on X · notes), Free Qwen-3.8 27B Model via Infron (Melvin Vivas on X · notes) and 24 more
Try this
- Have an agent launch Hugging Face Jobs to run batch workloads automatically.
More in LLMOps, Deployment & Monitoring
- Deploy Open-Source Models with Hugging Face Inference Endpoints
- Deploying Qwen3.8 27B on Hugging Face Inference Endpoints
- FireRouter: cost-aware model routing between open models and Claude Opus
- Run Laya Decision models locally with Unsloth on 4GB RAM
- AI Backends: A Production AI Workflow Engineering Site (Link Share)