AI Engineer Study Library

Run an abliterated Qwen3.8-27B GGUF locally with llama-server

Melvin Vivas · X post · 2026-09-09 · Open on X

Topics: AI Safety, Security & Guardrails, LLMOps, Deployment & Monitoring, LLM Fundamentals · Level: intermediate

Summary

The creator shares an uncensored (abliterated) GGUF build of Qwen3.8-27B on Hugging Face and says to use it only for education or red teaming. He shows how to serve it locally with llama.cpp's llama-server and a models.ini preset file. He tested it on an RTX 3090.

Key points

Resources mentioned

Try this

More in AI Safety, Security & Guardrails

All of AI Safety, Security & Guardrails