AI Engineer Study Library

Qwen 3.5 2B Runs On-Device on iPhone with MLX

Melvin Vivas · X video post · 2026-03-04 · Open on X

Topics: LLM Fundamentals, Industry Trends & Job Market · Level: beginner

Summary

Alibaba's Qwen 3.5 runs locally on an iPhone 17 Pro as a 2B model at 6-bit quantization, using MLX optimized for Apple Silicon. The original poster says it beats models four times its size, has strong visual understanding, and can toggle reasoning on or off. Melvin pitches it as a subscription-free alternative to ChatGPT.

Key points

Resources mentioned

Try this

More in LLM Fundamentals

All of LLM Fundamentals