Launching today
OptiQ

OptiQ

Quantize, serve and code with local models on your Mac

2 followers

OptiQ is the local-LLM stack for Apple Silicon. It measures each layer's sensitivity and gives sensitive layers more bits, so its mixed-precision quants beat a flat 4-bit quant at the same size. One MLX-native engine runs a CLI and OpenAI/Anthropic-compatible server, a local web workbench (Lab) and a terminal coding agent (Code). On Gemma-4-12B (M3 Max) it decodes 1.7x faster on prose and up to 2.9x on code edits, identical output. 200+ quants on Hugging Face. No PyTorch, no cloud.
OptiQ gallery image
OptiQ gallery image
OptiQ gallery image
OptiQ gallery image
Free
Launch Team