Launching today
VeloxQuant-MLX

VeloxQuant-MLX

Run bigger local LLMs in less memory, on Mac

2 followers

VeloxQuant lets you run bigger local AI models in less memory, fully private with no cloud required. One simple API compresses memory usage up to 16x while keeping generation fast on Apple Silicon.

VeloxQuant-MLX makers

Here are the founders, developers, designers and product people who worked on VeloxQuant-MLX