Launching today
VeloxQuant-MLX

VeloxQuant-MLX

Run bigger local LLMs in less memory, on Mac

2 followers

VeloxQuant lets you run bigger local AI models in less memory, fully private with no cloud required. One simple API compresses memory usage up to 16x while keeping generation fast on Apple Silicon.

VeloxQuant-MLX Reviews

Reviews