Launching today
Rondine

Rondine

Run the right local LLM for your hardware

1 follower

Rondine is an open-source control plane for local LLMs. It detects RAM and VRAM, recommends models that fit, applies hardware-tuned settings, downloads weights, and starts an OpenAI-compatible server. It supports Apple Silicon, NVIDIA GPUs, and DGX Spark through llama.cpp, MLX-LM, and vLLM. Instead of creating another inference engine, Rondine coordinates proven runtimes and shows every launch plan before execution.

Rondine makers

Here are the founders, developers, designers and product people who worked on Rondine