
Lizard AI Runtime
The Windows AI runtime built for real inference
5 followers
The Windows AI runtime built for real inference
5 followers
Most AI apps simply load a model. Lizard plans inference before the first token—optimizing memory, scheduling execution, and running GGUF models locally for faster, more stable AI on Windows.






Loaded a 7B GGUF model on my laptop and was honestly surprised how steady the tokens came through, no weird stuttering mid-response. The pre-planning angle actually shows in practice.
Loaded a 7B GGUF locally and was actually surprised it didn't choke my laptop, scheduling seems smart. The pre-token planning is a real difference if you care about stable inference.