Launching today

TokenAssemble
Assemble the right local AI stack for your hardware
2 followers
Assemble the right local AI stack for your hardware
2 followers
TokenAssemble tells you which local LLMs your GPU, Mac, or mini-PC can run before you download a model or buy new hardware. It evaluates VRAM or unified-memory fit, quantization, context length, expected speed, and runtime compatibility, then gives you a clear verdict and recommended setup for tools like Ollama, LM Studio, llama.cpp, and vLLM.





ran my macbook through it and honestly the context length breakdown was more useful than i expected, saved me from grabbing a 70b model that would've crawled.