Launching today

NobodyWho
Run AI models on any device
12 followers
Run AI models on any device
12 followers
NobodyWho is an inference engine for running LLMs fully on-device, built on llama.cpp. Open-source, free, no API keys, no cloud calls. We support Swift, Kotlin, Flutter, React Native, Python, and Godot. Includes type-safe tool calling with automatic grammar generation, multimodal input, Text-to-Speech & Speech-to-Text, GPU acceleration via Vulkan & Metal, and Hugging Face model downloads.




NobodyWho
Hey, β¨I'm Pierre from NobodyWho π
We've spent the last months getting local inference to be production-ready across six platforms and frameworks, not just a cool demo that works on one device.
With NobodyWho you can:
- Get answers from any open-weight AI models: Gemma, Qwen, LFM...
- Analyse images and audio through multimodal inputβ¨
- Transcribe speech to text with any Whisper models
- Generate natural-sounding speech with Supertonic, Pocket TTS and Kokoro models
β¨- Tool calling with guaranteed schema-valid output, the grammar is built from your function signature so the model can't return malformed JSON
β¨- Run long conversations without hitting a hard message-length wall, thanks to preemptive context shifting
Wanna try our work on your device? We've built a few demo apps: iOS, Android, Apple Watch & Vision Pro.
We've also built starter examples to get started in 5 minutes and a model selection page.
NobodyWho inference engine is open-source & free, please leave a star to support us on Github π
Happy to answer any questions :)