
NobodyWho
Run text, vision and speech models locally on any device
From the NobodyWho blog
Published by NobodyWho, not by us. Every card opens the original post.
Jev in 25 lines of Python
Everyone is talking about Jev - here it is in 25 lines of Python.
LLM inference vs. the OOM killer
Mobile memory warnings and handling them in Rust.
NobodyWho vs Cactus: On-Device Inference Engine Comparison
NobodyWho vs Cactus compared on engine design, model format, hardware, platforms, cloud, and licensing.
Announcing Expo support for NobodyWho
NobodyWho now works with Expo — run on-device LLMs in your Expo apps.
Announcing Voice Activity Detection
Detect when someone starts and stops speaking, on-device.
Use fewer threads for CPU inference
How many worker threads should you use for CPU inference? Not all of them.
Announcing Speech To Text & Text To Speech
STT & TTS in NobodyWho — easily generate and transcribe audio!
NobodyWho Chat app
NobodyWho Chat app is now available on mobile!
Apple Watch & Vision Pro apps
NobodyWho Swift bindings are now available! This article briefly introduces some of the interesting things we dealt with during development.
Announcing Kotlin bindings for NobodyWho
NobodyWho now ships Kotlin bindings — run LLMs fully on-device in your Android and JVM apps, no cloud or server required.
LLM, give me a JSON. Make no mistakes.
So how exactly do you make your LLM output a JSON? What happens under the hood? And how do you make it reliable and fast? Diving into constrained sampling.
Swift Bindings Release
NobodyWho Swift bindings are now available! This article briefly introduces some of the interesting things we dealt with during development.