WebLLM · 100% in-browser
Free AI chat, no sign-up — with local inference
Most "free" AI chats hand you a login form, a message cap, or both. WKO AI runs open LLMs like Qwen directly in your browser with WebLLM and WebGPU. Your chat content stays local unless you enable an external search provider.
Why it's different
Free means free — and private means private
-
Truly no sign-up
No email, no account, no "free trial" gate. The chat is ready the moment the model loads.
-
Private by architecture
The LLM runs on your device via WebLLM and WebGPU. AI content stays local unless you enable an external search provider.
-
No message limits
No daily cap, no throttling, no queue. Chat as long as your browser tab is open.
-
Free forever
Open models, cached in your browser after a one-time download. Nothing to subscribe to.
More than a blank prompt
Local documents and voice, with search when you choose
-
PDF knowledge (RAG)
Drop in a PDF and the chat answers from it — parsed and embedded locally, page by page.
-
Optional web search
Bring your own API key to let the assistant pull in live web results when you want fresh answers.
-
Voice mode
Speak your prompts and hear replies with local Whisper transcription and Kokoro TTS.
One click between you and the model
Pick a model, let it download once, and chat as much as you like. No account, no cap, and local inference.