On-device intelligence

Qwen2.5 0.5B

A compact multilingual assistant, running privately in your browser.

No server inference
Ready · model loads on first message
What can I help you explore? The first response downloads and caches a quantized model. WebGPU-capable browsers offer the best experience.
Enter to send · Shift+Enter for a new line
Generation settings