HomeAI & Machine LearningPersonal AI Assistants: On-Device vs Cloud LLM Inference Simulator

🤖 Personal AI Assistants: On-Device vs Cloud LLM Inference Simulator

Watch tokens flow through an on-device LLM (Phi-3, Gemma, Llama 3) versus a cloud round trip. Switch model, quantization and privacy mode and see throughput, latency and privacy score change live.

AI & Machine Learning3DModerate60 FPS
personal-ai-assistants-on-device-vs-cloud-llm-inference ↗ Open standalone
⚙ Under the hood

Watch tokens flow through an on-device LLM (Phi-3, Gemma, Llama 3) versus a cloud round trip. Switch model, quantization and privacy mode and see throughput, latency and privacy score change live.

Three.jsAI assistantson-device LLMprivacyInstancedMesh

3D · Three.js / WebGL renderer · 60 FPS target · runs fully client-side, no install

What did you find?

Add reproduction steps (optional)