🤖 Personal AI Assistants: On-Device vs Cloud LLM Inference Simulator
Watch tokens flow through an on-device LLM (Phi-3, Gemma, Llama 3) versus a cloud round trip. Switch model, quantization and privacy mode and see throughput, latency and privacy score change live.
AI & Machine Learning3DModerate60 FPS
⚙ Under the hood
Watch tokens flow through an on-device LLM (Phi-3, Gemma, Llama 3) versus a cloud round trip. Switch model, quantization and privacy mode and see throughput, latency and privacy score change live.
Three.jsAI assistantson-device LLMprivacyInstancedMesh
3D · Three.js / WebGL renderer · 60 FPS target · runs fully client-side, no install