Arena
arena.ai
Put one prompt to two anonymous models and vote for the better answer
chat.webllm.ai
Runs an open language model inside your browser, nothing sent to a server.
A chat interface that downloads an open model into your browser and runs inference on your own GPU through WebGPU. The conversation never leaves the machine — you can pull the network cable mid-chat and it keeps answering, which is the demonstration that makes the point. Model choices range from small instruct models up to a few billion parameters. First load pulls a multi-gigabyte weights file, and it needs a WebGPU-capable browser and a reasonably recent GPU.
Questions
Works the moment the page loads. No account, no email, no trial. The first run downloads gigabytes of model weights and it will not start at all without WebGPU.
No. WebLLM Chat does the work inside your browser, so the file never leaves your device — you can load the page, disconnect from the network, and it still works. That makes it a reasonable choice for material containing personal data.
The first run downloads gigabytes of model weights and it will not start at all without WebGPU.
Alternatives
Same category or same task, ordered by how closely they overlap.
arena.ai
Put one prompt to two anonymous models and vote for the better answer
chatgpt.com
The general assistant most people mean when they say they used AI
claude.ai
Anthropic's assistant, strongest on long documents and careful writing
deepseek.com
Open-weight reasoning model you can use free in a chat window
gemini.google.com
Google's assistant, wired into Search, Gmail, Drive and YouTube
huggingface.co
Thousands of runnable AI demos, no model installed on your machine