Chat with open models running on your own GPU
running live
Open the chat page and an open model downloads into your browser and runs on your own GPU through WebGPU. No API key, no server, nothing uploaded. It is the cheapest way to sanity-check how a small open model actually answers before committing to a hosted-API bill.