Detection, segmentation and depth estimation running in your own browser
running live
Open any of these pages and the model downloads into your browser and runs on your own GPU through WebGPU — nothing is uploaded to a server. The catalogue behind them holds 139 open models across 22 task types, including OCR. It is the fastest way to see what current open vision models do and do not handle on your kind of image.