Embeddings and extractive question answering you can test now
running live
Retrieval quality is decided by the embedding model and the reader on top of it. Both are open in your browser here, running on your own GPU through WebGPU with nothing uploaded. Try your own passage and your own question, and see how a given model behaves on your actual text before anyone builds a pipeline around it.