How these tools work
- Models are fetched from Hugging Face on first use and cached by your browser.
- Inference runs inside a Web Worker on your GPU; FastVLM.net never receives your images.
- Usage events record only fixed step names, the model and timings—never content.