How AI Works in Glassy
Understand how Glassy runs AI in your browser via WebGPU — two on-device runtimes, what they do, and when cloud AI is an option.
On-device intelligence, plus cloud and self-host paths.
How Glassy runs AI in your browser via WebGPU — related notes, smart tags, and library queries. Then fork into Glassy Cloud (managed AI, BYOK) or Glassy Self-Host (Ollama, no limits).
7 lessons
Understand how Glassy runs AI in your browser via WebGPU — two on-device runtimes, what they do, and when cloud AI is an option.
Discover how Glassy runs 384-dimensional embeddings in your browser via WebGPU to power Related Notes and Smart Tag Suggestions — no cloud, no data leaving your machine.
Use Glassy's AI assistant to ask questions about everything you've saved — with answers grounded in your actual notes, bookmarks, and captures, complete with source citations.
How Glassy Cloud routes AI queries to cloud providers when local models can't handle the job — and what that means for your privacy and costs.
Bring your own API keys for OpenAI, Anthropic, and Gemini to bypass Glassy Cloud's AI cost caps and use the models you already pay for.
Run capable LLMs like Llama and Qwen on your self-hosted Glassy server via Ollama — no API costs, no data leaving your machine, optional GPU acceleration.
How AI billing works on the self-hosted Glassy appliance — no metering, no cost caps, no system cloud keys. BYOK and Ollama are your paths to AI.