
Seer is a browser extension that automatically describes images without alt text using PaliGemma2-3b-mix-224 (a vision-language model), processing all images locally on the user's computer with nothing sent to a server.
Setup requires installing dependencies (fastapi, uvicorn) and running a local daemon on http://127.0.0.1:11435, then loading the extension in Chrome via developer mode; the model files total ~2.4 GB and require ~3 GB RAM with no GPU needed.
Built for people using screen readers (NVDA, JAWS, VoiceOver) who encounter images on websites without descriptions; no internet, account, or subscription required after initial setup.
Licensed under Apache 2.0 by Recursia Lab.
Ask the AI about this article →
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Hugging Face released @huggingface/kernels, a library for running optimized WebGPU kernels from the Hugging Fa…

Israeli startup DataAgent Ltd
Chinese large-model developer Z.ai says it can now support large-scale inference using roughly 100,000 domesti…

Broadcom announced VMware AI Factory, a software-defined foundation for VMware Private AI Cloud, at VMware Exp…

OpenClaw launched version 2.0, its largest update yet, with a version number of 2026.8.1
David Heinemeier Hansson (DHH), creator of Ruby on Rails, has released Omarchy 4.0 (Omarchy Quattro), the late…
