Vised.ai
Built at The Future of Agentic AI in Healthcare - Abridge x Anthropic x Lightspeed · Jul 18, 2026 · San Francisco, CA
The only API that legacy applications understand is screen / mouse and keyboard. We transcribe display screen to json text format what allows computer use agent to act quickly and at extreamly small cost. Filling a form like this cost 0.2 cents. This works on any interface surface - both browser and native apps. We use both DOM to read browser and AX tree to read native surfaces. Vised gives agents hands and eyes - allow to actuate on any software which can be operated with mouse and keyboard. We do not send screenshots and dont use vision models - just a cheap LLM like gemini 3.1 flash.