# Vised.ai

- **Event:** [The Future of Agentic AI in Healthcare - Abridge x Anthropic x Lightspeed](https://cerebralvalley.ai/e/abridge-hackathon)
- **When:** Sat, Jul 18 at 9:00 AM – 10:00 PM (PDT)
- **Where:** San Francisco, CA
- **Team:** [Konrad Sierzputowski](https://cerebralvalley.ai/u/Vised)
- **GitHub:** https://github.com/conrader/abridge-hackaton-vised
- **Demo video:** https://www.youtube.com/@visedai
- **Gallery:** https://cerebralvalley.ai/e/abridge-hackathon/hackathon/gallery
- **Page:** https://cerebralvalley.ai/e/abridge-hackathon/hackathon/gallery/11

The only API that legacy applications understand is screen / mouse and keyboard. We transcribe display screen to json text format what allows computer use agent to act quickly and at extreamly small cost. Filling a form like this cost 0.2 cents. This works on any interface surface - both browser and native apps. We use both DOM to read browser and AX tree to read native surfaces. Vised gives agents hands and eyes - allow to actuate on any software which can be operated with mouse and keyboard. We do not send screenshots and dont use vision models - just a cheap LLM like gemini 3.1 flash.

---

Markdown version of https://cerebralvalley.ai/e/abridge-hackathon/hackathon/gallery/11. Site index for agents: https://cerebralvalley.ai/llms.txt · full text: https://cerebralvalley.ai/llms-full.txt
