Skip to Main Content

VisionOps

Built at Zero to Agent: Vercel x Deepmind Hackathon SF · Mar 21, 2026 · San Francisco, CA

Demo video · www.loom.com/…

VisionOps is a multimodal AI application that turns first-person video into a structured, shareable field report. Users upload a POV clip from smart glasses or a phone, and the app samples key moments across the full video, detects visible objects, generates timestamped Gemini narration, and assembles everything into a synchronized playback experience with overlays and a live event feed. The product is useful because raw POV footage is hard to review, summarize, and share quickly. VisionOps makes that footage understandable by converting it into an AI-assisted replay that highlights what happened, when it happened, and what was visible in the scene. It also adds a share capsule layer so users can hand off the clip and recap into communication tools like Slack or WhatsApp, making the experience easier to consume for teammates, friends, or anyone who was not there.

Team