# Atlantis

- **Event:** [AI Engineer World's Fair Hackathon 2026](https://cerebralvalley.ai/e/aiewf-hackathon-2026)
- **When:** Jun 27 at 9:00 AM – Jun 28 at 5:00 PM (PDT)
- **Where:** San Francisco, CA
- **Team:** [Yiying Xie](https://cerebralvalley.ai/u/Irene_xie)
- **GitHub:** https://github.com/Bestpart-Irene/RL-persona-red-team-agent
- **Demo video:** https://youtu.be/DN-OdHUR5Z0
- **Gallery:** https://cerebralvalley.ai/e/aiewf-hackathon-2026/hackathon/gallery
- **Page:** https://cerebralvalley.ai/e/aiewf-hackathon-2026/hackathon/gallery/6

An RL agent (GRPO + LoRA) that learns the person-specific privacy attacks universal safety filters miss — eliciting "allowed-but-compromising" leaks judged against each person's contextual-integrity care vector, and recursively self-improving its own weights and its own curriculum of harder targets.

---

Markdown version of https://cerebralvalley.ai/e/aiewf-hackathon-2026/hackathon/gallery/6. Site index for agents: https://cerebralvalley.ai/llms.txt · full text: https://cerebralvalley.ai/llms-full.txt
