# BrownieSquad

- **Event:** [OpenEnv Hackathon SF](https://cerebralvalley.ai/e/openenv-hackathon-sf)
- **When:** Mar 7 at 9:00 AM – Mar 8 at 12:00 AM (PST)
- **Where:** Shack15, San Francisco, CA
- **Team:** [Sai Pranav Sripathi](https://cerebralvalley.ai/u/SaiPranav), [Sidhartha Mani](https://cerebralvalley.ai/u/wlan0), [Aman Nindra](https://cerebralvalley.ai/u/Amanmonster)
- **GitHub:** https://github.com/amannindra/RL_Hackathon
- **Website:** https://github.com/amannindra/RL_Hackathon/blob/main/server/softmax_surrogate_environment.py
- **Demo video:** https://drive.google.com/file/d/1pLhRHr2oHFP6DIwttfpk89Uyvipd-1AK/view?usp=sharing
- **Hugging Face:** https://huggingface.co/spaces/openenv-community/RL_Surrogate_ENV
- **Gallery:** https://cerebralvalley.ai/e/openenv-hackathon-sf/hackathon/gallery
- **Page:** https://cerebralvalley.ai/e/openenv-hackathon-sf/hackathon/gallery/72

We started with a surrogate discovery env and foudn that it is excellent and beats pytorch.compile (36x) and trioton.autotune (3x) using a gaussian process optimized by out oracle. 

Since the oracle was so powerful we used it for generative CUDA kernel optimization wth a online self-improvement loop using unsloth lora and DPO.

---

Markdown version of https://cerebralvalley.ai/e/openenv-hackathon-sf/hackathon/gallery/72. Site index for agents: https://cerebralvalley.ai/llms.txt · full text: https://cerebralvalley.ai/llms-full.txt
