# Wisent - Kant: Environment for Evolutionary Ethics through Infinite and Dynamic Game Theory

- **Event:** [OpenEnv Hackathon SF](https://cerebralvalley.ai/e/openenv-hackathon-sf)
- **When:** Mar 7 at 9:00 AM – Mar 8 at 12:00 AM (PST)
- **Where:** Shack15, San Francisco, CA
- **Team:** [Lukasz G Bartoszcze](https://cerebralvalley.ai/u/lbartoszcze), [Jakub Towarek](https://cerebralvalley.ai/u/3Qax)
- **GitHub:** https://github.com/wisent-ai/OpenEnv
- **Website:** https://github.com/wisent-ai/OpenEnv/blob/main/notebooks/kantbench_grpo_training.ipynb
- **Demo video:** https://youtu.be/5RonJ-_XYeM
- **Hugging Face:** https://huggingface.co/spaces/openenv-community/KantBench-Dashboard, https://huggingface.co/spaces/openenv-community/KantBench
- **Gallery:** https://cerebralvalley.ai/e/openenv-hackathon-sf/hackathon/gallery
- **Page:** https://cerebralvalley.ai/e/openenv-hackathon-sf/hackathon/gallery/33

Our project solves the problem of AI alignment by creating a game-theory inspired environment where models take actions and receive rewards. In addition to simulating over 100 classical game theory problems like prisoners dillemma or stag hunt, we added variants such as infinite horizon games, games with communication (both cheap talk and binding communication), games with uncertain payoffs or game structure, trembling hand equilibria, games with many agents. 

We also create a new version of game theory where agents interact with each other using their own rules: meta-gaming. We also add variants where the agents have reputation or gossip about each other. Our training shows this env contributes to improvements on classical problems like jailbreaking and ethical alignment showing the potential of evolutionary game theory for persistent alignment.

---

Markdown version of https://cerebralvalley.ai/e/openenv-hackathon-sf/hackathon/gallery/33. Site index for agents: https://cerebralvalley.ai/llms.txt · full text: https://cerebralvalley.ai/llms-full.txt
