- 🔭 I’m recently working on these repositories. Take a look if you are interested!
- 🚀 I’m always open for collaboration and you can find more about me via my personal website.
Forked from unimpor/T3
(ICLR 2026 Oral) Code for the paper: Reducing Belief Deviation in Reinforcement Learning for Active Reasoning of LLM Agents
Python 1
[NeurIPS 2024] Discovery of the Hidden World with Large Language Models