IterResearch. Python
66C-3PO. [ICML2025] The official implementation of "C-3PO: Compact Plug-and-Play Proxy Optimization to Achieve Human-like Retrieval-Augmented Generation"
44ReForm. Python
21SEER. Python
15ToolEVO. Python
12MPrompt. Python
10FCS-HGNN. Python
8CIE. Python
7code. Python
4IR_2021.
2AgentsMeetRL. An Awesome List of Reinforcement Learning-based Large Language Agent Works. Collect directly from official code base.
1Awesome-LLM-Strawberry. A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 and reasoning techniques.
1slime. slime is a LLM post-training framework aiming at scaling RL.
1Chen-GX.github.io. JavaScript
1rllm. Democratizing Reinforcement Learning for LLMs
1