This is your work, valued
PhD student@PolyU
SPA-RL-Agent. Official code for paper "SPA-RL: Reinforcing LLM Agent via Stepwise Progress Attribution"
89STeCa. (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"
29EVU. (ACL2026 Findings) Official code for the paper "Seeing Isn't Believing: Mitigating Belief Inertia via Active Intervention in Embodied Agents"
9COMP6704-Individual-Project. Individual Project for COMP6704: Optimization
4E2CL. (EMNLP24 Findings) Official code for the paper "E2CL: Exploration-based Error Correction Learning for Embodied Agents"
3moco. PyTorch implementation of MoCo: https://arxiv.org/abs/1911.05722
1perfect. PERFECT: Prompt-free and Efficient Few-shot Learning with Language Models
1LMFlow. An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Model for All.
1tuning_playbook_zh_cn. 一本系统地教你将深度学习模型的性能最大化的战术手册。
1event_detection_without_triggers. Python
1huggingface_hub. All the open source things related to the Hugging Face Hub.
1adapter-transformers. Huggingface Transformers + Adapters = ❤️
1eeqa. Event Extraction by Answering (Almost) Natural Questions
1