This is your work, valued

Shanghai, China

Shaobo (Steven) Wang

Expert
@gszfwsb

Intern @QwenLM, Ph.D Candidate @EPIC-Lab-sjtu. Previous Master Student @Thinklab-SJTU. Feel free to contact me or teach me.

NCFM. Official PyTorch implementation of the paper "Dataset Distillation with Neural Characteristic Function: A Minmax Perspective" (NCFM) in CVPR 2025 (Full Score, Highlight).

413

Awesome-HITWH-Resources-Sharing. Awesome Resource Sharing Project for HITWH-CS

102

Awesome-Dataset-Reduction. A curated list of awesome papers on dataset reduction, including dataset distillation (dataset condensation) and dataset pruning (coreset selection).

61

Data-Whisperer. Code for ACL 2025 Main paper "Data Whisperer: Efficient Data Selection for Task-Specific LLM Fine-Tuning via Few-Shot In-Context Learning".

53

OPUS. Code for ICML 2026 Oral paper "OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration."

29

AutoGnothi. Official PyTorch code for ICLR 2025 paper "Gnothi Seauton: Empowering Faithful Self-Interpretability in Black-Box Models"

23

ArtisticCloudBlog. Artistic Neural Style Transfer Software for DIY Stylized images and videos creations.

12

VideoCompressa. Code for VideoCompressa: Data-Efficient Video Understanding via Joint Temporal Compression and Spatial Reconstruction

4

EssenceBench. Official PyTorch implementation of the paper "Rethinking LLM Evaluation: Can We Evaluate LLMs with 200× Less Data" (EssenceBench) in ICLR 2026.

4

Socratic-Zero. Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning

3

nn-visualizer. PyTorch implementation for "Visualizing the Emergence of Intermediate Visual Patterns in DNNs", NeurIPS 2021

3

Q-tuning. Code for paper Q-Tuning: Winning the Pruning Gamble: A Unified Approach to Joint Sample and Token Pruning for Efficient Supervised Fine-Tuning

3

Unveiling-Induction-Heads. PyTorch implementation for "Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers", NeurIPS 2024

2

InfoUtil. Official PyTorch implementation of the paper "Grounding and Enhancing Informativeness and Utility in Dataset Distillation" (InfoUtil) in ICLR 2026.

2

gszfwsb. homepage

1

easy-rl. 强化学习中文教程(蘑菇书),在线阅读地址:https://datawhalechina.github.io/easy-rl/

1

Awesome-In-Context-Learning. Paper List for In-context Learning 🌷

1

Awesome-Dataset-Distillation. A curated list of awesome papers on dataset distillation and related applications.

1