This is your work, valued
DPE. From Blind Spots to Gains: Diagnostic-Driven Iterative Training for Large Multimodal Models
SymDPO. We have developed Symbol Demonstration Direct Preference Optimization (SymDPO) and validating its effectiveness across multiple benchmarks.