model refinement
- CoLT: The conditional localization test for assessing the accuracy of neural posterior estimates
- Multi-agent KTO: Enhancing Strategic Interactions of Large Language Model in Language Game
- SE-GUI: Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning
- The Surprising Effectiveness of Negative Reinforcement in LLM Reasoning