self-verification
- Generate, but Verify: Reducing Hallucination in Vision-Language Models with Retrospective Resampling
- Incentivizing LLMs to Self-Verify Their Answers
- ShorterBetter: Guiding Reasoning Models to Find Optimal Inference Length for Efficient Reasoning
- VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning