Dilek Hakkani-Tur
- MIRAGE: A Benchmark for Multimodal Information-Seeking and Reasoning in Agricultural Expert-Guided Conversations
- Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
- Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
- ToolRL: Reward is All Tool Learning Needs