instruction fine-tuning
- MASTER: Enhancing Large Language Model via Multi-Agent Simulated Teaching
- Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
- OWMM-Agent: Open World Mobile Manipulation With Multi-modal Agentic Data Synthesis
- Rethinking the Role of Verbatim Memorization in LLM Privacy
- Whose Instructions Count? Resolving Preference Bias in Instruction Fine-Tuning