cost-effectiveness
- Efficient Randomized Experiments Using Foundation Models
- Improving Task-Specific Multimodal Sentiment Analysis with General MLLMs via Prompting
- Neither Valid nor Reliable? Investigating the Use of LLMs as Judges
- Taccel: Scaling Up Vision-based Tactile Robotics via High-performance GPU Simulation
- Web-Shepherd: Advancing PRMs for Reinforcing Web Agents