NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
trustworthy ai
3 papers
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
CTRL-ALT-DECEIT Sabotage Evaluations for Automated AI R&D
Evaluating LLM-contaminated Crowdsourcing Data Without Ground Truth