NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
vulnerability detection
3 papers
CoRe: Benchmarking LLMs’ Code Reasoning Capabilities through Static Analysis Tasks
SECODEPLT: A Unified Benchmark for Evaluating the Security Risks and Capabilities of Code GenAI
Transforming Generic Coder LLMs to Effective Binary Code Embedding Models for Similarity Detection