Sida Wang
发表论文 4 篇 · 总被引 1984 次 · h-index 3
代表论文
- LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code (2024 · International Conference on Learning Representations · 被引 2097)
- CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution (2024 · International Conference on Machine Learning · 被引 346)
- What I cannot execute, I do not understand: Training and Evaluating LLMs on Program Execution Traces (2025 · arXiv.org · 被引 10)
- FairCoder: Evaluating Social Bias of LLMs in Code Generation (被引 3)