BigCodeBench: Benchmarking Large Language Models on Solving Practical and Challenging Programming Tasks 11 days ago • 30
BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions Paper • 2406.15877 • Published 7 days ago • 39 • 8