Rethinking Verification for LLM Code Generation: From Generation to Testing
Zihan Ma, Taolin Zhang, Maosong Cao +4 authors
A collaborative method combining human expertise and LLM reasoning enhances test-case generation for code evaluation, improving detection rates and verifier accuracy in benchmarks.