Author
Haitam Kadri
Recent research
- AI & ComputingOpen access
When Verification Explores Too Far: Semantic Coverage and Validity in LLM-Generated Code Checks
Large language models are increasingly used not only to generate code, but also to generate tests and other evidence intended to verify that code. This creates a methodological problem: a verifier may appear stronger when it explores behaviors beyond the public examples, while so...
- AI & ComputingOpen access
When Verification Explores Too Far: Semantic Coverage and Validity in LLM-Generated Code Checks
Large language models are increasingly used not only to generate code, but also to generate tests and other evidence intended to verify that code. This creates a methodological problem: a verifier may appear stronger when it explores behaviors beyond the public examples, while so...