Skip to main content

A comprehensive code domain benchmark review of LLM researches.

236
GitHub Stars
229
Curated Resources
2
Categories
17 hours ago
Last Refreshed
Surveys🚀 Benchmark Categories

Use this list with your AI agent

Add the Context Awesome MCP server to Claude, Cursor, or any MCP client, then ask:

"Show me program repair, testing & debugging resources from awesome-code-benchmark"

Installation instructions →

What's inside

Showing a sample of 229 resources. View the full list on GitHub →