-
nuprl/MultiPL-E
Viewer • Updated • 12.7k • 20.6k • 69 -
openai/openai_humaneval
Viewer • Updated • 164 • 219k • 396 -
Big Code Models Leaderboard
📈1.51kExplore code model leaderboard and submit evaluations
-
Copilot Evaluation Harness: Evaluating LLM-Guided Software Programming
Paper • 2402.14261 • Published • 10
Shaun
drgitt
AI & ML interests
None yet
Organizations
None yet
codegen_eval
-
nuprl/MultiPL-E
Viewer • Updated • 12.7k • 20.6k • 69 -
openai/openai_humaneval
Viewer • Updated • 164 • 219k • 396 - RunningAgents1.51k
Big Code Models Leaderboard
📈1.51kExplore code model leaderboard and submit evaluations
-
Copilot Evaluation Harness: Evaluating LLM-Guided Software Programming
Paper • 2402.14261 • Published • 10
Interesting LLMs
datasets 0
None public yet