Logo
Explore Help
Register Sign In
wuyang/zk-data-agent
1
0
Fork 0
You've already forked zk-data-agent
Code Issues Pull Requests Actions Packages Projects Releases Wiki Activity
Files
67e6dd6ae14e2c2251eadbcb76d63cd0fb5e024e
zk-data-agent/benchmarks/suites
T
History
Abdelrahman Abdallah a54c90b18f update the codebase and clean up it
2026-04-06 03:42:44 +02:00
..
__init__.py
Add 10 standard evaluation benchmark suites with CLI runner and README
2026-04-05 19:58:00 +00:00
aider.py
Add 10 standard evaluation benchmark suites with CLI runner and README
2026-04-05 19:58:00 +00:00
aime.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
base.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
bfcl.py
Fix review comments: remove empty string concat, fix variable scoping in BFCL evaluate
2026-04-05 19:59:16 +00:00
gsm8k.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
humaneval.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
ifeval.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
livecodebench.py
Add 10 standard evaluation benchmark suites with CLI runner and README
2026-04-05 19:58:00 +00:00
math_bench.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
mbpp.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
swe_bench.py
Add 10 standard evaluation benchmark suites with CLI runner and README
2026-04-05 19:58:00 +00:00
Powered by Gitea Version: 1.26.2 Page: 48ms Template: 1ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API