Logo
Explore Help
Register Sign In
wuyang/zk-data-agent
1
0
Fork 0
You've already forked zk-data-agent
Code Issues Pull Requests Actions Packages Projects Releases Wiki Activity
Files
c17c2768ebe734d92d1b498b6df0be17f7fd8146
zk-data-agent/benchmarks/suites
T
History
Abdelrahman Abdallah a54c90b18f update the codebase and clean up it
2026-04-06 03:42:44 +02:00
..
__init__.py
Add 10 standard evaluation benchmark suites with CLI runner and README
2026-04-05 19:58:00 +00:00
aider.py
Add 10 standard evaluation benchmark suites with CLI runner and README
2026-04-05 19:58:00 +00:00
aime.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
base.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
bfcl.py
Fix review comments: remove empty string concat, fix variable scoping in BFCL evaluate
2026-04-05 19:59:16 +00:00
gsm8k.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
humaneval.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
ifeval.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
livecodebench.py
Add 10 standard evaluation benchmark suites with CLI runner and README
2026-04-05 19:58:00 +00:00
math_bench.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
mbpp.py
update the codebase and clean up it
2026-04-06 03:42:44 +02:00
swe_bench.py
Add 10 standard evaluation benchmark suites with CLI runner and README
2026-04-05 19:58:00 +00:00
Powered by Gitea Version: 1.26.2 Page: 46ms Template: 1ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API