Explore/dataset/Taskbench
T

microsoft/TaskbenchUnknown

Taskbench is a collection of benchmark tasks hosted on Hugging Face, designed to evaluate and compare the performance of AI agents across a range of scenarios.

datasetHugging Face
GitHubCompare
Refreshed 15h ago
OverviewActivity52wAlternativesDocs
Stars0
Forks0
HF Downloads1.3k30d
Last commit
Refreshed15h ago
Project healthUnknownNo activity data.
Production readinessResearch / EarlyBest for exploration and prototyping.
Risk notesUnknown licenseVerify license before production use.
AgentHub Score
55 / 100
Composite score from 6 signals. How we score →
Active project
55Score
Growth
40C
Activity
30C
Documentation
70C+
Maturity
45C
Community
42C
Production
58C
GitHub stars · 42 days observed0 not enough history
snapshots
Repository activity · 42 days observedReal snapshots from pushed_at
inactivepushed
2026-07-262026-09-07
Practical assessment
Should you use it?

✓ Best for

  • Research and experimentation
  • Prototype development
  • Learning agentic patterns

◎ Strengths

  • Active community
  • Open source
  • Well-documented API

✕ Not ideal for

  • Untested at scale without validation
  • Teams without AI/ML expertise

⚠ Watch-outs

  • Review changelog before updating
  • Verify license for commercial use
Technical details
What's inside
Language
License
Sourcehf hub
Open source✗ No
Commercial use
Docs
Demo
Paper

AgentHub Score

55
Score 55/100
Below average

Alternatives

N
NexusBench-trajectories
0 · dataset
55
P
Paper2Arm-5135
0 · dataset
55
E
edge-agent-reasoning-websearch-260k
0 · dataset
55
G
ghostui
0 · dataset
55
Compare all →

Recent activity

Latest commit —
Indexed by AgentHub crawler15h ago
Monitor for new releasesongoing