A
AgentBench
@THUDM
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
Apache-2.0 Commit 8 maanden geleden ★ 3.753
Pijlers en weging
Waarom deze score
- Apache-2.0 (permissief)
- Laatste commit 231 dagen geleden
- Top 36% naar sterren in AI & ML › Grote taalmodellen
Feiten
- Licentie
- Apache-2.0 (Permissief)
- Taal
- Python
- GitHub-sterren
- 3.753
- Laatste commit
- 2026-02-08 (8 maanden geleden)
- Laatste release
- Onbekend
- OpenSSF Scorecard
- Nog niet gemeten
- Zelf hosten
- Niet vastgesteld
- Maintainer-locatie
- Onbekend
- Bron
- github-crawl
Alternatieven in Grote taalmodellen
01 88 02 87 03 84 04 80
D
Dify.ai
Build, test and deploy LLM applications.
L
Langfuse
LLM engineering platform for model tracing, prompt management, and application evaluation. Langfuse helps teams collaboratively debug, analyze, and iterate on their LLM applications such as chatbots or AI agents.
T
textgen
Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.
F
Firecrawl
The web data API to search, scrape, and interact at scale. 🔥