This small PR resolves the deprecation warnings of the `logger` library: ``` DeprecationWarning: The 'warn' method is deprecated, use 'warning' instead ``` |
||
---|---|---|
.. | ||
agbenchmark | ||
agbenchmark_config | ||
backend | ||
frontend | ||
reports | ||
tests | ||
.env.example | ||
.flake8 | ||
.gitignore | ||
LICENSE | ||
README.md | ||
agents_to_benchmark.json | ||
poetry.lock | ||
pyproject.toml | ||
run.sh |
README.md
Auto-GPT Benchmarks
Built for the purpose of benchmarking the performance of agents regardless of how they work.
Objectively know how well your agent is performing in categories like code, retrieval, memory, and safety.
Save time and money while doing it through smart dependencies. The best part? It's all automated.
Scores:
Ranking overall:
Detailed results:
Click here to see the results and the raw data!!
More agents coming soon !