GraphCanon updated 3w · GitHub synced 3w
Decision brief
athina-evals is a Python SDK developed for facilitating the evaluation of outputs from large language models through predefined metrics and frameworks.
Good fit when
- When comprehensive evaluation of LLM responses is required, leveraging athina's specific tools and metrics
- For users who need an API-based evaluation approach integrated into their workflow
Avoid when
- If open-source alternatives with transparent customization options are preferred over athina-evals' approach
- In scenarios where API access requirements limit the ability to perform evaluations offline or in private environments
Observed Jul 17, 2026 · Source: enrich:decision_facts
Verify the decision
Maintenance and security
Full trust report- Maintenance
- Dormant (417d since push)
- As of 3w
- Provenance
- Not a fork · Organization account
- As of 3w
- Security (OSV)
- No lockfile
- As of 1mo
Public GitHub metadata and optional OSV scans. Signals, not a guarantee. Trust methodology.
Install
pip install athina-evals PyPISimilar tools
Same-category neighbours. No typed graph edges are catalogued for this tool yet.
Evidence and technical details
Sourced facts, taxonomy, compatibility claims, README excerpt, and machine-readable endpoints.
Overview
athina-evals offers Python-based evaluation tools and metrics specifically for assessing large language model outputs.
Capability facts
- CLI
- CLI entrypoint
Source: pyproject.toml:[project.scripts] · Jul 28, 2026
- Languages
- python
Source: github.language+pyproject.toml · Jul 28, 2026
Categories
Tags
README
Quick Start
Follow this notebook for a quick start guide.
To get an Athina API key, sign up at https://app.athina.ai
For agents
This page has a .md twin and JSON over the API.