athina-evals logo

athina-evals

athina-ai/athina-evals

Python SDK for evaluating LLM generated responses

GraphCanon updated 3w · GitHub synced 3w

301 stars22 forksLast push 1y Python

Decision brief

athina-evals is a Python SDK developed for facilitating the evaluation of outputs from large language models through predefined metrics and frameworks.

Good fit when

  • When comprehensive evaluation of LLM responses is required, leveraging athina's specific tools and metrics
  • For users who need an API-based evaluation approach integrated into their workflow

Avoid when

  • If open-source alternatives with transparent customization options are preferred over athina-evals' approach
  • In scenarios where API access requirements limit the ability to perform evaluations offline or in private environments

Observed Jul 17, 2026 · Source: enrich:decision_facts

Verify the decision

Maintenance and security

Full trust report
Maintenance
Dormant (417d since push)
As of 3w
Provenance
Not a fork · Organization account
As of 3w
Security (OSV)
No lockfile
As of 1mo

Public GitHub metadata and optional OSV scans. Signals, not a guarantee. Trust methodology.

Install

pip install athina-evals
PyPI

Similar tools

Same-category neighbours. No typed graph edges are catalogued for this tool yet.

Evidence and technical details

Sourced facts, taxonomy, compatibility claims, README excerpt, and machine-readable endpoints.

Overview

athina-evals offers Python-based evaluation tools and metrics specifically for assessing large language model outputs.

Capability facts

CLI
CLI entrypoint

Source: pyproject.toml:[project.scripts] · Jul 28, 2026

Languages
python

Source: github.language+pyproject.toml · Jul 28, 2026

Categories

Tags

README

Quick Start

Follow this notebook for a quick start guide.

To get an Athina API key, sign up at https://app.athina.ai


For agents

This page has a .md twin and JSON over the API.

Was this helpful?

Anonymous feedback helps us improve pages and translations.