Compare frameworks
Ragas vs DeepEval vs CrewAI
Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that framework.
| Field | Ragas | DeepEval | CrewAI |
|---|---|---|---|
| Stars | 15,834Synced from sourceCurrent3h ago | 18,419Synced from sourceCurrent3h ago | 58,958 (highest of the compared values)Synced from sourceCurrent3h ago |
| Last commit | 2026-02-24Synced from sourceCurrent3h ago | 2026-09-23Synced from sourceCurrent3h ago | 2026-09-23Synced from sourceCurrent3h ago |
| Language | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago |
| Licence | Apache-2.0Synced from sourceCurrent3h ago | Apache-2.0Synced from sourceCurrent3h ago | MITSynced from sourceCurrent3h ago |
| Archived | noSynced from sourceCurrent3h ago | noSynced from sourceCurrent3h ago | noSynced from sourceCurrent3h ago |
| Category | evaluationSeeded, unreviewedwritten 1mo ago | evaluationSeeded, unreviewedwritten 1mo ago | multi-agentSeeded, unreviewedwritten 1mo ago |
| Contributors | 246Synced from sourceCurrent3h ago | 337Synced from sourceCurrent3h ago | 340Synced from sourceCurrent3h ago |
| Forks | 1725Synced from sourceCurrent3h ago | 1973Synced from sourceCurrent3h ago | 8560Synced from sourceCurrent3h ago |
| Maintainer | Vibrant LabsSeeded, unreviewedwritten 1mo ago | Confident AISeeded, unreviewedwritten 1mo ago | CrewAISeeded, unreviewedwritten 1mo ago |
| Open issues | 607Synced from sourceCurrent3h ago | 654Synced from sourceCurrent3h ago | 463Synced from sourceCurrent3h ago |
| PyPI package | ragasSynced from sourceCurrent3h ago | deepevalSynced from sourceCurrent3h ago | crewaiSynced from sourceCurrent3h ago |
| PyPI released | 2026-01-13T17:47:59.200116ZSynced from sourceCurrent3h ago | 2026-09-22T19:26:34.903554ZSynced from sourceCurrent3h ago | 2026-09-16T22:10:41.471882ZSynced from sourceCurrent3h ago |
| Requires Python | >=3.9Synced from sourceCurrent3h ago | <4.0,>=3.9Synced from sourceCurrent3h ago | <3.14,>=3.10Synced from sourceCurrent3h ago |
| PyPI version | 0.4.3Synced from sourceCurrent3h ago | 4.2.5Synced from sourceCurrent3h ago | 1.15.22Synced from sourceCurrent3h ago |
| Repository | https://github.com/vibrantlabsai/ragasSynced from sourceCurrent3h ago | https://github.com/confident-ai/deepevalSynced from sourceCurrent3h ago | https://github.com/crewAIInc/crewAISynced from sourceCurrent3h ago |
Sources: GitHub REST API, npm registry downloads, PyPI registry metadata, pypistats.org downloads. Most recent observation 3h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.
Change the comparison
Remove one, or open the directory to pick a different set. Up to 3 at a time.