Compare frameworks
Instructor vs AutoGen vs DeepEval
Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that framework.
| Field | Instructor | AutoGen | DeepEval |
|---|---|---|---|
| Stars | 13,939Synced from sourceCurrent4h ago | 61,130 (highest of the compared values)Synced from sourceCurrent4h ago | 18,419Synced from sourceCurrent4h ago |
| downloads | 13,080,962 (highest of the compared values)Synced from sourceStale1mo ago· past SLA | 47,487Synced from sourceAging6d ago | — |
| Last commit | 2026-09-22Synced from sourceCurrent4h ago | 2026-04-15Synced from sourceCurrent4h ago | 2026-09-23Synced from sourceCurrent4h ago |
| Language | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago |
| Licence | MITSynced from sourceCurrent4h ago | CC-BY-4.0Synced from sourceCurrent4h ago | Apache-2.0Synced from sourceCurrent4h ago |
| Archived | noSynced from sourceCurrent4h ago | noSynced from sourceCurrent4h ago | noSynced from sourceCurrent4h ago |
| Category | structured-outputSeeded, unreviewedwritten 1mo ago | multi-agentSeeded, unreviewedwritten 1mo ago | evaluationSeeded, unreviewedwritten 1mo ago |
| Contributors | 267Synced from sourceCurrent4h ago | 534Synced from sourceCurrent4h ago | 337Synced from sourceCurrent4h ago |
| Forks | 1262Synced from sourceCurrent4h ago | 9246Synced from sourceCurrent4h ago | 1973Synced from sourceCurrent4h ago |
| Maintainer | 567 LabsSeeded, unreviewedwritten 1mo ago | MicrosoftSeeded, unreviewedwritten 1mo ago | Confident AISeeded, unreviewedwritten 1mo ago |
| Open issues | 90Synced from sourceCurrent4h ago | 1098Synced from sourceCurrent4h ago | 654Synced from sourceCurrent4h ago |
| PyPI downloads, monthly | 27671654Synced from sourceStale1mo ago· past SLA | 263297Synced from sourceAging6d ago | — |
| PyPI downloads, weekly | 13080962Synced from sourceStale1mo ago· past SLA | 47487Synced from sourceAging6d ago | — |
| PyPI package | instructorSynced from sourceCurrent4h ago | pyautogenSynced from sourceCurrent4h ago | deepevalSynced from sourceCurrent4h ago |
| PyPI released | 2026-09-09T02:25:54.888746ZSynced from sourceCurrent4h ago | 2025-07-15T00:37:26.170442ZSynced from sourceCurrent4h ago | 2026-09-22T19:26:34.903554ZSynced from sourceCurrent4h ago |
| Requires Python | <4.0,>=3.9Synced from sourceCurrent4h ago | >=3.10Synced from sourceCurrent4h ago | <4.0,>=3.9Synced from sourceCurrent4h ago |
| PyPI version | 1.17.0Synced from sourceCurrent4h ago | 0.10.0Synced from sourceCurrent4h ago | 4.2.5Synced from sourceCurrent4h ago |
| Repository | https://github.com/567-labs/instructorSynced from sourceCurrent4h ago | https://github.com/microsoft/autogenSynced from sourceCurrent4h ago | https://github.com/confident-ai/deepevalSynced from sourceCurrent4h ago |
Sources: GitHub REST API, npm registry downloads, PyPI registry metadata, pypistats.org downloads. Most recent observation 4h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.
Change the comparison
Remove one, or open the directory to pick a different set. Up to 3 at a time.