Compare frameworks
DeepEval vs AutoGen vs LangChain
Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that framework.
| Field | DeepEval | AutoGen | LangChain |
|---|---|---|---|
| Stars | 18,476Synced from sourceCurrent3h ago | 61,194Synced from sourceCurrent3h ago | 147,178 (highest of the compared values)Synced from sourceCurrent3h ago |
| downloads | — | 47,487Synced from sourceStale10d ago | 3,201,198 (highest of the compared values)Synced from sourceCurrent3h ago |
| Last commit | 2026-09-25Synced from sourceCurrent3h ago | 2026-04-15Synced from sourceCurrent3h ago | 2026-09-27Synced from sourceCurrent3h ago |
| Language | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago |
| Licence | Apache-2.0Synced from sourceCurrent3h ago | CC-BY-4.0Synced from sourceCurrent3h ago | MITSynced from sourceCurrent3h ago |
| Archived | noSynced from sourceCurrent3h ago | noSynced from sourceCurrent3h ago | noSynced from sourceCurrent3h ago |
| Category | evaluationSeeded, unreviewedwritten 1mo ago | multi-agentSeeded, unreviewedwritten 1mo ago | orchestrationSeeded, unreviewedwritten 1mo ago |
| Contributors | 337Synced from sourceCurrent3h ago | 534Synced from sourceCurrent3h ago | 3741Synced from sourceCurrent3h ago |
| Forks | 1989Synced from sourceCurrent3h ago | 9256Synced from sourceCurrent3h ago | 24634Synced from sourceCurrent3h ago |
| Maintainer | Confident AISeeded, unreviewedwritten 1mo ago | MicrosoftSeeded, unreviewedwritten 1mo ago | LangChainSeeded, unreviewedwritten 1mo ago |
| npm downloads, monthly | — | — | 10682493Synced from sourceCurrent3h ago |
| npm downloads, weekly | — | — | 3201198Synced from sourceCurrent3h ago |
| npm package | — | — | langchainSynced from sourceCurrent3h ago |
| npm window ends | — | — | 2026-09-26Synced from sourceCurrent3h ago |
| Open issues | 676Synced from sourceCurrent3h ago | 1105Synced from sourceCurrent3h ago | 569Synced from sourceCurrent3h ago |
| PyPI downloads, monthly | — | 263297Synced from sourceStale10d ago | — |
| PyPI downloads, weekly | — | 47487Synced from sourceStale10d ago | — |
| PyPI package | deepevalSynced from sourceCurrent3h ago | pyautogenSynced from sourceCurrent3h ago | langchainSynced from sourceCurrent3h ago |
| PyPI released | 2026-09-24T09:31:55.310624ZSynced from sourceCurrent3h ago | 2025-07-15T00:37:26.170442ZSynced from sourceCurrent3h ago | 2026-09-18T17:31:10.056247ZSynced from sourceCurrent3h ago |
| Requires Python | <4.0,>=3.9Synced from sourceCurrent3h ago | >=3.10Synced from sourceCurrent3h ago | <4.0.0,>=3.10.0Synced from sourceCurrent3h ago |
| PyPI version | 4.2.6Synced from sourceCurrent3h ago | 0.10.0Synced from sourceCurrent3h ago | 1.4.2Synced from sourceCurrent3h ago |
| Repository | https://github.com/confident-ai/deepevalSynced from sourceCurrent3h ago | https://github.com/microsoft/autogenSynced from sourceCurrent3h ago | https://github.com/langchain-ai/langchainSynced from sourceCurrent3h ago |
Sources: GitHub REST API, npm registry downloads, PyPI registry metadata, pypistats.org downloads. Most recent observation 3h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.
Change the comparison
Remove one, or open the directory to pick a different set. Up to 3 at a time.