Compare frameworks
DeepEval vs Guardrails vs Letta
Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that framework.
| Field | DeepEval | Guardrails | Letta |
|---|---|---|---|
| Stars | 18,476Synced from sourceCurrent5h ago | 7,456Synced from sourceCurrent5h ago | 24,923 (highest of the compared values)Synced from sourceCurrent5h ago |
| Last commit | 2026-09-25Synced from sourceCurrent5h ago | 2026-09-25Synced from sourceCurrent5h ago | 2026-09-10Synced from sourceCurrent5h ago |
| Language | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago |
| Licence | Apache-2.0Synced from sourceCurrent5h ago | Apache-2.0Synced from sourceCurrent5h ago | Apache-2.0Synced from sourceCurrent5h ago |
| Archived | noSynced from sourceCurrent5h ago | noSynced from sourceCurrent5h ago | noSynced from sourceCurrent5h ago |
| Category | evaluationSeeded, unreviewedwritten 1mo ago | validationSeeded, unreviewedwritten 1mo ago | memorySeeded, unreviewedwritten 1mo ago |
| Contributors | 337Synced from sourceCurrent5h ago | 79Synced from sourceCurrent5h ago | 155Synced from sourceCurrent5h ago |
| Forks | 1989Synced from sourceCurrent5h ago | 707Synced from sourceCurrent5h ago | 2631Synced from sourceCurrent5h ago |
| Maintainer | Confident AISeeded, unreviewedwritten 1mo ago | Guardrails AISeeded, unreviewedwritten 1mo ago | LettaSeeded, unreviewedwritten 1mo ago |
| Open issues | 676Synced from sourceCurrent5h ago | 79Synced from sourceCurrent5h ago | 0Synced from sourceCurrent5h ago |
| PyPI package | deepevalSynced from sourceCurrent5h ago | guardrails-aiSynced from sourceCurrent5h ago | lettaSynced from sourceCurrent5h ago |
| PyPI released | 2026-09-24T09:31:55.310624ZSynced from sourceCurrent5h ago | 2026-08-14T14:22:27.750056ZSynced from sourceCurrent5h ago | 2026-09-28T04:32:31.896015ZSynced from sourceCurrent5h ago |
| Requires Python | <4.0,>=3.9Synced from sourceCurrent5h ago | <3.14,>=3.10Synced from sourceCurrent5h ago | >=3.9Synced from sourceCurrent5h ago |
| PyPI version | 4.2.6Synced from sourceCurrent5h ago | 0.11.0Synced from sourceCurrent5h ago | 0.33.5Synced from sourceCurrent5h ago |
| Repository | https://github.com/confident-ai/deepevalSynced from sourceCurrent5h ago | https://github.com/guardrails-ai/guardrailsSynced from sourceCurrent5h ago | https://github.com/letta-ai/lettaSynced from sourceCurrent5h ago |
Sources: GitHub REST API, npm registry downloads, PyPI registry metadata, pypistats.org downloads. Most recent observation 5h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.
Change the comparison
Remove one, or open the directory to pick a different set. Up to 3 at a time.