Compare frameworks
DSPy vs smolagents vs DeepEval
Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that framework.
| Field | DSPy | smolagents | DeepEval |
|---|---|---|---|
| Stars | 38,388 (highest of the compared values)Synced from sourceCurrent3h ago | 29,530Synced from sourceCurrent3h ago | 18,476Synced from sourceCurrent3h ago |
| downloads | — | 132,630Synced from sourceStale1mo ago· past SLA | — |
| Last commit | 2026-09-27Synced from sourceCurrent3h ago | 2026-09-23Synced from sourceCurrent3h ago | 2026-09-25Synced from sourceCurrent3h ago |
| Language | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago | pythonSeeded, unreviewedwritten 1mo ago |
| Licence | MITSynced from sourceCurrent3h ago | Apache-2.0Synced from sourceCurrent3h ago | Apache-2.0Synced from sourceCurrent3h ago |
| Archived | noSynced from sourceCurrent3h ago | noSynced from sourceCurrent3h ago | noSynced from sourceCurrent3h ago |
| Category | optimisationSeeded, unreviewedwritten 1mo ago | multi-agentSeeded, unreviewedwritten 1mo ago | evaluationSeeded, unreviewedwritten 1mo ago |
| Contributors | 461Synced from sourceCurrent3h ago | 208Synced from sourceCurrent3h ago | 337Synced from sourceCurrent3h ago |
| Forks | 3372Synced from sourceCurrent3h ago | 2999Synced from sourceCurrent3h ago | 1989Synced from sourceCurrent3h ago |
| Maintainer | Stanford NLPSeeded, unreviewedwritten 1mo ago | Hugging FaceSeeded, unreviewedwritten 1mo ago | Confident AISeeded, unreviewedwritten 1mo ago |
| Open issues | 739Synced from sourceCurrent3h ago | 853Synced from sourceCurrent3h ago | 676Synced from sourceCurrent3h ago |
| PyPI downloads, monthly | — | 614573Synced from sourceStale1mo ago· past SLA | — |
| PyPI downloads, weekly | — | 132630Synced from sourceStale1mo ago· past SLA | — |
| PyPI package | dspy-aiSynced from sourceCurrent3h ago | smolagentsSynced from sourceCurrent3h ago | deepevalSynced from sourceCurrent3h ago |
| PyPI released | 2026-09-25T04:04:30.339197ZSynced from sourceCurrent3h ago | 2026-05-29T05:08:43.280962ZSynced from sourceCurrent3h ago | 2026-09-24T09:31:55.310624ZSynced from sourceCurrent3h ago |
| Requires Python | >=3.9Synced from sourceCurrent3h ago | >=3.10Synced from sourceCurrent3h ago | <4.0,>=3.9Synced from sourceCurrent3h ago |
| PyPI version | 3.4.0Synced from sourceCurrent3h ago | 1.26.0Synced from sourceCurrent3h ago | 4.2.6Synced from sourceCurrent3h ago |
| Repository | https://github.com/stanfordnlp/dspySynced from sourceCurrent3h ago | https://github.com/huggingface/smolagentsSynced from sourceCurrent3h ago | https://github.com/confident-ai/deepevalSynced from sourceCurrent3h ago |
Sources: GitHub REST API, npm registry downloads, PyPI registry metadata, pypistats.org downloads. Most recent observation 3h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.
Change the comparison
Remove one, or open the directory to pick a different set. Up to 3 at a time.