Side by side
Compare two AI models
Pick two models and see what each costs, how well it works and what it's actually for. The scores are my own judgment on a 1–5 scale from using these — not vendor claims and not benchmark results. Your picks stay in the address bar, so you can send the comparison to someone else.
Choose a second model above to see them side by side.
Coding assistants
Devin
Cognition
VisitFree value
Performance
Range of uses
Listed for its free documentation — the product itself is a paid subscription.
- Published benchmark scores
- SWE-bench (original, unassisted)13.86%
share of real GitHub issues resolved end to end
Cognition's own technical report · 15 March 2024 · maker's own figures
The maker's own figure, and now two years old.
- SWE-bench Verified48.2% or 61.7%, depending on the source
share of human-checked GitHub issues resolved end to end
Dataku (48.2%) and Tensorfeed (61.7%) · 9 June 2025 / undated · independent
The two aggregators contradict each other and neither is the official leaderboard. Treat as unverified.
Devin's numbers are a mess to cite honestly: Cognition's only first-party report predates the SWE-bench Verified set, and the two aggregators that list a current figure disagree by 13 points. Both are below.
Figures copied from the sources above, checked 17 September 2026.
- Cost
- Free tier and free docs; Pro $20/month, Max $200/month.
- Performance
- Aims to take a whole ticket end to end rather than answer one question. Impressive on well-scoped tasks, expensive when it wanders.
- Use it for
- Understanding how autonomous coding agents are designed
- Keep in mind
- Listed here for the free documentation and free tier — the useful quotas are paid.
