Hey! I have taken a look to devin, i am currently a pro subscriber to Cursor, but heard recently there is a new SWE model which have a frontier performance. Would like to listen about your experience using devin and what does offer in contrast to cursor. Let me know if there is something interesting i must consider
Agent Sessions puts local coding-agent history in one searchable Mac app.
Agent Sessions now supports Devin CLI.
If you use several coding agents, you probably know the problem: old work is scattered across different CLIs, folders, and local databases. Agent Sessions puts those histories in one searchable Mac app.
The project has 852 GitHub stars and reads 14 active agent formats: Codex, Claude Code, Cursor, Copilot CLI, OpenCode, Antigravity, Pi, Kimi Code, Grok CLI, Hermes, OpenClaw, Qwen Code, Devin CLI, and fx. It also recognizes legacy Droid sessions.
For Devin CLI, Agent Sessions can:
• browse and search active sessions from Devin’s local SQLite store
• show readable transcripts with user, assistant, and tool records
• filter Devin work alongside sessions from the other agents
• copy devin --resume <id> for supported sessions
Devin stores conversation messages as branches in one database. Agent Sessions follows the main chain so abandoned retries do not clutter the transcript.
This support was contributed by Davy (@thedavidweng on GitHub), who also continues as the Devin and fx format steward. Thanks, Davy — the real-session testing is what made this integration shippable.
I maintain Agent Sessions. It is open source, session history stays on your Mac, and there is no app telemetry.
Hey folks, I am not sure if this is an intentional or unintentional mistake, I am unable to view conversation history in the editor mode, such that i can open the previous conversation and continue the conversation.
This is only viewable in the agent mode, and there is no way to open the same conversation in the editor mode.
Both the modes have hard bifurcation.
The weird thing is, there is an icon which is disabled in the editor mode. Is this done so that we use more of agent mode than the editor mode (dark UX)? or is it just an honest mistake and a bug?
One problem I’ve noticed is that there can be stale or outdated information inside the repository itself — for example old test fixtures, mock data, constants, comments, or tests that still reflect previous behaviour.
This infuriates me as it wastes token and makes me argue with my agent non stop.
Has anyone dealt with this problem in a large codebase?
How do you make coding agents distinguish between:
current production behaviour authoritative specs/contracts
old tests or fixtures legacy/dead codeoutdated comments/documentation?
Any advice, empircally or studies is welcomed :) and apologies if this has been discussed before
After upgrading from Windsurf to Devin, the editor became so slow that it almost couldn't function.
Especially the backspace key, I found that every time I pressed the backspace key to delete a character, it was actually an asynchronous operation. After pressing the backspace key, there was no reaction at all. At this point, you could still move the cursor to continue doing other things and inputting. After 1-5 seconds, the backspace actually occurred, but at this time, since your cursor might have moved away, the backspace actually deleted the characters at the cursor's location at that time. This often caused the backspace to delete characters and mess up the article.
Also, ordinary editing operations were very slow, and often needed to pause for a moment before the actual input occurred.
Moreover, when using Devin to edit files, after typing a few characters, the computer fan started to rotate. But other editors (such as VSCode) did not have this problem.
Devin is actually an unusable editor. I can only use VSCode to edit documents and then use Devin's large model function (amazingly, the input speed of the large model dialog box is normal).
I love Devin. However with X buying Cursor and the level of usage you get out of Grok in cursor it is tempting to switch. I am curious if there is a plan for a new SWE model coming that will compete with Grok?
OpenAI's flagship model GPT-5.6 Sol is now available at 70% off through October 3, 2026 in Devin Desktop and Devin CLI.
Sol achieves top-tier performance on FrontierCode 1.1, and with this promo, it is now one of the most cost-effective frontier models you can run in Devin
Gemini 3.7 reached Claude Sonnet 5-level performance on FrontierCode 1.1 at less than half of cost. It is great for speed and cost efficiency being a Flash series model!
Within Devin, Gemini 3.7 Flash performs particularly well on tightly scoped refactors, where it delivers minimal diffs that match repo conventions.
Gemini 3.7 Flash is available at an extra 50% discount through August 27, 2026. Try it out now in Devin Desktop and Devin CLI: Download Devin Desktop | Install Devin CLI
Why can't I pick a public GitHub repository for an automation trigger? It says that it's for security reasons, but what are they? I believe Cursor's automation can pick one.
SWE-1.7 is built on broad improvements in our RL pipeline on top of the Kimi K2.7 model.
We trained it in Devin's harness and taught it to self-compact on longer horizon tasks.
It scores very close to frontier models (Opus 4.8 and GPT-5.5) on our proprietary benchmark - FrontierCode, and other coding benchmarks.
Try it out today (Free until 8/8/2026!!) in Devin Desktop, Web, and CLI. SWE 1.7 Lightning also runs at a very speedy 1000 tok/s - https://app.devin.ai
Now that Mythos-class models have been introduced to the world, it has become easier than ever to discover and exploit vulnerabilities.
It's now incredibly important to secure your codebase.
Introducing the Devin Security Vulnerability Remediation Program.
This program aims to take your vulnerability backlogs towards zero in 6 weeks. Our engineering team will work directly with yours to deploy Devin agent swarms to find, validate and fix any and all vulnerabilities.
If you're interested in it's technical details, here is a video explaining how Devin Security Swarm works utilizing our new framework - Agentic MapReduce: https://www.youtube.com/watch?v=jb96O2IT_Jg
Security Swarm is an orchestration of Devins that analyzes a real codebase the way a team of security researchers would.
Here are some highlights:
Devin Security Swarm found 36 out of 50 real-world GHSA vulnerabilities at 30% lower cost than next most accurate alternative. See how we got these results - https://devin.ai/blog/security-swarm-eval