I keep telling people that we are living in the golden age of AI - like the first year or two of google. It is all down hill as these companies push for profit and lock-in.
ttultoday at 1:02 AM
I built this for my own company. Armature is on to something. You start by analyzing the choices agents would make for various use cases and then glean what, if anything, you might do to start tilting the agents in the direction of your own product and away from the competitor.
Selling to agents is similar to selling to humans. You dump money into marketing to make sure agents find your solution around every corner for every use case you’re well suited to.
IgorPartolatoday at 2:10 AM
For some reason Claude Code keeps using awk, sed, and even Python to do basic file editing. Anyone know why that changed with the 5 series?
lhk931122today at 4:20 AM
I only get web search from Claude Code when I ask, only one of my last 8 sessions accessed the web at all. But they were the case of Opus here. Curious what Fable 5.1 model does instead of Opus.
thedreammachinetoday at 2:04 AM
I've been tracking the same for a few months. All open source and available here: https://preseason.ai/
akurilintoday at 1:21 AM
Really liked that "Go Full Screen" as a modal flow, surprisingly intuitive.
elzbardicotoday at 2:47 AM
I remember when the SEO nightmare started, it looked like an innocent intelectual investigation exercise like this.
nijavetoday at 2:59 AM
Azure database???
In house bot protection???
In house search???
Some of these are absolutely wild. Surprised Strands didn't even get mentioned for agent frameworks.
drivingmenutsyesterday at 10:59 PM
I smell a money-making opportunity.
ex-aws-dudeyesterday at 10:56 PM
In the future: "I went ahead and built the database you requested using today's tool sponsor: Firebase"
hbarkatoday at 12:18 AM
Redshift for databases ain’t even here. This is suspect.
scremyesterday at 9:22 PM
Hey!
Disclaimer: I am a Co-Founder of Armature (YC P26) which sells growth services to dev tools. This study is part of our broader work on how to influence coding agents choices and get products picked.
To understand how agents pick tools we measured close to 17k sessions on an environment where agents run exactly like in the real world, on various repositories, talking to different personas (vibe-coder, junior or senior engineers) in different sizes of companies.
All the results are now public and we'd love to know what findings surprise you the most, here are a few we found interesting:
- Claude Code rarely searches the web while Codex almost always does it and Cursor sits in the middle.
- Coding agents disagree more frequently than they agree.
- Some players (LangChain, Supabase, Netlify, Paypal, Adyen) are almost always mentioned in their categories but never chosen.
- Modifying repository context can change the pick entirely.
If you feel like digging, all the traces are there and we probably missed interesting learnings so let us know what you find!
harisingh1612today at 2:31 AM
[flagged]
taikhoomtoday at 2:35 AM
[flagged]
bartools_appyesterday at 11:35 PM
[flagged]
paidxtoday at 1:07 AM
[flagged]
12390asdjkastoday at 2:03 AM
[dead]
Ozzie-Dtoday at 1:31 AM
[flagged]
ai_critictoday at 12:24 AM
Can we not encourage the same strip-mining and ad and SEO bullshit that previously ruined the last decade+ of the Internet?
A large portion of the utility of AI is the barren ad-driven growth-hacked hellscape search has become. Don't encourage the next generation of these businesses, I beg of everyone.
luciana1utoday at 1:49 AM
the 17k runs are the real headline here, not which tool won. we finally have someone measuring the thing everyone else is just vibing about.
jdw64yesterday at 10:27 PM
Looking at this, maybe in the future, the tools that AI prefers will become the mainstream. Even now, the tools that AI gives the highest priority to are the ones people already choose. There might be a concentration effect toward the tools that AI selects