Posts tagged: ai

All posts with the tag "ai"

110 posts latest post 2026-09-24
Publishing rhythm
Sep 2026 | 2 posts
Steve Yegge Steve Yegge: I was chatting with my buddy at Google, who's been a tech director there for about 20 years, about their AI adoption. Craziest convo I've had all year. … Simon Willison’s Weblog · simonwillison.net behind, yet positioned to completely dominate this race by hitting it with some sense. Making trends in what looks like longevity in the race that is not subsidising to simply get users, but to get by until they figure out how to 100x reduce the cost to a reasonable level. They feel like the guy sitting in the back with nothing big or flashy to say that is going to drop the hammer on their competition that overstretched itself taking on too much debt because it was necessary to change the game. There might be something to having a mix of hipsters, boomers, and luddites all trying to balance each other out.

An ai model created by Anthropic was announced as a closed preview on April 7, 2026 for critical security research and evaluation with its close partners with critical software such as operating systems and browsers. Anthropic claims that mythos is able to reason through so much more context that any model ever before. This enables it to find bugs that are 25 years old in the BSD, considered one of the most secure operating systems we have. Once it finds these zero day bugs never discovered before its able to use them together in malicious ways never expected. In ways the world is not ready for. At the time of writing these are claims without proof. It remains scary to know the potential this has and that there is only a few companies with this potential that will gatekeep who gets access.

5 star video, if you are going to watch one video to understand how harnesses and agents work, this is it. This really had my gears spinning on what tools do for agents and how big of a difference they make in their ability to manage context efficiently and accurately create changes. It’s crazy how good bash works, and that gives the agents the ability to do just about everything, but it could be better.

Agents Are Here

🌱 This post is still growing Late last year I started writing I'm Out On Agents. Agents sucked, the models were good, but there was still something missing between the harnesses and the models. They could write good code, they could do some debugging and exploring, but they were too good at fucking up the whole project to be useful. They could crank out Green Field POC’s like nobody’s business, but they created so much mess in brown field projects that it was easier to chat and edit yourself. The Inflection Point # It’s very well agreed on that the inflection point for most people happened with Anthropic Opus 4.5 in late Nov 2025. Early adopters probably noticed right away and shouted from the rooftops how good it was. But we’ve all heard that developers have 6 months before ai writes all the code for years, so this felt like the rest of the noise. Hitting the December slowdown many of us hit code freezes at work. We completely disconnect from work for the last Week and come back in Ja…
A really interesting long form interview with @simonwillison. If you follow him closely most of it is probably not new, but I found some interesting nuggets. Simon is writing most of his code from his phone these days using anthropic hosted platform. He mentioned that a lot of security risks go away when you don’t put secrets on the platform and you let them take the risk of running ai written code with ai chosen supply chain. He talked about the Pelican Riding a Bike benchmark for quite awhile. He was surprised at how well of a proxy it is for how capable a model is at just about everything. He also said that when he runs the benchmark he also runs half a dozen others that he’s never talked about so that He could see if they were to train a model specific to his benchmark he could catch them, but it seems they had caught on and if they were they seem that they would already be doing it on all of his others anyways. TDD is incredibly boring for humans, it strips so much creativity and joy from the process. Who cares if agents are bored they do better when doing TDD.
Laurie Voss (@seldo.com) Project Glasswing is a glimpse at an oncoming future in which agents do things humans could never have accomplished and the results are handled by other agents faster than humans could react and we... Bluesky Social · bsky.app Is Glasswing the next inflection point
What Happens When AI Stops Being Artificially Cheap The subsidy era is ending. Here danielmiessler.com I’ve been thinking about this for awhile and Daniel makes some great arguments here. Interestingly keeping inference cheap removes the incentives to make our tools better, help us choose the right model, lean on local models, open weight models. The frontier models are so affordable through subsidized subscription models why would you deal with anything less intelligent at this point. The tooling we use is not optimized for it, and why should it be.
ThePrimeagen (@ThePrimeagen) on X don't forget last time Anthropic, in their infinite PhD level wisdom, leaked their own source code (Feb 25) they DMCA'd all repos that had their code. Careful storing the code because Anthropic w… X (formerly Twitter) · x.com Everyone look away, nothing to see here.
Mete Polat (@metedata) on X @Fried_rice @Scobleizer Anthropic is now officially more open than OpenAI X (formerly Twitter) · x.com Anthropic safewords are the talk of the town today.
Josh Medeski (@joshmedeski) on X Did you know you can replace the spinning verbs in Claude Code. I'm having fun with it. X (formerly Twitter) · x.com The claude code source code leaked today and the tweets are great, maybe twitter is back. Did you know you can replace the spinning verbs in Claude Code. I’m having fun with it.
@nicknisi) — Y'all, I think I'm a convert to pi" loading="lazy"> Nick Nisi ( @nicknisi) Y'all, I think I'm a convert to pi Bluesky Social · bsky.app I’m about to be pi pilled.
To Live In A World Without AI | Nic Payne I'm finding lately that I wish we could go back to pre-ChatGPT... A world without a code-gen easy button, where "easy" was LSP autocomplete, wher pype.dev We f& #ing said @pype, well f& #ing said. I think a lot of us are feeling this, we’ve pitched our brain into a bucket and we are no longer stretching it in the same way. We still work in similar ways of old, with new ways of turning off and saying yes a bunch of times. the best thing I can hope for is that as things get better we have fewer yes loops, and more architectural design debates and deep thoughts. But I fear deep thoughts are gone to the way of “research the leading 10 frameworks and pick the best one for this project.” and letting the clankers do the deep thinking. Its signing us up for a weird distopia. I think a lot of us wish we could undo what has happened and go back to actually understanding what we are doing, but the world has changed, and if you are building average shit, like the average person, using models trained on average people doing average shit you cant keep up anymore.
My Thoughts on Beads | Nic Payne [Steve Yegge](https://en.wikipedia.org/wiki/Steve_Yegge) is a pretty well-known individual in the tech field, having been around for a long time at some of the pype.dev I’m in step with @pype here, I really want beads to work for me, but my systems for infra/platform work are all over the place, not one repo. I’m considering trying the env var but idk if it fits my workflow. For now, similar to @pype, I am rocking my own home vibed solution that I’ve intentionally put little effort in and its working great and I expect it to be broken and not working with the latest harnesses and models within a few months anyways, cause there is no predicting this train.
Vibe coding is going so far into the news sphere now that Adam Savage even weighs in with perspectives from someone who has built a life around building things with his hands, keeping up with new making techniques, discovering old techniques as they combine with new. He talks about 3d printing reviving his love of the pantograph as one automation technique eases the most difficult part of another.
Notes – 06:34 Mon 23 Mar 2026 Notes – 06:34 Mon 23 Mar 2026 dbushell.com · dbushell.com Does anyone think fast-code will continue to pay the same salary? The answer isn’t to switch your brain off during your McCode shift and write a poem after work. Your job will be replaced by a Banglasdeshi slop-shop if AI improves (which is inevitable, apparently). Possibly the same sweatshop that loomed my £3 T-shirt. The Luddites didn’t accept their fate so easily. David has some good points here, but I’m feeling the opposite direction a bit. Execs have always liked keeping the PM’s and the people steering the ship close by and were willing to farm out more and more grunt work. It feels like we are in a weird phase where there used to be a big group of people paid to write code. A few of them are exceptionally good at it and will remain. There will be a need for these people everywhere. Somehow we still need people hand editing assembly code optimizations, fortran, and cobol today. Those industries largely moved on, but a few great ones remain. I think this fast-code slop factory is going to be a short forgotten time in history, but no one yet knows what’s next. We are all waiting to find out.…

I don’t want someone else running my agents

I don't want to review the pr, I dont want to fight the mass of changes clobbered across the codebase. I want to own my platform. With everything changing with agents writing more code than I can imagine in a day work looks different now. I still want to work with real people. I want to collaborate on ideas. I want someone to bounce ideas off with. I want someone else in the war room with me on launch day, or when the whole thing goes down. But I don't them slopping in my sandbox, if someone is going to be stirring the slop in my product I want it to be me. Work is feeling different now. New lines need to be drawn in new directions. Expectations are changing, the way work is completed is changing, and we are all here trying to figure out what this looks like moving forward.
Very interesting takes from @thdxr in this interview. A lot has been hashed out by others all over the place, but a hot take here is that code quality is higher than ever right now. Codebases are becoming more consistent than ever. If you are not starting with a good consistent base from the start you are poising your context and doomed to fail and have all the common failures of ai written code. He still reads almost every PR, and will read all of the code eventually. There are a few cases where reading the PR is not worthwhile only when its low stakes, knows that good patterns have been established and followed. He argues that someone needs to be the expert of the code and of the product still and fears that too many people not looking at prs will fail companies.

Thinking about ai productivity again

Thinking about AI productivity again. It's allowing massive amounts of work to get done, to levels that humans cannot physically type out in some cases. But not all of this work is necessarily high value work. Right now I'm working on one of the biggest PRs to an internal cli library. Probably the largest PR I've ever done professionally. It touches all of the cli, refactors every command, reaches into the business logic layers to drive deeper separation. I reaches into the common layers to drive consistency. It ensures that every command (50 or so) has similar flags, supports --plain, --no-color. It specs out contracts to ensure that data goes out stdout, any extra goes out stderr. This makes everything unix pipe friendly. There was quite a bit of research and prep that went in, that turns out to already be distilled down into clig.dev. The point is that this is all good work. It will make the product consistent, repeatable, expected, and most of all boring. Most of the time, it wi...
Kids are leaving the party early, not drinking, cant watch netflix without the laptop open. They are leaving the party early to check on their agents. I get it, that feeling that you need to eek out one more prompt, keep your agents running. if they arent running what are you even doing. If not you 6 others are ready to pass you up. The timeline to be first has shrunk to nothing but unachievable.

The Ai Wars Are So Much Worse Than The Framework Wars

I’ve been thinking about this for awhile, the AI wars are so much worse and burnout prone than the framework wars of the 2010’s. I remember really starting my professional programming journey during the framework wars. It was a time when there were new and exciting js things every single month. Frameworks and meta frameworks came and went, the ones that lasted changed best practices yearly or so, often flip flopping on technique. I was deep in python and data engineering at the time and only experienced it adjacently. I was into webdev. I did a bit of react, gastby, vue, gave all the big ones a try in a demo level.
1 min read