Same here. And I was seriously impressed with how much noise cancelling the v4 can do. I will be seriously impressed again if the v5 is as much better as they claim.
I mean, there's certainly silly money involved, but there's surely no credible doubt that AI is already really good at many valuable tasks.
So if there's a crash, doesn't that just leave us with a bunch of useful AI hardware whose costs have been written off? I wouldn't have thought that much changes what social impacts AI is going to have.
Yes — configuring strongSwan as a bog-standard VPN server was so hard to fathom I made GitHub repo for it [1]. To be fair, some of the complexity comes from OS support that seems specifically designed to make secure setups difficult, presumably at the behest of various Three Letter Agencies.
I have now mostly switched to Wireguard for this, which is much more sane [2].
I just downloaded and paid for Codex this week because I want to stay on top of the AI tools and understand their capabilities. I've had some good results using 5.6 Sol, although it tends to never want to write any comments (despite modifying the project rules to tell it it MUST), and also occasionally just does a bit of thinking and stops. It'll say "Working through the remaining work" and just ends the chat until I tell it "continue" or something.
This is very anecdotal evidence but I have to rant about it... I tried 5.6 Terra (high) earlier today to fix a bug with a slow page. It just... removed the part of the page that was slow, and made it a client side request (still slow, but not blocking SSR I guess). I tried with Sonnet 5 and it correctly found the issue where an unhandled case was continuing to retry and failing. I am always telling people how the frontier models are SO much more capable than what they may think but this one thing today had me scratching my head at why it would ever do that. It was the first time I experienced the "great, I removed the failing test case" kind of issue.
I think the comment you're replying to is talking about cost and speed, and Opus 4.8 xhigh is pretty expensive and pretty slow. I work around the slowness by having multiple auto mode sessions going in parallel (review this, investigate that, plan this, etc.; and yes, I use a VM that can't touch prod for auto-mode work), and work around the cost by my employer pays for it and everyone else I work with uses more AI than I do :)
I really really like Fable for software engineering cases; add an alert and write a runbook, go pull metrics for this incident, write me a TUI that does blank, etc. It is astoundingly good and I am very picky. It costs A LOT of money though. I am not sure how I get away with my Fable usage.
My only issue with Codex (Sol 5.6) is that it sometimes finds tasks on its own to solve that are definitely out of scope. like it’s always aiming for the 1.0 release when we’re still on 0.1.0a
reply