Posts on X before Aug 5, 2025, 4:52:59 PM UTC

with opus 4.1 there's now no good reason to use sonnet 4 (given we're talking about max20 claude code plan here) it used to be that sonnet 4 could potentially outperform especially with speed opus 4, but seems like opus 4.1 is across the board better now
Media attached to this post
keep your memories clean store everything as markdown/plaintext this is a short term victory but a victory nonetheless in 5ish years though? it wont matter you'll just swarm xyz application you used to use, it will consume every bit of data it possibly can and migrate you easily
Media attached to this post
you can now live in the cli btw and do everything from there i mean you always could, but now you can make fully custom fine tuned golang tview interfaces perfected for your exact daily workflow... in about a few hours worth of tinkering
Media attached to this post
while i agree, the problem is until openai catches up with... 1. a coding model surpassing unlimited Opus 4 2. a CLI 2x better than Claude Code (codex is TRASH) 3. aforementioned unlimited coding model included in their already-existing $200 plan ...we're stuck with Daddy Dario
some of anthropic rugpulls so far: > windsurf no access to claude 4 > plus not getting opus 4 in claude code > 1.58-bit quantized models during daytime > max plans limits cut in half 2 weeks ago, no comms > weekly limits without concrete numbers > 5x/20x plans being actuall… MoreLess
some of anthropic rugpulls so far: > windsurf no access to claude 4 > plus not getting opus 4 in claude code > 1.58-bit quantized models during daytime > max plans limits cut in half 2 weeks ago, no comms > weekly limits without concrete numbers > 5x/20x plans being actually 3x/8x of plus > cutting off openai api access what an absolutely horrible company
this is an amazing example of thinking that you're being super productive and making use of your max 200 claude code plan the reality is this is much, much slower than one instance per project regarding the straight to main and skip perms - that's totally fine, i do that too
Media attached to this post
unfortunately they are still dogwater at coding yeah they can use a few tools and one shot things pretty ok but they totally fall apart further with every follow up prompt you make to it
Horizon Beta - Xbox 360 controller. The models are definitely getting better
Media attached to this post
claude code tip @Maciek_roboblog made the amazing claude-monitor most accurate and informationally useful tool i've found so far to quickly understand where you're at in terms of usage
Media attached to this post
on one hand, funding would speed up progress at least 5x on the other hand, there is no "other" hand, both hands are pro-funding
you have to be so rich to be rich, even if you're rich, you're not rich; you have to be AT LEAST wealthy to be rich nowadays.
Its insane how rich you have to be to be rich.
feels like the open source "you can run it at home" model that openai has in store for us all first impressions, yeah it's a SMALL model, but damn it if i say it's actually pretty good like, able to vibe code on the plane with no wifi good and actually get some work done
Introducing Horizon Alpha, a new stealth model 🌅 Try it for free and help shape its development:
Media attached to this post
every time i think im maybe shitposting a bit too much i see that u can still see the exact number of tweets when u look at my account profile vs. REAL shitposters have their numbers in the tens or hundreds of Ks surely i'll hit minimum 10K before EOY
they're coming quick, hide the thinking sand we made lets pretend we're still at the nukes stage
BREAKING: New study suggests a potentially hostile interstellar alien spacecraft measuring nearly 7 miles long could be on a trajectory toward Earth.
weird stuff going on with claude code's session limit resets twice today my claude 4 opus has reset at 4 hours, not 5 im not complaining but idk feels weird and no it's not setting itself as sonnet behind the scenes, i thought so too but nope, still opus
did mark see something that explains why he's throwing billion dollar offers for single persons? i have a feeling he has no care in the world when it comes to "paper money" as people call it his eyes have seen behind the curtain he's seen the loading screen for the next stage
Today Mark shared Meta’s vision for the future of personal superintelligence for everyone. Read his full letter here: meta.com/superintellige…
claude code tip: dont run agents in parallel. i've tried in many cases, the lack of them sharing information while in parallel really nerfs their efficacy. instead, let your agents know they're able to run themselves/other agents at another level deep:
Media attached to this post
i was checking out, saw the cashier using chatgpt. "ah, chatgpt! everyone's using it now instead of google, huh?" her: "haha yeah, it really helps me plan my homework out so i get the hard stuff done first" me: "cool! do you use regular 4o with that or o3?" her: "what's that?"
"contributing compute resources to advance critical research" so this is where my unlimited claude 4 opus went
We’re joining the UK AI Security Institute's Alignment Project, contributing compute resources to advance critical research. As AI systems grow more capable, ensuring they behave predictably and in line with human values gets ever more vital. alignmentproject.aisi.gov.uk
it's funny seeing people use tier 3/4 models (think gpt-4o-mini) to do mathematical predictions they repeat it with such confidence; as a child boasting to new classmates on their first day of kindergarten "did u know my dad is the smartest in the world? look what he said!"
Media attached to this post
BREAKING: Massive asteroid on potential impact path with the Moon could trigger destructive meteor shower on Earth.
why run multiple claude code instances when you can just run multiple parallel subagents unless you're splitting your attention between multiple projects which in that case i wouldn't recommend doing that in the first place if you want high quality maintainable results
uh today i learned that claude code sub agents can call sub agents of their own i wonder how many layers that can go
Media attached to this post
this will never be reality again btw it'll be mentioned the same way we mention how people handcrafted written letters, so carefully, lovingly, including a lipstick kiss or a dried daisy future generations will say "aw cute, you used to just, like, talk to strangers before?"
At a truck stop restaurant. Everyone is sitting together, on their phone, silently. In the village. Everyone is sitting together, on their phone, silently. Smartphones are a SCOURGE!!!
the bright side of this heartbreaking anthropic announcement of rate limiting claude code: openai perfectly positioned to add unlimited codex cli to their 200 plan
it's so over it's me. i'm the "one user"
Some of the biggest Claude Code fans are running it continuously in the background, 24/7. These uses are remarkable and we want to enable them. But a few outlying cases are very costly to support. For example, one user consumed tens of thousands in model usage on a $200 plan.
i remember watching nearly 2 decades ago, him being so calmly confident and matter-of-fact, all while polls against his predictions (of the top 1% researchers/industry leaders of the space) were all >80% against his predictions he never batted an eye, never panic-explained.
The most amazing thing about Ray Kurzweil is how completely and sincerely unsurprised about all of this he is He's not like "to tell you the truth I didn't really believe this would happen either, wow" Instead he pulls up the same graph he's been using for 30 years Legend
looks and feels like a gpt-4 moment for video models
The introduction of Runway Aleph, our state-of-the-art in-context video model setting a new frontier for multi-task visual generation. Chat mode now available on mobile. And our weekly community spotlight featuring incredible Act-Two creations. Get caught up on what happened This… MoreLess
The introduction of Runway Aleph, our state-of-the-art in-context video model setting a new frontier for multi-task visual generation. Chat mode now available on mobile. And our weekly community spotlight featuring incredible Act-Two creations. Get caught up on what happened This Week with Runway.
remember those videos where people would inpaint robots over people doing things? fooling boomers & co. on facebook, etc? is there much sci fi left to fake anymore? is there anything i could watch as a faked video when i'm 60+ yo and get tricked thinking it's real?
Well that's scary... China's Unitree has a new robot, R1
if you think chatgpt psychosis has messed with people's minds so far wait until these new storytelling mind-rooting models hit mainstream use
Once you get to intermediate / advanced level in Claude Code you start to see how there's very little tutorials on tehse topics and the internet is littered with beginner tutorials all teaching pretty much the same thing.
Once you get to intermediate / advanced level in game development you start to see how there's very little tutorials on these topics and the internet is littered with beginner tutorials all teaching pretty much the same thing.
this is all you need to see/know about the new subagents mode of claude code this ran start to finish for roughly over 1 hour the result is perfect, chef's kiss, cleanest code i've ever seen now imagine, i have at least 5 of these running at all times wbu anon?
Media attached to this post
literally the only problem with claude code now is tok/s im forced to use sonnet if i want to get a feature done using my subagents flow but damn is it worth it because so far it has a 100% success rate of whatever i throw at it
o3 in chatgpt feels very different feels like intelligence is oozing from it with every word i read what did they do?
re. claude code subagents update; commands are basically useless now, much better results with "run my _ agent to _"
new memetic virus installed vibe coding with claude code is basically just note-taking but your notes actually do something for you and you can run them
it's just me, you and our note-takers
if you're self sufficient (self employed), not knowing this ISN'T bad but if your livelihood depends on being an employee, not knowing this IS BAD if i asked an llm, it not knowing this makes it a bad coder llm tldr; treat hiring employees the same as testing out an llm
the interview question that i'd pull out to fail the vast majority of people was a simple double buffering question
it will literally become mainstream logic of "if xyz is so smart, why aren't they going viral on tiktok about what their plan/values are?"
Zohran is the first politician to do basically this. Culture flows from the cities outward. Today you can become mayor-designate of NYC if you poast well enough on TikTok. In a few years, governor of Ohio or perhaps superintendent of Kearney, Nebraska Public Schools … MoreLess
Zohran is the first politician to do basically this. Culture flows from the cities outward. Today you can become mayor-designate of NYC if you poast well enough on TikTok. In a few years, governor of Ohio or perhaps superintendent of Kearney, Nebraska Public Schools x.com/a16z/status/19…
crazy to see people still using llms to make single file manual edits, giving the llm no other context, just raw dogging + guessing + duplicate code must be insane
july 2025 is about to close out, seems like it was dominated by china i think i'll start an end of month usa vs china thing, seeing who won the month in the world of AI (when it comes to my niche llm/coding usefulness scale only, not really keeping up with image/video gen)
🚀 We’re excited to introduce Qwen3-235B-A22B-Thinking-2507 — our most advanced reasoning model yet! Over the past 3 months, we’ve significantly scaled and enhanced the thinking capability of Qwen3, achieving: ✅ Improved performance in logical reasoning, math, science & coding ✅ … MoreLess
🚀 We’re excited to introduce Qwen3-235B-A22B-Thinking-2507 — our most advanced reasoning model yet! Over the past 3 months, we’ve significantly scaled and enhanced the thinking capability of Qwen3, achieving: ✅ Improved performance in logical reasoning, math, science & coding ✅ Better general skills: instruction following, tool use, alignment ✅ 256K native context for deep, long-form understanding 🧠 Built exclusively for thinking mode, with no need to enable it manually. The model now natively supports extended reasoning chains for maximum depth and accuracy. Hugging Face: huggingface.co/Qwen/Qwen3-235… or huggingface.co/Qwen/Qwen3-235… ModelScope: modelscope.cn/models/Qwen/Qw… or modelscope.cn/models/Qwen/Qw… API Doc: alibabacloud.com/help/en/model-…
Media attached to this post
sonnet 4 wants to show off how amazing of a coder it is opus 4 could care less about impressing you opus 4 is like a genius only interested in growing its knowledge, not actually utilizing it
the only solution to this problem is to bring back civility (or teach it to the uninitiated) through kind but corrective confrontation different cases call for different handling usually a smile + eye contact + head nod until they step aside for you to leave the elevator works
Same in elevators now Increasingly worldwide I can't get out of an elevator anymore without people getting in first Like how do you expect this is going to work? x.com/Fremond_/statu…