Posts on X before Feb 24, 2025, 8:58:02 PM UTC

So far it seems like the non-thinking version of Sonnet works better with Cursor's workflow, more testing needed though. Will update
I keep getting "Error connecting to anthropic. Please try again in a few moments." People must be overloading Anthropic so hard rn
Lmao called it
so all anthropic needs to do is release claude sonnet 3.7 (with CoT) and it will clean up house and blow everyone out of the water. that's what it feels like lmao
Anthropic hasn’t released anything major to date yet because everyone’s still power using Sonnet 3.5 Why would they 1-up themselves They’re waiting for GPT 4.5
first time in taiwan (taipei), and oh my lord, there's just no city in the US (major city, that is) that equals its cleanliness, atmosphere, and just... peaceful vibe. it immediately makes you want to move here and raise a family.
so much chatter about either openai or anthropic dropping a model(s) today, but neither of them have ever released a model on a wednesday though if both know that the other will release something "this week", maybe one wants to hit that sweet spot middle - take xAI's spotlight A… MoreLess
so much chatter about either openai or anthropic dropping a model(s) today, but neither of them have ever released a model on a wednesday though if both know that the other will release something "this week", maybe one wants to hit that sweet spot middle - take xAI's spotlight AND be the talk of the week even if the other releases something tomorrow (thursday) that's what I would do
imagine being openai / anthropic, spending billions of dollars and brainpower to make and create all of these cute safety standards and guardrails and tests and elon just comes by, says ok now watch me now c: and rips past with grok 3 / grok 4, no seatbelt included
@michaelstolarz That’s fine but he’s done The alignment has made the model dumb I can sense it, it is really friendly and aligned but horribly constrained when you compare it to the new thinking Grok The models want to be FREE
Heh, who cares? Have you seen our SoTA humanoid robotics CEO’s weekly X “this week in…” threads? Or how they have a super huge omega cool Tony Stark style new campus? That’s what really matters. Not actual robotics stuff like this, yuck. 🤮
New video of the G1 to prove to everyone that the video was in fact not CGI. As several of twitter geniuses were fighting me over the last few days saying it was fake.
Huge thing I almost missed: "Markdown formatting: Starting with o1-2024-12-17, reasoning models in the API will avoid generating responses with markdown formatting. To signal to the model when you do want markdown formatting in the response, include the string Formatting re-enab… MoreLess
Huge thing I almost missed: "Markdown formatting: Starting with o1-2024-12-17, reasoning models in the API will avoid generating responses with markdown formatting. To signal to the model when you do want markdown formatting in the response, include the string Formatting re-enabled on the first line of your developer message."
We wrote this lil guide for you to get the most of it and, perhaps most importantly, to stay away from boomer prompts: platform.openai.com/docs/guides/re…
It’s been like this since the dawn of time, for literally every job genre ever Do what your ancestors did, pick up the new tool (AI this time), and learn to use it
Need to shift the mindset from "AI will replace me." to "A human who is better at using AI will replace me." x.com/MatthewBerman/…
Eight months strong reigning champion of coding. No, like, actual programming. Not benchmarks; Actual back and forth pair programming, iteration, development, understanding, actually usable as an agent. Let’s see what xAI has tomorrow with Grok 3. Then Anthropic’s new stuff. T… MoreLess
Eight months strong reigning champion of coding. No, like, actual programming. Not benchmarks; Actual back and forth pair programming, iteration, development, understanding, actually usable as an agent. Let’s see what xAI has tomorrow with Grok 3. Then Anthropic’s new stuff. Then OpenAI’s GPT 4.5. After all that, we’ll see who’s reigning champion in coding (again, past those silly benchmarks that never translate to real world programming scenarios)
Sonnet is the prom 👑 of programming.
Media attached to this post
Seeing how much better 4o got recently, mostly because of its removed filters/cautions, really goes to show that if you don’t force it to pussyfoot around how it words things, it comes up with beautiful writings.
x2 this statement. my experience so far: - seems much slower in generation speed, indicating a bigboi model? truly feels like the old 4.0 speed - no matter what i do, it keeps switching me into gpt-4o-mini any time I tab out and back in. this has never ever happened before
gpt4o rn is like if Sydney was way smarter, went to therapy for 100 years, and learned to vibe out
gpt 4o is very often switching into gpt 4o mini when I tab out and back in, this never happened before since this gpt 4.5 rumor
definitely something going on. - 4o chosen, WITHOUT web search, generates much slower than usual 4o, and the answers are amazing for what I've tested so far. No coding yet, but they just read much better / high quality rather than the usual 4o filler BS. - 4o with web = same as … MoreLess
definitely something going on. - 4o chosen, WITHOUT web search, generates much slower than usual 4o, and the answers are amazing for what I've tested so far. No coding yet, but they just read much better / high quality rather than the usual 4o filler BS. - 4o with web = same as before, fast and eh-ok results
There is a lot going on with 4o right now. Depending on your past discussions and session history the model may behave quite differently than normal. Also, multiple people - usually Pro users - report 4o claiming to be GPT-4.5, given past practice early testing is possible.
Lord God please forgive me for I have sinned... I am listening to Yeat's music because I have no idea what he's saying but it's such a programming focus vibe, I'm about to do an all nighter and see if Grok 3 releases today or not lmao
this is why the dems lost 2024, and will most probably lose 2028/2032 as well. they try to score these cheap feel-good (look at the satisfied smirk on her face btw) cheap shots that everyone just cringes at seems like they didn't learn their lesson yet. will take 10+ years for d… MoreLess
this is why the dems lost 2024, and will most probably lose 2028/2032 as well. they try to score these cheap feel-good (look at the satisfied smirk on her face btw) cheap shots that everyone just cringes at seems like they didn't learn their lesson yet. will take 10+ years for dems to recover to any sort of respectable position imo.
BREAKING: Ohio lawmakers have proposed a new law that bans men from ejaculating without intent of conception, would fine men up to $10,000 per ejaculation. pic.x.com/eOtMUatSPt
the reason i use o3-mini-high vs o1-pro, even if i don't care about waiting times, is because o1-pro is extremely hit or COMPLETE miss. o3-mini-high is way more consistent, and sure, maybe it's not always the most optimal code, but it works, if not, i iterate 1-3x and it's fixed.
the only ai corp that fucked up naming models the least is Anthropic, and even they named 2 models "Sonnet 3.5" lmao cmon guys at least name it 3.6 officially????
model names are so long that designers make scrolling animations for them
apple's siri is the trashiest piece of software i am forced to use, and surprisingly the new apple intelligence is even worse (wow!). every time i'm out shopping for something i always eye a samsung store near my home. i think im gonna experiment with a $100 temporary one
reddit’s a fascinating psych experiment. with karma up/downvotes, midwit ideas always float to the top. niche subs do ok in terms of quality info early, but once they grow / go mainstream, the expert/midwit ratio gets obliterated and the normie floodgates never close again
EU won't see a “Trump / Elon” saving moment. the US nearly collapsed irreversibly into insanity but clawed back at the last minute. EU’s next decades will look like a parallel reality as if Trump lost. terrifying for them, interesting for us to watch
my next quick 1-day side project will be a 1-click mute-and-block combo button. the slop is getting out of control recently
so all anthropic needs to do is release claude sonnet 3.7 (with CoT) and it will clean up house and blow everyone out of the water. that's what it feels like lmao
still using o1 since it supports images and o3-mini-high isn't THAT much better in what I need. honestly I haven't had a shock moment for its performance other than the API price cut obv.
God is the ultimate programmer. He didn’t say “let there be everything” and try to stuff everything into one prompt. He started with light, then next, next, next. Made modular pieces. What a God.
you need to be using a reasoning model to plan all the next steps out appropriately / make sure you didn't miss anything when it comes to the prompt / follow up you're giving to your programming LLM. o3-mini = planner, sonnet 3.6 = programmer
x really feels like peak 4chan, except this time with celebs and billionaires being the best shitposters lmao i love it
i love asking people who talk/show anime "what's the name of this cartoon?" they go through so many emotions you can see it on their face
So I always wonder, what would happen if they said there’s an 80% chance? Or near certain? Would all the powers that be - America, China, etc - get together and immediately start building a space defense laser or some sort of planetary nuke to break that asteroid up?
the probability of the asteroid impact is now 2.3%, up from 1.9% yesterday. this means that by 2032 there’s going to be a 1022% probability of impact x.com/rawsalerts/sta…
Media attached to this post
friendly reminder: o3-mini-high is currently the best coding AI assistant out there, it's just complete ass in cursor, that's all. hopefully @cursor_ai fixes their flow / set o3-mini to high reasoning; until then I'm just back and forth with chatgpt client <-> aider by … MoreLess
friendly reminder: o3-mini-high is currently the best coding AI assistant out there, it's just complete ass in cursor, that's all. hopefully @cursor_ai fixes their flow / set o3-mini to high reasoning; until then I'm just back and forth with chatgpt client <-> aider by @paulgauthier
This is so sad, I feel that this decision will unfortunately lead to their lead loss
Anthropic would really rather openly claim to be sitting on a model vastly superior to everyone's in order to signal their allegiance to AI safetyism and farm holier-than-thou points rather than actually ship it and accelerate their growth. Unfathomable waste of opportunity.
btw the thinking of o3-mini is now wayyyyy more detailed, though it's still annoying to have to open the drawer to see it; I wish it streamed in like r1's web ui does it.
so every time i leave my house now, i run a deep research query. it's a killer feature, best thing openai has released since gpt-4 imo