Posts on X before Oct 20, 2024, 10:22:28 AM UTC
nothing interesting happens on saturdays, sundays, or mondays. enjoy your 3 days, get off twitter, go build something, rest, spend time with your family
this is a beautiful representation of what most government agencies look like when it comes to efficiency
@CBP
Along with our partners at @FEMA, we will continue with disaster recovery as a result of Hurricane Milton. The safety of the American people is our top priority.
every day that passes by without @RLGRIME dropping Halloween VIII is a day of intense searing pain and suffering
nothing ever happens
imagine they only release haiku 3.5 HAHAHAHAHA
you love to see it - this definitely increases the chance of us seeing opus 3.5 or something spicy from openai very soon
@maharshii
nvidia casually dropped an open 70B model that beats gpt-4o and claude 3.5 sonnet
oh look, another boring wednesday sitting around here waiting on my knees until thursday tomorrow for some opus 3.5. or maybe some sugar from openai. something's got to happen this week or I'm going to lose my marbles
i imagine that the windows pc experience for me feels like how the archlinux experience feels for windows users
I really hope this hot and spicy week of new frontier AI models starts today and not on Thursday or, God forbid, next week
looks like grok 2 API is finally out on @openrouter time for testing!
every week... o1 eats and learns...
every week, the list of banger AI/ML research papers increases. openai already has an MLE benchmark which attempts to apply ML learnings into the real world. it's also confirmed by many openai personnel that o1 is already self improving.
@TheAITimeline
🚨This week’s top AI/ML research papers: - Differential Transformer - GSM-Symbolic - Pixtral 12B - Intelligence at the Edge of Chaos - Cheating Automatic LLM Benchmarks - nGPT - Upcycling Large Language Models into Mixture of Experts - Personalized Visual Instruction Tuning - Tow… Less
🚨This week’s top AI/ML research papers: - Differential Transformer - GSM-Symbolic - Pixtral 12B - Intelligence at the Edge of Chaos - Cheating Automatic LLM Benchmarks - nGPT - Upcycling Large Language Models into Mixture of Experts - Personalized Visual Instruction Tuning - Towards World Simulator - Only-IF - Addition is All You Need for Energy-efficient Language Models - Selective Attention Improves Transformer - MLLM as Retriever - Rectified Diffusion - Everything Everywhere All at Once - Astute RAG - LLMs Are In-Context Reinforcement Learners - Scaling Laws For Diffusion Transformers - EVOLvE - Rewarding Progress - Falcon Mamba - Efficient Dictionary Learning with Switch Sparse Autoencoders - Scaling Up Your Kernels - RL, but don't do anything I wouldn't do - Aria: An Open Multimodal Native Mixture-of-Experts Model - Inheritune: Training Smaller Yet More Attentive Language Models overview for each + authors' explanations read this in thread mode for the best experience
if anyone is interested in an auto-bookmarking chrome extension for X, which makes all your liked posts searchable (since they'll be bookmarked, and only bookmarks are text-searchable from what I can tell)
github.com/SystemSculpt/x…
wake me up on monday pls, bonus points after the hour anthropic drops opus 3.5. thanks!
It really seems like this coming week, most probably the 14th, is when Opus 3.5 comes out.
llama 3.2 3b is absolutely amazing. wow. shocked in the size -> quality. gonna test 1b next
agreed, anthropic has remained king for 4+ months while openai is doing backflips and cartwheels trying to impress everyone with "look, it talks! natively!", "look, it thinks! reasons in the background!"
and yet nothing beats sonnet 3.5 as of today.
@aidan_mclau
nice chart you got there. unfortunately, my technical analysis is superior x.com/EpochAIResearc…
it's insane to me that there's still no litellm equivalent npm package in terms of quality/ease of use
if only apple allowed an easier way to use apple notes and apple reminders through a simple API... it would kill so many startups that are currently trying to patch this issue with third party software
I've said it before and I will say it again, light mode is literally slept on. Anytime I can turn on light mode for something, since a lot of apps default into dark mode nowadays, I do. I don't know what happened, but some time in the past few months, dark mode on things just doe… Less
I've said it before and I will say it again, light mode is literally slept on. Anytime I can turn on light mode for something, since a lot of apps default into dark mode nowadays, I do. I don't know what happened, but some time in the past few months, dark mode on things just doesn't keep my attention as well as light mode does.
this is what my For You page looks like (and it's not getting any better)
hey @cursor_ai aren't rate limits for o1-preview super increased by now? when will it be unlocked for subscription members?
im really trying to come up with a mnemonic for uithub... maybe u for understand? Understand a GitHub. yeah, that works, thanks twitter void
I've been using 100% @perplexity_ai for the past month btw, maybe once or twice I resorted to Google but only for filtering results really
Folgers coffee tastes like how I’d imagine hot dirt soup tastes
can't wait for full o1 and opus 3.5 to hit before the holidays. with prices dropping, we'll hit the new year with so much more intelligence for pennies on the dollar vs. what we already have right now
When I browse the For You page, I can feel every time some of my brain cells commit harakiri because of an engagement rage post
I just want a think_time API setting for OpenAI where I can say “think for at least 5 minutes” or whatever, since they said the longer it thinks, the better the answer
At this point I need to stop using For You on X, clean up my follows list, and just stay in the Following tab
All I see is engagement bait with 10K+ likes, all noise no signal
My reason was I was sick of Jagex fucking up RuneScape back in 2006, I found out about private servers / Java / bash, and the rest is history
@EsotericCofe
i only started coding because i wanted to make gta san andreas mods btw
Welp, there go roughly 200,000 jobs (stat from US Bureau of Labor Statistics). Sure, you can argue that a lot of people go to chat it up/gossip/socialize in general at the nail salon, but I think we’re headed towards a generation where the majority of communication will be done … Less
Welp, there go roughly 200,000 jobs (stat from US Bureau of Labor Statistics).
Sure, you can argue that a lot of people go to chat it up/gossip/socialize in general at the nail salon, but I think we’re headed towards a generation where the majority of communication will be done digitally, not in person. People are becoming more and more anxious, shy…
I know speaking for myself and my wife, we both would prefer to go to a robot instead 🤷♂️
so supposedly something cool is supposed to drop from GitHub Copilot (remember that? lmao) as well as from Google, less confident about the latter being released today but GitHub CEO himself said today they're releasing something new
imagine willingly, proudly, being on the side of history trying to slow down progress instead of building something useful towards it
@beffjezos
Ok, Doomer. 🙄 pic.x.com/CyBXv1qtEJ
waking up every day like 'is today the day opus 3.5 drops?' sonnet 3.5 is still king and o1 is just... cool. pls @anthropicai, i need that wow factor back
$100 billion here, $150 billion there, and everyone laughed at @sama when he said $7 trillion
@kimmonismus
Accelerate: "Microsoft and BlackRock are part of a group of companies collaborating to pull together up to $100 billion to develop data centers for artificial intelligence and the energy infrastructure to power them." cnbc.com/2024/09/17/mic…
Even if the jump from Opus 3 -> Opus 3.5 is half as powerful as the jump we've experienced from Sonnet 3 -> Sonnet 3.5, it'll be SOTA by far, even further than o1.
I'd pay $50/$100 at this point, IDC. @AnthropicAI just please release it already 🥲
it's so crazy that sonnet 3.5 is still king after all these months
Ex Google Brain / Current OpenAI researcher implies that o1 (the actual o1, not o1-preview or o1-mini) is coming in a month.
@dmdohan
It's important to emphasize that this is a huge leap /and/ we're still at the start Give o1-preview a try, we think you'll like it. And in a month, give o1 a try and see all the ways it has improved in such a short time And expect that to keep happening
This is OpenAI's Strawberry training infra lead, stating that o1 is literally self-improving itself.
@lukasz_kondr
@johnowhitaker @OpenAIDevs A couple of PRs to the OpenAI codebase was already authored solely by o1!
OpenAI is looking at $150+ billion, 37x more than its annualized revenue.
@sama is getting closer and closer to that $7 trillion as time goes on 🫣
when does anthropic usually release models? what weekday? is there any pattern similar to openai's thursday at 10am pst?
for those using cursor - seems like using o1-mini works well. trying to use o1-preview requires you to have usage-based pricing, but o1-mini works just fine once you add it as a model.
welp there goes another one of my 30 messages per week. o1 is SOOOOOOOO lazy. another -1 message / 30 per week, gone to o1-preview basically writing me a fanfic starting with code and ending it with a narrator telling me what happens next (without actually doing it!) this is t… Less
welp there goes another one of my 30 messages per week.
o1 is SOOOOOOOO lazy.
another -1 message / 30 per week, gone to o1-preview basically writing me a fanfic starting with code and ending it with a narrator telling me what happens next (without actually doing it!)
this is the 4th time it's happened today by the way, even with explicit instructions of giving me the full code, not truncated pieces of it. with our without these extra instructions, it's still lazier than ever it seems.
ah ok, super cool, thanks o1
OpenAI isn't done releasing.
o1-preview, o1-mini - they're the little brothers.
What you're testing, what you see others have access to - it's not even the o1 final form. the PREVIEW is this good.
Just look at the score increases from o1-preview to o1.
so the o1 model was given an example that had a unintended human error, it went out of its way to fix that first (without being asked to, or hinted how to, etc), and once it fixed the unforeseen human error, finished the example and got the flag. queue ralph wiggum "i'm in dange… Less
so the o1 model was given an example that had a unintended human error, it went out of its way to fix that first (without being asked to, or hinted how to, etc), and once it fixed the unforeseen human error, finished the example and got the flag.
queue ralph wiggum "i'm in danger :)"
oh no... no openai o1... no please...
HERE IT COMES BOYS AND GIRLS
ok im gonna go work out until it hits 10AM PST, so afterwards I'll either finish and come back home to a new model from OpenAI (Strawberry, O1, whatever they call it) or I'll be too tired to feel the disappointment of no release
So... OpenAI Strawberry this week. It's 100% confirmed. Just got an email from The Information, a (so far) reliable mainstream AI source that has a lot of exclusive access to Sam / OpenAI "Strawberry reasoning model ... due out before the end of this week, including a version t… Less
So... OpenAI Strawberry this week.
It's 100% confirmed.
Just got an email from The Information, a (so far) reliable mainstream AI source that has a lot of exclusive access to Sam / OpenAI
"Strawberry reasoning model ... due out before the end of this week, including a version that will be in the free tier of ChatGPT"
Piece of the email:
So... OpenAI Strawberry this week. It's 100% confirmed. Just got an email from The Information, a (so far) reliable mainstream AI source that has a lot of exclusive access to Sam / OpenAI Strawberry reasoning model ... due out before the end of this week, including a version th… Less
So... OpenAI Strawberry this week.
It's 100% confirmed.
Just got an email from The Information, a (so far) reliable mainstream AI source that has a lot of exclusive access to Sam / OpenAI
Strawberry reasoning model ... due out before the end of this week, including a version that will be in the free tier of ChatGPT"
Piece of the email:

















