Posts on X before Dec 11, 2024, 6:59:20 PM UTC
This is already solved in the labs, they’re just packing it up for consumers rn.
We’ll see the first usable ones late Q1/Q2 which will still be clunky and buggy; then late 2025 it should be as “game changing” as the first useful LLMs felt
@GregKamradt
LLMs don't have long term memory - they're stateless Who's doing memory as a service? What's the best practice with incremental data and adding it to a long term memory?
Advanced Voice feels like I’m just talking with GPT 3.5 Turbo. Nothing of value is gained talking to it, it’s just a feel good loop that FEELS useful but it’s actually just a waste of time
@SmokeAwayyy
ChatGPT Advanced Voice Impressions After ~10 hours of 4o voice chats using ChatGPT Pro here are some initial thoughts: - Pretty good. Not great. - Responds too quickly. Turns out responding instantly isn't actually desirable. It fails to detect contextual pauses mid sentence/th… Less
ChatGPT Advanced Voice Impressions After ~10 hours of 4o voice chats using ChatGPT Pro here are some initial thoughts: - Pretty good. Not great. - Responds too quickly. Turns out responding instantly isn't actually desirable. It fails to detect contextual pauses mid sentence/thought so you need to have your entire thought formed before speaking. I would rather speak with an o1 type model that at least takes a second or two to reason. - Needs an optional toggle to speak up during long pauses. Currently it just sits there forever if you don't say anything. - Emotionless. I know "AI is a tool" and all that, but 4o feels completely emotionless and too rigid. - Way too strict. OpenAI models are now a minefield of compliance. 4o Voice is the most strict (won't output). 4o Text is less strict (warning label but outputs). o1 Text is now the least strict model because it is able to think around the OpenAI policies and self-moderate at a human level (outputs with less warning labels). - Incredible at speaking in different languages, accents, and styles. - Can't hold unique voices consistently. It constantly needs to be reminded to maintain a new voice/language. - Responses are all very short. Hard to get it to say more than a paragraph in one response. - Speaks much slower/differently at the end of long chats. - Chats are limited to 1 hour each so 24/7 voice mode is currently not possible even with ChatGPT Pro. Also have a 4 limit placed on advanced voice during a normal mid day chat so I'm guessing it's not actually unlimited. The ideal Voice Mode in my mind would look something like unlimited o1 with reasoning, custom instructions, "infinite" memory, search, and vision.
Insane how we still don’t have an online Search LLM API using the SotA models. @perplexity_ai @OpenAI please 🙏
when I'm on my Following tab, I feel my brain expanding; when I'm on the For You tab algo, I feel like I'm growing more and more irreversibly retarded
Holy fuck… I don’t think people understand what this means if it’s true. I’ve used this experimental model a lot and I thought it’s just another Gemini pro 1.5 upgrade they’re pushing out before 2.0. But nope, seems like it’s possibly FLASH 2.0. Basically free imo. … Less
Holy fuck… I don’t think people understand what this means if it’s true. I’ve used this experimental model a lot and I thought it’s just another Gemini pro 1.5 upgrade they’re pushing out before 2.0.
But nope, seems like it’s possibly FLASH 2.0. Basically free imo. x.com/legit_rumors/s…
for the past 2 days i've been like "damn the For You algo has gotten really, really good" but today I realized I've just been in the Following tab, lmao
surely Anthropic releases a new model / another sonnet update during these 12 days of OpenAI Christmas?
o1 pro just isn't that useful unless you're doing college level mathematics / science based reasoning.
it's not better for code, and the tradeoff in speed (~3-5 seconds with o1 vs 2-3 minutes with o1-pro) isn't worth it, especially if you need to do iterations/refactors.
my wet dream is chatgpt search with o1 - AS AN API.
holy fuck I'm trembling just thinking about it.
@SmokeAwayyy
o1 is a GPT-4 wrapper with October 2023 knowledge cutoff. I want a GPT-5 wrapper with no knowledge cutoff, infinite memory, and continual learning. Jk o1 is great. x.com/SmokeAwayyy/st…
in less than 12 hours we're going to see what @sama is so excited about
o1 pro shouldn't be your programmer, it should be your architect
yeah I mean $200 per month for pro is worth it even if the only benefit was unlimited o1.
even if there wasn't o1 pro, or unlimited voice mode - just having unlimited o1 is a game changer imo
one thing i've noticed so far, using nothing but o1-pro for an entire day, and then now using o1 for an entire day
is that iteration changes and your workflow in general is much faster with o1
I think I'm going to use o1 pro only for very complicated planning/decision trees
it's so interesting to see that sometimes when i try to do something in o1-pro, it just can't get it right but then i pop it into o1, and it's perfect on the first try i wonder what that means... is o1-pro just overthinking everything? gets stuck on something? maybe its contex… Less
it's so interesting to see that sometimes when i try to do something in o1-pro, it just can't get it right
but then i pop it into o1, and it's perfect on the first try
i wonder what that means... is o1-pro just overthinking everything? gets stuck on something?
maybe its context window with all of its behind-the-scenes think-through stuff gets all too big, and it just forgets what it's supposed to do?
very curious to keep testing this out 🤔
i have very mixed feelings regarding O1 Pro. on one hand, when it gets it right, it gets it right beyond what you expected. as in, it does things that you didn't even ask for, which is great for certain situations, but terrible and time-wasting for others, because then you have … Less
i have very mixed feelings regarding O1 Pro.
on one hand, when it gets it right, it gets it right beyond what you expected. as in, it does things that you didn't even ask for, which is great for certain situations, but terrible and time-wasting for others, because then you have to go back and modify the prompt and specify, hey, don't go ahead and do this, or don't change this.
so suffice to say, it's hit or miss, except when you hit, it's like you hit the jackpot. and when you miss, it's just a mild annoyance.
is it worth $200 a month? in my opinion, a quick good gauge for yes/no is if you are able to afford food delivery every day, if you're at that level of disposable income, then yes, you must find a way to rebudget a few things and get this.
of course, this should go without saying, but if you do end up buying the subscription, you should be using it as many times as you can every single day.
i'm slowly getting to the point where instead of browsing twitter, i'm browsing O1 Pro. or more specifically, i prompt O1 Pro and then open up my other tab into just typical O1, which responds much faster now. and then i chat there while i wait for O1 Pro, and then back and forth and back and forth, you get the idea.
Alright, there goes $200 buckaroos, I'll be using nothing but o1 Pro through the entire next 12ish hours, lets see when/if I hit a rate limit/notification 🫡
I'd be down for a $1,000 / month unlimited API access to o1 pro - (obviously agreeing to their caveat of not using this particular inf. API access for commercial purposes, only self-use/development, duh)
One thing I’ve come to realize (fortunately and unfortunately) is that we’re at a point where we can build anything software wise. And I mean that quite literally. If you know how to use agents, self improvement, and just replace your main driver model with whatever’s the latest … Less
One thing I’ve come to realize (fortunately and unfortunately) is that we’re at a point where we can build anything software wise. And I mean that quite literally. If you know how to use agents, self improvement, and just replace your main driver model with whatever’s the latest SOTA… then yeah, there’s nothing you can’t do
It’s still difficult, but for those that weren’t experts in certain domains, it would take years of study to apply things in a new domain. Now that same level of quality work can be done in a few months
as we’re nearing the edge of intelligence being available to the middest of wits, I propose a new giga-autist hyperfocus:
Architecture. More specifically, planning out pipelines and self-improving automations.
…even more specifically: system sculpting. 😏
Interesting to see that no one has reached limits yet with o1 pro…
super interesting. Might just have to try it out
The real white pill is realizing Safari is the best browser if you’re deep into the Apple ecosystem already
Few.
Hahahaha nope, THAT mode is for gooberment / internal use only
@MaxRovensky
@sama Does the $200 mode remove the thing where it lectures me on morality instead of answering a fucking question?
100%, anyone complaining about AI killing the software job market (it is) wasn't skilled enough to begin with - programming is not just about knowing how to write correct syntax or terminology, it's about building solutions given a set of tools and now one of the most revolution… Less
100%, anyone complaining about AI killing the software job market (it is) wasn't skilled enough to begin with - programming is not just about knowing how to write correct syntax or terminology, it's about building solutions given a set of tools
and now one of the most revolutionary tools just happen to be AI - learn to use it accordingly and the opposite will happen (you will become so GOATed that everyone will want to hire you)
personally i don't care much about chatgpt itself, i care about what i'll have access to regarding the API
fingers crossed that these new "o1-pro" and "gpt-4.5" will be available through API usage - idc at what price, I just want to test them in my own environment/sys prompts
it really feels like major softwares of our time with 100+ team members are actually just churning out unoptimized "uniquely designed" slop just so that they ship something
reminds me of game studios
@chamath
Is it just me or was 18.1.1 a total turd upgrade. Phone doesn’t work anymore.
this +:
- AR (ex. apple vision)
- speech to action
- 3d scene realtime rendering
- suno-level audio generation
-character consistency/“models”
then all you need is a dedicated empty room to build in; walk around, design a movie, make a game, i mean you’re literally tony stark
@krea_ai
announcing Krea Editor. our new editing tool feels like magic. who wants beta access? 👇
you can very clearly see when my baby was born, lmao
i've been seeing a lot of rumblings from openai employees (i analyze their post frequency via their notis being turned on for me)
it's matching the same frequencies as whenever there's a release coming, so i'm pretty sure something releases this week, maybe even today (tuesday)
my daughter will read my tweets ~15 years from now and wince from the cringe. im so excited
i mean what is even the purpose of spamming 99% offline (1% rate limited to oblivion) gemini experimental models @openrouter @alexatallah @OfficialLoganK
TIL perplexity API is pretty useless because there's no Perplexity Pro equivalent (ex. reasoning, ability to use Sonnet 3.6 as the search model, etc)
Just some outdated llama 3.1 perplexity online models. booooo 😭 @perplexity_ai @AravSrinivas pls implement, I'll pay w/e
bro, just one more all-nighter, bro. the api rewrite will be so worth it, bro. cleaner codebase, bro. just one more, bro.
i swear, bro
it's such a shame that o1-preview is retarded about anything "recent", and by recent I mean the last year of progress in AI/APIs/schemas/endpoint capabilities/documentation...
and then to add fuel to the fire, it doesn't support search lmao
once i hit $1M i'll start using my real name, i feel like $1M net worth (made from scratch from software) is a minimum to take an actual person seriously these days
surprisingly, the only 99.9% accurate speech-to-text, of course, after i do a post-transcription on it, is still OpenAI's Whisper API. all of these other providers that offer Turbo, like Groq, or any of the versions that are hostable on your own, if it's not that good ol' Whispe… Less
surprisingly, the only 99.9% accurate speech-to-text, of course, after i do a post-transcription on it, is still OpenAI's Whisper API.
all of these other providers that offer Turbo, like Groq, or any of the versions that are hostable on your own, if it's not that good ol' Whisper, it's just not as accurate.
what i find that happens very often is, it'll just not transcribe a portion of your speech. i'm not sure why this happens. it'll just skip over things, or it'll do the notorious "..." instead of actually transcribing out the words.
it's almost as if it's too lazy to do the full transcription sometimes. really interesting behavior. hopefully it's fixed in V4, or whatever.
but for now, if i want to have a good focus flow, without having to go back and try and fix transcriptions and such, or even worse, having to re-say everything that i initially said, i'm just going to stick with OpenAI's Whisper API for now. and then of course, Sonnet 3.6 for post-transcription touch-ups.
nice, hit 100 stars so far, just in time for the fully rewritten mobile friendly v1.0 release :)
when i first started using Cursor a little over a year ago, i really felt the "10x programmer" meme. especially since i was already so comfortable with the terminal and setting up automations in a pseudo-agent type of way with file watchers and procedures and loops... you get th… Less
when i first started using Cursor a little over a year ago, i really felt the "10x programmer" meme. especially since i was already so comfortable with the terminal and setting up automations in a pseudo-agent type of way with file watchers and procedures and loops...
you get the idea. now, with the latest Cursor update, where it has the agent feature, that initial 10x feels like another 10x. it's a steep learning curve, which isn't obvious from a first glance, even for those that use Cursor or code-gen tools every day.
i think 99% of programmers don't utilize Cursor efficiently. and then 99% of the 1% that do utilize Cursor don't really think through how this new agent feature is as big of a game changer as code generation AI is.
once people catch on and realize how powerful this is, i don't see any software limitations in terms of what can be created, even if tech progress magically stopped and we never have a better model than Sonnet 3.6 (lmao but you get the point)
what an amazing time to be alive, waiting for the next state-of-the-art LLM to further this productivity explosion, and more importantly imo, even better workflows that no one has even thought of yet (but AI will for you, especially when personalization becomes an integral part, not even mentioning continuous learning...)
"Almost every important piece of software from the last 10 years uses Go"
@wispem_wantex
Go is the programming language of the internet and the modern era. Almost every important piece of software from the last 10 years uses Go: - Docker - Kubernetes - Terraform - Prometheus - Grafana - Lets Encrypt - CockroachDB - etcd - Ethereum - Vault - Packer - Caddy
Remember threads?
wait until this guy finds out about @cursor_ai 😂 t.co/uvL6LH0TGi
are you even horsemaxxing, anon? come on, get with the times.
i may be wrong, but with the way that BS is built, there’s… no way for anyone to circumvent this? it’s literally part of their selling point regarding their protocol
you love to see it 😂
@vikhyatk
@latkins there’s a major controversy currently happening on the other app huggingface.co/datasets/alpin…
2024 isn't over yet, so i really hope i'm wrong with this next statement, but here it is: the only notable advancements that you can really feel in terms of output improvement have been Sonnet 3.5 and the o1/r1 "reasoning" models. everything else is cool, but tbh it just kind o… Less
2024 isn't over yet, so i really hope i'm wrong with this next statement, but here it is:
the only notable advancements that you can really feel in terms of output improvement have been Sonnet 3.5 and the o1/r1 "reasoning" models.
everything else is cool, but tbh it just kind of seems slopped together
TLDR: pretty much every frontier AI lab has perfected AI text generation slop and so now we can finally focus on what really matters
@_jasonwei
Prediction: within the next year there will be a pretty sharp transition of focus in AI from general user adoption to the ability to accelerate science and engineering. For the past two years it has been about user base and general adoption across the public. This is very natura… Less
Prediction: within the next year there will be a pretty sharp transition of focus in AI from general user adoption to the ability to accelerate science and engineering. For the past two years it has been about user base and general adoption across the public. This is very natural because user growth is a critical part any business model. But at this point I would say that there is widespread accessibility of LLMs: for most queries from the average person on earth, many LLMs can answer pretty well. In the upcoming five years I think the focal point will be the ability for AI to accelerate engineering and scientific research, which is the engine of progress in technology. At the frontier of innovation in any field, by definition there will be many open questions and a lot of headroom for better AI to make a difference. The stakes will be very high because progress compounds and also because AI accelerating AI research itself is a strong positive feedback loop. The other way of saying this is that there is somewhat limited headroom for improving the average user query, but massive headroom for improving the experience for the 1% of queries that would accelerate technological advancement, as well as on queries that people would want to ask the model but currently don’t because models are not smart enough to answer. AI research tends to improve where there is great headroom, and in scientific innovation there be substantial upside.
They're not wrong.
It's just not the "open" that we're all used to.
Open here means "open to vulnerabilities," not open source. Definitely not open source.
@tomwarren
quite the claim from Microsoft here 🤔
If you’re ever feeling down and want a stark reminder as to how far ahead you are of everyone else when it comes to AI-anything, just take a look at the replies in this thread
@_lizharvey
Maybe im stupid (im not) but there is literally not a single thing id need chatgpt for??????????? like does it do the dishes? x.com/ilikefufu_/sta…
It's so funny to see the annual cycle of people finding out about some new Twitter/X replacement... ...followed by them flocking over there, convincing their friends and family to use it... ...they come back to Twitter and post about it, saying "Oh my gosh, the discourse there … Less
It's so funny to see the annual cycle of people finding out about some new Twitter/X replacement...
...followed by them flocking over there, convincing their friends and family to use it...
...they come back to Twitter and post about it, saying "Oh my gosh, the discourse there is so amazing!"...
...and then this lasts about two-three weeks, after which they silently come back to being full-time on Twitter once again.
It takes them a while to realize how much they take for granted, and pair that with the fact that the honeymoon phase of using a new app / getting so many likes / comments / engagement with the thought of "wow, I just started, too, and I'm getting so much engagement already!"...
All of it fades away and yeah, X will always take you back in. Welcome back.
It never ceases to amaze me how people seem to avoid the reality of impending changes, whatever they may be. It seems like, even against all odds, they still hold on to the belief that everything will remain the same. In pretty much every field I can think of, whether it's po… Less
It never ceases to amaze me how people seem to avoid the reality of impending changes, whatever they may be.
It seems like, even against all odds, they still hold on to the belief that everything will remain the same.
In pretty much every field I can think of, whether it's politics, societal structure, software, biology - I change my mind frequently given new information. That's the whole point of progress, right?
I consistently check to see if I have any static beliefs, and I make sure that I always have worthy reasons to believe what I believe. If not, I go down a rabbit hole of research, playing both the defense and prosecutor positions before I finalize my decision.
I've trained myself to enjoy changing my mind on something that feels almost hard-coded into me - it honestly feels like a software update that fixed some bug in my code, to put it crudely.
Most people, from what I can tell, are the opposite - they'll take their opinions and decisions, even if they know they're wrong, to the grave.
They rather go never saying "you were right" - or worse, "I was wrong".
Big as always from DeepSeek.
@deepseek_ai
🚀 DeepSeek-R1-Lite-Preview is now live: unleashing supercharged reasoning power! 🔍 o1-preview-level performance on AIME & MATH benchmarks. 💡 Transparent thought process in real-time. 🛠️ Open-source models & API coming soon! 🌐 Try it now at chat.deepseek.com #DeepSeek
Anthropic's Haiku 3.5 would be justified with a 4x price increase if it supported images. At the very least, they should have kept the price the same as it was, stating that they're working on image capabilities and then increase the price once that goes into effect. All they'v… Less
Anthropic's Haiku 3.5 would be justified with a 4x price increase if it supported images.
At the very least, they should have kept the price the same as it was, stating that they're working on image capabilities and then increase the price once that goes into effect.
All they've really done is shoot themselves in the foot by neither releasing image capabilities nor keeping the price the same, which has led to negative publicity and, for lack of a better term, a failed launch for their "mass use" model.







