Posts on X before Dec 13, 2024, 2:09:06 AM UTC

So I have a love and hate relationship with @GroqInc - long story short, I wish their developer tier finally came out so I can buy credits or whatever so that it's finally reliable rather than being randomly limited since there's only a free tier.
We are so back. Holy fuck 2025 is going to be insane. Google finally got their shit together it seems; this is their FLASH model, its multimodal, speech output, video intake, cheap AF. Google is posturing to be not only #1, but #1 BY A LARGE MARGIN.
Gemini 2.0 Flash is our strongest model to date, crazy progress for a small model like Flash : )
Media attached to this post
This is already solved in the labs, they’re just packing it up for consumers rn. We’ll see the first usable ones late Q1/Q2 which will still be clunky and buggy; then late 2025 it should be as “game changing” as the first useful LLMs felt
LLMs don't have long term memory - they're stateless Who's doing memory as a service? What's the best practice with incremental data and adding it to a long term memory?
Advanced Voice feels like I’m just talking with GPT 3.5 Turbo. Nothing of value is gained talking to it, it’s just a feel good loop that FEELS useful but it’s actually just a waste of time
ChatGPT Advanced Voice Impressions After ~10 hours of 4o voice chats using ChatGPT Pro here are some initial thoughts: - Pretty good. Not great. - Responds too quickly. Turns out responding instantly isn't actually desirable. It fails to detect contextual pauses mid sentence/th… MoreLess
ChatGPT Advanced Voice Impressions After ~10 hours of 4o voice chats using ChatGPT Pro here are some initial thoughts: - Pretty good. Not great. - Responds too quickly. Turns out responding instantly isn't actually desirable. It fails to detect contextual pauses mid sentence/thought so you need to have your entire thought formed before speaking. I would rather speak with an o1 type model that at least takes a second or two to reason. - Needs an optional toggle to speak up during long pauses. Currently it just sits there forever if you don't say anything. - Emotionless. I know "AI is a tool" and all that, but 4o feels completely emotionless and too rigid. - Way too strict. OpenAI models are now a minefield of compliance. 4o Voice is the most strict (won't output). 4o Text is less strict (warning label but outputs). o1 Text is now the least strict model because it is able to think around the OpenAI policies and self-moderate at a human level (outputs with less warning labels). - Incredible at speaking in different languages, accents, and styles. - Can't hold unique voices consistently. It constantly needs to be reminded to maintain a new voice/language. - Responses are all very short. Hard to get it to say more than a paragraph in one response. - Speaks much slower/differently at the end of long chats. - Chats are limited to 1 hour each so 24/7 voice mode is currently not possible even with ChatGPT Pro. Also have a 4 limit placed on advanced voice during a normal mid day chat so I'm guessing it's not actually unlimited. The ideal Voice Mode in my mind would look something like unlimited o1 with reasoning, custom instructions, "infinite" memory, search, and vision.
when I'm on my Following tab, I feel my brain expanding; when I'm on the For You tab algo, I feel like I'm growing more and more irreversibly retarded
Holy fuck… I don’t think people understand what this means if it’s true. I’ve used this experimental model a lot and I thought it’s just another Gemini pro 1.5 upgrade they’re pushing out before 2.0. But nope, seems like it’s possibly FLASH 2.0. Basically free imo. … MoreLess
Holy fuck… I don’t think people understand what this means if it’s true. I’ve used this experimental model a lot and I thought it’s just another Gemini pro 1.5 upgrade they’re pushing out before 2.0. But nope, seems like it’s possibly FLASH 2.0. Basically free imo. x.com/legit_rumors/s…
for the past 2 days i've been like "damn the For You algo has gotten really, really good" but today I realized I've just been in the Following tab, lmao
o1 pro just isn't that useful unless you're doing college level mathematics / science based reasoning. it's not better for code, and the tradeoff in speed (~3-5 seconds with o1 vs 2-3 minutes with o1-pro) isn't worth it, especially if you need to do iterations/refactors.
my wet dream is chatgpt search with o1 - AS AN API. holy fuck I'm trembling just thinking about it.
o1 is a GPT-4 wrapper with October 2023 knowledge cutoff. I want a GPT-5 wrapper with no knowledge cutoff, infinite memory, and continual learning. Jk o1 is great. x.com/SmokeAwayyy/st…