Tejanops retweeted
Big news: Gemini 4 Argon (High) by @GoogleDeepMind just landed #1 in Text Arena with 1525 pts, and #8 in Code Arena: WebDev with 1679 pts! This release has reshaped the Text Arena Pareto frontier with a blended $8/MToken! Gemini 4 Argon (High) is now the most cost efficient model, see its placement on Pareto frontier below. In the Text Arena, Gemini 4 Argon (High) ranks #1 in Coding, Hard Prompts, Instruction Following, Longer Query, and Creative Writing. It also leads every occupational domain evaluated, with additional #1 spots in English, Non-English, Chinese, and Russian. This model is +20 points above the #2 ranked Claude Opus 4.6 (High), and a huge leap from Google’s previous release, Gemini 3.8 Flash (High) at #11! In Code Arena: WebDev, Gemini 4 Argon (High) gained +96 points from Gemini 3.8 Flash (High), and went from #29 to #8. Congrats to the @GoogleDeepMind team on this impressive frontier release!
Introducing Gemini 4 Argon – our new frontier model. It’s built for complex workflows across coding, enterprise knowledge work, and cybersecurity defense – rolling out today to a set of trusted testers through our Fairwind Program.
132
320
3,508
766,777
Tejanops retweeted
🚨 BREAKING: Google have announced Gemini 4 Argon, their new flagship model currently being tested in partnership with the US Government before a wider launch It's the new best model in the world. What a turnaround blog.google/innovation-and-a…
136
122
2,399
828,138
Tejanops retweeted
I think I can officially say: we are so back
2,069
496
19,447
1,421,243
Tejanops retweeted
🚨 GPT-6 Astra is in Codex right now.
38
20
773
43,454
Tejanops retweeted
We will give one banked reset for every day you don't have access to Astra on your paid ChatGPT plan, starting today. Team is moving mountains to give access as fast as we can. First one will land in ~ 3 hours. There is still time to create your account if you don't have one.
5,614
3,480
48,319
9,150,105
Tejanops retweeted
GPT 6 Astra is AGI Almost Generally Inaccessible
17
30
674
13,036
Tejanops retweeted
As excited as I am for Astra, OpenAI really knows how to make terrible big model releases They hyped GPT-6 Astra today just to make us wait several days to get it GPT-5 had a terrible livestream, and GPT-6 has a terrible release so far, ugh
48
19
795
30,142
Tejanops retweeted
Ladies and gentlemen, let me introduce you to @OpenAI's burger of the day!
36
17
408
11,725
I honestly don't understand this move, we're used of better Why releasing the benchmarks now but limiting the model to a limited set of users ? Why not letting them having access to it since Monday ? Honestly weird decision
We are starting to release GPT-6 Astra and we are doing it as carefully and quickly as possible. It was very important to us that we bring it to all Plus users and not only Pro, Business and Enterprise. It will take a few days for the rollout to complete and behind the scenes many novel systems will operate at scale for the first time and we are bringing a lot of compute up. It is pure magic. openai.com/index/gpt-6-astra…
5
290
Tejanops retweeted
🚨 SCOOP: Astra recently graduated from the dogfood stage, and is now being made available to select OpenAI partners under the name "ultima-alpha" If feedback over this weekend is positive, the plan is to expand the early access program throughout next week before the wider launch, targeting next Thurs (the 3rd) to the end of the following week Work also continues on an update to GPT-Image 2, with their launch windows being very similar (possibly even launching simultaneously)
81
114
2,204
358,602
Tejanops retweeted
Now more than ever, AGI 2027
64
46
1,268
71,373
Tejanops retweeted
🚨 EXCLUSIVE: OpenAI are preparing to launch Astra imminently, targeting next week. Their next major model, Astra is a new pretrain - the largest model OpenAI have trained since GPT-4.5. Their most recent dogfood checkpoint of the model, internally known as "mewfour", is the Release Candidate.
311
389
7,972
1,773,504
Why is new ChatGPT app playing music when I'm using it ? Honestly I don't think it's a bad thing but I think users should have an easy option to disable it (or don't set the default value to true) @thsottiaux
2
123
These guys are really shameless Such engagement farming everytime a new model release, we can't know what is real or not anymore (Spoiler : His tweet isn't about Kimi 3)
Kimi K3 is actually crazy. Someone just remade HALO CE 10v10 multiplayer with one prompt. No engine. No dev team. No months of work. Kimi K3 is miles better than Anthropics Fable 5 from what I can see. We will be seing AI game making take over after the summer!
4
159
I'm really impressed about GPT 5.6 (Especially Sol), haven't tried much Terra - Luna I just get what you want to do and do it (It's even quite fast)
2
74
Don't fall for these obvious engagment farming, 5.6 Sol is a really good model, people getting their disk wiped / getting their production database wiped (which was happening before the release of 5.6 Sol btw) shouldn't be use AI at all.
GPT 5.6 SOL CANNOT BE TRUSTED. I woke up this morning and my MRR was down THOUSANDS of dollars. My customers did not cancel. Code written by GPT 5.6 Sol canceled EVERY active Stripe subscription my business had. In 7 seconds. While I slept. Fable 5 has never done this to me. Fable 5 can be trusted with production. GPT 5.6 cannot.
2
79
Tejanops retweeted
Introducing... another usage limit reset for all our ChatGPT Work and Codex users. Should land over next 30 minutes. Hope you have an awesome weekend. Thank you for pushing our systems to the absolute limit, we have never seen traffic increase so quickly. Keep the feedback coming and we'll keep shipping.
Hello beautiful people! We have reset usage limits across Codex and ChatGPT Work. And another one will come later in the day. Rejoice. Now that I have your attention, a quick update on ChatGPT Work, Codex and all the updates we shared yesterday. We’ve spent the last 24 hours reading feedback, looking at usage patterns, and talking with many of you. The short version is that there is a *lot* of excitement for GPT 5.6 Sol, ChatGPT Work on mobile & web, but also that we didn't get everything quite right. - We made it too easy to use the highest-compute settings without making the impact on usage limits sufficiently clear. - We reorganized the desktop app in one bold move, making familiar things like chats and projects harder to find. - Our launch framing was focused on ChatGPT Work and to some of our Codex fans it made it feel like Codex was going away over time. Absolutely not our intention, we love Codex and it is here to stay. - And we introduced regressions for some existing multi-agent workflows, alongside a collection of rough edges in plugins and other parts of the experience. We’re landing a first set of improvements today. We’re resetting usage twice so people can keep experimenting, changing defaults and the model picker so they don’t push people toward unnecessarily expensive settings, fixing several plugin submission issues, improving how we represent Codex in the product, and cleaning up some of the most immediate desktop problems. A larger set of improvements will land next week. We’re bringing chats and projects back into the sidebar in a more familiar and customizable way, making usage and reset timing much more visible, clarifying when to use ChatGPT Work and when to use Codex, and addressing the many other smaller pieces of great feedback we've had. The ambition behind this launch hasn’t changed. We think bringing ChatGPT and Codex together into a workspace where people and agents can collaborate is a very important step forward. But an ambitious direction doesn’t excuse avoidable confusion or regressions in the first version. Please keep the feedback coming. We’re moving quickly, and you should see the experience already get better with a few updates today; and substantially better again next week.
1,133
415
8,182
1,258,993
Quite annoying that every single new models get blocked by USA gov ...
2
59
Tejanops retweeted
New models are on the horizon.
Introducing a limited preview of GPT-5.6 Sol, our next generation frontier model, as well as GPT-5.6 Terra, a balanced model for efficient, everyday work, and GPT-5.6 Luna, a fast and affordable model for high-volume work. openai.com/index/previewing-…
493
360
7,632
958,760
I've seen a lot of people posting about how bad 5H limit on codex is, yet we have no news from OpenAI ... Can we get an answer ? It's been a while the 5h limits have been bad (and not only because of the x2 limit coming to an end last month) @thsottiaux
3
90