M3 Ultra being tested 🔥🔥@Apple @exolabs
2
299
🚨 Confirmed: it’s a banked reset, and it’s already showing as available on my account 🔥 Huge thanks to @thsottiaux for the reset 🙌 and @codex_resets for keeping us updated 👀
42
Jaime O retweeted
DeepSeek silently updated their changelog with a new V4-Flash upgrade 1 hour ago. Their new Terminal-Bench score is 82.7, a massive +25.8 point leap from its initial April preview score of 56.9. Currently only available via their API, open weights release will follow shortly.
205
448
5,862
1,893,379
Jaime O retweeted
Releasing the model weights and technical report of Kimi K3. Kimi K3 is our most capable model: a 2.8T MoE model with native visual understanding and a 1M-token context window. New model architecture: 2.5x the intelligence per unit of compute, not just more params. Alongside Kimi K3, we're opening up more of the stack behind it — high-performance attention kernels, MoE communication library, and infrastructure for running agent environments at scale. Model weights: huggingface.co/moonshotai/Ki… Tech report: github.com/MoonshotAI/Kimi-K… Tech blog: kimi.com/blog/kimi-k3
1,534
7,202
45,924
14,614,717
Open weights. Open research. Open innovation.🫶 Marching for an open future.🤍
Preparation in progress for the San Francisco Open Weights mini-march tomorrow. Hope all these big tech CEOs won’t show up otherwise I might end up in jail (haven’t had time to ask for a permit 😅)
42
85
1,230
231,934
Jaime O retweeted
Jensen Huang on "distillation" On his new interview with axios, he was asked this question "Should open source model companies be allowed to distill closed models" "Distillation—learning from AI, learning from other people, and learning from other sources of knowledge, is fundamental to intelligence. We are constantly learning from other people. I am learning from you through the questions you are asking, and you are learning from me. All day long, we are learning from one another. AI also has to learn from something. The original AI models, whether they were open or closed, were trained on previously created knowledge from the internet. Now, AI is generating more content than humans. In a few more years, the internet could be 99% AI-generated content, and that content will have been created by some form of AI. As a result, AI systems will constantly be distilling knowledge and intelligence from other AI systems. The fact that AI can learn is a good thing. We want AI systems to be intelligent because a smarter AI can also be a safer AI." ---- From "Axios" YouTube channel, (full video link in comment)
173
514
3,464
444,075
Jaime O retweeted
For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. The world needs both frontier closed models and frontier open models. images.nvidia.com/pdf/Open-W…
16,039
29,414
171,999
66,407,522
Jaime O retweeted
Watch Starship's thirteenth flight test x.lingyaoai.com/i/broadcasts/1MKgNNXAZ…
1,122
3,244
16,437
6,888,165
Jaime O retweeted
689
1,127
15,514
3,238,027
Today, we are introducing Inkling. Inkling reasons efficiently across text, image, and audio modalities. We are making the full weights available. thinkingmachines.ai/news/int… Available today for fine-tuning on Tinker. Play with it in the Inkling Playground. 🧵
560
2,002
14,879
8,326,977
Jaime O retweeted
Replying to @yacineMTB
I was clearly wrong about Anthropic. They are obviously currently the leader in AI. No company has released a model as good as Mythos/Fable and they will undoubtedly have Mythos 2 ready soon. And I would never cut them off in a way that hurt them badly, even as a competitor. That’s not my style. Tesla open sourced its patents and we made the Supercharger network available to all competitors, even though we could have made it a walled garden. SpaceX launches competing satellite systems with no increase in price or use of unfair terms. Even my worst enemies can attack me on this platform. …
2,870
6,073
88,976
4,966,433
Jaime O retweeted
For today’s GPT-5.6 release day, I wore the only appropriate T-shirt. (If you know, you know.)
85
25
1,808
271,387
Jaime O retweeted
We’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5. We'll begin restoring access tomorrow, and will share an update soon. We’re grateful to our users for their patience, and to everyone who worked with us on redeploying the models.
4,012
12,702
83,982
15,166,708
Jaime O retweeted
We’ve been impressed with GLM-5.2 and so are introducing a $9.99/month subscription to give you 2-5x discounted access to it and other open weight models like DeepSeek, Kimi, MiniMax, Mimo, Qwen. Use it on Cline CLI & IDE with $1.99 special promo if sign up via: npm i -g cline
353
272
2,872
1,096,696
Jaime O retweeted
Mythos / Sol cybersecurity capabilities are equally useful in an offensive as well a defensive capacity. If adversaries get ahold of an equivalent offensive capability, it poses a serious threat to US companies that remain unaware of latent vulnerabilities. In the meantime, I strongly recommend running deepsec[1] or similar harnesses with the available frontier models. [1] github.com/vercel-labs/deeps…
JUST IN: A new Chinese AI model from Zhipu AI reportedly matches Claude Mythos’ performance at finding security bugs.
84
105
2,250
404,605
Jaime O retweeted
📣📣 Meet Qwen-AgentWorld — a native language world model that simulates 7 agent environments (MCP, Search, Terminal, SWE, Web, OS, Android) within a single model. Environment modeling is the training objective from day one, not a post-hoc adaptation. 🤔 LLMs are trained to be better agents — better at acting in environments. But nobody has trained them to model the environments themselves. 🗺️ Our roadmap: investigate how language world modeling can push the boundaries of general agent capabilities, along two routes: 1️⃣ Build a foundation model for environment simulation — outperforming Claude Opus 4.8 and GPT-5.4 on AgentWorldBench 2️⃣ Investigate how world modeling enhances agent training: 🔬 Controllable Sim RL (agentic RL with LWM as environments) surpasses training in real environments 🧠 Learning to predict environments (LWM warm-up) makes agents stronger — remarkably, even without any agent-specific training, this predictive knowledge transfers to agentic tasks with zero fine-tuning 📑 Paper: arxiv.org/abs/2606.24597 📖 Blog: qwen.ai/blog?id=qwen-agentwo… 💻 GitHub: github.com/QwenLM/Qwen-Agent… 🤗 HuggingFace: huggingface.co/collections/Q… 🧩 ModelScope: modelscope.cn/collections/Qw…
207
788
4,878
1,174,926
Jaime O retweeted
Introducing GLM-5.2: Frontier Intelligence, Open Weights - Significant improvements in coding and agentic tasks - Strong long-horizon capabilities with a 1M context window - Two levels of reasoning effort: GLM-5.2 (max) pushes the limits, while GLM-5.2 (high) strikes a strong balance between performance and token efficiency - MIT-licensed open weights - Same API pricing as GLM-5.1 Tech Blog: z.ai/blog/glm-5.2 Weights: huggingface.co/zai-org/GLM-5… API: docs.z.ai/guides/llm/glm-5.2 Coding Plan: z.ai/subscribe Chat: chat.z.ai
721
1,740
13,134
7,667,526
Jaime O retweeted
Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for general use. Its capabilities exceed those of any model we’ve ever made generally available.
4,934
14,174
103,717
57,648,218
Jaime O retweeted
You've been asking for this one... Now in preview: Codex in the ChatGPT mobile app. Start new work, review outputs, steer execution, and approve next steps, all from the ChatGPT mobile app. Codex will keep running on your laptop, Mac mini, or devbox.
1,665
2,523
21,588
5,071,767
Jaime O retweeted
GPT-5.5 Instant is starting to roll out in ChatGPT. It’s a big upgrade, giving you smarter, clearer, and more personalized answers in a warmer, more natural tone. And it's also more concise, which we heard you wanted. We think you'll love chatting with it.
813
1,029
10,494
2,059,641
SpaceX is preparing to launch next-generation Starlink V3 satellites using Starship. These satellites can handle massive data, with up to 1 Tbps downlink and 160 Gbps uplink speeds. With Starship, many V3 satellites can be launched at once, helping build a faster global internet network.
209
966
6,869
440,511