Getting open models production-ready takes constant software tuning.
Exciting to dive into this with Victor Su Ortiz (@MiniMax_AI ) & the @vllm_project team at Open Intelligence Summit by @digitalocean on Oct 13.
If you serve open models in prod, or build AI-native applications, come join us: do.co/4ApUvzV
Introducing the Open Intelligence Summit. We're bringing together AI founders, builders, and researchers to discuss everything open: weights, data, agents, infra, and tooling.
The intelligence layer is the product now. Own it.
October 12-13. San Francisco. Register here. 🔗 do.co/4A3yt69
Closed out Ray Summit with a happy hour our team co-hosted with @nvidia, @inferact, and @NousResearch.
Gave a talk on how @digitalocean & @vllm_project topped the Artificial Analysis leaderboard for model performance, and later Balaji demoed our Inference Router live!
Thanks to everyone who stopped by.
Valuemaxx your time after Ray Summit and the vLLM Conference. Join our happy hour, lightning talks, and live demos with @nvidia, @Inferact, and @NousResearch. 🍻
Open inference, open convos, open tab. 🍻
RSVP today. 🎟️ do.co/4g0oAhE
Inference routing is having its moment! 🚀
@DigitalOcean inference router is now cache-aware!
You don’t want to lose your warm cache for repeated context by switching models mid-conversation — that means higher latency and higher cost.
Cache-aware routing keeps the context where it belongs: warm.
digitalocean.com/blog/infere…
Inference routing is having its moment! 🚀
• OpenRouter announced joining Stripe this week! (congrats @alexatallah and team)
• At the Agentic AI Summit @ UC Berkeley, I broke down how we built the @DigitalOcean / Plano open-source inference router.
• Enjoyed reading "The Semantic Routing Moment" liuxunzhuo.com/semantic-rout… by @XunzhuoLiu — great piece on how the industry is converging toward letting users express a preference (speed/cost/accuracy) rather than picking a model.
Speaking at Ray Summit 2026 in SF!
@digitalocean + vLLM (with @inferact ) hit #1 output speed on DeepSeek, per Artificial Analysis — plus NVIDIA Blackwell Ultra.
This year's Summit co-locates the first vLLM Conference. Come say hi.
#RaySummit#vLLM
Amazon Personalize Now Generally Available
We are excited to announce the general availability of Amazon Personalize, a machine learning service that makes it easy for developers to create individualized recommendations for users of their applicati... aws.amazon.com/about-aws/wha…
Amazon Translate is now generally available! Whether it’s a few words or volumes of text, Translate provides fast, accurate & affordable translation that scales with your needs. amzn.to/2IuiMKm
From an article, “Collectively, the five FAANG stocks had a market capitalization of about $2.8 trillion at the end of last week. The entire value of the U.S. stock market was about $28 trillion”