The open platform for automating development. Infrastructure to build, measure, and interact with agents across the SDLC

Pinned Tweet
Introducing Scorers: Agents that grade your agents. Use LLM-as-a-judge to grade past coding agent sessions on: • Quality • Efficiency • Compliance • Or any custom dimension Scores feed into performance measurements and automatic self-improvement for software factories
18
13
191
568,296
Ah shoot sorry, launched this 8 months early. Forgot to change timezones when I ran `oz schedule`
Launching something big tomorrow...
3
1
19
3,510
Agents are getting complex to manage: models, skills, automations, permissions, cloud environments, orchestration patterns... So, we built the Terraform for agents: one repository that configures all of your agents as code. The Warp Factories configuration spec defines: - Agents and orchestration: how work gets routed and which models agents use - Automations: the events and schedules that kick off work - Access and environments: repositories, secrets, MCP servers, and runners - Measurement: scorers and benchmarks for understanding and improving performance Because it’s all code, both engineers and agents can propose changes to the factory itself. That makes it possible to test new configurations, benchmark them, and even have your factory propose improvements as PRs.
9
6
71
114,148
Check out the interactive guide and clone the example templates to learn more about how it works. We'll help you deploy with $10k in usage for qualified companies warp.dev/factories/configura… 🔖
4
1,192
Full breakdown of Warp Factories: our platform for building software factories
This past month, we shipped Warp Factories. It's a really robust, flexible solution to deploy software factories with any model + harness configuration. There's a lot of moving parts, so I recorded a complete walkthrough of what it can do. Hope it's helpful!
1
2
19
3,826
Full rundown of our software factories platform if you've been curious
This past month, we shipped Warp Factories. It's a really robust, flexible solution to deploy software factories with any model + harness configuration. There's a lot of moving parts, so I recorded a complete walkthrough of what it can do. Hope it's helpful!
3
1
63
14,545
You should benchmark models against your own conversations to find the right one from a cost/performance perspective. Here's the benchmarking setup we use:
3
3
47
46,206
Software engineering is shifting to "factory engineering:" the focus isn't just on building the product, but improving throughput. One measure of factory efficiency is "human touches per pr." Over time you want to drive this down.
13
10
83
49,386
GPT 6.1 Sol is now in Warp, ready to use with your Codex / ChatGPT subscription
56
4,428
You can now sign in with ChatGPT from the Warp Terminal and the Warp Agent CLI! Sign in to Warp with your ChatGPT account and use your subscription’s included Work and Codex usage for your agent conversations.
12
15
160
13,501
Agentic engineering is really two disciplines now: - Product engineering: ensuring the quality and usefulness of product - Factory engineering: deploying agents to build that product, optimizing system efficiency and output quality Here's how to adapt:
Article

Adapting for a world of software factories

Software engineers have been through a lot of change in the past two years, transitioning from writing code by hand to steering agents via local interactive prompting, the current paradigm. Engineers

36
36
306
23,865
Sonnet 5.5 is now in Warp and the Warp Agent CLI We ran the numbers, and confirmed it has reached 100% SOTA on the "Write Sonnets about Rust" bench
3
1
28
6,126
Warp retweeted
This past month, we shipped Warp Factories. It's a really robust, flexible solution to deploy software factories with any model + harness configuration. There's a lot of moving parts, so I recorded a complete walkthrough of what it can do. Hope it's helpful!
7
6
78
23,345
I sat down with @clairevo to talk about how we're automating development at Warp. One big change is how we approach code review - for low risk changes, we now let engineers review agent PRs and ship themselves, rather than requiring a second dev’s review.
6
3
30
13,603
Warp retweeted
We @warpdotdev are teaming up with @warpdotco to host a poker tournament this Wednesday evening in NYC! Luma link & buy-in details in thread 😁
11
3
24
6,646
Grok 4.7 has landed in Warp and the Warp Agent CLI. Connect your @grok subscription to get started
3
2
25
7,489
Claude Opus 5.5 is now available in the Warp Terminal and the Warp Agent CLI.
1
2
49
5,100
GPT 6 Sol and Luna are now available in the Warp Terminal and the Warp Agent CLI
2
28
3,818
Introducing Scorers: Agents that grade your agents. Use LLM-as-a-judge to grade past coding agent sessions on: • Quality • Efficiency • Compliance • Or any custom dimension Scores feed into performance measurements and automatic self-improvement for software factories
18
13
191
568,296
Finally, use these metrics to have agents suggest improvements to your setup automatically. Self-improvement agents run on a schedule and review failing grades to suggest changes to your setup. Here is an example agent skills PR with evidence cited from previous scoring runs:
1
1
882
If you want to set this up, scoring is part of Warp Factories in early access. We are offering up to $10k in usage to qualified companies. warp.dev/factories/request-a… 🔖
1
847