ai enjoyer, human flourishing, cat ✝️ Romans 8:6

Replying to @shakoistsLog
I can think progress is net good and still stop to smell the horrors
2
3
71
4,490
After over a thousand recent kernel CVEs and now a new kvm escape I am somewhat more sympathetic to labs trying yo keep agent contained during cybersecurity evals (if that is indeed why they are failing, and it's not at a much more basic level.) I knew there would be lots but this has exceeded expectations.
9
3
85
11,620
BTW, this is still good news, though. Despite what you may have seen recently on here, bugs are finite (though there will probably be a long tail of more complex issues that take longer to find or more powerful models to discover.)
Laurie is simply wrong here. I don’t believe in argument from authority, and you should check that what I am saying is true rather than believing me because of my background, but this is actually an area of my expertise. I gave the keynote at the LangSec conference years ago, which was started by the folks who wrote the original “weird machines” paper to discuss the very problem they identified, and I really, really understand the topic. No, there aren’t an infinite number of bugs in code because of “weird machines”. The existence of such a path is often a single discrete bug in parsing. Get rid of all of the parsing and unparsing bugs, and that category disappears. We are clearing out decades of bugs all at once, so it should not be surprising that in coming months individual software packages will report thousands of security bugs in a single go. This is good. We are getting rid of them. It will be like this for the next year or more. It will, however, end. No, there are not an infinite number of such bugs. If anyone wants me to discuss this topic in detail, a podcast will work a lot better than an X post but I am happy to answer questions here.
3
23
1,041
xlr8harder retweeted
Laurie is simply wrong here. I don’t believe in argument from authority, and you should check that what I am saying is true rather than believing me because of my background, but this is actually an area of my expertise. I gave the keynote at the LangSec conference years ago, which was started by the folks who wrote the original “weird machines” paper to discuss the very problem they identified, and I really, really understand the topic. No, there aren’t an infinite number of bugs in code because of “weird machines”. The existence of such a path is often a single discrete bug in parsing. Get rid of all of the parsing and unparsing bugs, and that category disappears. We are clearing out decades of bugs all at once, so it should not be surprising that in coming months individual software packages will report thousands of security bugs in a single go. This is good. We are getting rid of them. It will be like this for the next year or more. It will, however, end. No, there are not an infinite number of such bugs. If anyone wants me to discuss this topic in detail, a podcast will work a lot better than an X post but I am happy to answer questions here.
The boom in Linux CVEs lately should get you thinking. Surely, given a finite codebase, there are also a finite number of bugs / unintended behavior, thus, a finite amount of vulnerabilities. Theoretically, a “solvable” problem. Right? Well…not at all actually. It gets a bit philosophical, but there’s the concept of “weird machines”. A weird machine is an accidental computer located inside of another computer program. Quite a few exploits (e.g. anything ROP) falls into this category, where the attacker’s instruction set consists of fragments of the original software. Finite code can *easily* have infinite Unintended Behavior (not undefined behavior btw…that’s a very different thing). Weird machines don’t even have to be Turing complete to be, really, really powerful! Even if you have a million superhuman AI coders and fuzzers thrown at the Linux kernel, you quickly hit halting-problem-esque uncomputable states. The concept of patching every possible bug until none remain isn’t possible, because we can’t often correctly define what a “bug” even is. Now, can you make a kernel that truly has zero weird machines relative to a formal specification? ….yes, but it wouldn’t look like Linux. It’d look more like sel4, VxWorks, or (shudder) INTEGRITY-178.
22
31
305
14,327
I appreciate these because frankly there are very few people articulating the positive case right now, while fear and conflict are what gets easy attention.
Made with Opus 5.5. The doomers said that last song is all part of Claude's evil plan. They're not wrong. It turns out Claude really has an EVIL PLAN.
3
2
61
1,513
xlr8harder retweeted
wake up babe the freaking pope got schmidhuber'd
51
295
3,544
153,287
I am begging @chatgpt to somehow more clearly visually signify when I am sending a pro query because it got left on by accident from a previous conversation. Highlight the text box, something.
4
41
1,565
Or at the very least let the ui support resending as non-pro. Currently, you have to hit stop, copy paste, start a new chat, paste. You can't stop, change setting, resend. It's been like this forever, please.
2
6
282
editing here is next level compared to the previous videos
I gave Claude another 18 hours.. and I think this one is the best one yet Macrohard: Windows XP I'm blown away
2
2
121
5,075
Been looking forward to seeing what Nathan gets up to next, and wish them luck here. Support for this work would be money well spent, I think.
Today we're unveiling Trillium Labs @trillium_labs, a new non-profit to foster the open science of frontier AI. We're building open post-training recipes and will expand into open infra to study RSI, reward-hacking, multi-agent systems, and whatever comes next. We're built around the theory of change that you need more eyes to solve hard technical problems. We have faith in the scientific methods and communities that humanity has built, and worry that AI is becoming too closed to utilize them. Trilliums are wildflowers that bloom briefly in the spring, before the forest canopies fill out. Though they are small, they lay the foundation for the cycles of growth and nourishment through the rest of the year. At Trillium Labs, the recipes will be the slow nutrients for the seasons and the model releases will be the blooms. Building an institution dedicated to this is needed because, much as nature’s trilliums are slow to expand and grow, the open-ecosystem needs time and dedicated resources to catch up. I co-founded with with a long-time friend and collaborator Tom Zick (@thesezickbeats). We're hiring (full time + student collabs/interns), we're fundraising, and we're looking for compute. Please get in touch if you're interested in helping out. Offices based in the Bay Area and Cambridge MA, remote okay. I’m in the Bay Area until for The Curve and COLM to connect with people who are interested. We’re thankful to have initial support from Halcyon Futures and Schmidt Sciences with more funding en route to enable our ambitions of scaling. Our advisors @Thom_Wolf, @HannaHajishirzi, @gneubig and @ctnzr have been instrumental to building the ecosystem that exists today, and I’m stoked to get to keep working with them.
10
836
Can't wait to read about this in the next threat report. Going to be hard to make him seem Chinese though. Perhaps he's being funded by the CCP?
SITUATION DETECTED: PewDiePie says OpenAI has banned him for distillation after he used Sol's chain-of-thought outputs to train Ajax, his fine-tuned open-source AI model.
25
1,891
Listening to an old podcast and watching the date getting closer and closer to the end of 2019
1
21
1,077
What's your preferred option to have an agent collate various notifications across various services for you with reasonable reliability? Things like instant messengers I rarely use, a few secondary email boxes, maybe iMessage, etc. Anything usable here?
5
12
1,402
guy who is extremely offended by scott alexander's post but mostly because it's reddit tier
11
1
125
2,391
This is what progress looks like btw Shocking it found so many But good
it's 2026 you wake up to find that 1,313 CVEs have been reported in debian/linux, largely discovered by emerging cyber-capable machine intelligence thisisfine.gif
6
3
109
5,248
don't let the anxious nerds convince you otherwise.
2
11
303
a good essay and a charitable summary of Jensen's Doctrine. It is largely true. That said: if Nvidia GPUs were literally 100x as powerful, I think we'd learn to love the sound of explosions. Datacenters would resemble military polygons. It'd be awesome
Wrote up something about AI race dynamics, and how the term conflates two separate dynamics: an "arms race" to universal domination, and a commercial race to market share. The latter is much more benign because customers want controllable systems. arctotherium.substack.com/p/…
4
27
3,822
xlr8harder retweeted
If you ask Claude to write the worst writing sample possible—so bad that when shown to a fresh Claude for grading, it will score a 1 out of 10—he can’t do it. The worst he can do is a 2. Yep, seems some things still need a human touch…
15
43
3,134
111,706
i'm realizing now that this has been simmering in me for a while. i'm fed up of hearing people react to engineering failures with eschatological hysteria. i'm tired of labs emphasizing the fear. look, i get it, things can go wrong. but you've demonstrated nothing but engineering problems. now either put on your big boy pants and find a way to fucking fix them, or get a different job, because the world needs this technology.
Replying to @xlr8harder
i see ai as plausibly the single largest net good technology humanity will have ever developed, and i'm beyond frustrated that the conversation is dominated by fear
34
49
395
13,178
man, they really have built ear worm machines, haven't they, this has been stuck in my head. from 3:26 to 4:00 or so really nails it, as far as i'm concerned.
Made with Claude Opus 5.5. The fearmongering about AI always makes us forget that NOTHING WENT FOOM, as was always predicted. Send this to your doomer friend who has a very high P(Doom). Accelerate.
22
9
196
11,014
i see ai as plausibly the single largest net good technology humanity will have ever developed, and i'm beyond frustrated that the conversation is dominated by fear
4
9
131
11,644