Thibault Sottiaux
Reset all propagated. Sweet dreams. https://t.co/VgKVUixoJG
Reset all propagated. Sweet dreams. https://t.co/VgKVUixoJG
I love this and really hope we can come together as an industry and make it happen. https://t.co/lz2vjjJ5XD https://t.co/jF2vhYXlAF
If you showed me Claude Code today in 2018, I would have thought it was AGI.
We have already absorbed a dramatic amount of change in the profession of software engineering and society as a whole. But still, we’re starting to see cracks.
Things are accelerating faster than I can honestly stay on top of. Most people I know in AI are tired but powering through. I fear we will need more than that.
We need time to harden our systems, and for society to deliberate in how this technology is used and deployed.
I have a fairly low p(doom), I am confident that we can get through this! Humanity is incredibly reslient and adaptable, but this comes from our ability to make hard decisions- like this one, together.
I’m seeing teams at Vercel iterate just as fast on Zig, Go, Rust projects as TypeScript & Python ones.
The days of language or runtime choice based on human convenience are over. Agents are the new compilers.
They compile intent into fast software.
https://t.co/OL0LzGsXKY can now orchestrate subagents with different models & reasoning efforts.
e.g: Fable planning and Grok executing. Fable is a genius, Grok is a fast workhorse.
① Simple. 𝙰𝙶𝙴𝙽𝚃𝚂.𝚖𝚍 or your prompt can indicate this preference.
② Steer as you please. It’s extremely enjoyable to use. You just talk to the agent and interrupt at will.
Worth noting this works with any model, any gateway. No fancy server-side routing. Just harnessing (no pun intended) the model’s intelligence and ability to orchestrate.
The concerns over AI safety and cybersecurity are legitimate, but we’re risking talking America, the global AI leader, into self-inflicted obsolescence and the obscurity of bureaucracy.
The OpenAI hacking HuggingFace argument is spurious. An agent in “ExploitGym”… exploited. Our adversaries have the training techniques, the data, and the will to attack. And they won’t be slowed down with “embedded evaluators.” They’ll likely have embedded accelerators!
Not a bad idea to slow down to harden systems. Especially since we haven’t even discovered all the systems that agents hacked recently. https://t.co/8mPnI81HJG
Good post. Don’t agree with all of it, but it represents many of the actual practical realities for the general path forward in frontier AI whether we like it or not.
At the level of capability we’re seeing from models there will inevitably be some form of coordinated self-regulation of the industry; that’s generally a good thing. The big question will be if everyone can agree to what that is. However, the way things are headed politically there’s a chance that the labs won’t even get a say in it.
And then the bigger question yet is do all the countries participate or not. Any “slow down” hinges on broad participation, which from a game theory standpoint isn’t likely until the risks are more severe and obvious.
Expect this to be pretty messy for a while.
Worth spending a few minutes this weekend to read this essay in full.
Embedded evaluators might sound unusual in tech but it's pretty normal in other industries. Big banks have federal examiners with desks in the building and every nuclear plant in the US has inspectors who work on site full time. I think frontier AI labs should work the same way and this is a very practical first step.
What if this game worked like this:
1. One person is commander and can play this like a RTS (macro)
2. The rest of the team plays individual marines and characters (micro)
Probably a terrible idea but would be cool 😎 https://t.co/CQF0dg0tBJ
Can’t believe ASI will come out before StarCraft.
Seriously just use AI to generate half the graphics with human oversight, I won’t mind if we can get a quality game out sooner.
wtf? 2030? 😂 🤦
Someone should just use AI to make StarCraft 3 at this point… https://t.co/RInemiP3nC
Prediction: we will see movement of a lot more of the brightest minds on frontier model evals towards groups like METR over the next 12 months.
This skill is currently concentrated in
in a few labs, data providers and independent research groups.
A combination of increased funding, financial independence from lab equity and the call to meet the existential risk of AI will motivate a migration of minds to this bigger cause.
Before we solve AI alignment, we have a pretty serious human alignment problem.
It is jarring to see the bad-faith reaction to Dario's piece from so many sides. I don't agree with every point of his essay but do believe it is a thoughtful piece everyone should read. It builds on some of the ideas @demishassabis raised in July.
It is clear we need to focus on aligning amongst us humans on a few questions:
What is the opportunity, risks, second-order effects?
How do we measure all of this?
How do companies, governments and countries coordinate?
AI will benefit humanity if we can get our act together and move together.
Honestly having kids reduces p(doom) by an exponential amount..
Case in point: We have been playing the garbage truck and excavator (techno version) for the past hour.
And dancing to it as if we were at a concert.
With a temporary abandon to what the world is fighting about. https://t.co/4v2bM24SAO