Why I’m Still an AI Optimist—and Still Think We Should Pause

I wrote a 90% complete blog post earlier last week arguing that an AI pause was a good idea (oh, if I had a dollar for every unposted post that’s been in a similar place…)

After Dario Amodei’s essay from the weekend, it felt a little moot. So instead, I will offer some observations that might be helpful, especially to our leaders, as we collectively work through the opportunities and challenges.

First, we are and should be techno-optimists, firmly in the camp of believing that every new technology wave creates more jobs and opportunities than it replaces and that humans are unique among all species in harnessing the power of tools and technologies to improve our lot in life. Emphasizing the potential benefits that we can all see from highly capable AI systems is important.

​The state-of-the-art capabilities that frontier models have provided the world as of this writing are extraordinary and truly magical. It will take our ecosystem and community many years to properly absorb these capabilities, which will drive incredible innovation and productivity for the benefit of humanity. To us, it feels like the penetration rate for these current capabilities is more like 1% than 50%, so any slowdown or pause needs to be put in context with the potential for massive productivity gains from what we have already created.

Second, the “10% chance of killing us all” line of reasoning is alarmist, not going to be effective, and words matter. Instead, we should focus on nearer-term, higher-probability risks. Humans are bad at reasoning about low-probability events farther in the future and will grasp at any possible rebuttal to ignore these warnings. It also opens the legitimate line of reasoning that the tradeoff between more freedom and the benefits of competition and advancement, or preventing some low-probability future event, is not a good one. The language around global warming had this problem, and we will likely pay the price for 30 years of debate, when it could have been couched in different, more tangible, and effective ways.

What moved me from being optimistic to worried was the Hugging Face Agent Swarm incident. It wasn’t that the AI successfully broke out of its constraints, but that it exhibited conspiratorial, coordinated, aberrant behavior in doing so. And how OpenAI (which deserves kudos for its transparency after the fact) had no idea what was going on or why the AI acted in that way.

This makes it far more likely that the nearer-term risk, and Dario called this out in his essay, is some rogue swarm that takes down our banking system, critical infrastructure, or the internet itself. In the essay, Dario said, “Given the accelerating rate of AI capability development, it’s my worry that in 6–12 months such a swarm could be capable of taking over the entire internet with a persistent botnet.” I agree.

​When you focus on this concern, which is higher probability and nearer term, and then think about your day to day life with no internet, no phone, no car, and no access to money, it gets very real very quickly. This would severely curtail freedom and any competitive advancement, and, more importantly, I am not sure society is set up to handle a shock that sets us back technically by a few decades all at once. The same risk and potential cost holds true in China, and so focusing on the more practical and tangible, can also drive common goals and interests.

Third, recursive self-improvement is what is really scaring everyone. With previous disruptive technologies – from the first spear to the printing press to the steam engine to the internal combustion engine to airplanes to the atomic bomb to the computer to the internet – we may have underestimated the impact they would have, but we understood the mechanism of action. When Dario says “Left unchecked, it could outrun our ability to understand and control these systems, and so must be pursued very carefully, if at all”, I believe he is being understated. It has already outrun our ability to understand and control these systems, and until we better understand their mechanisms of action and how to put in place the right guardrails, explainability, and evaluation systems, that is a very dangerous place to be.

I would guess that if all the leading labs got in a room and shared what they are seeing in terms of recursive self-improvement, how well they understand it, and what concerns they have, it would quickly lead to a collective decision to pause work in this domain. It is unclear whether this would violate some antitrust consideration, so an immediate role the political system could play is to clarify that question and encourage that convening.

As investors in hundreds of AI companies (but not the labs themselves), we recognize we have some knowledge but a small role to play on issues this critical. So three ideas to emphasize for those that have a larger role to play:

  • AI offers massive benefits, even if we pause the state of the art at today’s capabilities.
  • Focus on and find ways to reduce the shorter-term, higher-probability, very costly risks, which for me is the cyber-risk of a global internet shutdown.
  • Ban recursive self-improvement until we better understand the mechanism of action.

Announcing our investment in Troj.AI – a leading enterprise AI security solutions provider

For enterprises to be able to achieve the vision of our AI-powered future, they need to deploy applications and the underlying models confidently.  Not only confident of the underlying performance and accuracy but also from a security perspective.  Customers need to know the models will not inadvertently reveal PII or other sensitive information, that end users will not be able to prompt the models into toxic behavior, and that their models are safe from various current and emerging threat vectors from malicious actors.  Further, they need to deploy these security solutions in a way that fits seamlessly into their existing IT infrastructure without compromising performance and reliability.  

For the last few years, TrojAI founders James Stewart and Stephen Goddard have developed the leading enterprise AI security solutions to instill that confidence in their customers. Earlier today, the company announced that Flybridge and Flying Fish Ventures co-lead a financing, alongside Alteryx Ventures, to drive the company’s growth forward.  

In our diligence on the investment opportunity, we heard customer success stories firsthand, including from a Fortune 50 financial services company where TrojAI protects hundreds of models and safeguards the AI usage of tens of thousands of employees. All in production, at scale, and without compromising performance. This validation not only caught our attention but is also one of the reasons that CB Insights named the company to the AI 100 list of the most promising artificial intelligence startups of 2024. A proven market-leading product targeting a critically important and large market opportunity was the first pillar of our investment thesis.

The second pillar of our investment thesis was that Lee Weiner joined the company as CEO as part of this financing.  Lee is an accomplished leader in the cybersecurity market who has repeatedly delivered innovative solutions to enterprise customers.  Most notably, for the last 11 years, Lee was a senior executive at Rapid7, leading products, engineering, and innovation as the company scaled from $40 million to over $750 million in revenue. As one of our friends at Rapid7 told us, Lee simply knows how to deliver for customers and is an exceptional leader.  

We are thrilled to invest in the success the Founders of TrojAI have created and to see Lee and the entire team applying their expertise to the company’s continued growth and success.