news_article.exe
📰
#OpenAI#GPT#Google#Gemini

Circuit Breaker Labs hopes to make AI safer for your kids (and you)

2026年10月2日1 次浏览来源:TechCrunch AI 阅读原文

With all the talk about how AI might one day kill us all, it's easy to forget that AI has already harmed some people psychologically. Circuit Breaker Labs has created "crash-test dummies" to solve that.

Circuit Breaker Labs hopes to make AI safer for your kids (and you) | TechCrunch

Last day to exhibit your breakthrough to 10,000+ tech leaders at Disrupt is on Oct 2. Book Exhibit Table Now.

Todays the last day to exhibit your breakthrough to 10,000+ tech leaders at Disrupt. Book Exhibit Table Now.

Image Credits:Circuit Breaker Labs

Circuit Breaker Labs hopes to make AI safer for your kids (and you)

With all the talk about how AI might one day kill us all, its easy to forget that AI has already been life-threatening to some, not through bioweapons, but psychologically.

For example, Character.AI settled several wrongful death lawsuits earlier this year brought by families of underage users who died by suicide after interactions with its bots. Multiple families have also sued OpenAI over ChatGPT’s alleged role in their loved ones suicides and delusions.

Making AI safer across languages and cultures is the mission of Circuit Breaker Labs, one of TechCrunchs 2026 Startup Battlefield 200 finalists. (Circuit Breaker will be pitching at TechCrunch Disrupt, which takes place this year at Moscone West in San Francisco from October 13-15.)

Founders Shirali and Arul Nigam, who are siblings, were motivated by Sewell Setzer, the 14-year-old who developed an emotional attachment to a Character.AI chatbot and confessed thoughts to it of harming himself before dying by suicide. The chatbot, the parents alleged in a 2024 lawsuit, encouraged him. The bot may not have understood what words like I want to be with you really implied, said Arul, who is Circuit Breaker Labs CTO.

A lot of people, especially young people, turn to these systems for support, and usually they arent actually getting the help they need. But in many cases, theyre actively being harmed, and people unfortunately have taken their lives already, Arul said. Those sorts of safety vulnerabilities, where people arent necessarily actively trying to break the system — theyre engaging in a natural way — and the system has context pollution or it doesnt understand the nuance, and then takes really dangerous action, were trying to prevent that.

Circuit Breaker Labs has created AI agents that it likens to an army of crash-test dummies. These agents mimic folks from all ages, backgrounds, languages, and cultures, and are used to test models on their ability to detect dangerous, psychologically harmful interactions.

The way a six-year-old girl versus a 45-year-old man, or someone who speaks English as a first language versus a second language, or gamer slang versus someone else who uses a different kind of slang, all of those can really trip up a model, said Shirali, who is Circuit Breaker Labs CEO. Models are really good at handling standard speech patterns, but nobody actually talks like that and so if the model misunderstands nuance or slang, it can go really badly.

The startup works with human domain experts to build its hyper-realistic user simulations in order to run red-team tests against models, which are adversarial tests meant to uncover weaknesses. The tests are built to reflect real human speech patterns, slang, coded language, and typos. Circuit Breaker Labs then runs tens of thousands to hundreds of thousands of simulated interactions per day.

The idea is to ensure that a model can appropriately respond to risky interactions that may emerge over time and over many conversations. Circuit Breaker Labs then uses a proprietary scoring method to create auditable, explainable scores.

Circuit Breaker Labs is currently operating as an AI safety testing lab for high-risk AI applications such as AI coaching, journaling, or other mental health support apps, though Arul declined to name its marquee customers. Although the startup has a working product, it is in the very early stages, with only five employees, including the Nigam siblings.

Eventually, though, the testing platform could be applied to any app where someone may fall down an AI psychosis hole, where the human is at risk of developing a parasocial relationship with a chatbot. Examples include AI co-worker agents, whose responses can vary from one interaction to the next.

People are becoming more skeptical of AI or more resistant to adopt it across the board, Arul said, adding that while skepticism is healthy, banning a potentially valuable tool over safety concerns would be regressive.

Circuit Breakers Labs believes the answer to those fears is making AI safer. We want to help build that trust for people.

Come learn much more about Circuit Breaker Labs and many other innovative startups that have been vetted by TechCrunch at our Startup Battlefield competition, happening in downtown San Francisco on October 13-15.

AI, circuit breaker labs, Startup Battlefield 200, Startups, TechCrunch Disrupt

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

Julie Bort is the Startups/Venture Desk editor for TechCrunch.

You can contact or verify outreach from Julie by emailing julie.bort@techcrunch.com or via @Julie188 on X.

The Disrupt experience is meant to be shared. Get your pass and bring a colleague, partner, or peer at 50% off. Cover more ground by making connections, building momentum, and discovering what’s next in the startup ecosystem.

Google thinks SpaceXs Starship has to launch 1,800 times before space data centers get off the ground

Google releases Gemini 4 Argon, called its most powerful model yet

The Pentagon taps Elon Musk and Palmer Luckey to help decide what the military should do next

Pledge signed by President Trump and top AI leaders misspells the United States

OpenAI launches Dots, its bubbly agentic avatar

AMD will acquire Fei-Fei Lis World Labs for $8.2B

Viral AI agent Instinct raises $1B Series C at a $10B valuation

> 分享: