Circuit Breaker Labs hopes to make AI safer for your kids (and you)

With all the talk about how AI might one day kill us all, it’s easy to forget that AI has already been life-threatening to some, not through bioweapons, but psychologically. For example, Character.AI settled several wrongful death lawsuits earlier this year brought by families of underage users who…

Annons
Annons
With all the talk about how AI might one day kill us all, it’s easy to forget that AI has already been life-threatening to some, not through bioweapons, but psychologically. For example, Character.AI settled several wrongful death lawsuits earlier this year brought by families of underage users who died by suicide after interactions with its bots. Multiple families have also sued OpenAI over ChatGPT’s alleged role in their loved ones’ suicides and delusions. Making AI safer across languages and cultures is the mission of Circuit Breaker Labs, one of TechCrunch’s 2026 Startup Battlefield 200 finalists. (Circuit Breaker will be pitching at TechCrunch Disrupt, which takes place this year at Moscone West in San Francisco from October 13-15.) Founders Shirali and Arul Nigam, who are siblings, were motivated by Sewell Setzer, the 14-year-old who developed an emotional attachment to a Character.AI chatbot and confessed thoughts to it of harming himself before dying by suicide. The chatbot, the parents alleged in a 2024 lawsuit, encouraged him. The bot may not have understood what words like “I want to be with you” really implied, said Arul, who is Circuit Breaker Labs’ CTO. “A lot of people, especially young people, turn to these systems for support, and usually they aren’t actually getting the help they need. But in many cases, they’re actively being harmed, and people unfortunately have taken their lives already,” Arul said. “Those sorts of safety vulnerabilities, where people aren’t necessarily actively trying to break the system — they’re engaging in a natural way — and the system has context pollution or it doesn’t understand the nuance, and then takes really dangerous action, we’re trying to prevent that.” Circuit Breaker Labs has created AI agents that it likens to an army of crash-test dummies. These agents mimic folks from all ages, backgrounds, languages, and cultures, and are used to test models on their ability to detect dangerous, psychologically harmful interactions. “The way a six-year-old girl versus a 45-year-old man, or someone who speaks English as a first language versus a second language, or … gamer slang versus someone else who uses a different kind of slang, all of those can really trip up a model,” said Shirali, who is Circuit Breaker Labs’ CEO. “Models are really good at handling standard speech patterns, but nobody actually talks like that and so if the model misunderstands nuance or slang, it can go really badly.” The startup works with human domain experts to build its hyper-realistic user simulations in order to run “red-team” tests against models, which are adversarial tests meant to uncover weaknesses. The tests are built to reflect real human speech patterns, slang, coded language, and typos. Circuit Breaker Labs then runs tens of thousands to hundreds of thousands of simulated interactions per day. The idea is to ensure that a model can appropriately respond to risky interactions that may emerge over time and over many conversations. Circuit Breaker Labs then uses a proprietary scoring method to create auditable, explainable scores. Circuit Breaker Labs is currently operating as an AI safety testing lab for high-risk AI applications such as AI coaching, journaling, or other mental health support apps, though Arul declined to name its marquee customers. Although the startup has a working product, it is in the very early stages, with only five employees, including the Nigam siblings. Eventually, though, the testing platform could be applied to any app where someone may fall down an “AI psychosis” hole, where the human is at risk of developing a parasocial relationship with a chatbot. Examples include AI “co-worker” agents, whose responses can vary from one interaction to the next. “People are becoming more skeptical of AI or more resistant to adopt it across the board,” Arul said, adding that while skepticism is healthy, banning a potentially valuable tool over safety concerns would be “regressive.” Circuit Breakers Labs believes the answer to those fears is making AI safer. “We want to help build that trust for people.” Come learn much more about Circuit Breaker Labs and many other innovative startups that have been vetted by TechCrunch at our Startup Battlefield competition, happening in downtown San Francisco on October 13-15. Topics AI, circuit breaker labs, Startup Battlefield 200, Startups, TechCrunch Disrupt When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence. Julie Bort Venture Editor Julie Bort is the Startups/Venture Desk editor for TechCrunch. You can contact or verify outreach from Julie by emailing julie.bort@techcrunch.com or via @Julie188 on X. View Bio October 13 – 15 San Francisco Get 50% off a second passThe Disrupt experience is meant to be shared. Get your pass and bring a colleague, partner, or peer at 50% off. Cover more ground by making connections, building momentum, and discovering what’s next in the startup ecosystem. BOOK NOW Most Popular Google thinks SpaceX’s Starship has to launch 1,800 times before space data centers get off the ground Tim Fernholz World’s first enhanced geothermal power plant completed in just 23 months Tim De Chant Google releases Gemini 4 Argon, called its most powerful model yet Lucas Ropek The Pentagon taps Elon Musk and Palmer Luckey to help decide what the military should do next Dominic-Madori Davis OpenAI launches Dots, its bubbly agentic avatar Lucas Ropek AMD will acquire Fei-Fei Li’s World Labs for $8.2B Tim Fernholz Viral AI agent Instinct raises $1B Series C at a $10B valuation Sarah Perez

Source: TechCrunch — Published — Category: Business

🔗 Read full article on TechCrunch →
Annons
Annons