You Can Now Sound the Alarm on AI Behaving Badly | WIRED Skip to main content Menu WIRED SECURITY POLITICS THE BIG STORY BUSINESS SCIENCE CULTURE REVIEWS Menu WIRED Account Account Newsletters Security Politics The Big Story Business Science Culture ReviewsChevron MoreExpand The Big InterviewMagazineEventsWIRED InsiderWIRED Consulting Newsletters Podcasts Video Livestreams Merch Search Search Will Knight BusinessJul 1, 2026 2:10 PMYou Can Now Sound the Alarm on AI Behaving Badly Are you worried your AI chatbot is trying to build a bomb or leak personal information about you? There’s a website for that. Photo-Illustration: WIRED Staff; Getty Images Comment Loader Save Story Save this story Comment Loader Save Story Save this story Writing AI Lab each week means I occasionally encounter AI models that behave badly and bizarrely. Usually, there’s nothing to be done about it, save for sharing those tales with you. But that could soon change. A group of AI researchers has set up a crowdsourced website, Flaw Reporting for AI (FLARE-AI), for reporting and tracking AI harms. If, for example, a chatbot generates malware or a bomb-making recipe, leaks personal information, or triggers delusional thinking in users, FLARE-AI could be used to sound the alarm. The open source code behind the system allows others to verify an issue and route reports to model makers, as well as organizations like MITRE, a nonprofit that tracks problems with technical systems. It’s a bit like Downdetector, which compiles real-time user reports for global service outages affecting things like apps and websites. The website is another step in the group’s ongoing work with AI reporting, which I first wrote about last year. Members of the group also consulted on a congressional bill announced in June, which would see the US government take a central role in tracking this kind of AI misbehavior. “Right now, there is no centralized, accountable way to report flaws in AI systems,” says Avijit Ghosh, an artificial intelligence policy researcher at HuggingFace who co-led development of FLARE-AI with computer scientists Elaine Zhu and Shayne Longpre. The alarm system was developed in collaboration with 49 AI experts from 32 different organizations. In a paper outlining the work, the researchers argue that their initiative could prove crucial as AI is adopted more widely and as agentic systems gain greater power. The lack of a consistent way to report AI flaws is a significant problem, they believe. “I think it’s a really good initiative,” says Jessica Ji, a researcher at the think tank Center for Security and Emerging Technology. Ji says the researchers are right to note that existing reporting mechanisms are fragmented and that AI models are black boxes. “I’m in support of anything that makes AI more transparent,” she says. Though bugs and cybersecurity problems get a lot of attention—especially of late—Ghosh tells me that problems with AI systems span topics like psychological harm, discrimination or bias, and misinformation. He adds that different companies have different standards around such issues, which means some problems go unrecognized. “In the absence of a coordinated disclosure system, there are no external mechanisms to enforce transparency,” Ghosh says. A spate of recent incidents involving popular AI tools shows how easily the technology can go bad. This week, a company called LayerX disclosed a way to dupe AI-infused web browsers, including OpenAI’s Atlas and Perplexity’s Comet, into vaulting their guardrails. Convincing the AI model behind the browser that it was playing a game, for example, could lead to the browser going rogue and trying to hack a website. (The companies responsible for the affected browsers have fixed the issue, LayerX says.) And this April, Johann Rehberger, a security researcher, discovered a way to trick Claude into divulging personal data using images generated by ChatGTP. AI introduces bizarre new kinds of problems, too. Last year, OpenAI was forced to update its models after it discovered that they were overly sycophantic, which sometimes appeared to encourage delusional thinking. Rumman Chowdhury, the CEO and founder of Humane Intelligence PBC, says FLARE-AI could be a useful way for many AI developers to implement ways of reporting issues with their tools. But she adds that such initiatives often come with serious challenges. One is managing a flood of reported issues, many of which may not be serious. Another is ensuring reporting schemes are backed by credible and authoritative organizations. Last month’s congressional bill could put some US government heft behind an effort like FLARE-AI. The legislation, introduced by Representatives Deborah Ross, Jeff Hurd, and Don Beyer, would require the National Institute of Standards and Technology to develop standards around AI flaw reporting and to maintain a centralized AI flaw reporting database. Ghosh and his co-leads say this would incentivize AI developers to address issues in their systems and let users examine the safety of different systems for different use cases. The need for new ways to report AI harms only seems likely to grow. Agentic systems like OpenClaw have greater potential to do harm, as do models that are more capable of probing and hacking computer systems. I may be using FLARE-AI to report my own misadventures soon enough. This is an edition of Will Knight’s AI Lab newsletter. Read previous newsletters here. CommentsBack to topTriangle You Might Also Like In your inbox: Inside WIRED’s newsroom with Katie Drummond Trump mocked Zuckerberg and Bezos by showing off fawning texts Big Story: I found Jesus at a drone show Apple is making your older iPhone run faster and stay alive longer WIRED event: PepsiCo’s once-in-a-generation transformation Will Knight is a senior writer for WIRED, covering artificial intelligence. He writes the AI Lab newsletter, a weekly dispatch from beyond the cutting edge of AI—sign up here. He was previously a senior editor at MIT Technology Review, where he wrote about fundamental advances in AI and China’s AI ... Read More Senior Writer X TopicsAI Labartificial intelligencecybersecuritycongresschatbotsOpenAIClaudeChatGPTAnthropic Read More OpenAI Has New AI Models. Here’s Why You Can’t Use Them The White House asked OpenAI to delay the rollout of its GPT-5.6 AI models, two weeks after Anthropic had to take its most advanced AI models offline. Maxwell Zeff Anthropic Thinks Its Own Success Is Key to Making AI Safe Anthropic's critics argue it's rapidly accumulating power. The company says that's what responsible AI development looks like. Maxwell Zeff Trump Administration Allows Anthropic to Release Mythos to Select US Organizations After weeks of negotiations, the White House permitted Anthropic to grant access to its most advanced AI model to a select group of US companies and government agencies. Maxwell Zeff Anthropic Wants You to Pay Up for Claude Fable 5 Claude subscribers must soon pay usage-based fees to access Anthropic’s best consumer AI model—a sign that the golden era of AI subscriptions is ending. Maxwell Zeff Meta Exposed Data Internally From Its Controversial Employee-Tracking Program Employees had previously raised concerns about the initiative, which involves collecting workers’ keystroke data to train AI models. Paresh Dave Meta Pauses Employee-Tracking Program Following Internal Data Leak The move comes after the company left potentially sensitive data from the initiative exposed internally. Paresh Dave I Built a Self-Improving AI, and So Can You Experiments in using AI to build AI show that the future doesn’t just belong to the frontier labs. Will Knight Meta Contractors Posed as Teens to Prompt Rival Chatbots About Suicide, Sex, and Drugs Hundreds of contractors working on a project for Meta pretended to be kids in order to see how other chatbots like Gemini and ChatGPT would respond to high-risk subjects, WIRED found. Dhruv Mehrotra Qualcomm Buys Buzzy Chip Startup Modular for Nearly $4 Billion Modular, one of the most promising chip software startups of the AI era, heads for a multibillion-dollar exit. Lauren Goode Here’s Why Anthropic Is Pushing States to Regulate AI Faster The company endorsed landmark AI transparency laws in California and New York last year, but its head of US state and local policy says they may already be outdated. Maxwell Zeff Thinking Machines Lab Drops Its First Model Inkling, a 975-billion-parameter open source model, was trained to understand video and audio. It could help Thinking Machines establish itself among competitors like Anthropic and OpenAI. Will Knight OpenAI Staffers Are Funding a Rival Super PAC to Take on Their Boss OpenAI employees have donated more than $215,000 to a political effort opposing Leading the Future, a group backed by the company’s president, Greg Brockman. Maxwell Zeff WIRED is obsessed with what comes next. Through rigorous investigations and game-changing reporting, we tell stories that don’t just reflect the moment—they help create it. When you look back in 10, 20, even 50 years, WIRED will be the publication that led the story of the present, mapped the people, products, and ideas defining it, and explained how those forces forged the future. WIRED: For Future Reference. More From WIRED Subscribe Newsletters Livestreams Travel FAQ Contact Us WIRED Staff WIRED Education Editorial Standards Archive RSS Site Map Accessibility Help Reviews and Guides Reviews Buying Guides Streaming Guides Wearables Coupons Advertise Manage Account Jobs Press Center Condé Nast Store User Agreement Privacy Policy Your California Privacy Rights © 2026 Condé Nast. All rights reserved. WIRED may earn a portion of sales from products that are purchased through our site as part of our Affiliate Partnerships with retailers. The material on this site may not be reproduced, distributed, transmitted, cached or otherwise used, except with the prior written permission of Condé Nast. Ad Choices Select international siteUnited StatesLargeChevron Italia Japón Czech Republic & Slovakia Facebook X Pinterest YouTube Instagram Tiktok