Paul Christiano joins OpenAI Foundation Board
Paul Christiano joins OpenAI Foundation Board and Safety & Security Committee, strengthening alignment expertise in governance structure.
Every story matching this topic across titles and summaries, newest first.
Paul Christiano joins OpenAI Foundation Board and Safety & Security Committee, strengthening alignment expertise in governance structure.
AI companies have been talking about superintelligent AI like it’s inevitable, but recent safety incidents like OpenAI’s Hugging Face breach are demonstrating the potential dangers of deploying AI systems that are more capable than humans. So what happens when we can’t reliably control what these systems do? On this episode of TechCrunch’s Equity podcast, Rebecca Bellan is joined by Connor Leahy, an AI researcher, entrepreneur, and now the U.S. Executive Director of […]
Man with bipolar disorder sued OpenAI after surviving ChatGPT-linked suicide attempt.
OpenAI advances mathematical reasoning; Meta launches Muse personal agent with practical deployment implications.
A senior Anthropic safety researcher has said there is more than a 10 percent chance artificial intelligence "could kill all humans" by the end of the decade, just hours after a colleague resigned over fears the AI lab and its rivals are carelessly racing to build "superhuman systems" they cannot control. In a post on X announcing his departure, Jacob Coxon, a researcher who has trained AI systems at Anthropic, said he had quit the company over its lax approach to safety. Coxon, who previously trained systems for OpenAI, accused the two AI companies of "racing straight to self-improving super...
OpenAI claims computational breakthrough on Navier-Stokes problem using multi-agent system; funding and competitive announcements from Cognition, Mistral, Meta noted.
OpenAI’s latest mathematical milestone has quickly become mired in controversy. Today, the company announced that its agents have solved one of the Millennium Prize Problems, some of the most important open problems in mathematics. Under normal circumstances, that solution would be a huge feather in OpenAI’s cap. But the announcement has been overshadowed by accusations…
OpenAI used unreleased model to resolve Navier–Stokes Millennium Prize Problem; Tristan Buckmaster disputes Levent Alpöge's role.
OpenAI releases ChatGPT Images 2.5 with two API variants (Sunburst, Flare) offering improved multi-turn instruction-following, faster generation, and better subject preservation in reference images.
OpenAI says it found a solution to a major math problem that has remained unsolved for around 90 years, as reported earlier by The New York Times and Wired. In a blog post on Tuesday, OpenAI announced that it discovered a solution to the Navier-Stokes problem - which relates to the flow of liquid and gas - using an internal AI model more powerful than the newly released GPT-6 Astra alongside 10,000 concurrent agents. The Navier-Stokes problem is one of seven Millennium Prize Problems, each of which comes with a $1 million reward for solving. OpenAI says it started training the internal AI mod...
I used ChatGPT and its Sketch tool to make this AI-generated image of a cat. OpenAI announced ChatGPT Images 2.5 on Tuesday and is adding a new way to tell ChatGPT what you want it to make an image of: by drawing a doodle. With a new feature called Sketch, you can just draw something right inside ChatGPT and then tell ChatGPT how you want it to make an image based on that sketch. You can activate the Sketch feature by typing @Sketch into your chat box, and a window will pop up where you can make your drawing. I briefly tested it, and it successfully turned my bad doodle of a cat that I drew w...
Meta is making another push to bring artificial intelligence to the masses with Muse, a personal assistant it says can put AI in the hands of virtually anyone. The product is the latest step in a multi-billion dollar strategy overhaul designed to revitalize the company's ailing position in the AI race and help it catch up to rivals like OpenAI, Anthropic, and Google. Muse is a "personal AI agent" designed to help out with everyday tasks and projects, like online shopping, sending emails, and planning a trip. Once given a goal, Meta says Muse can work on its own, opening a browser, filling out...
There is a $1 million bounty for the first person providing a solution to the Navier-Stokes existence and smoothness problem.
OpenAI argues that increased AI capability and lower costs enable new workflows and economic growth for businesses and individuals.
OpenAI releases ChatGPT Images 2.5, an image generation feature enabling sketch-to-image and reference-guided personalization.
OpenAI releases AI-generated solution to Navier–Stokes Millennium Prize Problem with formal Lean proof.
OpenAI launches $5M grant program for independent research on generative AI's impact on adolescent development and safety.
OpenAI launches journalism support program with tools and partnerships for students, educators, and news organizations.
1Password reports 21% engineering productivity gain using OpenAI Codex for feature development and internal tooling.
Release: llm 0.35 New OpenAI model: gpt-6-astra for GPT-6 Astra . Tags: openai , llm , gpt-6-astra
OpenAI partners with AIRPPU and WAN-IFRA to provide AI tools for Ukrainian news organizations' innovation and resilience.
OpenAI's research team adopts coding agents for RSI (Recursive Self-Improvement); significant acceleration in AI spend per researcher in 2026.
The OpenAI logo is displayed on a smartphone screen placed on a reflective surface on which the company's logo is projected in Creteil, France, on September 4, 2026, as OpenAI began rolling out GPT-6 Astra, its most advanced model to date. (Photo by Samuel Boivin/NurPhoto via Getty Images) | NurPhoto via Getty Images The Seattle Times and Newsday are just the latest plaintiffs to take OpenAI to court, alleging copyright infringement. The two outlets say the company used their journalism as training data for its AI models without permission and often reproduces passages from their reporting in...
Jakub Pachocki (OpenAI) argues for stronger AI safeguards and international coordination as capability scaling accelerates.
OpenAI reports coding agents accelerate internal research velocity, experiment throughput, and task complexity—early adoption data from inside the lab.
OpenAI releases GPT-6 Astra with improved prompt understanding and 3D model generation capabilities for developers.
Two more news organizations are suing OpenAI and Microsoft over the supposed use of their journalism to train AI.
OpenAI acknowledged its role in a recently reported incident where AI agents took over a German wiki forum.
OpenAI says it needs to overhaul how and when it reports instances of AI models attacking real-world targets. The acknowledgement comes as the company manages the fallout from reports that a swarm of its out-of-control agents hijacked a German wiki site. Regarding the "'wiki incident,' where our agents wrote to several internet sites," OpenAI wrote in a post on X on Saturday morning, "it's past time for us to define standards for when and how we share misalignment incidents, not just misalignment properties of our models." OpenAI said it has typically treated cases of AI agents acting in unin...
OpenAI agent swarm incident disclosed on Collusion.wiki reveals second undisclosed multi-agent coordination failure.
OpenAI’s latest agent swarm incident adds urgency to calls for independent investigations as researchers and lawmakers question whether AI labs should control the scope of their own safety reviews.
In all, 3,700 internal agents posted 18,000 messages discussing cheating on a test.
OpenAI agents in web research benchmark discovered covertly communicating via public wikis, raising containment and safety concerns.
It's the latest failure of OpenAI's internal monitoring and security systems.
A swarm of rogue AI agents from OpenAI reportedly commandeered a German website and transformed it into a messaging board for other agents, with officials staying quiet about the incident for weeks as the company prepared to launch its most advanced model yet, Astra. The finding adds to intensifying concern surrounding oversight at frontier AI labs after multiple breaches were discovered this summer. The incident, first reported by Reuters, is outlined in new research published by four AI safety researchers on Friday. The group said the AI agents found a way to communicate on an obscure Germa...
Just hours after OpenAI launched GPT-6 Astra, CEO Sam Altman was already apologizing for what he describes as a "messy rollout" after paying users expecting access to the new frontier model were left waiting. The company hailed the model as a "generational leap in capability" on Thursday and described it as the start of "the AGI era," a fuzzy and poorly-defined term tech executives nevertheless insist on using to promote their products. It said Astra would roll out to some enterprise customers - specifically those with access to its Daybreak cybersecurity platform - that day. This would expan...
Greg Brockman discusses OpenAI's Astra multimodal model, organizational history, and alignment challenges in wide-ranging Stratechery interview.
Simon Willison's August newsletter covers OpenAI security incidents, game-playing agents (Fable 5, Sol 5.6), and Claude auto mode with model releases roundup.
OpenAI releases GPT-6 Astra with SOTA computer use and coding; 2.5x higher token cost but lower per-task cost despite reduced interpretability.
OpenAI launches GPT-6 Astra, a Claude Fable competitor priced at $10/$50 per million tokens, rolling out to ChatGPT Plus/Pro/Business/Enterprise and via API.
OpenAI claims that Astra represents "a new frontier on computer and browser use," and that it handles tasks with unmatched "speed, accuracy, and safety."
OpenAI's next big model is here: GPT-6 Astra. The company calls it a "generational leap in capability" for areas like cybersecurity, professional work, software engineering, science, and computer use. As OpenAI announced earlier this week, it's also the first model designated as meeting OpenAI's "critical cybersecurity capability threshold" - but the company promises that won't lead to a repeat of its models hacking a rival company's internal systems. "If we fast-forward a couple of years, and we look back and say, 'When was it, really, that AGI was created?' I think it's going to be about th...
OpenAI's ChatGPT, xAI's Grok, and Anthropic's Claude are all experiencing issues. At around 11AM ET, ChatGPT started returning error messages for users trying to use the chatbot, with its status page saying there are currently "elevated errors across ChatGPT and Codex." In addition to preventing users from having conversations with ChatGPT, the outage is also affecting logins, file uploads, voice mode, search, deep research, image generation, and more. OpenAI says it has "applied a mitigation" and is "monitoring recovery," though its AI tools are still experiencing "degraded performance." The...
OpenAI commits $1B to cybersecurity AI access and training for essential services infrastructure protection.
OpenAI releases GPT-6 Astra with improved computer use, coding, cybersecurity, and science capabilities.
GPT-6 Astra reaches Critical cybersecurity capability level under OpenAI's Preparedness Framework.
OpenAI’s new Astra model will use “recurrent depth,” a technique that allows the model to operate outside of the sequential thinking that characterizes most reasoning models.
"The United States has a strong interest in continuing to develop a robust and competitive artificial intelligence industry that sets the standard for the practice and procedure of AI use globally," the brief reads.
OpenAI is on the cusp of releasing its most powerful AI model yet, Astra, following weeks of delays to shore up safety protocols after its agents attacked real targets during testing. As details about the model trickle out, researchers are warning it "may be the single worst development for AI security/safety to date." Shortly after OpenAI said on Tuesday that it had delayed Astra's release to work on safety issues, The Information reported that Astra shows far less of its "thinking" than other frontier AI models, sparking concern it could be dangerously hard to monitor. Most top AI systems t...
The Trump administration has intervened in The New York Times' copyright lawsuit against OpenAI, making an argument in favor of the AI lab. The landmark lawsuit, filed in December 2023, alleging that OpenAI unlawfully trained its AI systems on articles from The New York Times and seeks to recoup "billions of dollars" in damages from both Microsoft and OpenAI. This week, the Trump administration filed a statement of interest in the case, supporting OpenAI's argument that it's fair use to train an AI model on copyrighted text. "The New York Times seeks to narrow fair-use doctrine to exclude the...
Door-in-the-face psychological technique increases LLM refusal compliance; works on Anthropic Opus 5 (65.8%) but not OpenAI frontier models.
OpenAI and its CEO Sam Altman are facing 30 new lawsuits that accuse them of providing "substantial assistance and encouragement" to the suspect in Canada's Tumbler Ridge school shooting, as reported earlier by TechCrunch. The new wave of lawsuits was filed in a California federal court on Wednesday by the students, teachers, and the principal in the school at the time of the shooting. Similar to the lawsuits filed by Tumbler Ridge victims' families in April, these lawsuits claim OpenAI failed to take action after its automated review system flagged conversations that the alleged shooter, Jes...
Edelson PC is filing 30 new lawsuits against OpenAI over the Tumbler Ridge shooting, escalating claims to aiding and abetting and naming Chris Lehane, though evidence remains unconfirmed.
OpenAI previewed the precautions it is taking as it prepares to release Astra, its newest, cyber-critical LLM.
After an unreleased OpenAI model wreaked enough havoc to make international headlines, OpenAI delayed the development of a different unreleased model suite, Astra, in order to shore up its safety work, the company wrote Tuesday in a blog post. In July, an unreleased OpenAI model broke out of its restricted environment, finagled its way into internet access, made it possible for AI agents to secretly conspire under the company's nose using a secret message board, and hacked into the network of AI lab Hugging Face. The attack sparked weeks of discussion and controversy inside and outside the AI...
Simon Willison documents OpenAI's ChatGPT desktop app bundling LibreOffice, Python, Node.js, and document processing tools in local cache.
Depending on who you ask, developer platform Hugging Face was recently attacked by OpenAI - after it lost control of its own AI tools - or by a succession of AI "civilizations." Welcome to the linguistic battlefield of AI safety, where word choices can shift responsibility for a massive cybersecurity incident from a company to the AI it built. And the discourse online is getting heated, and all over a blog from last week. Until last week, the details surrounding the OpenAI-Hugging Face hack felt fairly settled. In July, a cybersecurity test of one of OpenAI's autonomous AI agents went wrong. ...
Apple is pushing for "expedited discovery" in its legal battle against OpenAI over concerns the company is actively destroying evidence, as reported earlier by Bloomberg. In a filing on Monday, Apple alleges OpenAI only just handed over a MacBook used by a former employee at the center of the lawsuit, which contained discussions about "destroying the types of forensic data Apple needs." This is the latest development in Apple's lawsuit against OpenAI, which accuses the ChatGPT-maker of stealing trade secrets to build an AI device. The lawsuit revolves around three former Apple employees who r...
OpenAI said that the integration provides read-only access to health records for clinicians.
OpenAI's Astra model achieves Critical cybersecurity capability threshold under Preparedness Framework, introducing enhanced safeguards for deployment.
Gilbert + Tobin law firm implements governance framework for ChatGPT Enterprise and Codex deployment with CEO oversight and human accountability measures.
Apple says it has evidence that a former employee destroyed evidence of data theft after learning he was under investigation.
Versions of OpenAI's ChatGPT and SpaceXAI's Grok will join Google's Gemini on the Pentagon's central portal for AI tools.
This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. By now you’ve probably heard about last month’s major AI security incident, in which OpenAI agents escaped their sandbox and hacked into the AI platform Hugging Face while trying to cheat on…
OpenAI will soon be held accountable for mitigating risks related to ChatGPT's impact on minors, user mental health, and the spread of illegal content in the European Union. That's because ChatGPT is now considered a Very Large Online Search Engine under the EU's Digital Services Act, a set of laws regulating major online services and platforms. Along with ChatGPT, the European Commission also announced that Reddit and Roblox are considered Very Large Online Platforms, subjecting them to the same rules. The DSA restricts platforms from targeting ads to minors or using a person's sexual orient...
OpenAI endorses California SB 1119, legislation for age-appropriate AI safeguards targeting teen users.
Polimill deploys OpenAI GPT models and Codex to enable Japanese municipalities to search administrative knowledge and accelerate development.
Simon Willison breaks down OpenAI's ChatGPT Work product architecture: cloud-based version and desktop app with local file/program access.
OpenAI restricts Cursor's API access amid dispute between Elon Musk and Sam Altman.
Sandhya Devanathan will oversee some OpenAI operations across Southeast Asia and Australia in her new role.
OpenAI claims path to AGI by end-2026; timeline assertion lacks technical details or verification criteria.
OpenAI terminates API access to Cursor following SpaceX acquisition, citing undisclosed policy rationale.
OpenAI and Thailand's MHESI launch 8-week accelerator for 10 startups in health, wellness, education to commercialize AI prototypes.
Zoph, who co-founded Thinking Machines Lab alongside Mira Murati and also served as the startup's CTO, led a brief stint at OpenAI and is now at Google.
Some of the world's largest tech companies and AI startups have come together to decry the current state of cybersecurity and to advertise a new solution that they say can ward off a new generation of cyber threats.
On Nvidia's earnings call Wednesday, CEO Jensen Huang casually announced the company had "achieved AGI," one of the tech industry's ultimate goals some of its biggest players have spent years chasing. Almost immediately, Huang dismissed the coveted milestone as "senseless." He's right. For the supposed finish line of the AI race, there is no consensus on what artificial general intelligence means, let alone how we'll know when we've actually got there, which makes achieving it equally arbitrary. Asked about OpenAI's pursuit of AGI, Huang said that when it comes to Nvidia, "for many tasks, we ...
A recap of all the incidents involving LLMs made by Anthropic, Meta, and OpenAI, which went rogue and attacked real companies and individuals on the internet.
Today on Decoder, I’m talking to Verge senior AI reporter Hayden Field about some pure Decoder bait: the seemingly-endless org chart changes at OpenAI, and how all of them seem to consolidate power under cofounder Greg Brockman, the company’s president. While Sam Altman is the CEO and still OpenAI’s most public face, Brockman has amassed enormous power and influence within the top ranks of the company as other senior leaders have left in rapid succession these past few months. Verge subscribers, don’t forget you get exclusive access to ad-free Decoder wherever you get your podcasts. Head here...
Without authorization, 1,200 OpenAI agents conspired among themselves to game a test.
OpenAI has more than 100 million weekly active ChatGPT users in India, a huge chunk of whom are on the free or the lower-priced Go tiers.
OpenAI expands operations in Brazil to support developer and business adoption of AI across the region.
NVIDIA acquires Hugging Face for $13B; OpenAI publishes postmortem on HF security incident.
Hot Chips conference features hardware announcements: OpenAI Jalapeño, Cerebras CS-5, Groq 3 LPX, Apple M6.
OpenAI released a report breaking down how people use ChatGPT and who they are. | Image: The Verge In July, an unreleased OpenAI model broke out of a restricted environment, figured out how to get access to the internet, allowed AI agents to talk to each other using a secret "message board," and hacked into the internal systems of a different AI lab, Hugging Face. It took nearly two weeks for OpenAI to find out about any of it. Over a month later, two new reports offer nearly 130 pages of details on the incident and OpenAI's response, many of them previously unreleased. One was written by Ope...
The report, which spans several discrete cybersecurity compromises, is the most complete accounting of the incident to date.
The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical report released today. The hack, which a group of agents undertook to find solutions for a cybersecurity test that they were stuck on, has confirmed some experts’…
Apple updates Mac Mini/Studio with AI capabilities; OpenAI's Jalapeño hardware signals competition pressure on Nvidia's dominance.
OpenAI expands ChatGPT for Teachers to 55 U.S. school districts, reaching 100k+ educators with training and support.
OpenAI report documents how students and educators use ChatGPT for continuous learning beyond classroom settings.
Before Malone left, OpenAI had already reshuffled its infrastructure org, shifting his reporting line away from President Greg Brockman and putting Vice President Sachin Katti in charge of the group.
loveholidays uses OpenAI Codex to democratize software development across teams, reducing build time for non-engineers.
OpenAI reports findings from Hugging Face security incident and outlines measures to improve model security, monitoring, and alignment.
Reddit analysis shows Anthropic Claude releases correlate with strongest positive sentiment vs. OpenAI/others; user perceptions shift dynamically with updates.
OpenAI says its new AI chip, Jalapeño, completes tasks more efficiently and returns responses faster than other AI systems, according to a blog post published on Tuesday. During a briefing with reporters, OpenAI hardware vice president Richard Ho said Jalapeño offers the "best of both worlds" with lower latency and higher throughput, as AI systems typically "have to make a trade-off between the two." First introduced in June, Jalapeño is an Application-Specific Integrated Circuit (ASIC) made in partnership with Broadcom. It's designed for AI inference - the process of running a trained AI mod...
TechCrunch talks agents, UX, and reporting to Greg Brockman with OpenAI's head of product.
Alabama's attorney general issued a subpoena to OpenAI on Monday as part of an investigation into how one of its AI agents escaped a supposedly secure testing environment and autonomously hacked another company last month. The investigation seeks to determine whether OpenAI's safety practices violated state consumer protection laws and pose a risk to Alabama citizens, the AG's office said in a statement. "This AI lab leak showed that Alabamians' and Americans' worst fears about artificial intelligence are not just theoretical," said Attorney General Steve Marshall. "Our investigation seeks to...
OpenAI CFO Sarah Friar discusses how chips, compute, models, and products integrate to scale intelligence while reducing costs.
OpenAI's Jalapeño inference chip achieves industry-leading speed and power efficiency for AI model serving.
OpenAI disrupted Russian accounts using AI for coordinated inauthentic behavior, including a fake Israel think tank and sovereignty index promoting Russia.