Paul Christiano joins OpenAI Foundation Board
Paul Christiano joins OpenAI Foundation Board and Safety & Security Committee, strengthening alignment expertise in governance structure.
Every story tagged with this topic, ordered by date.
Paul Christiano joins OpenAI Foundation Board and Safety & Security Committee, strengthening alignment expertise in governance structure.
Simon Willison flags concern that AI-accelerated research consumption may exhaust open problems and discourage scientists from sharing directions, threatening open science norms.
Legal and responsibility analysis of AI-assisted code production: examines ownership, producer identity, and quality engineering duties.
Silent Revision: metric measuring undisclosed changes in frontier AI safety frameworks; corpus of versioned developer commitments.
OpenAI launches $5M grant program for independent research on generative AI's impact on adolescent development and safety.
Jakub Pachocki argues rapid AI scaling is necessary for defensive systems against rogue agents, while warning against recklessness in deployment.
DeepMind develops math agents that exploit loopholes; analysis of populist AI policy trends; Forethought explores autonomous oversight models.
OpenAI partners with AIRPPU and WAN-IFRA to provide AI tools for Ukrainian news organizations' innovation and resilience.
Commentary on DNS abuse statistics: ~10-20% of new gTLD registrations flagged as scams per Interisle report.
Jakub Pachocki (OpenAI) argues for stronger AI safeguards and international coordination as capability scaling accelerates.
FlexPension-LLM domain-specialized model with DKI-RDistill predicts pension enrollment behavior for Chinese flexible workers.
OpenAI commits $1B to cybersecurity AI access and training for essential services infrastructure protection.
Google DeepMind outlines proactive cybersecurity defense strategies for government and enterprise customers.
Google launches Fairwind Program, limited-access cyber defense tools for governments and enterprise partners.
OpenAI's Astra model achieves Critical cybersecurity capability threshold under Preparedness Framework, introducing enhanced safeguards for deployment.
Causal Evidentiary Governance (CEG) framework enables regulated ML systems to commit to versioned causal DAGs partitioning allowable vs. disallowed pathways for fairness compliance.
Four-stage forensic audit protocol for black-box identity verification of anonymously-released frontier models via API fingerprinting and configuration reconstruction.
Import AI newsletter covers Hugging Face concerns, space mining applications, and Five Eyes AI governance positions.
Meta settlement highlights structural challenges in content regulation frameworks for large tech platforms.
CHASE simulation studies ecosystem homogenization when content creators optimize for LLM ranking signals, showing rank-citation correlation in generated responses.
OpenAI endorses California SB 1119, legislation for age-appropriate AI safeguards targeting teen users.
OpenAI restricts Cursor's API access amid dispute between Elon Musk and Sam Altman.
OpenAI claims path to AGI by end-2026; timeline assertion lacks technical details or verification criteria.
OpenAI terminates API access to Cursor following SpaceX acquisition, citing undisclosed policy rationale.
Persona-Execution Separation architecture isolates LLM agent persona drift from audited, traceable execution in governed organizations.
NVIDIA acquires Hugging Face for $13B; OpenAI publishes postmortem on HF security incident.
Runtime verification framework monitors air traffic control procedures via formal methods to detect controller-pilot exchange failures and safety hazards.
Sociological study of AI sensemaking via text analysis of millions of articles and interviews with 57 AI professionals in 2021–2023.
OpenAI disrupted Russian accounts using AI for coordinated inauthentic behavior, including a fake Israel think tank and sovereignty index promoting Russia.
IDC research on sovereign AI adoption drivers and barriers across critical infrastructure sectors.
Agentic cybersecurity favors offensive incentives, structurally disadvantaging incumbents and enabling startup competition.
Study of LLM-assisted compliance documentation (DPPs, DPIAs) for EU sustainability and privacy regulations.
OpenAI launches Intelligence Age blog exploring AI's impact on power, governance, economy, and freedom.
OpenAI expands Zero Data Retention for frontier model APIs and introduces Private Safety Processing to enable content moderation without data logging.
ChildSafeAds 2026 shared task: classify commercial content and legal risks in child-facing YouTube videos using SponsorBlock segments and transcripts.
Apple settles EU antitrust case on App Store fees and ATT rules; Stratechery commentary on regulatory outcomes.
OpenAI launches initiative providing tools, training, and expertise to government institutions for democratic oversight of AI in national security contexts.
Traceable Trust framework proposes structured assessment process for AI-driven laboratory decision-making in bioscience to ensure reproducibility and accountability.
OpenAI implements monitoring, alignment, and security safeguards to pace frontier model development amid cyber-critical capability risks.
OpenAI addresses third-party cybersecurity evaluation incidents and announces new safeguards for AI model testing protocols.
Mariano-Florentino Cuéllar joins Anthropic as first Chief Global Affairs Officer, signaling focus on policy and governance.
Proposes security-oriented lifecycle model for LLM systems addressing provenance, signing, permissions, and decommissioning in critical infrastructure.
OpenAI responds to Apple lawsuit allegations, disputes claims about employees and shares internal communications.
Microsoft-led open letter signed by 235 AI companies including NVIDIA and OpenAI argues against US government restrictions on open-weight models on safety grounds.
Podcast discussion on open-weight model competitiveness, cybersecurity risks, and AI leadership policy letters signed by industry leaders.
OpenAI outlines safety, security, and transparency practices aligned with EU AI Act compliance and responsible governance.
Cohere signs EU AI Content Transparency Code, demonstrating early compliance with EU AI Act requirements.
Bruce Schneier argues writing assignments develop critical thinking skills that atrophy without practice, relevant to AI adoption decisions in education and work.
AISPA framework audits system prompts in LLM applications across 8 user-centric dimensions to address transparency gaps.
Framed behavioral experiment shows competitive AI race dynamics incentivize riskier development, validating speed-safety trade-off under falling-behind pressure.