OpenAI says Hugging Face was breached by its own pre-release models
OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.
Buzz is a group chat platform for the workplace that puts humans and their AI agents in the same conversation.
Over the past decade, streaming platforms competed by dominating individual formats like music, video, podcasts, or audiobooks. Now, as AI makes it easier to create, organize, and recommend content, those distinctions are fading, pushing companies like Spotify, Netflix, YouTube, and TikTok to become all-purpose entertainment destinations instead.
Xaira Therapeutics builds causal models for drug discovery using synthetic data generation; Bo Wang and Ci Chu discuss data requirements for model training.
Substack will now help users determine whether what they're reading may have been written by AI. A new tool coming to the platform can scan posts, notes, replies, and comments to provide an estimate of how much text could be AI-generated or written with AI assistance, according to a blog post published on Tuesday. The tool is powered by an AI detection company called Pangram and is rolling out across the web and the iOS app, with an Android launch coming "soon." Readers can analyze content longer than 100 words by choosing the "Scan for AI text" option from the three-dot menu in the top-right...
New data centers built through 2033 could consume as much electricity as India uses today.
Long-context LLMs fail via repetitive copying rather than reasoning; RL-based evidence grounding improves step-by-step trace quality.
Appearance Pointers enable spatial region control in Diffusion Transformers via compact tokens, improving controllable image generation.
CodeRescue optimizes cost-aware routing for coding agents, determining when to retry vs. escalate after execution failures.
Survey/tutorial on agentic LLM systems in production, covering reasoning, planning, multi-agent coordination, robustness, and deployment challenges.
1-Lipschitz neural networks on Hadamard manifolds using Busemann functions for robustness in non-Euclidean geometry.
Theoretical analysis of K-class classification via O(log K) binary hyperplane classifiers in distributed settings under Gaussian assumptions.
PDDIM algorithm provides provable posterior sampling for linear inverse problems using diffusion priors with lightweight modifications.
ROMS-IMLE questions necessity of gradual noise-to-data transformation, proposing single-step generative modeling competitive with diffusion.
ISO optimization framework operationalizes spectral inheritance in RLVR, enabling weight-space updates via singular structure analysis.
Deep CNN model of visual valence processing inspired by Rescorla-Wagner associative learning, applied to natural scene classification.
MaLoRA and MaLOR introduce dynamic, recurrent state-space adapters for task and token-level LLM adaptation, improving reasoning via selective modulation.
GAMUT benchmark introduces two-level meta-rubrics to measure factual completeness in long-form LLM generation, addressing precision-recall gap in factuality eval.
ResearchArena framework evaluates AI control and monitoring for detecting sabotage in automated AI R&D agents across safety/capability post-training and optimization tasks.
CircuitKIT open-source library unifies circuit discovery, evaluation, and intervention workflows for mechanistic interpretability with automated contrastive prompts.
Anthropic blocks authors from opting out of $1.5B settlement at last minute.
Off-Context GRPO uses privileged training information (solution prefixes) to enable RL with verifiable rewards on hard reasoning problems avoiding zero-gradient plateaus.
Staypoint detection benchmark provides ground-truth annotations for semantic trajectory analysis from noisy GPS data, addressing lack of standardized evaluation.
Real-time SDF-based mapping and motion planning for autonomous UAVs unifies occupancy mapping and trajectory optimization using signed distance representations.
Framework generalizes batch normalization and neural modules to Riemannian manifolds, Lie groups, and gyrogroups for manifold-valued deep learning.
SHRED-ROM reduced-order modeling with shallow recurrent decoders enables real-time optimal control synthesis for high-dimensional dynamical systems.
Study models how imperfect LLM detectors distort user incentives and downstream metrics, showing counterintuitive effects of detection as behavioral intervention.
LangGraph practitioner guide with three executable recipes for stateful multi-step agent workflows in business processes.
Safety analysis showing distributed, normalized failures in deployed AI systems are harder to instrument than obvious errors.