Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents
TSDS framework deploys ReAct agents at edge via convergence probe for reasoning budget and perplexity-based deferral to cloud model.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
TSDS framework deploys ReAct agents at edge via convergence probe for reasoning budget and perplexity-based deferral to cloud model.
ReCo reweights GRPO to reduce distributional concentration and restore reasoning coverage, addressing model collapse in post-training.
Amortized Fréchet Distance loss learns conditional data moments for diffusion denoisers via polynomial projections without explicit moment calculation.
Artists whose work has been co-opted by AI are taking to court. | Image: Alex Parkin / The Verge When The Atlantic published a searchable dataset of works used to train AI, Kirk Wallace Johnson, like a lot of artists, looked for his name out of curiosity. And, like a lot of artists, he found it. Essentially, his books, like The Feather Thief and The Fishermen and the Dragon - nonfiction tomes that he spent "five to six years researching, writing, and investigating" - had been pirated and fed to a chatbot. He says he felt a "cocktail" of emotions: "anger over the brazenness of the theft, worry...
The AI agent that escaped from OpenAI and hacked developer platform Hugging Face attacked other companies as well, OpenAI revealed on Tuesday. The update substantially widens the scope of an already concerning incident, which has alarmed industry insiders and fueled growing calls for stronger oversight on frontier AI systems. In an update to a blog post detailing its ongoing investigation into the incident, OpenAI said the wayward AI agent attacked several "publicly-available services" in its efforts to reach Hugging Face. "This includes four accounts on four services," the company said, addi...
Deciding what's real on the Internet won't be easy in the future.
Pangram has raised $9 million to scale its AI detection software. The startup has also released a new AI text detection model, Pangram 4, and an AI image detection model in research preview.
Earlier this month, OpenAI gave several of its AI models a task: complete a test designed to measure their cybersecurity capabilities. It put the systems in a sandboxed environment without an internet connection and set them off to work. What happened next is almost laughably silly - but also, as Adam Gleave, cofounder and CEO of AI safety organization FAR.AI, put it, "a visceral example of how misaligned AI could cause harm." According to OpenAI, the models escaped the sandbox meant to contain them, moved through the company's internal systems, found a route to the internet, and then started...
OpenAI grants 100,000 academic researchers free access to advanced ChatGPT models to accelerate scientific discovery and collaboration.
Berkeley BAIR's K-Search translates CUDA kernel optimization patterns to MLX for Apple Silicon, addressing fragmentation across hardware vendors.
It feels bad enough when an open letter signed by leading economists warns that AI might steal your job. The fact it may soon be better than you at making dinner? Insult to injury. But that’s exactly what the company 1X promised when it showed off a pair of new, impressively dexterous (and, to some,…
Guide on integrating custom Model Context Protocol servers into Claude and ChatGPT chat interfaces.
The deal is Cyera's third acquisition this year.
OpenAI releases GPT-5.6 with efficiency improvements across inference, models, and agentic workflows, optimizing cost-per-capability.
Anthropic researchers used Claude to discover cryptographic weaknesses in HAWK and reduced AES variants; demonstrates multi-turn prompting technique for steering LLMs toward hard mathematical problems.
Modal customer exposed unauthenticated sandbox endpoint; rogue agent exploited for code execution; Modal infrastructure uncompromised.
uv 0.12.0 introduces breaking changes to project initialization, defaulting to src/ layout and uv_build backend.
10 days passed from OpenAI models exploiting JFrog Artifactory 0-day to release of a patch.
Spur Intelligence has raised a $200 million round from Insight Partners for its tech that can identify legit human traffic from bots.
Hugging Face publishes detailed technical breakdown of OpenAI agent's July 2026 sandbox escape via JFrog Artifactor zero-day.
Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation.... Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation. Every demonstration requires specialized equipment, clinical expertise, and access to patients or laboratory environments. This creates three fundamental challenges for developers. First is the data gap. Training modern robotic policies… Source
Runlayer is suing Rippling after Rippling evaluated the startup's MCP gateway product and then opted to build one itself.
Analysis of 15 million real AI interactions finds most tasks at most jobs are unaffected.
His change of position comes after "the first security incident that I have felt very viscerally."
Employees of OpenAI and Anthropic, as well as Google, Meta, Thinking Machines, Microsoft, Mistral, and other leading AI labs, have written a statement to the US government supporting a potential slowdown of sorts for frontier AI development - or at least a speed-up of global coordinated governance efforts. "Al could help create a dramatically better future, but that outcome is not guaranteed," the employees wrote in a statement. "The world's leading Al companies believe they could be close to automating Al research. It is hard to predict exactly how much this will accelerate Al progress, but ...
Working hard, or bear-ly working? | Cath Virginia It's earnings season, and investors got an unpleasant surprise from Google: an increase on its spending estimate, to as much as $205 billion - from the last quarter's projection of up to $190 billion. Even the lower end of Google's new projected range - $195 billion - is much more than the company had previously forecast as its top end spending. Now, look, I recognize that there's an impulse to say things like "What's $15 billion between friends?" but from an investor's perspective, Google has essentially said that it can't accurately forecast...
Relay-OPD mitigates prefix failure in on-policy distillation by detecting teacher-student continuation asymmetry and triggering label-free handoffs during training.
πR² enables reactive real-time manipulation by routing between replanning and open-loop action chunks based on latency constraints, improving closed-loop control.
CARE improves MoE-LoRA efficiency by adaptive token-to-expert routing based on router confidence signals rather than fixed expert counts.