VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation
VAD counterfactual algorithm isolates visually attributable components in multimodal on-policy distillation corrections.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
VAD counterfactual algorithm isolates visually attributable components in multimodal on-policy distillation corrections.
MixFrag uses KL-divergence-guided mixed-precision quantization to deploy Vision Transformers efficiently on resource-constrained devices.
PAIChecker reveals 13.6% misalignment between PR-issue pairs in SWE-bench-like benchmarks, questioning LLM code-resolution eval validity.
β-OPSD generalizes on-policy self-distillation via tunable KL regularization to improve reasoning language models with better stability.
DualG-MRAG decouples macro reasoning and micro-matching in multimodal retrieval-augmented generation to handle complex multi-hop tasks.
Study shows repeated sampling outperforms self-refinement and reflexion methods across 1.5B-7B models at equivalent token budgets with confidence intervals.
Computational complexity analysis of committee election winner determination under Thiele voting rules with FPT algorithms.
Empirical study of inference-time scaling strategies for local computer-use agents under hardware constraints across multiple dimensions.
Frontis-MA1 (35B) meta-evolution agent trained for recursive self-improvement in ML engineering via OpenMLE system with program-evolution operators.
DR-FRL applies doubly robust estimation to irregular longitudinal functional data for causal inference with efficient influence function targeting.
APO applies unsupervised policy optimization to 3D atomic structure prediction, eliminating need for ground-truth coordinate alignment.
Apptronik’s Apollo 2 robot takes a baseball glove off of a shelf. | Image: Google Google DeepMind says the latest version of its Gemini Robotics AI model can "control entire humanoid robots." While the previous model focused on controlling a humanoid robot's upper body, Gemini Robotics 2 now supports "whole-body motions" ranging from its feet to fingertips, according to an announcement on Thursday. The new model will allow humanoid robots to perform a wider range of actions, as it allows them to walk, crouch, stretch, and manipulate objects. Videos shared by Google show how Apptronik's Apollo...
ORCA-bench evaluates LLM agents on production oncall root-cause analysis using real telemetry (Prometheus, Jaeger, OpenSearch) and code in realistic incident scenarios.
ScaFE uses LLM-generated feature programs to classify pathological scars (keloids vs. hypertrophic) from clinical photos with limited labeled data and local data governance.
Graph neural network magnetic force fields learn effective energy functionals for itinerant spin dynamics from electronic structure, analogous to ML interatomic potentials.
Study shows LLMs reproduce standard language ideologies, privileging Inner Circle English norms and marginalizing non-dominant World Englishes variants.
MANTA enables LLM-based multi-agent systems to self-adapt communication topology at inference time based on task and real-time collaboration performance.
DAR-Net addresses dual ambiguity (semantic and spatial) in all-in-one image restoration by disentangling degradation cues from scene content.
Proposes leakage-free evaluation protocol for same-graph cross-task transfer in GNNs across node classification and link prediction with fixed splits and negatives.
Introduces selective credibility-limited belief update framework extending Katsuno-Mendelzon model to handle partial compound epistemic inputs.
CS-RNR method enables agents in imperfect-information games to safely exploit flawed opponents with provable certificates on deployed strategy.
Multi-level framework models literary creativity as selective transformation across lexical, semantic, conceptual, structural, and narrative dimensions.
Sociolinguists analyze how generative AI impacts academic writing norms and linguistic diversity in global scholarly publishing.
Live benchmark measuring LLM resilience to state-backed information operations using 2,100+ real disinformation campaigns.
Target-conditioned abstraction method for scientific inspiration retrieval using analogical reasoning to transfer problem-solving principles.
Framework using causal reasoning to distinguish genuine individual improvement from strategic gaming in algorithmic recourse systems.
Structured extraction of event type, impact scope, and temporal horizon from financial news using LLaMA-3.1-70B outperforms sentiment-only prediction.
Diagnostic analysis of KV cache precision effects on token prefix reconstruction in Qwen2.5-derived language model inference.
Multi-agent reinforcement learning approach coupling assortment, sourcing, and routing decisions for end-to-end supply chain optimization.
Do you remember Friend? The Friend that launched an AI pendant, spent $1.8 million of its $2.5 million to acquire friend.com, and plastered the NYC subway with ads promoting artificial companionship? Yeah well, if you didn't remember, Friend is back. And now, it's twice the price. Introducing friend. pic.twitter.com/T5vUCj7iqv - friend (@friend) July 30, 2026 This morning, Friend launched a new ad showing two people talking to their Friend pendants about very personal life problems. It didn't say it explicitly, but the ad showcased a new feature the pendant now includes: a speaker. Before, Fr...