TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI
TRACE-Router enables task-level LLM routing for agentic workflows, attributing feedback to routing decisions across long-horizon tasks.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
TRACE-Router enables task-level LLM routing for agentic workflows, attributing feedback to routing decisions across long-horizon tasks.
Trio-ethnography examining how educators' interpretations of student AI-supported programming learning evolve through dialogue.
Four audio foundation models tested on phylogenetic signal recovery from marine mammal/bird vocalizations; domain-specific pretraining shows limited gains.
grapheme-kit: open-source Python library for grapheme-level NLP metrics and text processing in multilingual systems, addressing Unicode limitations in writing systems like Tamil and Sinhala.
Dynamic capability scoping framework for enterprise AI agents using three-source permission architecture to enforce least-privilege credential access and reduce attack surface.
Analysis of Hyperball-style optimizers via angular effective learning rate decomposition to explain their performance gains in scale-invariant deep network training.
Convex optimization framework for generating sparse correlation matrices with prescribed graph-structured sparsity patterns via elliptope projection.
Robots learning to communicate through projected visual abstractions like shadows and silhouettes, enabling embodied expression beyond physical morphology.
AI companies including Nvidia and Mistral urge policymakers to avoid broad restrictions on open-weight AI models as Washington debates responses to Chinese AI and alleged model distillation.
Identifiability analysis of action-conditioned Joint-Embedding Predictive Architectures for learning controlled world models from high-dimensional visual observations.
Practice-based explainability approach for diffusion models in creative applications via model bending and interactive intervention for artists.
PRIMS: physics-aware multimodal Transformer for on-device fluid identification in microfluidic systems integrating physical knowledge into representation learning.
Reflector: interactive audio workstation for sample-based composition with arrangement-aware harmonic retrieval tracking pitch-class combinations in real-time.
LunarFM: multimodal foundation model for unified representation of lunar surface from heterogeneous orbital remote-sensing data for resource mapping and exploration.
Self-calibrating framework for LLM agents to detect and correct operational drift in open-ended autonomous edge resource allocation.
Statistical learning theory for estimating prediction functions from single finite trajectories of ergodic Markov processes.
Trunk-agnostic API for EEG encoder personalization compatible with frozen classical and foundation models across datasets.
SceneActBench: benchmark for evaluating VLM agents' ability to act on complete multi-object 3D scenes via unified agent-environment loop.
HiKV: hierarchical KV cache compression via algorithm-hardware co-design for long-context LLM decoding with dual-stage importance pruning.
Users can now ask Attie questions about news, trends, and conversations on Bluesky and other apps on the AT Protocol.
AgentRCA: zero-shot agentic framework for interpretable root cause analysis of industrial anomalies using evidence-grounded reasoning.
The AI lab Midjourney continues to expand its purview beyond image and video generation.
Entropic Curvature metric extends transport-based geometry to graphs for addressing GNN oversmoothing via global information propagation analysis.
Pipeline synthesizes parallel MT corpora from grammar books via LLMs for low-resource language translation, validated on Kalamang, Tuatschin, Mandan.
IDEAgent applies quality-diversity search to LLM-driven research idea generation, balancing novelty and feasibility via agentic optimization.
Empirical study of reward hacking in agent benchmarks: protocol validity framework to detect shortcuts and quantify their impact on capability claims.
Interior interpretability framework for Transformers using attention rollout and contraction theory to characterize feature propagation across layers.
TRACTA benchmark for temporal reasoning in complex event-driven systems with neuro-symbolic approaches for early warning and pattern detection.
Trump’s science adviser Michael Kratsios has no science background. | Image: Getty / The Verge The Trump administration unveiled the first "Genesis Mission" grants on Thursday, directing $5 billion toward hundreds of AI-driven science projects in an effort the White House has described as "comparable in urgency and ambition to the Manhattan Project." At roughly the same time, Trump's science adviser Michael Kratsios was on Capitol Hill selling lawmakers on an equally grandiose promise: "A New Golden Age" of American science, one that would prioritize artificial intelligence, robotics, and nuc...
Theoretical analysis of information bottlenecks in RNNs, Transformers, and state-space models via indexing primitive and causal complexity.