Moral Attitudes of Sentient ASI towards Humanity and Implications for AGI Development
Philosophical paper proposes inverting AI ethics to examine how future sentient ASI may morally evaluate humanity, suggesting post-human moral principles.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
Philosophical paper proposes inverting AI ethics to examine how future sentient ASI may morally evaluate humanity, suggesting post-human moral principles.
Semantic-aware contrastive learning framework for 3D medical imaging addresses false negatives by recognizing high-level semantic similarity in negative samples.
OmniaBench: unified benchmark evaluating LLM-based agents across diverse scenarios with explicit state spaces for systematic capability characterization.
Demographically-conditioned Stable Diffusion 2.1 synthetic images mitigate bias in medical classifiers and detect per-subgroup performance disparities.
Lila Sciences argues scientific labs as data sources for AI training, positioning robotics and experimental workflows as frontier training data beyond internet corpora.
CFM-Bench: unified multi-domain benchmark for channel foundation models enabling fair comparison across wireless tasks and pretraining approaches.
Linus Torvalds states Linux will adopt AI tools; rejects anti-AI stance as maintainer policy.
GradientSHAP with implicit differentiation and LLM narrative generation explains industrial process optimization recommendations to operators.
Geometric trajectory discrimination detects AI-generated text by modeling latent representation evolution across sequences rather than static embeddings.
Multi-axis max@K reinforcement learning improves target-mode coverage and demographic diversity in diffusion-based text-to-image generation.
LLM-based sequential detector identifies online firestorms by contextual chunk assessment on Reddit threads, outperforming volume/sentiment baselines.
LongStraw: GPU-efficient execution stack for million-token RL post-training with GRPO, bridging inference context length vs. post-training gap for agents.
1Password has launched a new browser integration for Claude that allows the Anthropic chatbot to access stored security credentials like usernames and passwords. The 1Password for Claude feature means that users can authorize Claude to complete multi-step tasks like booking travel and managing online accounts on their behalf without having to manually input their login credentials, but without actually exposing your security information to Anthropic's AI models, according to 1Password. That's made possible by a new "zero-exposure security framework" developed by 1Password, which works by inje...
Closed-form optimal mixing coefficient for self-distillation in rectified flow guarantees strict improvement over suboptimal teacher velocity fields.
Mechanistic interpretability via contrastive activation directions steers World Action Models toward robustness under distribution shift.
Causal inference method for sequential observational data with outcome interference and latent confounders using low-rank factor models.
DynaBase: minimal two-parameter interpretable architecture for zero-shot dynamical systems reconstruction achieving comparable performance to DynaMix.
Study evaluating whether synthetic face datasets can replace real benchmarks for face recognition evaluation across 12 synthetic vs 7 real datasets.
Random Logit Scaling defense against black-box score-based adversarial attacks on deep neural networks with query-efficient robustness.
Graph neural network approach for LLM authorship attribution using reasoning structures rather than surface-level linguistic features.
FlashDecoder: Transformer-based streaming video decoder with fixed temporal window enabling constant-latency real-time generation at high resolutions.
StructureClaw: artifact-centered benchmark for evaluating LLM agents on complete structural engineering workflows with verifiable evidence chains.
Method using instruction tuning and model merging to adapt reasoning language models to unverifiable domains with existing human-written solutions.
Analysis of domain mismatch effects in plug-and-play proximal gradient descent image reconstruction when denoisers are deployed outside training domains.
Google must give rival AI assistants and search engines greater access to key parts of Android and Google Search after the European Union ordered the company to comply with the bloc's digital antitrust rules. The two decisions, handed down Thursday, could weaken Google's control over two of the tech industry's most important platforms and have far-reaching consequences for the company, shape the future of its AI tool Gemini, and open up new opportunities for rivals to gain ground. Google has until January 2027 to begin sharing search data and July 2027 to implement changes to Android. The rul...
Proof-or-Stop: lifecycle control framework for autonomous coding agents using mechanically verifiable evidence gates for transitions between states.
I stood before a hulking glass and brick structure in the heart of Fort Worth, Texas. Thousands gathered inside to see what had been billed as "the future of policing in the digital age." As press, I was prohibited from entering, but from a number of nearby locations, I met with attendees who told me what was being sold within. And I learned that AI is threatening to seize the very heart of policing in America. The promise of AI at this year's International Association of Chiefs of Police (IACP) Technology Conference focused on automating routine parts of the job, which also happen to be crit...
Google DeepMind and Isomorphic Labs outline joint approach to applying AI models for biological resilience research.
OpenAI documents internal use of Codex for custom tool development, prototyping, and creative ideation workflows.