Cohere Deepens Partnership with Government of Canada
Cohere partners with Government of Canada to develop sovereign AI capabilities for public-sector services.
Every story matching this topic across titles and summaries, newest first.
Cohere partners with Government of Canada to develop sovereign AI capabilities for public-sector services.
Cohere releases W4A8 quantization with vLLM integration; compares W4A16 and W8A8 schemes on NVIDIA Hopper.
Spatial normalization technique for cross-domain retinal OCT layer segmentation in clinical neurodegenerative disease analysis.
Cohere partners with University of Toronto on multi-year AI adoption and responsibility initiative.
Proposes structure-aware dance generation via atomic movements to improve choreographic coherence and controllability.
Cohere outlines enterprise AI total cost of ownership framework covering token pricing, model selection, private deployment, and vendor infrastructure.
ESFP benchmark measures whether LLMs distinguish and coherently shift between neutral attribution and self-stance epistemic registers in contested claims.
Cohere releases Tiny Aya Expedition, a multilingual model supporting 70+ languages for on-device and educational AI applications.
Cohere introduces Dynamic Speculative Decoding (DSD) that optimizes K parameter selection based on hardware constraints to improve inference efficiency.
Cohere guide on multi-agent system architectures, patterns, enterprise use cases, and adoption challenges.
Hierarchical Acoustic-Semantic Modeling addresses modality interference in full-duplex Spoken Language Models via separation and semantic coherence techniques.
Cohere releases open-source Arabic speech recognition model for enterprise transcription across Arabic dialect variants.
Quantum algorithm study on stabilizer state testing with limited coherent memory; theoretical contribution tangential to LLM frontier.
CheckRLM detects and corrects factual inconsistencies in reasoning chains via retrieval-augmented claim checking during LLM inference.
LuckyStar 111B hybrid reasoning model from Cohere and LG CNS enables efficient multilingual tool-using agents with Korean-English support.
LeVo 2 combines LLM and diffusion models to generate full-length songs with coherent vocals, accompaniment, and lyric adherence via hierarchical track modeling.
Cohere publishes practical guide on building AI agents for enterprise automation, covering reliability, security, and deployment patterns.
Learn how the Cohere team uses North, Wiz, and a custom MCP server to automate incident response workflows with AI.
This theoretical note studies the finite axiomatizability of strict majority reasoning in finite social decision frames. Moss and Pedersen (2026) introduce a coherence criterion that characterizes exactly when qualitative majority judgments are representable by a finitely additive measure. The question addressed here is whether that coherence criterion can be replaced, in the finite setting, by any bounded finite fragment. We prove that it cannot. For every $k\ge 1$, we construct a maximal standard frame whose shortest coherence violation has length exactly $2k+2$. Hence there is no uniform f...
Cohere partners with Aston Martin Aramco Formula One™ as official Generative AI provider, bringing enterprise AI solutions to enhance performance and innovation across the racing team. Starting Australian GP 2026.
In open-ended generation, LLMs frequently fall into the "likelihood trap", marked by repetitive degeneration and vocabulary dullness, creating a discrepancy between machine-generated and human-written text. While post-hoc tail truncation (e.g., Top-$p$, Min-$p$) avoids sampling from the unreliable tail, it can over-sample from the uncalibrated head and misalign generation with human lexical preferences; fixed scalar repetition penalties likewise ignore variation in logit scale across inference steps, potentially disrupting semantic coherence. To address both limitations, we propose Variance-C...
Remote sensing is increasingly relied upon to deliver actionable science for forest and wildfire risk management across large landscapes. Wall-to-wall, annually updated maps are a persistent need for effective forest management. Many planning systems and data collections combine disparate data sources with different purposes, vintages, and prediction quality, which leads to confounding behavior in operational planning systems. We introduce the VibrantForests framework, developed and applied to map forest attributes and provide a coherent foundation for effective forest and wildfire planning. ...
Enhancing the formal math reasoning capabilities of Large Language Models (LLMs) has become a key focus in both mathematical and computer science communities in recent years. While significant progress has been made in using state-of-the-art Auto-Regressive (AR) LLMs for formal theorem proving, these models suffer from inherent limitations. Their next-token prediction generation methods may yield suboptimal performance due to the challenges of long-range coherence and the compounding of errors over long sequences. Recent advancements in diffusion LLMs (dLLMs), which generate text through iter...
The knowledge encoded in large language models (LLMs) can serve as a substrate for structured reasoning over variables describing a complex world, but accessing this knowledge in a probabilistically coherent manner poses a difficult inference problem. We propose Large Language Gibbs, a scheme for structured probabilistic inference that uses conditional distributions of an LLM as transition operators. Rather than sampling structured objects through single-pass autoregressive generation, we iteratively resample individual variables conditioned on others using an LLM's next-token conditionals. T...
Score- and flow-matching models often rely on preference-based reinforcement learning for two purposes: aligning with subjective preferences and, surprisingly, recovering properties such as visual realism and coherent object structure that matching-based training is intended to learn from the data itself. We argue that this reflects a structural mismatch. Matching losses measure $\ell_2$ regression error on the velocity or score field under training-time marginals, a proxy poorly aligned with the visual and semantic properties that determine sample quality at inference. Given a reward aligned...
Game generation is an emerging application of coding agents, requiring models to transform natural-language specifications into playable interactive systems. Unlike traditional coding tasks, game generation takes place within a game engine, where scripts, scenes, assets, rendering, and runtime interactions must jointly produce coherent gameplay. We formalize end-to-end game generation as the problem of producing a complete game artifact that realizes a specification through observable player-game interaction in a target environment. We argue that evaluating this setting requires three desider...
Large language models (LLMs) are often hypothesized to perform implicit Bayesian inference, yet a key coherence condition, the martingale property of predictive beliefs, has been shown to fail in controlled synthetic in-context learning settings. We revisit this question in a more typical usage regime: generic multiple-choice question answering. Exploiting the discrete answer space, we compute exact predictive distributions and study belief dynamics induced by autoregressive answer resampling. We introduce prompted predictive resampling (PPR), where an LLM generates a sequence of answers to t...
The new office places Cohere at the centre of London’s AI growth story with a growing global research hub based out of the city.
Narrative question answering (NQA) is a challenging task in natural language processing that requires models to understand long textual contexts, capture relationships across events, and generate coherent responses. Despite recent advances in pretrained language models, most existing approaches rely on a single decoding output during inference, making them sensitive to generation variability and often resulting in incomplete or inconsistent answers .To address this limitation, we propose a self-ensemble Self-Consistency-Based reranking framework for narrative question answering. The proposed ...
North enables enterprises that prioritize data security to deploy AI agents and automations at scale within their own infrastructure.
We study generative modeling of Bach-style symbolic piano music using a shared MIDI corpus and three model families: autoregressive LSTMs with attention, latent-variable models including recurrent VAEs and vector-quantized VAEs, and generative adversarial networks. We compare their ability to model polyphonic note sequences, learn useful latent representations, and generate stylistically coherent compositions. Our experiments show that the autoregressive LSTM with attention produces the most musically coherent samples, while vector quantization helps mitigate posterior collapse and yields mor...
Recent advances in large language models (LLMs) have prompted claims that such systems exhibit agency or qualify as moral agents. This paper argues that these attributions are misguided. We maintain that moral responsibility requires commitment-bearing agency grounded in intrinsic intentionality and self-attributed action, and that such agency constitutes the form of free will relevant to responsibility. Although LLMs generate coherent and normatively evaluable outputs, their operation is fully characterized by probabilistic input-output mappings learned from data. Their apparent intentionali...
Existing approaches for multimodal variational autoencoders (VAEs) face a trade-off between generative quality and coherence-i.e., they struggle to generate realistic and diverse samples that, at the same time, are semantically consistent across modalities. A recent work shows that using a simple approximation to Hölder pooling as an aggregation method improves coherence over the SOTA MMVAE+, despite assuming a single shared representation across all modalities. Yet, it slightly compromises sample diversity. Inspired by this insight, we propose Hölder++, a novel multimodal VAE that improves t...
Computational creativity in Interactive Fiction faces a fundamental tension: Large Language Models (LLM) may produce creative narratives but struggle with world coherence, while symbolic systems ensure consistency but lack creative flexibility. We present IVIE (Incremental & Validated Interactive Experiences), a neuro-symbolic approach to generating complete and playable interactive fiction worlds from scratch. Building upon PAYADOR's neuro-symbolic framework, IVIE implements a four-stage incremental generation pipeline that delegates creative decisions--setting and character creation, puzzle...
Reinforcement Learning with Verifiable Rewards (RLVR) is a central technique for improving long-horizon reasoning in Large Language Models (LLMs). However, existing RLVR methods often encourage unnecessarily long reasoning rollouts, which can degrade reasoning coherence and exhaust the available context budget. Existing approaches to long-context organization often depend on external mechanisms to organize rollouts, rather than enabling the model to manage its own reasoning trajectory. To address this limitation, we propose ReSum, a novel RLVR framework that enables LLMs to compress and organ...
Current Large Reasoning Models (LRMs) exhibit remarkable general capabilities but significantly underperform in spatial reasoning tasks. Existing approaches treat this gap as a knowledge deficit, relying on supervised fine-tuning (SFT) to ingest labeled spatial data from external vision sources or synthetic engines. In contrast, we argue that for many tasks, spatial reasoning capabilities are already present in pre-trained LRMs but require alignment through logical coherence under geometric 2D and 3D constraints. In this work, we propose a self-supervised reinforcement learning (RL) framework...
Introducing North Mini Code: Cohere's first open-source agentic coding model. Built for sovereign developers, this efficient 30B MoE model delivers strong software development performance with minimal hardware requirements.
Generating coherent and controllable long-form content remains a persistent challenge for Large Language Models (LLMs). While reasoning-enhanced models have demonstrated success in logic-intensive domains, our evaluation reveals that they suffer from a severe length collapse in open-ended writing, where performance degrades sharply as target lengths exceed 2,000 words. We attribute this failure to the limitation of static hierarchical planning, which struggles to provide dynamic guidance over extended contexts. To bridge this gap, we introduce the Interleaved Structural Chain-of-Thought (IS-C...
Whisper, a widely adopted ASR model, is known to suffer from hallucinations - coherent transcriptions generated for non-speech audio entirely disconnected from the input. We investigate whether hallucinations can be detected and mitigated through Whisper's internal representations. We extract audio encoder activations and evaluate two representation spaces: raw Whisper activations and Sparse AutoEncoder (SAE) latents. We show that both spaces encode linearly separable hallucination-related information, with discriminative power concentrated in a sparse feature subset and increasing toward dee...
Indoor scene generation is crucial for robot simulation and modern interior design. However, complex layouts together with scarce 3D scene data make learning-based generation challenging. Existing methods often rely on hand-crafted rules or focus on isolated sub-tasks (e.g., floorplan synthesis or single-room furnishing), producing whole-home scenes that lack global coherence, realism, and simulation readiness. To mitigate these limitations, we propose a unified hierarchical framework that decomposes indoor scene synthesis into controllable stages. First, we curate a large-scale dataset of 30...
A blog about how Cohere Labs built coplot, a data visualization tool that not only helps their releases, but also their research process.
Vision language models (VLMs) excel at many tasks but still struggle with spatial reasoning when critical information is not directly observable. Many such problems require imaginative perception: inferring what would be seen from an unseen viewpoint, tracing paths through occluded spaces, or integrating partial observations into a coherent spatial representation. We introduce Imaginative Perception Tokens (IPT), intermediate perceptual representations that externalize what a VLM would perceive under alternative spatial configurations while remaining consistent with the observed input. To stu...
Conversational AI agents require memory systems that are both scalable and semantically coherent across long interaction horizons. Existing approaches rely predominantly on large language model (LLM)-based summarisation at write time, which introduces non-determinism, escalating token costs, and opacity in pruning decisions. We present the Deterministic Memory Framework (DMF), a CPU-first approach that replaces generative memory compression with a fully deterministic pipeline grounded in classical NLP analysis, vector geometry, and mathematical scoring. DMF assigns each conversational interac...
Effective "all-team" summarization in high-complexity settings like the Neonatal Intensive Care Unit (NICU) requires aggregating insights from diverse disciplines (physicians, nurses, therapists) spread across hundreds of clinical free-text notes. Simply pooling heterogeneous text often leads to incoherent outputs. Structured summarization therefore first requires accurate categorization of sentence-level provenance across multi-source notes. This pilot study introduces a clinical provenance categorization pipeline using supervised fine-tuning (SFT) of large language models (LLMs). We adapted...
Order-agnostic language models (OALMs), including discrete diffusion language models (dLLMs), are trained to predict masked tokens under arbitrary conditioning sets, allowing sequences to be generated or scored under arbitrary reveal orders at inference time. In LLaDA-2.1, we report three findings. First, the learned conditionals are not exact factorizations of a coherent joint distribution: changing only the reveal order shifts target log-likelihood by up to 0.49 nats/token, so likelihood alone mixes content difficulty with path-dependent artifacts. Second, although confidence-first (CF) dec...
Unsupervised skill discovery (USD) aims to learn diverse behaviors without reward functions, but often results in task-irrelevant or hazardous behaviors due to uniform exploration. Guided skill discovery (GSD) addresses this issue by incorporating human intent to focus exploration on meaningful regions. However, existing GSD methods typically require training additional guidance models, and rely on pre-defined rules or expert demonstration, which can be ineffective under sparse, online-collected human feedback. To overcome this, we propose COLLIE, a GSD framework that leverages dense unsuperv...
Multi-component LLM agents assemble probabilistic claims from components that each see only part of a joint problem; the composition can violate basic probability axioms even when every component is locally coherent. We formalise this locally coherent, globally incoherent failure via the compositional residual eps*, the L2 distance from the composed quote to the joint coherent polytope, computable at runtime from system output and the declared cross-component coupling constraints. A product-structure dichotomy characterises when local coherence suffices, and a Rayleigh-quotient prediction mat...
Delivers advanced reasoning with a minimal compute footprint. Command A+ offers full data sovereignty for governments and regulated industries worldwide.
A specialized translation model leverages RWS’ global language and cultural expertise and Cohere’s Command A+ model to power the new Language Weaver Pro.
Cohere and Mila announced plans for a new academic research collaboration focused on improving AI evaluation across languages and cultures, starting with French-language cultural context in Quebec.
Recent generative models have largely closed the gap on low-level artifacts - pixel fingerprints, frequency anomalies, upsampling traces - particularly in person-centric and partial-edit settings where the manipulated region is small and surrounded by photometrically authentic content. We introduce Social Gaze Consistency, a high-level semantic cue defined as the mutual coherence of gaze direction, head-eye alignment, and pupil placement between interacting individuals, and show that it constitutes a previously underutilized detection axis orthogonal to existing low-level paradigms. We instan...
Browse the new Cohere Merch Store and view the latest collection for sale, along with an archive of our merch and swag history.
Developer fine-tuned Cohere Transcribe to add diarization and timestamp support, extending open-source speech-to-text capabilities.
Cohere launches Command A+, first open-weights MoE model emphasizing efficiency and latency over peak performance.
Cohere releases Command-A-Plus-05-2026 bfloat16 model weights on Hugging Face Hub.
Cohere signs MOUs with Indra Group and Multiverse Computing to advance AI deployment with focus on sovereignty, security, and accessibility.
Cohere releases Command A+, an open-source model optimized for enterprise agent deployment with improved speed and capability.
Cohere acquires Reliant AI to strengthen sovereign AI capabilities for regulated healthcare and life sciences sectors.
PDI-Bench: Quantitative framework for auditing geometric coherence in generated video via perspective distortion and point-tracking metrics.
Value-filtered decoding selectively applies safety steering at test-time, avoiding unnecessary interventions that degrade helpfulness and coherence.
Cohere argues private AI deployment offers banks control, security, and customization advantages.
Cohere positions LLMs as solutions for financial services knowledge work and customer service automation.
Cohere advises financial firms on scaling generative AI from pilot programs to production with risk management.
Cohere discusses metrics and integration strategies for achieving AI-native enterprise operations.
Cohere showcases AI agents for financial services compliance, efficiency, and customer trust.
Cohere outlines governance frameworks for responsible AI development and deployment.
Proposes segment-level supervision for LLM-based Lean 4 theorem proving, balancing dense local signals of step-level training with coherence of whole-proof generation.
Google's leaked video model 'Omni' shows improved text coherence in generated video content.
Cohere publishes enterprise AI governance guide covering monitoring, responsibility frameworks, and innovation-risk balance.
NVIDIA GB200 NVL72 introduces a fundamentally new way to build GPU clusters by extending NVIDIA NVLink coherence across an entire rack. This design enables... NVIDIA GB200 NVL72 introduces a fundamentally new way to build GPU clusters by extending NVIDIA NVLink coherence across an entire rack. This design enables exascale performance, but it also changes the assumptions that many scheduling systems were built on. As a result, “rack-scale locality” becomes a hard constraint. When workloads cross domain boundaries, performance drops sharply… Source
Cohere and Aleph Alpha merge to form AI company positioned on sovereign/on-premise deployment.
Cohere outlines five-phase enterprise AI maturity framework identifying common production blockers.
Learning to Defer framework extended to hierarchical multi-label medical imaging with coherence constraints preventing taxonomic contradictions.
Motivated by sensing modalities in modern autonomous systems that involve hardware-constrained spatial sampling over large arrays with limited coherence time, we develop a novel framework for rapid super-resolution multi-signal direction-of-arrival (DoA) estimation based on Hankel-structured sensing and data matrix decomposition of arbitrary rank, under both the $L_2$ and $L_1$-norm formulation. The resulting $L_2$-norm estimator is shown to be maximum-likelihood optimal in white Gaussian noise. The $L_1$-norm estimator is shown to be maximum-likelihood optimal in independent, identically dis...
AI Council framework mitigates artificial consensus in multi-LLM policy simulation via architectural heterogeneity and coherence validation across value perspectives.
Differentiable physics-informed approach for phase retrieval in coherent transition radiation spectroscopy diagnostics.
Dependency-driven multi-stage prompt pipeline for coherent RPG world and narrative generation using LLMs.
Cohere publishes guidance on enterprise AI security: deployment best practices, vulnerability classes, and secure configuration for production systems.
Talker-T2AV decouples semantic and low-level modeling in autoregressive audio-video generation for improved talking head synthesis coherence.
Canadian AI startup Cohere is taking over Germany-based Aleph Alpha with support from Lidl’s owner, Schwarz Group. With the blessing of their governments, the companies intend to offer a sovereign alternative to enterprises in an AI landscape dominated by American players.
Cohere preparing new MoE model with vLLM optimization support, signaling open-weights or community-accessible release.
Gated context projectors improve cross-stage coherence in autonomous driving VQA by reducing perception-planning inconsistencies by 42.6%.
User reports Image Gen 2 demonstrates advanced reasoning in color composition, generating cinematic palettes and tonal coherence across panels.
Cohere explains MoE models' efficiency gains with speculative decoding via expert routing correlation and bandwidth optimization.
Ensemble deploys Cohere's custom LLM to add agentic AI automation to healthcare RCM platform.
Cohere explains enterprise AI search capabilities for data integration and workflow automation.
Cohere discusses emerging multimodal AI search capabilities for information discovery.
Cohere releases Rerank 3 model with improved search precision and enterprise retrieval performance.
Cohere provides FAQ guide addressing common enterprise RAG search and retrieval questions.
Cohere explores how rerankers and embeddings improve search and retrieval for AI applications.
Notion integrates Cohere Rerank to improve workspace search and retrieval accuracy.
Cohere publishes practical guide for implementing agentic RAG systems in enterprise settings.
Cohere educational overview of AI infrastructure components: hardware, software, networking layers for AI stack construction.
Cohere marketing guide covering generative AI use cases and integration strategies for campaigns.
Cohere Labs releases Tiny Aya, open-weight multilingual model optimized for on-device inference across 200+ languages.
Cohere publishes enterprise deployment guidance covering cost, security, and scaling challenges for AI model operations.
Cohere launches Transcribe speech-to-text API with accuracy/speed claims for audio data search and automation.
Cohere conceptual overview defining enterprise AI, its business applications, and role in driving automation and growth.