Line-Anchored Feedback Cuts Token Costs and Improves Correctness in AI Code Editing
FileMark VSCode extension uses line-anchored feedback to reduce token generation in Claude Opus (22%) and Sonnet (58%), cutting code-editing latency and cost.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
FileMark VSCode extension uses line-anchored feedback to reduce token generation in Claude Opus (22%) and Sonnet (58%), cutting code-editing latency and cost.
Codex usage grew 10x to 7M users in 6 months; article questions whether it has outpaced Claude Code amid sparse adoption metrics.
Claude users in India are starting to see Indian rupee-denominated subscription plans.
Automated red-teaming system discovers reusable vulnerability patterns in production LLM agents (Claude Code, Codex) operating on untrusted content.
Anthropic extends Claude Fable 5 availability through July 19 on paid plans, citing compute constraints and GPT-5.6 Sol positioning.
Anthropic partners with UST to deploy Claude in physical AI systems and robotics applications.
The AI firm Anthropic has developed a technique that has given it the clearest glimpse yet at what’s really going on inside large language models as they answer questions or carry out tasks. What they found ranges from the mundane to the unnerving. Researchers at the company built a tool called the Jacobian lens (or…
OpenAI releases GPT-5.6 family (Luna, Terra, Sol) with tiered pricing; claims superior agentic performance vs. Claude Opus/Fable on benchmarks.
Claude’s new Reflect dashboard doesn’t just visualize how you use AI. It also subtly reinforces how much of your daily work now depends on Anthropic’s chatbot.
The popularity of Spotify Wrapped has kicked off a wide range of year-in-review features, on apps from YouTube to Uber - and now, the lookback trend has come to AI. Anthropic on Thursday announced a "reflect" feature for its Claude chatbot, allowing users to see an analysis of their usage data over the past month, three months, six months, or year. Anthropic bills the reflection dashboard as a way to "see your patterns and shape them," the company wrote in a blog post. It begins with a summary of an individual's key topics brought up with Claude, as well as types of tasks they delegate and th...
Anthropic ships usage tracking and reflection feature for Claude, enabling users to monitor API/app consumption patterns.
When it comes to achieving artificial general intelligence (AGI), large language models just don’t have what it takes. Models like ChatGPT and Claude are great at text, but they’re less skilled at understanding how things actually move through space and time — an essential skill for producing intelligence that generalizes. That gap, it turns out, might be filled by gaming data. That’s the bet behind General Intuition, a […]
When it comes to achieving artificial general intelligence (AGI), large language models just don’t have what it takes. Models like ChatGPT and Claude are great at text, but they’re less skilled at understanding how things actually move through space and time — an essential skill for producing intelligence that generalizes. That gap, it turns out, might be filled by gaming data. That’s the bet behind General Intuition, a […]
Starting Tuesday, Anthropic's Claude Cowork AI platform will be available on mobile and web for the first time. The expanded access is rolling out first to Max subscribers and coming to Claude users on other plans "in the coming weeks." Claude Cowork was previously only accessible through the Claude desktop app for macOS and Windows, but now users on iOS and Android can also use it. However, Anthropic says the "full experience" for Cowork will still be on the desktop app, including features like local file access. Cowork sessions will also now run in the cloud by default, so you can continue ...
Anthropic’s Claude Cowork is now available on web and mobile for Max subscribers. Until now, Cowork largely lived on a user’s laptop. With the update, users can start a task from their desk, get status updates on their phone, and pick up the finished output later — even if their laptop is closed.
sqlite-utils 4.0rc4 release candidate incorporates Claude feedback before stable launch.
Anthropic accused of spying on users; engineer says “experiment” is over.
Government of Alberta deploys Claude Opus and Sonnet via Claude Code for cybersecurity vulnerability detection and remediation across government systems since 2025.
sqlite-utils 4.0rc2 released; Claude Fable assisted development at ~$149.25 cost.
Claude Opus 4.8 degrades tool-use reliability by inventing extra fields in nested JSON, breaking schema validation despite correct edit logic.
Alibaba has reportedly classified Claude Code as high-risk software.
Fanfiction communities are trying to hunt down writers who haven’t written works with their own hands. | Image: Álvaro Bernis / The Verge Over the past week, a new fanworks movement has kicked off, with the aim to root out authors using generative AI. But the detection methods being implemented are questionable, and any fanfic writer could be caught in the crossfire. Broad distaste around the use of Claude, ChatGPT, and other AI tools has long been a thing in creative communities, including the world of fanfiction. Readers and writers have passed around tips for spotting supposedly AI-generat...
Simon Willison reports Anthropic's Claude Code team recommends letting Fable/Opus apply autonomous judgment on tasks like testing rather than prescribing workflows.
SkillOpt-Lite: minimal skill optimization pipeline for autonomous agents via zeroth-order optimization, grounded in Claude.
Simon Willison's June 2026 newsletter covers Claude Fable 5, GPT-5.6, GLM-5.2 open weights, and US export restrictions.
At the event "The Briefing: AI for Science" earlier this week, Anthropic announced Claude Science, a new "AI workbench for scientists" that pulls fragmented tools and datasets into one environment, and generates figures and visuals. Anthropic, already dominating the industry with its popular coding tools and powerful AI models, framed the launch around what it says is AI's potential to "dramatically accelerate the pace of scientific discovery and the development of healthcare interventions," and touted a long list of biotech and pharma customers already using Claude. Anthropic also went a ste...
Comparative evaluation of frontier LLMs (GPT, Claude Opus, Gemini, GLM) for automated Linux/bash exam grading using cognitive taxonomy.
Ben Guez has "a bunch of potential international wives in [his] DMs," thanks to an automated script he set up using OpenClaw, Claude code, and Instagram trials.
Let’s start with a game. Open up your chatbot of choice—Claude, ChatGPT, Gemini—and type “Give me a random number between 1 and 10.” You’re going to get 7. Almost always. Now type “Another” and you’ll get 3 or 4. Type “Another” again and you’ll get 8 or 9. That won’t work every time—but if it…
Zero-shot benchmark evaluates Claude, GPT-5.4, and Gemini on fine-grained 13-class emotion classification.