Source-linked reporting

AI news in plain English—and why it matters

Simple updates on the AI tools, companies, and changes that may affect your work or everyday life. Every story explains the practical takeaway and links to the original source.

Read Anthropic Adds Plugin Evals to Claude Code: 6 Grader Types, a No-Plugin Baseline, and a CI Gate for Skills
AI toolsBuilding with AISep 11, 2026

Anthropic Adds Plugin Evals to Claude Code: 6 Grader Types, a No-Plugin Baseline, and a CI Gate for Skills

Plugin developers can use these checks to catch problems before release and add them to continuous integration, which automatically tests changes. The feature is aimed at Claude Code users building or maintaining plugins, rather than everyday users.

MarkTechPostSee why it matters
Read Cognition helps Devin test its own work with GPT‑6 Astra
New AI featuresAI toolsSep 11, 2026

Cognition helps Devin test its own work with GPT‑6 Astra

This could make Devin more useful for engineers by giving it a stronger role in checking its own work. The feed does not provide details about availability, pricing, or access.

OpenAI NewsSee why it matters
Read Rapidly scaling online storage to serve over 1 billion ChatGPT users
AI toolsBuilding with AISep 11, 2026

Rapidly scaling online storage to serve over 1 billion ChatGPT users

This shows the infrastructure needed to operate ChatGPT at very large scale, but it does not announce a new user feature, pricing change, or broader availability. The update is mainly relevant to people building or operating large online services.

OpenAI NewsSee why it matters
Read New Deepseek model V4.1-Flash cuts memory needs for AI agents
New AI featuresFree toolsSep 10, 2026

New Deepseek model V4.1-Flash cuts memory needs for AI agents

The MIT-licensed release is aimed at developers building AI agents, which may be able to run with lower memory needs and potentially lower costs. The feed does not state where the model can be downloaded or provide confirmed pricing.

The DecoderSee why it matters
Read Google Open-Sources Mantis: A Modular Skills Toolkit That Lets Coding Agents Find, Reproduce and Patch Vulnerabilities
Building with AIFree toolsSep 9, 2026

Google Open-Sources Mantis: A Modular Skills Toolkit That Lets Coding Agents Find, Reproduce and Patch Vulnerabilities

Developers can use the Apache 2.0 toolkit to automate more of the vulnerability-review process across different coding setups. It is currently documented as demonstration-only, so users should not treat it as a production-ready security solution.

MarkTechPostSee why it matters
Read OpenBMB Releases MiniCPM5-2B: A 2.52B Dense Model Averaging 53.9 Across 34 Benchmarks and Built to Run On Device
New AI featuresFree toolsSep 7, 2026

OpenBMB Releases MiniCPM5-2B: A 2.52B Dense Model Averaging 53.9 Across 34 Benchmarks and Built to Run On Device

Developers can download and run the model locally through tools including Ollama, llama.cpp, vLLM, SGLang, and MLX, without changing the model code. Its relatively small file size may make local use more practical, though the feed does not state which consumer devices can run it or how fast it will be.

MarkTechPostSee why it matters
Read H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder
New AI featuresIdeas & discoveriesSep 6, 2026

H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder

NeoMME could help build faster, smaller search systems for documents that mix writing and images. The feed does not say whether the models are publicly available, under what license, or what they cost, so practical access remains unclear.

MarkTechPostSee why it matters
Read OpenAI Releases GPT-6 Astra: A 1.05M-Context Computer-Use Model Gated Behind a ‘Critical’ Cyber Threshold
New AI featuresBuilding with AISep 3, 2026

OpenAI Releases GPT-6 Astra: A 1.05M-Context Computer-Use Model Gated Behind a ‘Critical’ Cyber Threshold

Astra may help eligible users automate computer-based tasks, but access and permitted uses are limited by OpenAI’s “Critical” cybersecurity classification. Its listed price is $10 for input and $50 for output per million text units processed.

MarkTechPostSee why it matters
Read Anthropic Released Claude Commerce Agents: An Apache-2.0 Blueprint for Shopping and Merchant Agents Across Retail, Travel, Telecom and Entertainment
AI toolsFree toolsSep 3, 2026

Anthropic Released Claude Commerce Agents: An Apache-2.0 Blueprint for Shopping and Merchant Agents Across Retail, Travel, Telecom and Entertainment

Teams can use the published code as a starting point instead of rebuilding this basic infrastructure themselves. It is available under a permissive open-source license, but the feed does not specify how ready it is for production use or what it costs to operate.

MarkTechPostSee why it matters
Read Meta is paying to peek at how you use their latest AI model
New AI featuresBuilding with AISep 3, 2026

Meta is paying to peek at how you use their latest AI model

Developers can use the model at a much lower cost, but the discount requires contributing their interactions to help Meta develop future models. The offer may be most useful for people testing coding or other automated agents who are comfortable sharing this data.

TechCrunch AISee why it matters
Read GPT-6 Astra Is Here—and OpenAI Thinks It May Kick Off the AGI Era
New AI featuresBuilding with AISep 3, 2026

GPT-6 Astra Is Here—and OpenAI Thinks It May Kick Off the AGI Era

If available, the model could help people automate computer-based tasks and create software more effectively. The feed does not specify its release status, price, or access requirements.

WIRED AISee why it matters
Read Perplexity Open Sources Lily: A Rust + Metal Inference Engine for Qwen3.6-35B-A3B on Apple Silicon
Free toolsBuilding with AISep 3, 2026

Perplexity Open Sources Lily: A Rust + Metal Inference Engine for Qwen3.6-35B-A3B on Apple Silicon

Developers with compatible Apple hardware can now access the code behind a local component of Perplexity Computer and potentially run this model more efficiently. Lily is narrowly optimized for one model and Apple’s chip family, so its usefulness may not extend to other hardware or models.

MarkTechPostSee why it matters