Claude Dynamic Workflows: 1,000 Parallel Agents Now Available
Anthropic added dynamic workflows to Claude Managed Agents, enabling up to 1,000 AI agents to run in parallel per execution through managed agent infrastructure.
Mohamed Aminar is the founder and editor of AI Tool Herald. He sets the publication's editorial standards and is responsible for every story it publishes. Read how stories are sourced and corrected in the editorial policy.
Anthropic added dynamic workflows to Claude Managed Agents, enabling up to 1,000 AI agents to run in parallel per execution through managed agent infrastructure.
Microsoft releases Decision-1, a Qwen3.5-9B decision-scoring model for routing and classification tasks, available in Foundry and OpenRouter.
Nace AI open-sources Drex 1.5, a 9B decision model that scores multiple options in one forward pass, ranking first under 10B parameters on Decision Index.
Cloudflare launches Clef-omni, a decision model processing audio, video, images, and text in one API call, with price cuts and speed gains for existing models.
CoreWeave Forge connects training, inference, evaluation and agent development in one environment, letting teams run the entire AI improvement loop continuously.
JetBrains released Mellum2.1, a 12B mixture-of-experts open model for coding agents trained on real repositories under Apache 2.0 license.
Liquid AI's Liquid Context software enables on-device personal AI agents to improve after deployment within fixed hardware constraints, without cloud dependency.
Natura launched Interface, a $99 smart ring combining AI agent control with health tracking, letting users summon tasks hands-free via voice and access agents 24/7.
NVIDIA and Microsoft co-engineered RTX Spark hardware and Windows infrastructure to run AI agents locally on PCs, with laptop preorders open now.
Perplexity AI released pplx-embed-v2-late embedding models: a 0.6B edge-deployable model and a 9B high-performance variant for on-device and server RAG.
Atlassian announced the Agentic Multiplayer Protocol at Team '26 Europe to enable humans and AI agents to work together with shared context and governed permissions.
Anthropic released Claude Haiku 5.5 with up to 90% lower token costs and 72.4% computer-use benchmark performance, available now on AWS, Google Cloud, and Azure.
OpenAI's Decisions API classifies text and images 10x faster than Responses API for yes/no decisions, category picks, and scaled ratings.
Musubi announced PolicyLM-1.7B, an open-weight decision model for real-time content moderation that applies policies in under 50 milliseconds.
Cohere North 2 orchestrates multi-step AI agent workflows across organizations, supporting any model and deployment option.
Google DeepMind launches EmbeddingGemma 2, a 740M open embedding model mapping text, images, video, and audio into a unified space for on-device retrieval.
Mistral AI releases public preview of Mistral Large 4, a 1 trillion-parameter model trained in Europe with strengths in cybersecurity, coding, and multimodal tasks.
Reka released a 19B omni-reasoning model that handles text, images, video, and robot actions in a single neural network, replacing multi-model pipelines.
GitHub released ReviewBench, an open benchmark for evaluating AI code review agents on 219 realistic pull requests across 19 languages with independent validation.
Together AI released Together Link, a free CLI to run open-source models like Kimi K3 inside Claude Code and other coding agents.
OpenAI and Synopsys are building GPT-Synopsys, a model meant to reason about chip design and operate Synopsys EDA tools. What is known and what to verify.
Microsoft's September 25, 2026 announcement unifies Copilot around Home, Code, and Autopilot. Here is what each piece does and what teams should test first.
A measured reading of Anthropic's September 2026 threat report: how AI misuse is changing, where its evidence stops, and what defenders can do now.
OpenAI's Data agent in ChatGPT Work connects to warehouses, semantic layers, and BI tools. Here is what business teams should test before rollout.
OpenAI's ChatGPT for Financial Services pairs GPT-6 Astra with licensed data and citations. What it offers regulated teams and what to verify.
What GPT-Live-1 changes for voice apps, what OpenAI claims it costs and scores, and how to test latency, interruptions, accessibility, and cost before launch.
A plain-language guide to OpenAI's Agents API: the Codex harness it exposes, where agents run, what it costs, and the safety checks needed before agents act.
A sourced guide to GPT-6 Astra: what OpenAI claims, what it costs, the new admin controls, and the checks teams should run before adopting it.
OpenAI's ChatGPT Images 2.5 promises faster generation and steadier edits. What changed, the new API models, and how creators should test it.
Claude Fable 5.1 and Mythos 5.1 are one model with two safeguard levels. Here is what Anthropic claims, what it costs, and how to evaluate it.
What Anthropic's Model Hardware Standard research preview does, what early lab partners reported, and the safety checks teams need before agents touch hardware.
Anthropic is offering 10,000 free or discounted Claude team seats to scientists. Here is who qualifies, what the limits are, and how to evaluate it.
How Anthropic's Claude text watermark works, why it differs from AI detectors, and why a positive or missing signal is evidence, not proof.
What Anthropic's Claude Opus 5 launch means for coding and professional work, with pricing, safeguards, and the controls longer autonomous tasks need.
How to choose among GPT-5.6 Sol, Terra, and Luna using OpenAI's published prices and benchmarks, plus the routing tests to run on your own work.
Google says AI search visibility comes from useful, crawlable content, not special files or tricks. Here is a practical checklist based on its guide.
A practical automation model for AI-assisted publishing, built on Google's generative AI guidance, with human gates for evidence, originality, and approval.
Google launched Gemini 3.5 Flash as an agentic model family. This guide separates the launch claims from the tests teams should run themselves.
How to write clear affiliate disclosures for AI tool reviews, based on FTC staff guidance on placement, wording, and free or vendor-supplied access.
A repeatable method for comparing AI models on factual quality, instruction following, review time, latency, and total cost, with checklists you can reuse.
AI Tool Herald's method for reviewing AI tools: dated tests, preserved evidence, privacy checks, visible scoring weights, and clear limitations.
What Meta announced with Muse Spark, what its multimodal reasoning and multi-agent Contemplating mode imply, and which claims need independent checks.