AgentsThe story, in brief

Microsoft Open Sources code-testing-generator: a Polyglot Unit-Test Agent That Hits 92.1% Task Completion Versus 78.9% for Stock Copilot

92.1% vs 78.9%. Microsoft open-sourced a unit-test agent that outperforms stock Copilot on vague prompts — and it's MIT-licensed.

Illustration of independent geometric mechanisms passing paper tasks along branching amber tracks.
AI agents and the coordination of work.AI illustration by KeyNews
The KeyNews take

Why it matters

A concrete agent deployment showing how agentic patterns (read → plan → execute → validate) beat prompt-and-pray on real developer workflows. The polyglot capability and internal benchmark data matter for practitioners evaluating agent frameworks.

The key facts

7 to know
  1. Microsoft open-sourced code-testing-generator in dotnet/skills (MIT license)

  2. 92.1% task completion (140/152 tasks) vs 78.9% for stock Copilot (120/152) on internal benchmark

  3. Agent reads repo first: detects language, test framework, existing conventions, build/test commands

  4. Workflow: plan → write → run → validate tests

  5. Polyglot support (language-agnostic)

  6. Gain concentrated in vague prompts and diff-targeted requests

  7. Published August 6, 2026

Go to the source

MarkTechPostmarktechpost.com

Publisher excerpt: Microsoft has open sourced code-testing-generator, a polyglot unit-test agent shipping in the MIT-licensed dotnet/skills repository. It reads a repository before writing anything — detecting the language, test framework, existing conventions, and the real build and test commands — then plans,…
Read original report
Back to today's editionMore agents news

Keep reading

Related stories

More from Agents