Skip to content
All master reviews

Review Machine · Master Review · AI platforms

GitHub Copilot coding agent

The announcement against the pull requests

Across 12 sources read on 10 October 2026, GitHub's Copilot coding agent does what its announcement said on narrow, well-specified tasks, and the later reports are mostly about the review work its pull requests create, which no source we read sizes.

Preview 19 May 2025, Pro+ and EnterpriseGenerally available 25 September 2025, all paid plansRead 10 October 2026Skip to the verdictRead as text

Verdict age

ProvisionalMakers detailed · independents thin

The .NET report covers to 22 March 2026 and the launch-week thread is from May 2025. GitHub changes the product often, so later reports may differ.

  • Early · editors only
  • Provisional · editors in, owners thin
  • Settled · owner reviews read over months

From 12 sources to one verdict

Seven levels, in the same order in every Master Review. Each line is the takeaway; open a level for the evidence under it, or stop when you have enough.

What we read

12 sources, 19 May 2025 – 10 Oct 2026. News and forum tiers thin. Editorial tier empty.

By tier, heaviest first

T1Official7

Seven GitHub and Microsoft pages. The .NET posts are Microsoft's own team on its own use, so they are first-party claims.

T2Editorial0None readable

No independent hands-on review could be read.

T2News3Thin

Three pieces, 20 May 2025 to 7 May 2026. One is about agent pull requests in general.

T3Buyers0None readable
Forum1Thin

One Hacker News thread from launch week. Themes read, not counted.

Other1

One independent write-up of the .NET report.

Date window

19 May 2025 to 10 Oct 2026

News 20 May 2025 – 7 May 2026Sources published 19 May 2025 to 26 May 2026. Two pages read and one search run on 10 Oct 2026.

Couldn’t read, so not used

  • Visual Studio Magazine, 29 Jan 2026HTTP 403
  • Visual Studio Magazine, 3 Jun 2026HTTP 403

Nothing was counted that rested only on an unreadable page. No review list was paged and no filter applied.

Read, then rejected

  • PRarena tracker · a third-party tracker with no date shown, so its merge figures are not repeated
  • GitHub Community discussion 165382 · about Copilot code review errors, not the coding agent
  • Hacker News resubmissions, Mar 2026 · four posts of the .NET report with no comments

What they measured

In one repository, 535 merged out of 878 opened. GitHub's search counts 1,528,432 merged out of 2,118,197.

dotnet/runtime pull requests, 19 May 2025 to 22 Mar 2026

878opened by the agent

.NET blog · 23 Mar

535merged

.NET blog · 23 Mar

Success rate by task type, dotnet/runtime

.NET blog · 23 Mar

With a person's commits against autonomous

.NET blog · 23 Mar

Success rate before and after an instructions file

38.1%before

.NET blog · 23 Mar

69%after

.NET blog · 23 Mar

Start Debugging gives 41.7% and 72% for the same change.

Public pull requests by the agent's bot account

2,118,197opened, GitHub search totalIncludes open ones

GitHub API · searched 10 Oct

1,528,432merged, GitHub search total

GitHub API · searched 10 Oct

Longest session

59minone pull request per task

GitHub Docs · read 10 Oct

Where they agree

Narrow, well-described tasks, with a person reviewing before anything merges.

Best on bounded, clearly described tasks

Cleanup 84.7%, tests 75.6%, bug fixes 69.4% in the .NET report. T1Official T2News Other

A person reviews before anything merges

T1Official T2News

Draft pull requests and logs make the work followable

T2News

Setup matters most

38.1% before an instructions file, 69% after, in one repository. T1Official

What keeps coming up against it

The review work is the lead complaint, reported but not sized. Some pull requests never merge.

Review capacity

Reported · size unknown
Documented byOtherT2NewsForumHow commonUnknown: no source we read counts it
Other

Reads the .NET report as saying one developer can create five to nine hours of review work remotely.

Start Debugging · 29 Mar

T2News

Hackernoon reports unease that human work becomes approving machine output. Let's Data Science, on agent pull requests in general, warns that tidy changes can hide duplicated logic.

Hackernoon · 26 Oct 2025Let's Data Science · 7 May

Forum

The launch-week thread doubts that busy teams will scrutinise every change. Not counted

Hacker News · 19 May 2025

Review capacity

T2News Forum Other

Not every pull request lands

535 of 878 in .NET; 1,528,432 of 2,118,197 in GitHub search. T1Official Other

Limits on harder work

Performance work 54.5%; Linux only. T1Official Forum Other

Where they split

Four places the evidence pulls two ways, starting with how much gets merged.

How much gets merged

T1Official

In one repository, 535 of 878 pull requests were merged.

T1Official

Against that: Across public GitHub, 1,528,432 merged out of 2,118,197 opened, open ones included.

Our readDifferent populations, and neither says how carefully merged changes were read.

What setup is worth

T1Official

38.1% before an instructions file, 69% after.

Other

Against that: 41.7% before, 72% after, writing about the same report.

SoWe use the report's own figures.

The gate against the habit

T2News

Every pull request needs human approval before CI workflows run.

Forum

Against that: Launch-week commenters ask whether people will do more than approve.

Both trueOne describes a gate, the other a habit.

Then and now

T1GitHub · 19 May 2025

Preview for Pro+ and Enterprise.

T1GitHub · now

Against that: Every paid plan, mention-only replies from write-access users, 59-minute sessions.

SoEarly reports describe a narrower product.

Who it's for, who should pass

For teams with tests, written instructions and review time. Pass if review time is the thing you lack.

It suits you if

  • You maintain a GitHub-hosted repository with working tests and written build and test commands.The .NET report's success rate rose from 38.1% to 69% after an instructions file.
  • You have a backlog of cleanup, test additions and well-described bug fixes.Those are the strongest categories in the .NET report.

Pass if

  • You cannot spare reviewer time.Review capacity is the lead complaint, and no source sizes it.
  • You need Windows or macOS-specific changes checked before merge.The .NET report says the agent runs on Linux only.
  • Your work is performance tuning or design decisions.Performance work is the weakest category in the .NET report, at 54.5%.

The verdict

It does what the announcement said. The review cost is reported and unsized.
VerdictProvisionalMakers detailed · independents thin

A background agent that suits bounded, well-described tasks and hands a person a draft to review, with a review cost that sources describe and none of them can size.

Confidence, by tier

Officialstrong

Detailed, but first-party claims from the makers and their parent company.

Newsthin

Three pieces, one about agent pull requests in general.

Forumthin

One launch-week thread, themes not counted.

Editorialnone

No independent hands-on review could be read.

Editorial evidence
None read
Newest report
7 May 2026
Read
10 Oct · month 16

Share line

It opens the pull request while you sleep. You review it awake.
Review MachineGitHub Copilot coding agent · Master Review

Rests onGitHub's launch text says the agent works in the background and tags you for review; four sources raise review capacity, size unknown.

Sources

12 sources, heaviest tier first. Every figure above comes from one of these.

  1. T1OfficialGitHub changelog, public preview19 May 2025
  2. T1OfficialGitHub changelog, review experience5 Aug 2025
  3. T1OfficialGitHub changelog, generally available25 Sep 2025
  4. T1Official.NET blog, ten months in dotnet/runtime23 Mar
  5. T1Official.NET blog, doing more with Copilot26 May
  6. T1OfficialGitHub Docs, about the coding agentread 10 Oct
  7. T1OfficialGitHub search, pull requests by the bot accountone public search call, total onlysearched 10 Oct
  8. T2NewsInfoWorld, GitHub unveils coding agent20 May 2025
  9. T2NewsHackernoon on the cloud agent26 Oct 2025
  10. T2NewsLet's Data Science on agent pull requests7 May
  11. ForumHacker News, launch-week threadone ordinary thread read, themes not counted19 May 2025
  12. OtherStart Debugging on the ten-month data29 Mar

MethodOn 10 October 2026 we read 12 sources from 19 May 2025 to 26 May 2026 by web search and plain page fetch, and synthesised them with AI; we collected no developer reviews, ran nothing and tested nothing.

Dates without a year are 2026.

The review as text

The same review as one piece of writing · 6 min read

GitHub Copilot's coding agent, which GitHub's documentation now calls the Copilot cloud agent, takes a task, works in the background and opens a pull request for a person to review. GitHub announced it in public preview on 19 May 2025 for Copilot Pro+ and Copilot Enterprise. It became generally available to all paid Copilot plans on 25 September 2025, with an administrator switching it on for Business and Enterprise accounts.

We read 12 sources: seven official pages from GitHub and Microsoft, three news pieces, one independent write-up and one Hacker News thread. They were published from 19 May 2025 to 26 May 2026, and we read two pages and ran one public search on 10 October 2026. The independent hands-on tier is empty: two trade-press hands-on articles returned HTTP 403 and could not be read. The announcement promised a pull request. This review holds that promise against what developers report about merging them.

Consensus

Across the sources we read, the agent does well on narrow, well-specified work, and its pull requests are drafts for a person to approve rather than finished changes. Five of the 12 sources describe a person reviewing before anything merges as part of the design: three official pages (GitHub's preview and general availability announcements, and the .NET team's May 2026 advice) and two news pieces (InfoWorld and Hackernoon).

The announcement said the agent suits low-to-medium complexity tasks in well-tested code. The later reports agree with that scope. Where they pull away from the announcement is review: the launch text says the agent tags you for review, and none of the sources we read sizes how much review that creates. We read this as an announcement that was accurate about what the agent attempts and quiet about the cost on the reviewer's side. That is our interpretation, not a finding any source states.

Confidence is mixed to thin. The official evidence is detailed but comes from the makers and their parent company. The independent evidence is three news pieces, one write-up and one forum thread, and nothing from owners after months of use that we could read.

Recurring strengths

Six of the 12 sources say the agent is best on bounded, clearly described tasks: three official (GitHub's preview announcement and the .NET team's two posts), two news (InfoWorld and Hackernoon) and one independent write-up (Start Debugging). The .NET team's ten-month report, covering 19 May 2025 to 22 March 2026 in the dotnet/runtime repository, gives its success rate by task: 84.7% for removal and cleanup, 75.6% for tests, and 69.4% for bug fixes.

Two news sources, InfoWorld and Hackernoon, say the draft pull request and session logs make the agent's work easy to follow. One official source, the .NET report, says setup mattered most: its success rate rose from 38.1% to 69% once the repository held a written instructions file with build commands and test patterns. That is one team's account of one repository.

Recurring complaints

Four of the 12 sources raise review capacity: one independent write-up (Start Debugging), one forum thread (Hacker News) and two news pieces (Hackernoon and Let's Data Science). Start Debugging reads the .NET report as saying one developer can create five to nine hours of review work remotely. Hackernoon reports unease that human work becomes approving machine output. The Hacker News thread, from launch week, doubts that busy teams will scrutinise every change. Let's Data Science writes about agent-written pull requests in general, not this agent, and warns that tidy-looking changes can hide duplicated logic. No source counts how many teams are affected.

Three sources, two official and one independent, show that not every pull request lands. In the .NET report, 535 of 878 pull requests were merged. On 10 October 2026, GitHub's public search reported 1,528,432 merged out of 2,118,197 pull requests opened by the agent's bot account, a count that includes pull requests still open.

Three sources, one official, one independent and the Hacker News thread, describe limits on harder work. The .NET report gives 54.5% for performance optimisation, says the agent runs on Linux only so it cannot check Windows or macOS code, and says pull requests with a person's own commits on top reached 86% against 55% for fully autonomous ones. Developers in the launch-week thread said the code works but is structured poorly. We did not count how many.

Where reviewers split

How much gets merged. The .NET report counts one repository. GitHub's search counts every public pull request from the bot account. Neither says how carefully the merged ones were read, and a merge count is not a quality measure.

What setup is worth. The .NET post puts the success rate at 38.1% before an instructions file and 69% after. Start Debugging, writing about the same report, gives 41.7% and 72%. We use the report's own figures, and treat the difference as a reason to read the original.

Announcement against the thread. InfoWorld reports that every pull request needs human approval before CI workflows run. The launch-week Hacker News thread asks whether people will do more than approve. Both can be true, since the first describes a gate and the second describes habit.

Then and now. The preview announcement named two plans, and the general availability announcement names every paid plan. GitHub's changelog of 5 August 2025 says the agent now answers only to people with write access who mention it, and the documentation read on 10 October 2026 limits a session to 59 minutes and one pull request per task.

Who it suits

You maintain a GitHub-hosted repository with working tests and can write down build and test commands for the agent. You have a backlog of cleanup, test additions and well-described bug fixes, and someone with time to read each pull request.

Who should pass

You cannot spare reviewer time, or you need Windows or macOS-specific changes checked before merge. Performance tuning and design decisions are also where the .NET report records its weakest results. If your code lives outside GitHub, the documentation says the agent does not work there.

One verdict a week: the most useful Master Review we finished, the complaint that kept appearing, and who should skip it.

Sources

On 10 October 2026 we read 12 sources from 19 May 2025 to 26 May 2026 by web search and plain page fetch, and synthesised them with AI; we collected no developer reviews, ran nothing and tested nothing.

One verdict a week.

Every week, the most useful Master Review we finished: what the internet agrees on, the complaint that kept appearing, and who should skip it.

By subscribing you agree to receive one weekly email from Review Machine. You can unsubscribe at any time.