Review Machine · Master Review · AI platforms
GitHub Copilot coding agent
The announcement against the pull requests
Across 12 sources read on 10 October 2026, GitHub's Copilot coding agent does what its announcement said on narrow, well-specified tasks, and the later reports are mostly about the review work its pull requests create, which no source we read sizes.
Preview 19 May 2025, Pro+ and EnterpriseGenerally available 25 September 2025, all paid plansRead 10 October 2026Skip to the verdictRead as text

Verdict age
ProvisionalMakers detailed · independents thinThe .NET report covers to 22 March 2026 and the launch-week thread is from May 2025. GitHub changes the product often, so later reports may differ.
- Early · editors only
- Provisional · editors in, owners thin
- Settled · owner reviews read over months
From 12 sources to one verdict
Seven levels, in the same order in every Master Review. Each line is the takeaway; open a level for the evidence under it, or stop when you have enough.
What we read
12 sources, 19 May 2025 – 10 Oct 2026. News and forum tiers thin. Editorial tier empty.
By tier, heaviest first
Seven GitHub and Microsoft pages. The .NET posts are Microsoft's own team on its own use, so they are first-party claims.
No independent hands-on review could be read.
Three pieces, 20 May 2025 to 7 May 2026. One is about agent pull requests in general.
One Hacker News thread from launch week. Themes read, not counted.
One independent write-up of the .NET report.
Date window
19 May 2025 to 10 Oct 2026
News 20 May 2025 – 7 May 2026Sources published 19 May 2025 to 26 May 2026. Two pages read and one search run on 10 Oct 2026.
Couldn’t read, so not used
- Visual Studio Magazine, 29 Jan 2026HTTP 403
- Visual Studio Magazine, 3 Jun 2026HTTP 403
Nothing was counted that rested only on an unreadable page. No review list was paged and no filter applied.
Read, then rejected
- PRarena tracker · a third-party tracker with no date shown, so its merge figures are not repeated
- GitHub Community discussion 165382 · about Copilot code review errors, not the coding agent
- Hacker News resubmissions, Mar 2026 · four posts of the .NET report with no comments
What they measured
In one repository, 535 merged out of 878 opened. GitHub's search counts 1,528,432 merged out of 2,118,197.
dotnet/runtime pull requests, 19 May 2025 to 22 Mar 2026
878opened by the agent
.NET blog · 23 Mar535merged
.NET blog · 23 MarSuccess rate by task type, dotnet/runtime
.NET blog · 23 MarWith a person's commits against autonomous
.NET blog · 23 MarSuccess rate before and after an instructions file
38.1%before
.NET blog · 23 Mar69%after
.NET blog · 23 MarStart Debugging gives 41.7% and 72% for the same change.
Public pull requests by the agent's bot account
2,118,197opened, GitHub search totalIncludes open ones
GitHub API · searched 10 Oct1,528,432merged, GitHub search total
GitHub API · searched 10 OctLongest session
59minone pull request per task
GitHub Docs · read 10 OctWhere they agree
Narrow, well-described tasks, with a person reviewing before anything merges.
Best on bounded, clearly described tasks
Cleanup 84.7%, tests 75.6%, bug fixes 69.4% in the .NET report. T1Official T2News Other
A person reviews before anything merges
T1Official T2News
Draft pull requests and logs make the work followable
T2News
Setup matters most
38.1% before an instructions file, 69% after, in one repository. T1Official
What keeps coming up against it
The review work is the lead complaint, reported but not sized. Some pull requests never merge.
Review capacity
Reported · size unknownReads the .NET report as saying one developer can create five to nine hours of review work remotely.
Start Debugging · 29 Mar
Hackernoon reports unease that human work becomes approving machine output. Let's Data Science, on agent pull requests in general, warns that tidy changes can hide duplicated logic.
Hackernoon · 26 Oct 2025Let's Data Science · 7 May
The launch-week thread doubts that busy teams will scrutinise every change. Not counted
Hacker News · 19 May 2025
Review capacity
T2News Forum Other
Not every pull request lands
535 of 878 in .NET; 1,528,432 of 2,118,197 in GitHub search. T1Official Other
Limits on harder work
Performance work 54.5%; Linux only. T1Official Forum Other
Where they split
Four places the evidence pulls two ways, starting with how much gets merged.
How much gets merged
In one repository, 535 of 878 pull requests were merged.
Against that: Across public GitHub, 1,528,432 merged out of 2,118,197 opened, open ones included.
Our readDifferent populations, and neither says how carefully merged changes were read.
What setup is worth
38.1% before an instructions file, 69% after.
Against that: 41.7% before, 72% after, writing about the same report.
SoWe use the report's own figures.
The gate against the habit
Every pull request needs human approval before CI workflows run.
Against that: Launch-week commenters ask whether people will do more than approve.
Both trueOne describes a gate, the other a habit.
Then and now
Preview for Pro+ and Enterprise.
Against that: Every paid plan, mention-only replies from write-access users, 59-minute sessions.
SoEarly reports describe a narrower product.
Who it's for, who should pass
For teams with tests, written instructions and review time. Pass if review time is the thing you lack.
It suits you if
- You maintain a GitHub-hosted repository with working tests and written build and test commands.The .NET report's success rate rose from 38.1% to 69% after an instructions file.
- You have a backlog of cleanup, test additions and well-described bug fixes.Those are the strongest categories in the .NET report.
Pass if
- You cannot spare reviewer time.Review capacity is the lead complaint, and no source sizes it.
- You need Windows or macOS-specific changes checked before merge.The .NET report says the agent runs on Linux only.
- Your work is performance tuning or design decisions.Performance work is the weakest category in the .NET report, at 54.5%.
The verdict
It does what the announcement said. The review cost is reported and unsized.
A background agent that suits bounded, well-described tasks and hands a person a draft to review, with a review cost that sources describe and none of them can size.
Confidence, by tier
Detailed, but first-party claims from the makers and their parent company.
Three pieces, one about agent pull requests in general.
One launch-week thread, themes not counted.
No independent hands-on review could be read.
- Editorial evidence
- None read
- Newest report
- 7 May 2026
- Read
- 10 Oct · month 16
It opens the pull request while you sleep. You review it awake.
Rests onGitHub's launch text says the agent works in the background and tags you for review; four sources raise review capacity, size unknown.
Sources
12 sources, heaviest tier first. Every figure above comes from one of these.
- T1OfficialGitHub changelog, public preview19 May 2025
- T1OfficialGitHub changelog, review experience5 Aug 2025
- T1OfficialGitHub changelog, generally available25 Sep 2025
- T1Official.NET blog, ten months in dotnet/runtime23 Mar
- T1Official.NET blog, doing more with Copilot26 May
- T1OfficialGitHub Docs, about the coding agentread 10 Oct
- T1OfficialGitHub search, pull requests by the bot accountone public search call, total onlysearched 10 Oct
- T2NewsInfoWorld, GitHub unveils coding agent20 May 2025
- T2NewsHackernoon on the cloud agent26 Oct 2025
- T2NewsLet's Data Science on agent pull requests7 May
- ForumHacker News, launch-week threadone ordinary thread read, themes not counted19 May 2025
- OtherStart Debugging on the ten-month data29 Mar
MethodOn 10 October 2026 we read 12 sources from 19 May 2025 to 26 May 2026 by web search and plain page fetch, and synthesised them with AI; we collected no developer reviews, ran nothing and tested nothing.
Dates without a year are 2026.
The review as text
GitHub Copilot's coding agent, which GitHub's documentation now calls the Copilot cloud agent, takes a task, works in the background and opens a pull request for a person to review. GitHub announced it in public preview on 19 May 2025 for Copilot Pro+ and Copilot Enterprise. It became generally available to all paid Copilot plans on 25 September 2025, with an administrator switching it on for Business and Enterprise accounts.
We read 12 sources: seven official pages from GitHub and Microsoft, three news pieces, one independent write-up and one Hacker News thread. They were published from 19 May 2025 to 26 May 2026, and we read two pages and ran one public search on 10 October 2026. The independent hands-on tier is empty: two trade-press hands-on articles returned HTTP 403 and could not be read. The announcement promised a pull request. This review holds that promise against what developers report about merging them.
Consensus
Across the sources we read, the agent does well on narrow, well-specified work, and its pull requests are drafts for a person to approve rather than finished changes. Five of the 12 sources describe a person reviewing before anything merges as part of the design: three official pages (GitHub's preview and general availability announcements, and the .NET team's May 2026 advice) and two news pieces (InfoWorld and Hackernoon).
The announcement said the agent suits low-to-medium complexity tasks in well-tested code. The later reports agree with that scope. Where they pull away from the announcement is review: the launch text says the agent tags you for review, and none of the sources we read sizes how much review that creates. We read this as an announcement that was accurate about what the agent attempts and quiet about the cost on the reviewer's side. That is our interpretation, not a finding any source states.
Confidence is mixed to thin. The official evidence is detailed but comes from the makers and their parent company. The independent evidence is three news pieces, one write-up and one forum thread, and nothing from owners after months of use that we could read.
Recurring strengths
Six of the 12 sources say the agent is best on bounded, clearly described tasks: three official (GitHub's preview announcement and the .NET team's two posts), two news (InfoWorld and Hackernoon) and one independent write-up (Start Debugging). The .NET team's ten-month report, covering 19 May 2025 to 22 March 2026 in the dotnet/runtime repository, gives its success rate by task: 84.7% for removal and cleanup, 75.6% for tests, and 69.4% for bug fixes.
Two news sources, InfoWorld and Hackernoon, say the draft pull request and session logs make the agent's work easy to follow. One official source, the .NET report, says setup mattered most: its success rate rose from 38.1% to 69% once the repository held a written instructions file with build commands and test patterns. That is one team's account of one repository.
Recurring complaints
Four of the 12 sources raise review capacity: one independent write-up (Start Debugging), one forum thread (Hacker News) and two news pieces (Hackernoon and Let's Data Science). Start Debugging reads the .NET report as saying one developer can create five to nine hours of review work remotely. Hackernoon reports unease that human work becomes approving machine output. The Hacker News thread, from launch week, doubts that busy teams will scrutinise every change. Let's Data Science writes about agent-written pull requests in general, not this agent, and warns that tidy-looking changes can hide duplicated logic. No source counts how many teams are affected.
Three sources, two official and one independent, show that not every pull request lands. In the .NET report, 535 of 878 pull requests were merged. On 10 October 2026, GitHub's public search reported 1,528,432 merged out of 2,118,197 pull requests opened by the agent's bot account, a count that includes pull requests still open.
Three sources, one official, one independent and the Hacker News thread, describe limits on harder work. The .NET report gives 54.5% for performance optimisation, says the agent runs on Linux only so it cannot check Windows or macOS code, and says pull requests with a person's own commits on top reached 86% against 55% for fully autonomous ones. Developers in the launch-week thread said the code works but is structured poorly. We did not count how many.
Where reviewers split
How much gets merged. The .NET report counts one repository. GitHub's search counts every public pull request from the bot account. Neither says how carefully the merged ones were read, and a merge count is not a quality measure.
What setup is worth. The .NET post puts the success rate at 38.1% before an instructions file and 69% after. Start Debugging, writing about the same report, gives 41.7% and 72%. We use the report's own figures, and treat the difference as a reason to read the original.
Announcement against the thread. InfoWorld reports that every pull request needs human approval before CI workflows run. The launch-week Hacker News thread asks whether people will do more than approve. Both can be true, since the first describes a gate and the second describes habit.
Then and now. The preview announcement named two plans, and the general availability announcement names every paid plan. GitHub's changelog of 5 August 2025 says the agent now answers only to people with write access who mention it, and the documentation read on 10 October 2026 limits a session to 59 minutes and one pull request per task.
Who it suits
You maintain a GitHub-hosted repository with working tests and can write down build and test commands for the agent. You have a backlog of cleanup, test additions and well-described bug fixes, and someone with time to read each pull request.
Who should pass
You cannot spare reviewer time, or you need Windows or macOS-specific changes checked before merge. Performance tuning and design decisions are also where the .NET report records its weakest results. If your code lives outside GitHub, the documentation says the agent does not work there.
Sources
- GitHub changelog, Copilot coding agent in public preview, official, 19 May 2025
- InfoWorld, GitHub unveils coding agent for GitHub Copilot, news, 20 May 2025
- Hacker News, launch-week thread on the coding agent, forum, 19 May 2025
- GitHub changelog, improved pull request review experience, official, 5 August 2025
- GitHub changelog, coding agent generally available, official, 25 September 2025
- Hackernoon, GitHub's Copilot adds cloud agent to draft pull requests, news, 26 October 2025
- .NET blog, Ten months with Copilot coding agent in dotnet/runtime, official, 23 March 2026
- Start Debugging, the dotnet/runtime ten-month data, independent write-up, 29 March 2026
- Let's Data Science, engineers review agent-generated pull requests, news, 7 May 2026
- .NET blog, doing more with GitHub Copilot, official, 26 May 2026
- GitHub Docs, about the Copilot coding agent, official, read 10 October 2026
- GitHub search API, pull requests by the agent's bot account, official, searched 10 October 2026
On 10 October 2026 we read 12 sources from 19 May 2025 to 26 May 2026 by web search and plain page fetch, and synthesised them with AI; we collected no developer reviews, ran nothing and tested nothing.
One verdict a week.
Every week, the most useful Master Review we finished: what the internet agrees on, the complaint that kept appearing, and who should skip it.