Pr Af
PR-AF is the #1 open-source code reviewer on Martian Code-Review-Bench. It is built for deep code review, not shallow diff summaries: turn each PR into a task-specific review plan, spawn focused reviewer agents, ground findings in code evidence, challenge the results, and squeeze more useful review intelligence out of cheaper models. Run DeepSeek-class models for routine PRs, GLM-5.2 for deep open-model reviews, or Opus-class frontier models for major PRs, where PR-AF tops the benchmark by a wide margin. On the 38 runnable Martian Code-Review-Bench PRs, PR-AF with GLM-5.2 is the #1 open-source reviewer in golden recall: 0.706 across 42 compared tools. It is ahead of cubic-v2 and every qodo, coderabbit, greptile, copilot, and devin variant in this snapshot. strength result ------ Known bug recall 0.706 golden recall, #1 open source across 42 compared tools. More real issues found 595 independently valid findings, ~3× more than the leading commercial tools in the adjusted comparison. Open + reproducible Single open model (GLM-5.2), public results, per-PR judge verdicts, and reproduction scripts. Self-hosted API Run locally with Docker; trigger reviews by CLI, curl, CI, or other agents. Model-flexible Use cheaper models for regular PRs, GLM-5.2 for open-model CI gates, and Opus-class frontier models for highest-stakes reviews. Frontier ceiling With Opus-class commercial models, PR-AF tops the benchmark by a wide margin. Cost position About 10× cheaper per review than closed-source tools.
View Pr Af on GitHub
PR-AF is the #1 open-source code reviewer on Martian Code-Review-Bench. It is built for deep code review, not shallow diff summaries: turn each PR into a task-specific review plan, spawn focused reviewer agents, ground findings in code evidence, challenge the results, and squeeze more useful review intelligence out of cheaper models. Run DeepSeek-class models for routine PRs, GLM-5.2 for deep open-model reviews, or Opus-class frontier models for major PRs, where PR-AF tops the benchmark by a wide margin.
On the 38 runnable Martian Code-Review-Bench PRs, PR-AF with GLM-5.2 is the #1 open-source reviewer in golden recall: 0.706 across 42 compared tools. It is ahead of cubic-v2 and every qodo, coderabbit, greptile, copilot, and devin variant in this snapshot.
strength result ------ Known bug recall 0.706 golden recall, #1 open source across 42 compared tools. More real issues found 595 independently valid findings, ~3× more than the leading commercial tools in the adjusted comparison. Open + reproducible Single open model (GLM-5.2), public results, per-PR judge verdicts, and reproduction scripts. Self-hosted API Run locally with Docker; trigger reviews by CLI, curl, CI, or other agents. Model-flexible Use cheaper models for regular PRs, GLM-5.2 for open-model CI gates, and Opus-class frontier models for highest-stakes reviews. Frontier ceiling With Opus-class commercial models, PR-AF tops the benchmark by a wide margin. Cost position About 10× cheaper per review than closed-source tools.
Pr Af at a glance
| Stars | 640 |
|---|---|
| Forks | 71 |
| Language | Go |
| Last update | 2026-09-21 |
| Contributors | 6 |
How to install Pr Af
bash git clone https://github.com/Agent-Field/pr-af
Where Pr Af is listed
Navid.me is reader-supported. When you buy through links on this site, I may earn an affiliate commission. Learn more.
GitHub repos like this
More repo topics
More free tools
Related MCP servers & CLIs
The most actionable AI newsletter for founders
Every week, get proven AI strategies, curated tools, and step-by-step systems to grow your audience, create better content, and build a profitable creator business.
No fluff, no filler, no BS. Just five minutes each week that might level up your online business and life.
P.S. Sign up now to get free access to my ultimate AI tools guide for creators.







































