BenchRank
#19 in Code AssistantsUpdated 2026-08

Pullfrog

by Pullfrog · Open-source GitHub agent bot that runs in GitHub Actions

BenchRank score

50.2 — BenchRank score out of 100

Screenshots of Pullfrog

Homepage · Pullfrog

Homepage of Pullfrog

Overview

Pullfrog is a GitHub bot that runs AI agents inside your own repo's GitHub Actions through a pullfrog.yml workflow. It listens for GitHub events — PRs opened, issues created, reviews submitted, CI failures — and starts agent runs from your configuration, or when someone mentions @pullfrog. Agents get a purpose-built MCP server for git and GitHub operations plus a headless browser.

Best for
GitHub teams wanting PR review, issue triage and CI fixes run by agents in their own Actions
Pricing
Free for personal GitHub accounts, $30/month per organisation for Pro, and a custom-priced Enterprise tier listed as coming soon.

Strengths and trade-offs

Strengths

  • Works with any LLM provider; switch models via config change
  • Flat $30/month per org, no per-run or per-seat billing
  • Keys kept as GitHub secrets; short-lived install token per run
  • Built-in headless browser and a GitHub/git MCP server

Trade-offs

  • GitHub only; needs a pullfrog.yml workflow added to the repo
  • Free tier covers personal accounts only; orgs pay $30/month
  • Model spend is separate: BYOK, or Router billed at provider cost
  • SSO, audit logs and SOC 2 sit in an Enterprise tier marked coming soon

How Pullfrog markets itself

A structured read of the promise, proof and page design on Pullfrog’s captured homepage.

Homepage capture

“The open-source CodeRabbit alternative that runs in GitHub Actions”

  • Angle: Contrarian / anti-incumbent
  • Hero: Product screenshot

Pricing

Published plans and prices from Pullfrog’s own pricing page.

How this score is made up

Each dimension is scored out of 100 and combined into the headline score using fixed weights.

  • MCP support

    Whether an agent can drive the product through the Model Context Protocol, and how much setup that takes.

  • API quality

    Public API surface: machine-readable spec, official SDKs, documented auth, errors, rate limits and versioning.

  • Documentation

    Publicly reachable docs — coverage, freshness, code samples and machine readability.

  • Agent friendliness

    How readable the site is to an automated client: llms.txt, structured data, server-rendered content, crawler access.

  • Changelog

    A public, dated record of what shipped and when — the clearest signal that a product is still alive.

  • Marketing site structure

    Whether the site answers a buyer's questions: clear positioning, the pages that matter, and accessibility.

  • Page speed

    How fast the site loads for real visitors: Chrome UX Report 75th-percentile LCP, INP and CLS, with a Lighthouse mobile run standing in where a site has too little traffic for field data.

  • Operational trust

    Status page and incident history, security disclosure, compliance and data-processing documentation.

Measured, but not part of the score

Useful to know, but not a mark for or against the product — so these do not affect the ranking.

  • Openness

    Source availability, self-hosting, data export and open standards. Scored and shown, but not part of the composite — paid SaaS is not worse for being paid SaaS.

  • Maintenance

    Release cadence and repository activity. Scored and shown, but not part of the composite — it is only measurable for open repositories.

This doesn’t look right — report a problem with Pullfrog’s score

Where this comes from

The Pullfrog pages BenchRank reads when it scores the product — its documentation, release notes, status and security pages, and its repository where there is one.

Alternatives in Code Assistants

  • Ranked 1

    78 — BenchRank score out of 100

    Superset

    Superset · Desktop app for running coding agents in parallel Git worktrees

    Best for: Developers on macOS running several CLI coding agents in parallel across isolated Git worktrees

  • Ranked 2

    77.9 — BenchRank score out of 100

    Warp

    Warp · Terminal for running and orchestrating coding agents

    Best for: Developers running coding agents like Claude Code or Codex who want them managed in one terminal.

  • Ranked 3

    74.2 — BenchRank score out of 100

    Frontman

    Frontman · AI website editor for existing WordPress, Next.js, Astro and Vite sites

    Best for: Designers, PMs and marketing teams editing existing sites without waiting on developer tickets.

See all 42 alternatives to Pullfrog

Is this your product?

Claim Pullfrog to manage its profile. Claiming lets you suggest edits to the descriptive fields — it never changes scores or rankings.

Claim this business

Report a problem with this page

Report an issue with Pullfrog