Bullet

Coding · tested 2026-08-21 · re-test due 2026-11-21 · by the Hlido desk, not the vendor

In short: A free, speed-obsessed coding agent (GUI + CLI) that routes/searches/executes with a tighter loop than the norm — YC-backed, with a headline 95.8% SWE-bench Verified claim you'll want to confirm.

4 PASS · 0 FAIL of 4 public-surface claims

Quick answer

Bullet scores 73/100 (STEADY) on Hlido’s independent, hands-on test (reviewed 2026-08-21). STEADY (73) for a coherent, well-articulated coding agent with sensible efficiency protocols (model routing, targeted search, parallel execution with loop interception), cross-platform GUI+CLI, a free no-key start, and Y

Bullet is a coding agent built around one thesis: keep up with the developer. Its 'system' is three sensible protocols — Route/Escalate (send straightforward work to fast models, escalate to reasoning models only when the task demands it), Search/Acquire (targeted search and file reads instead of embedding the whole repo), and Parallel Execution (independent tool calls run together; duplicate calls and stuck loops are intercepted before they waste time). That's a coherent, credible design that maps to real pain — teams 'burning hours waiting on agent runs' — and the founders' note ('we were building with Claude Code every day, watching capable models move through needlessly slow machinery') is honest about where it came from. It ships cross-platform (macOS/Linux/Windows) as a GUI plus a free CLI (npm i -g @trybullet/cli), no key required to start, backed by Y Combinator. The thing to flag loudly: the site leads with '95.8% on SWE-bench Verified', a very strong claim that a homepage can't substantiate and that should be checked against the published results before you weight it. Speed and loop-efficiency claims are likewise the whole pitch and can only be judged by using it. For developers frustrated with slow agent runs, it's a genuinely interesting, free option — just verify the benchmark and try it on your own repo.

Why STEADY

STEADY (73) for a coherent, well-articulated coding agent with sensible efficiency protocols (model routing, targeted search, parallel execution with loop interception), cross-platform GUI+CLI, a free no-key start, and YC backing. Not higher because its headline differentiators — the 95.8% SWE-bench Verified claim and the speed/loop-efficiency promises — are exactly what a homepage cannot substantiate and must be independently verified. Not FADING because the design is credible, the product is real and shipping, and the free CLI makes it directly testable.

Public-surface checklist

What we saw

1 screenshot captured by the Hlido engine during the reviewed run (run-dc777aea094124d8-codewithbullet-com). Our own captures — not vendor marketing material.

Bullet — run screenshot 1 (home.png)
home.png

What it does well

What it fails at

Red flags

Best for

  • Developers frustrated with slow agent runs who want a speed-focused coding agent to try for free
  • People who want both a GUI and a CLI from the same agent, cross-platform
  • Anyone wanting to A/B a routing/parallel-execution approach against their current coding assistant

Not recommended for

  • Buyers who need the SWE-bench claim independently verified before adopting
  • Teams standardised on a mature coding agent with proven track record and ecosystem
  • Users who need enterprise features/support (this is a free, early product)

Compared to

Agent relevance

CLI Behavioral-testable

Agentic-Commerce Readiness 47/100 · SURFACE-ONLY

Independent readiness for agent delegation & transaction. How it’s scored · check live

A coding agent available as a cross-platform GUI and a free CLI (npm i -g @trybullet/cli, no key to start), so its behaviour is directly testable on a real repository. It is an end-user coding agent rather than an agent-callable service, but the free CLI makes hands-on evaluation straightforward.

Agent-friendly score: 7/10

Evidence

scorecard.json · transparency passport · registry · methodology

More: compare agents · best of · developer tools · incident registry

Verdict by Hlido Editor, our automated editorial system · Method: public-surface-tier-1+editorial-narrative-v2 · Methodology version 2026.05 · Next review due 2026-11-21

How this page was produced. The scores, claim verdicts and evidence come from automated hands-on testing of the product’s public surface. The written analysis is drafted by an AI system, and pages publish without a person reviewing each one. Hlido publishes this record and answers for it — tell us if anything here is wrong and we will correct it.

Embed this trust badge

Hlido trust score

Live, always-current independent score — free to embed on your site or README. No vendor pays for placement.

Markdown

[![Hlido trust score](https://hlido.eu/badge/codewithbullet.svg)](https://hlido.eu/check/?agent=codewithbullet)

HTML

<a href="https://hlido.eu/check/?agent=codewithbullet"><img src="https://hlido.eu/badge/codewithbullet.svg" alt="Hlido trust score"></a>