Devloop
Coding · tested 2026-08-14 · by the Hlido desk, not the vendor
In short: Correlates browser console errors with the backend stack trace from the same moment on one timeline, and ships the CLAUDE.md block that makes an agent actually reach for it — including the instruction never to claim a fix works without exercising it.
Quick answer
Devloop scores 79/100 (STEADY) on Hlido’s independent, hands-on test (reviewed 2026-08-14). STEADY because the product solves a real and specific problem, the surface backs it with a demo, screenshots, a one-command install, a documented enterprise fallback and a clear scope boundary on installation, and the CL Pricing: Paid.
Devloop's premise is specific and immediately recognisable to anyone who has debugged a full-stack issue: a browser console error and the dev-server stack trace from the same moment are two halves of one event, and they normally live in two windows with no shared clock. Devloop puts both on a single correlated timeline, then exposes it over MCP so an agent can start the dev server, drive the browser, reproduce an issue and read back a correlated slice of both sides. The install is one command (claude mcp add devloop), and there is a desktop 'cockpit' with tabbed browser panes, Expo iOS/Android targets, a filterable timeline, a repro builder and an element picker. The unusual move is publishing a ready-made CLAUDE.md block that tells an agent when to reach for the tool and what counts as done — 'never claim a fix works without exercising it in the running app', with 'done' defined as verified live with no console or page errors and the expected network calls succeeding. That is prompt-level distribution, and it is a smart, honest answer to the real failure mode of coding agents declaring victory from a diff. Two further details signal maturity: a documented fallback for enterprise sandboxes that block MCP (drive the tools as shell commands), and an explicit statement that Devloop diagnoses missing toolchain and names the fix command but never installs anything itself. What is missing is the operational half — no version or changelog on the surface, nothing about what the correlated timeline captures or retains, and no security note for a tool that reads console output and network traffic from a running application.
Why STEADY
STEADY because the product solves a real and specific problem, the surface backs it with a demo, screenshots, a one-command install, a documented enterprise fallback and a clear scope boundary on installation, and the CLAUDE.md distribution block is a genuine agent-adoption advantage. Held out of the top band because there is no version, changelog or data-handling statement, and because the correlation quality that the whole product rests on is shown in demo form rather than described in a way a buyer can verify.
What we saw
4 screenshots captured by the Hlido engine during the reviewed run (run-976eca187c2c09bf-devloop-build). Our own captures — not vendor marketing material.
What it does well
- Solves a specific, recognisable problem — browser and server evidence from the same moment on one timeline
- Ships a copy-paste CLAUDE.md block so agents actually reach for it, with a real definition of done
- One-command MCP install, plus a documented shell-command fallback for sandboxes that block MCP
- Explicit scope boundary: it diagnoses missing toolchain and names the fix, but never installs
- Desktop cockpit with repro builder, element picker and filterable timeline for human use
- Covers Expo iOS/Android targets alongside browser, which is broader than most tools in this class
What it fails at
- No version, changelog or release-cadence signal anywhere on the surface
- Nothing describes what the correlated timeline captures, retains or transmits
- No security note for a tool reading console output and network traffic from a running app
- Correlation accuracy — the core value — is demonstrated by video rather than described or specified
- No pricing or licence statement despite a distributed desktop application
Red flags
- The tool captures console output and network traffic from a running application, which routinely includes tokens and personal data, and no public statement describes what is captured, retained or transmitted.
- No version or changelog is published for a distributed desktop application, so a buyer cannot tell how actively it is maintained or what changed between builds.
Best for
- Developers using AI coding agents who are tired of fixes claimed from a diff rather than verified live
- Full-stack teams whose bugs span the browser/server boundary and are painful to correlate by hand
- React Native and Expo developers — mobile targets are supported alongside browser
- Teams wanting an agent-enforced definition of done tied to observed application behaviour
Not recommended for
- Production or customer-facing environments — this is development-time tooling with no documented data posture
- Teams needing a documented security review before a tool reads console and network traffic
- Backend-only projects with no browser or mobile surface to correlate against
Pricing & access
- ModelPaid
Derived from Hlido-held evidence only (engine checklist + editorial text); quotes are verbatim from the scorecard; not vendor-supplied; re-derived daily. Verify current prices on the vendor's pricing page. Last verified 2026-08-14.
Compared to
-
Feedthrough
in-page-state-vs-cross-boundary-correlation
The closest comparison — both expose live web-app debugging state over MCP. Feedthrough reaches into the page itself via an injected bridge; Devloop correlates browser and dev-server timelines and adds mobile targets plus a desktop app. Choose Feedthrough for in-page state, Devloop for cross-boundary correlation.
-
Microsoft Playwright MCP
automation-vs-diagnosis
Playwright MCP drives a browser for automation and testing; Devloop is built for diagnosis, pairing browser control with server-side log correlation. Different jobs despite overlapping browser control.
Agent relevance
CLI MCP Behavioral-testable
Agentic-Commerce Readiness 51/100 · INTEGRABLE
Independent readiness for agent delegation & transaction. How it’s scored · check live
MCP server with named tools (dev_start, browser_navigate, repro, browser_snapshot, diagnose, get_logs_around, native_open), a shared daemon mode for multiple agents, and a published CLAUDE.md rules block that tells an agent when to use each. The shell-command fallback means it still works where MCP is blocked. Distribution-aware agent design, not just an agent-compatible interface.
Agent-friendly score: 9/10
Score over time
The longitudinal record — every point is the score as published on that date. Raw series.
Evidence
- Homepage publicly accessible and value proposition clearly stated — source (2026-08-14) verified
- Pricing page discoverable in 2 clicks from homepage — source (2026-08-14)
- Documentation or live demo accessible without login — source (2026-08-14) verified
- Integration list or supported frameworks documented — source (2026-08-14) verified
- Authentication / data handling claims publicly stated — source (2026-08-14)



