Naseem
Coding · tested 2026-09-11 · re-test due 2026-12-10 · by the Hlido desk, not the vendor
In short: A Mac-native coding agent with an unusually rigorous 'prove, don't assert' philosophy and transparent BYO-key pricing — genuinely differentiated, though young and untested by us behind the download.
5 PASS · 0 FAIL of 5 public-surface claims
Quick answer
Naseem scores 79/100 (STEADY) on Hlido’s independent, hands-on test (reviewed 2026-09-11). STEADY (79) — one of the stronger public surfaces we've seen for a young agent: specific, falsifiable claims rather than vague AI marketing, a genuine architectural differentiator (native Swift + tree-first Simulator/Mac Pricing: Free tier (free entry point documented).
Naseem is a native macOS coding agent (macOS 14+), written entirely in Swift with no Electron or embedded browser, that positions itself against the heavy chat-first assistants: it acts on your actual Mac — real files, a real terminal (git, xcodebuild, package managers, tests, Python), the iOS Simulator, and even GUI-only apps via Apple's Accessibility APIs. The most distinctive part of the pitch is a verification discipline that most agents only gesture at. Naseem claims it 'proves, it doesn't assert': a 'done' that says a file was created is rejected if the file isn't on disk, a Simulator tap on a changed screen is refused, and a scheduled rule that ran no check is refused too — with the menu bar and an on-disk status file as the source of truth ('the app is the status page — no server to trust'). It pairs that with concrete engineering claims: completed tool outputs collapse to compact receipts so the prompt stays bounded over hundreds of steps, a byte-identical request prefix to exploit provider prompt caching, memory that survives restarts, crash-safe resumption, and a stated-cost window. Commercially it is honest and low-friction: a free tier with no account and no subscription (bring your own AI key), and a one-time Pro license (launch price $29, normally $49). The caveats are about maturity and verifiability rather than honesty — it is a young product (v1.4.24) whose public GitHub repo and Product Hunt presence are encouraging but not a track record, it is Mac-only, and Hlido reviewed the landing surface without downloading and running the app, so the strong verification and context-bounding claims are credible-on-their-face but not independently confirmed here.
Why STEADY
STEADY (79) — one of the stronger public surfaces we've seen for a young agent: specific, falsifiable claims rather than vague AI marketing, a genuine architectural differentiator (native Swift + tree-first Simulator/Mac automation), a rigorous verification philosophy, an open GitHub repo, and transparent low-friction pricing. Held below VITAL because it is a young product with no long-run track record, Mac-only, and — being a downloadable native app — outside what Hlido's public-surface engine can functionally verify, so the headline verification and context claims remain unconfirmed by us. Confidence medium.
Public-surface checklist
- PASS Homepage loads (required)
- PASS Primary value prop (required) — 'The Mac-native AI agent that does the work, not just the chat'
- PASS Cta present (required) — 'Download for macOS'
- PASS Pricing or access — Free plan (BYO key, no account); Pro one-time $29 launch / $49
- PASS Evidence or demo — Multiple in-action demos (iOS Simulator build/debug, parallel sub-agents) + open GitHub repo
What it does well
- Acts on the real machine — files, a real terminal, the iOS Simulator, and GUI-only Mac apps via Accessibility APIs — not a sandbox
- Rigorous 'prove, don't assert' verification: claimed actions are rejected unless the on-disk/UI state confirms them
- Native Swift, no Electron — stays light while Xcode, the Simulator, and the model do the heavy work
- Engineered for long tasks: bounded context via collapsed receipts, cache-friendly stable prefix, restart-surviving memory, crash-safe resume
- Honest, low-friction commercials: free tier, no account, bring-your-own key, one-time Pro license with visible cost tracking
What it fails at
- macOS 14+ only — no Windows, Linux, or mobile
- Young product (v1.4.24) with no long-run reliability track record yet
- Bring-your-own-key means the user carries model cost and setup
- No public API/MCP surface — it's an end-user agent, not a component other agents can drive
- Strong verification and context-bounding claims are not independently confirmed on the public surface
Best for
- macOS developers who want an agent that actually builds, runs, tests, and debugs on their real machine
- iOS developers who want native Simulator driving (build, install, launch, tap, screenshot)
- Engineers wary of Electron-heavy assistants who value a lightweight native app
- Users who want to bring their own model key and avoid subscriptions
- Anyone who values an agent that verifies its own 'done' rather than asserting it
Not recommended for
- Windows, Linux, or mobile users (Mac-only)
- Teams needing a programmatic/agent-drivable API or MCP surface
- Buyers who require a proven, multi-year reliability record before adoption
- Users who want a fully managed, no-key, hosted experience
Pricing & access
- ModelFree tier
- Free entry pointYes — a free tier or open-source edition is documented
- Price points we recorded
…subscription (bring your own AI key), and a one-time Pro license (launch price $29, normally $49). The caveats are about maturity and verifiability rather than…
- Pricing findable on the public surfacePASS Free plan (BYO key, no account); Pro one-time $29 launch / $49 (tested 2026-09-11)
Derived from Hlido-held evidence only (engine checklist + editorial text); quotes are verbatim from the scorecard; not vendor-supplied; re-derived daily. Verify current prices on the vendor's pricing page. Last verified 2026-09-11.
Related agents
Agent relevance
No programmatic surfaces
Naseem is itself an end-user agent harness for macOS (LLM + harness + your Mac), driven by a human via a native app with bring-your-own model key. It exposes no public API, CLI, or MCP for another agent to drive it programmatically. Its agentic value is as a capable local coding/automation agent, not as an integration component in a larger agent pipeline.
Agent-friendly score: 3/10
Score over time
The longitudinal record — every point is the score as published on that date. Raw series.
Evidence
- Mac-native (macOS 14+) AI agent that acts on files, terminal, and apps — source (2026-09-11) verified
- Drives the iOS Simulator natively via the accessibility tree; automates GUI-only Mac apps via Accessibility APIs — source (2026-09-11) verified
- 'Proves, doesn't assert' — rejects a claimed action unless on-disk/UI state confirms it — source (2026-09-11) verified
- Free plan (no account, no subscription, BYO key); Pro one-time license $29 launch / $49 — source (2026-09-11) verified
- Written in Swift, no Electron; open GitHub repo; launched on Product Hunt; version 1.4.24 — source (2026-09-11) verified