e2e (tester-army/e2e), an Apache-2.0 AI end-to-end testing framework, gained 1,430 stars today to top GitHub Trending (3,828 total; 60,933 npm downloads last week). Tests mix natural-language agent.act/agent.assert steps with locator assertions; verified agent steps are recorded and replayed with no model calls until the app changes. v0.17.0 adds Copilot Responses-API models, OpenCode Console sign-in, and a decision-model executor for bounded actions.
Key Takeaways
- ✓#1 on GitHub daily trending: +1,430 stars today, 3,828 total (Apache-2.0)
- ✓npm package e2e: 60,933 weekly downloads (Sep 28–Oct 4); latest 0.17.0 released Oct 4
- ✓Record & replay: verified agent steps rerun with zero model calls until the app changes
- ✓v0.16.0 cut install footprint from 117 to 29 packages (~36MB to ~31MB)
- ✓7 packages: Playwright (Chromium/Firefox/WebKit), iOS/Android, PR reporter, hosted browsers/simulators, decision-model executor

Developer 3s Key Decision Metrics
Heavy Claude Code use: compare subscription limits and API bills
Compare 29+ dev plans & simulate token costs vs $20/mo subscriptions
Project Links & Resources
Direct AccessIn-Depth Technical Analysis
Background
Coding agents ship code faster than teams can write brittle Playwright/Cypress selectors, while pure-LLM 'AI testing' is costly and non-deterministic. TesterArmy's open-source e2e (Apache-2.0) gained 1,430 stars today to hit #1 on GitHub daily trending (3,828 total).
How it works
- Hybrid tests: natural-language
agent.act()/agent.assert()steps sit next to deterministic locator assertions in the same test. - Record & replay: an agent step that a later assertion verifies is recorded; the next run replays it with no model calls until the app changes. v0.17.0 hardens replay keys (agent + redacted-context hash), requires observable effects, and
--strict-cachefails stale recordings withREPLAY_STALE. - Engines:
@e2e-dev/web(Playwright: Chromium/Firefox/WebKit),@e2e-dev/mobile(iOS/Android simulators), plus a GitHub PR reporter, Kernel hosted browsers, and EAS hosted simulators. - Decision-model executor:
@e2e-dev/decisionuses a decision model to pick operations/elements from probability distributions, with a small text model only for field values.
Numbers
No official accuracy/latency benchmark is published. Verifiable figures: 3,828 stars (+1,430 today); 60,933 npm weekly downloads (Sep 28–Oct 4); v0.16.0 cut install deps from 117 to 29 packages (~36MB to ~31MB). Pre-1.0: APIs may change between minors.
Getting started
Run npx e2e init, pick web or mobile and a model source (subscription, API key, or local model; tests without agent steps need no model). v0.17.0 adds GitHub Copilot Responses-API models and OpenCode Console sign-in. Docs ship inside node_modules/e2e/docs for coding agents, with migration guides from Playwright, Cypress, Selenium, Detox, and Maestro. Anonymous CLI telemetry is on by default (E2E_TELEMETRY_DISABLED=1 to opt out); docs recommend Node.js 24.8+ (22.22.3+ on Node 22), and WSL on Windows.
Benchmark side-by-side against alternatives, or calculate monthly token cost vs subscription break-even.
Discussion & Comments
0Sign in to join the discussion
Connect with AI developers to exchange benchmark insights.