e2e (tester-army/e2e), an Apache-2.0 AI end-to-end testing framework, gained 1,430 stars today to top GitHub Trending (3,828 total; 60,933 npm downloads last week). Tests mix natural-language agent.act/agent.assert steps with locator assertions; verified agent steps are recorded and replayed with no model calls until the app changes. v0.17.0 adds Copilot Responses-API models, OpenCode Console sign-in, and a decision-model executor for bounded actions.

Key Takeaways

  • ✓#1 on GitHub daily trending: +1,430 stars today, 3,828 total (Apache-2.0)
  • ✓npm package e2e: 60,933 weekly downloads (Sep 28–Oct 4); latest 0.17.0 released Oct 4
  • ✓Record & replay: verified agent steps rerun with zero model calls until the app changes
  • ✓v0.16.0 cut install footprint from 117 to 29 packages (~36MB to ~31MB)
  • ✓7 packages: Playwright (Chromium/Firefox/WebKit), iOS/Android, PR reporter, hosted browsers/simulators, decision-model executor
TesterArmy's open-source e2e tops GitHub daily trending: natural-language agent steps for web and mobile tests, replayed with zero model calls
🖼️Official Media
Click to view high-res
🧭

Heavy Claude Code use: compare subscription limits and API bills

Compare 29+ dev plans & simulate token costs vs $20/mo subscriptions

🔬

In-Depth Technical Analysis

Background

Coding agents ship code faster than teams can write brittle Playwright/Cypress selectors, while pure-LLM 'AI testing' is costly and non-deterministic. TesterArmy's open-source e2e (Apache-2.0) gained 1,430 stars today to hit #1 on GitHub daily trending (3,828 total).

How it works

  • Hybrid tests: natural-language agent.act() / agent.assert() steps sit next to deterministic locator assertions in the same test.
  • Record & replay: an agent step that a later assertion verifies is recorded; the next run replays it with no model calls until the app changes. v0.17.0 hardens replay keys (agent + redacted-context hash), requires observable effects, and --strict-cache fails stale recordings with REPLAY_STALE.
  • Engines: @e2e-dev/web (Playwright: Chromium/Firefox/WebKit), @e2e-dev/mobile (iOS/Android simulators), plus a GitHub PR reporter, Kernel hosted browsers, and EAS hosted simulators.
  • Decision-model executor: @e2e-dev/decision uses a decision model to pick operations/elements from probability distributions, with a small text model only for field values.

Numbers

No official accuracy/latency benchmark is published. Verifiable figures: 3,828 stars (+1,430 today); 60,933 npm weekly downloads (Sep 28–Oct 4); v0.16.0 cut install deps from 117 to 29 packages (~36MB to ~31MB). Pre-1.0: APIs may change between minors.

Getting started

Run npx e2e init, pick web or mobile and a model source (subscription, API key, or local model; tests without agent steps need no model). v0.17.0 adds GitHub Copilot Responses-API models and OpenCode Console sign-in. Docs ship inside node_modules/e2e/docs for coding agents, with migration guides from Playwright, Cypress, Selenium, Detox, and Maestro. Anonymous CLI telemetry is on by default (E2E_TELEMETRY_DISABLED=1 to opt out); docs recommend Node.js 24.8+ (22.22.3+ on Node 22), and WSL on Windows.

Action HubReady to adopt this in production?

Benchmark side-by-side against alternatives, or calculate monthly token cost vs subscription break-even.