CodeariaAcademy
GitHub repositorycommit this monthDevelopmentAutomation

TesterArmy e2e

8.1k

A test step is a sentence like "upgrade the workspace to the Pro plan", and an agent finds the buttons itself on a website, an iPhone or an Android app. Successful actions are recorded and replayed later without calling the model. Open source, Apache-2.0.

We have not run it ourselves: described from the docs and the repository.

In numbers

GitHub stars as of 05.10.2026
on the daily GitHub Trending list on 05.10.2026
browser engines: Chromium, Firefox, WebKit
licence

What it is

A classic end-to-end test knows exactly what it verified, and you pay for that precision by rewriting a selector every time someone renames a button or changes a user flow. e2e from TesterArmy lets you describe the goal of a step in words, agent.act('upgrade the workspace to the Pro plan'), and the agent works out the path on its own, so a reworked page is far less likely to break the test. The same test keeps ordinary locators and expect calls, the way Playwright does, so the things that matter are still checked precisely.

What sets it apart from Playwright MCP, its neighbour in this catalogue, is the job. There an agent gets a browser to do work; here it gets a test suite you run before every release. Hence the cache: a confirmed action is replayed next time without the model, so your token spend follows how often the UI changes rather than how often the tests run. It suits anyone building a site or an app with an agent who wants checkout and sign-up clicked through automatically before every release.

How it works

npx e2e init
The wizard asks for the engine (Web or Mobile) and a model provider, and None is an option for tests with no AI. It writes the config and a sample test; run it with npx e2e run.
agent.act · agent.assert
act carries out a goal described in words, assert checks a condition in words, extract pulls data off the screen, waitFor waits for a state. Plain expect(screen.getByRole(...)) sits alongside.
.e2e/cache
An action confirmed by the next check is recorded and replayed next time without the model. If the screen has changed, a live agent takes the step over. Checks in words (assert, extract, waitFor) call the model every time.
web · iOS · Android
Browsers through Playwright, iOS simulators and Android emulators, plus a real phone plugged into your machine. One test runs once per target, and the report lands as a pull request comment.
npx skills add
npx skills add tester-army/e2e installs a skill and an MCP server for Claude Code, Cursor and Codex. And e2e explore clicks through the app hunting for bugs before a PR.

Know before installing

A Claude subscription is not supported: Claude works only through an API key or a model in your Copilot plan. Without a key you can sign in with a ChatGPT, Copilot, OpenCode or SuperGrok subscription.
On Windows it runs only inside WSL, and iOS tests need macOS with Xcode. Desktop apps are not covered yet.
Version 0.17.0 as of 05.10.2026, still pre-1.0: the API and config can change between minor releases. Versions move fast, so check the current one. Telemetry is on by default and turns off with one command.

Comments