Open source developer CLI tools, proven to build and run
Command-line tools that build and run out of the box, or do not. Argusic followed each project's own instructions on a blank Linux machine, and the table records the result with a link to the session log.
Tested between and . Each row shows its own test date; a project can change after that day.
164 of 173 tested projects run. 111 more waiting for a test.
In short: 143 of the 173 tested projects started as-is on a fresh machine: lazygit, career-ops, ripgrep, bat, herdr, coreutils, chalk, and ratatui, and 135 more. 21 more started once a stand-in replaced a service they expect, such as a database: gh-dash, shell_gpt, App-Store-Connect-CLI, coreutils, petdex, dockly, jevgrep, and snapai, and 13 more. 9 could not be verified: go-recipes, httm, TUI-ConsoleLauncher, CodeMachine-CLI, intelligent-terminal, Codewhale, kyanos, and agent-of-empires, and 1 more; the log shows where each one stopped.
Measured by Argusic on a fresh machine every time. Every number links to its evidence.
| # | project | verdict | Argusic Score | language | stars | tested on |
|---|---|---|---|---|---|---|
| 1 | lazygit | Runs | 100 / 100 | Go | 82,996 | |
simple terminal UI for git commands What the test found: lazygit builds, all 32 unit-test packages pass, all integration tests pass headlessly, and the TUI binary launches and runs on the virtual display. 12 minutes. | ||||||
| 2 | career-ops | Runs | 100 / 100 | JavaScript | 73,771 | |
Open-source AI job search agent and job finder: scan job boards, score each job 1-5 against your CV before you apply, tailor an ATS-friendly resume and cover... What the test found: All core tests pass on Node 22 with --experimental-sqlite. Playwright Chromium 151.0.7922.34 is installed and launches. The doctor and update-system checks work. The one remaining failure is in web/tests/lib/apply-cv-resolver.test.mjs which needs --experimental-strip-types for .ts imports. 66 minutes. | ||||||
| 3 | ripgrep | Runs | 100 / 100 | Rust | 68,925 | |
ripgrep recursively searches directories for a regex pattern while respecting your gitignore What the test found: ripgrep 15.2.0 built from source, all 332 tests pass, binary searches correctly. 3 minutes. | ||||||
| 4 | bat | Runs | 100 / 100 | Rust | 60,702 | |
A cat(1) clone with wings. What the test found: bat 0.26.1 builds from source, all 293 tests pass, and the binary runs and syntax-highlights input correctly. 5 minutes. | ||||||
| 5 | herdr | Runs | 100 / 100 | Rust | 42,896 | |
the runtime your coding agents live on What the test found: herdr 0.8.2 builds, runs its server, passes all Rust unit tests, Python maintenance tests, Bun integration tests, and plugin marketplace tests with no failures. 19 minutes. | ||||||
| 6 | coreutils | Runs | 100 / 100 | Rust | 24,224 | |
Cross-platform Rust rewrite of the GNU coreutils What the test found: uutils coreutils 0.11.0 builds cleanly with Rust 1.98.1 on Ubuntu 24.04, produces a working multi-call binary (verified: --version, --help, cat, sort, echo, true, false, ls, head, md5sum, wc), and passes 5078 of 5129 functional tests, all 51 failures are container-environment issues (no TTY colors, no /dev devices... 24 minutes. | ||||||
| 7 | chalk | Runs | 100 / 100 | JavaScript | 23,322 | |
๐ Terminal string styling done right What the test found: chalk v6.0.0 installs, builds, imports as ESM, and passes all 58 tests with Node 22. The npm registry entry validates as published. 4 minutes. | ||||||
| 8 | ratatui | Runs | 100 / 100 | Rust | 22,902 | |
A Rust crate for cooking up terminal user interfaces (TUIs) ๐จโ๐ณ๐ https://ratatui.rs What the test found: The Ratatui workspace builds cleanly and all 3176 tests pass on Rust 1.98.1 (stable); interactive examples require a real TTY which is unavailable in the container. 8 minutes. | ||||||
| 9 | witr | Runs | 100 / 100 | Go | 22,627 | |
Why is this running? Trace any process, port, container, or file back to what started it - CLI + TUI. What the test found: witr builds, passes all 8 test packages (including -race), and traces real processes in the container. 4 minutes. | ||||||
| 10 | nnn | Runs | 100 / 100 | C | 22,057 | |
nยณ The unorthodox terminal file manager What the test found: nnn v5.3 builds from source and runs as a terminal file manager, successfully listing directories, navigating into subdirectories, and returning; the test harness confirms all assertions pass. 8 minutes. | ||||||
| 11 | vhs | Runs | 100 / 100 | Go | 21,071 | |
Your CLI home video recorder ๐ผ What the test found: VHS builds, all 4 Go packages pass tests (lexer, parser, token, main), and runs end-to-end against a real auto-downloaded Chromium browser + real ttyd + real ffmpeg, producing valid GIF and text output files. 16 minutes. | ||||||
| 12 | ast-grep | Runs | 100 / 100 | Rust | 16,141 | |
โกA CLI tool for code structural search, lint and rewriting. Written in Rust What the test found: ast-grep 0.45.3 installed via npm, pip, and cargo build from source; all 202 unit/integration tests pass; search, scan, rewrite, and test subcommands verified working on real code. 9 minutes. | ||||||
| 13 | bottom | Runs | 100 / 100 | Rust | 14,090 | |
Yet another cross-platform graphical process/system monitor. What the test found: The bottom system monitor compiles from source, reports its version, and passes all 342 non-ignored tests on Linux x86_64. 8 minutes. | ||||||
| 14 | ccstatusline | Runs | 100 / 100 | TypeScript | 13,222 | |
๐ Beautiful highly customizable statusline for Claude Code CLI with powerline support, themes, and more. What the test found: Dependencies installed, project builds to dist/ccstatusline.js, all 1897 tests pass, and piped JSON input renders a formatted status line with model name and git info. 6 minutes. | ||||||
| 15 | broot | Runs | 100 / 100 | Rust | 13,053 | |
A new way to see and navigate directory trees What the test found: The Rust project broot v1.59.0 builds from source with cargo build, the broot binary runs and reports version and help, and all 132 tests pass (125 unit tests, 7 integration tests, 0 failures). 3 minutes. | ||||||
| 16 | console | Runs | 100 / 100 | PHP | 9,814 | |
Eases the creation of beautiful and testable command line interfaces What the test found: Symfony Console component library installed with composer dependencies; PHP 8.4.26 runs all 2036 unit tests; 1997 pass, 25 skipped, 17 known environmental issues unrelated to the library code. 33 minutes. | ||||||
| 17 | Graft | Runs | 100 / 100 | TypeScript | 9,771 | |
Turbocharge Claude Code, Cursor, Codex, Gemini & every coding agent: faster, cheaper, with contextual understanding specific to your codebase. What the test found: npm install completes, TypeScript compiles, the CLI builds/checks/queries a structural code graph on a real repo without any keys, and all 1220 tests run with 0 failures. 10 minutes. | ||||||
| 18 | npkill | Runs | 100 / 100 | TypeScript | 9,464 | |
List any node_modules ๐ฆ dir in your system and how heavy they are. You can then select which ones you want to erase to free up space ๐งน What the test found: npm install completes, TypeScript compiles cleanly, 236/236 unit tests pass, and the npkill CLI binary runs (version 0.12.2, help text, and JSON scan mode all verified). 2 minutes. | ||||||
| 19 | bubbles | Runs | 100 / 100 | Go | 8,978 | |
TUI components for Bubble Tea ๐ซง What the test found: The project compiles and all tests pass using the standard Go toolchain with Charm's Go 1.26.0 bootstrap. 7 minutes. | ||||||
| 20 | presenterm | Runs | 100 / 100 | Rust | 8,904 | |
A markdown terminal slideshow tool What the test found: presenterm 0.16.1 builds, all 520 unit tests pass, non-interactive commands (--help, --version, --current-theme, --list-comment-commands, --acknowledgements, --export-html) function correctly; HTML export of demo presentation produces valid 324KB output. 7 minutes. | ||||||
| 21 | websocat | Runs | 100 / 100 | Rust | 8,710 | |
Command-line client for WebSockets, like netcat (or curl) for ws:// with advanced socat-like functions What the test found: websocat 1.14.1 compiles from source, binary serves WebSocket connections and passes all 7 integration tests (trivial, tcp, ws, ws_ll, ws_persist, unix, abstract_). 2 minutes. | ||||||
| 22 | jc | Runs | 100 / 100 | Python | 8,691 | |
CLI tool and python library that converts the output of popular command-line tools, file-types, and common strings to JSON, YAML, or Dictionaries. This allows... What the test found: jc v1.25.7 is installed in a Python 3.12 venv at /tmp/jc_venv, all 1557 non-skipped tests pass, and the CLI parses command output to JSON correctly. 1 minute. | ||||||
| 23 | grex | Runs | 100 / 100 | Rust | 8,216 | |
A command-line tool and Rust library with Python bindings for generating regular expressions from user-provided test cases What the test found: grex 1.4.6 builds from source, its CLI prints version 1.4.6 and generates correct regexes from stdin input, all 693 Rust tests and 23 Python tests pass, and the Python extension wheel installs and builds regexes via RegExpBuilder. 4 minutes. | ||||||
| 24 | xh | Runs | 100 / 100 | Rust | 8,122 | |
Friendly and fast tool for sending HTTP requests What the test found: xh 0.26.2 builds with cargo, cargo test passes 281 tests (1 ignored: expired test certificate), and the installed binary answered HTTP requests against a local Python server (GET 200 with JSON body, POST echo with JSON object, --curl translation). 5 minutes. | ||||||
| 25 | miniserve | Runs | 100 / 100 | Rust | 7,895 | |
๐ For when you really just want to serve some files over HTTP right now! What the test found: miniserve 0.35.0 builds from source, runs all 276 integration tests across 16 modules, serves HTTP on localhost, and generates valid tar/zip archives with correct content. 67 minutes. | ||||||
| 26 | himalaya | Runs | 100 / 100 | Rust | 7,407 | |
CLI to manage emails What the test found: Rust toolchain 1.99.0 installed, himalaya v2.1.0 built from source with all default backends (imap, smtp, jmap, gmail, msgraph, maildir, mbox, sieve, pimdir), binary executes and reports version, full test suite of 161 tests passes cleanly. 3 minutes. | ||||||
| 27 | sd | Runs | 100 / 100 | Rust | 7,386 | |
Intuitive find & replace CLI (sed alternative) What the test found: sd CLI version 1.0.0 builds, passes all tests, and correctly performs find-and-replace operations from stdin and in-place file modification. 2 minutes. | ||||||
| 28 | ANUS | Runs | 100 / 100 | JavaScript | 6,550 | |
A free coding agent in your terminal. It runs on the smartest free model that is up today. What the test found: The ANUS CLI installs cleanly, all 28 node --test tests pass, and the --version flag returns 0.2.2. 2 minutes. | ||||||
| 29 | pastel | Runs | 100 / 100 | Rust | 6,516 | |
A command-line tool to generate, analyze, convert and manipulate colors What the test found: pastel v0.12.0 built from source, all 74 tests pass, binary executes and correctly converts colors between formats. 2 minutes. | ||||||
| 30 | alive-progress | Runs | 100 / 100 | Python | 6,315 | |
A new kind of Progress Bar, with real-time throughput, ETA, and very cool animations! What the test found: alive-progress 3.3.0 is installed in a venv, passes all 563 tests, and the progress bar renders correctly via alive_bar. 1 minute. | ||||||
| 31 | viddy | Runs | 100 / 100 | Rust | 5,425 | |
๐ A modern watch command. Time machine and pager etc. What the test found: viddy builds in debug and release profiles, all 35 unit tests pass, the help and version commands respond, and the TUI launches and runs under a pseudo-terminal (script) without panicking or errors. 10 minutes. | ||||||
| 32 | colorls | Runs | 100 / 100 | Ruby | 5,140 | |
A Ruby gem that beautifies the terminal's ls command, with color and font-awesome icons. :tada: What the test found: colorls gem 1.5.0 installed from source, all 40 integration tests and 90/91 RSpec tests pass, binary lists directory contents with colorized icons. 19 minutes. | ||||||
| 33 | gitlogue | Runs | 100 / 100 | Rust | 5,089 | |
A cinematic Git commit replay tool for the terminal, turning your Git history into a living, animated story. What the test found: The gitlogue binary builds from source, all 228 unit/integration tests pass, and the CLI responds with correct help text and version 0.11.0. 4 minutes. | ||||||
| 34 | tuios | Runs | 100 / 100 | Go | 5,030 | |
A terminal window manager that knows what your agents are doing. Tiling panes, workspaces, sessions that survive restarts, and one Inbox for every coding agent. What the test found: tuios and tuios-web binaries build from source, the full Go test suite (49 packages) passes, and the daemon creates sessions, accepts typed input via send-text/send-keys, and captures pane output showing shell command execution. 22 minutes. | ||||||
| 35 | fallow | Runs | 100 / 100 | Rust | 5,021 | |
Codebase intelligence for TypeScript and JavaScript. Health, complexity hotspots, duplication, architecture boundaries, circular dependencies, design-system... What the test found: Fallow CLI installs via npm and cargo build, runs on this repo producing valid analysis output, and passes 17018 tests across all workspace crates. 58 minutes. | ||||||
| 36 | sqlit | Runs | 100 / 100 | Python | 4,881 | |
A user friendly TUI for SQL databases. Written in python. Supports SQL server, Mysql, PostreSQL, SQLite, Turso and more. What the test found: sqlit is installed in a venv at /home/runner/.venv, the CLI and TUI launch correctly, the SQLite test suite passes with 2018 passing tests, and CLI queries against a real SQLite file return correct JSON and CSV output. 18 minutes. | ||||||
| 37 | YouPlot | Runs | 100 / 100 | Ruby | 4,859 | |
A command line tool that draw plots on the terminal. What the test found: YouPlot 0.5.0 installed and working: uplot bar/hist/line/scatter/density/boxplot all produce terminal plots; 109 unit tests pass with 0 failures. 10 minutes. | ||||||
| 38 | tmuxp | Runs | 100 / 100 | Python | 4,586 | |
๐ฅ๏ธ Session manager for tmux, built on libtmux. What the test found: tmuxp 1.74.0 is installed in a Python venv at ~/venv, tmux 3.4 extracted from Debian packages runs via LD_LIBRARY_PATH and PATH, all 770 tests pass (2 skipped: suppress_history test and AI-generated skip), tmuxp load creates sessions from YAML workspaces, tmuxp freeze snapshots them, and debug-info reports all... 11 minutes. | ||||||
| 39 | zsh-vi-mode | Runs | 100 / 100 | Shell | 4,460 | |
๐ป A better and friendly vi(vim) mode plugin for ZSH. What the test found: The zsh-vi-mode ZSH plugin v0.12.0 sources cleanly in an interactive ZSH 5.9 with ZLE modules, creating vi-mode keymaps with 95 normal-mode, 43 insert-mode, 19 visual-mode, and 13 operator-pending-mode bindings, plus 103 zvm_* functions for all documented features. 6 minutes. | ||||||
| 40 | gptme | Runs | 100 / 100 | Python | 4,443 | |
Your agent in your terminal, equipped with local tools: writes code, uses the terminal, browses the web. Make your own persistent autonomous agent on top! What the test found: gptme 0.34.0 installed in a Poetry venv; `gptme --version` returns 0.34.0+unknown; `gptme-doctor` reports 22 checks passed, system operational; 1746 out of 1746 core tests pass. 34 minutes. | ||||||
| 41 | doxx | Runs | 100 / 100 | Rust | 3,771 | |
Expose the contents of .docx files without leaving your terminal. Fast, safe, and smart, no Office required! What the test found: doxx 0.1.4 builds and runs successfully on Ubuntu 24.04 with Rust 1.99.0. All 144 tests pass. The binary exports .docx files to text, markdown, JSON, CSV, and ANSI formats correctly. NO_COLOR=1 env var must be cleared for color-dependent tests. 9 minutes. | ||||||
| 42 | curlie | Runs | 100 / 100 | Go | 3,730 | |
The power of curl, the ease of use of httpie. What the test found: curlie builds from source, passes its args tests, and makes real HTTP requests to the internet returning correct responses. 17 minutes. | ||||||
| 43 | abtop | Runs | 100 / 100 | Rust | 3,707 | |
Like htop, but for AI coding agents. Monitor Claude Code & Codex CLI sessions, tokens, context window, rate limits, and ports in real-time. What the test found: abtop 0.5.5 builds and passes all 222 tests; --once and --json modes produce correct live terminal-monitor snapshots from local process/file state. 9 minutes. | ||||||
| 44 | walk | Runs | 100 / 100 | Go | 3,642 | |
Terminal file manager What the test found: walk v1.13.0 builds from source, passes its 2 Go tests, answers --version and --help, and launches as a TUI under a pseudo-terminal. 22 minutes. | ||||||
| 45 | Surge | Runs | 100 / 100 | Go | 3,586 | |
Blazing fast TUI download manager built in Go for power users What the test found: Surge builds from source and all 26 test packages pass cleanly with zero failures. 21 minutes. | ||||||
| 46 | hl | Runs | 100 / 100 | Rust | 3,302 | |
A fast and powerful log viewer and processor that converts JSON logs or logfmt logs into a clear human-readable format. What the test found: The hl binary v0.36.3 builds from source, all 1796 unit tests pass, and the tool correctly parses and renders JSON log files into human-readable output with working field filtering and level filtering. 6 minutes. | ||||||
| 47 | viu | Runs | 100 / 100 | Rust | 3,289 | |
Terminal image viewer with native support for iTerm and Kitty What the test found: viu 1.6.1 builds, installs, and runs; 1 unit test passes; binary displays images as ANSI terminal output. 19 minutes. | ||||||
| 48 | organize | Runs | 100 / 100 | Python | 3,158 | |
The file management automation tool. What the test found: The project installs with pip, its full test suite of 266 tests passes, and the CLI runs and reports v3.3.0. 2 minutes. | ||||||
| 49 | asciigraph | Runs | 100 / 100 | Go | 3,100 | |
Go package to make lightweight ASCII line graph โญโโฏ in command line apps with no other dependencies. What the test found: Go 1.22 extracted from debs; asciigraph library compiles and all 27 test groups pass; the CLI binary /tmp/asciigraph builds and renders line graphs from stdin data. 7 minutes. | ||||||
| 50 | dnote | Runs | 100 / 100 | Go | 3,087 | |
A simple command line notebook What the test found: Dnote CLI and server compile from source, their test suites pass entirely, and the server starts serving HTTP on port 3001 with a working /health endpoint. 12 minutes. | ||||||
| 51 | opensrc | Runs | 100 / 100 | Rust | 3,010 | |
Fetch source code for npm packages to give AI coding agents deeper context What the test found: The opensrc CLI and docs site both install, build, and run. CLI fetches source from npm, PyPI, and crates.io and provides cached paths for grep/search. Docs site builds and serves on port 3456 returning HTTP 200. Rust compilation tests fail due to missing cargo (no root), but the pre-built native binary is downloaded... 13 minutes. | ||||||
| 52 | readme-ai | Runs | 100 / 100 | Python | 3,002 | |
README file generator, powered by AI. What the test found: readmeai installs in a venv, passes its full 373-test suite, and generates README output files via the CLI in offline mode. The pyproject.toml dependency bounds were relaxed from caret to minimum versions to avoid Rust-compiler requirements for newer prebuilt wheels. 5 minutes. | ||||||
| 53 | fast-cli | Runs | 100 / 100 | TypeScript | 2,884 | |
Test your download and upload speed using fast.com What the test found: npm install and tsc build succeed; CLI help and download speed measurement work end-to-end against fast.com; upload speed tests time out due to fast.com rate-limiting from the container IP. 56 minutes. | ||||||
| 54 | tach | Runs | 100 / 100 | Rust | 2,835 | |
A Python tool to visualize + enforce dependencies, using modular architecture ๐ Open source ๐ Installable via pip ๐ง Able to be adopted incrementally - โก... What the test found: tach 0.35.1 is fully installed and working: pip install from PyPI succeeds, the Rust extension loads, dependency checking passes on the repo itself, and both the full Python test suite (88/88) and Rust test suite (70/70) pass. 8 minutes. | ||||||
| 55 | iredis | Runs | 100 / 100 | Python | 2,762 | |
Interactive Redis: A Terminal Client for Redis with AutoCompletion and Syntax Highlighting. What the test found: IRedis installed, all 791 tests pass against a real Redis 7.2 server, and non-interactive CLI produces correct output (PING -> PONG). 7 minutes. | ||||||
| 56 | dekit | Runs | 100 / 100 | Rust | 2,746 | |
Process manager for dev and prod What the test found: dekit v0.10.0 builds from source on x86_64 Linux; all 359 tests pass and the CLI (help, up --help) produces valid output. 26 minutes. | ||||||
| 57 | webcmd | Runs | 100 / 100 | TypeScript | 2,661 | |
Self-learning agent browser What the test found: Webcmd 0.8.4 is installed, builds from source, runs unit tests (3571 passed), connects to Cloak browser daemon, and responds to CLI commands. 5 minutes. | ||||||
| 58 | mpb | Runs | 100 / 100 | Go | 2,513 | |
multi progress bar for Go cli applications What the test found: Built and installed Go 1.26.0, compiled the project, and ran 250 tests, all passing. 5 minutes. | ||||||
| 59 | MovieBox-Tui | Runs | 100 / 100 | Rust | 2,455 | |
Terminal interface to find, download, and stream movies, TV shows, and live TV using local media players. What the test found: MovieBox-TUI 0.1.26 builds from source on Linux x86_64, all 394 unit tests pass, and the binary responds correctly to --help and --version flags. 21 minutes. | ||||||
| 60 | oh-my-mermaid | Runs | 100 / 100 | TypeScript | 2,297 | |
Turn complex codebases into clear, navigable architecture diagrams with Claude Code. What the test found: The oh-my-mermaid package installed, built, passed all 37 tests, and all CLI commands (init, list, status, config, read, tree, validate, diff, refs, setup) produce correct output. 1 minute. | ||||||
| 61 | tabulate | Runs | 100 / 100 | C++ | 2,177 | |
Table Maker for Modern C++ What the test found: tabulate builds with CMake from source, all 18 tests pass (empty_rows, format_propagation, nested_tables, column_operations, alignment_and_wrapping, trim_mode, iterators, border_corner_styles, unicode_borders with single_header variants), and sample binaries produce correct table output. 2 minutes. | ||||||
| 62 | serie | Runs | 100 / 100 | Rust | 2,135 | |
A rich git commit graph in your terminal, like magic ๐ What the test found: Rust 1.99.0 installed, dependencies fetched, project compiles with cargo build, all 94 unit tests pass, binary prints version and help. TUI cannot initialize without /dev/tty (expected in headless CI). 5 minutes. | ||||||
| 63 | terminal-code | Runs | 100 / 100 | TypeScript | 2,132 | |
VS Code in the terminal What the test found: Build succeeds with Node 22, all 118 tests pass, and `tode --serve` starts a code-server instance that responds to HTTP requests. 7 minutes. | ||||||
| 64 | circumflex | Runs | 100 / 100 | Go | 2,095 | |
๐ฟ It's Hacker News in your terminal What the test found: circumflex (clx) builds, runs, passes all 29 test suites, and responds to --version (5.1-dev) and help commands. 6 minutes. | ||||||
| 65 | sad | Runs | 100 / 100 | Rust | 2,045 | |
CLI search and replace | Space Age seD What the test found: sad v0.4.32 builds from source with cargo, its 2 unit tests pass, and the binary performs batch file edits: piping a file path to 'sad PATTERN REPLACE --commit' writes the edited file in-place. 2 minutes. | ||||||
| 66 | ov | Runs | 100 / 100 | Go | 2,033 | |
๐Feature-rich terminal-based text viewer. It is a so-called terminal pager. What the test found: The ov pager compiles, all unit tests pass, the binary runs and responds to --version, --help, --generate-config, and pipe input via --quit-if-one-screen. 6 minutes. | ||||||
| 67 | graphql-cli | Runs | 100 / 100 | TypeScript | 2,017 | |
๐ Command line tool for common GraphQL development workflows What the test found: graphql-cli v5.0.0 builds all 6 core packages, the CLI binary works (--help, --version, discover all succeed), and the integration test passes using real generate and codegen commands against a local schema project. 15 minutes. | ||||||
| 68 | TokenTracker | Runs | 100 / 100 | JavaScript | 1,996 | |
Local-first AI token usage & cost tracker for 31 coding tools incl. Claude Code, Codex, Cursor, Gemini & DeepSeek Harness, with native apps. Never reads... What the test found: All 3052 CLI tests pass with 0 failures; the tokentracker serve command starts and returns HTTP 200 on port 7680. 26 minutes. | ||||||
| 69 | mac-cleaner-cli | Runs | 100 / 100 | TypeScript | 1,991 | |
Free macOS CLI to clean disk space, caches, logs, Homebrew, Xcode. Open-source alternative to CleanMyMac What the test found: Installation and build succeed with pnpm + Node 22. All 344 tests pass across 40 files. CLI runs correctly: --version (1.3.5), --help prints full usage, scan finds cleanable files, categories lists all categories. 6 minutes. | ||||||
| 70 | amazon-q-developer-cli | Runs | 100 / 100 | Rust | 1,983 | |
โจ Agentic chat experience in your terminal. Build applications using natural language. What the test found: The Amazon Q CLI builds, passes its 328-unit test suite (0 failures, 13 pre-existing ignores), and the chat_cli binary prints its help text with all subcommands. 6 minutes. | ||||||
| 71 | carapace-bin | Runs | 100 / 100 | Go | 1,982 | |
A multi-shell completion binary. What the test found: carapace binary installed at /home/runner/go/bin/carapace, version 'carapace-bin develop', reports 12,587 completers, all 13 test cases pass, formatting/lint/staticcheck all clean, shell snippets and MCP server functional. 12 minutes. | ||||||
| 72 | resterm | Runs | 100 / 100 | Go | 1,981 | |
Terminal API client for HTTP, GraphQL and gRPC. Plain .http files you can diff and version, with workflows, mocks, profiling, tracing, OpenAPI import, SSH... What the test found: resterm builds from source with Go 1.26.8, all 64 test packages pass, the CLI runner executes real HTTP requests (verified 204 from httpbin.org), the mock server serves responses (verified 200 on /health), and the init command creates a new workspace. 21 minutes. | ||||||
| 73 | gita | Runs | 100 / 100 | Python | 1,947 | |
Manage many git repos with sanity ไปๅฎน็ฎก็ๅคไธชgitๅบ What the test found: gita v0.16.8.2 is installed in a Python 3.12 venv, all 72 tests pass, and the CLI correctly lists repos and displays multi-repo status. 5 minutes. | ||||||
| 74 | fastmod | Runs | 100 / 100 | Rust | 1,930 | |
A fast partial replacement for the codemod tool. Assists with large-scale codebase refactors via regex-based find and replace with human oversight and... What the test found: fastmod v0.4.5 builds, passes all 15 tests, and runs correctly as a find-and-replace tool from the command line and via 'cargo install'. 7 minutes. | ||||||
| 75 | tldx | Runs | 100 / 100 | Go | 1,929 | |
Bulk domain availability checking via RDAP, DNS, and WHOIS, with prefix/suffix permutations, regex patterns, MCP, and multiple output formats What the test found: tldx builds from source, its 294-test suite passes, and the binary answers version/help, performs real RDAP availability checks, and runs its MCP server over stdio without any third-party credentials. 32 minutes. | ||||||
| 76 | grepai | Runs | 100 / 100 | C | 1,901 | |
Semantic Search & Call Graphs for AI Agents (100% Local) What the test found: grepai binary builds and runs; all 16 test packages pass with zero failures. 5 minutes. | ||||||
| 77 | sharing | Runs | 100 / 100 | JavaScript | 1,839 | |
Sharing is a command-line tool to share directories and files from the CLI to iOS and Android devices without the need of an extra client app What the test found: npm install completed cleanly, all 32 tests pass, the server starts on a specified port and serves HTTP content with 200 OK. 1 minute. | ||||||
| 78 | env-cmd | Runs | 100 / 100 | TypeScript | 1,814 | |
Setting environment variables from a file What the test found: env-cmd installs, builds (dist/ pre-built), all 127 tests pass, CLI outputs --version (11.0.0), --help, and correctly loads and injects environment variables from env files into child processes. 7 minutes. | ||||||
| 79 | CoreCoder | Runs | 100 / 100 | Python | 1,796 | |
Minimal AI coding agent (~1,000 lines of Python) inspired by Claude Code. Works with any LLM. Think NanoGPT for coding agents. Formerly NanoCoder. What the test found: CoreCoder installed in a virtual environment; 185 tests all green; CLI --help and --demo mode both work; ruff reports one pre-existing EXE001 lint on examples/plan_hooks_demo.py. 1 minute. | ||||||
| 80 | cloudburn | Runs | 100 / 100 | TypeScript | 1,795 | |
Open-source policy engine that blocks bad AWS spending patterns before they ship and remediates what's already burning. What the test found: CloudBurn monorepo installs, builds, lints, typechecks, and passes all 1804 unit tests, 46 e2e tests, and 13 package-install tests across 5 packages; the CLI binary runs and correctly scans Terraform/CloudFormation fixtures. 5 minutes. | ||||||
| 81 | npq | Runs | 100 / 100 | JavaScript | 1,795 | |
safely install npm packages by auditing them pre-install stage What the test found: npq is installed with Node 24, all 777 tests pass, the CLI runs and audits packages via OSV/Snyk returning JSON output with real security findings. 2 minutes. | ||||||
| 82 | gpg-tui | Runs | 100 / 100 | Rust | 1,770 | |
Manage your GnuPG keys with ease! ๐ What the test found: gpg-tui 0.11.2 builds from source, answers --version (gpg-tui 0.11.2) and --help, and all 21 unit tests pass with a real GPG keyring. 35 minutes. | ||||||
| 83 | OpenOSINT | Runs | 100 / 100 | Python | 1,719 | |
AI-powered OSINT agent with interactive REPL, MCP server, and CLI. 20 tools. Works with Claude, GPT-4, or local models. For authorized security research only. What the test found: OpenOSINT 2.29.0 installs from source, CLI shows help with 20 tools, DNS tool resolves real domains, playbook generates real investigation reports, web server serves on port 9877 with HTTP 200 on / and /api/health, and 712 automated tests pass. 7 minutes. | ||||||
| 84 | zero | Runs | 100 / 100 | Go | 1,697 | |
The coding agent that answers to you, your model, your machine, your rules. What the test found: Zero v0.9.0 builds from source, passes all 87 test packages, passes go vet, format check, vulncheck, smoke test, and answers all CLI commands (help, models list, providers list, doctor) on linux/amd64. 9 minutes. | ||||||
| 85 | zero | Runs | 100 / 100 | Go | 1,697 | |
The coding agent that answers to you, your model, your machine, your rules. What the test found: Zero builds from source, all 87 test targets pass, release build produces version 0.9.0 and passes smoke test, and the CLI launches and responds to commands. 9 minutes. | ||||||
| 86 | sshs | Runs | 100 / 100 | Rust | 1,619 | |
Terminal user interface for SSH What the test found: sshs builds from source, all 34 unit tests pass, and the TUI binary launches on the virtual display printing its version and starting the curses interface. 4 minutes. | ||||||
| 87 | color | Runs | 100 / 100 | Go | 1,607 | |
๐จ Terminal color rendering library, support 8/16 colors, 256 colors, RGB color rendering output, support Print/Sprintf methods, compatible with Windows. GO... What the test found: All tests pass and the example demo runs, producing colored terminal output from the command-line color library. 11 minutes. | ||||||
| 88 | ttyper | Runs | 100 / 100 | Rust | 1,600 | |
Terminal-based typing test. What the test found: ttyper v1.6.0 builds from source, passes all 7 unit tests, installs via cargo install, and responds correctly to --version, --help, and --list-languages. The interactive TUI cannot be exercised in this headless environment, which is expected for a terminal-based app. 3 minutes. | ||||||
| 89 | ruby-progressbar | Runs | 100 / 100 | Ruby | 1,598 | |
Ruby/ProgressBar is a text progress bar library for Ruby. What the test found: ruby-progressbar 1.13.0 builds and passes all 244 RSpec tests on Ruby 3.4.8 (self-compiled from source) in a rootless Ubuntu 24.04 container. 12 minutes. | ||||||
| 90 | rang | Runs | 100 / 100 | C++ | 1,595 | |
A Minimal, Header only Modern c++ library for terminal goodies ๐โจ What the test found: The rang header-only library builds with CMake on Linux (GCC 13.3, C++11). All 6 doctest-based test cases pass with 21 assertions. The two standalone tests (colorTest, envTermMissing) also run and exit 0. cmake --install produces correct headers, cmake config, and pkg-config files in the install prefix. 3 minutes. | ||||||
| 91 | iocraft | Runs | 100 / 100 | Rust | 1,558 | |
A Rust crate for beautiful, artisanally crafted CLIs, TUIs, and text-based IO. What the test found: iocraft v0.9.1 and iocraft-macros v0.2.4 build and all 195 tests pass; example programs (hello_world, borders, use_output) execute correctly producing the expected terminal output and ANSI-rendered layouts. 27 minutes. | ||||||
| 92 | rumdl | Runs | 100 / 100 | Rust | 1,558 | |
Fast Markdown linter and formatter written in Rust What the test found: rumdl v0.2.77 (cargo build --release) and v0.2.78 (npm) both installed. All 13662 tests pass. The binary lints markdown files, detects violations like MD009 trailing spaces, and auto-fixes with --fix. 29 minutes. | ||||||
| 93 | lstr | Runs | 100 / 100 | Rust | 1,544 | |
A fast, minimalist directory tree viewer, written in Rust. What the test found: Rust toolchain installed via rustup, lstr v0.4.0 compiles from source, all 93 tests pass, and the binary runs correctly in classic mode (text, JSON, HTML) with all flags functional. 4 minutes. | ||||||
| 94 | jwt-cli | Runs | 100 / 100 | Rust | 1,516 | |
A super fast CLI tool to decode and encode JWTs built in Rust What the test found: jwt-cli builds from source, all 63 tests pass, the binary encodes and decodes JWTs with HS256 correctly. 4 minutes. | ||||||
| 95 | spotatui | Runs | 100 / 100 | Rust | 1,416 | |
A fast, standalone terminal music player in Rust: native Spotify streaming plus local, Subsonic, radio, Qobuz and Youtube sources What the test found: spotatui 0.41.0 builds, passes 1282 tests, and runs (--help, --version) in a container with no root access, after installing Rust via rustup and extracting libasound2-dev from its .deb. 18 minutes. | ||||||
| 96 | codesight | Runs | 100 / 100 | TypeScript | 1,414 | |
Universal AI context generator. Saves thousands of tokens per conversation in Claude Code, Cursor, Copilot, Codex, and more. What the test found: codesight v1.19.0 builds with tsc, all 149 tests pass, CLI scans and generates CODESIGHT.md and wiki output using real detectors (no mocks required). 4 minutes. | ||||||
| 97 | vibeyard | Runs | 100 / 100 | TypeScript | 1,387 | |
The IDE built for AI coding agents. What the test found: npm install and npm run build succeed; all 1957 tests pass (143 files); the app launches in Xvfb under electron . --no-sandbox, prints Provider detection+state init, and stays up until signal-terminated by timeout. 5 minutes. | ||||||
| 98 | ttyplot | Runs | 100 / 100 | C | 1,381 | |
a realtime plotting utility for terminal/console with data input from stdin What the test found: ttyplot 1.7.6 and stresstest binaries build from source, run without crash, and the ttyplot binary renders real-time text-mode plots from piped stdin data via ncurses. 17 minutes. | ||||||
| 99 | oh-my-agent | Runs | 100 / 100 | TypeScript | 1,337 | |
Mechanical verification for AI coding agents, skills pack or full harness (stop-hook gates, artifact checks, independent judges). What the test found: oh-my-agent v13.2.1 builds successfully, the oma CLI binary prints version 13.2.1, and all 4264 tests in 323 test files pass with no failures. 53 minutes. | ||||||
| 100 | shelve | Runs | 100 / 100 | TypeScript | 458 | |
Open-source secret & environment management. Secure, simple, collaborative. CLI & Github Sync What the test found: All 5 workspace packages build successfully, 158 tests pass (117 unit + 41 e2e), the Nuxt app serves HTTP 200 on port 3000, and the CLI binary responds to --help. 12 minutes. | ||||||
| 101 | plandex | Runs | 98.7 / 100 | Go | 15,702 | |
Open source AI coding agent. Designed for large projects and real world tasks. What the test found: Plandex server starts on port 8099 with PostgreSQL 16.4 and LiteLLM proxy on port 4000; server tests pass; CLI builds and authenticates against local server; full CRUD API for projects, plans, branches, settings is operational. 18 minutes. | ||||||
| 102 | gitui | Runs | 97.8 / 100 | Rust | 22,551 | |
Blazing ๐ฅ fast terminal-ui for git written in rust ๐ฆ What the test found: GitUI 0.28.1 builds from source, its full test suite passes (314 tests), and the binary launches as a terminal UI displaying real git repositories on the configured virtual display. 8 minutes. | ||||||
| 103 | cheat.sh | Runs | 97.3 / 100 | Python | 41,792 | |
the only cheat sheet you need What the test found: cheat.sh HTTP server is running on port 8002 with no-cache mode, serving cheat sheets from 6 fetched upstream repositories across all configured adapters (tldr, cheat, cheat.sheets, learnxiny, rosetta, late.nz). 16 minutes. | ||||||
| 104 | zoxide | Runs | 97.3 / 100 | Rust | 39,966 | |
A smarter cd command. Supports all major shells. What the test found: zoxide 0.10.0 builds from source, runs without errors, passes all 16 unit tests, and successfully adds and queries directories in its database. 3 minutes. | ||||||
| 105 | DeepSeek-Reasonix | Runs | 97.3 / 100 | Go | 35,748 | |
A reliable coding agent for complex software engineering tasks. What the test found: Reasonix builds as a single Go binary (bin/reasonix), answers --help with full usage including CLI/TUI, web, serve, ACP, and bot subcommands, passes all 147+ test packages with 0 failures, and produces valid JSON diagnostics via doctor --json. 19 minutes. | ||||||
| 106 | hyperfine | Runs | 97.3 / 100 | Rust | 28,966 | |
A command-line benchmarking tool What the test found: hyperfine 1.20.0 builds, all 102 tests pass, and the binary produces correct benchmark timing for a real command (sleep 0.1). 2 minutes. | ||||||
| 107 | ipatool | Runs | 97.3 / 100 | Go | 11,516 | |
Command-line tool that allows you to search for iOS, iPadOS, tvOS, visionOS, and macOS apps on the App Store, and download .ipa or macOS .pkg app packages. What the test found: ipatool binary builds successfully, all 15 unit test packages pass, and the search command reaches the live Apple App Store API and returns real results when provided with a keychain passphrase and a minimal mock account entry. 6 minutes. | ||||||
| 108 | posting | Runs | 96.7 / 100 | Python | 12,493 | |
The modern API client that lives in your terminal. What the test found: Posting 2.10.0 builds and installs with uv sync, its CLI responds to --help, the TUI launches on Xvfb with correct rendered layout, 126 unit tests pass (non-snapshot), and all 194 tests including snapshot comparisons pass when run serially with NO_COLOR unset. 31 minutes. | ||||||
| 109 | crit | Runs | 96.7 / 100 | Go | 1,186 | |
Review AI coding agents' plans, diffs and running apps in the browser. Local-first, works with any agent. What the test found: Built crit binary compiles, launches an HTTP server on 127.0.0.1, serves /api/health returning 200, auto-detects git changes in feature branches, serves embedded frontend, and the full test suite passes all unit and JS tests except one pre-existing intermittent flake in session lazy-threshold ordering. 42 minutes. | ||||||
| 110 | agent-deck | Runs | 96 / 100 | Go | 1,037 | |
Terminal session manager for AI coding agents. One TUI for Claude, Gemini, OpenCode, Codex, and more. What the test found: Agent-deck vdev builds and runs in the container. The binary responds to version, help, and ls commands. 34 of 34 directly testable internal packages pass. The WebSocket terminal bridge functions correctly. Only pre-existing environment-specific test failures: keepalive client-count assertions (container PTY behavior)... 62 minutes. | ||||||
| 111 | zeroshot | Runs | 95 / 100 | Rust | 1,931 | |
Runs coding agents as a graph: one agent implements, independent agents review, failures go to repair, and nothing ships until the checks pass. Works with... What the test found: Zeroshot installs via npm and builds from source; the CLI version, template list, UI server (200 on /ui/), and run validation all work, with 1607/1608 Rust tests passing in this container. 25 minutes. | ||||||
| 112 | terragrunt | Runs | 94.7 / 100 | Go | 9,872 | |
Terragrunt is a flexible orchestration tool that allows Infrastructure as Code written in OpenTofu/Terraform to scale. What the test found: Terragrunt v1.1.4 builds from source, the `terragrunt --version` and `--help` commands work, all 49 internal/ and pkg/ unit test packages pass, and catalog/scaffold integration tests pass. 13 minutes. | ||||||
| 113 | agmsg | Runs | 93.3 / 100 | Shell | 1,541 | |
Cross-vendor messaging for CLI AI coding agents, let Claude Code, Codex, Gemini & Copilot talk to each other in one team. Bash + SQLite, no daemon, no... What the test found: agmsg installed to ~/.agents/skills/agmsg/ (version f5a72de), sends/inboxes/history work correctly between 2 agents on the same team through SQLite store, and the test suite runs with all core functional tests passing. 20 minutes. | ||||||
| 114 | navi | Runs | 92.2 / 100 | Rust | 17,743 | |
An interactive cheatsheet tool for the command-line What the test found: navi builds from source, passes all 26 Rust unit tests, and its core cheatsheet query, info, widget generation, tldr integration, and cheatsh integration produce correct output when run with a controlling terminal. 26 minutes. | ||||||
| 115 | inshellisense | Runs | 90 / 100 | TypeScript | 10,725 | |
IDE style command line auto complete What the test found: inshellisense 0.0.4 on Node 18.19.1 builds, installs, runs doctor, and provides shell autocompletions via the complete command and passes all unit tests; e2e UI tests skipped due to native binary dependency requiring Node >=20. 5 minutes. | ||||||
| 116 | squad | Runs | 90 / 100 | TypeScript | 3,258 | |
Squad: AI agent teams for any project What the test found: Squad CLI v0.13.1-build.3 builds from source, runs `version` and `--help`, and passes 3454+ tests across 128 test files on Node 22.13.0. 27 minutes. | ||||||
| 117 | ripwire | Runs | 90 / 100 | C++ | 2,423 | |
The ripgrep of AI context: a zero-dependency C++23 CLI + MCP server for coding agents. Find what you want without reading the repo, then check you built what... What the test found: Build succeeds (plain dev build/ and Release build-install/), binary parses the repo source tree and emits deterministic minified XML with ranked symbols via Personalized PageRank, installed to /home/runner/.local/bin/ripwire with skills and hooks, --version/--help work, Python indexing confirms correct function... 21 minutes. | ||||||
| 118 | devspace | Runs | 87.3 / 100 | Go | 5,197 | |
DevSpace - The Fastest Developer Tool for Kubernetes โก Automate your deployment workflow with DevSpace and develop software directly inside Kubernetes. What the test found: DevSpace CLI builds from source and all unit tests pass; the binary prints help output correctly and is ready for use with a Kubernetes cluster. 16 minutes. | ||||||
| 119 | cc-safety-net | Runs | 86.7 / 100 | TypeScript | 1,582 | |
A pre-execution guard for AI coding agents. It blocks destructive Git and file system commands, plus common attempts to access sensitive files, before a tool... What the test found: bun install and bun run build succeed. bun run check passes lint, formatting, typecheck, knip, and duplication checks, with 3673 tests passing out of 3677 total across 213 files. Coverage is 98.91%. All E2E tests pass (packed-runtime, hermes-openclaw, and protection contracts). 15 minutes. | ||||||
| 120 | spotify-player | Runs | 86.2 / 100 | Rust | 7,267 | |
A Spotify player in the terminal with full feature parity What the test found: spotify_player 0.25.1 builds and passes all 45 tests, clippy (all features + no-features), and cargo fmt; the binary starts on Xvfd and produces the expected OAuth authentication URL on first launch. All requested features (streaming, rodio-backend, media-control, image, notify, fzf, daemon) compile and run without... 15 minutes. | ||||||
| 121 | claudexor | Runs | 86 / 100 | TypeScript | 496 | |
Multi-harness control plane for Claude Code, Codex, Cursor, and OpenCode: quota-aware rotation across multiple Claude/Codex subscriptions, shared thread... What the test found: pnpm install completed, pnpm build succeeded (31 packages in 18s), the CLI runs and reports version 3.9.8, the full test suite passes (350 test files, 4692 tests), canary stories pass (6 files, 50 tests), and typecheck passes (61 tasks). 20 minutes. | ||||||
| 122 | rtk | Runs | 80 / 100 | Rust | 82,684 | |
CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies What the test found: RTK (Rust Token Killer) v0.42.4 builds from source, passes all 2936 tests with 0 failures, and runs as a working CLI proxy producing token-compressed filtered output for ls, git, cargo, and other commands. 8 minutes. | ||||||
| 123 | hunk | Runs | 80 / 100 | TypeScript | 9,538 | |
Review-first terminal diff viewer for agentic coders What the test found: Hunk 0.23.0 builds from source with Bun 1.4.2, passes typecheck, runs its CLI (help, version, log, diff commands), and passes 2372/2383 default tests. 1 pre-existing daemon timing test fails, website tests require optional playwright dependencies, and the changeset versioning test requires Node 22+. 37 minutes. | ||||||
| 124 | codex-keysmith | Runs | 80 / 100 | Python | 4,703 | |
Versioned Codex instruction deployment with preview, ownership manifests, hook isolation, scenario evaluation, and recovery. What the test found: The codex-instruct.py CLI v0.6.0 installs, deploys overlay/unrestricted/contract prompts into Codex config directories with timestamped backups and manifest transactions, reports accurate multi-field status, and uninstalls with full rollback, verified by 1198 passing tests and a complete end-to-end... 11 minutes. | ||||||
| 125 | fd | Runs | 73.3 / 100 | Rust | 44,666 | |
A simple, fast and user-friendly alternative to 'find' What the test found: fd 10.5.0 builds from source, all 268 tests pass, and the binary correctly searches files by pattern, extension, and glob. 6 minutes. | ||||||
| 126 | textual | Runs | 73.3 / 100 | Python | 37,413 | |
The lean application framework for Python. Build sophisticated user interfaces with a simple Python API. Run your apps in the terminal and a web browser. What the test found: Textual 8.2.8 with syntax extras is installed; the full test suite (3459 tests) passes; the demo app and a minimal app both run and render correctly. 14 minutes. | ||||||
| 127 | typer | Runs | 73.3 / 100 | Python | 20,056 | |
Typer, build great CLIs. Easy to code. Based on Python type hints. What the test found: Typer v0.27.2 is installed, all 1372 tests pass, and the CLI tool runs Python scripts as CLI apps. 20 minutes. | ||||||
| 128 | trippy | Runs | 70.7 / 100 | Rust | 7,996 | |
A network diagnostic tool What the test found: The Trippy Rust workspace (trippy/privacy/traceroute/ping tool) compiles, passes all 807 unit tests, and the CLI binary reports its version. The project is ready for development. 33 minutes. | ||||||
| 129 | yazi | Runs | 66.7 / 100 | Rust | 42,689 | |
๐ฅ Blazing fast terminal file manager written in Rust, based on async I/O. What the test found: Yazi 26.9.1 builds from source with Rust 1.98.1, both the TUI file manager (yazi) and CLI (ya) binaries produce correct version output, the TUI launches and renders the file listing interactively, and the CLI executes subcommands without error. 22 minutes. | ||||||
| 130 | goaccess | Runs | 66.7 / 100 | C | 21,009 | |
GoAccess is a real-time web log analyzer and interactive viewer that runs in a terminal in *nix systems or through your browser. What the test found: GoAccess v1.11 builds from source and runs. It parses Apache/Nginx Combined log format files into valid JSON output with correct request counts (4 requests, 3 unique visitors, 2xx status codes, browser/user-agent breakdown) and generates a complete self-contained HTML report. The binary at /work/repo/goaccess executes... 14 minutes. | ||||||
| 131 | codeburn | Runs | 66.7 / 100 | TypeScript | 11,355 | |
Free, local tool to track AI coding token usage and cost across 37 tools and agents (Claude Code, Cursor, Codex, Gemini and more), by model, project, and task... What the test found: CodeBurn v0.9.23 builds, passes 3546 of 3567 tests, TypeScript compiles clean, and the CLI prints help and version on demand. 8 minutes. | ||||||
| 132 | pueue | Runs | 66.7 / 100 | Rust | 6,366 | |
:stars: Manage your shell commands. What the test found: Pueue v4.0.4 builds from source, all 186 tests pass, and the daemon-client pair runs correctly: adding tasks, querying status (plain & JSON), viewing logs, killing, cleaning, and waiting all work over the Unix socket. 10 minutes. | ||||||
| 133 | jscpd | Runs | 66.7 / 100 | Rust | 6,357 | |
Copy/paste detector for source code. 220+ languages, Rust engine, SARIF/HTML/badge reporters, GitHub Action, MCP server for AI agents. What the test found: jscpd 5.1.1 builds from source via cargo, installs via npm as a prebuilt binary, and runs duplication detection on 224 language formats with 15 reporters; the full Rust test suite passes. 9 minutes. | ||||||
| 134 | television | Runs | 66.7 / 100 | Rust | 6,330 | |
A very fast, portable and hackable fuzzy finder. What the test found: The television fuzzy finder builds from source with cargo and all 386 tests pass including 125 CLI integration tests exercising the TUI via phantom-test on Xvfb. 20 minutes. | ||||||
| 135 | xplr | Runs | 66.7 / 100 | Rust | 4,837 | |
A hackable, minimal, fast TUI file explorer What the test found: xplr 1.1.1 TUI file explorer built from source via cargo build --locked --release; binary responds to --version and --help; all 22 tests pass. 4 minutes. | ||||||
| 136 | Archon | Runs | 64 / 100 | TypeScript | 23,644 | |
The first open-source harness builder for AI coding. Make AI coding deterministic and repeatable. What the test found: Bun 1.4.2 installed, 1282 dependencies installed; all 9 workspace packages' tests pass with 0 failures; server starts on port 3090 and serves Archon web UI (HTTP 200); CLI prints full help output; one import-boundary test expectation updated in packages/cli/src/cli.test.ts for bun 1.4.2 compatibility. 36 minutes. | ||||||
| 137 | f3d | Runs | 60 / 100 | C++ | 4,743 | |
Fast and minimalist 3D viewer. What the test found: F3D Python bindings (pip) create a rendering engine and produce valid 300x300 PNG images; the F3D CLI binary reads OBJ files and renders them to valid PNG output on the virtual display. 36 minutes. | ||||||
| 138 | inquire | Runs | 60 / 100 | Rust | 2,635 | |
A Rust library for building interactive prompts What the test found: inquire 0.9.4 Rust workspace compiles with rustc 1.82.0, all 250 tests pass when NO_COLOR is unset, all 3 workspace crates (inquire, inquire-derive, examples) build successfully. 33 minutes. | ||||||
| 139 | spinner | Runs | 60 / 100 | Go | 2,530 | |
Go (golang) package with 90 configurable terminal spinner/progress indicators. What the test found: The Go+ compiler v1.24.13 runs on this system, the spinner package builds successfully, and all 17 tests pass including spinner start/stop/reverse/color/update/backspace/line-computation tests. 8 minutes. | ||||||
| 140 | jcode | Runs | 53.3 / 100 | Rust | 20,351 | |
High performance coding agent harness written in rust What the test found: jcode v0.81.7-dev is built from source, installed via symlinks in ~/.jcode/builds/current and ~/.local/bin, and works correctly with OpenRouter provider. jcode run and jcode repl both complete successfully with real AI responses. 20 minutes. | ||||||
| 141 | code2prompt | Runs | 33.3 / 100 | Rust | 7,721 | |
A CLI tool to convert your codebase into a single LLM prompt with source tree, prompt templating, and token counting. What the test found: code2prompt v4.3.0 builds, passes all 113 tests, and the CLI runs correctly, generating formatted prompts from any codebase path. 10 minutes. | ||||||
| 142 | mirrord | Runs | 33.3 / 100 | Rust | 5,358 | |
Run any process, on your machine or in an AI agent's environment, as if it were a pod in your Kubernetes cluster: real env vars, DNS, network, traffic. What the test found: mirrord CLI binary and layer build, run, and pass all unit tests and clippy checks on x86_64 Linux with nightly-2026-08-13 Rust toolchain. 30 minutes. | ||||||
| 143 | lnav | Runs | 25 / 100 | C++ | 10,726 | |
Log file navigator What the test found: Lnav v0.14.0 is running correctly with full functionality; all tested log formats, SQL queries, filters, and headless operations work; source from the repository can be configured with zig cc as the C++ compiler after provisioning development headers manually. 39 minutes. | ||||||
| 144 | gh-dash | Runs with mocks | 92 / 100 | Go | 12,605 | |
A rich terminal UI for GitHub that doesn't break your flow. What the test found: The gh-dash binary builds from source, runs on Linux amd64, prints help and version output, and passes all 243 unit tests across 19 Go packages. 16 minutes. | ||||||
| 145 | shell_gpt | Runs with mocks | 92 / 100 | Python | 12,292 | |
A command-line productivity tool powered by AI large language models like GPT-5, will help you accomplish your tasks faster and more efficiently. What the test found: ShellGPT 1.5.1 installs into a Python venv, its `sgpt --version` reports the version, and all 29 unit tests pass with mocked OpenAI completions. 8 minutes. | ||||||
| 146 | App-Store-Connect-CLI | Runs with mocks | 92 / 100 | Go | 7,693 | |
Fast, scriptable CLI for the App Store Connect API. Automate TestFlight, builds, submissions, signing, analytics, screenshots, subscriptions, and more What the test found: The asc binary builds from source, prints version and help, and all 110+ Go test packages pass with zero failures under ASC_BYPASS_KEYCHAIN=1 (mock/internal HTTP, no real API credentials). 11 minutes. | ||||||
| 147 | coreutils | Runs with mocks | 92 / 100 | Rust | 5,235 | |
Coreutils for Windows: Installer & Packaging What the test found: The coreutils multi-call binary builds and runs 79 UNIX utilities on Linux after adapting the Windows-native project to compile without Windows SDK. 27 minutes. | ||||||
| 148 | petdex | Runs with mocks | 92 / 100 | TypeScript | 4,210 | |
A public gallery of animated pets for Codex, Claude Code, DeepSeek Harness, Hermes, OpenCode, Gemini CLI, and more. What the test found: All dependencies installed, Next.js app builds and serves pages (200 on /, /en/, /robots.txt, /sitemap.xml), CLI builds/typechecks/installs pets, 537 of 539 tests pass. 24 minutes. | ||||||
| 149 | dockly | Runs with mocks | 92 / 100 | JavaScript | 4,033 | |
Immersive terminal interface for managing docker containers and services What the test found: Dockly installs, lints pass, --version and --help work, and the full TUI initializes against a mock Docker daemon with all dockerode API calls responding correctly. 7 minutes. | ||||||
| 150 | jevgrep | Runs with mocks | 92 / 100 | TypeScript | 2,436 | |
Find code by asking what it does. A CLI for coding agents that uses Jev to discover relevant files and source context. What the test found: Repository builds, types check, lints clean. All 4 test suites pass (70 pass, 0 fail). End-to-end CLI search works against a mock provider: auth saves credentials, doctor verifies access, search returns file locations and source excerpts. 6 minutes. | ||||||
| 151 | snapai | Runs with mocks | 92 / 100 | TypeScript | 1,941 | |
AI-powered icon generation CLI for React Native & Expo developers. Generate stunning app icons in seconds using OpenAI's latest models. What the test found: SnapAI CLI installs, builds (tsc + webpack), and runs; the icon and feature-graphic commands successfully generate valid PNG images when pointed at a mock OpenAI-compatible API endpoint. Gemini model requires a real Google API key. 5 minutes. | ||||||
| 152 | cli | Runs with mocks | 92 / 100 | Go | 1,833 | |
A command-line interface for Hetzner Cloud What the test found: hcloud CLI binary builds and runs, printing version 1.69.0-dev and full help output. The complete internal test suite (34 packages) passes with zero failures. 5 minutes. | ||||||
| 153 | jiratui | Runs with mocks | 92 / 100 | Python | 1,712 | |
A Textual User Interface for interacting with Atlassian Jira from your shell What the test found: The JiraTUI v1.16.0 project installs cleanly with uv sync, the CLI tool (jiratui) responds correctly to version, help, config, and themes commands, and the full test suite passes over 300 tests covering actions (standard and legacy key bindings), models, search, attachments, git, JQL editor, confirmation screens... 34 minutes. | ||||||
| 154 | yt-x | Runs with mocks | 92 / 100 | Shell | 1,668 | |
Posix script to browse youtube plus other yt-dlp supported sites from your terminal (fzf) or app launcher (rofi) with optional previews. (supports bash, zsh... What the test found: yt-x v0.8.7 installed with yt-dlp, fzf, jq, and mock-mpv; performs YouTube searches via yt-dlp, displays search results, auto-selects first result, and correctly invokes the configured media player with the video URL. 7 minutes. | ||||||
| 155 | ttl | Runs with mocks | 92 / 100 | Rust | 1,472 | |
Fast, modern traceroute with real-time TUI, per-hop stats, ASN/geo lookup, ECMP detection, and MPLS label parsing. A better mtr. What the test found: The ttl v0.23.0 Rust project builds from source and passes all 345 automated tests (unit, integration, doc-tests), the binary prints version 0.23.0 and help text, but raw network probing cannot run due to container-level CAP_NET_RAW restriction. 12 minutes. | ||||||
| 156 | jira-cli | Runs with mocks | 88 / 100 | Go | 6,020 | |
๐ฅ Feature-rich interactive Jira command line. What the test found: The jira-cli binary is built, runs help/version commands without error, and all unit tests pass. Real Jira API credentials are required for operational commands. 5 minutes. | ||||||
| 157 | dblab | Runs with mocks | 86 / 100 | Go | 3,242 | |
The database client every command line junkie deserves. What the test found: dblab builds from source and all unit tests pass; the interactive TUI app cannot launch because /dev/tty is unavailable in the headless container. 9 minutes. | ||||||
| 158 | ospec | Runs with mocks | 86 / 100 | JavaScript | 452 | |
Spec-driven, agentic workflow framework for AI coding agents. Turn a request into a verifiable goal loop, plan, act, verify, with durable specs and evidence in... What the test found: npm install completed without errors (Node 18 engine warnings only); ospec CLI v2.1.0 responds to --version, --help, init, status, change, session commands; release smoke test and content scan both pass; index rebuild tool works; global install via --prefix succeeds. 5 minutes. | ||||||
| 159 | lazycodex | Runs with mocks | 85.3 / 100 | TypeScript | 3,747 | |
The one and only agent harness for complex codebases. Project memory, planning, execution, and verified completion inside Codex. What the test found: LazyCodex plugin (omo 5.0.0-beta.90) is installed, enabled, and verified via the Codex marketplace path with 9/9 repo tests passing, 288/319 plugin tests passing, 27 skills, 21 hooks, and 4 components (rules, comment-checker, lsp, telemetry) responding to CLI commands. 37 minutes. | ||||||
| 160 | termscp | Runs with mocks | 82 / 100 | Rust | 3,117 | |
๐ฅ A feature rich terminal UI file transfer and explorer with support for SCP/SFTP/FTP/S3/SMB/WebDAV What the test found: termscp builds from source (without SMB), its binary responds to --version and --help, and all 335 unit tests pass. 9 minutes. | ||||||
| 161 | terminal-browser | Runs with mocks | 56 / 100 | TypeScript | 3,706 | |
A browser inside your terminal What the test found: pnpm install and pnpm -r build succeed. Node 22 with --experimental-sqlite passes all browser and shared tests (42/43 pixel tests). CLI binary runs and responds to help/version/ls commands. 9 minutes. | ||||||
| 162 | smenu | Runs with mocks | 56 / 100 | C | 2,494 | |
smenu started as a lightweight and flexible terminal menu generator, but quickly evolved into a powerful and versatile CLI selection tool for interactive or... What the test found: smenu 1.5.0 binary builds and runs, reads words from stdin/file, displays in scrollable terminal window with ncurses-based cursor/selection, supports keyboard navigation, and outputs selected words on stdout with exit 0. The automated test harness (tests/test.sh) cannot execute because ptylie requires root for... 20 minutes. | ||||||
| 163 | radian | Runs with mocks | 46 / 100 | Python | 2,291 | |
A 21 century R console What the test found: radian 0.7.1 installed via pip, uses R 4.3.3 extracted from Ubuntu debs. 15 of 22 automated tests pass. R console launches interactively (test_aaa confirms prompt). CRAN-based features (reticulate Python mode, askpass, renv) untested due to no CRAN mirror access. 30 minutes. | ||||||
| 164 | ouroboros | Runs with mocks | 44.3 / 100 | Python | 6,194 | |
Agent OS: the agent gets smarter on its own. We just hold the line: Interview-gated, staged evaluation, budgeted evolution loop. MCP server, 14 runtimes... What the test found: All unit tests pass (0 failures in ~21000 tests across 34 unit directories), ruff lint and format clean, CLI boots with correct version, MCP server help displays correctly, 2 Windows-specific bugs were caught and fixed. 73 minutes. | ||||||
| 165 | go-recipes | Could not verify | 80 / 100 | Go | 4,782 | |
๐ฆฉ Tools for Go projects What the test found: Go 1.23.4 installed; project builds without error; README generated from page.yaml via go generate; binary runs and exits successfully. 3 minutes. | ||||||
| 166 | httm | Could not verify | 80 / 100 | Rust | 1,668 | |
Interactive, file-level Time Machine-like tool for ZFS/btrfs/nilfs2 (and even Time Machine and Restic backups!) What the test found: httm 0.51.0 builds and runs on Rust 1.99.0; the binary prints its version, help text, and debug config correctly, but requires ZFS/BTRFS/NILFS2 snapshots for data output. 4 minutes. | ||||||
| 167 | TUI-ConsoleLauncher | Could not verify | 80 / 100 | Java | 1,427 | |
Linux CLI Launcher for Android What the test found: Both fdroid-debug and playstore-debug APKs build successfully from source with Gradle 8.2, Android SDK 34, and a self-signed dev keystore. 9 minutes. | ||||||
| 168 | CodeMachine-CLI | Could not verify | 75 / 100 | TypeScript | 2,511 | |
CodeMachine is an open-source tool that orchestrates AI coding agents into repeatable, long-running workflows. โก๏ธ What the test found: CodeMachine CLI v0.8.0 installs via 'bun install' (591 packages), responds to all CLI subcommands (help, version, auth, run, step, templates, agents, mcp, import, export), and launches its TUI interface. No test suite exists in the repository. 5 minutes. | ||||||
| 169 | intelligent-terminal | Could not verify | 25 / 100 | C++ | 2,043 | |
A fork of Windows Terminal with native agent integration, right in your command line. What the test found: Rust toolchain installed and working. Installer bootstrap crate builds and its 2 unit tests pass. WTA crate type-checks for Windows MSVC target. C++ component (Visual Studio/MSBuild) cannot build on Linux. Full test suite (1997 WTA tests) requires Windows to execute. 32 minutes. | ||||||
| 170 | Codewhale | Could not verify | 20 / 100 | Rust | 41,083 | |
Open-source Rust agent engine and terminal client for Codewhale, with provider choice, tools, approvals and receipts. What the test found: could not be verified; the log shows where it stopped. 19 minutes. | ||||||
| 171 | kyanos | Could not verify | 20 / 100 | C | 5,071 | |
Kyanos is a networking analysis tool using eBPF. It can visualize the time packets spend in the kernel, capture requests/responses, makes troubleshooting more... What the test found: could not be verified; the log shows where it stopped. 38 minutes. | ||||||
| 172 | agent-of-empires | Could not verify | 20 / 100 | Rust | 3,328 | |
Manage multiple Claude Code, OpenCode agents from either TUI or Web for easy access on mobile. Also supports Mistral Vibe, Codex CLI, Gemini CLI, Pi.dev... What the test found: could not be verified; the log shows where it stopped. 2 minutes. | ||||||
| 173 | zsh-z | Could not verify | 20 / 100 | Shell | 2,459 | |
Jump quickly to directories that you have visited "frecently." A native Zsh port of z.sh with added features. What the test found: could not be verified; the log shows where it stopped. 15 minutes. | ||||||
| - | gh-skyline | Not yet tested | - | Go | 1,356 | - |
| - | sttr | Not yet tested | - | Go | 1,354 | - |
| - | install-nothing | Not yet tested | - | Rust | 1,347 | - |
| - | tcping | Not yet tested | - | Go | 1,343 | - |
| - | autoprompt-skill | Not yet tested | - | JavaScript | 1,298 | - |
| - | intelli-shell | Not yet tested | - | Rust | 1,297 | - |
| - | OpenContext | Not yet tested | - | JavaScript | 1,261 | - |
| - | PyPDFForm | Not yet tested | - | Python | 1,251 | - |
| - | multi-gitter | Not yet tested | - | Go | 1,239 | - |
| - | thClaws | Not yet tested | - | Rust | 1,238 | - |
| - | k2tf | Not yet tested | - | Go | 1,230 | - |
| - | lla | Not yet tested | - | Rust | 1,230 | - |
| - | libtmux | Not yet tested | - | Python | 1,214 | - |
| - | dstask | Not yet tested | - | Go | 1,210 | - |
| - | SearchCLI | Not yet tested | - | TypeScript | 1,194 | - |
| - | claude-powerline | Not yet tested | - | TypeScript | 1,174 | - |
| - | progressbar | Not yet tested | - | Java | 1,173 | - |
| - | argc | Not yet tested | - | Rust | 1,169 | - |
| - | lurk | Not yet tested | - | Rust | 1,158 | - |
| - | xq | Not yet tested | - | Go | 1,158 | - |
| - | ai-agent-skills | Not yet tested | - | JavaScript | 1,149 | - |
| - | MiniCode | Not yet tested | - | TypeScript | 1,137 | - |
| - | best-claude-hud | Not yet tested | - | Rust | 1,111 | - |
| - | sonar | Not yet tested | - | Go | 1,093 | - |
| - | dao-code | Not yet tested | - | TypeScript | 1,082 | - |
| - | cs | Not yet tested | - | Go | 1,079 | - |
| - | Ix | Not yet tested | - | TypeScript | 1,031 | - |
| - | themes | Not yet tested | - | Python | 1,029 | - |
| - | dscode | Not yet tested | - | JavaScript | 1,021 | - |
| - | easy-agent | Not yet tested | - | TypeScript | 1,008 | - |
| - | codealmanac | Not yet tested | - | TypeScript | 997 | - |
| - | git-graph | Not yet tested | - | Rust | 989 | - |
| - | invoice_printer | Not yet tested | - | Ruby | 978 | - |
| - | amber | Not yet tested | - | Rust | 952 | - |
| - | axe | Not yet tested | - | Go | 895 | - |
| - | kepubify | Not yet tested | - | Go | 893 | - |
| - | deja | Not yet tested | - | Go | 888 | - |
| - | webkubectl | Not yet tested | - | Go | 882 | - |
| - | fcp | Not yet tested | - | Rust | 858 | - |
| - | tcping | Not yet tested | - | Go | 857 | - |
| - | herdr-reviewr | Not yet tested | - | Rust | 851 | - |
| - | mcp-client-for-ollama | Not yet tested | - | Python | 827 | - |
| - | tokentap | Not yet tested | - | Python | 814 | - |
| - | keifu | Not yet tested | - | Rust | 810 | - |
| - | tui-journal | Not yet tested | - | Rust | 787 | - |
| - | claude-code-zh-cn | Not yet tested | - | JavaScript | 782 | - |
| - | halp | Not yet tested | - | Rust | 768 | - |
| - | agentacct | Not yet tested | - | Python | 766 | - |
| - | music-dl | Not yet tested | - | PHP | 763 | - |
| - | summon | Not yet tested | - | Go | 762 | - |
| - | codemap | Not yet tested | - | Go | 704 | - |
| - | fontbakery | Not yet tested | - | Python | 704 | - |
| - | shotgun | Not yet tested | - | Python | 689 | - |
| - | azure-devops-cli-extension | Not yet tested | - | Python | 688 | - |
| - | cmd2 | Not yet tested | - | Python | 688 | - |
| - | klog | Not yet tested | - | Go | 686 | - |
| - | comics-downloader | Not yet tested | - | Go | 680 | - |
| - | jsongrep | Not yet tested | - | Rust | 677 | - |
| - | aislop | Not yet tested | - | TypeScript | 675 | - |
| - | sigmap | Not yet tested | - | JavaScript | 645 | - |
| - | wtp | Not yet tested | - | Go | 636 | - |
| - | python-launcher | Not yet tested | - | Rust | 631 | - |
| - | purrcrypt | Not yet tested | - | Rust | 627 | - |
| - | agentty | Not yet tested | - | C++ | 615 | - |
| - | paper-age | Not yet tested | - | Rust | 615 | - |
| - | ClaudeCodeStatusLine | Not yet tested | - | Shell | 614 | - |
| - | gup | Not yet tested | - | Go | 611 | - |
| - | session-kit | Not yet tested | - | Python | 602 | - |
| - | cc-lens | Not yet tested | - | TypeScript | 598 | - |
| - | go-cover-treemap | Not yet tested | - | Go | 596 | - |
| - | ubi | Not yet tested | - | Rust | 595 | - |
| - | hostc | Not yet tested | - | TypeScript | 592 | - |
| - | laravel-top | Not yet tested | - | PHP | 584 | - |
| - | onchain-dex-cli | Not yet tested | - | TypeScript | 583 | - |
| - | outsmart-cli | Not yet tested | - | TypeScript | 583 | - |
| - | agent-manager | Not yet tested | - | Go | 577 | - |
| - | lowfat | Not yet tested | - | Rust | 577 | - |
| - | xcov | Not yet tested | - | Ruby | 573 | - |
| - | brain.md | Not yet tested | - | JavaScript | 564 | - |
| - | geek-life | Not yet tested | - | Go | 563 | - |
| - | jmxterm | Not yet tested | - | Java | 563 | - |
| - | hcom | Not yet tested | - | Rust | 561 | - |
| - | renamer | Not yet tested | - | JavaScript | 560 | - |
| - | castor | Not yet tested | - | PHP | 556 | - |
| - | remarshal | Not yet tested | - | Python | 553 | - |
| - | harness-score | Not yet tested | - | TypeScript | 548 | - |
| - | cli | Not yet tested | - | TypeScript | 542 | - |
| - | foremerge | Not yet tested | - | Rust | 538 | - |
| - | patent | Not yet tested | - | Rust | 534 | - |
| - | claude-code-hooks | Not yet tested | - | JavaScript | 533 | - |
| - | mco | Not yet tested | - | Python | 529 | - |
| - | token-tracker | Not yet tested | - | Python | 528 | - |
| - | clsh | Not yet tested | - | TypeScript | 526 | - |
| - | Extract | Not yet tested | - | Shell | 525 | - |
| - | agentbox | Not yet tested | - | TypeScript | 523 | - |
| - | performance | Not yet tested | - | PHP | 520 | - |
| - | roam-code | Not yet tested | - | Python | 517 | - |
| - | heh | Not yet tested | - | Rust | 504 | - |
| - | innodb-java-reader | Not yet tested | - | Java | 495 | - |
| - | Kiri | Not yet tested | - | Rust | 490 | - |
| - | Flowtrace | Not yet tested | - | TypeScript | 485 | - |
| - | grok-keysmith | Not yet tested | - | Python | 484 | - |
| - | octocode | Not yet tested | - | Rust | 482 | - |
| - | chief | Not yet tested | - | Go | 473 | - |
| - | rapidgzip | Not yet tested | - | Python | 469 | - |
| - | koji | Not yet tested | - | Rust | 465 | - |
| - | snip | Not yet tested | - | Go | 464 | - |
| - | circleci-cli | Not yet tested | - | Go | 461 | - |
| - | dsh-hub-cli | Not yet tested | - | TypeScript | 458 | - |
| - | linode-cli | Not yet tested | - | Python | 442 | - |
| - | pipe-rename | Not yet tested | - | Rust | 437 | - |
runs installed and started with its real dependencies. runs with mocks started after stand-ins replaced external services such as a database or a third-party API. could not verify neither the standard agent nor the stronger one got it running within the time limit; the log shows where it stopped.
Before you choose
Most of this list builds with one command. The Rust and Go tools dominate the "runs" group: ripgrep built and passed all 332 of its tests, broot 132 of 132, gitlogue 173, and the binaries answered version, help, and a real command inside the container. Where the source build needed a compiler the machine lacked, the agent used the project's pre-built binary instead: websocat (its log names the missing gcc and binutils), television, and jscpd.
Decide whether the tool needs a service or an account. plandex launched with a real PostgreSQL connected and all migrations applied. Tools that front a remote API could only be verified against stand-ins or without credentials, and are "runs with mocks": jira-cli passed its 15 test packages against mocked HTTP servers, App-Store-Connect-CLI answered every subcommand without Apple credentials. ipatool is "runs" because it reports the missing App Store credentials correctly and everything else passes.
Look at the minutes. The npm and pip installs were up in 3 to 9 minutes (npkill, jc, ccstatusline, ipatool, codeburn); the larger Rust builds took 24 to 41 (broot, websocat, grex, coreutils, xh, gitlogue, ast-grep, gitui, trippy, miniserve, xplr, television).
Two could not be verified within the time limit, even by the stronger agent: code2prompt and pueue; their logs show where each attempt stopped. rtk is waiting for a finished attempt, and lnav and mirrord have not been tested yet; all three sit unranked.
How we tested
A command-line tool counts as working when it builds from source or installs by the project's own instructions, its test suite passes, and the binary answers: version, help, and at least one real command inside the container, such as ripgrep searching files, jc converting command output to JSON, or miniserve serving a directory over HTTP.
Most of this list is Rust and Go, which build with one command and need no services, so "runs" dominates. Tools that front a remote service got "runs with mocks": jira-cli passed its tests against mocked HTTP servers, App-Store-Connect-CLI built and answered every subcommand without Apple credentials. Interactive terminal applications were verified through headless tests where the project provides them; lazygit is the example.


On this list as of the latest test: 143 projects ran as-is, 21 with mocks, 9 could not be verified, 111 still waiting. Languages tested: C, C++, Go, Java, JavaScript, PHP, Python, Ruby, Rust, Shell, TypeScript. Every attempt used a clean single-use machine, the subject at a pinned version, and a 45-minute limit; the complete procedure is on the methodology page.
Frequently asked questions (FAQs)
What does "runs" mean for a CLI tool?
It built from source or installed by the project's own instructions, its test suite passed, and the binary answered version, help, and at least one real command inside the container. ripgrep, jc, and miniserve are examples in the logs.
Why is a tool "runs with mocks"?
It fronts a remote service and was verified against a stand-in. jira-cli passed its tests against mocked HTTP servers; App-Store-Connect-CLI answered every subcommand without Apple credentials.
Are interactive terminal apps tested?
Yes, through the headless tests the project provides. lazygit passed its headless integration tests; where a project has no such tests, the verification is the binary starting and answering commands.
Does a fast build help the rank?
No. Build time is in the log but not in the score. Rank follows the verdict, then the Argusic Score, then GitHub stars.
How current is this list?
Each row carries its own test date and the page states the window. A project can change after its test; a new attempt replaces the verdict only when it finishes with complete evidence.
How is this list ranked?
By measurement, not opinion: projects Argusic installed and launched on a fresh machine come first, then those that ran with mocks in place of external services, then those it could not verify. Ties go to the Argusic Score, then how popular it is on its own source.
Why are some projects unranked?
111 projects are still waiting for a test or for a finished attempt. They are listed without a rank until Argusic has measured them.
More lists in this category
- open source coding agents (shares ANUS, Codewhale, CoreCoder, DeepSeek-Reasonix, MiniCode, agent-deck, agentacct, agentbox, agentty, autoprompt-skill, brain.md, cc-safety-net, claudexor, codealmanac, crit, dao-code, dscode, easy-agent, foremerge, herdr, herdr-reviewr, jcode, jevgrep, oh-my-agent, ospec, ouroboros, ripwire, rtk, tuios, zero, zero with this list)
- MCP servers (shares Graft, OpenOSINT, jscpd, mcp-client-for-ollama, ouroboros, ripwire, thClaws with this list)
- self-hosted AI apps (shares TokenTracker, cc-lens, ccstatusline, crit, plandex, tokentap with this list)
- AI agent frameworks (shares DeepSeek-Reasonix, Graft, herdr, plandex with this list)
- observability tools (shares TokenTracker, agentacct, codeburn, witr with this list)
- API clients (shares curlie, resterm, xh with this list)
- CI/CD tools (shares circleci-cli, xcov with this list)
- Open source music servers, installed and played (shares spotatui, spotify-player with this list)
- project management tools (shares codemap, geek-life with this list)
- wikis (shares brain.md, codealmanac with this list)
- workflow automation tools (shares CodeMachine-CLI, autoprompt-skill with this list)
- browser automation tools (shares webcmd with this list)
- self-hosted password managers (shares shelve with this list)
- search engines (shares SearchCLI with this list)
- self-hosted analytics (shares goaccess with this list)
- API gateways
- CMS platforms
- LLM gateways
- VPN tools
- backup tools
- code editors
- Open source databases, installed and queried
- e-commerce platforms
- ebook readers
- game engines
- open source games
- home automation tools
- low-code platforms
- map tools
- message queues
- Open source office suites, installed and launched
- screen recorders
- speech tools
- static site generators
- uptime monitors
- vector databases
- open source video editors
- video players
- web scraping tools
- whiteboard tools
- self-hosted Notion alternatives
- self-hosted dashboards
- self-hosted git servers