Open source MCP servers that come up on first run
Servers for the Model Context Protocol, ordered by what happened when Argusic cloned each repository and tried to start it. A passing verdict means the server started; the recording shows exactly how.
Tested between and . Each row shows its own test date; a project can change after that day.
162 of 172 tested projects run. 48 more waiting for a test.
In short: 104 of the 172 tested projects started as-is on a fresh machine: playwright-mcp, mcp-for-blender, Skill_Seekers, SeleniumBase, QuantDinger, fastapi_mcp, Graft, and git-mcp, and 96 more. 58 more started once a stand-in replaced a service they expect, such as a database: rea, xiaozhi-esp32-server, lamda, waha, magic-mcp, mcp-atlassian, klavis, and zotero-mcp, and 50 more. 10 could not be verified: deep-research, azure-skills, FinanceToolkit, open-seo-mcp-skills, httprunner, dbhub, brave-search-mcp-server, and drawio-mcp-server, and 2 more; the log shows where each one stopped.
Measured by Argusic on a fresh machine every time. Every number links to its evidence.
| # | project | verdict | Argusic Score | language | stars | tested on |
|---|---|---|---|---|---|---|
| 1 | playwright-mcp | Runs | 100 / 100 | TypeScript | 37,924 | |
Playwright MCP server What the test found: The Playwright MCP server v0.0.80 installs, builds, passes all 8 tests (browser navigation, click, capabilities, CLI, and CommonJS import), and responds to --version and --help correctly using Chromium in headless mode. 6 minutes. | ||||||
| 2 | mcp-for-blender | Runs | 100 / 100 | Python | 30,231 | |
Community plugin to control Blender 3D with any LLM of your choice. Not affiliated with the official Blender Foundation. What the test found: mcp-for-blender v2.0.0 builds with uv, all 152 tests pass, and the MCP server binary starts and correctly responds to JSON-RPC initialize requests; the install-addon CLI reports Blender not found (expected, Blender is not installed in this container), and no third-party API keys were needed for any test to pass. 2 minutes. | ||||||
| 3 | Skill_Seekers | Runs | 100 / 100 | Python | 15,115 | |
Convert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection What the test found: skill-seekers 3.10.0.dev0 installed and all 3857 unit/adaptor/scraper tests pass without external API keys or network dependencies. 6 minutes. | ||||||
| 4 | SeleniumBase | Runs | 100 / 100 | Python | 13,053 | |
📊 Browser automation framework for scraping, testing, and completing tasks with Python. Supports pytest. Stealth options. Over 100 examples. What the test found: SeleniumBase 4.54.7 is installed in a Python 3.12 venv; 8 framework unit tests, 7 offline HTML tests, and 2 real web tests passed using headless Chrome 154. 5 minutes. | ||||||
| 5 | QuantDinger | Runs | 100 / 100 | Python | 12,554 | |
Open-source AI Trading OS, agent trading, and vibe trading, with Jev System One integration. Research, build Python strategies, backtest, and paper/live trade... What the test found: QuantDinger Python API v5.4.1 is installed and fully functional, all 3050 unit tests pass, the Flask app creates and serves the health endpoint on HTTP 200, and the only skipped tests require PostgreSQL/integration Docker containers that are not available in this environment. 7 minutes. | ||||||
| 6 | fastapi_mcp | Runs | 100 / 100 | Python | 12,019 | |
Expose your FastAPI endpoints as Model Context Protocol (MCP) tools, with Auth! What the test found: The fastapi-mcp package installs, its full test suite (89 tests) passes, and the MCP StreamableHTTP endpoint responds with a valid initialize handshake (HTTP 200, protocol version 2024-11-05, server capabilities). 5 minutes. | ||||||
| 7 | Graft | Runs | 100 / 100 | TypeScript | 9,771 | |
Turbocharge Claude Code, Cursor, Codex, Gemini & every coding agent: faster, cheaper, with contextual understanding specific to your codebase. What the test found: npm install completes, TypeScript compiles, the CLI builds/checks/queries a structural code graph on a real repo without any keys, and all 1220 tests run with 0 failures. 10 minutes. | ||||||
| 8 | git-mcp | Runs | 100 / 100 | TypeScript | 8,456 | |
Put an end to code hallucinations! GitMCP is a free, open-source, remote MCP server for any GitHub project What the test found: The GitMCP app installs, builds, and serves its web UI on the react-router dev server at HTTP 200, and all 37 unit tests pass. 7 minutes. | ||||||
| 9 | browser-tools-mcp | Runs | 100 / 100 | TypeScript | 7,327 | |
Monitor browser logs directly from Cursor and other MCP compatible IDEs. What the test found: The BrowserTools MCP workspace installs, builds, and passes all 356 unit and integration tests with Node 22.14.0. The MCP server binary prints version 2.0.2, help with 16 tools, and runs --doctor setup checks. 12 minutes. | ||||||
| 10 | engram | Runs | 100 / 100 | Go | 7,087 | |
Persistent memory system for AI coding agents. Agent-agnostic Go binary with SQLite + FTS5, MCP server, HTTP API, CLI, and TUI. What the test found: Engram v2.0.0 builds from source, all 29 test packages pass, CLI commands (save/search/context/timeline/doctor/test) work, and the HTTP API (health/sessions/observations/search) returns correct responses with real SQLite storage. 8 minutes. | ||||||
| 11 | MobileBuildMCP | Runs | 100 / 100 | TypeScript | 6,467 | |
A Model Context Protocol (MCP) server and CLI that provides tools for agent use when working on iOS and macOS projects. What the test found: MobileBuildMCP 2.7.1 builds, typechecks, and passes all 2591 unit tests (posttest also passes: 16/16 warden-watchdog tests). The CLI binary responds with version 2.7.1, lists 72 canonical tools, and starts the MCP server on stdio. All of this was done on Linux (no Xcode/macOS), so the macOS/Xcode-specific tools cannot... 10 minutes. | ||||||
| 12 | gemini-notebook-mcp-cli | Runs | 100 / 100 | Python | 6,252 | |
Programmatic access to Gemini Notebook - via command-line interface (CLI), Model Context Protocol (MCP) server, and AI agent skills. What the test found: Dependencies installed, both CLIs (nlm, notebooklm-mcp) produce help output, MCP server binds to HTTP port, and 1899/1899 unit tests pass. 8 minutes. | ||||||
| 13 | bb-browser | Runs | 100 / 100 | TypeScript | 6,242 | |
Your browser is the API. CLI + MCP server for AI agents to control Chrome with your login state. What the test found: bb-browser pnpm install and build succeed; CLI version 0.14.2 prints help; daemon auto-downloads and launches Chrome 149 headed on Xvfb; HTTP API on port 19826 responds 200 to /status, tab_list, eval, snap, open commands; 129 of 143 tests pass, 13 skipped (need Chrome), 1 flaky timing failure. 22 minutes. | ||||||
| 14 | semble | Runs | 100 / 100 | Python | 6,192 | |
Fast and Accurate Code Search for Agents. Uses 99% fewer tokens than grep+read What the test found: semble 0.5.5 installed in a virtual environment at /work/repo/.venv; all 330 tests pass; CLI and MCP server run correctly on CPU. 2 minutes. | ||||||
| 15 | godot-mcp | Runs | 100 / 100 | JavaScript | 5,976 | |
MCP server for interfacing with Godot game engine. Provides tools for launching the editor, running projects, and capturing debug output. What the test found: Install builds cleanly. MCP server starts and responds to all 13 tools. Real Godot v4.3-stable binary was downloaded and used for verification, 10 tools work end-to-end, 2 (load_sprite) fail due to headless rendering limitation, and 2 (get_uid, update_project_uids) correctly report needing Godot 4.4+. 11 minutes. | ||||||
| 16 | tradingview-mcp | Runs | 100 / 100 | Python | 4,965 | |
TradingView MCP server, real-time market data, technical analysis, screeners & backtesting for Claude, ChatGPT, Cursor & any MCP client. Stocks, crypto, forex... What the test found: The tradingview-mcp-server Python package is installed in a .venv at /work/repo, all 286 unit tests pass, and the MCP server starts and responds to initialize, tools/list (37+ tools), and tools/call on stdio transport. 2 minutes. | ||||||
| 17 | notion-mcp-server | Runs | 100 / 100 | TypeScript | 4,664 | |
Official Notion MCP Server What the test found: npm install completes, npm run build produces bin/cli.mjs, all 109 tests pass, and the HTTP server starts and responds to health checks and JSON-RPC requests on the MCP endpoint. 3 minutes. | ||||||
| 18 | mcp-server-chart | Runs | 100 / 100 | TypeScript | 4,393 | |
🤖 A visualization mcp & skills contains 25+ visual charts using @antvis. Using for chart generation and data analysis. What the test found: All 43 tests pass, the MCP server builds and starts on stdio/SSE/streamable transports, and generates real chart images via the AntV GPT-Vis API. 38 minutes. | ||||||
| 19 | mcpo | Runs | 100 / 100 | Python | 4,393 | |
A simple, secure MCP-to-OpenAPI proxy server What the test found: mcpo installs and runs correctly, proxying a real stdio MCP server (FastMCP) over REST with auto-generated OpenAPI schemas; all 27 unit tests pass, both tool endpoints return correct results via HTTP. 5 minutes. | ||||||
| 20 | mcp-server-cloudflare | Runs | 100 / 100 | TypeScript | 4,362 | |
What the test found: Monorepo installs cleanly with Node 22; full test suite passes; demo-day MCP server starts locally and answers Streamable HTTP tools/list requests. 6 minutes. | ||||||
| 21 | flint-chart | Runs | 100 / 100 | TypeScript | 4,349 | |
🪄 Flint is a visualization language that lets AI agents reliably create expressive, good-looking charts from simple, human-editable chart specs. What the test found: npm install succeeds, all 1739 Vitest tests pass, both JS library and MCP server build to working dist output with functioning CLI binaries. 14 minutes. | ||||||
| 22 | CodeGraphContext | Runs | 100 / 100 | Python | 4,248 | |
An MCP server plus a CLI tool that indexes local code into a graph database to provide context to AI assistants. What the test found: CodeGraphContext 0.6.13 installed and working: CLI, FalkorDB/KuzuDB/LadybugDB backends, indexing, MCP server, and 1615 of 1627 tests pass. 36 minutes. | ||||||
| 23 | excel-mcp-server | Runs | 100 / 100 | Python | 4,215 | |
A Model Context Protocol server for Excel file manipulation What the test found: Install builds cleanly via pip install -e . in a venv. All 5 existing tests pass. The MCP server starts in stdio, SSE, and streamable-http modes. Tools/list returns all 25 tools, and create_workbook/write_data/read_data work end-to-end via the MCP stdio protocol. 4 minutes. | ||||||
| 24 | anything-analyzer | Runs | 100 / 100 | TypeScript | 3,741 | |
全能协议分析工具:浏览器抓包 + MITM 代理 + 指纹伪装 + AI 分析 + MCP Server 无缝对接 AI Agent/IDE | All-in-one protocol analysis toolkit, built-in browser capture, MITM proxy, JS... What the test found: Installation, build, full test suite (185/189 passing), Electron binary, and built app all work in the container. Only expected dbus/GPU errors appear in container. 9 minutes. | ||||||
| 25 | java-sdk | Runs | 100 / 100 | Java | 3,723 | |
The official Java SDK for Model Context Protocol servers and clients. Maintained in collaboration with Spring AI What the test found: MCP Java SDK compiles and all 771 non-Docker tests pass with 0 failures; 12 Docker-dependent tests are skipped. 15 minutes. | ||||||
| 26 | boost | Runs | 100 / 100 | PHP | 3,644 | |
Laravel-focused MCP server for augmenting your AI powered local development experience. What the test found: Laravel Boost library installed with PHP 8.3.33 and Composer 2.10.3; all 1100 tests pass (3257 assertions) and all 5 architecture checks pass. 8 minutes. | ||||||
| 27 | neo | Runs | 100 / 100 | JavaScript | 3,285 | |
Neo.mjs is a self-evolving software organism: a professional end-to-end AI engineering team whose cross-model swarm inhabits live apps via Neural Link, Active... What the test found: Project installed, bundled browser deps, built CSS themes, and ran all 4662 unit tests passing with no failures. 6 minutes. | ||||||
| 28 | fastmcp | Runs | 100 / 100 | TypeScript | 3,273 | |
A TypeScript framework for building MCP servers. What the test found: The project builds and 685+ tests pass across auth, edge, bin, openapi, and FastMCP sub-suites. The edge WebStreamableHTTPServerTransport now correctly imports `randomUUID` from the `crypto` module instead of relying on the global `crypto` variable. Remaining failures are pre-existing: openapi loadSpec hits Node 18... 57 minutes. | ||||||
| 29 | cortex | Runs | 100 / 100 | TypeScript | 3,233 | |
Cortex - Generates interactive API documentation, typed SDKs, and MCP servers from OpenAPI, AsyncAPI, GraphQL, gRPC, OpenRPC, and Markdown. What the test found: Cortex Docs installs, builds, passes its full unit test suite (377 tests, 0 failures), and the CLI successfully initializes a project, validates specs, and generates SDKs in 11 languages plus an MCP server from OpenAPI, AsyncAPI, GraphQL, and OpenRPC sources. 6 minutes. | ||||||
| 30 | arxiv-mcp-server | Runs | 100 / 100 | Python | 3,202 | |
A local MCP server for agent literature work. Original-LaTeX section reads, BibTeX from arXiv metadata, and topic watches. Papers stay on disk. Search is... What the test found: The arxiv-mcp-server Python package builds, all 447 tests pass, and the server launches via HTTP transport responding 200 on /healthz without any modifications to the repository. 2 minutes. | ||||||
| 31 | shadcn-ui-mcp-server | Runs | 100 / 100 | TypeScript | 3,033 | |
A mcp server to allow LLMS gain context about shadcn ui component structure,usage and installation,compaitable with react,svelte 5,vue & React Native What the test found: The shadcn-ui-mcp-server builds from source, starts in both stdio and SSE modes, responds to MCP protocol requests, and fetches real component data from the shadcn-ui/ui GitHub repository without any modifications or mocks. 5 minutes. | ||||||
| 32 | markdownify-mcp | Runs | 100 / 100 | TypeScript | 3,002 | |
A Model Context Protocol server for converting almost anything to Markdown What the test found: Install (bun install, bun run build) complete. All 79 tests pass. MCP server launches via bun start/bun dist/index.js and responds to tools/list and tools/call over stdio. 9 minutes. | ||||||
| 33 | vexa | Runs | 100 / 100 | Python | 2,865 | |
Open-source meeting transcription API for Google Meet, Microsoft Teams & Zoom. Auto-join bots, real-time WebSocket transcripts, MCP server for AI agents... What the test found: 11 Python service test suites and the JS terminal client test suite run successfully in a container with Node 22.13 and Python 3.12.3; the compose/Docker deployment path cannot execute without Docker engine being available. 17 minutes. | ||||||
| 34 | solon | Runs | 100 / 100 | Java | 2,794 | |
🔥 Java enterprise application development framework for full scenario: Restrained, Efficient, Open, Ecologicalll!!! 700% higher concurrency 50% memory savings... What the test found: Solon Java framework: all 23 bundles compile and install; 46 selected unit tests pass cleanly on JDK 17 with Maven 3.9.6. 27 minutes. | ||||||
| 35 | mcp-proxy | Runs | 100 / 100 | Python | 2,775 | |
A bridge between Streamable HTTP and stdio MCP transports What the test found: mcp-proxy 0.12.0 installs cleanly, CLI responds to --version and --help, 91/92 unit tests pass covering CLI argument parsing, config loading, MCP server, HTTP/SSE transports, progress forwarding, and proxy server request forwarding. 3 minutes. | ||||||
| 36 | metamcp | Runs | 100 / 100 | TypeScript | 2,698 | |
MCP Aggregator, Orchestrator, Middleware, Gateway in one docker What the test found: MetaMCP backend server starts on port 12009 with PostgreSQL, serves SSE/MCP/OpenAPI endpoints, authenticates users via email/password, bootstraps users/keys/namespaces/endpoints from env, and responds to all standard MCP protocol methods (initialize, tools/list, prompts/list) successfully. 11 minutes. | ||||||
| 37 | korean-law-mcp | Runs | 100 / 100 | TypeScript | 2,648 | |
법제처 국가법령정보를 LLM에서 바로 조회하는 MCP 서버. 법령·판례·조례 검색과 인용 검증 | MCP server for Korean law, search statutes, precedents, and ordinances, and verify citations What the test found: Korean Law MCP server v4.14.2 builds, all 989 tests pass, STDIO and HTTP modes run correctly with MCP protocol, CLI reports version, 10 tools registered via tools/list. 3 minutes. | ||||||
| 38 | nunu | Runs | 100 / 100 | Go | 2,607 | |
A CLI tool for building Go applications. What the test found: Go 1.27.1 installed at /tmp/go, nunu CLI installed at $HOME/go/bin/nunu, source repo builds and all 4 test packages pass, created testapp project compiles and serves HTTP 200 on :8000, hot-reload and wire injection both functional. 5 minutes. | ||||||
| 39 | mcp_excalidraw | Runs | 100 / 100 | TypeScript | 2,509 | |
MCP server and Claude Code skill for Excalidraw, programmatic canvas toolkit to create, edit, and export diagrams via AI agents with real-time canvas sync. What the test found: Backend TypeScript compiles, npm test suite (MCP wire protocol + canvas bind checks) passes 7/7 checks, canvas server runs serving REST API on :3000, CLI commands (add, describe, export, status) all function correctly creating/managing Excalidraw elements. 27 minutes. | ||||||
| 40 | tavily-mcp | Runs | 100 / 100 | TypeScript | 2,424 | |
Production ready MCP server with real-time search, extract, map & crawl. What the test found: npm install succeeds, TypeScript compiles without errors, all 6 unit tests pass, and the MCP server starts and advertises all 6 tools (tavily_search, tavily_extract, tavily_crawl, tavily_map, tavily_research, tavily_feedback) via JSON-RPC over stdio. 2 minutes. | ||||||
| 41 | kordoc | Runs | 100 / 100 | TypeScript | 2,371 | |
모두 파싱해버리겠다, HWP·HWPX·PDF·Office 문서를 Markdown으로. 양식 자동 채우기와 신구대조를 갖춘 CLI·MCP 서버 | Convert Korean documents (HWP, HWPX, PDF, Office) to Markdown, CLI and MCP... What the test found: kordoc 4.13.1 builds, all 1691 automated tests pass, the CLI parses HWPX and XLS files to Markdown correctly, and the library API returns structured parse results with fileType, markdown, blocks, and metadata. 28 minutes. | ||||||
| 42 | 12306-mcp | Runs | 100 / 100 | JavaScript | 2,319 | |
This is a 12306 ticket search server based on the Model Context Protocol (MCP). What the test found: 12306-mcp v0.3.10 builds and runs: stdio MCP server responds with 8 tools, loads live 12306 station data, and the HTTP/SSE server starts on port 8080. 3 minutes. | ||||||
| 43 | stealth-browser-mcp | Runs | 100 / 100 | Python | 2,216 | |
Stealth browser automation for AI agents over MCP and the Chrome DevTools Protocol: navigation, network hooks, DOM extraction, and pixel-accurate UI cloning. What the test found: Stealth Browser MCP server runs on stdio transport, spawns a headless Chromium 157.0.8086.0, navigates to data URLs, extracts page HTML content, and closes cleanly -- all via MCP protocol over stdio. 15 minutes. | ||||||
| 44 | skybridge | Runs | 100 / 100 | TypeScript | 2,151 | |
Skybridge is a full-stack TypeScript framework for MCP Apps and ChatGPT Apps. Type-safe. React-powered. Platform-agnostic. What the test found: Skybridge monorepo installs, builds, and all 476 unit tests pass across all packages (core, vite-plugin, devtools, create-skybridge, test). The CLI reports skybridge/1.0.0 and responds to --version and --help commands. 44 minutes. | ||||||
| 45 | mcp-memory-service | Runs | 100 / 100 | Python | 1,995 | |
Open-source persistent memory for AI agent pipelines (LangGraph, CrewAI, AutoGen) and Claude. REST API + knowledge graph + autonomous consolidation. What the test found: mcp-memory-service 11.13.0 installed from source via pip install -e '.[dev]', the uvicorn server starts on 127.0.0.1:9791 (HTTP) with sqlite_vec backend, health endpoint returns 200, memory CRUD operations work end to end, and the full test suite passes except for one benchmark test that fails due to hash-embedding... 20 minutes. | ||||||
| 46 | contextplus | Runs | 100 / 100 | TypeScript | 1,983 | |
Semantic Intelligence for Large-Scale Engineering. Context+ is an MCP server designed for developers who demand 99% accuracy. By combining RAG, Tree-sitter... What the test found: npm install + npm run build succeed; all 271 tests pass; the MCP server starts on stdio and shuts down cleanly. 1 minute. | ||||||
| 47 | gortex | Runs | 100 / 100 | Go | 1,870 | |
High-performance code-intelligence engine for AI agents and IDE, supports 257 languages, multi repositories, based on graph, with access via CLI, MCP Server... What the test found: Go 1.27.0 installed locally. Gortex builds from source (v0.64.5) and the daemon indexes this repo into a knowledge graph of 182,637 nodes and 822,458 edges. Core graph, search, parser, resolver, query, analysis, semantic, serverstack, and agent test suites pass. The pre-built release binary v0.64.7 was also verified. 69 minutes. | ||||||
| 48 | trpc-agent-go | Runs | 100 / 100 | Go | 1,850 | |
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability. What the test found: The tRPC-Agent-Go root module (trpc.group/trpc-go/trpc-agent-go) builds and all 235 test suites pass, the test/ module passes with Go 1.24.4 auto-downloaded, and the examples/ module builds successfully, all from a fresh container with Go installed from the upstream tarball. 14 minutes. | ||||||
| 49 | mcp-brasil | Runs | 100 / 100 | Python | 1,807 | |
MCP Server para 70 APIs públicas brasileiras What the test found: All 2566 tests pass, the server responds correctly to MCP initialize (2024-11-05 protocol) with 363 tools registered across 51 features, no code modifications required. 3 minutes. | ||||||
| 50 | docs-mcp-server | Runs | 100 / 100 | TypeScript | 1,789 | |
Grounded Docs MCP Server: Open-Source Alternative to Context7, Nia, and Ref.Tools What the test found: Node 22.23.3, npm install, build, and all 2325 tests pass; CLI --help prints correctly. 8 minutes. | ||||||
| 51 | MeiGen-AI-Design-MCP | Runs | 100 / 100 | TypeScript | 1,781 | |
Supports GPT Image 2, Seedance & ComfyUI, with a 1,400+ prompt library, carefully crafted hooks and a multi-task orchestration system What the test found: Project builds and all 128 library + tool tests pass on Node 18.19.1. The MCP server module loads and the CLI init command responds. The release CI script requires Node 22+ and pnpm which are not available in this container. 56 minutes. | ||||||
| 52 | OpenOSINT | Runs | 100 / 100 | Python | 1,719 | |
AI-powered OSINT agent with interactive REPL, MCP server, and CLI. 20 tools. Works with Claude, GPT-4, or local models. For authorized security research only. What the test found: OpenOSINT 2.29.0 installs from source, CLI shows help with 20 tools, DNS tool resolves real domains, playbook generates real investigation reports, web server serves on port 9877 with HTTP 200 on / and /api/health, and 712 automated tests pass. 7 minutes. | ||||||
| 53 | mcpvault | Runs | 100 / 100 | TypeScript | 1,681 | |
A lightweight Model Context Protocol (MCP) server for safe Obsidian vault access What the test found: All 279 tests pass, the server builds and runs correctly, and read_note returns content from a real vault note via MCP stdio transport. 4 minutes. | ||||||
| 54 | llmwiki | Runs | 100 / 100 | Python | 1,671 | |
Open Source Implementation of Karpathy's LLM Wiki. Upload documents, connect your Claude account via MCP, and have it write your wiki ! What the test found: Python API serves health endpoint on localhost:8000, CLI initializes workspaces, all unit and integration tests pass on SQLite and Postgres backends. 18 minutes. | ||||||
| 55 | mcptools | Runs | 100 / 100 | Go | 1,628 | |
A command-line interface for interacting with MCP (Model Context Protocol) servers using both stdio and HTTP transport. What the test found: The mcp binary builds and runs, all tests pass, and the mock server, tools listing, and tool calling work end-to-end from the CLI. 13 minutes. | ||||||
| 56 | rulego | Runs | 100 / 100 | Go | 1,620 | |
⛓️RuleGo is a lightweight, high-performance, embedded, next-generation component orchestration rule engine framework for Go. What the test found: The rulego Go library compiles with go build and all its test suite passes (go test exits 0, PASS). 8 minutes. | ||||||
| 57 | mcp-remote | Runs | 100 / 100 | TypeScript | 1,618 | |
Connect an MCP Client that only supports local (stdio) servers to a Remote MCP Server. What the test found: All 490 unit/integration tests pass, all 8 concurrent-instance OAuth tests pass with the local OAuth simulator, all 2 real E2E tests connect to Cloudflare and Hugging Face MCP servers, all 4 protocol-era bridge tests pass, and the dist build runs without toWellFormed errors on Node 18. 40 minutes. | ||||||
| 58 | datagouv-mcp | Runs | 100 / 100 | Python | 1,601 | |
Official data.gouv.fr Model Context Protocol (MCP) server that allows AI chatbots to search, explore, and analyze datasets from the French national Open Data... What the test found: Python 3.12 environment with relaxed requires-python; package installs and builds; all 147 tests pass; MCP server starts and responds HTTP 200 on /health. 5 minutes. | ||||||
| 59 | code-mode | Runs | 100 / 100 | TypeScript | 1,592 | |
🔌 Plug-and-play library to enable agents to call MCP and UTCP tools via code execution. What the test found: TypeScript library @utcp/code-mode builds and passes all 29 tests using [email protected] native sandbox. Python library builds and passes all 17 tests using RestrictedPython sandbox. CLI and MCP bridge both start and respond correctly. 17 minutes. | ||||||
| 60 | pi-mcp-adapter | Runs | 100 / 100 | TypeScript | 1,579 | |
Token-efficient MCP adapter for Pi coding agent What the test found: Installed and running on Node 22.23.3. Tests pass 1997/2000, typecheck passes, public exports verified, CLI works. 9 minutes. | ||||||
| 61 | mcp-server-qdrant | Runs | 100 / 100 | Python | 1,543 | |
An official Qdrant Model Context Protocol (MCP) server implementation What the test found: The MCP server builds, all 24 tests pass, and the server responds correctly over stdio transport with initialize, tools/list, tools/call (store and retrieve using Qdrant :memory: mode with fastembed). 3 minutes. | ||||||
| 62 | mysql_mcp_server | Runs | 100 / 100 | Python | 1,398 | |
A Model Context Protocol (MCP) server that enables secure interaction with MySQL databases What the test found: MySQL MCP server installs, builds, and runs with all 26 tests passing against a real MySQL 8.0 instance. STDIO and SSE transport modes verified. All MCP tools (execute_sql, get_schema_info, get_table_sample), prompts (explore_database, analyze_table), and resources are functional. 7 minutes. | ||||||
| 63 | deepwiki-mcp | Runs | 100 / 100 | TypeScript | 1,387 | |
📖 MCP server for fetch deepwiki.com and get latest knowledge in Cursor and other Code Editors What the test found: All 16 tests pass, the MCP server builds and runs on both stdio and HTTP transports, and responds to tools/list with the deepwiki_fetch tool. 4 minutes. | ||||||
| 64 | apple-docs-mcp | Runs | 100 / 100 | TypeScript | 1,381 | |
MCP server for Apple Developer Documentation - Search iOS/macOS/SwiftUI/UIKit docs, WWDC videos, Swift/Objective-C APIs & code examples in Claude, Cursor & AI... What the test found: npm install and tsc build succeed, all 462 tests pass, the MCP server binary launches on stdio and correctly responds to initialize, tools/list, and tools/call with bundled WWDC data (1266 videos, 12 years). 4 minutes. | ||||||
| 65 | Douyin_TikTok_Download_API | Runs | 97.3 / 100 | Python | 20,511 | |
🚀 抖音、TikTok 数据采集与无水印视频下载 API,自托管,支持 MCP 调用与 Docker 一键部署。| Self-hosted TikTok & Douyin scraper and no-watermark video downloader, async REST API, MCP server... What the test found: All 2960 tests pass (2450 unit + 510 integration), the CLI outputs version 5.0.1 and lists 10 commands, the API server starts and serves healthz/readyz/swagger/console root with 200, the web console SPA is built, and both the PostgreSQL+TimescaleDB and Redis backends and the Go downloader sidecar run. 48 minutes. | ||||||
| 66 | drawio-skill | Runs | 97.3 / 100 | Python | 10,003 | |
Agent skill that turns natural language, code, Terraform/K8s, SQL, OpenAPI, AsyncAPI, Protobuf and GraphQL sources into editable, tested draw.io architecture... What the test found: The drawio-skill project runs fully: 174 unit tests pass, the diagramctl CLI (12 subcommands) works end-to-end for core semantic workflows, all importers produce correct IR, the MCP server responds with 9 tools, and the shape/index/validate infrastructure is operational. Only the draw.io desktop binary (for native... 10 minutes. | ||||||
| 67 | github-mcp-server | Runs | 94.7 / 100 | Go | 33,447 | |
GitHub's official MCP Server What the test found: The GitHub MCP Server builds from source, all unit tests pass, and the HTTP server responds to MCP initialize requests with HTTP 200 and a complete capabilities object. 7 minutes. | ||||||
| 68 | Figma-Context-MCP | Runs | 94.7 / 100 | TypeScript | 15,965 | |
MCP server to provide Figma layout information to AI coding agents like Cursor What the test found: The Figma MCP server installs, builds, type-checks, passes all 250 tests, and runs an HTTP MCP server that accepts initialize and tools/list requests on the configured port. 4 minutes. | ||||||
| 69 | hexstrike-ai | Runs | 94.7 / 100 | Python | 12,525 | |
HexStrike AI MCP Agents is an advanced MCP server that lets AI agents (Claude, GPT, Copilot, etc.) autonomously run 150+ cybersecurity tools for automated... What the test found: HexStrike AI MCP Server v6.0 starts on any port, serves Flask API with 30+ endpoints including command execution, file management, intelligence analysis, process management, caching, and telemetry. FastMCP client connects and registers 16+ security tools. 34 minutes. | ||||||
| 70 | Windows-MCP | Runs | 94.7 / 100 | Python | 8,245 | |
MCP Server for Computer Use in Windows What the test found: 285 tests passing, 321 skipped (Windows-only); CLI entry point responds to --help; MCP server with stdio/SSE/streamable-http transport available. 64 minutes. | ||||||
| 71 | gentle-ai | Runs | 93.3 / 100 | Go | 7,599 | |
Gentle-AI configures the AI coding agents you already use: Claude Code, Cursor, OpenCode, Codex, Pi, and more. Choose persistent memory, Organic-Driven... What the test found: gentle-ai 2.0.0 builds from source, prints version (2.0.0-20260906031619-2c25e878eea4) and help text, all 53 Go test packages pass. 17 minutes. | ||||||
| 72 | mcp-language-server | Runs | 93.3 / 100 | Go | 1,608 | |
mcp-language-server gives MCP enabled clients access semantic tools like get definition, references, rename, and diagnostics. What the test found: The MCP language server builds and runs; it initializes as an MCP server, connects to gopls, pyright, typescript-language-server, and rust-analyzer via stdio, and exposes 6 tools (definition, references, diagnostics, hover, rename_symbol, edit_file) that work correctly with real language servers across Go, Python... 62 minutes. | ||||||
| 73 | xiaohongshu-mcp | Runs | 90.7 / 100 | Go | 16,146 | |
MCP for xiaohongshu.com What the test found: The HTTP server starts on port 18060, the /health endpoint returns 200, the built-in Chromium browser launches via go-rod, and the /api/v1/feeds/list endpoint returns real feed data from xiaohongshu.com. All 10 Go test packages pass (0 failures). The browser extraction bug is fixed by using system tar instead of the... 18 minutes. | ||||||
| 74 | firecrawl-mcp-server | Runs | 90.2 / 100 | JavaScript | 7,569 | |
🔥 Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM clients. What the test found: Firecrawl MCP Server v3.24.1 builds, installs, and starts successfully with 25 tools in stdio mode; 81 of 84 tests pass. 24 minutes. | ||||||
| 75 | kubefwd | Runs | 90 / 100 | Go | 4,175 | |
Bulk port forwarding Kubernetes services for local development. What the test found: The kubefwd Go project builds, all 24 test packages pass with race detection and coverage profiling, and the binary reports its version and help text successfully. 7 minutes. | ||||||
| 76 | mcp-grafana | Runs | 90 / 100 | Go | 3,532 | |
MCP server for Grafana What the test found: Go toolchain installed from source, binary builds cleanly, all 973 unit tests pass, and the MCP server starts and responds to tool discovery requests. 7 minutes. | ||||||
| 77 | ripwire | Runs | 90 / 100 | C++ | 2,423 | |
The ripgrep of AI context: a zero-dependency C++23 CLI + MCP server for coding agents. Find what you want without reading the repo, then check you built what... What the test found: Build succeeds (plain dev build/ and Release build-install/), binary parses the repo source tree and emits deterministic minified XML with ranked symbols via Personalized PageRank, installed to /home/runner/.local/bin/ripwire with skills and hooks, --version/--help work, Python indexing confirms correct function... 21 minutes. | ||||||
| 78 | paperbanana | Runs | 90 / 100 | Python | 2,388 | |
Open source implementation and extension of Google Research’s PaperBanana for automated academic figures, diagrams, and research visuals, expanded to new... What the test found: PaperBanana v0.3.0 is installed and its full test suite (901 tests) passes without errors. The CLI responds to all commands. The only missing pieces are external API keys (GOOGLE_API_KEY, OPENAI_API_KEY) which are expected for a fresh install. 16 minutes. | ||||||
| 79 | pg-aiguide | Runs | 90 / 100 | Python | 1,860 | |
MCP server and Claude plugin for Postgres skills and documentation. Helps AI coding tools generate better PostgreSQL code. What the test found: bun i installs 356 packages, tsc -p tsconfig.build.json compiles to dist/, bun test runs 66 passing tests, bun run validate-skills validates all 10 skills, and ./check passes (skipping Python checks for missing uv). 6 minutes. | ||||||
| 80 | ghidra-mcp | Runs | 89 / 100 | Java | 4,838 | |
Ghidra MCP Server, 200+ MCP tools for AI-powered reverse engineering. GUI plugin + headless server, lazy tool loading, convention enforcement, batch... What the test found: Python bridge-mcp-ghidra 7.0.0 installs, all 552 non-slow unit tests pass, MCP server starts on stdio/sse/streamable-http transports and responds to initialize and tools/list with 8 static bridge tools; integration tests require a live Ghidra instance not available in this container. 5 minutes. | ||||||
| 81 | blender-mcp | Runs | 88 / 100 | Python | 30,231 | |
Community plugin to control Blender 3D with any LLM of your choice. Not affiliated with the official Blender Foundation. What the test found: The blender-mcp package is fully installed under uv, all 113 tests pass, and the MCP server launches correctly over stdio. 4 minutes. | ||||||
| 82 | serena | Runs | 88 / 100 | Python | 30,098 | |
A powerful MCP toolkit for coding, providing semantic retrieval and editing capabilities - the IDE for your agent What the test found: Serena is installed, initialized, and the MCP server starts successfully with all 708 tests passing. 12 minutes. | ||||||
| 83 | toolhive | Runs | 88 / 100 | Go | 2,248 | |
ToolHive is an enterprise-grade platform for running and managing Model Context Protocol (MCP) servers. What the test found: ToolHive v0.51.3 builds from source with the standalone Go 1.27.1 compiler. The thv binary starts, prints its help page, and outputs version/build information. Unit tests pass at 198/201 with the 3 failures all due to missing Docker or system keyring in this container. 55 minutes. | ||||||
| 84 | code-graph-rag | Runs | 86.7 / 100 | Python | 5,239 | |
The ultimate RAG for your monorepo. Query, understand, and edit multi-language codebases with the power of AI and knowledge graphs What the test found: code-graph-rag 0.0.985 installed and building from source; CLI responds with version 0.0.985; 185 non-integration tests pass, 45 skip due to missing optional grammars; Docker+Memgraph unavailable without Docker. 55 minutes. | ||||||
| 85 | davinci-resolve-mcp | Runs | 80 / 100 | Python | 3,406 | |
MCP server integration for DaVinci Resolve Studio What the test found: Python compound MCP server starts and answers MCP initialize/tools/list returning 37 tools. Python test suite: 3554 offline tests pass. Node.js resolve-advanced test suite: 913/961 pass (remaining failures from Node.js version gap v18 vs >=20.9). npm install and venv install complete. DaVinci Resolve not available in... 22 minutes. | ||||||
| 86 | ddgs | Runs | 80 / 100 | Python | 3,005 | |
A metasearch library that aggregates results from diverse web search services What the test found: The ddgs package installs, lints, and formats cleanly; the CLI and Python API produce real results for text, images, news, and extract; videos and books tests fail because DuckDuckGo's v.js endpoint and Anna's Archive return 403 from this container. 12 minutes. | ||||||
| 87 | agent-toolkit-for-aws | Runs | 80 / 100 | Python | 2,827 | |
Official, AWS-supported MCP servers, skills, and plugins to help AI agents build on AWS What the test found: All repository validation scripts (manifest, spec conformance, skill sync, markdown lint) pass. 275 embedded skill tests across agent-advisor, cloudwatch-omni, and openclaw-bridge suites all pass. 3 pre-existing test gaps exist in agents-pay (httpx isolation, exit code expectations, Typescript build requirement). 4 minutes. | ||||||
| 88 | paper-search-mcp | Runs | 80 / 100 | Python | 2,765 | |
MCP, CLI, Skills for searching and downloading academic papers from multiple sources like arXiv, PubMed, bioRxiv, etc. What the test found: Package installs, imports, CLI search/results, and MCP server handshake all work. Tests pass modulo transient rate limits from external API servers. 9 minutes. | ||||||
| 89 | apify-mcp-server | Runs | 73.3 / 100 | TypeScript | 10,121 | |
The Apify MCP server enables your AI agents to extract data from social media, search engines, maps, e-commerce sites, or any other website using thousands of... What the test found: Dependencies install, type-check, lint, format, unit tests, and build all pass; the dev server starts on port 3001 and responds to MCP initialize with HTTP 200 and valid server capabilities. 7 minutes. | ||||||
| 90 | codex-with-chatgpt | Runs | 73.3 / 100 | TypeScript | 7,129 | |
ChatGPT thinks. Codex works. Use ChatGPT as the planning brain while keeping the Codex harness. What the test found: The project installs, builds, and passes all 164 tests; the bridge server starts on localhost port 48765 and responds to health checks with status ok. 5 minutes. | ||||||
| 91 | jscpd | Runs | 66.7 / 100 | Rust | 6,357 | |
Copy/paste detector for source code. 220+ languages, Rust engine, SARIF/HTML/badge reporters, GitHub Action, MCP server for AI agents. What the test found: jscpd 5.1.1 builds from source via cargo, installs via npm as a prebuilt binary, and runs duplication detection on 224 language formats with 15 reporters; the full Rust test suite passes. 9 minutes. | ||||||
| 92 | mcp | Runs | 64.8 / 100 | Python | 9,761 | |
Open source MCP Servers for AWS What the test found: 58 AWS MCP server projects install and build with uv; ~15,000 unit tests pass across 50+ servers; the aws-documentation-mcp-server launches, responds to MCP initialize over stdio, and makes real API calls to AWS Documentation Search and docs.aws.amazon.com. 35 minutes. | ||||||
| 93 | fast-agent | Runs | 60 / 100 | Python | 3,928 | |
Code, Build and Evaluate agents - excellent Model and Skills/MCP/ACP/A2A Support What the test found: All 8730 unit tests pass, CLI boots and responds, playback LLM agent runs end-to-end, format/lint/typecheck all pass. 30 minutes. | ||||||
| 94 | nitrostack | Runs | 60 / 100 | TypeScript | 2,465 | |
The full-stack TypeScript framework to build, test, and deploy production-ready MCP servers and AI-native apps. What the test found: All three NitroStack packages (core, cli, widgets) build and all 892 tests pass on Node.js 18.19.1. The CLI binary responds to --help. The core server starts over stdio, initializes, lists tools, and executes tools successfully. 75 minutes. | ||||||
| 95 | fusio | Runs | 60 / 100 | PHP | 2,125 | |
Self-Hosted API Management for Builders What the test found: Fusio v8.8.7 API server running on PHP 8.4 with SQLite at http://127.0.0.1:8084; CLI commands work; backend admin UI at /apps/fusio/ returns 200; system/health endpoint returns healthy: true. 18 minutes. | ||||||
| 96 | mex | Runs | 60 / 100 | TypeScript | 1,757 | |
Team memory for engineers and their AI agents. Lives in your repo. Shared through Git. What the test found: MEX 0.8.2 built, typechecked, and tested with Node 24.0.0 (downloaded). CLI version 0.8.2 confirmed. Test suite passes 4288/4317 tests with 6 pre-existing failures (4 timeouts on constrained hardware, 2 test bugs treating directories as files). npm install, build, and typecheck all succeed. FTS5 full-text search... 26 minutes. | ||||||
| 97 | phantom | Runs | 60 / 100 | TypeScript | 1,480 | |
An AI co-worker with its own computer. Self-evolving, persistent memory, MCP server, secure credential collection, email identity. Built on the Claude Agent... What the test found: 2639 tests pass, the HTTP server starts on port 3100 and returns 302 on GET /, and the CLI phantom doctor command reports the project configuration is valid. 34 minutes. | ||||||
| 98 | cli-agent-orchestrator | Runs | 48 / 100 | Python | 1,397 | |
Multi-agent orchestration for AI coding CLIs, Claude Code, Kiro, Codex, and more, coordinated in isolated tmux sessions What the test found: CAO server starts, responds 200 on /health and /docs, and passes all unit tests (2986+ passed) except for 4 pre-existing test-isolation wiki_lint failures and Node version-gated web UI build. 43 minutes. | ||||||
| 99 | DesktopCommanderMCP | Runs | 40 / 100 | TypeScript | 9,964 | |
This is MCP server for Claude that gives it terminal control, file system search and diff file editing capabilities What the test found: The @wonderwhy-er/desktop-commander MCP server v0.2.48 builds, installs, starts, responds to MCP protocol initialize, and passes its full test suite including real file operations, Excel, PDF, REPL, and process management tests. 9 minutes. | ||||||
| 100 | lean-ctx | Runs | 40 / 100 | Rust | 3,870 | |
LeanCTX, Context Gateway for AI Systems. Control what your AI can see. Open-source Engine for context selection, supported controls, and evidence. What the test found: lean-ctx 3.10.1 release binary is built, installed at ~/.cargo/bin/lean-ctx, and fully functional with all CLI commands working. 65 minutes. | ||||||
| 101 | chrome-devtools-mcp | Runs | 33.3 / 100 | TypeScript | 53,112 | |
Chrome DevTools for coding agents What the test found: The project builds, all tests pass (including e2e), and both CLI and MCP server start successfully using the puppeteer-installed Chrome browser. 51 minutes. | ||||||
| 102 | codebase-memory-mcp | Runs | 33.3 / 100 | C | 46,127 | |
High-performance code intelligence MCP server. Indexes codebases into a persistent knowledge graph, average repo in milliseconds. 158 languages, sub-ms... What the test found: codebase-memory-mcp v0.10.8 is installed, the MCP server is configured in Codex CLI config.toml with hooks, and the repository is indexed with a knowledge graph of 21,005 nodes and 119,361 edges, all 15 MCP tools verified working. 15 minutes. | ||||||
| 103 | whodb | Runs | 33.3 / 100 | Go | 5,031 | |
Where data access meets operational intelligence What the test found: WhoDB backend and CLI binaries build and run, all core unit tests pass, and frontend compiles successfully. 30 minutes. | ||||||
| 104 | mcp-context-forge | Runs | 33.3 / 100 | Python | 4,584 | |
An AI Gateway, registry, and proxy that sits in front of any MCP, A2A, or REST/gRPC APIs, exposing a unified endpoint with centralized discovery, guardrails... What the test found: ContextForge gateway installed in .venv, server launches and responds on HTTP, 592+ unit tests pass against real SQLite and auth infrastructure. 32 minutes. | ||||||
| 105 | rea | Runs with mocks | 92 / 100 | TypeScript | 19,434 | |
Reverse engineer anything with agents, from app behavior down to native binaries. What the test found: REA compiles and runs on Node.js 22.19.0 with npm 11.16.0. The full TypeScript build succeeds, all 353 Vitest test files (1756 individual tests) pass covering domain logic, adapters, composition, MCP boundary, acceptance, conformance, and evaluation layers, and the CLI lists all 60+ analysis commands without error. 8 minutes. | ||||||
| 106 | xiaozhi-esp32-server | Runs with mocks | 92 / 100 | JavaScript | 10,754 | |
本项目为xiaozhi-esp32提供后端服务,帮助您快速搭建ESP32设备控制服务器。Backend service for xiaozhi-esp32, helps you quickly build an ESP32 device control server. What the test found: Python xiaozhi-server installs and runs on ports 8000/8003 with 59/64 pytest tests passing (5 skipped due to placeholder API keys), manager-web Vue UI installs and passes all 62 tests (9 contract + 53 unit), Java manager-api cannot be built (no JDK). 7 minutes. | ||||||
| 107 | lamda | Runs with mocks | 92 / 100 | Python | 8,553 | |
Android Full-Stack Device Control Platform: WebRTC/H.264 remote desktop, UI/OCR/image-matching automation, one-click MITM, built-in Frida, proxy/VPN/frp/P2P... What the test found: lamda Python client v10.8 installs, all modules import, and the gRPC protocol stack works end-to-end via a mocked server. 7 minutes. | ||||||
| 108 | waha | Runs with mocks | 92 / 100 | TypeScript | 7,567 | |
WAHA - WhatsApp HTTP API (REST API) that you can configure in a click! Multiple engines: WEBJS (browser based), NOWEB (websocket nodejs), GOWS (websocket go)... What the test found: The project builds, all 308 unit tests pass across 31 suites, and the server starts and responds to HTTP API calls (GET /api/sessions returns [] with 200 status, logged 'Nest application successfully started' and 'WhatsApp HTTP API is running on: http://[::1]:3000'). 62 minutes. | ||||||
| 109 | magic-mcp | Runs with mocks | 92 / 100 | TypeScript | 5,979 | |
It's like v0, but in your Cursor / Claude Code / Windsurf: search 10,000+ React/Tailwind components, generate new UI with AI, and publish your own, right from... What the test found: The MCP compatibility proxy builds, installs, and runs as a stdio-to-HTTP proxy, correctly handling initialize, tools/list, auth-failure, and auth-latch scenarios against a local mock server. 3 minutes. | ||||||
| 110 | mcp-atlassian | Runs with mocks | 92 / 100 | Python | 5,977 | |
MCP server for Atlassian tools (Confluence, Jira) What the test found: Project installs with uv sync, 3938 unit tests pass, and the MCP server binary starts via SSE transport and prints its version banner. 7 minutes. | ||||||
| 111 | klavis | Runs with mocks | 92 / 100 | Python | 5,807 | |
Klavis AI: MCP integration platforms that let AI agents use tools reliably at any scale What the test found: The strata-mcp package installs and imports cleanly. The CLI parses all commands (add, remove, list, enable, disable, auth, run, tool). Stdio mode starts without errors. HTTP server mode starts on port 8080, and the StreamableHTTP MCP endpoint /mcp/ responds 200 to tools/list requests, returning the Strata tool... 15 minutes. | ||||||
| 112 | zotero-mcp | Runs with mocks | 92 / 100 | Python | 5,284 | |
Zotero MCP: Connects your Zotero research library with Claude and other AI assistants via the Model Context Protocol to discuss papers, get summaries, analyze... What the test found: Zotero MCP v0.13.3 installed in a Python venv, all modules load, CLI responds to help/version commands, and pytest reports 3785 passed, 0 failed, 1 skipped across the non-live test suite after fixing one environment-leak bug in a test fixture. 11 minutes. | ||||||
| 113 | exa-mcp-server | Runs with mocks | 92 / 100 | TypeScript | 5,094 | |
Exa MCP for web search and web crawling! What the test found: The Exa MCP server builds, all 154 unit tests pass, and the stdio binary starts and correctly serves the JSON-RPC protocol with web_search_exa and web_fetch_exa tools declared. 3 minutes. | ||||||
| 114 | aci | Runs with mocks | 92 / 100 | Python | 4,907 | |
ACI.dev is the open source tool-calling platform that hooks up 600+ tools into any agentic IDE or custom AI agent through direct function calling or a unified... What the test found: The ACI backend services (FastAPI server, PostgreSQL with pgvector, AWS KMS mock, Propelauth auth mock) and frontend (Next.js dev portal) install and test successfully in this container without Docker. 159 backend tests pass covering health checks, apps CRUD, functions execution, linked accounts, projects, app... 37 minutes. | ||||||
| 115 | mcp-obsidian | Runs with mocks | 92 / 100 | Python | 4,468 | |
MCP server that interacts with Obsidian via the Obsidian rest API community plugin What the test found: All 81 unit tests pass, the MCP server responds to initialize and tools/list over stdio, and the Obsidian client makes successful HTTP calls against a mock REST API server. 4 minutes. | ||||||
| 116 | hyperresearch | Runs with mocks | 92 / 100 | Python | 3,804 | |
Convert Claude Code or Codex into the most intelligent Deep Research Agent. Collect, search, and synthesize web research into a persistent, searchable wiki... What the test found: All 1719 pytest tests pass, CLI reports version 0.12.0, vault initialization and web server launch succeed, and both `hyperresearch` and `hpr` entry points are functional. 3 minutes. | ||||||
| 117 | py-xiaozhi | Runs with mocks | 92 / 100 | Python | 3,494 | |
Open-source AI assistant ecosystem with MCP integrations, multimodal workflows, IoT support, and cross-platform voice interaction. What the test found: 80 test cases pass in the project test suite; CLI help and CLI mode launch successfully. 16 minutes. | ||||||
| 118 | fli | Runs with mocks | 92 / 100 | Python | 3,213 | |
Google Flights MCP, CLI and Python Library What the test found: fli installs cleanly with uv, all 1404 offline tests pass (29 skipped in container config), the CLI parses and displays airport data, the MCP HTTP server starts and serves /health (200) and /mcp/ (307). 8 minutes. | ||||||
| 119 | ableton-mcp | Runs with mocks | 92 / 100 | Python | 3,151 | |
Control Ableton Live with any LLM: create tracks, arrange clips & compose music via MCP What the test found: Ableton MCP v1.4.5 is installed in a Python 3.12 venv, all 49 unit tests pass, and the MCP server starts and responds to MCP protocol messages over stdio transport against a mock Ableton backend. 5 minutes. | ||||||
| 120 | agent-scan | Runs with mocks | 92 / 100 | Python | 3,125 | |
Security scanner for AI agents, MCP servers and agent skills. What the test found: Project installs with `uv sync`, CLI prints help and version, and the full test suite passes 2293/2326 collected tests. 34 minutes. | ||||||
| 121 | skills | Runs with mocks | 92 / 100 | TypeScript | 3,092 | |
Skills, MCP servers, Custom Agents, Agents.md for SDKs to ground Coding Agents What the test found: The test harness installs and evaluates all 136 skills with mock mode, scoring 96.3% pass rate across 1227 scenarios; vitest unit tests are blocked by a missing native binary in the rolldown dependency. 8 minutes. | ||||||
| 122 | js-reverse-mcp | Runs with mocks | 92 / 100 | TypeScript | 2,910 | |
AI Agent-first JS 逆向 MCP Server:有头 Chrome 调试、断点、网络/WebSocket 分析、Patchright 反检测,可选 CloakBrowser。 What the test found: The js-reverse-mcp v4.0.5 TypeScript project compiles to JavaScript, its full test suite of 114 tests passes, and the MCP server binary responds to JSON-RPC initialize on stdin/stdout. A Node.js v22.12.0 upgrade was applied (portable) and the cloakbrowser optional dependency was installed. Google Chrome is not... 4 minutes. | ||||||
| 123 | brightdata-mcp | Runs with mocks | 92 / 100 | JavaScript | 2,661 | |
A powerful Model Context Protocol (MCP) server that provides an all-in-one solution for public web access. What the test found: The Bright Data MCP server installs and all 21 unit tests pass under Node.js v22.12.0; the server starts listening when API_TOKEN is set. 4 minutes. | ||||||
| 124 | KiCAD-MCP-Server | Runs with mocks | 92 / 100 | Python | 2,618 | |
KiCAD MCP is a Model Context Protocol (MCP) implementation that enables Large Language Models (LLMs) like Claude to directly interact with KiCAD for printed... What the test found: Installation, TypeScript build, and full test suite (112 TS + 2611 Python tests) all pass. Server registers all 244 tools at startup but requires KiCAD's pcbnew module for the Python backend, which is unavailable in this container and is mocked by the test suite's conftest.py. 5 minutes. | ||||||
| 125 | modelcontextprotocol | Runs with mocks | 92 / 100 | TypeScript | 2,553 | |
The official MCP server implementation for the Perplexity API Platform What the test found: npm install builds successfully, all 90 tests pass, HTTP server starts and serves the /health endpoint and MCP tools/list endpoint correctly. 2 minutes. | ||||||
| 126 | mcphub | Runs with mocks | 92 / 100 | TypeScript | 2,503 | |
Self-hosted MCP gateway and control plane for connecting, controlling, and operating MCP servers. What the test found: Backend compiles (tsc exits 0), frontend builds (1902 modules, vite exits 0), 99+ test suites pass (919+ tests, no failures), app server starts on :3000 and responds with 200 on /health, /, and /login; MCP servers report 'degraded' because uvx/npx for fetch/amap/playwright/slack are not installed in this container... 76 minutes. | ||||||
| 127 | agency-orchestrator | Runs with mocks | 92 / 100 | TypeScript | 2,335 | |
🚀 One sentence → your one-person company of AI experts → complete deliverable in minutes. 276 CN + 184 EN + 5 more languages (ko/ru/pt-BR/id/ar) · zero-code... What the test found: Agency Orchestrator 0.19.2 installed and built successfully on Node.js v22, CLI answers version/validate/roles commands, web Studio API serves health/workflows/roles endpoints, and the majority of unit tests pass. 49 minutes. | ||||||
| 128 | AssetOpsBench | Runs with mocks | 92 / 100 | Python | 2,330 | |
AssetOpsBench - Industry 4.0: A unified benchmark and framework for building, orchestrating, and evaluating domain-specific AI agents for Industry 4.0 asset... What the test found: Project installs, builds, and 675 of 702 tests pass; 27 integration tests skip due to missing CouchDB/LLM credentials; all CLI entry points import and display help. 25 minutes. | ||||||
| 129 | kubernetes-mcp-server | Runs with mocks | 92 / 100 | Go | 2,151 | |
Model Context Protocol (MCP) server for Kubernetes and OpenShift What the test found: Go 1.26.4 installed, kubernetes-mcp-server binary built (102MB), all 39 test packages pass, binary runs and prints help/version, runtime requires a Kubernetes kubeconfig. 29 minutes. | ||||||
| 130 | mcp-server-mysql | Runs with mocks | 92 / 100 | JavaScript | 2,145 | |
A Model Context Protocol server that provides read-only access to MySQL databases. This server enables LLMs to inspect database schemas and execute read-only... What the test found: 151 unit tests pass; MCP server launches in both stdio and remote HTTP modes, responding to initialize requests with proper capabilities. 17 minutes. | ||||||
| 131 | DevDocs | Runs with mocks | 92 / 100 | TypeScript | 2,108 | |
Completely free, private, UI based Tech Documentation MCP server. Designed for coders and software developers in mind. Easily integrate into Cursor, Windsurf... What the test found: Backend (FastAPI on :24125) and frontend (Next.js on :3001) both start and respond 200. The discover/crawl/crawl-status API endpoints successfully process a job from initialization through discovery_complete to completed, writing consolidated markdown and metadata files with crawl results including HTTP status codes. 23 minutes. | ||||||
| 132 | boss-agent-cli | Runs with mocks | 92 / 100 | Python | 2,054 | |
🤖 Local-assist BOSS Zhipin CLI for AI agents, search, welfare filtering, shortlist, JSON-envelope output; low-risk & compliant by default. What the test found: The project installs and builds under uv, the whole offline test suite passes 2084/2084, ruff and mypy pass, the boss CLI and the MCP server start and answer, and the CLI's status, detail, search, welfare-filter and wizard commands return correct ok:true JSON envelopes against a local mock of the zhipin API that... 14 minutes. | ||||||
| 133 | azure-devops-mcp | Runs with mocks | 92 / 100 | TypeScript | 2,049 | |
The MCP server for Azure DevOps, bringing the power of Azure DevOps directly to your agents. What the test found: Project builds and all 1349 tests pass with Node.js v20.20.2. The MCP server binary reports version 2.10.0 and responds to --help with all options (organization, domains, authentication, tenant). The dist/ directory contains compiled JS output. 3 minutes. | ||||||
| 134 | ADR | Runs with mocks | 92 / 100 | Python | 1,940 | |
ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber. What the test found: All three Python packages (adr-benchmark, adr-sensor, adr-discovery) install and pass their test suites (802 total tests, 0 failures). The llamaFirewall detector pipeline runs against packed benchmark data with a mock API key, scoring conversations and writing analysis results. CLI entry points for Discovery and... 11 minutes. | ||||||
| 135 | mcp-gsc | Runs with mocks | 92 / 100 | Python | 1,879 | |
Google Search Console Insights with Claude AI for SEOs What the test found: All 52 unit tests pass with mocked GSC API. The server loads, registers 21 MCP tools, completes MCP initialization handshake, responds to tools/list and tools/call via stdio transport, and serves 200 on the SSE endpoint. 3 minutes. | ||||||
| 136 | slack-mcp-server | Runs with mocks | 92 / 100 | Go | 1,861 | |
The most powerful MCP Slack Server with no permission requirements, Apps support, GovSlack, DMs, Group DMs and smart history fetch logic. What the test found: slack-mcp-server builds, unit tests pass (30+ tests), and the server starts in demo mode serving SSE and HTTP MCP endpoints with all tools registered. 15 minutes. | ||||||
| 137 | Dive | Runs with mocks | 92 / 100 | TypeScript | 1,827 | |
Dive is an open-source MCP Host Desktop Application that seamlessly integrates with any LLMs supporting function calling capabilities. ✨ What the test found: Installed with Node v22.14.0 and uv 0.12.23; `npm run build:electron` and `npm run check` exit 0; 287 mcp-host pytest pass; Electron app launches on Xvfb display with host reporting state UP and its HTTP API answering 200 on three endpoints. 27 minutes. | ||||||
| 138 | telegram-mcp | Runs with mocks | 92 / 100 | Python | 1,792 | |
Telegram MCP server powered by Telethon to let MCP clients read chats, manage groups, and send/modify messages, media, contacts, and settings. What the test found: The project installed via uv sync with all 70 dependencies, the 791-test suite passes, the session generator and migration scripts produce valid help output, and the MCP server imports and initializes its Telegram client infrastructure correctly when env vars are set (real Telegram credentials required for full... 4 minutes. | ||||||
| 139 | cve-mcp-server | Runs with mocks | 92 / 100 | Python | 1,619 | |
Production-grade MCP server giving Claude 27 security intelligence tools across 21 APIs, CVE lookup, EPSS scoring, CISA KEV, MITRE ATT&CK, Shodan, VirusTotal... What the test found: 30/30 pytest tests pass, server starts in both stdio and streamable-http transport, NVD rate-limit warning confirmed expected without API key. 2 minutes. | ||||||
| 140 | mcp-server-kubernetes | Runs with mocks | 92 / 100 | TypeScript | 1,596 | |
MCP Server for kubernetes management commands What the test found: 26 unit test files pass (297 tests), the build compiles cleanly, the MCP server responds to initialize and tool calls, and all tool interactions are handled through a mock kubectl since no real cluster is available. 14 minutes. | ||||||
| 141 | MiniMax-MCP | Runs with mocks | 92 / 100 | Python | 1,583 | |
Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs. What the test found: Package installed in a venv, all 9 pre-existing tests pass (3 server + 6 utils), and the MCP server binary launches and waits for stdio input. 2 minutes. | ||||||
| 142 | terraform-mcp-server | Runs with mocks | 92 / 100 | Go | 1,545 | |
The Terraform MCP Server provides seamless integration with Terraform ecosystem, enabling advanced automation and interaction capabilities for Infrastructure... What the test found: Build compiles, binary runs, stdio MCP mode responds correctly to JSON-RPC, streamable-http mode starts and binds port. 16/16 test packages pass, 1 e2e package fails due to missing Docker in environment. 4 minutes. | ||||||
| 143 | duckduckgo-mcp-server | Runs with mocks | 92 / 100 | Python | 1,526 | |
A Model Context Protocol (MCP) server that provides web search capabilities through DuckDuckGo, with additional features for content fetching and parsing. What the test found: All 136 tests pass (128 unit + 8 e2e), ruff lint passes, and the server starts on streamable-http transport and prints its configuration banner and listen address. 1 minute. | ||||||
| 144 | ros-mcp-server | Runs with mocks | 92 / 100 | Python | 1,490 | |
Connect AI models like Claude & GPT with robots using MCP and ROS. What the test found: ros-mcp 3.1.2 installed in /tmp/venv, CLI responds with usage, MCP server runs over stdio/http with full tool/resource/prompt registry, 12/12 unit tests pass, tools verified against mock rosbridge WebSocket server. 60 minutes. | ||||||
| 145 | gpt-researcher | Runs with mocks | 89.8 / 100 | Python | 29,952 | |
An autonomous agent that conducts deep research on any data using any LLM providers What the test found: The GPT Researcher package is installed in a Python venv. The FastAPI backend server launches, listens on 0.0.0.0:8000, and serves the GPT Researcher frontend (HTTP 200). The CLI help works. 410 of 411 tests pass under --forked isolation; 2 integration tests require real OpenAI credentials and are skipped. No code was... 23 minutes. | ||||||
| 146 | AiSOC | Runs with mocks | 88 / 100 | Python | 2,389 | |
Open-source AI Security Operations Center: alert fusion, LLM-agent triage, MITRE ATT&CK investigation, and a replayable decision ledger for every agent step... What the test found: All 7 Python packages install and their 249 tests pass. aisoc-sandbox runs a 4-step investigation offline in 1.3s. aisoc-lite CLI triages 200 demo alerts and translates Sigma rules. Docker is unavailable so the full production stack (make up -> make smoke) cannot run. 11 minutes. | ||||||
| 147 | ios-simulator-mcp | Runs with mocks | 86 / 100 | JavaScript | 2,190 | |
MCP server for interacting with the iOS simulator What the test found: The ios-simulator-mcp server builds, starts, and responds to all 17 MCP tools correctly on Linux with mocked external dependencies. Real functionality requires macOS with Xcode, iOS Simulators, and Facebook IDB. 5 minutes. | ||||||
| 148 | ida-pro-mcp | Runs with mocks | 82 / 100 | Python | 12,534 | |
AI-powered reverse engineering assistant that bridges IDA Pro with language models through MCP. What the test found: Package installs cleanly, all 255 unit tests pass, both stdio and HTTP MCP server modes respond with valid initialize responses, CLI help/config/list-clients all work. 4 minutes. | ||||||
| 149 | obsidian-local-rest-api | Runs with mocks | 82 / 100 | TypeScript | 3,007 | |
A secure REST API and Model Context Protocol (MCP) server for your vault. What the test found: npm install and npm run build succeed. All 1015 unit tests pass after fixing certificate key size; 1 test remains blocked on Node 18 ESM limitation. 17 minutes. | ||||||
| 150 | open-webSearch | Runs with mocks | 77 / 100 | TypeScript | 1,846 | |
Multi-engine MCP server, CLI, and local daemon for agent web search and content retrieval, skill-guided workflows, no API keys. What the test found: npm install and npm run build compile cleanly; 35/38 tests pass (3 live tests excused for external network limits); the daemon starts, health endpoint returns HTTP 200, and DuckDuckGo search returns real results; Playwright-dependent features are unavailable because the container runs Node 18 (requires 20+). 16 minutes. | ||||||
| 151 | buildwithclaude | Runs with mocks | 76 / 100 | Python | 3,604 | |
A single hub to find Claude Skills, Agents, Commands, Hooks, Plugins, and Marketplace collections to extend Claude Code, Claude Desktop, Agent SDK and OpenClaw What the test found: npm install, validation of 294 agents/249 commands/40 hooks/89 skills, and 22 unit tests all pass after patching the test runner for Node 18 compatibility; the web-ui Next.js build and generate-registry scripts are blocked by missing Node 20+ and 3 absent script modules. 11 minutes. | ||||||
| 152 | google_workspace_mcp | Runs with mocks | 72 / 100 | Python | 3,303 | |
Control Gmail, Google Calendar, Docs, Sheets, Slides, Chat, Forms, Tasks, Search & Drive with AI - Comprehensive Google Workspace MCP Server & CLI Tool What the test found: workspace-mcp 1.29.0 installs, builds, launches in both stdio and streamable-http modes, serves a working MCP endpoint with 43 registered tools, passes 2660+ tests with mock Google API credentials, and both CLI entry points (workspace-mcp and workspace-cli) respond to --help. 31 minutes. | ||||||
| 153 | pipeshub-ai | Runs with mocks | 68 / 100 | Python | 3,818 | |
The open-source context layer for AI agents. PipesHub turns your company's knowledge (Slack, Drive, Jira, GitHub, Microsoft 365 and 40+ connectors) into a... What the test found: 286 Python unit tests pass, 9633 JavaScript unit tests pass (4 pending). Both suites execute against mocked external services without real infrastructure dependencies. 62 minutes. | ||||||
| 154 | XcodeBuildMCP | Runs with mocks | 61.3 / 100 | TypeScript | 6,467 | |
A Model Context Protocol (MCP) server and CLI that provides tools for agent use when working on iOS and macOS projects. What the test found: XcodeBuildMCP v2.7.0 installs, builds (npm run build), passes all 2591 unit/integration tests across 243 test files, passes typecheck, and responds to CLI commands --help and tools on Node 18.19.1. 14 minutes. | ||||||
| 155 | linkedin-mcp-server | Runs with mocks | 61.3 / 100 | Python | 3,784 | |
Open-source MCP server for LinkedIn. Give Claude and any MCP-compatible AI agent access to profiles, companies, jobs, and messages. What the test found: Project installs cleanly with uv sync, server imports OK, test suite passes all non-environment-dependent tests (2921 of 3014 selected pass with the fix applied; the 11 failures are pre-existing container/TERM mismatches). 30 minutes. | ||||||
| 156 | context7 | Runs with mocks | 58.5 / 100 | TypeScript | 62,797 | |
Context7 Platform -- Up-to-date code documentation for LLMs and AI code editors What the test found: Build succeeds on all 6 packages. Test suites pass for CLI (361), MCP (59), SDK (27 with mock), and PI (4). The tools-ai-sdk has 11 of 16 tests passing; 5 require real AWS Bedrock credentials. The ctx7 CLI binary and MCP server both launch and respond correctly. 10 minutes. | ||||||
| 157 | agentset | Runs with mocks | 56 / 100 | TypeScript | 2,105 | |
The open-source RAG platform: built-in citations, deep research, 22+ file formats, partitions, MCP server, and more. What the test found: Agentset monorepo installs dependencies with bun 1.4.2 (using Bun's bundled Node 26.3.0), generates Prisma client with all enums, and passes 616 of 643 tests across web and engine packages. 42 minutes. | ||||||
| 158 | ouroboros | Runs with mocks | 44.3 / 100 | Python | 6,194 | |
Agent OS: the agent gets smarter on its own. We just hold the line: Interview-gated, staged evaluation, budgeted evolution loop. MCP server, 14 runtimes... What the test found: All unit tests pass (0 failures in ~21000 tests across 34 unit directories), ruff lint and format clean, CLI boots with correct version, MCP server help displays correctly, 2 Windows-specific bugs were caught and fixed. 73 minutes. | ||||||
| 159 | cursor-talk-to-figma-mcp | Runs with mocks | 44 / 100 | JavaScript | 7,047 | |
TalkToFigma: MCP integration between AI Agent (Cursor, Claude Code, Codex) and Figma, allowing Agentic AI to communicate with Figma for reading designs and... What the test found: Dependencies installed, the project builds with tsup, the WebSocket relay starts on port 3055, and the MCP server starts on stdio, initializes, connects to the relay, and responds to all MCP protocol messages (initialize, tools/list, prompts/list, tools/call). 8 minutes. | ||||||
| 160 | mcp-chrome | Runs with mocks | 30.7 / 100 | TypeScript | 12,473 | |
Chrome MCP Server is a Chrome extension-based Model Context Protocol (MCP) server that exposes your Chrome browser functionality to AI assistants like Claude... What the test found: The monorepo builds successfully (packages/shared, app/chrome-extension, app/native-server), all 642 tests pass, the CLI responds to help/doctor commands, and the MCP stdio server starts without errors. 12 minutes. | ||||||
| 161 | radar | Runs with mocks | 30.7 / 100 | Go | 3,684 | |
The missing open-source Kubernetes UI with a built-in MCP server for AI agents. See what's broken, why, and what changed. Issues, Topology, event timeline... What the test found: Radar builds from source and runs; frontend is built and embedded, Go binary serves the UI at localhost:9282, and all unit tests pass. 21 minutes. | ||||||
| 162 | ha-mcp | Runs with mocks | 23 / 100 | Python | 4,976 | |
The Unofficial and Awesome Home Assistant MCP Server What the test found: Dependencies installed, smoke test passes, and the full unit test suite passes (11,790 passed, 357 skipped, 0 failed). 70 minutes. | ||||||
| 163 | deep-research | Could not verify | 80 / 100 | JavaScript | 4,694 | |
Use any LLMs (Large Language Models) for Deep Research. Support SSE API and MCP server. What the test found: pnpm dependencies installed, next build completed with all 25 route handlers compiled, and the production server serves the Deep Research Next.js app at localhost:3000 responding HTTP 200 with a fully rendered page. 17 minutes. | ||||||
| 164 | azure-skills | Could not verify | 80 / 100 | Shell | 1,549 | |
Official agent plugin providing skills and MCP server configurations for Azure scenarios. What the test found: Landing page builds with 0 errors and serves HTTP 200; all JSON metadata files validate; all 56 SKILL.md files are present; npm dependencies install cleanly with node 22. 4 minutes. | ||||||
| 165 | FinanceToolkit | Could not verify | 50 / 100 | Python | 5,404 | |
Transparent and Efficient Financial Analysis What the test found: FinanceToolkit 2.2.0 is installed in a Python 3.12 venv, all 1472 tests pass, and both MCP CLI entry points (financetoolkit-mcp, financetoolkit-mcp-setup) respond correctly. 27 minutes. | ||||||
| 166 | open-seo-mcp-skills | Could not verify | 50 / 100 | Shell | 4,554 | |
Free SEO MCP server + open-source SEO and GEO skills for Claude: keyword research, rank tracking, audits, backlinks, AI visibility on your real GSC/GA4/ads... What the test found: All 8 SKILL.md files installed to ~/.claude/skills/ with correct content. The repository is a collection of Markdown skill definitions for Claude Code that rely on the external Ryze MCP connector, no local application to launch or tests to execute. 6 minutes. | ||||||
| 167 | httprunner | Could not verify | 20 / 100 | Go | 4,297 | |
HttpRunner 是一款开源的 API/UI 测试框架,简单易用,功能强大,具有丰富的插件化机制和高度的可扩展能力。 What the test found: could not be verified; the log shows where it stopped. 86 minutes. | ||||||
| 168 | dbhub | Could not verify | 20 / 100 | TypeScript | 3,617 | |
Token conscious database MCP server for Postgres, MySQL, SQL Server, Oracle, MariaDB, SQLite. What the test found: could not be verified; the log shows where it stopped. 12 minutes. | ||||||
| 169 | brave-search-mcp-server | Could not verify | 20 / 100 | TypeScript | 1,488 | |
What the test found: could not be verified; the log shows where it stopped. | ||||||
| 170 | drawio-mcp-server | Could not verify | 20 / 100 | TypeScript | 1,479 | |
Draw.io Model Context Protocol (MCP) Server What the test found: could not be verified; the log shows where it stopped. | ||||||
| 171 | agentgateway | Could not verify | 0 / 100 | Rust | 5,228 | |
Next Generation Agentic Proxy for AI Agents and MCP servers What the test found: could not be verified; the log shows where it stopped. 87 minutes. | ||||||
| 172 | Unla | Could not verify | 0 / 100 | TypeScript | 2,239 | |
🧩 MCP Gateway - A lightweight gateway service that instantly transforms existing MCP Servers and APIs into MCP servers with zero code changes. Features Docker... What the test found: could not be verified; the log shows where it stopped. 87 minutes. | ||||||
| - | llm-wiki-compiler | Not yet tested | - | TypeScript | 2,169 | - |
| - | MCPJungle | Not yet tested | - | Go | 1,293 | - |
| - | thClaws | Not yet tested | - | Rust | 1,238 | - |
| - | grafbase | Not yet tested | - | Rust | 1,228 | - |
| - | free4chat | Not yet tested | - | TypeScript | 1,211 | - |
| - | northcinder | Not yet tested | - | JavaScript | 1,208 | - |
| - | agentdock | Not yet tested | - | Go | 1,205 | - |
| - | tuui | Not yet tested | - | TypeScript | 1,155 | - |
| - | innovation-lab-examples | Not yet tested | - | Python | 1,148 | - |
| - | solidworks-automation-skill | Not yet tested | - | Python | 1,106 | - |
| - | atlassian-mcp-server | Not yet tested | - | JavaScript | 1,086 | - |
| - | stackql | Not yet tested | - | Go | 1,066 | - |
| - | minima | Not yet tested | - | Python | 1,051 | - |
| - | arcade-mcp | Not yet tested | - | Python | 1,047 | - |
| - | anymd | Not yet tested | - | Rust | 1,028 | - |
| - | figwright | Not yet tested | - | TypeScript | 987 | - |
| - | mcpc | Not yet tested | - | TypeScript | 986 | - |
| - | mcp-sequential-thinking | Not yet tested | - | Python | 953 | - |
| - | MCP-Bridge | Not yet tested | - | Python | 927 | - |
| - | mcp-notion-server | Not yet tested | - | TypeScript | 920 | - |
| - | skills | Not yet tested | - | Python | 918 | - |
| - | iai-personal-memory-engine | Not yet tested | - | Python | 902 | - |
| - | awesome-mcp-servers | Not yet tested | - | TypeScript | 885 | - |
| - | projectmem | Not yet tested | - | Python | 854 | - |
| - | mcp-nixos | Not yet tested | - | Python | 851 | - |
| - | reddit-mcp-buddy | Not yet tested | - | TypeScript | 845 | - |
| - | memorix | Not yet tested | - | TypeScript | 836 | - |
| - | supabase-mcp-server | Not yet tested | - | Python | 831 | - |
| - | mcp-client-for-ollama | Not yet tested | - | Python | 827 | - |
| - | scira-mcp-chat | Not yet tested | - | TypeScript | 825 | - |
| - | second-brain-cloudflare | Not yet tested | - | TypeScript | 803 | - |
| - | jadx-mcp-server | Not yet tested | - | Python | 793 | - |
| - | okf-agent-memory | Not yet tested | - | Go | 756 | - |
| - | memora | Not yet tested | - | Python | 731 | - |
| - | VibeUE | Not yet tested | - | C++ | 726 | - |
| - | after-effects-mcp | Not yet tested | - | JavaScript | 714 | - |
| - | MCP-Nest | Not yet tested | - | TypeScript | 714 | - |
| - | sandbase-harness | Not yet tested | - | TypeScript | 704 | - |
| - | maverick-mcp | Not yet tested | - | Python | 701 | - |
| - | obsidian-mcp-server | Not yet tested | - | TypeScript | 693 | - |
| - | Deuz-SDK | Not yet tested | - | TypeScript | 687 | - |
| - | mcp-client-cli | Not yet tested | - | Python | 675 | - |
| - | Adobe_Premiere_Pro_MCP | Not yet tested | - | TypeScript | 666 | - |
| - | world-intel-mcp | Not yet tested | - | Python | 657 | - |
| - | python-utcp | Not yet tested | - | Python | 653 | - |
| - | biomcp | Not yet tested | - | Rust | 648 | - |
| - | trinity | Not yet tested | - | Python | 629 | - |
| - | ironcurtain | Not yet tested | - | TypeScript | 613 | - |
runs installed and started with its real dependencies. runs with mocks started after stand-ins replaced external services such as a database or a third-party API. could not verify neither the standard agent nor the stronger one got it running within the time limit; the log shows where it stopped.
Before you choose
Decide what the server needs to talk to. Servers that work on a local resource started for real and are "runs": serena (708 tests passing), blender-mcp, semble, DesktopCommanderMCP (real file, Excel, PDF, and process operations in its suite), and codebase-memory-mcp, which indexed the repository into a graph of 21,005 nodes. Wrappers around a remote service could only answer the protocol with a placeholder credential, because the agent never receives real third-party keys; they are "runs with mocks": github-mcp-server, exa-mcp-server, magic-mcp, Figma-Context-MCP, and cursor-talk-to-figma-mcp.
Browser-driven servers are the hard case. playwright-mcp ran with its bundled Chromium once the agent resolved the missing system libraries, while xiaohongshu-mcp's pre-compiled Chromium crashed in the container, so its browser features stayed unverified. Seven servers could not be verified within the time limit, even by the stronger agent: context7, chrome-devtools-mcp, mcp-chrome, XcodeBuildMCP, whodb, ha-mcp, and radar; each log shows where the attempt stopped.
Check the transport you need. Figma-Context-MCP and tradingview-mcp ran in both stdio and HTTP modes; firecrawl-mcp-server started in stdio mode with 25 tools; github-mcp-server answered over HTTP.
Look at the minutes. Most Python and Node servers were up in 4 to 9 minutes (blender-mcp, magic-mcp, tradingview-mcp, exa-mcp-server, Figma-Context-MCP, cursor-talk-to-figma-mcp, DesktopCommanderMCP, ghidra-mcp); hexstrike-ai and gpt-researcher took 34 and 38. mcp-context-forge and linkedin-mcp-server have not been tested yet and sit unranked.
How we tested
An MCP server counts as working when the process starts on the clean machine and answers the protocol: the agent sent an initialize request and asked for the tool list over stdio or HTTP, and the project's own tests ran.
Servers that work on a local resource started for real: serena, blender-mcp, semble, and hexstrike-ai. Servers that exist to wrap a remote service could only be exercised with a placeholder credential or a stand-in backend, so they are "runs with mocks": github-mcp-server answered initialize and tools/list with a fake token, exa-mcp-server and magic-mcp responded to JSON-RPC initialization without a live account, and ghidra-mcp ran its Python bridge without the Ghidra SDK.

The agent never receives real third-party keys, so a wrapper around a paid API cannot score higher than "runs with mocks" here. That is a limit of the method, stated on purpose, not a flaw in the project.

On this list as of the latest test: 104 projects ran as-is, 58 with mocks, 10 could not be verified, 48 still waiting. Languages tested: C, C++, Go, Java, JavaScript, PHP, Python, Rust, Shell, TypeScript. Every attempt used a clean single-use machine, the subject at a pinned version, and a 45-minute limit; the complete procedure is on the methodology page.
Frequently asked questions (FAQs)
What does "runs" mean for an MCP server?
The process started on a clean machine and answered the protocol: an initialize request and a tools/list request over stdio or HTTP, with the project's own tests run. serena and blender-mcp are examples.
Why are wrappers around GitHub, Exa, or similar services "runs with mocks"?
The agent never receives real third-party keys. github-mcp-server answered initialize and tools/list with a fake token; exa-mcp-server responded to JSON-RPC initialization without a live account. The server works; the remote side was not exercised.
Does the test check the tools the server exposes?
It checks that the server starts, answers initialize, and lists its tools, plus whatever the project's tests cover. It does not call every tool against a real backend.
Are SDKs and frameworks for building MCP servers on this list?
No. Libraries with nothing to launch on their own were removed from the list on purpose; the list is for servers you can start.
How current is this list?
Each row carries its own test date and the page states the window. A project can change after its test; a new attempt replaces the verdict only when it finishes with complete evidence.
How is this list ranked?
By measurement, not opinion: projects Argusic installed and launched on a fresh machine come first, then those that ran with mocks in place of external services, then those it could not verify. Ties go to the Argusic Score, then how popular it is on its own source.
Why are some projects unranked?
48 projects are still waiting for a test or for a finished attempt. They are listed without a rank until Argusic has measured them.
More lists in this category
- self-hosted AI apps (shares Skill_Seekers, VibeUE, anything-analyzer, gortex, mcp-nixos, obsidian-mcp-server, pipeshub-ai, projectmem, stealth-browser-mcp, trinity, world-intel-mcp with this list)
- open source coding agents (shares cli-agent-orchestrator, gentle-ai, lean-ctx, memorix, okf-agent-memory, ouroboros, pi-mcp-adapter, rea, ripwire, serena with this list)
- AI agent frameworks (shares Deuz-SDK, Graft, fast-agent, gentle-ai, lamda, rea, trpc-agent-go with this list)
- developer CLI tools (shares Graft, OpenOSINT, jscpd, mcp-client-for-ollama, ouroboros, ripwire, thClaws with this list)
- web scraping tools (shares Douyin_TikTok_Download_API, SeleniumBase, Skill_Seekers, brightdata-mcp, firecrawl-mcp-server, stealth-browser-mcp with this list)
- API gateways (shares agentgateway, fusio, mcp-context-forge with this list)
- browser automation tools (shares brightdata-mcp, js-reverse-mcp, stealth-browser-mcp with this list)
- observability tools (shares mcp-context-forge, trpc-agent-go with this list)
- speech tools (shares MiniMax-MCP, vexa with this list)
- vector databases (shares iai-personal-memory-engine, mcp-memory-service with this list)
- open source video editors (shares Adobe_Premiere_Pro_MCP, after-effects-mcp with this list)
- wikis (shares llm-wiki-compiler, projectmem with this list)
- workflow automation tools (shares agentdock, rulego with this list)
- API clients (shares cortex with this list)
- LLM gateways (shares agentgateway with this list)
- home automation tools (shares ha-mcp with this list)
- low-code platforms (shares rulego with this list)
- Open source office suites, installed and launched (shares kordoc with this list)
- whiteboard tools (shares drawio-skill with this list)
- CI/CD tools
- CMS platforms
- VPN tools
- backup tools
- code editors
- Open source databases, installed and queried
- e-commerce platforms
- ebook readers
- game engines
- open source games
- map tools
- message queues
- Open source music servers, installed and played
- self-hosted password managers
- project management tools
- screen recorders
- search engines
- self-hosted analytics
- static site generators
- uptime monitors
- video players
- self-hosted Notion alternatives
- self-hosted dashboards
- self-hosted git servers