Open source workflow automation, installed and run

Workflow automation tools you can run yourself. The list is ordered by whether each project came up when Argusic installed it fresh, with the recording of that attempt one click away.

Tested between and . Each row shows its own test date; a project can change after that day.

34 of 36 tested projects run. 61 more waiting for a test.

In short: 26 of the 36 tested projects started as-is on a fresh machine: conductor, invisible_dots, temporal, cadence, flowgram.ai, loopx, vibe-coding-prompt-template, and argo-events, and 18 more. 8 more started once a stand-in replaced a service they expect, such as a database: nanobot, ODS, open-claude-cowork, social-media-research-skills, danghuangshang, ai-moive-studio, n8n-skills, and harbor. 2 could not be verified: CodeMachine-CLI and selfhost-ai; the log shows where each one stopped.

Measured by Argusic on a fresh machine every time. Every number links to its evidence.

#projectverdictArgusic Scorelanguagestarstested on
1conductorRuns100 / 100Java32,268

Conductor is an event driven agentic workflow engine providing durable and highly resilient execution engine for applications and AI Agents

What the test found: Conductor server runs on port 8080 with SQLite backend, listens on /api, answers health checks, registers and executes workflows, and all 1,365 non-Docker-dependent tests pass across core, common, rest, grpc, grpc-server, grpc-client, server, and sqlite-persistence modules. 13 minutes.

2invisible_dotsRuns100 / 100TypeScript31,843

Open-source, self-hosted alternative to OpenAI Dots, Grok Bot. Built to be undetectable by anti-bot systems.

What the test found: npm ci installs 428 packages without error; npm run build produces the CLI binary (apps/cli/dist/invisible-dots.mjs) and the server binary (apps/web/.next/standalone/apps/web/server.js). The test suite passes 2103 tests across 150 files (0 failures, 3 skipped). The server starts, creates its PGlite database, applies... 49 minutes.

3temporalRuns100 / 100Go23,537

Temporal service

What the test found: Temporal server 1.33.0 builds from source with Go 1.27.0, starts with SQLite in-memory persistence, exposes gRPC on 4 ports (frontend:7233, history:7234, matching:7235, worker:7239), responds SERVING on health check, and launches system workflows. 5 binaries built. Core unit tests pass. 55 minutes.

4cadenceRuns100 / 100Go9,480

Cadence is a distributed, scalable, durable, and highly available orchestration engine to execute asynchronous long-running business logic in a scalable and...

What the test found: Cadence server compiled all 8 binaries, installed SQLite persistence schemas, started all 4 services (frontend, matching, history, worker) into RUNNING state, and passed all unit tests across the codebase. 23 minutes.

5flowgram.aiRuns100 / 100TypeScript8,487

FlowGram is an extensible workflow development framework with built-in canvas, form, variable, and materials that helps developers build AI workflow platforms...

What the test found: FlowGram monorepo builds successfully (73/73 Rush projects), all tests pass across 9 packages with vitest (0 failures), and the demo-free-layout app serves correctly via Rsbuild dev server answering HTTP 200. 15 minutes.

6loopxRuns100 / 100Python6,202

A control plane with a durable state kernel for long-horizon agents and teams. Keep work moving and improving across sessions, with less human attention.

What the test found: LoopX 1.2.0 installed from source in venv with Node.js 22; CLI responds, doctor passes, demo workspace serves on localhost, 125/125 TypeScript control plane tests pass, canary and smoke suites pass, main Python test suite collects 12206 tests with 3 pre-existing failures (stale registry manifest, missing git worktree... 59 minutes.

7vibe-coding-prompt-templateRuns100 / 100TypeScript3,131

Templates and workflow for generating PRDs, Tech Designs, and MVP and more using LLMs for AI IDEs

What the test found: The vibeworkflow CLI builds and passes all 27 unit tests; the scaffold correctly fills app name, one-liner, target users, phase, tech stack, and commands from PRD/TechDesign metadata; the CLI detects tools, generates AGENTS.md, CLAUDE.md, agent_docs, and skills, and surfaces remaining placeholders. 6 minutes.

8argo-eventsRuns100 / 100Go2,698

Event-driven Automation Framework for Kubernetes

What the test found: Argo Events CLI binary builds and runs, all 78 unit test suites pass with 0 failures. 12 minutes.

9yao-meta-skillRuns100 / 100Python2,692

YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.

What the test found: All 87 CI tests pass, the CLI validates the project correctly, and the Review Studio reports decision 'review' with score 86. 21 minutes.

10shepherdRuns100 / 100Python2,476

A runtime substrate that turns an agent's execution into a reversible, Git-like trace, so meta-agents can observe, fork, replay, and revert any run. Couples...

What the test found: All 50 workspace packages installed via uv sync; integration test suite reports 77/82 passed (4 skipped due to missing native jail, 1 expected failure for fuse-overlayfs check); core package tests 553/554 passed (1 slow-hypothesis flake); shepherd2 tests 186/186 passed; offline quickstart runs successfully end-to-end. 33 minutes.

11doitRuns100 / 100Python2,090

CLI task management & automation tool

What the test found: doit v0.38.dev0 installed and fully functional, all 830 tests pass and the CLI executes incremental build tasks correctly with file dependency tracking, up-to-date skipping, and clean operations. 2 minutes.

12YoutarrRuns100 / 100JavaScript1,730

Self-hosted web app that automates downloading, organizing, and scheduling YouTube channel content with support for Plex, Kodi, Emby and Jellyfin

13rulegoRuns100 / 100Go1,620

⛓️RuleGo is a lightweight, high-performance, embedded, next-generation component orchestration rule engine framework for Go.

What the test found: The rulego Go library compiles with go build and all its test suite passes (go test exits 0, PASS). 8 minutes.

14n8n-as-codeRuns100 / 100TypeScript1,597

Give your AI agent n8n superpowers. 537 nodes with full schemas, 7,700+ templates, Git-like sync, and TypeScript workflows.

What the test found: npm install --ignore-scripts completes with warnings only, tsc -b TypeScript compilation succeeds, CLI version 2.7.0 responds to all commands, and 6 test suites pass (922 total tests, 0 failures across transformer, CLI, skills, MCP, openclaw, and vscode packages). 19 minutes.

15domain-lockerRuns100 / 100TypeScript1,536

🌐 The all-in-one tool, for keeping track of your domain name portfolio. Got domain names? Get Domain Locker!

What the test found: Dependencies installed, all 37 tests pass, production build completes, SSR server serves /api/health (200), / (200), /login (200) correctly. 9 minutes.

16obseiRuns100 / 100Python1,434

Obsei is a low code AI powered automation tool. It can be used in various business flows like social listening, AI based alerting, brand image analysis...

What the test found: Obsei library installs, all 45 tests pass (including real model inference with transformers), and example scripts run successfully producing analysis output. 12 minutes.

17utaskRuns100 / 100Go1,408

µTask is an automation engine that models and executes business processes declared in yaml. ✏️📋

What the test found: uTask 1.24.0 builds into a working ELF binary (44MB), starts an HTTP server on port 8081 that serves the OpenAPI spec, responds 200 pong on health endpoint, and all 14 test packages pass on a fresh PostgreSQL 16 database. 16 minutes.

18sequential-workflow-designerRuns96.7 / 100TypeScript1,475

Customizable no-code component for building flow-based programming applications or workflow automation. 0 external dependencies. Check out https://nocode-js.com

What the test found: Core designer builds completely (lib/cjs, lib/esm, dist/umd) and passes all 112 unit tests via karma+jasmine on Chrome Headless 130. 13 minutes.

19Vibe-SkillsRuns93.5 / 100Python3,606

Intelligent Skill routing and workflow orchestration for AI agents, +21.12 pp reward, −29.6% tokens on SkillsBench with DeepSeekV4Flash-VE.

What the test found: VibeSkills v4.1.0 installs successfully, check verifies all 292 receipt-owned files, the vgo_cli Python launcher works end-to-end, and 129+ core unit tests pass with 3 isolated test failures due to missing pwsh in this container. 37 minutes.

20hiveRuns90 / 100Python11,087

Multi-Agent Harness for Production AI

What the test found: All Python dependencies installed and both framework and tools packages build successfully. Core test suite (2500 tests) passes cleanly. Tools test suite (6840 of 7261 tests) passes with one pre-existing endpoint mismatch failure. The hive CLI launches and displays help text. 20 minutes.

21LaunchStackRuns90 / 100TypeScript890

AI-powered StartUp Accelerator Engine built with Next.js, LangChain, PostgreSQL + pgvector. Upload, organize, and chat with documents. Includes predictive...

What the test found: LaunchStack monorepo installs cleanly after Node upgrade to v20.11.0; package-level vitest (327/327) and web jest (hundreds of unit+integration tests) all pass against a standalone PostgreSQL 16 with pgvector; only missing real external credentials for chat, S3, and provider APIs. 33 minutes.

22astron-rpaRuns88.2 / 100Python5,256

Agent-ready RPA suite with out-of-the-box automation tools. Built for individuals and enterprises.

What the test found: Engine Python 3.13.15 dependencies installed, engine modules import correctly, 207 pytest tests passing across data processing and encryption components, 1 source bug fixed (string fill slice type). Frontend shared and CLI packages build. Can't run backend services or Electron desktop app without Windows/Docker. 22 minutes.

23OpenAdaptRuns80 / 100Python1,762

Compiles a demonstrated GUI task into a program that reports VERIFIED only if an independent check agrees. pip install openadapt; openadapt flow tutorial...

What the test found: Package installs and all 288 tests pass; CLI, doctor, version commands work; recording generates valid artifacts; compile is CPU-bound by onnxruntime in this container but eventually produces bundles. 20 minutes.

24edictRuns60 / 100Python16,976

🏛️ 三省六部制 · OpenClaw Multi-Agent Orchestration System, 9 specialized AI agents with real-time dashboard, model config, and full audit trails

What the test found: All 68 tests pass on Python 3.12.3; the dashboard HTTP server serves the React frontend, health check, and all API endpoints correctly. 6 minutes.

25Auto-CompanyRuns60 / 100Python3,120

An auto-company works for 24/7 on your own PC - Windows/Linux/macOS.

What the test found: All 419 Python tests pass, all 4 shell test suites pass, all 19 frontend JS tests pass, dashboard server starts and serves HTTP 200 with structured JSON on / and /api/status. 9 minutes.

26certimateRuns50 / 100Go9,355

An open-source and free self-hosted SSL certificates ACME tool, automates the full-cycle of issuance, deployment, renewal, and monitoring visually. 完全开源免费的自托管...

What the test found: Certimate v0.4.33 builds from source with Go 1.26.8, serves web UI responding HTTP 200, and runs all 3 test packages successfully. 17 minutes.

27nanobotRuns with mocks92 / 100Python48,864

Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-agent workflows, automation, and chat...

What the test found: nanobot v0.3.0 installed from source in a venv, all 6014 automated tests pass, the gateway launches successfully, and ruff linting passes cleanly. 28 minutes.

28ODSRuns with mocks92 / 100Python7,147

ODS V3 Pre-Release: Public testing and refinement ahead of the official V3 launch. Turn your PC, Mac, or Linux box into a private AI server.

What the test found: Project builds correctly, all 418 bats tests, 14/14 smoke tests (after ai_err fix), ~300 shell/Python contract tests, and the vite frontend build pass. The Docker runtime install cannot complete in this environment, but the dry-run installer, CLI (-v2.6.0), all standalone test suites, and the frontend are functional. 15 minutes.

29open-claude-coworkRuns with mocks92 / 100JavaScript4,288

Open Source version of Claude Cowork with 500+ SaaS app integrations

What the test found: Backend server starts and responds on all API endpoints (health 200, chat 200 with SSE streaming, abort 200, validation errors 400); Electron desktop app launches and prints 'Electron app ready'. 10 minutes.

30social-media-research-skillsRuns with mocks92 / 100Python3,319

AI agent skills for social media research. Outlier posts, comment mining, competitor teardowns, ad libraries & trends across TikTok, Instagram, YouTube...

What the test found: All 13 SKILL.md files have valid frontmatter and non-empty bodies, the npx skills CLI installs the repo correctly (13 skills listed under .agents/skills/), and the validate-skills.py script passes cleanly. 3 minutes.

31danghuangshangRuns with mocks92 / 100TypeScript2,703

Open-source multi-agent collaboration system inspired by Chinese governance, deploy and coordinate specialized AI agents with OpenClaw.

What the test found: npm dependencies install cleanly, all 19 Jest tests pass, both manual test suites pass, OpenClaw Gateway runs and responds HTTP 200 on dashboard and health endpoints with a mock LLM provider configured. 19 minutes.

32ai-moive-studioRuns with mocks92 / 100Python1,613

自然语言驱动的无限画布工作流 Agent,让 AI 视频创作第一次真正变成可编辑的工作流。 AICON 面向创作者,提供从剧本拆解、分镜生成、素材生成、视频合成到内容分发的一整套能力。 不是只给你一个输入框,而是让你用自然语言和无限画布一起驱动创作,把文本、图片、视频节点组织成完整链路,真正把“从灵感到...

What the test found: Backend FastAPI server starts and serves root, health, and API v1 endpoints on port 8000. Frontend builds to dist/. 57 backend unit tests and 15 frontend utility tests pass. 22 minutes.

33n8n-skillsRuns with mocks86 / 100Shell6,393

n8n skillset for Claude Code to build flawless n8n workflows

What the test found: n8n-skills builds 16 valid distribution zips (98 files, 15 skills), n8n-mcp v2.90.0 MCP server responds to all 7 tools correctly via stdio protocol, skills install into ~/.codex/skills/, and all pre/post-tool-use hooks are executable. 8 minutes.

34harborRuns with mocks66 / 100Python3,240

Stop configuring your AI stack. Start using it. One command brings a complete pre-wired LLM stack with hundreds of services to explore.

What the test found: Harbor CLI v0.5.6 installed from source, frontend web app builds, all Deno-based unit tests (41 total) pass, lint self-test green (12 bash rules, 9 compose rules, 1 boost rule, 3 orchestrator checks). Docker-dependent container test matrix blocked. 16 minutes.

35CodeMachine-CLICould not verify75 / 100TypeScript2,511

CodeMachine is an open-source tool that orchestrates AI coding agents into repeatable, long-running workflows. ⚡️

What the test found: CodeMachine CLI v0.8.0 installs via 'bun install' (591 packages), responds to all CLI subcommands (help, version, auth, run, step, templates, agents, mcp, import, export), and launches its TUI interface. No test suite exists in the repository. 5 minutes.

36selfhost-aiCould not verify15 / 100Shell942

One-command installer for a self-hosted AI stack on your own server: n8n, Ollama, Open WebUI, OpenClaw, Dify, Flowise, Supabase, ComfyUI, Qdrant & 30+ tools...

What the test found: The repository is structurally valid (Docker Compose YAML, shell scripts, JSON configs, Python files) but cannot install or launch because Docker is absent and cannot be installed without root privileges in this container. 13 minutes.

-amneziawg-installerNot yet tested-Shell1,353-
-SuggestArrNot yet tested-Python1,334-
-autoprompt-skillNot yet tested-JavaScript1,298-
-workflowNot yet tested-PHP1,247-
-agentdockNot yet tested-Go1,205-
-gh-action-pypi-publishNot yet tested-Python1,184-
-BubbleLabNot yet tested-TypeScript1,097-
-sdk-goNot yet tested-Go981-
-eclaireNot yet tested-TypeScript923-
-dbos-transact-golangNot yet tested-Go850-
-sleepless-agentNot yet tested-Python833-
-compass-skillsNot yet tested-Python751-
-romeNot yet tested-TypeScript728-
-second-brainNot yet tested-Python666-
-Vibe-WorkflowNot yet tested-JavaScript613-
-smart-ralphNot yet tested-Shell557-
-Open-Workflow-LibraryNot yet tested-Python556-
-n8n-workflow-builderNot yet tested-JavaScript547-
-forge-filmNot yet tested-Python538-
-sdk-rustNot yet tested-Rust529-
-WA-AKGNot yet tested-TypeScript505-
-vibe-check-mcp-serverNot yet tested-TypeScript502-
-opencrewNot yet tested-Shell498-
-10xProductivityNot yet tested-Python478-
-QuantumNot yet tested-TypeScript475-
-LinghunNot yet tested-TypeScript474-
-agentsNot yet tested-Python451-
-quickyNot yet tested-JavaScript438-
-sdk-javaNot yet tested-Java434-
-GameDesignOSNot yet tested-Python413-
-sutandoNot yet tested-Python396-
-graspNot yet tested-Go395-
-workflowbuilderNot yet tested-TypeScript391-
-codifyNot yet tested-C387-
-roteNot yet tested-Go376-
-UniEmployeeNot yet tested-Python358-
-flowctlNot yet tested-Go341-
-homelabNot yet tested-TypeScript323-
-pneumaticworkflowNot yet tested-Python320-
-TwitchDropsMinerNot yet tested-Python316-
-metisNot yet tested-TypeScript313-
-open-vettaNot yet tested-TypeScript291-
-lazyskillsNot yet tested-Go270-
-foraNot yet tested-Python264-
-foraNot yet tested-Python264-
-huddolNot yet tested-Python264-
-GoGogotNot yet tested-Go263-
-workflowNot yet tested-Go260-
-awesome-hermes-usecasesNot yet tested-Python253-
-deplao-builderNot yet tested-TypeScript247-
-mcp-n8n-workflow-builderNot yet tested-JavaScript234-
-portable-hermes-agentNot yet tested-Python232-
-open-flowNot yet tested-TypeScript226-
-ComfyUI-fal-APINot yet tested-Python223-
-LoomFlowNot yet tested-TypeScript222-
-grok-reg-toolNot yet tested-TypeScript215-
-dsh-research-reportNot yet tested-TypeScript214-
-openclaw-setupNot yet tested-JavaScript214-
-dsh-industry-researchNot yet tested-TypeScript213-
-compartmentNot yet tested-TypeScript205-
-flowbakerNot yet tested-Go205-

runs installed and started with its real dependencies. runs with mocks started after stand-ins replaced external services such as a database or a third-party API. could not verify neither the standard agent nor the stronger one got it running within the time limit; the log shows where it stopped.

How we tested

On this list as of the latest test: 26 projects ran as-is, 8 with mocks, 2 could not be verified, 61 still waiting. Languages tested: Go, Java, JavaScript, Python, Shell, TypeScript. Every attempt used a clean single-use machine, the subject at a pinned version, and a 45-minute limit; the complete procedure is on the methodology page.

Frequently asked questions (FAQs)

How is this list ranked?

By measurement, not opinion: projects Argusic installed and launched on a fresh machine come first, then those that ran with mocks in place of external services, then those it could not verify. Ties go to the Argusic Score, then how popular it is on its own source.

Why are some projects unranked?

61 projects are still waiting for a test or for a finished attempt. They are listed without a rank until Argusic has measured them.

Where is the evidence?

Every row links to the project's Argusic page, where each run has a full log and a terminal recording stored with a sha256 fingerprint. The same pages exist for every one of the tested projects, on this list or not.

More lists in this category

All lists: Best. All tested projects: subjects.