mrkeyoor.com_
Tue 29 Sept 18:25 UTC
AI Toolsevaluationupdated 27 Aug 2026

CLI-Anything review

CLI-Anything is a collection of agent skills, plugins, generated command-line wrappers, and a registry for controlling existing software. It asks a coding agent to inspect an application, build a Python CLI around it, test the commands, and package the result for tools such as Claude Code, Codex, Cursor, and OpenClaw.

+969stars / 7d
Verdict

Our CLI-Hub build passed in 11 seconds, but 7 of 164 tests failed because analytics code could not determine a home directory in the sandbox. CLI-Anything is worth studying when one stubborn desktop or specialist app needs an agent-facing interface. Treat every generated harness as application code that needs review, destructive-path testing, and a real end-to-end run before an agent touches important data.

We ran it

Lab card: what happened when we ran CLI-AnythingScreenshot of CLI-Anything (clianything.cc)
Install✓ · 26s37 packages · 38 MB
Build✓ · 3s
Tests✗ · 10s157 passed · 7 failed of 164 (pytest)
Known vulns0(pip-audit)
Repo1872 files~321,103 lines of source · 63 MB · 7 CI workflows · tests dir

Answers from our run

Does CLI-Anything build from source?

Dependencies installed in 26 seconds (37 packages), and the build succeeded in 3 seconds. We cloned commit 34f5195 into a clean Debian container with 3 CPUs and no project-specific setup.

Do CLI-Anything's tests pass?

Not all of them: 157 of 164 passed and 7 failed when we ran the project's own test command (pytest). Some failures need services or credentials a bare container does not have.

Does CLI-Anything have known vulnerabilities in its dependencies?

pip-audit found none in the dependency tree at the time of our run.

Who should not use CLI-Anything?

Anyone expecting generated code to be trustworthy without review: the README says weaker models can produce incomplete or incorrect CLIs and repeated refinement may be needed.

What are the alternatives to CLI-Anything?

Open Interpreter, Aider, OpenHands. Our CLI-Hub build passed in 11 seconds, but 7 of 164 tests failed because analytics code could not determine a home directory in the sandbox.

Setup3/5Build passed; 7 analytics tests failed in the sandbox
Docs4/5Agent installs and the 7-phase method are documented in depth
Community5/548,322 stars with active August fixes and bug reports
Maturity2/5v0.4.0 spans many wrappers, but damaging gaps remain open

Discussed on

  1. hnCLI-Anything4 points

Who it’s for

Agent developers who need structured JSON commands around software that exposes no suitable CLI.
Teams willing to inspect and repair generated Python before giving it access to real files or applications.
Claude Code, Codex, Cursor, OpenClaw, and Pi users who want a shared harness method.
Contributors prepared to test each wrapper against the exact upstream application and operating system they use.

Who it’s NOT for

Anyone expecting generated code to be trustworthy without review: the README says weaker models can produce incomplete or incorrect CLIs and repeated refinement may be needed.
Users wrapping closed binaries with no accessible source: the documented method says coverage degrades when decompilation is required.
Teams treating an exit code of 0 as proof of a completed action: open issue 451 reports a Blender render command that produced no image and still exited successfully.
Users putting important notes behind the current Obsidian wrapper: open issue 454 reports that note open can overwrite the active note.

Setup reality

Our sandbox installed 35 packages in 28 seconds from the cli-hub project, using 36 MB on disk, and built it in 11 seconds. Tests failed after 14 seconds: 157 passed and 7 failed of 164 because analytics tests could not determine a home directory.

Using the hub needs Python 3.10 or newer. Generating a new wrapper also needs the target source and a supported coding agent; using a generated wrapper may require the real upstream application, credentials, plugins, or local services it controls.

Installation differs by agent, and Windows Claude Code needs Bash through Git for Windows or WSL. Generated code can write files or call application APIs, so run validation and inspect destructive commands before exposing a harness to valuable projects.

The output is a reusable CLI, not a one-off agent action

CLI-Anything turns source access into an application-specific Python command line. Its method asks an agent to map GUI actions to callable APIs, design command groups and state, implement Click commands with JSON output, write unit and end-to-end checks, document the result, and package it. The generated harness can then be called by a person, a shell script, or another agent without repeating the source-analysis work for every task.

The companion CLI-Hub handles discovery, installation, updates, launching, and removal. It mixes project-built harnesses with selected public CLIs, and v0.4.0 added matrices that install several tools for a workflow. Many wrappers still depend on the real application or backend. Installing a GIMP, Blender, LibreOffice, or Obsidian entry does not install away the upstream program, its plugins, API keys, file formats, or platform quirks.

Claude Code gets a plugin and each harness gets a skill

Claude Code users add the GitHub marketplace and install the CLI-Anything plugin, then invoke /cli-anything against a path or repository. Codex, Cursor, Pi, OpenCode, OpenClaw, Hermes, and other agents have separate adapters. Each generated harness also receives a SKILL.md, which helps compatible agents discover its commands and constraints rather than guessing from --help alone.

That packaging has had real drift. Open issue 433 reports that an npx skills add installation omitted the references directory containing HARNESS.md and command specifications. The installed skill instructed agents to read files that were not present, leaving only condensed guidance. The repository's Codex installer vendors the reference material differently. Before generating anything, verify that the chosen agent installation includes the full method and mode-specific command instructions.

What happened when we ran it

Our sandbox tested commit 810c18b from the cli-hub subproject. Installation took 28 seconds, adding 35 Python packages and 36 MB on disk. The build succeeded in 11 seconds. The full checkout contained 1,872 files, about 321,103 source lines, and occupied 63 MB. We found 7 CI workflow files, no Dockerfile, and a tests directory. Pip-audit reported 0 known vulnerabilities.

Pytest exited with code 1 after 14 seconds: 157 tests passed and 7 failed of 164. Every listed failure was in TestAnalytics, and every traceback ended with RuntimeError: Could not determine home directory. The failures covered event sending, install and uninstall event names, launch, and human or agent visits. The log does not establish why a home directory was unavailable. It does establish that the measured commit's hub suite did not pass in our fresh Python 3.12 Debian container.

A successful command can still produce no artifact

The project's own method says export commands should verify output magic bytes, archive structure, pixels, audio levels, or duration rather than trusting exit status. Open issue 451 shows why. In the reported Blender wrapper, render execute generated a script and printed the command it would run, returned exit code 0, but never started Blender and never produced the requested PNG. Pull request 453 was opened to invoke Blender during execution.

Issue 454 is more serious because it concerns data loss. The reported Obsidian note open implementation sent a PUT request to the active-note endpoint with the target path as the request body. That endpoint replaces the current note's contents, so the active note became the literal path string instead of opening the requested file. A wrapper around a familiar application can make a wrong API call just as easily as any handwritten integration.

The generator depends on source access and model judgment

Python 3.10 or newer is the base requirement. A new harness also needs the target program or repository and a supported coding agent. The README explicitly says strong foundation models are needed for dependable generation, weaker models may emit incomplete or incorrect CLIs, closed-source binaries reduce coverage, and one pass may not cover the application. Those limitations are unusually candid and should shape the workflow.

Generated code should enter the same review process as a human contribution. Check authentication, path handling, command quoting, output verification, rollback, and what delete, overwrite, publish, or send really do. Run it first on disposable data. The 7-phase method provides useful structure, yet phase names cannot certify that a generated test asserted the right behavior against a current upstream application.

August bug activity shows a living but uneven catalog

GitHub recorded 48,322 stars, 83 combined issues and pull requests, and a last push on August 21, 2026. Current work includes fixes for Blender execution, preview bundle selection, skill output paths, and harness CI. Release v0.4.0 arrived on June 25 with new catalog entries, workflow matrices, security fixes, and contributions from many new authors.

Catalog size creates a maintenance burden because every upstream application can change independently. Open issue 403 argued that advertised harness results were not reproduced by CI; a later watchdog CI pull request addressed suite drift, but the issue remained open. CLI-Anything is most credible as a method and starting codebase. The installed wrapper still has to earn trust against the precise app version, operating system, files, and side effects that matter to you.

Alternatives

ProjectWhat it isPick it when
Open Interpreter gh↗An agent that executes code locally to operate the computer and complete tasks.pick this instead when direct local code execution is enough and you do not need a reusable application-specific CLI.
Aider gh↗A terminal coding agent focused on editing and testing source repositories.pick this instead when the job is changing code rather than packaging another application's controls.
OpenHands gh↗A development-agent platform with a sandboxed runtime and extensible tools.pick this instead when you need an agent runtime for software tasks rather than a catalog of generated CLIs.

What people are saying

  1. [github-trending] HKUDS/CLI-Anything

Sources

  1. CLI-Anything README
  2. CLI-Anything v0.4.0 release
  3. Harness CI issue 403
  4. Blender execution issue 451
  5. Obsidian overwrite issue 454

More ai tools reviews

voltagent · InferenceX · Bonsai-demo · qwen-audio-agent · wechat-intelligence-hub · dlss5-visual-enhancer · the whole board →