cli-anything-ollama

Command-line interface for Ollama - Local LLM inference and model management via Ollama REST API. Designed for AI agents and power users who need to manage models, generate text, chat, and create embeddings without a GUI.

By hkuds · 463 installs

npx skills add hkuds/cli-anything --skill cli-anything-ollama

Source repository · Upstream listing

cli anything ollama Local LLM inference and model management via the Ollama REST API. Designed for AI agents and power users who need to manage models, generate text, chat, and create embeddings without a GUI. Installation This CLI is installed as part of the cli anything ollama package: Prerequisites: Python 3.10+ Ollama must be installed and running ( ollama serve ) Usage Basic Commands REPL Mode When invoked without a subcommand, the CLI enters an interactive REPL session: Command Groups Model Model management commands. Command Description list List locally available models show Show model details (parameters, template, license) pull Download a model from the Ollama library rm Delete a model from local storage copy Copy a model to a new name ps List models currently loaded in memory Generate Text generation and chat commands. Command Description text Generate text from a prompt chat Send a chat completion request Embed Embedding generation commands. Command Description text Generate embeddings for text Server Server status and info commands. Command Description status Check if Ollama server is running version Show Ollama server version Session Session state commands. Command Description status Show current session state history Show chat history for current session Examples List and Pull Models Generate Text Chat Embeddings Interactive REPL Session Start an interactive session for exploratory use. Connect to Remote Host State Management The CLI maintains lightweight session state: Current host URL : Configurable via host Chat history : Tracked for multi turn conversations in REPL Last used model : Shown in REPL prompt Output Formats All commands support dual output modes: Human readable (default): Tables, colors, formatted text Machine readable ( json flag): Structured JSON for agent consumption For AI Agents When using this CLI programmatically: 1. Always use json flag for parseable output 2. Check return codes 0 for success, non zero for errors 3. Parse stderr for error messages on failure 4. Use no stream for generate/chat to get complete responses 5. Verify Ollama is running with server status before other commands More Information Full documentation: See README.md in the package Test coverage: See TEST.md in the package Methodology: See HARNESS.md in the cli anything plugin Version 1.0.1