cli-anything-ollama
Command-line interface for Ollama - Local LLM inference and model management via Ollama REST API. Designed for AI agents and power users who need to manage models, generate text, chat, and create embeddings without a GUI.
By hkuds · 463 installs
npx skills add hkuds/cli-anything --skill cli-anything-ollama
Source repository · Upstream listing
cli anything ollama
Local LLM inference and model management via the Ollama REST API. Designed for AI agents and power users who need to manage models, generate text, chat, and create embeddings without a GUI.
Installation
This CLI is installed as part of the cli anything ollama package:
Prerequisites:
Python 3.10+
Ollama must be installed and running ( ollama serve )
Usage
Basic Commands
REPL Mode
When invoked without a subcommand, the CLI enters an interactive REPL session:
Command Groups
Model
Model management commands.
Command Description
list List locally available models
show Show model details (parameters, template, license)
pull Download a model from the Ollama library
rm Delete a model from local storage
copy Copy a model to a new name
ps List models currently loaded in memory
Generate
Text generation and chat commands.
Command Description
text Generate text from a prompt
chat Send a chat completion request
Embed
Embedding generation commands.
Command Description
text Generate embeddings for text
Server
Server status and info commands.
Command Description
status Check if Ollama server is running
version Show Ollama server version
Session
Session state commands.
Command Description
status Show current session state
history Show chat history for current session
Examples
List and Pull Models
Generate Text
Chat
Embeddings
Interactive REPL Session
Start an interactive session for exploratory use.
Connect to Remote Host
State Management
The CLI maintains lightweight session state:
Current host URL : Configurable via host
Chat history : Tracked for multi turn conversations in REPL
Last used model : Shown in REPL prompt
Output Formats
All commands support dual output modes:
Human readable (default): Tables, colors, formatted text
Machine readable ( json flag): Structured JSON for agent consumption
For AI Agents
When using this CLI programmatically:
1. Always use json flag for parseable output
2. Check return codes 0 for success, non zero for errors
3. Parse stderr for error messages on failure
4. Use no stream for generate/chat to get complete responses
5. Verify Ollama is running with server status before other commands
More Information
Full documentation: See README.md in the package
Test coverage: See TEST.md in the package
Methodology: See HARNESS.md in the cli anything plugin
Version
1.0.1