Simple CLI to chat with GGUF models locally (no Ollama/LM Studio required)
Project description
ai-launcher-cli
Simple Python CLI to chat with .gguf models locally using llama-cpp-python. No Ollama, llama.cpp, LM Studio, or online providers required.
Install
pip install ai-launcher-cli
Usage
# Basic usage
ailaunch path/to/model.gguf
# With custom settings
ailaunch model.gguf -c 8192 -t 0.8 --max-tokens 1024
# Disable streaming (wait for full response)
ailaunch model.gguf --no-stream
# Custom system prompt
ailaunch model.gguf --system "You are a coding assistant."
# Use a built-in system prompt template
ailaunch model.gguf --system-template coder
# List available models
ailaunch --list-models
# Auto-select model from common directories
ailaunch auto
# Options:
# -c, --ctx-size Context window size (default: 4096)
# -g, --gpu-layers GPU layers to offload (-1 = all, default: -1)
# -t, --threads CPU threads (0 = auto, default: 0)
# --temperature Sampling temperature (default: 0.7)
# --max-tokens Max tokens to generate (default: 512)
# --no-stream Disable streaming output
# --system Custom system prompt
# --system-template Built-in template (coder, reviewer, teacher, creative, analyst, translator, shell)
# --list-models List available GGUF models and exit
# --save-config Save current options as defaults
# --benchmark Run benchmark after loading
# --export Export conversation on exit (markdown/json)
# --export-file File to export conversation to
# --no-history Disable loading/saving chat history
# --clear-history Clear chat history for this model
# -v, --version Show version
Interactive Commands
While chatting, type any of these commands:
| Command | Description |
|---|---|
/help |
Show help |
/save |
Save conversation to history |
/export [fmt] |
Export conversation (markdown/json) |
/clear |
Clear conversation (keep system prompt) |
/system <prompt> |
Change system prompt |
/template <name> |
Use built-in template |
/config |
Show current configuration |
/bench |
Run benchmark |
/models |
List available models |
/switch [path] |
Switch to another model |
exit/quit/q |
Exit |
Configuration
Config is saved to ~/.config/ailaunch/config.yaml. Use --save-config to save current options.
Model Auto-Detection
Models are automatically searched in these directories:
~/.lmstudio/models~/.lmstudio/.internal/bundled-models~/.cache/huggingface/hub~/models~/Downloads~/OneDrive/Downloads~/OneDrive/Documents/Downloads
GPU Acceleration
Install with GPU extras for acceleration:
# NVIDIA CUDA
pip install ai-launcher-cli[cuda]
# Apple Metal
pip install ai-launcher-cli[metal]
Then use -g -1 to offload all layers to GPU.
Requirements
- Python 3.8+
llama-cpp-python>=0.3.0(installs automatically)
Exit
Type exit, quit, q or press Ctrl+C to exit.
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Filter files by name, interpreter, ABI, and platform.
If you're not sure about the file name format, learn more about wheel file names.
Copy a direct link to the current filters
File details
Details for the file ai_launcher_cli-0.1.6.tar.gz.
File metadata
- Download URL: ai_launcher_cli-0.1.6.tar.gz
- Upload date:
- Size: 9.8 kB
- Tags: Source
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/6.2.0 CPython/3.13.13
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
f6f48f58bfd9068cdc0f751877dda213192808fdb55699755e448df92dca9e90
|
|
| MD5 |
22627d94464f53e62508591c150904e0
|
|
| BLAKE2b-256 |
08f372ac80d70b63050791350aa2df6978d7c70b1e3fc8fb9b7a5c133f9f45bd
|
File details
Details for the file ai_launcher_cli-0.1.6-py3-none-any.whl.
File metadata
- Download URL: ai_launcher_cli-0.1.6-py3-none-any.whl
- Upload date:
- Size: 9.1 kB
- Tags: Python 3
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/6.2.0 CPython/3.13.13
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
c84bd8878812b49b47174099f8d33a786699126f8511dc331c4c0919b88c69ed
|
|
| MD5 |
8a462752425c8ae610e21fac62a79be9
|
|
| BLAKE2b-256 |
43845e5978d1be163a0fe0cb352b910e78bf52948a101fb1ada6bb8190decadd
|