Skip to main content

A unified Python connector for multiple LLM providers with YAML-driven configuration.

Project description

LLM Connector

A clean, modular Python connector providing a unified interface for various Large Language Model (LLM) providers, including OpenRouter, OpenAI, Anthropic, Google (Gemini), Groq, and local instances like llama.cpp and Ollama.

Features

  • Multiple Providers: Connect natively to OpenRouter, Google, OpenAI, Anthropic, Groq, local llama.cpp, and Ollama.
  • YAML Configuration: Manage all URLs, timeouts, retry logic, logging structures, and default models strictly through transparent YAML files (llm.yaml, security.yaml, logs.yaml).
  • Secure Overrides: Inject local offline IPs and enterprise pricing seamlessly via a .gitignore restricted override.yaml and keep API keys strictly in .env.
  • Resilient Connections: Features internal connection pooling, exponential backoff logic (for Rate Limits & 50x errors), and dynamically tracks OpenRouter internet pricing.
  • Dynamic Client Loading: Adapters load their heavy SDKs exclusively on-demand.

Installation

Option A: Install from PyPI (recommended)

pip install llm-connector

Then scaffold your project workspace:

llm-connector init

This creates an llm-connector/ directory with conf/, logs/, .env.template, and override.yaml.template — everything you need to configure the connector.

CLI Reference

The llm-connector command-line tool helps you set up and manage your workspace.

llm-connector --help

llm-connector init

Scaffolds a new llm-connector/ workspace in the current directory with all necessary configuration files.

llm-connector init [--force]
Flag Description
--force Overwrite an existing llm-connector/ directory

Generated structure:

llm-connector/
├── .env.template              # Copy to .env and add your API keys
├── conf/
│   ├── llm.yaml               # Default provider, model, temperature, max_tokens
│   ├── security.yaml          # Connection pooling, retry limits, backoff
│   ├── logs.yaml              # Log rotation and formatting
│   └── override.yaml.template # Copy to override.yaml for local endpoints
└── logs/                      # Runtime logs (auto-created)

After scaffolding:

cd llm-connector
cp .env.template .env          # Add your API keys
cp conf/override.yaml.template conf/override.yaml  # Add local endpoints (optional)

Option B: Clone as a submodule / standalone repo

git clone https://github.com/isbogdanov/llm_connector.git
cd llm_connector
pip install -e .

Configuration

  1. API Keys (.env): Copy the environment template and insert your secret keys.

    cp .env.template .env
    

    API keys should exclusively live in .env and never be committed.

  2. Local Environment Overrides (conf/override.yaml): If you are running local inference servers (like Ollama or LLaMA.cpp), or if you want to hardcode specific contract prices, create an override file:

    cp conf/override.yaml.template conf/override.yaml
    

    This file is safely untracked by Git, meaning your local IP addresses won't leak into the remote repository.

  3. Base System Configuration (conf/llm.yaml, conf/logs.yaml, conf/security.yaml): These files define the core tracked architecture. You can tune defaults like default_provider, log rotation limits, and API retry backoff-factors directly inside them.

Note: When installed via pip, all paths resolve relative to your scaffolded llm-connector/ directory. You can override this with the LLM_CONNECTOR_HOME environment variable.

Usage

Import the chat_completion function from the connector.

from llm_connector import chat_completion

# Example messages
messages = [
    {"role": "system", "content": "You are a highly capable AI assistant."},
    {"role": "user", "content": "Explain the importance of context windows in LLMs."},
]

# 1. Using the default provider & model (defined in llm.yaml)
response, p_t, c_t, t_t, latency = chat_completion(messages)
print(f"Default Response: {response}")

# 2. Specifying a specific provider and native model
response_openrouter, _, _, _, _ = chat_completion(
    messages,
    provider=("openrouter", "anthropic/claude-3.5-sonnet"),
    temperature=0.5,
)

# 3. Hitting a local offline model
response_local, _, _, _, _ = chat_completion(
    messages,
    provider=("ollama", "llama3.1:8b"),
)

Creating Custom Adapters

This package is designed to be infinitely extensible. To add a brand-new API provider, follow these 3 steps:

  1. Create the Adapter Class: Add a new file in llm_connector/adapters/ (e.g., my_adapter.py). It must inherit from AdapterBase and strictly implement the core chat_completion signature:

    from .adapter import AdapterBase
    
    class MyCustomAdapter(AdapterBase):
        def chat_completion(self, messages, model, temperature, max_tokens, top_p, **kwargs):
            # 1. Initialize your specific SDK client here
            # 2. Translate the generic 'messages' array into your provider's requested format
            # 3. Await the response
            # 4. Extract token usage and latency
            
            return response_text, prompt_tokens, completion_tokens, total_tokens, latency
    
  2. Export the Adapter: Expose your new adapter class inside llm_connector/adapters/__init__.py:

    from .my_adapter import MyCustomAdapter
    __all__ = [..., "MyCustomAdapter"]
    
  3. Register it in the Router: Add your provider string internally into the get_adapter network factory inside llm_connector/connector.py:

    elif provider_name == "my-custom-api":
        _adapters[provider_name] = MyCustomAdapter()
    

Testing

The testing suite natively utilizes pytest to rigidly guarantee the dynamic YAML hierarchy maps flawlessly to the internal routing logic without triggering infinite fallback loops.

To execute the entire engine diagnostic comprehensively, run:

pytest tests/

To include local model tests (requires a running llama.cpp or Ollama server):

pytest tests/ --run-local

The local suite actively asserts:

  1. Integration Verification: Dynamically attempts to route offline dummy prompts safely through OpenRouter, OpenAI, Anthropic, Google, Groq, and your local/Ollama networking blocks.
  2. Security Validation: Explicitly deconstructs the requests.Session() engine upon boot and validates your exact security.yaml limits (Connection Pools, HTTP Max Retries, and network backoff_factors) are clamped physically to the memory pipeline.
  3. Adapter Isolation: Prevents SDK crosstalk by verifying requests generated for "groq" cannot accidentally bleed over or trigger "local" internal logic modules.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

llm_connector-1.1.7.tar.gz (28.3 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

llm_connector-1.1.7-py3-none-any.whl (37.8 kB view details)

Uploaded Python 3

File details

Details for the file llm_connector-1.1.7.tar.gz.

File metadata

  • Download URL: llm_connector-1.1.7.tar.gz
  • Upload date:
  • Size: 28.3 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/6.1.0 CPython/3.13.7

File hashes

Hashes for llm_connector-1.1.7.tar.gz
Algorithm Hash digest
SHA256 9cdda1d0510a390849e21ae2fefef7ae184ea0058d904c3c4ee21ff69113b0c0
MD5 fa14b9f840a72dc7040c4447ba2a2b9a
BLAKE2b-256 ec33319941a5ce232668991ce3675e8532b3bf6585b72b986b0ca1b19ba942f9

See more details on using hashes here.

Provenance

The following attestation bundles were made for llm_connector-1.1.7.tar.gz:

Publisher: publish.yml on isbogdanov/llm_connector

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file llm_connector-1.1.7-py3-none-any.whl.

File metadata

  • Download URL: llm_connector-1.1.7-py3-none-any.whl
  • Upload date:
  • Size: 37.8 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/6.1.0 CPython/3.13.7

File hashes

Hashes for llm_connector-1.1.7-py3-none-any.whl
Algorithm Hash digest
SHA256 96e4ef0c0539c69177694d32a9d83742938cc770c6e0109dd77057bb9b531431
MD5 abe1f63ca665cb72c32226157807faa2
BLAKE2b-256 70860aaea77f69c92d021421eefc629abb4969d628957d65a7d9b647093afc1f

See more details on using hashes here.

Provenance

The following attestation bundles were made for llm_connector-1.1.7-py3-none-any.whl:

Publisher: publish.yml on isbogdanov/llm_connector

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page