Skip to main content

A unified Python connector for multiple LLM providers with YAML-driven configuration.

Project description

LLM Connector

A clean, modular Python connector providing a unified interface for various Large Language Model (LLM) providers, including OpenRouter, OpenAI, Anthropic, Google (Gemini), Groq, and local instances like llama.cpp and Ollama.

Features

  • Multiple Providers: Connect natively to OpenRouter, Google, OpenAI, Anthropic, Groq, local llama.cpp, and Ollama.
  • YAML Configuration: Manage all URLs, timeouts, retry logic, logging structures, and default models strictly through transparent YAML files (llm.yaml, security.yaml, logs.yaml).
  • Secure Overrides: Inject local offline IPs and enterprise pricing seamlessly via a .gitignore restricted override.yaml and keep API keys strictly in .env.
  • Resilient Connections: Features internal connection pooling, exponential backoff logic (for Rate Limits & 50x errors), and dynamically tracks OpenRouter internet pricing.
  • Dynamic Client Loading: Adapters load their heavy SDKs exclusively on-demand.

Installation

Option A: Install from PyPI (recommended)

pip install llm-connector

Then scaffold your project workspace:

llm-connector init

This creates an llm-connector/ directory with conf/, logs/, .env.template, and override.yaml.template — everything you need to configure the connector.

Option B: Clone as a submodule / standalone repo

git clone https://github.com/isbogdanov/llm_connector.git
cd llm_connector
pip install -e .

Configuration

  1. API Keys (.env): Copy the environment template and insert your secret keys.

    cp .env.template .env
    

    API keys should exclusively live in .env and never be committed.

  2. Local Environment Overrides (conf/override.yaml): If you are running local inference servers (like Ollama or LLaMA.cpp), or if you want to hardcode specific contract prices, create an override file:

    cp conf/override.yaml.template conf/override.yaml
    

    This file is safely untracked by Git, meaning your local IP addresses won't leak into the remote repository.

  3. Base System Configuration (conf/llm.yaml, conf/logs.yaml, conf/security.yaml): These files define the core tracked architecture. You can tune defaults like default_provider, log rotation limits, and API retry backoff-factors directly inside them.

Note: When installed via pip, all paths resolve relative to your scaffolded llm-connector/ directory. You can override this with the LLM_CONNECTOR_HOME environment variable.

Usage

Import the chat_completion function from the connector.

from llm_connector import chat_completion

# Example messages
messages = [
    {"role": "system", "content": "You are a highly capable AI assistant."},
    {"role": "user", "content": "Explain the importance of context windows in LLMs."},
]

# 1. Using the default provider & model (defined in llm.yaml)
response, p_t, c_t, t_t, latency = chat_completion(messages)
print(f"Default Response: {response}")

# 2. Specifying a specific provider and native model
response_openrouter, _, _, _, _ = chat_completion(
    messages,
    provider=("openrouter", "anthropic/claude-3.5-sonnet"),
    temperature=0.5,
)

# 3. Hitting a local offline model
response_local, _, _, _, _ = chat_completion(
    messages,
    provider=("ollama", "llama3.1:8b"),
)

Creating Custom Adapters

This package is designed to be infinitely extensible. To add a brand-new API provider, follow these 3 steps:

  1. Create the Adapter Class: Add a new file in connector/adapters/ (e.g., my_adapter.py). It must inherit from AdapterBase and strictly implement the core chat_completion signature:

    from .adapter import AdapterBase
    
    class MyCustomAdapter(AdapterBase):
        def chat_completion(self, messages, model, temperature, max_tokens, top_p, **kwargs):
            # 1. Initialize your specific SDK client here
            # 2. Translate the generic 'messages' array into your provider's requested format
            # 3. Await the response
            # 4. Extract token usage and latency
            
            return response_text, prompt_tokens, completion_tokens, total_tokens, latency
    
  2. Export the Adapter: Expose your new adapter class inside connector/adapters/__init__.py:

    from .my_adapter import MyCustomAdapter
    __all__ = [..., "MyCustomAdapter"]
    
  3. Register it in the Router: Add your provider string internally into the get_adapter network factory inside connector/connector.py:

    elif provider_name == "my-custom-api":
        _adapters[provider_name] = MyCustomAdapter()
    

Testing

The testing suite natively utilizes pytest to rigidly guarantee the dynamic YAML hierarchy maps flawlessly to the internal routing logic without triggering infinite fallback loops.

To execute the entire engine diagnostic comprehensively, run:

pytest tests/

To include local model tests (requires a running llama.cpp or Ollama server):

pytest tests/ --run-local

The local suite actively asserts:

  1. Integration Verification: Dynamically attempts to route offline dummy prompts safely through OpenRouter, OpenAI, Anthropic, Google, Groq, and your local/Ollama networking blocks.
  2. Security Validation: Explicitly deconstructs the requests.Session() engine upon boot and validates your exact security.yaml limits (Connection Pools, HTTP Max Retries, and network backoff_factors) are clamped physically to the memory pipeline.
  3. Adapter Isolation: Prevents SDK crosstalk by verifying requests generated for "groq" cannot accidentally bleed over or trigger "local" internal logic modules.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

llm_connector-1.1.0.tar.gz (27.6 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

llm_connector-1.1.0-py3-none-any.whl (37.2 kB view details)

Uploaded Python 3

File details

Details for the file llm_connector-1.1.0.tar.gz.

File metadata

  • Download URL: llm_connector-1.1.0.tar.gz
  • Upload date:
  • Size: 27.6 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.11.7

File hashes

Hashes for llm_connector-1.1.0.tar.gz
Algorithm Hash digest
SHA256 f107618ef1c418c646fee8ca1b48bf98928b8452829e76572034823da4be54e9
MD5 a507d25aa7ec0a929c3cbab4166eb2b2
BLAKE2b-256 3c850b7104026062bf89d84f49d1d2ab33570fce5c6ffc28f0161ea3c4d1a2b7

See more details on using hashes here.

File details

Details for the file llm_connector-1.1.0-py3-none-any.whl.

File metadata

  • Download URL: llm_connector-1.1.0-py3-none-any.whl
  • Upload date:
  • Size: 37.2 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.11.7

File hashes

Hashes for llm_connector-1.1.0-py3-none-any.whl
Algorithm Hash digest
SHA256 9ccd8793c3177cf60b5723b4d67857b351c6044b3bf5ac0a344922d78b66df52
MD5 ea504ae33928d259336b6dd6cff475f1
BLAKE2b-256 3659f304c9b665c979552c2492d380c45a1519840e95bab7667adfe3b4ea31d7

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page