lmux-azure-foundry
Azure AI Foundry provider for lmux. Talks to the Azure OpenAI REST API directly with httpx.
Supports chat completions, streaming, embeddings, and the Responses API.
Part of the lmux ecosystem: standardized interface, cost tracking on every response, and registry-based routing across providers.
Optional Extras
lmux-azure-foundry[identity]: Azure AD token authentication viaazure-identity
Auth
Three authentication methods:
API Key (default)
Set AZURE_FOUNDRY_API_KEY in your environment:
from lmux_azure_foundry import AzureFoundryProvider
provider = AzureFoundryProvider(endpoint="https://your-resource.openai.azure.com")
Azure AD Token
from lmux_azure_foundry import AzureFoundryProvider, AzureAdToken
provider = AzureFoundryProvider(
endpoint="https://your-resource.openai.azure.com",
auth=my_auth_returning_azure_ad_token,
)
Token Provider
from lmux_azure_foundry import AzureFoundryTokenAuthProvider
provider = AzureFoundryProvider(
endpoint="https://your-resource.openai.azure.com",
auth=AzureFoundryTokenAuthProvider(), # uses azure-identity DefaultAzureCredential
)
Usage
Chat
from lmux import UserMessage
response = provider.chat("gpt-4o", [UserMessage(content="Hello")])
print(response.content)
print(response.cost)
Streaming
for chunk in provider.chat_stream("gpt-4o", [UserMessage(content="Hello")]):
if chunk.delta:
print(chunk.delta, end="")
Embeddings
response = provider.embed("text-embedding-3-small", "Hello")
print(response.embeddings)
Responses API
Required for models that are only served through the Responses API:
response = provider.create_response("gpt-5-pro", "Hello")
print(response.output_text)
Async
All methods have async variants: achat, achat_stream, aembed, acreate_response.
Registry
Use with the lmux registry to route across multiple providers:
from lmux import Registry
registry = Registry()
registry.register("azure", provider)
response = registry.chat("azure/gpt-4o", messages)
Provider Params
from lmux_azure_foundry import AzureFoundryParams
response = provider.chat(
"gpt-4o",
messages,
provider_params=AzureFoundryParams(deployment_type="data_zone"),
)
| Parameter | Type | Description |
|---|---|---|
reasoning_effort |
"low" | "medium" | "high" |
Reasoning effort for o-series models |
seed |
int |
Deterministic sampling seed |
user |
str |
End-user identifier |
prompt_cache_key |
str |
Cache key for Azure's automatic prompt caching (chat + responses) |
prompt_cache_retention |
"in_memory" | "24h" |
Prompt cache retention policy (chat + responses) |
deployment_type |
"global" | "data_zone" | "regional" |
Affects cost calculation only, not sent to API |
Constructor Options
AzureFoundryProvider(
endpoint=..., # required, Azure resource endpoint
auth=..., # AuthProvider, default: AzureFoundryKeyAuthProvider()
api_version=..., # API version (default: "2025-04-01-preview")
timeout=..., # Request timeout in seconds
max_retries=..., # Max retry attempts
default_headers=..., # Optional headers included with every request
transport=..., # Optional httpx.BaseTransport for the sync client (proxies, testing)
async_transport=..., # Optional httpx.AsyncBaseTransport for the async client
)
default_headers is useful for gateway authentication, tracing, and routing. Foundry-managed authentication and
content-type headers take precedence over caller values, case-insensitively.
Metadata
Release files for lmux-azure-foundry 0.10.2
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| lmux_azure_foundry-0.10.2.tar.gz | 18.3 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| lmux_azure_foundry-0.10.2-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 40.1 kB
Release files / lmux_azure_foundry-0.10.2.tar.gz
| Download URL | lmux_azure_foundry-0.10.2.tar.gz |
|---|---|
| Size | 18.3 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
cacd57ede77f01a7dd9cc2701cf3c8c5ef0ae2f875fb2e00b6cc6f7858dd821c
|
|
BLAKE2b-256 checksum How to use checksums |
d4a79106c79677de5dda1576cff7a82417b868fb119d3dfcb37042d1fb268cfb
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/6.1.0 CPython/3.13.13
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Sep 1, 2026.
Transparency logRelease files / lmux_azure_foundry-0.10.2-py3-none-any.whl
| Download URL | lmux_azure_foundry-0.10.2-py3-none-any.whl |
|---|---|
| Size | 21.9 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
aeacecbe4e6a3aa2c0d6724f89141c657ea5d00664280542254c176a474242ae
|
|
BLAKE2b-256 checksum How to use checksums |
9c3f3f3ca2d5ec085508b76365fedad4933ece6d572553f41f7fdabad44cbe58
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/6.1.0 CPython/3.13.13
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Sep 1, 2026.
Transparency log