Structured analysis for voice-agent call recordings — transcript, quality, sentiment, latency, cost, and events from raw audio.

These details have not been verified by PyPI

Project links

Project description

AudioTrace

CI/CD for voice AI

What is AudioTrace?

Voice Agents AI plumbing tool every team rebuilds from scratch — until now. Drop in a call recording. Get back everything: transcript, quality scores, sentiment shifts, latency breakdown, cost attribution, compliance flags. Normalized. Structured. Queryable. Works with any provider, any stack. One integration. Zero plumbing. Ship faster.

import audiotrace

report = audiotrace.analyze(
    audio    = "call_recording.wav",
    metadata = {"agent_version": "v2.1", "provider": "vapi"}
)

print(report.quality.overall_score)        # 0.87
print(report.sentiment.caller_frustration) # False
print(report.latency.llm_first_token_ms)   # 420
print(report.events.drop_off)              # False
print(report.cost.total_usd)               # 0.063

📺 Watch the demo:

Why AudioTrace?

Every team building voice agents faces the same problem: raw audio is a black box. You can listen to recordings manually, or you can build your own signal extraction pipeline from scratch — but no open-source framework normalizes the full call into a structured, queryable object.

AudioTrace exists to be that shared layer. It handles the hard parts so you can focus on what you're building:

Transcription with speaker diarization
Silence gaps, interruptions, speaking pace, and pitch analysis
Per-turn sentiment tracking and frustration detection
Per-stage latency breakdown (STT → LLM → TTS → telephony)
Unified cost calculation across any provider mix
Compliance flag detection (PII leakage, consent gaps)

Installation

pip install audiotrace

# With specific provider adapter
pip install audiotrace[vapi]
pip install audiotrace[retell]
pip install audiotrace[twilio]

# Full install
pip install audiotrace[all]

Docker

docker build -f docker/Dockerfile -t audiotrace .
docker run -it audiotrace

Requirements: Python 3.9+, FFmpeg installed on system

Quick start

Analyze a single call

import audiotrace

report = audiotrace.analyze(
    audio    = "call.wav",
    metadata = {
        "call_id":       "abc123",
        "agent_version": "v2.1",
        "provider":      "vapi",
        "campaign":      "healthcare_intake"
    }
)

# Media
print(report.media.duration_ms)          # int
print(report.media.codec)                # str

# Transcript
print(report.transcript.full_text)
for turn in report.transcript.turns:
    print(f"{turn.speaker}: {turn.text}")

# Quality
print(report.quality.overall_score)       # float 0.0–1.0
print(report.quality.interruptions)       # int
print(report.quality.silence_gaps)        # List[Gap]
print(report.quality.speaking_pace_wpm)   # float

# Sentiment
print(report.sentiment.overall)           # float -1.0 to 1.0
print(report.sentiment.shift_points)      # List[int] — turn indices
print(report.sentiment.caller_frustration)# bool

# Latency
print(report.latency.stt_ms)             # int
print(report.latency.llm_first_token_ms) # int
print(report.latency.tts_ms)             # int
print(report.latency.total_ms)           # int

# Cost
print(report.cost.stt_usd)               # float
print(report.cost.llm_usd)               # float
print(report.cost.total_usd)             # float

# Events
print(report.events.outcome)             # "completed" | "dropped" | "failed"
print(report.events.drop_off_turn)       # int | None
print(report.events.compliance_flags)    # List[str]

Use provider adapters

from audiotrace.adapters import VapiAdapter

adapter = VapiAdapter(api_key="...")
call    = adapter.fetch_call(call_id="abc123")
report  = audiotrace.analyze(call.audio, call.metadata)

Output — CallReport

CallReport
├── media
│   ├── duration_ms: int
│   ├── sample_rate_hz: int
│   ├── channels: int
│   ├── codec: str
│   ├── file_size_bytes: int
│   ├── file_format: str
│   └── bitrate_kbps: float
├── transcript
│   ├── full_text: str
│   ├── turns: List[Turn]   # speaker · text · start_ms · end_ms · confidence · words[]
│   └── language: str
├── quality
│   ├── overall_score: float
│   ├── interruptions: int
│   ├── silence_gaps: List[Gap]
│   ├── speaking_pace_wpm: float
│   ├── pitch_variance: float
│   └── turn_length_avg_ms: float
├── sentiment
│   ├── by_turn: List[float]
│   ├── overall: float
│   ├── shift_points: List[int]
│   └── caller_frustration: bool
├── latency
│   ├── stt_ms: int
│   ├── llm_first_token_ms: int
│   ├── llm_full_response_ms: int
│   ├── tts_ms: int
│   ├── total_ms: int
│   └── waterfall: List[LatencySpan]
├── cost
│   ├── stt_usd: float
│   ├── llm_usd: float
│   ├── tts_usd: float
│   ├── telephony_usd: float
│   └── total_usd: float
└── events
    ├── outcome: str
    ├── drop_off: bool
    ├── drop_off_turn: int | None
    ├── intent_detected: str
    ├── failure_type: str | None
    └── compliance_flags: List[str]

Provider support

Provider	Adapter	Status
Vapi	`audiotrace[vapi]`	Stable
Retell	`audiotrace[retell]`	Stable
Twilio	`audiotrace[twilio]`	Stable
ElevenLabs	`audiotrace[elevenlabs]`	Beta
Deepgram	`audiotrace[deepgram]`	Stable
Custom webhook	`CustomAdapter`	Stable

How it works

AudioTrace builds on top of best-in-class audio libraries so you don't have to:

Raw audio file
      │
      ▼
  FFmpeg              — format normalization, turn splitting
      │
      ├── Whisper     — transcription
      ├── pyannote    — speaker diarization
      ├── Librosa     — silence gaps, pace, pitch, energy
      └── Transformers — sentiment, intent detection
      │
      ▼
  CallReport (Pydantic)

Part of the Lang ecosystem TBD

AudioTrace is the open-source foundation that powers two commercial products:

Product	What it does	Built on
LangTrace	Live call observability & analytics dashboards	AudioTrace
LangGate	Pre-deploy simulation & CI/CD quality gate	AudioTrace

AudioTrace is free and MIT-licensed. The commercial products are optional hosted layers on top.

Running locally

For quick testing or interactive analysis, you can use the provided runner script. It automatically handles virtual environment setup and dependency validation.

# Analyze default golden data fixture
./scripts/run.sh

# Concise per-section summary tables instead of the raw JSON
./scripts/run.sh --summary

# Playback, inferring speakers by pitch (no pyannote token needed)
./scripts/run.sh --playback --skip-pyannote

# Analyze a specific file
./scripts/run.sh path/to/your/audio.wav

Development & Validation

Before submitting changes, ensure everything passes the local validation suite (formatting, linting, type-checking, and tests):

./scripts/test_local.sh test

Contributing

Contributions are welcome — especially new provider adapters, persona definitions for simulation, and compliance rule sets.

git clone https://github.com/audiotrace/audiotrace
cd audiotrace
./scripts/test_local.sh test  # Run all checks (formatting, lint, types, tests)

See CONTRIBUTING.md for guidelines.

License

MIT — see LICENSE

Project details

These details have not been verified by PyPI

Project links

Release history Release notifications | RSS feed

This version

1.1.1

Jun 26, 2026

1.1.0

Jun 25, 2026

1.0.1

Jun 21, 2026

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

audiotrace-1.1.1.tar.gz (1.8 MB view details)

Uploaded Jun 26, 2026 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

audiotrace-1.1.1-py3-none-any.whl (25.4 kB view details)

Uploaded Jun 26, 2026 Python 3

File details

Details for the file audiotrace-1.1.1.tar.gz.

File metadata

Download URL: audiotrace-1.1.1.tar.gz
Upload date: Jun 26, 2026
Size: 1.8 MB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.2.0 CPython/3.14.3

File hashes

Hashes for audiotrace-1.1.1.tar.gz
Algorithm	Hash digest
SHA256	`fa94e06a2ca025b4615180f2ed3a13715e1b47adc9d31061e373c3954fbd25b3`
MD5	`852be97cd4d074ef8996212429720a01`
BLAKE2b-256	`27eab754d7deb95846347d2b1e8f664020b9caea6fdbdb660459c88df6806464`

See more details on using hashes here.

File details

Details for the file audiotrace-1.1.1-py3-none-any.whl.

File metadata

Download URL: audiotrace-1.1.1-py3-none-any.whl
Upload date: Jun 26, 2026
Size: 25.4 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.2.0 CPython/3.14.3

File hashes

Hashes for audiotrace-1.1.1-py3-none-any.whl
Algorithm	Hash digest
SHA256	`4feb44faf03087d296f947e31e42db92a2b505195faeb09f86947d4e214d93b1`
MD5	`89e45ecf38c8166db348cf9ae670304d`
BLAKE2b-256	`8436ce55fc7aa17a7f654d917d1fe30b5c637c967f494410c3919c1a8bc55fbe`

See more details on using hashes here.

audiotrace 1.1.1

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

AudioTrace

CI/CD for voice AI

What is AudioTrace?

Why AudioTrace?

Installation

Docker

Quick start

Analyze a single call

Use provider adapters

Output — CallReport

Provider support

How it works

Part of the Lang ecosystem TBD

Running locally

Development & Validation

Contributing

License

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes