txtai

All-in-one open-source AI framework for semantic search, LLM orchestration and language model workflows

These details have not been verified by PyPI

Project links

Project description

All-in-one AI framework

txtai is an all-in-one AI framework for semantic search, LLM orchestration and language model workflows.

architecture

The key component of txtai is an embeddings database, which is a union of vector indexes (sparse and dense), graph networks and relational databases.

This foundation enables vector search and/or serves as a powerful knowledge source for large language model (LLM) applications.

Build autonomous agents, retrieval augmented generation (RAG) processes, multi-model workflows and more.

Summary of txtai features:

🔎 Vector search with SQL, object storage, topic modeling, graph analysis and multimodal indexing
📄 Create embeddings for text, documents, audio, images and video
💡 Pipelines powered by language models that run LLM prompts, question-answering, labeling, transcription, translation, summarization and more
↪️️ Workflows to join pipelines together and aggregate business logic. txtai processes can be simple microservices or multi-model workflows.
🤖 Agents that intelligently connect embeddings, pipelines, workflows and other agents together to autonomously solve complex problems
⚙️ Web and Model Context Protocol (MCP) APIs. Bindings available for JavaScript, Java, Rust and Go.
🔋 Batteries included with defaults to get up and running fast
☁️ Run local or scale out with container orchestration

txtai is built with Python 3.10+, Hugging Face Transformers, Sentence Transformers and FastAPI. txtai is open-source under an Apache 2.0 license.

[!NOTE]

NeuML is the company behind txtai and we provide AI consulting services around our stack. Schedule a meeting or send a message to learn more.

We're also building an easy and secure way to run hosted txtai applications with txtai.cloud.

Why txtai?

why

New vector databases, LLM frameworks and everything in between are sprouting up daily. Why build with txtai?

Up and running in minutes with pip or Docker

# Get started in a couple lines
import txtai

embeddings = txtai.Embeddings()
embeddings.index(["Correct", "Not what we hoped"])
embeddings.search("positive", 1)
#[(0, 0.29862046241760254)]

Built-in API makes it easy to develop applications using your programming language of choice

# app.yml
embeddings:
    path: sentence-transformers/all-MiniLM-L6-v2

CONFIG=app.yml uvicorn "txtai.api:app"
curl -X GET "http://localhost:8000/search?query=positive"

Run local - no need to ship data off to disparate remote services
Work with micromodels all the way up to large language models (LLMs)
Low footprint - install additional dependencies and scale up when needed
Learn by example - notebooks cover all available functionality

Use Cases

The following sections introduce common txtai use cases. A comprehensive set of over 70 example notebooks and applications are also available.

Semantic Search

Build semantic/similarity/vector/neural search applications.

demo

Traditional search systems use keywords to find data. Semantic search has an understanding of natural language and identifies results that have the same meaning, not necessarily the same keywords.

Get started with the following examples.

Notebook	Description
Introducing txtai ▶️	Overview of the functionality provided by txtai
Similarity search with images	Embed images and text into the same space for search
Build a QA database	Question matching with semantic search
Semantic Graphs	Explore topics, data connectivity and run network analysis

LLM Orchestration

Autonomous agents, retrieval augmented generation (RAG), chat with your data, pipelines and workflows that interface with large language models (LLMs).

llm

See below to learn more.

Notebook	Description
Prompt templates and task chains	Build model prompts and connect tasks together with workflows
Integrate LLM frameworks	Integrate llama.cpp, LiteLLM and custom generation frameworks
Build knowledge graphs with LLMs	Build knowledge graphs with LLM-driven entity extraction
Parsing the stars with txtai	Explore an astronomical knowledge graph of known stars, planets, galaxies

Agents

Agents connect embeddings, pipelines, workflows and other agents together to autonomously solve complex problems.

agent

txtai agents are built on top of the smolagents framework. This supports all LLMs txtai supports (Hugging Face, llama.cpp, OpenAI / Claude / AWS Bedrock via LiteLLM). Agent prompting with agents.md and skill.md are also supported.

Check out this Agent Quickstart Example. Additional examples are listed below.

Notebook	Description
Granting autonomy to agents	Agents that iteratively solve problems as they see fit
TxtAI got skills	Integrate skill.md files with your agent
Agent Tools ▶️	Learn about the txtai agent toolkit
Analyzing LinkedIn Company Posts with Graphs and Agents	Exploring how to improve social media engagement with AI

Retrieval augmented generation

Retrieval augmented generation (RAG) reduces the risk of LLM hallucinations by constraining the output with a knowledge base as context. RAG is commonly used to "chat with your data".

rag

Check out this RAG Quickstart Example. Additional examples are listed below.

Notebook	Description
Build RAG pipelines with txtai ▶️	Guide on retrieval augmented generation including how to create citations
RAG is more than Vector Search	Context retrieval via Web, SQL and other sources
GraphRAG with Wikipedia and GPT OSS	Deep graph search powered RAG
Speech to Speech RAG ▶️	Full cycle speech to speech workflow with RAG

Language Model Workflows

Language model workflows, also known as semantic workflows, connect language models together to build intelligent applications.

flows

While LLMs are powerful, there are plenty of smaller, more specialized models that work better and faster for specific tasks. This includes models for extractive question-answering, automatic summarization, text-to-speech, transcription and translation.

Check out this Workflow Quickstart Example. Additional examples are listed below.

Notebook	Description
Run pipeline workflows ▶️	Simple yet powerful constructs to efficiently process data
Building abstractive text summaries	Run abstractive text summarization
Transcribe audio to text	Convert audio files to text
Translate text between languages	Streamline machine translation and language detection

Installation

install

The easiest way to install is via pip and PyPI

pip install txtai

Python 3.10+ is supported. Using a Python virtual environment is recommended.

See the detailed install instructions for more information covering optional dependencies, environment specific prerequisites, installing from source, conda support, lightweight minimal installation and how to run with containers.

Model guide

models

See the table below for the current recommended models. These models all allow commercial use and offer a blend of speed and performance.

Component	Model(s)
Embeddings	all-MiniLM-L6-v2
Image Captions	BLIP
Labels - Zero Shot	DeBERTa v3 Zeroshot
Labels - Fixed	Fine-tune with training pipeline
Large Language Model (LLM)	Gemma 4 31B
Summarization	DistilBART
Text-to-Speech	ESPnet JETS
Transcription	Whisper
Translation	OPUS Model Series

Models can be loaded as either a path from the Hugging Face Hub or a local directory. Model paths are optional, defaults are loaded when not specified. For tasks with no recommended model, txtai uses the default models as shown in the Hugging Face Tasks guide.

See the following links to learn more.

Powered by txtai

The following applications are powered by txtai.

apps

Application	Description
rag	Retrieval Augmented Generation (RAG) application
ncoder	Open-Source AI coding agent
paperai	AI for medical and scientific papers
annotateai	Automatically annotate papers with LLMs

In addition to this list, there are also many other open-source projects, published research and closed proprietary/commercial projects that have built on txtai in production.

Documentation

Full documentation on txtai including configuration settings for embeddings, pipelines, workflows, API and a FAQ with common questions/issues is available.

Contributing

For those who would like to contribute to txtai, please see this guide.

Project details

These details have not been verified by PyPI

Project links

Release history Release notifications | RSS feed

This version

9.11.0

Jul 1, 2026

9.10.0

Jun 4, 2026

9.9.0

May 12, 2026

9.8.0

Apr 29, 2026

9.7.0

Mar 20, 2026

9.6.0

Feb 25, 2026

9.5.0

Feb 12, 2026

9.4.1

Jan 23, 2026

9.4.0

Jan 21, 2026

9.3.0

Dec 23, 2025

9.2.0

Nov 21, 2025

9.1.0

Nov 4, 2025

9.0.1

Sep 15, 2025

9.0.0

Aug 28, 2025

8.6.0

Jun 10, 2025

8.5.0

Apr 14, 2025

8.4.0

Mar 11, 2025

8.3.1

Feb 12, 2025

8.3.0

Feb 11, 2025

8.2.0

Jan 9, 2025

8.1.0

Dec 10, 2024

8.0.0

Nov 18, 2024

7.5.1

Oct 25, 2024

7.5.0

Oct 14, 2024

7.4.0

Sep 5, 2024

7.3.0

Jul 15, 2024

7.2.0

May 31, 2024

7.1.0

Apr 19, 2024

7.0.0

Feb 21, 2024

6.3.0

Jan 2, 2024

6.2.0

Nov 8, 2023

6.1.0

Sep 26, 2023

6.0.0

Aug 10, 2023

5.5.1

Apr 27, 2023

5.5.0

Apr 20, 2023

5.4.0

Mar 6, 2023

5.3.0

Feb 7, 2023

5.2.0

Dec 20, 2022

5.1.0

Oct 18, 2022

5.0.0

Sep 27, 2022

4.6.0

Aug 15, 2022

4.5.0

May 17, 2022

4.4.0

Apr 20, 2022

4.3.1

Mar 11, 2022

4.3.0

Mar 10, 2022

4.2.1

Feb 28, 2022

4.2.0

Feb 24, 2022

4.1.0

Feb 3, 2022

4.0.0

Jan 11, 2022

3.7.0

Nov 23, 2021

3.6.0

Nov 8, 2021

3.5.0

Oct 18, 2021

3.4.0

Oct 7, 2021

3.3.0

Sep 10, 2021

3.2.0

Aug 17, 2021

3.1.0

May 22, 2021

3.0.0

May 4, 2021

2.0.0

Jan 13, 2021

1.5.0

Nov 21, 2020

1.4.0

Nov 3, 2020

1.3.0

Oct 11, 2020

1.2.1

Sep 11, 2020

1.2.0

Sep 10, 2020

1.1.0

Aug 18, 2020

1.0.0

Aug 11, 2020

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

txtai-9.11.0.tar.gz (238.8 kB view details)

Uploaded Jul 1, 2026 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

txtai-9.11.0-py3-none-any.whl (324.5 kB view details)

Uploaded Jul 1, 2026 Python 3

File details

Details for the file txtai-9.11.0.tar.gz.

File metadata

Download URL: txtai-9.11.0.tar.gz
Upload date: Jul 1, 2026
Size: 238.8 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.2.0 CPython/3.10.20

File hashes

Hashes for txtai-9.11.0.tar.gz
Algorithm	Hash digest
SHA256	`2f463aa98e1bb70c14be5767ca137fb9808b0c07e02778875ec62693d9a57c35`
MD5	`8711bd9de6d6bab630395c2102c936fc`
BLAKE2b-256	`4d56ca8185030a741fb7f884a70c5be25d8ba666543a8fd2cf6422c104a62af8`

See more details on using hashes here.

File details

Details for the file txtai-9.11.0-py3-none-any.whl.

File metadata

Download URL: txtai-9.11.0-py3-none-any.whl
Upload date: Jul 1, 2026
Size: 324.5 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.2.0 CPython/3.10.20

File hashes

Hashes for txtai-9.11.0-py3-none-any.whl
Algorithm	Hash digest
SHA256	`bc6193e167da0899873e652b97c1c2a9d0bdaffb3dd908fbb372a106bc8473b1`
MD5	`fc584906e561ec9d93c743295a7fe1b8`
BLAKE2b-256	`9e2c2396a30770514583393df1abe00824d13eeb37b27b31ab9e98cdede13715`

See more details on using hashes here.

txtai 9.11.0

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Why txtai?

Use Cases

Semantic Search

LLM Orchestration

Agents

Retrieval augmented generation

Language Model Workflows

Installation

Model guide

Powered by txtai

Further Reading

Documentation

Contributing

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes