These details have not been verified by PyPI

Project links

Homepage

Development Status
- 3 - Alpha
Intended Audience
- Developers
License
- OSI Approved :: MIT License
Programming Language

Project description

Pattern Based Question and Answer

Description

Pattern Based Question and Answer (PBQA) is a Python library that provides tools for querying LLMs and managing text embeddings. It combines guided generation with multi-shot prompting to improve response quality and ensure consistency. By enforcing valid responses, PBQA makes it easy to combine the flexibility of LLMs with the reliability and control of symbolic approaches.

Installation
Usage
Roadmap
Relevant Literature
Contributing
Support
License

Installation

PBQA requires Python 3.9 or higher, and can be installed via pip:

pip install PBQA

Additionally, PBQA requires a running instance of llama.cpp to interact with LLMs. For instructions on installation, see the llama.cpp repository.

Usage

llama.cpp

For instructions on hosting a model with llama.cpp, see the following page. Optionally, caching can be enabled to speed up generation.

Python

PBQA provides a simple API for querying LLMs.

from PBQA import DB, LLM
from time import strftime

# First, we set up a database at a specified path
db = DB(path="db")
# Then, we load a pattern file into the database
db.load_pattern("examples/weather.yaml")

# Next, we connect to the LLM server
llm = LLM(db=db, host="127.0.0.1")
# And connect to the model
llm.connect_model(
    model="llama",
    port=8080,
    stop=["<|eot_id|>", "<|start_header_id|>"],
    temperature=0,
)

# Finally, we query the LLM and receive a response based on the specified pattern
# Optionally, external data can be provided to the LLM which it can use in its response
weather_query = llm.ask(
        "Could I see the stars tonight?",
        "weather",
        "llama",
        external={"now": strftime("%Y-%m-%d %H:%M")},
    )

Using the weather.yaml pattern file and llama 3 running on 127.0.0.1:8080, the response should look something like this:

{
    "latitude": 51.51,
    "longitude": 0.13,
    "time": "2024-06-18 01:00",
}

For more information, see the examples directory.

Pattern Files

Pattern files are used to guide the LLM in generating responses. They are written in YAML and consist of three parts: the system prompt, component metadata, and examples.

# The system prompt is the main instruction given to the LLM telling it what to do
system_prompt: Your job is to translate the user's input into a weather query. Reply with the json for the weather query and nothing else.
now:  # Each component of the response needs to have it's own key, "component:" at minimum
  external: true  # Optionally, specify whether the component requires external data
latitude:
  grammar: |  # Or define a GBNF grammar
    root         ::= coordinate
    coordinate   ::= integer "." integer
    integer      ::= digit | digit digit
    digit        ::= "0" | "1" | "2" | "3" | "4" | "5" | "6" | "7" | "8" | "9"
longitude:
  grammar: ...
time:
  grammar: ...
examples:  # Lastly, examples can be provided for multi-shot prompting
- input: What will the weather be like tonight
  now: 2019-09-30 10:36
  latitude: 51.51
  longitude: 0.13
  time: 2019-09-30 20:00
- input: Could I see the stars tonight?
  ...

For more examples, look at the pattern files in the examples directory. Information on the GBNF grammar format can be found here.

Cache

Unless overridden, queries using the same pattern will use the same system prompt and base examples, allowing a large part of the response to be cached and speeding up generation. This can be disabled by setting use_cache=False in the ask() method.

PBQA allocates a slot/process for each pattern-model pair in the llama.cpp server. Set -np to the number of unique combinations of patterns and models you want to enable caching for. Slots are allocated in the order they are requested, and if the number of available slots is exceeded, the last slot is reused for any excess pattern-model pairs.

You can manually assign a cache slot to a specific pattern-model pair using the link method. Optionally, a specific cache slot can be provided, up to the number of available processes. The cache slot used for a query can also be overridden by passing the cache_slot parameter to the llm.ask() method.

from PBQA import DB, LLM


db = DB(path="db")
db.load_pattern("examples/weather.yaml")

llm = LLM(db=db, host="127.0.0.1")
llm.connect_model(
    model="llama",
    port=8080,
    stop=["<|eot_id|>", "<|start_header_id|>"],
    temperature=0,
)
llm.link(pattern="weather", model="llama")

Once a pattern-model pair is linked, the "model" parameter in the ask() method may also be omitted. The query will instead use the model assigned during the last appropriate link call.

Roadmap

Future features in no particular order with no particular timeline:

Reranking
Preset grammars for common data types
Parallel query execution
Combining multi-shot prompting with message history
Multimodal support
Further speed improvements (possibly batching)
Support for more LLM backends

Relevant Literature

Contributing

Contributions are welcome! If you have any suggestions or would like to contribute, please open an issue or a pull request.

Support

If you want to support the development of PBQA, consider buying me a coffee. Any support is greatly appreciated!

License and Acknowledgements

This project is licensed under the terms of the MIT License. For more details, see the LICENSE file.

Qdrant is a vector database that provides an API for managing and querying text embeddings. PBQA uses Qdrant to store and retrieve text embeddings.

llama.cpp is a C++ library that provides an easy-to-use interface for running LLMs on a wide variety of hardware. It includes support for Apple silicon, x86 architectures, and NVIDIA GPUs, as well as custom CUDA kernels for running LLMs on AMD GPUs via HIP. PBQA uses llama.cpp to interact with LLMs.

PBQA was originally developed by Bart Haagsma as part of different project. If you have any questions or suggestions, please feel free to contact me at dev.baagsma@gmail.com.

Project details

These details have not been verified by PyPI

Project links

Homepage

Development Status
- 3 - Alpha
Intended Audience
- Developers
License
- OSI Approved :: MIT License
Programming Language

Release history Release notifications | RSS feed

1.3.2

Apr 4, 2026

1.3.1

Feb 27, 2026

1.3.0

Feb 10, 2026

1.2.9

Dec 27, 2025

1.2.8

Dec 27, 2025

1.2.7

Oct 24, 2025

1.2.6

Oct 20, 2025

1.2.5

Oct 8, 2025

1.2.4

Jun 6, 2025

1.2.3

Jun 6, 2025

1.2.2

Feb 11, 2025

1.2.1

Feb 2, 2025

1.2.0

Dec 15, 2024

1.1.0

Dec 12, 2024

1.0.0

Dec 10, 2024

This version

0.2.7

Dec 1, 2024

0.2.4

Sep 27, 2024

0.2.3

Aug 3, 2024

0.2.2

Aug 2, 2024

0.2.1

Aug 2, 2024

0.2.0

Aug 2, 2024

0.1.12

Jul 30, 2024

0.1.11

Jul 4, 2024

0.1.10

Jul 3, 2024

0.1.9

Jun 25, 2024

0.1.8

Jun 24, 2024

0.1.7

Jun 24, 2024

0.1.6

Jun 23, 2024

0.1.5

Jun 23, 2024

0.1.4

Jun 23, 2024

0.1.3

Jun 19, 2024

0.1.2

Jun 18, 2024

0.1.1

Jun 18, 2024

0.1.0

Jun 18, 2024

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

pbqa-0.2.7.tar.gz (16.8 kB view details)

Uploaded Dec 1, 2024 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

PBQA-0.2.7-py3-none-any.whl (17.4 kB view details)

Uploaded Dec 1, 2024 Python 3

File details

Details for the file pbqa-0.2.7.tar.gz.

File metadata

Download URL: pbqa-0.2.7.tar.gz
Upload date: Dec 1, 2024
Size: 16.8 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.0.1 CPython/3.12.7

File hashes

Hashes for pbqa-0.2.7.tar.gz
Algorithm	Hash digest
SHA256	`d721faac01bc7b2fb5b83f145ef0fce546f57f9a46c65002fae7ff3c9fd1fee6`
MD5	`11d3aeb46d778aab86d86b540c87a404`
BLAKE2b-256	`df8869221a03c13fe02560a0cec9e953165469b60cf16b99afa08997b5d00222`

See more details on using hashes here.

File details

Details for the file PBQA-0.2.7-py3-none-any.whl.

File metadata

Download URL: PBQA-0.2.7-py3-none-any.whl
Upload date: Dec 1, 2024
Size: 17.4 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.0.1 CPython/3.12.7

File hashes

Hashes for PBQA-0.2.7-py3-none-any.whl
Algorithm	Hash digest
SHA256	`456d6586a5498ee6e5a2605f9aa029e69d88981283b036bfc9386f60da5bf1fc`
MD5	`d225e259bbdd94700dae6a85400c3364`
BLAKE2b-256	`d3a59749d7dc16da68fc5baa493da19696cc28f0e381bdf141e8582785c17571`

See more details on using hashes here.

PBQA 0.2.7

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Pattern Based Question and Answer

Description

Installation

Usage

llama.cpp

Python

Pattern Files

Cache

Roadmap

Relevant Literature

Contributing

Support

License and Acknowledgements

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes