Interact with the Databricks Foundation Model API from python

These details have not been verified by PyPI

Project links

Homepage

Project description

Databricks Generative AI Inference SDK (Beta)

The Databricks Generative AI Inference Python library provides a user-friendly python interface to use the Databricks Foundation Model API.

It includes a pre-defined set of API classes Embedding, Completion, ChatCompletion with convenient functions to make API request, and to parse contents from raw json response.

We also offer a high level ChatSession object for easy management of multi-round chat completions, which is especially useful for your next chatbot development.

You can find more usage details in our SDK onboarding doc.

[!IMPORTANT]
We're preparing to release version 1.0 of the Databricks GenerativeAI Inference Python library.

Installation

pip install databricks-genai-inference

Usage

Embedding

from databricks_genai_inference import Embedding

Text embedding

response = Embedding.create(
    model="bge-large-en", 
    input="3D ActionSLAM: wearable person tracking in multi-floor environments")
print(f'embeddings: {response.embeddings[0]}')

Text embedding with instruction

response = Embedding.create(
    model="bge-large-en", 
    instruction="Represent this sentence for searching relevant passages:", 
    input="3D ActionSLAM: wearable person tracking in multi-floor environments")
print(f'embeddings: {response.embeddings[0]}')

Text embedding (batching)

[!IMPORTANT]
Support max batch size of 16

response = Embedding.create(
    model="bge-large-en", 
    input=[
        "3D ActionSLAM: wearable person tracking in multi-floor environments",
        "3D ActionSLAM: wearable person tracking in multi-floor environments"])
print(f'response.embeddings[0]: {response.embeddings[0]}\n')
print(f'response.embeddings[1]: {response.embeddings[1]}')

Text embedding with instruction (batching)

[!IMPORTANT]
Support one instruction per batch Batch size

response = Embedding.create(
    model="bge-large-en", 
    instruction="Represent this sentence for searching relevant passages:",
    input=[
        "3D ActionSLAM: wearable person tracking in multi-floor environments",
        "3D ActionSLAM: wearable person tracking in multi-floor environments"])
print(f'response.embeddings[0]: {response.embeddings[0]}\n')
print(f'response.embeddings[1]: {response.embeddings[1]}')

Text completion

from databricks_genai_inference import Completion

Text completion

response = Completion.create(
    model="mpt-7b-instruct", 
    prompt="Represent the Science title:")
print(f'response.text:{response.text:}')

Text completion (streaming)

[!IMPORTANT]
Only support batch size = 1 in streaming mode

response = Completion.create(
    model="mpt-7b-instruct", 
    prompt="Count from 1 to 100:",
    stream=True)
print(f'response.text:')
for chunk in response:
    print(f'{chunk.text}', end="")

Text completion (batching)

[!IMPORTANT]
Support max batch size of 16

response = Completion.create(
    model="mpt-7b-instruct", 
    prompt=[
        "Represent the Science title:", 
        "Represent the Science title:"])
print(f'response.text[0]:{response.text[0]}')
print(f'response.text[1]:{response.text[1]}')

Chat completion

from databricks_genai_inference import ChatCompletion

[!IMPORTANT]
Batching is not supported for ChatCompletion

Chat completion

response = ChatCompletion.create(model="llama-2-70b-chat", messages=[{"role": "system", "content": "You are a helpful assistant."},{"role": "user", "content": "Knock knock."}])
print(f'response.text:{response.message:}')

Chat completion (streaming)

response = ChatCompletion.create(model="llama-2-70b-chat", messages=[{"role": "system", "content": "You are a helpful assistant."},{"role": "user", "content": "Count from 1 to 30, add one emoji after each number"}], stream=True)
for chunk in response:
    print(f'{chunk.message}', end="")

Chat session

from databricks_genai_inference import ChatSession

[!IMPORTANT]
Streaming mode is not supported for ChatSession

chat = ChatSession(model="llama-2-70b-chat")
chat.reply("Kock, kock!")
print(f'chat.last: {chat.last}')
chat.reply("Take a guess!")
print(f'chat.last: {chat.last}')

print(f'chat.history: {chat.history}')
print(f'chat.count: {chat.count}')

Project details

These details have not been verified by PyPI

Project links

Homepage

Release history Release notifications | RSS feed

0.2.3

Mar 27, 2024

0.2.2

Mar 21, 2024

0.2.1

Feb 17, 2024

0.2.0

Feb 13, 2024

0.1.3

Dec 20, 2023

0.1.2

Dec 7, 2023

This version

0.1.1

Nov 13, 2023

0.1.0

Nov 10, 2023

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

databricks-genai-inference-0.1.1.tar.gz (20.8 kB view details)

Uploaded Nov 13, 2023 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

databricks_genai_inference-0.1.1-py3-none-any.whl (15.7 kB view details)

Uploaded Nov 13, 2023 Python 3

File details

Details for the file databricks-genai-inference-0.1.1.tar.gz.

File metadata

Download URL: databricks-genai-inference-0.1.1.tar.gz
Upload date: Nov 13, 2023
Size: 20.8 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/4.0.2 CPython/3.9.6

File hashes

Hashes for databricks-genai-inference-0.1.1.tar.gz
Algorithm	Hash digest
SHA256	`5763a841714e7cc561d0029dca8a8a50b15323186199dc6d2a466d13827f97e6`
MD5	`ee1e89e90e6a9467b647fd2352cf8c7c`
BLAKE2b-256	`abb14b4bbc75a8b7626cb97f8e6ae34def51eee52c4f2639bcd847cb062c779f`

See more details on using hashes here.

File details

Details for the file databricks_genai_inference-0.1.1-py3-none-any.whl.

File metadata

Download URL: databricks_genai_inference-0.1.1-py3-none-any.whl
Upload date: Nov 13, 2023
Size: 15.7 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/4.0.2 CPython/3.9.6

File hashes

Hashes for databricks_genai_inference-0.1.1-py3-none-any.whl
Algorithm	Hash digest
SHA256	`df167ff40457053b2a338bb60288fc4ce54a12d9bad64557ad9632ec5eacf07a`
MD5	`32d14fbcd7f19583bb912ea2f3c76c4f`
BLAKE2b-256	`35c78b699c69f8dfdd646bada45f9c4e507992a5edf0ea49694e25f07045fa81`

See more details on using hashes here.

databricks-genai-inference 0.1.1

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Databricks Generative AI Inference SDK (Beta)

Installation

Usage

Embedding

Text embedding

Text embedding with instruction

Text embedding (batching)

Text embedding with instruction (batching)

Text completion

Text completion

Text completion (streaming)

Text completion (batching)

Chat completion

Chat completion

Chat completion (streaming)

Chat session

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes