Skip to main content

Bookworm - A LLM-powered bookmark search engine

Project description

bookworm 📖

main PyPI version

LLM-powered bookmark search engine

bookworm allows you to search from your local browser bookmarks using natural language. For times when you have a large collection of bookmarks and you can't quite remember where you put that one website you need at the moment.

Install

python -m pip install bookworm_genai

[!TIP] If you are using uvx then you can also just run this:

uvx --from bookworm_genai bookworm --help

Usage

export OPENAI_API_KEY=

# Run once and then anytime bookmarks across supported browsers changes
bookworm sync

# Sync bookmarks only from a specific browser
bookworm sync --browser-filter chrome

# Ask questions against the bookmark database
bookworm ask

# Ask questions against the bookmark database
# Specify the query when invoking the command
# If you omit this then you will be asked for a query when the tool is running
bookworm ask -q pandas

# Ask questions against the bookmark database and specify the number of results that should come back
bookworm ask -n 1

The sync process currently supports the following configurations:

Operating System Google Chrome Mozilla Firefox Brave Microsoft Edge
Linux
macOS
Windows

[!TIP] ✨ Want to contribute? See the adding an integration section.

Processes

bookworm sync

python -m bookworm sync
graph LR

subgraph Bookmarks
    Chrome(Chrome Bookmarks)
    Brave(Brave Bookmarks)
    Firefox(Firefox Bookmarks)
end

Bookworm(bookworm sync)

EmbeddingsService(Embeddings Service e.g OpenAIEmbeddings)

VectorStore(Vector Store e.g DuckDB)

Chrome -->|load bookmarks|Bookworm
Brave -->|load bookmarks|Bookworm
Firefox -->|load bookmarks|Bookworm

Bookworm -->|vectorize bookmarks|EmbeddingsService-->|store embeddings|VectorStore
Details

The vector database depicted above is stored locally on your machine. You can check it's location by running the following after installing this project:

from platformdirs import PlatformDirs

print(PlatformDirs('bookworm').user_data_dir)

bookworm ask

python -m bookworm ask
graph LR

query
Bookworm(bookworm ask)

subgraph _
    LLM(LLM e.g OpenAI)
    VectorStore(Vector Store e.g DuckDB)
end

query -->|user queries for information|Bookworm

Bookworm -->|simularity search|VectorStore -->|send similar docs + user query|LLM
LLM -->|send back response|Bookworm

Developer Setup

# LLMs
export OPENAI_API_KEY=

# Langchain (optional, but useful for debugging)
export LANGCHAIN_API_KEY=
export LANGCHAIN_TRACING_V2=true
export LANGCHAIN_PROJECT=bookworm

# Misc (optional)
export LOGGING_LEVEL=INFO

Recommendations:

poetry env use 3.9 # or path to your 3.9 installation

poetry shell
poetry install

bookworm --help
Running Linux tests on MacOS/Windows

If you are running on a non-linux machine, it may be helpful to run the provided Dockerfile to verify it's working on that environment.

You can build this via:

make docker_linux

You will need to have Docker installed to run this.

Adding an Integration

As you can see from usage, bookworm supports various integrations but not all. If you find one that you want to support one, then a change is needed inside integrations.py.

You can see in that file there is a variable called browsers that follows this structure:

browsers = {
    "BROWSER": {
        "PLATFORM": {
            ...
        }
    }
}

So say you wanted to add Chrome support in Windows then you would go under the Chrome key and then add a win32 key which has all the details. You can refer to existing examples but generally the contents of those details are where to find the bookmarks on the user's system along with how to interpret them.

You can also find a full list of the document loaders supported here.

Project details


Release history Release notifications | RSS feed

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

bookworm_genai-0.12.1b86.tar.gz (12.4 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

bookworm_genai-0.12.1b86-py3-none-any.whl (14.0 kB view details)

Uploaded Python 3

File details

Details for the file bookworm_genai-0.12.1b86.tar.gz.

File metadata

  • Download URL: bookworm_genai-0.12.1b86.tar.gz
  • Upload date:
  • Size: 12.4 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/1.8.5 CPython/3.12.3 Linux/6.8.0-1017-azure

File hashes

Hashes for bookworm_genai-0.12.1b86.tar.gz
Algorithm Hash digest
SHA256 955a37a0030533c3870deee04c792222be03822f48088eb4417878e3f9380dcc
MD5 93ed339f91208e61d06a18e736cab7a5
BLAKE2b-256 9a8b800f17cdaa31d4a071fa59e91eb62a1e8c74abf0cc7259fec20a2ea8efd1

See more details on using hashes here.

File details

Details for the file bookworm_genai-0.12.1b86-py3-none-any.whl.

File metadata

  • Download URL: bookworm_genai-0.12.1b86-py3-none-any.whl
  • Upload date:
  • Size: 14.0 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/1.8.5 CPython/3.12.3 Linux/6.8.0-1017-azure

File hashes

Hashes for bookworm_genai-0.12.1b86-py3-none-any.whl
Algorithm Hash digest
SHA256 e834883c2d3bdcebcf48938e4ea3685c8fc8b9fe86cf1e19ecda4af143fd9de4
MD5 5ed6d2c1a8416e21d11f5230993207bb
BLAKE2b-256 d2353562e7336d08c0c5524c4f8079cb771af75f920e61cc8c334c00b869981b

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page