AI-powered document perception and analysis MCP server with intelligent provider selection

These details have not been verified by PyPI

Project links

Project description

🔍 Docsray MCP Server

Docsray is a powerful Model Context Protocol (MCP) server that gives AI assistants like Claude advanced document perception capabilities. Extract text, navigate pages, analyze structure, and understand any document with ease.

✅ Status: Published to PyPI and TestPyPI - Working in Cursor, Claude Desktop, and other MCP clients

✨ Features

🎯 Five Powerful Tools

docsray_peek - Quick document overview with format detection and provider capabilities
docsray_map - Generate comprehensive document structure maps with caching
docsray_xray - AI-powered deep analysis extracting entities, relationships, and insights
docsray_extract - Extract content in multiple formats (markdown, text, JSON, tables)
docsray_seek - Navigate to specific pages, sections, or search for content

🔌 Multi-Provider Architecture

PyMuPDF4LLM - Lightning-fast PDF processing (✅ Implemented)
- Fast markdown extraction
- Basic table detection
- Multi-page support
- Always enabled as fallback
LlamaParse - Deep document understanding with LLMs (✅ Implemented)
- AI-powered entity extraction
- Custom analysis instructions
- Comprehensive caching in .docsray directories
- Rich format preservation (markdown, images, tables)
PyTesseract - OCR for scanned documents (🔄 Planned)
Mistral OCR - AI-powered OCR and analysis (🔄 Planned)

🚀 Key Benefits

Universal Input Support - Local files (./path, ../path, /absolute) and URLs (https://)
Intelligent Provider Selection - Automatically chooses the best tool for each task
Smart Caching - LlamaParse results cached in .docsray directories for instant access
Dynamic Discovery - Tools report actual capabilities based on what's enabled
Production Ready - Comprehensive error handling, logging, and 56 tests
Self-Documenting - Built-in resources for discovery by MCP clients

📦 Installation

Quick Start with uvx (Recommended)

# Run directly without installation
uvx docsray-mcp start

# Or install globally
uv tool install docsray-mcp
# Then run with:
docsray start
# or
docsray-mcp start

Alternative: Install with pip

# Basic installation (PyMuPDF4LLM only)
pip install docsray-mcp

# With LlamaParse for AI analysis
pip install "docsray-mcp[ai]"

# Development installation
pip install -e ".[dev]"

🚀 Quick Start

1. Set up API Keys (Optional but Recommended)

Create a .env file in your project:

# For AI-powered analysis with LlamaParse
LLAMAPARSE_API_KEY=llx-your-key-here

# Or use environment variables
export LLAMAPARSE_API_KEY=llx-your-key-here

Get your free LlamaParse API key at cloud.llamaindex.ai

2. Configure with Your MCP Client

For Cursor

Add to your Cursor settings:

{
  "mcpServers": {
    "docsray": {
      "command": "uvx",
      "args": ["docsray-mcp"],
      "env": {
        "LLAMAPARSE_API_KEY": "llx-your-key-here"
      }
    }
  }
}

For Claude Desktop

Add to ~/Library/Application Support/Claude/claude_desktop_config.json:

{
  "mcpServers": {
    "docsray": {
      "command": "uvx",
      "args": ["docsray-mcp"],
      "env": {
        "LLAMAPARSE_API_KEY": "llx-your-key-here"
      }
    }
  }
}

📚 Usage Examples

Basic Document Overview

Peek at ./document.pdf to see its structure and available formats

Extract Entities from Contracts

Xray ./contract.pdf and extract all parties, dates, payment terms, and obligations

Navigate Documents

Map the complete structure of ./manual.pdf including all sections and subsections

Extract Specific Content

Extract pages 10-20 from ./report.pdf as markdown

Analyze Web Documents

Analyze https://arxiv.org/pdf/2301.00234.pdf for methodology and key findings

Compare Providers

Extract text from document.pdf with provider pymupdf4llm (fast)
Xray document.pdf with provider llama-parse (AI analysis)

🛠️ Advanced Configuration

Environment Variables

# Provider Configuration
DOCSRAY_PYMUPDF4LLM_ENABLED=true  # Always true by default
DOCSRAY_LLAMAPARSE_ENABLED=true
LLAMAPARSE_API_KEY=llx-your-key

# Performance Tuning
DOCSRAY_CACHE_ENABLED=true
DOCSRAY_CACHE_TTL=3600
DOCSRAY_MAX_CONCURRENT_REQUESTS=5
DOCSRAY_TIMEOUT_SECONDS=30

# Logging
DOCSRAY_LOG_LEVEL=INFO

Provider Capabilities

PyMuPDF4LLM (Always Available)

✅ Fast text extraction
✅ Markdown formatting
✅ Basic table detection
✅ Multi-page support
❌ No AI analysis
❌ No OCR

LlamaParse (When API Key Configured)

✅ AI-powered analysis
✅ Entity extraction
✅ Custom instructions
✅ Table extraction
✅ Image extraction
✅ Layout preservation
✅ Relationship mapping
✅ Result caching

🧪 Testing

# Run all tests
pytest tests/

# Run only unit tests (no API calls)
pytest tests/unit/

# Run integration tests
pytest tests/integration/

# Run with coverage
pytest tests/ --cov=src/docsray --cov-report=html

Current test coverage: 52 tests passing with comprehensive coverage across all components

📖 API Reference

Tool: docsray_peek

Get quick document overview and metadata.

{
  "document_url": "path/to/document.pdf",
  "depth": "structure",  # metadata | structure | preview
  "provider": "auto"     # auto | pymupdf4llm | llama-parse
}

Tool: docsray_map

Generate comprehensive document structure map.

{
  "document_url": "path/to/document.pdf",
  "include_content": false,
  "analysis_depth": "deep",  # basic | deep | comprehensive
  "provider": "auto"
}

Tool: docsray_xray

Deep AI-powered document analysis.

{
  "document_url": "path/to/document.pdf",
  "analysis_type": ["entities", "key-points"],
  "custom_instructions": "Extract all dates and amounts",
  "provider": "llama-parse"
}

Tool: docsray_extract

Extract content in various formats.

{
  "document_url": "path/to/document.pdf",
  "extraction_targets": ["text", "tables"],
  "output_format": "markdown",  # markdown | text | json
  "pages": [1, 2, 3],  # Optional: specific pages
  "provider": "auto"
}

Tool: docsray_seek

Navigate to specific document locations.

{
  "document_url": "path/to/document.pdf",
  "target": {"page": 5},  # or {"section": "Introduction"} or {"query": "search text"}
  "extract_content": true,
  "provider": "auto"
}

🏗️ Architecture

docsray-mcp/
├── src/docsray/
│   ├── server.py           # FastMCP server with discovery resources
│   ├── providers/          # Provider implementations
│   │   ├── base.py        # Provider interface
│   │   ├── pymupdf4llm.py # Fast PDF extraction
│   │   └── llamaparse.py  # AI-powered analysis
│   ├── tools/             # MCP tool implementations
│   │   ├── peek.py        # Document overview
│   │   ├── map.py         # Structure mapping
│   │   ├── xray.py        # Deep analysis
│   │   ├── extract.py     # Content extraction
│   │   └── seek.py        # Navigation
│   └── utils/             # Utilities
│       ├── cache.py       # Document caching
│       └── llamaparse_cache.py  # LlamaParse .docsray cache
├── tests/
│   ├── unit/              # Fast isolated tests
│   ├── integration/       # Component interaction tests
│   └── manual/            # Debugging scripts
└── PROMPTS.md            # Example prompts for all use cases

🤝 Contributing

We welcome contributions! See CONTRIBUTING.md for guidelines.

Development Setup

# Clone the repository
git clone https://github.com/docsray/docsray-mcp.git
cd docsray-mcp

# Install in development mode
pip install -e ".[dev]"

# Run tests
pytest tests/

# Run linting
ruff check src/

📄 License

This project is licensed under the Apache License 2.0 - see the LICENSE file for details.

🙏 Acknowledgments

Built on FastMCP framework
Document processing powered by PyMuPDF4LLM
AI analysis powered by LlamaParse
Inspired by the Model Context Protocol specification

📬 Support

Made with ❤️ for the MCP ecosystem

Project details

These details have not been verified by PyPI

Project links

Release history Release notifications | RSS feed

This version

0.3.3

Aug 6, 2025

0.3.2

Aug 6, 2025

0.3.1

Aug 6, 2025

0.3.0

Aug 6, 2025

0.2.0

Aug 6, 2025

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

docsray_mcp-0.3.3.tar.gz (46.5 kB view details)

Uploaded Aug 6, 2025 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

docsray_mcp-0.3.3-py3-none-any.whl (50.8 kB view details)

Uploaded Aug 6, 2025 Python 3

File details

Details for the file docsray_mcp-0.3.3.tar.gz.

File metadata

Download URL: docsray_mcp-0.3.3.tar.gz
Upload date: Aug 6, 2025
Size: 46.5 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.12.11

File hashes

Hashes for docsray_mcp-0.3.3.tar.gz
Algorithm	Hash digest
SHA256	`f813428b2f23c1833249752933197afde7ba3145dc9438371091a42657913155`
MD5	`44b944ced4ad673e82d68d0c3b39d301`
BLAKE2b-256	`f49ba927f03cb84d539434d067cd335cd77fed64f6c188c20defee6a64758050`

See more details on using hashes here.

File details

Details for the file docsray_mcp-0.3.3-py3-none-any.whl.

File metadata

Download URL: docsray_mcp-0.3.3-py3-none-any.whl
Upload date: Aug 6, 2025
Size: 50.8 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.12.11

File hashes

Hashes for docsray_mcp-0.3.3-py3-none-any.whl
Algorithm	Hash digest
SHA256	`8ecd57a743a1c473c298884e2ef95be4bcaa3d538226fa4f8c8ca6253ef0f7a6`
MD5	`2c62971759cdf000a698726843f78b76`
BLAKE2b-256	`ae6bc28592742595d74b0a4cb09737456a73b2f655a271f34f06b1fc669bf8e7`

See more details on using hashes here.

docsray-mcp 0.3.3

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

🔍 Docsray MCP Server

✨ Features

🎯 Five Powerful Tools

🔌 Multi-Provider Architecture

🚀 Key Benefits

📦 Installation

Quick Start with uvx (Recommended)

Alternative: Install with pip

🚀 Quick Start

1. Set up API Keys (Optional but Recommended)

2. Configure with Your MCP Client

For Cursor

For Claude Desktop

📚 Usage Examples

Basic Document Overview

Extract Entities from Contracts

Navigate Documents

Extract Specific Content

Analyze Web Documents

Compare Providers

🛠️ Advanced Configuration

Environment Variables

Provider Capabilities

PyMuPDF4LLM (Always Available)

LlamaParse (When API Key Configured)

🧪 Testing

📖 API Reference

Tool: docsray_peek

Tool: docsray_map

Tool: docsray_xray

Tool: docsray_extract

Tool: docsray_seek

🏗️ Architecture

🤝 Contributing

Development Setup

📄 License

🙏 Acknowledgments

📬 Support

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes