Skip to main content

A package for evaluating PDFs with text, image, and fact-checking modules.

Project description

Evaluate Blogs

Evaluate Blogs is a Python package designed to evaluate the content of PDF documents comprehensively. It combines advanced text analysis, image evaluation, and fact-checking capabilities to ensure the quality, relevance, and credibility of the document. The package leverages Azure OpenAI and other state-of-the-art tools to provide accurate and insightful evaluations.

Features

1. Text Content Evaluation

  • Analyzes the grammar, structure, and coherence of the text.
  • Provides feedback on the readability and relevance of the content.
  • Detects potential issues such as redundancy or lack of clarity.

2. Image Relevance Analysis

  • Evaluates the quality and relevance of images in the document.
  • Ensures that images align with the context and purpose of the document.
  • Detects low-quality or irrelevant images.

3. Fact-Checking

  • Verifies the factual accuracy of the content using external sources.
  • Highlights potential inaccuracies or unsupported claims.
  • Ensures the credibility of the information presented.

4. Modular Design

  • Each feature is implemented as a separate module, allowing for flexible usage.
  • Users can choose to run specific evaluations or combine them as needed.

Install

pip install vbi-evaluate-blogs
playwright install

Usage

Basic Example

Here is an example of how to use the vbi_evaluate_blogs package to analyze a PDF:

from evaluate_module import evaluate
from langchain_openai import AzureChatOpenAI

# Initialize the Azure OpenAI model
model = AzureChatOpenAI(api_key="your_api_key")

# Path to the PDF file
pdf_path = "path/to/your/pdf_file.pdf"

# Evaluate the PDF
result = evaluate(pdf_path, model=model)

# Print the evaluation result
print(result)

Detailed Usage

1. Text Evaluation

The text evaluation module analyzes the PDF's text content for grammar, structure, and relevance. It provides insights into the quality of the written content.

from evaluate_module import evaluate_text

# Path to the PDF file
pdf_path = "path/to/your/pdf_file.pdf"

# Analyze text content
text_result = evaluate_text(pdf_path, model=model)
print("Text Evaluation Result:", text_result)

2. Image Analysis

The image analysis module checks the relevance and quality of images in the PDF. It ensures that images align with the document's context.

from evaluate_module import evaluate_images

# Path to the PDF file
pdf_path = "path/to/your/pdf_file.pdf"

# Analyze images in the PDF
image_result = evaluate_images(pdf_path)
print("Image Analysis Result:", image_result)

3. Fact-Checking

The fact-checking module verifies the factual accuracy of the content using external sources. This ensures the credibility of the information presented.

from evaluate_module import evaluate_facts

# Path to the PDF file
pdf_path = "path/to/your/pdf_file.pdf"

# Perform fact-checking
fact_result = evaluate_facts(pdf_path, model=model)
print("Fact-Checking Result:", fact_result)

Command-Line Usage

You can also use the package via the command line for quick evaluations:

python -m evaluate_module --file path/to/your/pdf_file.pdf

Additional Options

  • --text: Perform only text evaluation.
  • --images: Perform only image analysis.
  • --facts: Perform only fact-checking.

Example:

python -m evaluate_module --file path/to/your/pdf_file.pdf --text --images

Advanced Usage

Customizing the Model

You can customize the Azure OpenAI model by providing additional parameters during initialization:

model = AzureChatOpenAI(api_key="your_api_key", temperature=0.7, max_tokens=1000)

Combining Modules

You can combine multiple modules to perform a comprehensive evaluation:

from evaluate_module import evaluate_text, evaluate_images, evaluate_facts

# Path to the PDF file
pdf_path = "path/to/your/pdf_file.pdf"

# Perform evaluations
text_result = evaluate_text(pdf_path, model=model)
image_result = evaluate_images(pdf_path)
fact_result = evaluate_facts(pdf_path, model=model)

# Combine results
combined_result = {
    "text": text_result,
    "images": image_result,
    "facts": fact_result
}

print("Combined Evaluation Result:", combined_result)

License

This project is licensed under the MIT License. See the LICENSE file for details.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

vbi_evaluate_blogs-0.1.6.tar.gz (24.5 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

vbi_evaluate_blogs-0.1.6-py3-none-any.whl (28.4 kB view details)

Uploaded Python 3

File details

Details for the file vbi_evaluate_blogs-0.1.6.tar.gz.

File metadata

  • Download URL: vbi_evaluate_blogs-0.1.6.tar.gz
  • Upload date:
  • Size: 24.5 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.1.0 CPython/3.10.12

File hashes

Hashes for vbi_evaluate_blogs-0.1.6.tar.gz
Algorithm Hash digest
SHA256 d0897d501fa54be156b8999557f51f51d4cf2d3d335dbba5f20b6af6cb139ab2
MD5 b1cd457e869aa47b8a8e3087d3604ec1
BLAKE2b-256 acbe527e68f97c363886d83fb567fb6c5f75b9aa6c1cb8096fdb8c3107c5dfa7

See more details on using hashes here.

File details

Details for the file vbi_evaluate_blogs-0.1.6-py3-none-any.whl.

File metadata

File hashes

Hashes for vbi_evaluate_blogs-0.1.6-py3-none-any.whl
Algorithm Hash digest
SHA256 75b30e6cfa9d3f9c49368605acfb728fcc5037052c6d5f85d34a432c2aeadb74
MD5 3abc1a955d821971f0d8975663309386
BLAKE2b-256 361efc12c9d1428a02eb9d7b6627529e7af6a5cdbdb1e13613301577857c9f5e

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page