Skip to main content

A short description of the package

Project description

VBI Evaluate Blogs

vbi_evaluate_blogs is a Python package designed to evaluate Vietnamese crypto blog content quality. It uses Azure OpenAI to analyze text quality, image relevance, and fact accuracy with a focus on Web3/DeFi content.

Features

1. Text Content Evaluation (check_text_module.py)

  • Analyzes article structure and organization
  • Evaluates content quality and technical accuracy
  • Checks grammar and writing style for Vietnamese crypto content
  • Provides SEO optimization recommendations
  • Generates comprehensive quality reports

2. Image Analysis (check_image_module.py)

  • Analyzes image relevance and quality
  • Evaluates alt text and metadata
  • Checks image-text alignment
  • Provides visual accessibility recommendations
  • Supports common image formats (jpg, png, webp, etc.)

3. Fact Checking (check_fact_module.py)

  • Verifies claims using web search
  • Analyzes source credibility
  • Provides evidence-based verification
  • Uses SearxNG for research
  • Supports Vietnamese language validation

Installation

pip install vbi-evaluate-blogs
playwright install

Quick Start

  1. Set up environment variables:
AZURE_OPENAI_API_KEY="your_api_key"
AZURE_OPENAI_ENDPOINT="your_endpoint"
SEARXNG_URL="your_searx_instance" # For fact checking
  1. Basic usage:
from vbi_evaluate_blogs import check_text, check_image, check_fact
from langchain_openai import AzureChatOpenAI
from dotenv import load_dotenv
import os

load_dotenv()

# Initialize models
text_llm = AzureChatOpenAI(
    api_key=os.getenv("AZURE_OPENAI_API_KEY"),
    azure_endpoint=os.getenv("AZURE_OPENAI_ENDPOINT"),
    model="o3-mini",
    api_version="2024-12-01-preview"
)

image_llm = AzureChatOpenAI(
    api_key=os.getenv("AZURE_OPENAI_API_KEY"),
    azure_endpoint=os.getenv("AZURE_OPENAI_ENDPOINT"),
    model="gpt-4o-mini", # Vision model required
    api_version="2024-08-01-preview",
    temperature=0.7,
    max_tokens=16000
)

# Example content
content = """
# Sample Vietnamese Crypto Blog
Content with ![](image.jpg) and technical claims...
"""

# Get comprehensive analysis
text_report = check_text(text_llm, content)
image_report = check_image(text_llm, image_llm, content)
fact_report = check_fact(text_llm, content)

Module Details

Text Analysis Module

# Evaluate text content quality
result = check_text(text_llm, content)
print(result)
"""
Returns detailed report covering:
- Article structure analysis
- Content quality evaluation
- Grammar and style check
- SEO recommendations
"""

Image Analysis Module

# Analyze images in content
result = check_image(text_llm, image_llm, content)
print(result)
"""
Returns comprehensive report including:
- Image relevance scores
- Alt text evaluation
- Visual accessibility analysis
- Image-text alignment check
"""

Fact Checking Module

# Verify factual claims
result = check_fact(text_llm, content)
print(result)
"""
Returns fact check report with:
- Claim extraction
- Evidence analysis
- Source credibility
- Verification results
"""

Advanced Configuration

Custom Evaluation Criteria

You can customize the evaluation criteria by modifying the prompt templates in each module:

from vbi_evaluate_blogs.check_text_module import check_article_structure

# Custom structure analysis
result = check_article_structure(
    llm=text_llm,
    text=content,
    custom_criteria="Your custom evaluation criteria..."
)

Language Settings

The modules default to Vietnamese but support other languages:

from vbi_evaluate_blogs.check_image_module import ImageAnalyzer

analyzer = ImageAnalyzer(
    text_llm=text_llm,
    image_llm=image_llm,
    language="en"  # Change output language
)

Command Line Usage

Evaluate content directly from files:

# Full analysis 
python -m vbi_evaluate_blogs --file blog.md

# Specific checks
python -m vbi_evaluate_blogs --file blog.md --text --images
python -m vbi_evaluate_blogs --file blog.md --facts

License

MIT License. See LICENSE file for details.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

vbi_evaluate_blogs-0.1.20.tar.gz (37.1 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

vbi_evaluate_blogs-0.1.20-py3-none-any.whl (37.2 kB view details)

Uploaded Python 3

File details

Details for the file vbi_evaluate_blogs-0.1.20.tar.gz.

File metadata

  • Download URL: vbi_evaluate_blogs-0.1.20.tar.gz
  • Upload date:
  • Size: 37.1 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.1.0 CPython/3.10.12

File hashes

Hashes for vbi_evaluate_blogs-0.1.20.tar.gz
Algorithm Hash digest
SHA256 285c6c9fbe14cf78dd35cbcd137984e0e21713071f8cfff9379c4fa3141f0cba
MD5 44cc8e91706edb6e5fb59a21d3686a95
BLAKE2b-256 b92ab289a27c652155092d7e5c82352ac250aa05e1d89f43a0aea71a68a5ced0

See more details on using hashes here.

File details

Details for the file vbi_evaluate_blogs-0.1.20-py3-none-any.whl.

File metadata

File hashes

Hashes for vbi_evaluate_blogs-0.1.20-py3-none-any.whl
Algorithm Hash digest
SHA256 09fc5d034d55152ef2cfd5bd337d7ffaff6e0b65e3418f2e882cbc268aaae838
MD5 5c7fc7e928b81fdd1d747909f2f8ebba
BLAKE2b-256 eb93eed453ac66bf12545c69ff4dd336a434e563eca2b00daf398694adcd5dd7

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page