Skip to main content

AI Research Assistant

An automated research tool that uses web scraping, AI, and natural language processing to gather, analyze, and synthesize information on specified topics.

Features

  • Automated web searching and content extraction
  • AI-powered content analysis and summarization
  • Duplicate content detection and removal
  • Progress tracking and research history
  • Customizable research parameters
  • Structured output in various formats
  • File-based research organization

Requirements

  • Python 3.x
  • OpenAI API key
  • Internet connection

Dependencies

openai
googlesearch-python
requests
beautifulsoup4
datetime

Installation

  1. Clone the repository
  2. Install required packages:
pip install openai googlesearch-python requests beautifulsoup4
  1. Set up your API key

Usage

researchBot = ResearchSession()
researchBot.apiKey = 'your-api-key'
researchBot.topic = 'Your Research Topic'
researchBot.numSources = 3  # Number of desired sources
researchBot.outputFormat = 'formal essay'  # Or other format
researchBot.startResearch()

File Structure

The program creates a research folder with timestamped subfolders containing:

  • links.txt: List of discovered URLs
  • history.txt: Research session history
  • extracted_data.txt: Processed content from sources
  • final_research.txt: Final compiled research output

Key Methods

  • webSearch(): Performs Google searches
  • readWebpage(): Extracts content from URLs
  • getAIResponse(): Interfaces with AI for analysis
  • cleanLinksFileForDuplicates(): Removes duplicate sources
  • cleanExtractedDataFileForDuplicateData(): Removes redundant content
  • finalize(): Generates final research document

Research Process

  1. Conducts web searches for relevant sources
  2. Extracts and processes content from sources
  3. AI analyzes and summarizes information
  4. Removes duplicates and organizes data
  5. Continues until research criteria are met
  6. Generates final formatted document

Notes

  • Requires valid API key for AI services
  • Research quality depends on source availability
  • Internet connectivity required throughout process
  • Output format can be customized

Release files for GPTResearch 5.0.3

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for GPTResearch 5.0.3
File Size Uploaded
gptresearch-5.0.3.tar.gz 18.4 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for GPTResearch 5.0.3
File Interpreter ABI Platform
gptresearch-5.0.3-py3-none-any.whl Python 3 none any Details

Total release size: 37.6 kB

Release files / gptresearch-5.0.3.tar.gz

Download URL gptresearch-5.0.3.tar.gz
Size 18.4 kB
Tags Source
SHA-256 checksum
How to use checksums
c4719a0f976b31b351196a7fb62adaf4e2769341d7cc66ae0450687601531178
BLAKE2b-256 checksum
How to use checksums
038418a8ff13bc8acecf982b30bf8a4c8818646a58a142f6b25a3de6a81b44e8
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via twine/6.0.1 CPython/3.11.9

Release files / gptresearch-5.0.3-py3-none-any.whl

Download URL gptresearch-5.0.3-py3-none-any.whl
Size 19.1 kB
Tags Python 3
SHA-256 checksum
How to use checksums
8e930696f663d9beb3d9e70279e8f37d9b7f264d1a2a6aa9ea884b5ac101dbb8
BLAKE2b-256 checksum
How to use checksums
9bcd27ec3087ae6f9014900d12385f2eb976a1348dc92b61863ec0620e621ba5
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via twine/6.0.1 CPython/3.11.9

Release history Release notifications | RSS feed

This release

5.0.3 This release

2 release files

5.0.2

2 release files

5.0.1

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page