Skip to main content

A crawler for PsychonautWiki wrote in python.

Project description

PsychonautCrawler

What is this? Why?

This project is a crawler, or scraper if you prefer, for the website PsychonautWiki. I plan to be very extensive with it since this is my first big project with Python, as I learn it in college. A web crawler is a bot that goes on websites and indexes them; search engines use them for this reason. However, they can be controversial because they can go and index websites meant to be private, thus leading to websites including robots.txt that request them not to index some, or all, of the website.

Like I stated above this is a learning project for myself, thus it will be going to be going through a lot of changes and there probably are going to be mistakes.

Currently these are my plans (are subject to change):

  • Fix the filing project structure
  • Understand how to recognize if there is a robots.txt warding me away
  • Fix the name from Psychonaught -> Psychonaut
  • Enable searching for specific substances
  • Print to documents
    • JSON
    • CSV
    • Markdown
  • Ability to see other aspects
    • Experience reports
    • Subjective effects
    • Research
    • Dangerous interactions
    • Legality
  • Documentation
    • In /docs/ would be nice
    • Use the projects on GitHub more
    • Fix the PyPi resource

Project Structure

This project also is available on pypi, I use upload it there so that I can import the project globally and also people can install it using pip.

PsychonautCrawler/
├── bin/
├── docs/
├── psychonautcrawler/
│	 └── crawler.py
└──  tests/
Directory Description
PsychonautCrawler/ Root
bin/ Programs and scripts wrote with the package.
docs/ Folder to hold documentation that is generated by Read the Docs
tests/ Folder for conducting tests on variables, functions, libraries etc.

Documentation

Documentation Status

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

PsychonautCrawler-1.0.2.tar.gz (4.2 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

PsychonautCrawler-1.0.2-py3-none-any.whl (5.2 kB view details)

Uploaded Python 3

File details

Details for the file PsychonautCrawler-1.0.2.tar.gz.

File metadata

  • Download URL: PsychonautCrawler-1.0.2.tar.gz
  • Upload date:
  • Size: 4.2 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.4.1 importlib_metadata/4.4.0 pkginfo/1.7.0 requests/2.25.1 requests-toolbelt/0.9.1 tqdm/4.61.0 CPython/3.9.5

File hashes

Hashes for PsychonautCrawler-1.0.2.tar.gz
Algorithm Hash digest
SHA256 15503dfe1e57b06939d9e263aa9eb425670d0c931ac30cb08eeb6085455a8989
MD5 ca78f88371344c66232b46a14672911c
BLAKE2b-256 8f22f7e4eee5747eefcb42b5ad1ec5418609a2a39730b417d56501004e722a4f

See more details on using hashes here.

File details

Details for the file PsychonautCrawler-1.0.2-py3-none-any.whl.

File metadata

  • Download URL: PsychonautCrawler-1.0.2-py3-none-any.whl
  • Upload date:
  • Size: 5.2 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.4.1 importlib_metadata/4.4.0 pkginfo/1.7.0 requests/2.25.1 requests-toolbelt/0.9.1 tqdm/4.61.0 CPython/3.9.5

File hashes

Hashes for PsychonautCrawler-1.0.2-py3-none-any.whl
Algorithm Hash digest
SHA256 7bdcd355442567159e267881148db980ffd697bebe7f5e15871276afaba3f8ca
MD5 555654a58ab38cd4ba672cb388ef70e4
BLAKE2b-256 2df9e049618f566f02850f37ba58f30ffcc140a1c14c5193103ecb85697b7367

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page