Skip to main content

instascrape logo

instascrape: Instagram scraping for humans

What is it?

instascrape is a powerful, lightweight Python library for scraping Instagram data and content with no configurations necessary! It is designed with flexibility and developer productivity in mind so you can stop wasting valuable time preparing Instagram data and just start analyzing it :muscle:

Official website

Version Code style: black Release License

Downloads Activity Dependencies Issues

Example showing tech profile scrapes

Key features

  • :muscle: Powerful, object-oriented scraping tools
  • :dancer: Flexibly determines whether you want to scrape HTML, JSON, BeautifulSoup, or request and scrape the URL itself
  • :floppy_disk: Download content to your computer as png, jpg, mp4, and mp3
  • :art: Dynamically retrieve HTML embed code for posts
  • :musical_score: Expressive and consistent API for concise and elegant code
  • :bar_chart: Designed for seamless integration with Selenium, Pandas, and other industry standard tools for data collection and analysis
  • :hammer: Lightweight: you don't have to build a hammer factory when all you need is the hammer
  • :spider_web: The only hard dependencies are Requests and Beautiful Soup; no more worrying about configurations or webdrivers
  • :watch: Proven to work as of December, 2020

Table of Contents


:computer: Installation

Minimum Python version

This library currently requires Python 3.7 or higher.

pip

Install from PyPI using

$ pip3 install insta-scrape

WARNING: make sure you install insta-scrape and not a package with a similar name!


:mag_right: Sample Usage

All top-level, ready-to-use features can be imported using:

from instascrape import *

instascrape uses clean, consistent, and expressive syntax to make the developer experience as painless as possible.

# Instantiate the scraper objects 
google = Profile('https://www.instagram.com/google/')
google_post = Post('https://www.instagram.com/p/CG0UU3ylXnv/')
google_hashtag = Hashtag('https://www.instagram.com/explore/tags/google/')

# Scrape their respective data 
google.scrape()
google_post.scrape()
google_hashtag.scrape()

After being scraped, relevant attributes can be accessed with dot or bracket notation

print(google.followers)
print(google_post['hashtags'])
print(google_hashtag.amount_of_posts)
>>> 12262794
>>> ['growwithgoogle']
>>> 9053408

:books: Documentation

The official documentation can be found on Read The Docs :newspaper:


:newspaper: Blog Posts

Check out blog posts on the official site or DEV for ideas and tutorials!


:pray: Contributing

All contributions, bug reports, bug fixes, documentation improvements, enhancements, and ideas are welcome!

Feel free to open an Issue, check out existing Issues, or start a discussion.

Beginners to open source are highly encouraged to participate and ask questions if you're unsure what to do/where to start :heart:


:spider_web: Dependencies

Instascrape primarily relies on two third-party libraries for requesting and scraping Instagram HTML content:

  1. Requests: HTTP requests
  2. BeautifulSoup: Scraping and parsing HTML data.

The rest of its functionality is provided directly from Python 3's standard library for unobtrusive code under the hood with little to no overhead.


:credit_card: License

MIT


:grey_question: Support

Check out the FAQ

Reach out to me if you have questions or ideas!


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

insta-scrape-1.4.0.tar.gz (19.2 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

insta_scrape-1.4.0-py3-none-any.whl (25.0 kB view details)

Uploaded Python 3

File details

Details for the file insta-scrape-1.4.0.tar.gz.

File metadata

  • Download URL: insta-scrape-1.4.0.tar.gz
  • Upload date:
  • Size: 19.2 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.1.1 pkginfo/1.4.2 requests/2.22.0 setuptools/45.2.0 requests-toolbelt/0.8.0 tqdm/4.30.0 CPython/3.8.5

File hashes

Hashes for insta-scrape-1.4.0.tar.gz
Algorithm Hash digest
SHA256 ee3c75f37b58c7b699680ccef121da3054c5b3475c12da168c51f15732957340
MD5 03f2b5cf524ad731d00bfc5d3c3b35e1
BLAKE2b-256 25f665b73d61e7081f840c9f23d0a265a13482554b3e5345bba5ff4796c786d2

See more details on using hashes here.

File details

Details for the file insta_scrape-1.4.0-py3-none-any.whl.

File metadata

  • Download URL: insta_scrape-1.4.0-py3-none-any.whl
  • Upload date:
  • Size: 25.0 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.1.1 pkginfo/1.4.2 requests/2.22.0 setuptools/45.2.0 requests-toolbelt/0.8.0 tqdm/4.30.0 CPython/3.8.5

File hashes

Hashes for insta_scrape-1.4.0-py3-none-any.whl
Algorithm Hash digest
SHA256 909f47f7923f0ea3e5a276f3c59dc6e2e5e0386f2727213440fe715276ccf5bc
MD5 5b15c608320eff7fc78d2e426606afd8
BLAKE2b-256 ac674b5586cd4d2f77d0c9ace61ce0da55bda14dbdf510a06631d35974b9f1af

See more details on using hashes here.

Release history Release notifications | RSS feed

2.1.2

2 files

2.1.1

2 files

2.1.0

2 files

2.0.2

2 files

2.0.0

1 file

1.7.1

2 files

1.7.0

1 file

1.6.1

2 files

1.6.0

1 file

1.5.0

1 file

This release

1.4.0 This release

2 files

1.3.4

1 file

1.3.3

1 file

1.3.2

1 file

1.3.1

1 file

1.3.0

1 file

1.2.8

1 file

1.2.7

1 file

1.2.6

1 file

1.2.5

1 file

1.2.4

1 file

1.2.3

1 file

1.2.2

1 file

1.2.1

1 file

1.2.0

1 file

1.1.0

1 file

1.0.1

1 file

1.0.0

1 file

0.11.0

1 file

0.10.0

1 file

0.9.0

1 file

0.8.1

1 file

0.8.0

1 file

0.7.1

1 file

0.7.0

1 file

0.6.7

1 file

0.6.6

1 file

0.6.5

1 file

0.6.4

1 file

0.6.3

1 file

0.6.2

1 file

0.6.1

1 file

0.6.0

1 file

0.5.1

1 file

0.5.0

1 file

0.4.0

1 file

0.3.0

1 file

0.2.1

2 files

0.2.0

2 files

0.1.0

2 files

0.0.7

2 files

0.0.6

2 files

0.0.5

2 files

0.0.4

2 files

0.0.0

1 file

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page