Skip to main content

fedmsg consumer that extracts hashes of source files

Project description

A fedmsg consumer that extracts and stores hashes of source files.

summershum is composed of two components:

  • A fedmsg consumer plugin that listens for org.fedoraproject.prod.git.lookaside.new messages. Whenever a contributor uploads a new source tarball to the lookaside cache, summershum will download that tarball, unpack it, and calculate the sha1 sum of every file in the tarball. Those hashes are then stored in a database to be queried later.
  • A cli tool summershum-cli that queries datagrepper for the fedmsg history. It then crawls through old lookaside messages to fill in data where it was missed.

With the summershum database, we can then make some interesting queries in short time:

  • how many files have this hash sum in all of fedora? and for which packages ?
  • we can easily find what is bundling what and generate a programatic list
  • we could check the db in taskotron tests
  • we could check to see how many packages include the full GPL license
  • how many packages have that license but with the old FSF address

Project details


Release history Release notifications

History Node

0.1.5

This version
History Node

0.1.4

History Node

0.1.3

History Node

0.1.2

History Node

0.1.1

History Node

0.1

History Node

0.0.1

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Filename, size & hash SHA256 hash help File type Python version Upload date
summershum-0.1.4.tar.gz (17.2 kB) Copy SHA256 hash SHA256 Source None Feb 21, 2014

Supported by

Elastic Elastic Search Pingdom Pingdom Monitoring Google Google BigQuery Sentry Sentry Error logging CloudAMQP CloudAMQP RabbitMQ AWS AWS Cloud computing Fastly Fastly CDN DigiCert DigiCert EV certificate StatusPage StatusPage Status page