Skip to main content

data wrangling for lists of tuples and dictionaries

Project description

https://img.shields.io/pypi/v/pytups.svg https://img.shields.io/pypi/l/pytups.svg https://img.shields.io/pypi/pyversions/pytups.svg https://travis-ci.org/pchtsp/pytups.svg?branch=master

What and why

The idea is to allow sparse operations to be executed in matrix data.

I grew used to the chained operations in R’s tidyverse packages or, although not a great fan myself, python’s pandas . I find myself using dictionary and list comprehensions all the time to pass from one data format to the other efficiently. But after doing it for the Nth time, I thought of automaticing it.

In my case, it helps me construct optimisation models with PuLP. I see other possible uses not related to OR.

I’ve implemented some additional methods to regular dictionaries, lists and sets to come up with interesting methods that somewhat quickly pass from one to the other and help with data wrangling.

In order for the operations to make any sense, the assumption that is done is that whatever you are using has the same ‘structure’. For example, if you a have a list of tuples: every element of the list is a tuple with the same size and the Nth element of the tuple has the same type, e.g. [(1, 'red', 'b', '2018-01'), (10, 'ccc', 'ttt', 'ff')]. Note that both tuples have four elements and the first one is a number, not a string. We do not check that this is consistent.

They’re made to always return a new object, so no “in-place” editing, hopefully.

Right now there are three classes to use: dictionaries, tuple lists and ordered sets.

Python versions

Python 3.5 and up.

Quick example

We index a tuple list according to some index positions.:

import pytups as pt
some_list_of_tuples = [('a', 'b', 'c', 1), ('a', 'b', 'c', 2), ('a', 'b', 'c', 45)]
tp_list = pt.TupList(some_list_of_tuples)
tp_list.to_dict(result_col=3)
# {('a', 'b', 'c'): [1, 2, 45]}
tp_list.to_dict(result_col=3).to_dictdict()
# {'a': {'b': {'c': [1, 2, 45]}}}
tp_list.to_dict(result_col=[2, 3])
# {('a', 'b'): [('c', 1), ('c', 2), ('c', 45)]}

We do some operations on dictionaries with common keys.:

import pytups as pt
some_dict = pt.SuperDict(a=1, b=2, c=3, d=5)
some_other_dict = pt.SuperDict(a=5, b=7, c=1)
some_other_dict + some_dict
# {'a': 6, 'b': 9, 'c': 4}
some_other_dict.vapply(lambda v: v**2)
# {'a': 25, 'b': 49, 'c': 1}
some_other_dict.kvapply(lambda k, v: v/some_dict[k])
# {'a': 5.0, 'b': 3.5, 'c': 0.3333333333333333}

Installing

pip install pytups

or, for the development version:

pip install https://github.com/pchtsp/pytups/archive/master.zip

Testing

Run the command:

python -m unittest discover -s tests

if the output says OK, all tests were passed.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

pytups-0.81.0.tar.gz (11.1 kB view details)

Uploaded Source

Built Distribution

pytups-0.81.0-py3-none-any.whl (12.1 kB view details)

Uploaded Python 3

File details

Details for the file pytups-0.81.0.tar.gz.

File metadata

  • Download URL: pytups-0.81.0.tar.gz
  • Upload date:
  • Size: 11.1 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.4.2 importlib_metadata/4.8.1 pkginfo/1.7.1 requests/2.26.0 requests-toolbelt/0.9.1 tqdm/4.62.2 CPython/3.9.7

File hashes

Hashes for pytups-0.81.0.tar.gz
Algorithm Hash digest
SHA256 d53be0c8a4acf0ddced672cef66033d968ac42e12f128f46f0ff954e81236e0a
MD5 7d2c3276df65f51fb9733761d7265088
BLAKE2b-256 cd231ad4eaa614b0134335ce34747af125c8c8efbc3665704eacee18aae1efaf

See more details on using hashes here.

File details

Details for the file pytups-0.81.0-py3-none-any.whl.

File metadata

  • Download URL: pytups-0.81.0-py3-none-any.whl
  • Upload date:
  • Size: 12.1 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.4.2 importlib_metadata/4.8.1 pkginfo/1.7.1 requests/2.26.0 requests-toolbelt/0.9.1 tqdm/4.62.2 CPython/3.9.7

File hashes

Hashes for pytups-0.81.0-py3-none-any.whl
Algorithm Hash digest
SHA256 d39d4264445b79a826cb782858f687064d1b3345fc24adb87a4fac4f18200834
MD5 a0e269b6ee11bbbed8b5b88c1f571253
BLAKE2b-256 2262f17b1e32fef2d2865aa87e82d6ac941abb9345f72bd721dcc762fcb3ae0c

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page