Skip to main content

Classes for data manipulation

Project description

Python Classes for Data Manipulation

Test Documentation Status PyPI Downloads

Dataiter currently includes the following classes.

DataFrame is a class for tabular data similar to R's data.frame or pandas.DataFrame. It is under the hood a dictionary of NumPy arrays and thus capable of fast vectorized operations. You can consider this to be a light-weight alternative to Pandas with a simple and consistent API. Performance-wise Dataiter relies on NumPy and Numba and is likely to be at best comparable to Pandas.

ListOfDicts is a class useful for manipulating data from JSON APIs. It provides functionality similar to libraries such as Underscore.js, with manipulation functions that iterate over the data and return a shallow modified copy of the original. attd.AttributeDict is used to provide convenient access to dictionary keys.

GeoJSON is a simple wrapper class that allows reading a GeoJSON file into a DataFrame and writing a data frame to a GeoJSON file. Any operations on the data are thus done with methods provided by the data frame class. Geometry is read as-is into the "geometry" column, but no special geometric operations are currently supported.

Installation

# Latest stable version
pip install -U dataiter

# Latest development version
pip install -U git+https://github.com/otsaloma/dataiter

# Numba (optional)
pip install -U numba

Dataiter optionally uses Numba to speed up certain operations. If you have Numba installed and importing it succeeds, Dataiter will use it automatically. It's currently not a hard dependency, so you need to install it separately.

Documentation

https://dataiter.readthedocs.io/

If you're familiar with either dplyr (R) or Pandas (Python), the comparison table in the documentation will give you a quick overview of the differences and similarities.

https://dataiter.readthedocs.io/en/latest/comparison.html

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

dataiter-0.99.tar.gz (49.1 kB view details)

Uploaded Source

Built Distribution

dataiter-0.99-py3-none-any.whl (67.8 kB view details)

Uploaded Python 3

File details

Details for the file dataiter-0.99.tar.gz.

File metadata

  • Download URL: dataiter-0.99.tar.gz
  • Upload date:
  • Size: 49.1 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/5.1.1 CPython/3.12.5

File hashes

Hashes for dataiter-0.99.tar.gz
Algorithm Hash digest
SHA256 395c0bb5ea2f8f53e2f9c3adb929952c731fdb281c61fa61a68acb5e15ee1b75
MD5 8a219da4b7f6580adf3fed677af6eea5
BLAKE2b-256 edecd3d97ab996e845a225961d65366512119e4aaa16f7ca11899098b8b3e53b

See more details on using hashes here.

File details

Details for the file dataiter-0.99-py3-none-any.whl.

File metadata

  • Download URL: dataiter-0.99-py3-none-any.whl
  • Upload date:
  • Size: 67.8 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/5.1.1 CPython/3.12.5

File hashes

Hashes for dataiter-0.99-py3-none-any.whl
Algorithm Hash digest
SHA256 7ffc0f814968ec86507e5ca35f602f3d39841ef9a2c745e762583f9886e352b6
MD5 f1a95bbea544727e2fe86a9a8d56ff66
BLAKE2b-256 00872c58a4909b2b33562b60ae1b14d86e1d207bbd3557a19836ff7e93bb7730

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page