Skip to main content

Supplementary library for pandas that processes dataframes derived from CSV files.

Project description

datasoap

What is it?

datasoap is a supplementary library for pandas that processes dataframes derived from CSV files. The module checks cell data for correct numerical formatting and converts mismatched data to the correct data type (ex. str > float64).

Main Features

  • Strips unnecessary characters from numerical data fields in pandas dataframes to ensure consistent data formatting
  • Provides before and after representations of dataframes to allow for comparison

Repository

Source code is hosted on: github.com/snake-fingers/data-soap

Dependencies

pandas - Python package that provides fast, flexible, and expressive data structures designed to make working with “relational” or “labeled” data both easy and intuitive.

Installation

poetry add datasoap

Documentation

Documentation to come.

Background

datasoap originated from a Code Fellows 401 Python midterm project. The project team includes Alex Angelico, Grace Choi, Robert Carter, Mason Fryberger, and Jae Choi. After working with a few painful datasets using, we wanted to create a library that allows users to more efficiently manipulate clean datasets extracted from CSVs that may have inconsistent formatting.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

datasoap-1.0.2.tar.gz (4.8 kB view details)

Uploaded Source

Built Distribution

datasoap-1.0.2-py3-none-any.whl (5.8 kB view details)

Uploaded Python 3

File details

Details for the file datasoap-1.0.2.tar.gz.

File metadata

  • Download URL: datasoap-1.0.2.tar.gz
  • Upload date:
  • Size: 4.8 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/1.1.4 CPython/3.9.0 Linux/4.19.128-microsoft-standard

File hashes

Hashes for datasoap-1.0.2.tar.gz
Algorithm Hash digest
SHA256 6fbb8426cadbbd4de97ad7eb0699b1e8fb5bb1b450a43798c2c85f9d7369d179
MD5 98a6fca37be6ed64238bfa63019c4354
BLAKE2b-256 6187eebbab450675f1ffc6e45d1ed562692cdc59564816ef00dde091feddec97

See more details on using hashes here.

File details

Details for the file datasoap-1.0.2-py3-none-any.whl.

File metadata

  • Download URL: datasoap-1.0.2-py3-none-any.whl
  • Upload date:
  • Size: 5.8 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/1.1.4 CPython/3.9.0 Linux/4.19.128-microsoft-standard

File hashes

Hashes for datasoap-1.0.2-py3-none-any.whl
Algorithm Hash digest
SHA256 b50c10cd7feb16f440288cc35d355e151e8d7b32b77e4ed3b7fe2e2d0a3fac2a
MD5 38844ca5195e939d3d9e8f436ed24ca7
BLAKE2b-256 67c30995baaf7b5c183867c54fd834f1a8f3e9a23050ae5f9086f24d0f99b0c0

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page