Skip to main content

Scalable Objects Persistence for Python.

Project description

What is SOP?

Scalable Objects Persistence (SOP) is a raw storage engine that bakes together a set of storage related features & algorithms in order to provide the most efficient & reliable (ACID attributes of transactions) technique (known) of storage management and rich search, as it brings to the application, the raw muscle of "raw storage", direct IO communications w/ disk drives. In a code library form factor.

SOP supported Hardware/OS

SOP supports popular architectures & Operating Systems such as Linux, Darwin & Microsoft Windows, in both ARM64 & AMD64 architectures. For Windows, only AMD64 is supported since it is the only architecture Windows is available in.

SOP Dependencies

  • Redis, you will need to have one of the recent or latest version of Redis for use in SOP caching.
  • More than one Disk Drives(recommended is around four or more, for replication) with plenty of drive space available, for storage management. Example:
    /disk1
    /disk2
    /disk3
    /disk4

SOP for Python package

Following steps outlines how to use the Scalable Objects Persistence code library for Python:

  • Install the package using: pip install sop-python-beta-3
  • Follow standard Python package import and start coding to use the SOP for Python code library for data management. Import the sop package in your python code file.
  • Specify Home base folders where Store info & Registry data files will be stored.
  • Specify Erasure Coding (EC) configuration details which will be used by SOP's EC based replication.
  • Create a transaction
  • Begin a transaction
  • Create a new B-tree, or Open an existing B-tree
  • Manage data, do some CRUD operations
  • Commit the transaction

Below is an example code black for illustrating the above steps. For other SOP B-tree examples, you can checkout the code in the unit tests test_btree.py & test_btree_idx.py files that comes w/ the SOP package you downloaded from pypi.

import sop.transaction
import sop.btree
import sop.context

stores_folders = ("/disk1", "/disk2")
ec = {
    # Erasure Config default entry(key="") will allow different B-tree(tables) to share same EC structure.
    "": transaction.ErasureCodingConfig(
        2,  # two data shards
        2,  # two parity shards
        (
            # 4 disk drives paths
            "/disk1",
            "/disk2",
            "/disk3",
            "/disk4",
        ),
        # False means Auto repair of failed reads from (shards') disk drive will not get repaired.
        False,
    )
}

# Transaction Options (to).
to = transaction.TransationOptions(
    transaction.TransactionMode.ForWriting.value,
    # commit timeout of 5mins
    5,
    # Min Registry hash mod value is 250, you can specify higher value like 1000. A 250 hashmod
    # will use 1MB sized file segments. Good for demo, but for Prod, perhaps a bigger value is better.
    transaction.MIN_HASH_MOD_VALUE,
    # Store info & Registry home base folders. Array of strings of two elements, one for Active & another, for passive folder.
    stores_folders,
    # Erasure Coding config as shown above.
    ec,
)

# Context object.
ctx = context.Context()

# initialize/open SOP global Redis connection
ro = RedisOptions()
Redis.open_connection(ro)

t = transaction.Transaction(ctx, to)
t.begin()

cache = btree.CacheConfig()

# "barstoreec" is new b-tree name, 2nd parameter set to True specifies B-tree Key field to be native data type
bo = btree.BtreeOptions("barstoreec", True, cache_config=cache)
bo.set_value_data_size(btree.ValueDataSize.Small)

# create the new "barstoreec" b-tree store.
b3 = btree.Btree.new(ctx, bo, t)

# Since we've specified Native data type = True in BtreeOptions, we can use "integer" values as Key.
l = [
    btree.Item(1, "foo"),
]

# Add Item to the B-tree,
b3.add(ctx, l)

# Commit the transaction to finalize the new B-tree (store) change.
t.commit(ctx)

SOP in Github

SOP open source project (MIT license) is in github. You can checkout the "...sop/jsondb/" package which contains the Go code enabling general purpose JSON data management & the Python wrapper, coding guideline of which, was described above.

Please feel free to join the SOP project if you have the bandwidth and participate/co-own/lead! the project engineering.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

sop_python_beta_3-3.0.3.tar.gz (22.4 MB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

sop_python_beta_3-3.0.3-py3-none-any.whl (22.7 MB view details)

Uploaded Python 3

File details

Details for the file sop_python_beta_3-3.0.3.tar.gz.

File metadata

  • Download URL: sop_python_beta_3-3.0.3.tar.gz
  • Upload date:
  • Size: 22.4 MB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.1.0 CPython/3.13.5

File hashes

Hashes for sop_python_beta_3-3.0.3.tar.gz
Algorithm Hash digest
SHA256 d3aa4decc56ceaee06100b9ac554ab5e1f7352b73c1bb25127d5198adf6e8478
MD5 25578e1536efc5d5ee19daf81d2d87fd
BLAKE2b-256 eb12ef0bc32dca91e9a4c9f66b52c035462daaf701e017f73950ea6afee2828a

See more details on using hashes here.

File details

Details for the file sop_python_beta_3-3.0.3-py3-none-any.whl.

File metadata

File hashes

Hashes for sop_python_beta_3-3.0.3-py3-none-any.whl
Algorithm Hash digest
SHA256 1a038440f9603974ab39894731d93d5dfb199dccda7b8dc8805e292f3ba85930
MD5 bb3c44b783798438a6fd81f51eef3303
BLAKE2b-256 bad98964bea003dee33e5f07a4bc7ec38fc6ab263436056cd37c4c2e55967d80

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page