Skip to main content
# torchvision-enhance

torchvision-enhance is used to enhance the offical PyTorch vision library torchvision. Here is the enhanced parts:
- support multi-channel(> 4 channels, e.g. 8 channels) images
- support 16-bit TIF file
- more easier to semantic segmentation transform



## Support transforms
- RandomFlip
- RandomVFlip
- RandomHFlip
- RandomRotate
- RandomShift
- RandomCrop
- CenterCrop
- Resize
- Pad
- GaussianBlur
- PieceTransform
- Lambda
- ToTensor
- Normalize

## Install
```
pip install torchvision-enhance
```

or install from the source

```
git clone
pip install -r requirements.txt
python setup.py install
```
## Dependencies
- numpy
- scipy
- Pillow
- PyTorch
- opencv
- scikit-image

## Usage
For more useage, check out the [example-classification.py](./test/example-classification.py) and [example-segmentation.py](./test/example-segmentation.py)

``` python
from torchvision_x.datasets import image_loader
from torchvision_x.transforms import transforms_seg,functional

transform = transforms_seg.SegCompose([
# transforms_seg.SegFlip(),
transforms_seg.SegVFlip(),
# transforms_seg.SegHFlip(),
# transforms_seg.SegRandomFlip(),
# transforms_seg.SegRandomRotate(90),
# transforms_seg.SegRandomShift(40),
# transforms_seg.SegRandomCrop((256,256)),
# transforms_seg.SegCenterCrop(224),
# transforms_seg.SegResize(224),
# transforms_seg.SegPad(20),
# transforms_seg.SegNoise(dtype='uint16', var=0.001), #TODO
# transforms_seg.SegGaussianBlur(sigma=2, dtype='uint8', multichannel=False),
# transforms_seg.SegPieceTransform(),
# transforms_seg.SegLambda(lambda x: functional.to_tensor(x))
transforms_seg.SegToTensor(),
transforms_seg.SegNormalize((0.5,0.5,0.5),(0.5,0.5,0.5)),
])

trainset = image_loader.SemanticSegmentationLoader(
rootdir='sample-data/', lstpath='sample-data/segmentation_jpg.lst',
filetype='jpg', transform=transform,
)
trainloader = DataLoader(dataset=trainset,batch_size=batch_size,shuffle=False)

for step, (inputs, targets) in enumerate(trainloader):
print('batch: {} ........'.format(step))
print(type(inputs), inputs.shape)
print(type(targets), targets.shape)
```

## TODO
- Noise

Metadata

Release files for torchvision-enhance 0.1.3

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for torchvision-enhance 0.1.3
File Size Uploaded
torchvision-enhance-0.1.3.tar.gz 13.4 kB Details

Release files / torchvision-enhance-0.1.3.tar.gz

Download URL torchvision-enhance-0.1.3.tar.gz
Size 13.4 kB
Tags Source
SHA-256 checksum
How to use checksums
3f03d638216b33d299d4238fb8f9a5c9968373c33c651e9f8620fd1bf0980eee
BLAKE2b-256 checksum
How to use checksums
a4ae7e1ac9784927b4ae5174c6f6533acacfa964982c30f77cd379ebfbaa7fd6
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No

Release history Release notifications | RSS feed

This release

0.1.3 This release

1 release file

0.1.2

1 release file

0.1.1

1 release file

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page