Skip to main content

This is a library that implements methods to aggregate local features (mainly for multimedia) into a single global feature that can be used easily with any classifier.

Dependencies

The library depends on scikit-learn and all the feature aggregation methods extend the scikit-learn BaseEstimator class.

Example

import numpy as np
from feature_aggregation import BagOfWords, FisherVectors

X = np.random.rand(1000, 2)
bow = BagOfWords(10)
fv = FisherVectors(10)

bow.fit(X)
fv.fit(X)

G1 = bow.transform(np.random.rand(10, 100, 2))
G2 = fv.transform([
    np.random.rand(int(np.random.rand()*100), 2) for _ in range(10)
])

A more complex example using OpenCV to extract dense SIFT and then transform them using Bag Of Words and train an SVM with chi square additive kernel.

import numpy as np
import cv2
from sklearn.datasets import fetch_olivetti_faces
from sklearn.kernel_approximation import AdditiveChi2Sampler
from sklearn.metrics import classification_report
from sklearn.pipeline import Pipeline
from sklearn.svm import LinearSVC

from feature_aggregation import BagOfWords

def sift(*args, **kwargs):
    try:
        return cv2.xfeatures2d.SIFT_create(*args, **kwargs)
    except:
        return cv2.SIFT()

def dsift(img, step=5):
    keypoints = [
        cv2.KeyPoint(x, y, step)
        for y in range(0, img.shape[0], step)
        for x in range(0, img.shape[1], step)
    ]
    features = sift().compute(img, keypoints)[1]
    features /= features.sum(axis=1).reshape(-1, 1)
    return features

# Generate dense SIFT features
faces = fetch_olivetti_faces()
features = [
    dsift((x.reshape(64, 64, 1)*255).astype(np.uint8))
    for x in faces.data
]

# Aggregate those features with bag of words using online training
bow = BagOfWords(100)
for i in range(2):
    for j in range(0, len(features), 10):
        bow.partial_fit(features[j:j+10])
faces_bow = bow.transform(features)

# Split in training and test set
train = np.arange(len(features))
np.random.shuffle(train)
test = train[200:]
train = train[:200]

# Train and evaluate
svm = Pipeline([("chi2", AdditiveChi2Sampler()), ("svm", LinearSVC(C=10))])
svm.fit(faces_bow[train], faces.target[train])
print(classification_report(faces.target[test], svm.predict(faces_bow[test])))

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distributions

No source distribution files available for this release.See tutorial on generating distribution archives.

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

feature_aggregation-0.1.1-py2.py3-none-any.whl (10.5 kB view details)

Uploaded Python 2Python 3

File details

Details for the file feature_aggregation-0.1.1-py2.py3-none-any.whl.

File metadata

File hashes

Hashes for feature_aggregation-0.1.1-py2.py3-none-any.whl
Algorithm Hash digest
SHA256 c866576e044bc2eb2dfd4e05cfe2781a393d8a183c4602d197a78e4d6c12abaf
MD5 66adf5ec3cb3f4dba0a5dcf4ddcff5be
BLAKE2b-256 20f0dd56e251f62e5a0f8a3d04496de6847059a44deb6e51a566091104c7f7a6

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page