PifPaf: Composite Fields for Human Pose Estimation

Project description

openpifpaf

Continuously tested on Linux, MacOS and Windows:
CVPR 2019 paper

PifPaf: Composite Fields for Human Pose Estimation

We propose a new bottom-up method for multi-person 2D human pose estimation that is particularly well suited for urban mobility such as self-driving cars and delivery robots. The new method, PifPaf, uses a Part Intensity Field (PIF) to localize body parts and a Part Association Field (PAF) to associate body parts with each other to form full human poses. Our method outperforms previous methods at low resolution and in crowded, cluttered and occluded scenes thanks to (i) our new composite field PAF encoding fine-grained information and (ii) the choice of Laplace loss for regressions which incorporates a notion of uncertainty. Our architecture is based on a fully convolutional, single-shot, box-free design. We perform on par with the existing state-of-the-art bottom-up method on the standard COCO keypoint task and produce state-of-the-art results on a modified COCO keypoint task for the transportation domain.

Demo

example image with overlaid pose predictions

Image credit: "Learning to surf" by fotologic which is licensed under CC-BY-2.0.
Created with python3 -m openpifpaf.predict docs/coco/000000081988.jpg --show --image-output --json-output which also produces json output.

More demos:

openpifpafwebdemo project (best performance)
OpenPifPaf running in your browser: https://vita-epfl.github.io/openpifpafwebdemo/ (experimental)
the openpifpaf.video command (requires OpenCV)
Google Colab demo

Install

Python 3 is required. Python 2 is not supported. Do not clone this repository and make sure there is no folder named openpifpaf in your current directory.

pip3 install openpifpaf

For a live demo, we recommend to try the openpifpafwebdemo project. Alternatively, openpifpaf.video (requires OpenCV) provides a live demo as well.

For development of the openpifpaf source code itself, you need to clone this repository and then:

pip3 install numpy cython
pip3 install --editable '.[train,test]'

The last command installs the Python package in the current directory (signified by the dot) with the optional dependencies needed for training and testing. If you modify functional.pyx, run this last command again which recompiles the static code.

Interfaces

python3 -m openpifpaf.predict --help: help screen
python3 -m openpifpaf.video --help: help screen
python3 -m openpifpaf.train --help: help screen
python3 -m openpifpaf.eval_coco --help: help screen
python3 -m openpifpaf.logs --help: help screen

Tools to work with models:

python3 -m openpifpaf.migrate --help: help screen
python3 -m openpifpaf.export_onnx --help: help screen

Pre-trained Models

Performance metrics with version 0.11 on the COCO val set obtained with a GTX1080Ti:

Backbone	AP	APᴹ	APᴸ	t_{total} [ms]	t_{dec} [ms]
shufflenetv2k16w	67.1	62.0	75.3	54	25
shufflenetv2k30w	71.1	65.9	79.1	94	22

Command to reproduce this table: python -m openpifpaf.benchmark --backbones shufflenetv2k16w shufflenetv2k30w.

Pretrained model files are shared in the openpifpaf-torchhub repository and linked from the backbone names in the table above. The pretrained models are downloaded automatically when using the command line option --checkpoint backbonenameasintableabove.

For comparison, old v0.10:

Backbone	AP	APᴹ	APᴸ	t_{total} [ms]	t_{dec} [ms]
shufflenetv2x2 v0.10	60.4	55.5	67.8	56	33
resnet50 v0.10	64.4	61.1	69.9	76	32
resnet101 v0.10	67.8	63.6	74.3	97	28

Train

See datasets for setup instructions.

The exact training command that was used for a model is in the first line of the training log file.

ShuffleNet models are trained without ImageNet pretraining:

time CUDA_VISIBLE_DEVICES=0,1 python3 -m openpifpaf.train \
  --lr=0.1 \
  --momentum=0.9 \
  --epochs=150 \
  --lr-warm-up-epochs=1 \
  --lr-decay 120 \
  --lr-decay-epochs=20 \
  --lr-decay-factor=0.1 \
  --batch-size=32 \
  --square-edge=385 \
  --lambdas 1 1 0.2   1 1 1 0.2 0.2    1 1 1 0.2 0.2 \
  --auto-tune-mtl \
  --weight-decay=1e-5 \
  --update-batchnorm-runningstatistics \
  --ema=0.01 \
  --basenet=shufflenetv2k16w \
  --headnets cif caf caf25

# for improved performance, take the epoch150 checkpoint and train with
# extended-scale and 10% orientation invariance:
time CUDA_VISIBLE_DEVICES=0,1 python3 -m openpifpaf.train \
  --lr=0.05 \
  --momentum=0.9 \
  --epochs=250 \
  --lr-warm-up-epochs=1 \
  --lr-decay 220 \
  --lr-decay-epochs=30 \
  --lr-decay-factor=0.01 \
  --batch-size=32 \
  --square-edge=385 \
  --lambdas 1 1 0.2   1 1 1 0.2 0.2    1 1 1 0.2 0.2 \
  --auto-tune-mtl \
  --weight-decay=1e-5 \
  --update-batchnorm-runningstatistics \
  --ema=0.01 \
  --checkpoint outputs/shufflenetv2k16w-200504-145520-cif-caf-caf25-d05e5520.pkl --extended-scale --orientation-invariant=0.1

You can refine an existing model with the --checkpoint option.

To visualize logs:

python3 -m openpifpaf.logs \
  outputs/resnet50block5-pif-paf-edge401-190424-122009.pkl.log \
  outputs/resnet101block5-pif-paf-edge401-190412-151013.pkl.log \
  outputs/resnet152block5-pif-paf-edge401-190412-121848.pkl.log

To produce evaluation metrics every five epochs and check the directory for new checkpoints every 5 minutes:

while true; do \
  CUDA_VISIBLE_DEVICES=0 find outputs/ -name "shufflenetv2k16w-200504-145520-cif-caf-caf25.pkl.epoch??[0,5]" -exec \
    python3 -m openpifpaf.eval_coco --checkpoint {} -n 500 --long-edge=641 --skip-existing \; \
  ; \
  sleep 300; \
done

Person Skeletons

COCO / kinematic tree / dense:

Created with python3 -m openpifpaf.datasets.constants.

Video

Requires OpenCV.

python3 -m openpifpaf.video --checkpoint shufflenetv2k16w myvideotoprocess.mp4 --video-output --json-output

Replace myvideotoprocess.mp4 with 0 for webcam0 or other OpenCV compatible sources.

Documentation Pages

Related Projects

monoloco: "Monocular 3D Pedestrian Localization and Uncertainty Estimation" which uses OpenPifPaf for poses.
openpifpafwebdemo: web front-end.

Citation

@InProceedings{kreiss2019pifpaf,
  author = {Kreiss, Sven and Bertoni, Lorenzo and Alahi, Alexandre},
  title = {PifPaf: Composite Fields for Human Pose Estimation},
  booktitle = {The IEEE Conference on Computer Vision and Pattern Recognition (CVPR)},
  month = {June},
  year = {2019}
}

Project details

Release history Release notifications | RSS feed

0.13.11

Feb 5, 2023

0.13.10

Feb 1, 2023

0.13.9

Jan 29, 2023

0.13.8

Jan 15, 2023

0.13.7

Nov 11, 2022

0.13.6

Nov 2, 2022

0.13.5

Aug 31, 2022

0.13.4

Jun 1, 2022

0.13.3

Mar 22, 2022

0.13.2

Mar 21, 2022

0.13.1

Dec 23, 2021

0.13.0

Sep 26, 2021

0.12.14

Sep 15, 2021

0.12.13

Jul 22, 2021

0.12.12

Jul 7, 2021

0.12.11

Jun 11, 2021

0.12.10

May 19, 2021

0.12.9

Apr 21, 2021

0.12.8

Apr 11, 2021

0.12.7

Apr 9, 2021

0.12.6

Mar 24, 2021

0.12.5

Mar 4, 2021

0.12.4

Mar 2, 2021

0.12.2

Feb 24, 2021

0.12.1

Feb 23, 2021

0.12

Feb 23, 2021

0.12b4 pre-release

Feb 14, 2021

0.12b2 pre-release

Feb 9, 2021

0.12b1 pre-release

Dec 17, 2020

0.12a6 pre-release

Nov 6, 2020

0.12a4 pre-release

Oct 15, 2020

0.12a3 pre-release

Oct 15, 2020

0.12a2 pre-release

Oct 15, 2020

0.12a1 pre-release

Oct 14, 2020

0.11.9

Aug 25, 2020

0.11.8

Aug 19, 2020

0.11.7

Jul 14, 2020

0.11.6

Jun 15, 2020

0.11.5

Jun 4, 2020

0.11.4

Jun 2, 2020

0.11.3

Jun 1, 2020

0.11.2

May 29, 2020

0.11.1

May 15, 2020

This version

0.11.0

May 13, 2020

0.10.1

Dec 9, 2019

0.10.0

Oct 21, 2019

0.9.0

Jul 30, 2019

0.8.0

Jul 8, 2019

0.7.0

Jun 6, 2019

0.6.3

May 28, 2019

0.6.2

May 23, 2019

0.6.1

May 13, 2019

0.6.0

May 7, 2019

0.5.2

May 7, 2019

0.5.1

May 1, 2019

0.5.0

Apr 26, 2019

0.4.0

Apr 8, 2019

0.3.2

Apr 5, 2019

0.3.1

Apr 3, 2019

0.3.0

Apr 3, 2019

0.2.3

Mar 19, 2019

0.2.2

Mar 17, 2019

0.2.1

Mar 16, 2019

0.2.0

Mar 15, 2019

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

openpifpaf-0.11.0.tar.gz (201.8 kB view details)

Uploaded May 13, 2020 Source

File details

Details for the file openpifpaf-0.11.0.tar.gz.

File metadata

Download URL: openpifpaf-0.11.0.tar.gz
Upload date: May 13, 2020
Size: 201.8 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/1.13.0 pkginfo/1.5.0.1 requests/2.21.0 setuptools/40.6.2 requests-toolbelt/0.9.1 tqdm/4.31.1 CPython/3.7.4

File hashes

Hashes for openpifpaf-0.11.0.tar.gz
Algorithm	Hash digest
SHA256	`510a4e643e9461a28103606af4b3976756a868ee2137b7d495e9d787836856ec`
MD5	`dc59587f4e057fa07ed0e569e50415a5`
BLAKE2b-256	`88d4bc22584864a101ce6d806a58bd512325549e3268f251a0bb6f018f5f4b23`

See more details on using hashes here.

openpifpaf 0.11.0

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Project description

openpifpaf

Demo

Install

Interfaces

Pre-trained Models

Train

Person Skeletons

Video

Documentation Pages

Related Projects

Citation

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Release history Release notifications | RSS feed

Download files

Source Distribution

File details

File metadata

File hashes