Efficient 3-D Human Pose Estimation: A Synergy of Classical Computer Vision and Deep Learning

This repository is the implementation of a research project for the 2023 Fall Semester Computer Vision class by Team 16, based on the H3WB repository.

Install Dependencies

We conducted all experiments with Python 3.9 with dependencies listed in requirements.txt.

conda create -n [env name] python=3.9
conda activate [env name]
pip install -r requirements.txt

Data Preparation and Preprocessing

Run data.sh in a preferred directory (takes several GB and takes 30~45 minutes.)
Put RGBto3D_train.json and RGBto3D_test_img.json to ./data/h3wb/annotations
Run python resize.py to resize images to 224x224.
Run python split_dataset.py to split the data into pre-defined train, dev, and test sets.

For further details, refer to ./Readme.txt.

Training

We implemented our models in models/ClassicalModel.py and models/CombinedModel.py

output_path=/path/to/model/checkpoint
model_name=resnet50 
# one of {"resnet50", "resnet18"} for baseline models
# one of {"resnet50_4_with_sobel", "resnet18_4_with_sobel"} for sobel operator models

python train.py \
    --learning_rate 1e-5 --batch_size 16 --num_epochs 20 \
    --model_name ${model_name} --use_pretrained \
    --save_path ${output_dir}

Evaluation

checkpoint_path=/path/to/model/checkpoint
python evaluate.py \
    --model_path ${checkpoint_path} --model_name ${model_name}

Name		Name	Last commit message	Last commit date
Latest commit History 182 Commits
datasets		datasets
imgs		imgs
models		models
utils		utils
.gitignore		.gitignore
Dataset.md		Dataset.md
LICENSE.md		LICENSE.md
README.md		README.md
Readme.txt		Readme.txt
benchmark.md		benchmark.md
data.sh		data.sh
evaluate.py		evaluate.py
file_not_found_list.txt		file_not_found_list.txt
h3wb.py		h3wb.py
rename.py		rename.py
requirements.txt		requirements.txt
resize.py		resize.py
split_dataset.py		split_dataset.py
test_leaderboard.py		test_leaderboard.py
train.py		train.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Repository files navigation

Efficient 3-D Human Pose Estimation: A Synergy of Classical Computer Vision and Deep Learning

Install Dependencies

Data Preparation and Preprocessing

Training

Evaluation

About

Uh oh!

Releases

Packages

Uh oh!

Contributors 2

Uh oh!

Languages

License

willystumblr/fa23-cv

Folders and files

Latest commit

History

Repository files navigation

Efficient 3-D Human Pose Estimation: A Synergy of Classical Computer Vision and Deep Learning

Install Dependencies

Data Preparation and Preprocessing

Training

Evaluation

About

Resources

License

Uh oh!

Stars

Watchers

Forks

Releases

Packages 0

Uh oh!

Contributors 2

Uh oh!

Languages

Packages