Marwan Ajem

28
Mar
ViTPose : A simple yet powerful transformer baseline for Human Pose Estimation

ViTPose : A simple yet powerful transformer baseline for Human Pose Estimation

ViTPose: Simple Vision Transformer Baselines for Human Pose EstimationAlthough no specific domain knowledge is considered in the design, plain vision
5 min read
27
Jun
MultiFacetEval

MultiFacetEval

Multifaceted Evaluation to probe LLMs in mastering medical knowledge Yuxuan Zhou et al., June 2024 Large language models (LLMs) have
4 min read
15
Mar
Depth Anything

Depth Anything

Unleashing the Power of Large-Scale Unlabeled Data Lihe Yang et al., January 2024 Monocular Depth Estimation (MDE) The goal is
4 min read
07
Dec
Skeleton project with DWPose

Skeleton project with DWPose

Goal The goal of this project is to predict a user's keypoints in real time and guide them
4 min read