Video-Based Markerless Motion Capture for Clinical and Rehabilitation Biomechanics: A PRISMA-ScR Scoping Review of Validated Architectures, Clinical Readiness, and Emerging Methods
cs.CV, physics.med-ph
Submitted: 2026-09-16
Updated: 2026-09-16
Terminology
Sources
- BlazePose: On-device Real-time Body Pose tracking
- BioPose: Biomechanically-accurate 3D Pose Estimation from Monocular Videos
- OpenCap Monocular: 3D Human Kinematics and Musculoskeletal Dynamics from a Single Smartphone Video
- TRAM: Global Trajectory and Motion of 3D Humans from in-the-wild Videos
- RapidPoseTriangulation: Multi-view Multi-person Whole-body Human Pose Triangulation in a Millisecond
- DETRPose: Real-Time End-to-End Multi-Person Pose Estimation via Modified Transformer Decoder and Novel Denoising Keypoints
- FastHMR: Accelerating Human Mesh Recovery via Token and Layer Merging with Diffusion Decoding
- AJAHR: Amputated Joint Aware 3D Human Mesh Recovery
- Mamba-Driven Topology Fusion for Monocular 3D Human Pose Estimation
- DreamPose3D: Hallucinative Diffusion with Prompt Learning for 3D Human Pose Estimation
- HyperDiff: Hypergraph Guided Diffusion Model for 3D Human Pose Estimation
- Validation of Human Pose Estimation and Human Mesh Recovery for Extracting Clinically Relevant Motion Data from Videos
- Biomechanically Accurate Gait Analysis: A 3d Human Reconstruction Framework for Markerless Estimation of Gait Parameters
- Physics-Informed Learning for Human Whole-Body Kinematics Prediction via Sparse IMUs
- Efficient Vision Transformer for Human Pose Estimation via Patch Selection
- Beyond Static Frames: Temporal Aggregate-and-Restore Vision Transformer for Human Pose Estimation
- BiomechGPT: Extending Motion-Language Models to Clinical Motion Understanding
- BiomechAgent: AI-Assisted Biomechanical Analysis Through Code-Generating Agents
- GroundLink: A Dataset Unifying Human Body Movement and Ground Reaction Dynamics
- RTMPose: Real-Time Multi-Person Pose Estimation based on MMPose
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models