Back to Academic/Research
Open WeightsSpecializedvideo3dUpdated June 29, 2026

VidiHand

Model Overview

VidiHand is a undisclosed-parameter model developed by Academic/Research. Released on 2026-06-29.


📊 Quick Specs

Specification Table
SpecificationValue
Parametersundisclosed
Taskother
Modalityvideo, 3d
LicenseOther/Custom
Typeopen-weights

✨ Key Features

  • Detector-Free Full-Frame Processing without localized cropping or hand detectors
  • Leverages Internet-Scale Pretrained Video Diffusion Models (Wan2.1-VACE)
  • Extreme Robustness to Heavy Hand-Object and Hand-Hand Occlusions
  • Eliminates Test-Time Optimization (TTO) and post-hoc temporal infilling
  • State-of-the-Art Temporal Smoothness (jitter down to 3.18 mm/frame) and Pose Accuracy (21.668 mm MPJPE-p on ARCTIC)

🔗 Resources


📜 License & Access

Other/Custom — See repository for specific license details.

Key Features

Detector-Free Full-Frame Processing without localized cropping or hand detectors

Feature 01

Leverages Internet-Scale Pretrained Video Diffusion Models (Wan2.1-VACE)

Feature 02

Extreme Robustness to Heavy Hand-Object and Hand-Hand Occlusions

Feature 03

Eliminates Test-Time Optimization (TTO) and post-hoc temporal infilling

Feature 04

State-of-the-Art Temporal Smoothness (jitter down to 3.18 mm/frame) and Pose Accuracy (21.668 mm MPJPE-p on ARCTIC)

Feature 05

You might also want to compare

Verified Sources

Tags

research-preview4d-hand-trackingvideo-diffusionntusjtuegocentric-vision

Model Specs

open-weights

Parameters

Undisclosed

Context Window

undisclosed

License

Other/Custom

Deployment

self-hostable

Resources & Links

Curator Notes

Partially enriched via migration on 2026-07-25. Manual review recommended.

Compare Specs

Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.

Compare Model