Switch language한국어
Back to the list

GenPrior: Unleashing Text-to-Motion Generative Priors for Zero-Shot Skeleton-based Action Recognition

TL;DR AI

Key summary

2 min read
  1. Researchers introduced GenPrior, a zero-shot skeleton action recognition framework that transfers knowledge from pre-trained text-to-motion models.

  2. It uses dispersion-gated feature fusion to inject structural motion cues into text embeddings, plus generative prototype refinement to better fit unseen actions.

  3. GenPrior tackles the semantic-kinematic gap that weakens text-prototype methods and improves recognition of unseen skeleton actions.

  4. The method outperforms prior approaches on NTU-60, NTU-120, and PKU-MMD in both zero-shot and generalized zero-shot settings.

Read the original