Search: プロジェクトページとコードが公開されています。 - ai.jp.net

Paper #Computer Vision 🔬 ResearchAnalyzed: Jan 3, 2026 16:27

Video Gaussian Masked Autoencoders for Video Tracking

Published:Dec 27, 2025 06:16

•

1 min read

•

ArXiv

Analysis

This paper introduces a novel self-supervised approach, Video-GMAE, for video representation learning. The core idea is to represent a video as a set of 3D Gaussian splats that move over time. This inductive bias allows the model to learn meaningful representations and achieve impressive zero-shot tracking performance. The significant performance gains on Kinetics and Kubric datasets highlight the effectiveness of the proposed method.

Key Takeaways

•Proposes Video-GMAE, a self-supervised approach for video representation learning.
•Represents videos as moving 3D Gaussian splats.
•Achieves strong zero-shot tracking performance.
•Significantly improves performance on Kinetics and Kubric datasets.
•Project page and code are publicly available.

Reference

“Mapping the trajectory of the learnt Gaussians onto the image plane gives zero-shot tracking performance comparable to state-of-the-art.”

Permalink ArXiv

Video Gaussian Masked Autoencoders for Video Tracking

Analysis

Key Takeaways

📬 Get AI News Delivered

Browse by Category

Trending Topics

📬 Get AI News Delivered

Browse by Category

Trending Topics