Home Knowledge Base TimeSformer

TimeSformer is the factorized video transformer that separates spatial attention and temporal attention to reduce computation while preserving long-range modeling - by decomposing full 3D attention into two simpler steps, it scales better to longer clips and higher resolutions.

What Is TimeSformer?

Why TimeSformer Matters

Factorization Variants

Space Then Time:

Time Then Space:

Divided Attention Blocks:

How It Works

Step 1:

Step 2:

TimeSformer is a divide-and-conquer transformer design that delivers strong video modeling without full joint attention cost - factorized attention makes long-clip processing significantly more practical.

timesformervideo understanding

Explore 500+ Semiconductor & AI Topics

From EUV lithography to CUDA optimization — search the full knowledge base or chat with our AI assistant.