AIToday
Large Language ModelsVideo GenerationAI Safety & AlignmentarXiv cs.CVPublished: Mar 25, 2026, 13:13 JST1 min read

New TrajLoom framework predicts future motion in videos by modeling dense point trajectories using flow matching and variational autoencoders.

New TrajLoom framework predicts future motion in videos by modeling dense point trajectories using flow matching and variational autoencoders.

3 Key Points

  1. TrajLoom uses Grid-Anchor Offset Encoding to reduce location bias by representing points as offsets from pixel-center anchors

  2. TrajLoom-VAE learns compact spatiotemporal latent space for dense trajectories with masked reconstruction and consistency regularization

  3. TrajLoom-Flow generates future trajectories via flow matching with boundary cues and K-step fine-tuning for stable sampling

  4. Framework predicts both future trajectories and visibility from observed video context for improved motion representation

  5. Introduces TrajLoomBench, a unified benchmark dataset combining real and synthetic video data for evaluation

Ask the AI about this article →

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Nvidia revives Rubin CPX chip with major redesignYahoo Finance AI · 2h ago
  • AI advice followed by 79%, but well-being unchangedITmedia AI+ · 5h ago
  • Enterprises face agent governance gapSiliconANGLE AI · 8h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleNew evaluation framework reveals that only 35% of LLM responses on sexual and reproductive health in Nepali meet quality standards, highlighting gaps in low-resource language support for sensitive health topics.