The engine Every edition RSS Refreshed every 30 min 24m ago
PRISM The internet, refracted.

Home Research

Unified Video Dense Prediction from Disjoint Data

Signal strength 17/100

Scene understanding requires simultaneous prediction about geometry, appearance, and semantics. However, existing task-specific annotations are fragmented across incompatible, domain-specific datasets. Current unified systems circumvent this by restricting…

PRISM indexes and ranks — it never republishes. The full piece lives with its author on arxiv.org.

Read on arxiv.org

Same wavelength

Stories the engine considers adjacent to this one.

Research arXiv

GraphVid: Interactive Graph-Controllable Video Generation

Controllable video generation remains challenging due to the difficulty of specifying precise multi-object interactions using text prompts or motion-control inputs that primarily constrain pixel movement. In practice, trajectory-based control often requires…

1 min 0 views

Research arXiv

Self-Supervised Learning of Structured Dynamics from Videos

Understanding motion in video is a fundamental challenge for visual learning, as frame-to-frame change entangles two sources of dynamics: camera motion and object motion. This decomposition has remained underexplored in representation learning, partly because…

1 min 0 views