Results and benchmarks
Towards Long-Form Video Understanding focuses on video understanding / reasoning.
Benchmark evidence is limited
Evidence graph: 4 refs, 4 links.
Utility signals: depth 100/100, grounding 95/100, status high.
Implementation
Historical official implementation (not recommended for new builds)
Only a historical official implementation is available
Use with caution for new projects; verify against current tooling and maintained community alternatives.
chaoyuaw/lvu · 87 stars · Last push Mar 4, 2024
md-mohaiminul/ViS4mer is the closest maintained adjacent implementation (Community adoption signal (58 stars)). It is not paper-verified; validate algorithm and evaluation setup against the paper before trusting reported metrics. Community adoption signal: 58 GitHub stars.
Open chaoyuaw/lvu- Adjacent implementations are not paper-verified
- Recommended repository is adjacent and not paper-verified.
- Adjacent implementation match confidence is low.
- No direct maintained implementation is currently verified.
- Only historical official repository was found: chaoyuaw/lvu.
- No maintained paper-verified implementation met reliability thresholds.
Reproduction readiness
Major work
No dependency manifest, manual reconstruction required
- chaoyuaw/lvu has no requirements.txt, environment.yml, pyproject.toml, or Dockerfile.
- You will need to reverse-engineer dependencies from import statements in the source code.
- Last push was 904 days ago.
Hardware requirements
- Expect multi-day setup/compute for meaningful reproduction based on current guidance.
Repositories and ecosystem
Closest related implementations
These are not paper-verified. Use them as reference points when no direct implementation is available.
- md-mohaiminul/ViS4mer Adjacent · Confidence: Low · 58 stars
Community adoption signal (58 stars)
No additional verified repositories beyond the primary recommendation.
Hugging Face artifacts
No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.
Models
- 2264K/vjepa2-qwen3.5-video-understanding
20 downloads · 3 likes
Broaden model search
Datasets
- filwsyl/video_understanding
15 downloads · 1 likes · Updated May 6, 2022
Broaden dataset search
Research context
Tasks
Video understanding / reasoning
Methods
None detected
Domains
Computer vision
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXJump to Paper2Code search queries derived from this paper's research context.
Data includes links from Papers with Code ( CC-BY-SA-4.0 ).