Balanced Multimodal Learning via On-the-fly Gradient Modulation
Results and benchmarks
Balanced Multimodal Learning via On-the-fly Gradient Modulation is the primary contribution described in this paper.
| Task | Dataset | Metric | Value | Source |
|---|---|---|---|---|
| On-the-fly Gradient Modulation | Audio-only | CREMA-D | 52.5 | paper-derived |
| On-the-fly Gradient Modulation | Visual-only | CREMA-D | 41.9 | paper-derived |
| On-the-fly Gradient Modulation | Concatenation | CREMA-D | 51.7 | paper-derived |
| On-the-fly Gradient Modulation | Modality-Drop (audio) | CREMA-D | 54.4 | paper-derived |
| Optimization | SGD | CREMA-D | 51.7 | paper-derived |
Audit each benchmark finding before selecting an implementation path. Evidence refs map to the disclosure below.
Evidence graph: 4 refs, 4 links.
Utility signals: depth 100/100, grounding 95/100, status high.
Implementation
Best maintained implementation now
The repo for "Balanced Multimodal Learning via On-the-fly Gradient Modulation", CVPR 2022 (ORAL)
319 stars · 25 forks · Last push Sep 22, 2025 · MIT license
- License
- CI
- Dependencies
- Docker
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata · Strong overlap with paper title keywords
gewu-lab/ogm-ge_cvpr2022 is the strongest maintained implementation based on ranking signals. License is declared (MIT).
Open gewu-lab/ogm-ge_cvpr2022- No CI workflows detected
- Dependency manifest is missing
- Selected gewu-lab/ogm-ge_cvpr2022 as the strongest maintained implementation for new work.
- Repository activity is within the last 24 months.
Reproduction readiness
Major work
No dependency manifest, manual reconstruction required
- gewu-lab/ogm-ge_cvpr2022 has no requirements.txt, environment.yml, pyproject.toml, or Dockerfile.
- You will need to reverse-engineer dependencies from import statements in the source code.
- Last push was 338 days ago.
Hardware requirements
- Expect multi-day setup/compute for meaningful reproduction based on current guidance.
Hugging Face artifacts
No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.
Models
- gradientai/Llama-3-8B-Instruct-Gradient-1048k
30,422 downloads · 681 likes
- crusoeai/Llama-3-8B-Instruct-Gradient-1048k-GGUF
1,428 downloads · 71 likes
- PrunaAI/Llama-3-8B-Instruct-Gradient-1048k-GGUF-smashed
2,843 downloads · 32 likes
Broaden model search
Datasets
No trustworthy datasets matches right now.
Search datasets on Hugging FaceSpaces
No trustworthy spaces matches right now.
Search spaces on Hugging FaceResearch context
Tasks
On-the-fly Gradient Modulation, Optimization
Methods
None detected
Domains
None detected
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXJump to Paper2Code search queries derived from this paper's research context.
Data includes links from Papers with Code ( CC-BY-SA-4.0 ).