DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models
Results and benchmarks
DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models is the primary contribution described in this paper.
Benchmark evidence is limited
Evidence graph: 4 refs, 4 links.
Utility signals: depth 65/100, grounding 85/100, status medium.
Implementation
Best maintained implementation now
DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models (ICCV 2023)
142 stars · 7 forks · Last push Jun 10, 2025 · MIT license
- License
- CI
- Dependencies
- Docker
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata · Strong overlap with paper title keywords
j-min/dalleval is the strongest maintained implementation based on ranking signals. License is declared (MIT).
Open j-min/dalleval- No CI workflows detected
- Dependency manifest is missing
- Selected j-min/dalleval as the strongest maintained implementation for new work.
- Repository activity is within the last 24 months.
Compare implementation paths
Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Limited
- Stars
- 142
- Last push
- Jun 10, 2025 (441d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No CI pipeline detected
- No tagged releases
- Maintenance
- Stale
- Confidence
- Low
- Reproducibility
- Moderate
- Stars
- 1,030
- Last push
- Jan 3, 2024 (965d)
Partial overlap with paper title keywords · Community adoption signal (1030 stars)
- No push in 12+ months
- No CI pipeline detected
- No tagged releases
- Maintenance
- Stale
- Confidence
- Low
- Reproducibility
- Limited
- Stars
- 4
- Last push
- Nov 15, 2022 (1379d)
Matched via arXiv identifier search · Strong overlap with paper title keywords
- No push in 12+ months
- No CI pipeline detected
- Dependency manifest missing
Reproduction readiness
Major work
No dependency manifest, manual reconstruction required
- j-min/dalleval has no requirements.txt, environment.yml, pyproject.toml, or Dockerfile.
- You will need to reverse-engineer dependencies from import statements in the source code.
- Last push was 441 days ago.
Hardware requirements
- Expect multi-day setup/compute for meaningful reproduction based on current guidance.
Validation caveat
Repositories and ecosystem
No additional verified repositories beyond the primary recommendation.
These repositories had low-confidence matching signals and are hidden by default.
- kakaobrain/rq-vae-transformer
Confidence: Low · 1,030 stars
- aszala/PaintSkills-Simulator
Confidence: Low · 4 stars
Hugging Face artifacts
No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.
Models
- Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-GGUF
59,458 downloads · 679 likes
- Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-v2-GGUF
26,062 downloads · 616 likes
- nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4
1,570,072 downloads · 178 likes
Broaden model search
Datasets
- FreedomIntelligence/medical-o1-reasoning-SFT
16,287 downloads · 1,168 likes · Updated Apr 22, 2025
- Qyrou/reasoning-corpus-4K-5M-v1
14,678 downloads · 204 likes · Updated Jul 31, 2026
Broaden dataset search
Spaces
No trustworthy spaces matches right now.
Search spaces on Hugging FaceResearch context
Tasks
None detected
Methods
None detected
Domains
Computer vision
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXJump to Paper2Code search queries derived from this paper's research context.
Data includes links from Papers with Code ( CC-BY-SA-4.0 ).