M4: Multi-generator, Multi-domain, and Multi-lingual Black-Box Machine-Generated Text Detection
Results and benchmarks
M4: Multi-generator, Multi-domain, and Multi-lingual Black-Box Machine-Generated Text Detection is the primary contribution described in this paper.
Benchmark evidence is limited
Evidence graph: 3 refs, 3 links.
Utility signals: depth 65/100, grounding 75/100, status medium.
Implementation
Historical official implementation (not recommended for new builds)
Only a historical official implementation is available
Use with caution for new projects; verify against current tooling and maintained community alternatives.
mbzuai-nlp/semeval2024-task8 · 82 stars · Last push Apr 22, 2024
mbzuai-nlp/m4 is the closest maintained adjacent implementation (Official implementation from Papers with Code). It is not paper-verified; validate algorithm and evaluation setup against the paper before trusting reported metrics. Community adoption signal: 52 GitHub stars.
Open mbzuai-nlp/semeval2024-task8- Adjacent implementations are not paper-verified
- Recommended repository is adjacent and not paper-verified.
- Adjacent implementation match confidence is low.
- No direct maintained implementation is currently verified.
- Only historical official repository was found: mbzuai-nlp/semeval2024-task8.
- No maintained paper-verified implementation met reliability thresholds.
Compare implementation paths
Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Limited
- Stars
- 82
- Last push
- Apr 22, 2024 (855d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No CI pipeline detected
- No tagged releases
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Limited
- Stars
- 52
- Last push
- Apr 9, 2024 (868d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No CI pipeline detected
- No tagged releases
- Maintenance
- Recently updated
- Confidence
- Low
- Reproducibility
- Limited
- Stars
- 1
- Last push
- Jul 4, 2026 (52d)
Matched via arXiv identifier search
- No CI pipeline detected
- No tagged releases
- No Docker setup
Reproduction readiness
Major work
No dependency manifest, manual reconstruction required
- mbzuai-nlp/semeval2024-task8 has no requirements.txt, environment.yml, pyproject.toml, or Dockerfile.
- You will need to reverse-engineer dependencies from import statements in the source code.
- Last push was 855 days ago.
Hardware requirements
- Expect multi-day setup/compute for meaningful reproduction based on current guidance.
Validation caveat
Repositories and ecosystem
Closest related implementations
These are not paper-verified. Use them as reference points when no direct implementation is available.
- mbzuai-nlp/m4 Adjacent · Confidence: Low · 52 stars
Official implementation from Papers with Code
Official
- mbzuai-nlp/m4Confidence: High
M4: Multi-generator, Multi-domain, and Multi-lingual Black-Box Machine-Generated Text Detection
52 stars · 5 forks · Last push Apr 9, 2024
Community
No additional community repositories detected yet.
These repositories had low-confidence matching signals and are hidden by default.
Showing top 6 by score. 1 additional low-confidence matches are hidden.
- bluestar0037-cmyk/aigc-text-detection-demo
Confidence: Low · 1 stars
- ahm1129/Msc-AI-Detection
Confidence: Low · 0 stars
- Elena-2521/Week7-8-AI-Text-Detect-Experiment
Confidence: Low · 0 stars
- ValentinVinka/SemEval-2024-Task-8
Confidence: Low · 1 stars
- andricValdez/semeval
Confidence: Low · 0 stars
- codefire53/NLP701-Assignment2
Confidence: Low · 0 stars
Hugging Face artifacts
No trustworthy direct or curated related Hugging Face artifacts were found yet. Use targeted searches to quickly locate candidate models, datasets, and demos.
Datasets
Spaces
Tip: start with models, then check datasets and spaces if you need evaluation data or demos.
Research context
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXData includes links from Papers with Code ( CC-BY-SA-4.0 ).