Skip to content
OpenTrain AIFor AI Companies

Multimodal Image Evaluation Annotator

Create evaluation materials for multimodal AI by researching real-world images, verifying provenance, editing visual clues, and writing prompts and rubrics. Work remotely for at least 20 hours per week with Pacific Time overlap.

OpenTrain AI

Image & Video Annotation

100% Remote

Worldwide

Eligibility

Entry

Experience

Sep 17, 2026

Posted

Open worldwide

About OpenTrain

OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. OpenTrain AI hires and contracts contributors for projects that help shape how modern AI systems understand and respond to real-world information.

You can create a free OpenTrain account, build a professional AI training profile, and apply in minutes. This role offers an opportunity to develop practical experience in multimodal evaluation, research, image handling, and quality-focused documentation.

About AI Training Work

AI training is the human side of building artificial intelligence. Contributors prepare examples, evaluate model outputs, and create reliable reference materials that help AI systems learn from images, text, audio, and other information.

Multimodal evaluation work focuses on how AI interprets visual information alongside web research and written instructions. Careful human judgment helps make these systems more accurate, useful, and grounded in evidence.

The Role

OpenTrain is recruiting a Multimodal Image Evaluation Annotator to create evaluation materials for AI systems working with visual information and web research. You will work with authentic photographs of real-world places, products, and media while preserving realistic visual clues and concealing key identifying details.

The role combines image sourcing, copyright and licensing verification, investigative research, image editing, prompt writing, rubric development, and documentation of ground-truth answers. The engagement is remote and intended for an entry-level contributor with strong English and careful analytical skills.

  • Remote contractor engagement
  • Part-time schedule of at least 20 hours per week
  • At least 4 hours per day required
  • Includes 4 hours of overlap with Pacific Time
  • Engagement may last up to 16 weeks
  • Compensation follows standard OpenTrain project defaults for this opportunity

What You'll Do

You will source and prepare visual evaluation tasks based on relevant, authentic photographs. Your work should retain realistic secondary details while preventing the image or prompt from directly revealing the target identity or answer.

You will also research incomplete visual clues, document evidence, and create objective materials that allow an AI system's answer to be evaluated consistently.

  • Source high-quality photographs for an assigned locale, including landmarks, universities, municipal buildings, products, and domestic film or media.
  • Verify whether images are original, publicly available, copyright-free, or explicitly licensed.
  • Document each image's provenance and license type.
  • Use Google Search, Google Lens, reverse image search, and other reliable sources to identify subjects from partial visual clues.
  • Create cropped, occluded, masked, or degraded image versions while retaining realistic secondary details.
  • Write time-bound English prompts that do not reveal the target identity or answer.
  • Produce concise, objective grading criteria and solution notes.
  • Document the research process, evidence, expected outcome, and ground-truth identity.

Requirements

This role requires native or highly fluent English and the ability to communicate clearly in prompts, rubrics, and solution notes. You should be comfortable researching unfamiliar subjects from incomplete visual information and documenting your reasoning in a structured way.

Strong familiarity with culturally significant but non-touristy places, institutions, products, or media in the assigned locale is important. You should also understand practical image licensing, copyright, and provenance verification, and be able to perform basic image edits.

  • Native or highly fluent English writing ability
  • Investigative web research skills using search and reverse-image tools
  • Strong local cultural knowledge relevant to the assigned locale
  • Practical understanding of image licensing, copyright, and provenance
  • Basic ability to crop, mask, or edit images
  • Careful rubric-writing and solution-note skills
  • Attention to detail with shared Google Drive folders, file naming, and structured submissions

Who Should Apply

This opportunity may suit people who enjoy visual investigation, online research, cultural discovery, and precise written analysis. Prior AI evaluation, data annotation, content review, quality assurance, or other analytical experience is helpful but not required.

A bachelor's degree or equivalent practical experience in any field is welcome. The listed experience level is entry level, so demonstrated care, strong judgment, and the ability to follow structured requirements are especially valuable.

  • Researchers who can identify subjects from partial visual clues
  • Writers who can create clear prompts, rubrics, and evidence-based notes
  • Contributors familiar with local places, institutions, products, or media
  • Detail-oriented workers comfortable organizing files and documenting sources
  • Applicants seeking flexible, remote experience in AI training and evaluation

How to Get Started

Create a free OpenTrain account to build your AI training profile and apply for this opportunity in minutes. OpenTrain helps contributors discover and grow through work that spans data labeling, evaluation, research, and other ways people teach AI.

If selected, you will complete the project according to its remote contractor schedule, time-zone overlap requirement, and submission standards. This role can help you build a credible portfolio of multimodal AI evaluation work.

  • Language: English
  • Work arrangement: Remote and worldwide
  • Engagement type: Contractor and part time
  • Availability: 20 or more hours per week
  • Pacific Time overlap: 4 hours required
  • Maximum engagement length: 16 weeks

Ready to apply?

Create a free OpenTrain account and apply for this role in minutes.

Keep exploring

Similar jobs

View all AI training jobs

Japanese Multimodal Image Evaluation Annotator

Help evaluate multimodal AI by researching Japanese places, products, and media, creating realistic image variations, and writing clear prompts and rubrics. This remote contractor role requires fluent Japanese, strong English, and 20 or more hours weekly.

Image & Video Annotation
Image
Remote · Worldwide
Japanese, English
Part-time · Flexible
Entry level

Posted Sep 17, 2026

Spanish Multimodal Image Evaluation Annotator

Research and evaluate authentic images of places, products, and media while creating realistic masked or degraded visuals for multimodal AI systems. Use Spanish cultural knowledge, English writing, web research, and image editing in a remote contract role.

Image & Video Annotation
Image
Remote · Worldwide
Spanish, English
Part-time · Flexible
Entry level

Posted Sep 17, 2026

Arabic Multimodal Image Evaluation Annotator

Create Arabic-focused multimodal image evaluation materials for AI systems by sourcing, editing, researching, and documenting partially hidden real-world images. This remote contractor role requires 20+ hours per week and strong Arabic and English skills.

Image & Video Annotation
Image
Remote · Worldwide
Arabic, English
Part-time · Flexible
Entry level

Posted Sep 17, 2026