Ara Zeta Project - AI Safety Evaluator - Korean (Korea)

2 months, 4 weeks ago
Freelance
Mid Level
Artificial Intelligence and Machine Learning
Welo Global

Welo Global

Welo Global specializes in providing enterprise localization, AI training data, and multilingual content solutions, leveraging a portfolio of specialized global brands to meet diverse customer needs in the AI and language services industry.

Professional Services
Founded 1997

Description

  • Evaluate AI-generated responses using a structured safety rubric.
  • Complete two independent evaluations for each prompt-image pair.
  • Write concise, well-structured rationales in English.
  • Participate in calibration sessions with the project team.
  • Support arbitration when evaluation discrepancies occur.
  • Maintain quality and throughput targets throughout the evaluation period.

Requirements

  • Fluency in the target language and English.
  • Deep cultural understanding of the target locale.
  • Strong written English skills for documentation and rationales.
  • Prior experience in safety evaluation, policy review, content moderation, or rubric-based assessment preferred.
  • Ability to apply detailed guidelines consistently.
  • Strong analytical skills and attention to nuance.
  • Reliable availability during the production window.
  • Experience with prompt-image or similar evaluation projects preferred to reduce onboarding time.
  • Comfort working with explicit and sensitive adult material in a professional capacity.

Benefits

  • Freelance remote engagement with complete autonomy over schedule.
  • Weekly commitment options of 20–40 hours per week.
  • Project duration of 2–3 weeks with an immediate start.
  • Hourly rate of 22 USD.
  • Optional access to AI and Large Language Model workshops.
  • Support from a global contributor community.
  • Opportunity to contribute to AI systems through applied expertise.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

QA Contractor

FUTURESEARCH Internet Software & Services

FUTURESEARCH is hiring a meticulous quality checker to verify research evaluation tasks used for training and benchmarking frontier AI models.

13 hours, 49 minutes ago

AI / ML Consultant

Crosslake 251-1K Capital Markets

Crosslake is hiring a US-based remote consultant to help technology and private equity clients evaluate, improve, and implement AI and software-related initiatives across the investment lifecycle.

Agile Computer Vision Generative AI LLM Machine Learning MLOps NLP Statistics
1 day, 14 hours ago

Test Lead - (Temenos T24 - Mortgage & Lending)

Unison Group Technology consulting

Test Lead for Temenos T24 mortgage enhancements at a banking platform, responsible for end-to-end testing across core banking and workflow systems.

Agile
1 day, 14 hours ago

Senior Python Engineer - AI Coding Agent Evaluation (Freelance)

Mindrift.ai: Be the “I” in AI Internet Software & Services

Mindrift is seeking an experienced software development specialist to create realistic coding-agent evaluation tasks and tests for project-based AI opportunities.

Docker FastAPI JavaScript Kafka PostgreSQL Python React Redis TypeScript
2 days, 13 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers