AI Safety & Red Teaming Specialist

1 month, 3 weeks ago
Contract
Junior
Artificial Intelligence and Machine Learning
Weekday

Weekday

Weekday helps companies hire engineers who are vouched by other software engineers, enabling passive income for engineers. They offer services like drafting outreach messages, shortlisting candidates, and conducting reference checks. Backed by Y Combin...

Construction & Engineering
11-50
Founded 2020

Description

  • Design and implement advanced evaluation methodologies for AI system safety, including ethical jailbreak testing, prompt injection detection, LLM red teaming, and tool-use abuse scenarios.
  • Develop cross-domain adversarial testing strategies to uncover complex, multi-turn attack patterns and model vulnerabilities.
  • Build, maintain, and enhance regression test suites to continuously assess jailbreak susceptibility and prompt injection risks.
  • Create comprehensive evaluation frameworks that simulate real-world adversarial threats to improve AI robustness and reliability.
  • Collaborate with technical teams to translate security findings into actionable recommendations for AI safety improvements.
  • Document testing methodologies, findings, and best practices in clear technical reports and presentations for technical and non-technical stakeholders.

Requirements

  • 2+ years of experience in AI Safety, Adversarial Machine Learning, LLM Red Teaming, AI Security, or a related field.
  • Hands-on experience researching, testing, or identifying vulnerabilities involving prompt injection, ethical jailbreaks, adversarial attacks, or tool-use exploitation.
  • Strong understanding of modern LLM architectures, prompt engineering, and AI safety evaluation methodologies.
  • Experience developing structured security assessments, regression testing frameworks, and adversarial evaluation strategies.
  • Excellent analytical, documentation, and communication skills with the ability to explain complex technical findings clearly.
  • Ability to collaborate effectively within cross-functional technical teams.
  • Master's or PhD in Computer Science, Cybersecurity, Machine Learning, Artificial Intelligence, or a related discipline is preferred.
  • Contributions to AI security research, open-source AI safety tools, conference presentations, or published research are preferred.
  • Experience with AI model evaluation frameworks, prompt engineering techniques, and AI security assessment tools is preferred.
  • Background in multidisciplinary AI safety, cybersecurity, or adversarial machine learning projects is preferred.
  • AI Safety.
  • LLM Red Teaming.
  • Prompt Injection.
  • Adversarial Machine Learning.
  • Ethical Jailbreaking.
  • AI Security.
  • Prompt Engineering.
  • AI Evaluation Frameworks.

Benefits

  • $50 - $90/hour pay.
  • Remote work location.
  • Contractor role.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Generative AI Analyst | Arabic (Morocco)

Welo Global Professional Services

Welo Data is seeking a remote freelance Generative AI Analyst fluent in Moroccan Arabic to evaluate, annotate, and quality-review AI-generated content across text, audio, images, video, and related data projects.

Generative AI
9 hours, 50 minutes ago

French Quality Control Specialist - Project Lyra

Welo Global Professional Services

Welo Data is seeking French-speaking linguistics experts for a four-week remote freelance project evaluating and improving AI-generated language outputs for accuracy, grammar, and cultural relevance.

LLM NLP
10 hours, 20 minutes ago

Generative AI Analyst | Traditional Chinese (US)

Welo Global Professional Services

Welo Data, part of Welocalize, is seeking remote freelance Traditional Chinese Generative AI Analysts in the US to evaluate, annotate, and improve AI-generated content across text, audio, images, video, and other data types.

Generative AI LLM
2 days, 9 hours ago

Freelance Agent Evaluation Engineer

Mindrift.ai: Be the “I” in AI Internet Software & Services

Mindrift seeks experienced software developers for a project-based opportunity creating and evaluating realistic coding tasks used to assess AI coding agents for leading technology companies.

Cybersecurity Docker FastAPI JavaScript Kafka Machine Learning PostgreSQL Python React Redis TypeScript
2 days, 10 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers