OpenTrain AI
This is a part-time AI Safety Red Teamer position at OpenTrain AI, based in Anywhere, with remote work available. The role offers $0.1k - $0.1k.
$0.1k - $0.1k
About OpenTrain
OpenTrain is the #1 platform for finding and building careers in AI training and data labeling. We help people discover specialized projects, build a credible AI training profile, and apply for work that matches their experience.
Creating an OpenTrain account is free. Your profile can help you present your experience in AI safety, red teaming, trust and safety, and related fields as you pursue a longer-term career in AI training.
About AI Safety Red Teaming
AI training is the human work behind modern artificial intelligence. Contributors write prompts, evaluate model responses, identify failures, and provide structured feedback that helps AI systems become more reliable and safer.
Red teaming applies this process adversarially. Instead of testing only expected behavior, you will probe model boundaries, surface vulnerabilities, and examine how systems respond to ambiguous, sensitive, or high-risk scenarios.
OpenTrain is recruiting an AI Safety Red Teamer to perform adversarial testing of frontier AI systems. The work combines structured evaluation, safety judgment, prompt design, and clear documentation to support model robustness and alignment.
This is an expert-level, part-time contractor opportunity requiring 20 or more hours per week. The role is available to candidates in the listed countries and requires English-language work.
You will design challenging prompts and evaluate model behavior across complex, high-risk, and ambiguous topics. Findings must be recorded clearly so researchers and safety teams can understand the issue, assess its significance, and strengthen safeguards.
Your evaluations may cover cyber, biosecurity, fraud, political content, scientific safety, and related sensitive domains.
Required Qualifications
A bachelor's degree or higher in computer science, cybersecurity, journalism, communications, psychology, biology, chemistry, public policy, or a related discipline is required. You must also have at least five years of professional experience in AI safety, AI red teaming, trust and safety, cybersecurity, investigative journalism, life sciences, or a related field.
Demonstrated experience designing adversarial prompts or evaluating frontier AI systems is required. Strong analytical reasoning, prompt design, and written communication skills are essential for this work.
Helpful Background
Experience with AI red teaming, reinforcement learning from human feedback, supervised fine-tuning, AI alignment, trust and safety, jailbreak testing, prompt engineering, or adversarial evaluation methodologies is valuable.
Specialized knowledge of cyber, biosecurity, political content, misinformation, or scientific safety can help you assess difficult cases with appropriate context and care.
Why This Work Matters
Every major AI system depends on human contributors who prepare examples, assess outputs, and identify weaknesses. By testing how frontier models behave under pressure, red teamers help shape safer and more dependable AI systems.
AI training is a fast-growing field that can offer flexible, remote work for people with specialized expertise. This role gives experienced professionals a direct way to apply their judgment to cutting-edge AI development.
Eligible Locations
This opportunity is open to candidates located in the United States, Denmark, Estonia, Finland, Ireland, Latvia, Lithuania, Norway, Sweden, Austria, Belgium, France, Germany, the Netherlands, Switzerland, the United Kingdom, Albania, Bosnia and Herzegovina, Croatia, Greece, Italy, Malta, Portugal, Serbia, Slovenia, Spain, Bulgaria, Czechia, Hungary, Moldova, Poland, Romania, or Slovakia.
© 2025 JustJobs. All rights reserved.