Remote (India) · Mid Level · Remote
Applicants who checked fit first are 3.1× more likely to hear back
Your score for this role already exists
ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.
No credit card · 1 tap with Google
HIGH
Jobgether is actively reviewing profiles and moving candidates through the pipeline right now.
First 72 hours
Still inside it - posted 2h agoEarly applicants get seen before the pile builds.
Not a repost
The first time we've seen this listing - it hasn't been closed and reopened.
You almost certainly match several of these already. Unlock your skill map to see the matches, the gaps, and what to fix first.
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Content Adversarial Red Team Analyst - English based in India.
As a Content Adversarial Red Team Analyst, you will help evaluate AI systems against content safety policies and compliance requirements.
You will design challenging test scenarios that explore how AI platforms respond to complex, unexpected, or ambiguous inputs.
The role involves examining edge cases, testing different forms of language and intent, and identifying gaps in safety controls or policy enforcement.
You will use creativity and analytical judgment to approach systems from different user perspectives and uncover potential weaknesses.
Your findings will help teams understand where AI behavior can be improved and support the development of safer, more reliable systems.
This freelance opportunity offers the potential to transition into a full-time role and provides hands-on exposure to AI evaluation and safety testing.
Design and execute authorized adversarial test scenarios to evaluate AI systems against defined content safety policies and compliance requirements.
Develop diverse, challenging, and unexpected prompts, inputs, and situations designed to test system behavior.
Explore edge cases, unusual inputs, complex contexts, and different user behaviors that may expose gaps in safety controls.
Evaluate AI-generated responses and identify potential weaknesses, inconsistencies, policy concerns, or failures in enforcement.
Test system behavior across different forms of language, context, intent, and communication styles.
Identify recurring patterns and scenarios that may require additional investigation, testing, or system improvement.
Clearly document test scenarios, system responses, findings, and supporting evidence.
Apply project requirements, testing methodologies, and content safety guidelines consistently.
Review complex or ambiguous cases and apply sound judgment when evaluating system behavior.
Maintain high standards of accuracy, consistency, documentation quality, and attention to detail across assigned evaluations.
Educational background or equivalent experience in Trust & Safety, Content Safety, Policy, Linguistics, Communications, Journalism, Research, AI Evaluation, or a related field.
Relevant experience in content safety, trust and safety, AI evaluation, content moderation, policy enforcement, quality assurance, or a related discipline.
Strong understanding of content safety principles, policy enforcement, and common safety risks.
Strong written English comprehension and communication skills, with the ability to understand nuanced language and context.
Strong understanding of user intent, contextual meaning, and the different ways people may communicate.
Ability to think creatively and develop challenging, unusual, or unexpected test scenarios.
Strong analytical and critical-thinking skills, with the ability to identify patterns, weaknesses, inconsistencies, and behavioral gaps.
Comfort working with complex or ambiguous situations and making informed decisions based on defined requirements.
Ability to understand and consistently apply detailed testing guidelines and policies.
Strong documentation skills and exceptional attention to detail.
Familiarity with AI systems, large language models, adversarial testing, red teaming, or AI safety is preferred.
Ability to work independently while maintaining consistent quality across a high volume of evaluation tasks.
Compensation: US$13 per hour.
Contract type: Freelance, with potential opportunity to transition into a full-time role.
Location: India.
Language: English.
Flexibility: Freelance structure offering flexibility in managing assigned work.
Professional exposure: Hands-on experience evaluating AI systems, content safety controls, and policy compliance.
Impact: Contribute to identifying safety gaps and improving the reliability of AI-powered systems.
Skill development: Opportunity to strengthen expertise in AI evaluation, adversarial testing, trust and safety, and policy analysis.
Free · no signup
Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.
No spam. Just jobs and resources.
Why people use ASAI
Scored, not searched. Every role ranked against your actual profile.
Alerts as often as hourly. Reach new roles while the pile is still small.
Skill gaps, spelled out. See exactly which requirements you don't meet yet.
Verified jobs, only. Say no to ghost jobs. Your time deserves respect.
Keep browsing