Talent.com
24-MAG
Remote | Urdu-English AI Safety Red Team Evaluator —$20–$30/hour24-MAG • New York, United States
Remote | Urdu-English AI Safety Red Team Evaluator —$20–$30/hour

Remote | Urdu-English AI Safety Red Team Evaluator —$20–$30/hour

24-MAG • New York, United States
30+ days ago
Salary
$20.00 hourly
Job type
  • Part-time
  • Remote
Job description
We are sharing a specialised part-time consulting opportunity for Urdu-English bilingual professionals experienced in AI safety evaluation, red team testing, adversarial review, vulnerability classification, and structured feedback on sensitive text-based AI outputs. This role supports current and upcoming remote consulting opportunities focused on AI safety evaluation, bilingual red team testing, conversational model assessment, misuse-risk review, vulnerability annotation, and high-quality project execution. Selected professionals will test AI systems using structured adversarial scenarios, identify safety weaknesses, classify risks, and produce clear English-language evaluation artifacts across English and Urdu contexts. Key Responsibilities Professionals in this role may contribute to: Bilingual AI Safety & Red Team Testing Review English and Urdu AI outputs for safety, reliability, bias, misinformation, and harmful-behavior risks Stress-test conversational AI models and agents using structured adversarial scenarios Evaluate model behavior across multi-turn conversations, sensitive topics, and edge-case prompts Identify vulnerabilities that require stronger safety controls, clearer refusals, or improved response quality Vulnerability Classification & Risk Review Annotate failures, classify vulnerabilities, and flag recurring safety patterns Apply taxonomies, benchmarks, and project-specific playbooks to keep testing consistent Assess misuse cases, bias exploitation, prompt-injection scenarios, and socio-technical risk patterns at a high level Generate high-quality human evaluation data through careful review and structured judgment Reproducible Documentation & Evaluation Artifacts Produce clear reports, datasets, test cases, and written summaries that support model improvement Document findings reproducibly so results can be reviewed, compared, and acted upon Explain risks clearly for both technical and non-technical audiences Maintain accuracy, consistency, and strong attention to detail across submitted evaluations Ideal Profile Strong candidates may have: Native-level fluency in both English and Urdu Prior experience in AI red teaming, adversarial testing, cybersecurity, trust and safety, socio-technical risk review, or conversational AI evaluation Ability to think adversarially while staying structured, careful, and methodical Experience using frameworks, benchmarks, or rubrics rather than unstructured testing alone Strong written communication skills and ability to explain safety findings clearly Comfort reviewing text-based content involving sensitive topics under clear guidelines Adaptability across project types, safety categories, and evaluation workflows Educational Background Formal degree requirements may vary based on project needs Backgrounds in AI safety, cybersecurity, linguistics, policy, trust and safety, social science, psychology, writing, data evaluation, or technical analysis may be highly relevant Practical experience in red team testing, model evaluation, content risk analysis, or structured review work may also be valuable Nice to Have Experience with adversarial ML concepts, jailbreak datasets, prompt injection, RLHF/DPO attack patterns, or model behavior testing Cybersecurity experience such as penetration testing, exploit analysis, reverse engineering, or security assessment Socio-technical risk experience involving harassment, misinformation, abuse analysis, bias testing, or conversational AI safety Creative probing background, including psychology, acting, writing, role-play design, or unconventional adversarial thinking Experience producing reproducible reports, labeled datasets, structured risk notes, or benchmark-style evaluation artifacts Why This Opportunity Apply Urdu-English bilingual expertise to structured AI safety and red team evaluation work Contribute to stronger, safer, and more reliable AI systems through careful adversarial testing Work on flexible assignments aligned with language skills, safety judgment, and structured analysis Build experience in human data-driven AI safety evaluation and bilingual risk review Remote structure with competitive hourly compensation Contract Details Independent contractor role Fully remote with flexible scheduling Eligible professionals may be based in approved project locations depending on project needs Native-level English and Urdu fluency are required for project work Work is text-based and may involve sensitive topics such as bias, misinformation, harassment, or harmful-behavior risks Topic areas will be communicated before exposure to content, and participation in higher-sensitivity projects may depend on candidate comfort and project fit Part-time commitment depending on project availability Competitive rates between $20–$30 per hour depending on expertise and project scope Weekly payments via Stripe or Wise Projects may be extended, shortened, or adjusted depending on scope and performance Work will not involve access to confidential or proprietary information from any employer, client, or institution About the Platform This opportunity is available through 24-MAG LLC. We connect experienced professionals with remote consulting opportunities across technical, evaluation, and project-based workstreams. By submitting this application, you acknowledge that your information may be processed by 24-MAG LLC for recruitment and opportunity matching in accordance with our Privacy Policy: https://www.24-mag.com/privacy-policy .
Create a job alert for this search

Remote | Urdu-English AI Safety Red Team Evaluator —$20–$30/hour • New York, United States

Similar jobs

Remote | Turkish Voice AI Evaluation Specialist Up to $25/hour

24-MAG LLCNew York, NY, United States
Remote
Part-time

Specialised Part-Time Consulting Opportunity.We are sharing a specialised part-time consulting opportunity for native Turkish speakers experienced in bilingual communication, voice-based task compl... Show more

 • Promoted

Clojure AI / LLM Engineer - REMOTE

StyliticsNew York City, NY, United States
Remote
Full-time

We are looking for a versatile and experienced AI / LLM Data Engineer to join our team and help shape the future of how Stylitics leverages large language models.In this role, you'll combine your e... Show more

 • Promoted

Earn Side Money at Home Testing Products

Product Review JobsLEONARDO, NJ, United States
Full-time

Compensation: Varies per assignment.Location: Remote (USA) Company: ProductReviewJobs Thank you for your interest in becoming a Paid Product Tester.This opportunity is for completing market res... Show more

 • Promoted

FT / PT Remote AI Prompt Engineering & Evaluation - Will Train

DataAnnotation.techNew York City, NY, United States
Remote
Full-time +1

This is a full-time or part-time REMOTE position.You'll be able to choose which projects you want to work on, and you can work on your own schedule.Projects are paid hourly, starting at $15-20 per ... Show more

 • Promoted

AI Workflow Developer (Remote) - Media & Entertainment

Tandym GroupEnglewood Cliffs, NJ, United States
Remote
Full-time

AI Workflow Developer (Remote) - Media & EntertainmentGet AI-powered advice on this job and more exclusive features.A recognized media organization offers a remote opportunity for an innovative... Show more

 • Promoted

AI Adoption Lead -Remote

Finn PartnersNew York City, NY, United States
Remote
Full-time

AI Adoption LeadThe AI Adoption Lead plays a critical role in bridging advanced AI capabilities with practical, everyday business solutions.This position focuses on driving AI adoption across the o... Show more

 • Promoted

Become a Luxury Brand Evaluator in New Jersey, US

CXGRed Bank, NJ, United States
Full-time

Turn your passion for luxury into a career opportunity.Explore the world of premium brands and make a lasting impact in fashion, beauty, jewelry, or automobiles.Join CXG, the global leader in custo... Show more

 • Promoted

AI Safety Expert - Red Teaming

MercorNew York, New York, United States
$20.00 hourly
Remote
Part-time
Quick Apply

Headquartered in San Francisco, our investors include.AI Safety Experts — English & Odia.Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging syste... Show more

Remote Product Tester - Flexible Studies

BuzzTestersOakhurst, NJ
Remote
Full-time

Earn up to $400/week, depending on the number and type of studies you complete.Share honest feedback on products and services from independent brands.Assignments include short online surveys (10–20... Show more

 • Promoted

AI Red Team Specialist - Remote | Upto $22/hr

MercorNew York, New York, United States
$20.00 hourly
Remote
Part-time
Quick Apply

Headquartered in San Francisco, our investors include.AI Safety Experts — English & Malayalam.Perform jailbreaks, prompt injections, misuse cases, and bias exploitation.Generate high-quality hu... Show more

Flexible remote AI work. Your schedule. Paid weekly, straight to your bank account.

Meridian.aiSpring Valley, NY, US
Remote
Full-time

Review and label digital content including text, images, and documents.Every task you complete helps improve how technology interprets information and performs in practical settings.Detail-oriented... Show more

 • Promoted

Web Browsing Evaluator remote

ESRhealthcare and EXEC STAFF RECRUITERSNYC, New York, United States
Remote
Full-time
Quick Apply

Role Title: Web Browsing Evaluator.Web Browsing Evaluators to contribute high-quality web interaction data for a valued customer project.In this role, you'll apply your expertise to help train next... Show more

Remote Referral Partnerships Lead

Micro1Bradley Gardens, New Jersey, US
$30.00 hourly
Remote
Full-time

AI data lab for training frontier models and evaluating AI agents.Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more.AI... Show more

 • Promoted

Consumer Insights Analyst

Earn HausOceanport, New Jersey, United States
Full-time +1

We are urgently seeking people interested in taking market research studies for well known brands.If you are a self-starter, looking for flexible hours throughout the week, this may be for you! Ear... Show more

 • Promoted

AI Safety Expert - Red Team

MercorNew York, New York, United States
$20.00 hourly
Remote
Part-time
Quick Apply

Headquartered in San Francisco, our investors include.AI Safety Experts — English & Bengali.Focus on jailbreaks, prompt injections, misuse cases, and bias exploitation.Generate high-quality hum... Show more

Hybrid Remote Solutions Engineer (Pre-Sales) AI Safety

intenseyeNew York City, NY, United States
Remote
Full-time

A tech firm focused on AI-powered safety is seeking a Solutions Engineer (Pre Sales) to support sales expansion into Fortune 500 companies.The ideal candidate will interface with clients to provide... Show more

 • Promoted

Engineering Expert remote

ESRhealthcare and EXEC STAFF RECRUITERSNYC, New York, United States
Remote
Full-time
Quick Apply

Code=1761f2a1-0970-4c4f-a273-a1b557d5ad9f&utm_source=referral&utm_medium=share&utm_campaign=job_referral.Engineering Experts to contribute expert knowledge and analytical capabilities t... Show more

Technical SEO Engineer - Remote Impact & AI-Driven Growth

iPullRankNew York City, NY, United States
Remote
Full-time

Join a dynamic digital marketing agency as a Search Engine Optimization Engineer, where your expertise will help clients enhance their organic visibility.This role involves conducting technical aud... Show more

 • Promoted

English (U.S. Native) AI Trainer & Evaluator (Remote, Hourly Contractor)

CNTXT AIBrooklyn, new york, US
Remote
Full-time

This is a fully remote, hourly contractor role supporting AI data and language projects on a project-based, flexible hour basis.Project scopes vary and may include:.AI learning across diverse topic... Show more

Freelance Luxury Brand Evaluator - Central & Southern New Jersey, US

CXGWest Long Branch, NJ, United States
Full-time

Turn your passion for luxury into a career opportunity.Explore the world of premium brands and make a lasting impact in fashion, beauty, jewelry, or automobiles.Join CXG, the global leader in custo... Show more