Mercor is a San Francisco-based company that connects top creative and technical talent with leading AI research labs. Backed by prominent investors such as Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey, Mercor is seeking AI Safety Experts fluent in both English and Punjabi for a remote, contract-based role. The position offers competitive compensation ranging from $20 to $22 per hour and focuses on ensuring the safety and robustness of conversational AI models through rigorous adversarial testing and analysis.

Key Responsibilities
- Conduct red teaming exercises on conversational AI models and agents, with an emphasis on identifying jailbreaks, prompt injections, misuse cases, and bias exploitation.
- Generate and annotate high-quality human data by classifying vulnerabilities, documenting failures, and flagging systemic risks.
- Apply structured methodologies using taxonomies, benchmarks, and playbooks to ensure consistent and thorough testing across projects.
- Produce reproducible documentation including detailed reports, datasets, and attack case studies to support customer action and remediation efforts.

Required Qualifications
- Native fluency in both English and Punjabi is essential for this role.
- Prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing is required.
- Strong communication skills with the ability to clearly explain risks to both technical and non-technical stakeholders.
- Flexibility and adaptability to work across multiple projects and client environments.

Preferred Qualifications
- Experience with adversarial machine learning techniques such as jailbreak datasets, prompt injection, RLHF/DPO attacks, and model extraction.
- Background in cybersecurity, including penetration testing, exploit development, and reverse engineering.
- Expertise in socio-technical risk areas like harassment and disinformation probing, abuse analysis, and conversational AI testing.
- Creative probing skills drawing from psychology, acting, or writing to foster unconventional adversarial thinking.

Application Process
The application process is designed to be efficient, taking approximately 20 to 30 minutes to complete. Candidates will be required to upload their resume, participate in an AI-driven interview based on their resume, and submit a form to finalize their application. Mercor’s team reviews applications daily, so timely completion of all steps is encouraged to be considered for the opportunity.

Additional Information
For detailed information about the interview process and platform, candidates are encouraged to visit Mercor’s official talent documentation site. Support resources are available for any questions or assistance needed during the application process.

This role offers a unique opportunity to contribute to cutting-edge AI safety research in a dynamic, remote work environment, supported by a company with a strong reputation and influential backers.

Job Details

Total Positions:
1 Post
Job Shift:
Work From Home
Job Type:
Job Location:
Gender:
No Preference
Age:
18 - 65 Years
Career Level:
Mid-Level
Maximum Experience:
5 Years
Apply Before:
Sep 29, 2026
Posting Date:
Sep 23, 2026

Mercor

· 11-50 employees - Karachi

What is your Competitive Advantage?

Get quick competitive analysis and professional insights about yourself
Talk to our expert team of counsellors to improve your CV!
Try Rozee Premium
I found a job on Rozee!