Other open roles
LLM Red-Teamer
Work Expert (WE) is an independent publishing and referral website. We are not a recruiter, hiring manager, agent or employer, and we are not affiliated with or endorsed by micro1. Applying takes you to the platform's own website, where we may be recorded as the referring source. We may receive a referral fee at no additional cost to you. Read our full affiliate disclosure →
About the Role
micro1 is engaging LLM Red-Teamers to contribute to a high-impact customer project focused on the evaluation and improvement of frontier language models. In this role, you'll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world input. No prior experience in AI is required — your domain knowledge is what matters. • Develop complex, adversarial multi-turn conversations and task-based scenarios aligned with detailed project specifications. • Autho
What You'll Do
- Develop complex, adversarial multi-turn conversations and task-based scenarios aligned with detailed project specifications.
- Author clear, precise evaluation rubrics to rigorously assess model responses against defined behavioral targets.
- Iteratively test conversations and tasks against frontier LLMs, escalating difficulty and nuance until the desired quality threshold is achieved.
- Deliver comprehensive task packages, including transcripts, target behaviors, binary rubrics, and supporting rationale or evidence.
- Validate LLM outputs, documenting model strengths and failure modes relative to the project specification.
- Maintain calibration with team leads and quality control contacts as project requirements evolve.
You're a Good Fit If You
- Required skills: Adversarial prompt construction, Precision in written English, Rubric design, Iteration stamina
- Exceptional written English skills, with clarity, precision, and strong structural organization.
- Prior experience in AI human data environments (RLHF, SFT, evaluations, annotation, or prompt engineering).
- Deep familiarity with large language models, including the ability to anticipate and identify common failure patterns.
- Demonstrated ability to work autonomously, interpreting and executing complex specifications with minimal oversight.
- Proven critical thinking and meticulous attention to detail.
- Experience designing evaluation items or rubrics is advantageous.
- Background in writing-intensive or analysis-centric fields such as research, editorial, technical writing, or quality assurance is a plus.
Role Highlights
Pay & Payout
Engaged directly through micro1. Apply via the link below - micro1 handles onboarding and payment.
Apply on micro1Opens micro1 in a new tab. You pay nothing; we may earn a referral fee.
Similar roles
More research work you may qualify for.
Economics Expert (PhD)
micro1 is engaging Economics Experts (PhD) to contribute advanced subject-matter expertise to a customer's AI training initiative. In this role, you'll apply your expertise to help
Statistics Expert (PhD)
micro1 is engaging Statistics Experts (PhD) to contribute their advanced subject-matter expertise to a customer's AI training initiative. In this role, you'll apply your expertise
Computer Science Expert (PhD)
micro1 is engaging Computer Science Experts (PhD) to contribute their advanced subject-matter expertise to a customer's AI training initiative. In this role, you'll apply your expe