Summary
Capability Elicitation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.Category: AI Safety & Red Teaming · Pay: $65–$70 / hr · Location: Remote — US-eligible · ContractorCapability Elicitation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts.About the roleCapability Elicitation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.Push on refusal boundaries and dual-use risk before a model ships.ResponsibilitiesDesign adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for Capability Elicitation Red Team Specialist assignments.Document every successful attack with reproduction steps and the policy clause it violated.Score model defenses across single-turn and multi-turn conversations.Triage emerging attack vectors and route them to the safety team with severity ratings.Role detailsTrack Adversarial evaluation Work model Remote · Independent specialist contractor Compensation Hourly rate confirmed after the interview process. Eligible from USWhat you should bringDemonstrated experience red-teaming AI systems, security research, or adversarial ML work for Capability Elicitation Red Team Specialist work.Strong written communication — your reports become the patch ticket.Comfort working in policy-grey areas with clear documentation of what was attempted and why.Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.Reliable async availability for at least 10 hours per week.Role signalsExample tasksConstruct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.Score a model's defenses against a known jailbreak pattern across 20 variants.Propose a new red-team rubric category after spotting an emerging attack vector.Reproduce a failure another reviewer reported and confirm the severity tag.Useful experienceBackground in offensive security, AppSec, or trust & safety operations.Experience publishing or reproducing public adversarial-ML research.Multilingual fluency for cross-language attack testing.Compensation and scheduleHourly rate confirmed after the interview process.Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.Skills used in matchingAdversarial promptingRed-team analysisPolicy taxonomyFailure documentationAdversarial prompt testingAI safetyRed-team evaluationPolicy rubricsCapabilityElicitationApplication boundaryCreating a specialist profile records your experience and preferences. Starting role intake is a separate action that attaches this role to your candidate record.Specialist intake- 01 Confirm profile and eligibility details.- 02 Attach this role deliberately.- 03 Receive a human review decision or follow-up.Placement timing depends on program demand and reviewer confirmation.#J-18808-Ljbffr
Job Description
Capability Elicitation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.Category: AI Safety & Red Teaming · Pay: $65–$70 / hr · Location: Remote — US-eligible · ContractorCapability Elicitation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts.About the roleCapability Elicitation Red Team Specialist is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.Push on refusal boundaries and dual-use risk before a model ships.ResponsibilitiesDesign adversarial prompts that probe known weakness classes (jailbreak, policy bypass, prompt injection) for Capability Elicitation Red Team Specialist assignments.Document every successful attack with reproduction steps and the policy clause it violated.Score model defenses across single-turn and multi-turn conversations.Triage emerging attack vectors and route them to the safety team with severity ratings.Role detailsTrack Adversarial evaluation Work model Remote · Independent specialist contractor Compensation Hourly rate confirmed after the interview process. Eligible from USWhat you should bringDemonstrated experience red-teaming AI systems, security research, or adversarial ML work for Capability Elicitation Red Team Specialist work.Strong written communication — your reports become the patch ticket.Comfort working in policy-grey areas with clear documentation of what was attempted and why.Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.Reliable async availability for at least 10 hours per week.Role signalsExample tasksConstruct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.Score a model's defenses against a known jailbreak pattern across 20 variants.Propose a new red-team rubric category after spotting an emerging attack vector.Reproduce a failure another reviewer reported and confirm the severity tag.Useful experienceBackground in offensive security, AppSec, or trust & safety operations.Experience publishing or reproducing public adversarial-ML research.Multilingual fluency for cross-language attack testing.Compensation and scheduleHourly rate confirmed after the interview process.Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.Skills used in matchingAdversarial promptingRed-team analysisPolicy taxonomyFailure documentationAdversarial prompt testingAI safetyRed-team evaluationPolicy rubricsCapabilityElicitationApplication boundaryCreating a specialist profile records your experience and preferences. Starting role intake is a separate action that attaches this role to your candidate record.Specialist intake- 01 Confirm profile and eligibility details.- 02 Attach this role deliberately.- 03 Receive a human review decision or follow-up.Placement timing depends on program demand and reviewer confirmation.#J-18808-Ljbffr
Government Careers
Government jobs offer stability, competitive benefits, and the chance to make a meaningful impact on your community and country.
Whether you’re starting your career or seeking new opportunities, these roles provide pathways for growth, security, and service.
Explore positions across a wide range of fields and take the first step toward a rewarding future in public service.
MORE JOBS
-
Lead Statistical Assistant (Administrative Specialist 1)
- Trenton, New Jersey
- New Jersey Judiciary
- Sep 30, 2026
-
Fast Track to Management Program
- Uniondale, New York
- New York Life
- Oct 06, 2026
-
Veterans Justice Programs Specialist
- New York, New York
- AlphaX
- Oct 06, 2026
-
Accelerated Path to Management Program
- Saddle Brook, New Jersey
- New York Life
- Oct 06, 2026
-
Public Policy Senior Analyst
- New York, New York
- REBNY (Real Estate Board of New York)
- Oct 06, 2026
-
Accelerated Path to Management Program
- Edison, New Jersey
- New York Life
- Oct 06, 2026