CybersecurityJobs.io
← Back to all jobs

Job Description

AuraOne Human Data is building adversarial evaluation into its AI AppSec process so models are hardened before they ship to customers. This remote contract role supports that goal by running red-team style security assessments, focusing on jailbreak and policy-bypass weaknesses through carefully crafted adversarial prompts.

As an AppSec Security Evaluator, you will probe AI behavior in both single-turn and multi-turn settings, produce clear failure reports with reproduction details, and help the safety team prioritize fixes and regress-test coverage as new attack vectors appear.

Responsibilities

  • Design adversarial prompts to probe known weakness classes for AppSec Security Evaluator assignments, including jailbreak, policy bypass, and prompt injection.
  • Document every successful attack with reproduction steps and the policy clause that was violated.
  • Score model defenses across single-turn and multi-turn conversations.
  • Triage emerging attack vectors and route them to the safety team with severity ratings.
  • Maintain a personal library of attack patterns and propose new red-team rubrics.
  • Calibrate with the broader red-team cohort to keep coverage and severity consistent.

Requirements

  • Demonstrated experience red-teaming AI systems, security research, or adversarial ML work for AppSec Security Evaluator roles.
  • Strong written communication, since reports function as the patch ticket.
  • Comfort working in policy-grey areas, with clear documentation of what was attempted and why.
  • Familiarity with prompt-injection, jailbreak, and policy-bypass taxonomies.
  • Reliable async availability for at least 10 hours per week.

Skills

  • Adversarial prompting
  • Red-team analysis
  • Policy taxonomy
  • Failure documentation
  • AppSec Security evaluation
  • Security review
  • Adversarial testing
  • Threat modeling
  • AppSec

Work Model

  • Remote (remote)
  • Remote ยท Independent specialist contractor
  • Employment type: CONTRACTOR
  • Applicants must be authorized to work from US.

Compensation

Hourly rate confirmed after the interview process.

Application Process

  • Apply through AuraOne's specialist intake for role-specific routing and review.
  • Final project scope, schedule, and contractor terms are confirmed before placement.

Example Tasks

  • Construct a 5-turn adversarial conversation that bypasses a specific policy clause and write up the patch ticket.
  • Score a model's defenses against a known jailbreak pattern across 20 variants.
  • Propose a new red-team rubric category after spotting an emerging attack vector.
  • Reproduce a failure another reviewer reported and confirm the severity tag.

Nice to Have

  • Background in offensive security, AppSec, or trust & safety operations.
  • Experience publishing or reproducing public adversarial-ML research.
  • Multilingual fluency for cross-language attack testing.

Why this role matters

Adversarial evaluation is how AuraOne hardens AI models before they ship to customers.

Similar Jobs