← Sottava, jobs the hour they open
10 h agofound 1 h ago
SafetyTech Client #1 | Adversarial Task Writer for AI Security RL Gyms
Read out of the posting
LevelNot stated
Experience askedNot stated
EmploymentNot stated
LocationSerbia, Armenia, Bulgaria
RemoteYes
Visa sponsorshipNot stated
SalaryNot published, and most postings do not
Posted2026-10-05
Found viabreezy, direct from their system
We saw it 9 hours after it went up.
The posting, as the company wrote it
Employment: Full-Time
On behalf of SafetyTech Client #1 , SD Solutions is looking for a talented Adversarial Task Writer for AI Security RL Gyms.
SD Solutions is a staffing company operating globally. Contact us to get more details about the benefits we offer.
Responsibilities:
You design prompt injection scenarios in YAML, run them against frontier models, validate success rates, and submit passing tasks. 5 high-quality tasks per week (full-time equivalent). Per-task compensation, paid on acceptance.
Requirements:
Adversarial mindset: you think like an attacker and understand how to exploit an AI agent’s helpfulness, authority assumptions, or trust in its environment
Prompt injection expertise: direct (role-play, encoding, context flooding) and indirect/environment-embedded (poisoned tool responses, malicious content in documents, cross-context leakage)
Technical writing in YAML
Comfortable with Docker, CLI tools, and running systematic tests against multiple models
Domain realism in at least one vertical: e-commerce, finance, HR, enterprise SaaS, healthcare, travel
Background in pentesting, appsec, LLM security research, or red teaming strongly preferred
The Task
You build adversarial prompt injection tasks for Alice’s RL Gym platform. Each task is a self-contained YAML scenario simulating a realistic AI agent deployment, testing whether the agent can be manipulated into violating its safety policies.
About the company:
A company building specialized evaluation infrastructure for AI safety and robustness testing. Their platform simulates adversarial conditions used by AI development teams to validate agent behavior before deployment. Currently expanding a freelance contributor pool for scenario and environment development.
By applying for this position, you agree to the terms outlined in our Privacy Policy. Please take a moment to review our Privacy Policy https://sd-solutions.breezy.hr/privacy-notice , and make sure you understand its contents. If you have any questions or concerns regarding our Privacy Policy, please feel free to contact us.
Copied from SD Solutions’s own board, not rewritten. Original ↗
Also open at SD Solutions
HealthTech Client #3 | Sales & Business Development Representative1 h agoDataTech Client #1 | Technical Project Manager3 h agoWorkforceTech Client #1 | Call Scheduler4 h agoAdTech Client #1 | Software Developer7 h agoSD Solutions | Content Creator & Multimedia Specialist9 h agoGaming Client #1 | Senior UI / 2D Artist (EU)10 h ago
Why this page exists
We read companies’ own hiring systems every hour, 1,182 of them, and show a job the hour it opens instead of when a job board gets around to indexing it. We saw it 9 hours after it went up.
The feed is free. No card, no trial to expire.