Sottava
Scanning 1,123 companies · 22,552 open jobs · last pass 46 min ago

← Sottava, jobs the hour they open

8 mo agofound 5 d ago

Prompt Engineer (LLM Systems, Evals & Safety)

webook.com·Amman, Jordan·via workable
remotemid5+ yrsFull timepythontypescript
What the posting is about

Design high-quality prompts for LLM features. Own evaluation and improvement. Collaborate with engineers for integration.

Read out of the posting
Levelmid
Experience asked5+ years
EmploymentFull time
LocationAmman, Jordan
RemoteYes
Visa sponsorshipNot stated
SalaryNot published, and most postings do not
Posted2026-01-12
Found viaworkable, direct from their system

We saw it 9 months after it went up.

The posting, as the company wrote it
Experience: Mid-Senior level Education: Bachelor's Degree Do you want to love what you do at work? Do you want to make a difference, an impact, and transform peoples lives? Do you want to work with a team that believes in disrupting the normal, boring, and average? If yes, then this is the job you are looking for , webook.com is Saudi’s #1 event ticketing and experience booking platform in terms of technology, features, agility, revenue serving some of the largest mega events in the Kingdom surpassing over 2 billion in sales. Role Overview Design high-quality prompts, system instructions, and tooling that make our LLM features accurate, safe, and cost-effective. You’ll own evaluation, prompt versioning, and continuous improvement. Key Responsibilities : Author, refactor, and chain prompts (system/tool/policy) for varied tasks. Create offline/online evaluation harnesses (rubrics, golden sets, metrics). Build prompt libraries with versioning, A/B testing, and telemetry. Reduce hallucinations via verification, constrained decoding, and tool use. Implement safety: jailbreak/prompt-injection tests, content policy checks, PII handling. Partner with engineers to integrate prompts into production features. Requirements Demonstrated prompt design across multiple task types and models. Experience building eval datasets and automated scoring (e.g., accuracy, faithfulness, utility, cost/latency). Familiarity with retrieval-augmented generation concepts and tool/function calling. Strong scripting (Python/TypeScript) for data prep, evals, and analysis. Clear writing; ability to translate business goals into measurable prompt specs. Nice-to-Haves Experience with LangChain/LLM orchestration, vector stores, and rerankers. Knowledge of safety tooling and red-teaming techniques. Experiment platforms (feature flags, A/B tests), analytics.

Copied from webook.com’s own board, not rewritten. Original ↗

Also open at webook.com

Why this page exists

We read companies’ own hiring systems every hour, 1,123 of them, and show a job the hour it opens instead of when a job board gets around to indexing it. We saw it 9 months after it went up.

The feed is free. No card, no trial to expire.