Frontier AI Safety Evaluator
Mercor
pMercor seeks experienced AI Safety Practitioners to evaluate frontier AI models across complex, policy-sensitive topics.brZorg ervoor dat u solliciteert met alle gevraagde informatie, zoals uiteengezet in het functieoverzicht hieronder.brYou will assess AI-generated responses for safety, factual accuracy, and alignment, and contribute to improving model behavior through structured evaluations. xgiwjmb /ppYou will define and apply evaluation rubrics for RLHF and SFT, identify unsafe outputs, and provide actionable feedback.brThe role collaborates with AI researchers and safety teams on ongoing evaluation programs and /pbr
- pMercor is seeking experienced AI Safety Practitioners to evaluate frontier model safety, quality, and alignment across policy-sensitive topics.brAlle kandidaten dienen de volgende functieomschrijving en informatie zorgvuldig te lezen alvorens te solliciteren.brYou will...Aanbevolen
- pObsidian is seeking experienced AI Safety Practitioners in Amsterdam to evaluate frontier models for safety, quality, and alignment across nuanced topics.brLees verder om te ontdekken wat u nodig heeft om te slagen in deze functie, inclusief vaardigheden, kwalificaties...Aanbevolen
- pObsidian in Amsterdam seeks experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing.brDenkt u de juiste... ...prompts, uncover model weaknesses, and evaluate AI behavior across high-risk topics. /ppJoin a team...Aanbevolen
- pWe are seeking experienced bAI Safety Practitioners /b to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve...Aanbevolen
- pMercor in Amsterdam collaborates with a leading AI research lab on Frontier Code Agents, evaluating and improving frontier AI coding models through structured technical assessments.brLigt uw cv klaar? Zo ja, en bent u ervan overtuigd dat dit de functie voor u is, zorg...Aanbevolen
- ...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position: AI Safety Practitioner Type: Contract... ...Remote Role Responsibilities Evaluate AI-generated responses for safety, factual...Werken op afstandMet contractZomerbaan
- pMercor collaborates with a leading AI research lab on a Frontier Code Agents project to assess infrastructure engineering tasks and model outputs.brLigt uw cv klaar? Zo ja, en bent u ervan overtuigd dat dit de functie voor u is, zorg er dan voor dat u zo snel mogelijk...
- ...technical talent with leading AI research labs. Headquartered in... ...Dorsey . Position: AI Safety Red Teamer Type: Contract... ...prompts to stress-test frontier AI models . Identify jailbreaks... ..., and policy failures. Evaluate model robustness across...Werken op afstandMet contractZomerbaan
- pWe are seeking experienced bAI Safety Red Teamers /b to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-...
- pCognizant is hiring a Frontier Engineer in the Netherlands to design, deploy and continuously improve AI-native systems for leading organizations. You will work in a hybrid... ...and LLMOps, with a focus on responsible AI, safety, and production readiness. /p #J-18808-Ljbffr
- ph3About the Role /h3 ul liMercor is partnering with a leading AI research lab to support a Frontier Code Agents project. /li liContributors help evaluate and improve frontier AI coding models through structured technical assessments. /li liThe work focuses on realistic...
$400 per dag
...connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our... ...: Remote Role Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review...Werken op afstandMet contractZomerbaan$20 per uur
...Dutch AI Product Evaluator The Project Productive Playhouse is building a talent pool of Dutch speakers for an upcoming project testing... ...directly with various AI models to assess their capabilities, safety, and helpfulness. Your insights and data will directly...Met contractVoor uitvoerders- ...creative and technical talent with leading AI research labs. Headquartered in San... ...Dutch (Netherlands) Generalist Expert — AI Safety Type: Contract Compensation... ...reasoning behind your judgments. Evaluate and strengthen AI model handling of sensitive...Werken op afstandMet contractPer direct beginnenZomerbaan
- ...creative and technical talent with leading AI research labs. Headquartered in San... ...Dutch (Netherlands) STEM Expert (PhD) — AI Safety Type: Contract Compensation... ...specialized scientific topics. Evaluate and annotate model responses for scientific...Werken op afstandMet contractPer direct beginnenZomerbaan
- ...based in the Netherlands to help make advanced AI models safer. You\u2019ll apply your scientific expertise to evaluate and strengthen how these models handle specialized... .../p /li /br lipSound judgment around scientific safety and the responsible handling of dual-use...Per direct beginnen
- ...Learning and Inference Research, AI Alignment Science /b role who... .... /ppThis role will pursue frontier, foundational research that shapes... ..., agent scaffolding and evaluation, multi-agent systems, memory and... ...interpretability, robustness, or safety. /liliPh.D. in a relevant area...Stage
- pMercor is seeking experienced musicians to evaluate generative music AI models in partnership with a leading AI lab.brAlle kandidaten dienen de volgende functieomschrijving en informatie zorgvuldig te lezen alvorens te solliciteren.brYou will assess AI-generated music...Per direct beginnen
- pMercor is hiring experienced musicians to evaluate generative music AI models, in partnership with a leading AI lab.brSolliciteer (door op de betreffende knop te klikken) na het doornemen van alle gerelateerde vacature-informatie hieronder.brYou will assess AI-generated...Per direct beginnen
- pMercor is seeking experienced musicians to evaluate generative music AI models in collaboration with a leading AI lab.brVoor de volgende functie kunnen diverse soft skills en ervaring vereist zijn. Zorg ervoor dat u het onderstaande overzicht zorgvuldig controleert.brYou...
- pMercor is hiring experienced Musicians to evaluate generative musical AI models in partnership with a leading AI lab, with part-time hours. /pbrVaardigheden, ervaring, kwalificaties: als u de juiste match bent voor deze functie, zorg er dan voor dat u vandaag nog solliciteert...Met contractParttimePer direct beginnen
- pMercor is seeking experienced music producers and audio engineers to evaluate generative music AI models in partnership with a leading AI lab.brKandidaten dienen de tijd te nemen om alle onderdelen van deze vacature zorgvuldig te lezen. Gelieve snel te solliciteren.brYou...Met contractPer direct beginnen
- ...review complex Excel workbooks, reports, and presentations, aligning outputs with Fortune 500 standards. /ppThe role emphasizes evaluating AI-generated documents, providing feedback, and collaborating on AI prompts. Strong communication and data modelling skills are essential...Met contractOp afstand werken
- pYO AI Labs is seeking Turkish Bilingual Experts to support a language and AI training project.brLees de informatie in deze vacature... ...wat er van potentiële kandidaten wordt verwacht.brYou will evaluate Turkish audio for nativeness, fluency, pronunciation, and intonation...Voor uitvoerdersOp afstand werken
- pAccenture Netherlands in Amsterdam seeks an Agentic AI Quality Engineer to lead quality and performance evaluation for complex AI systems. You will drive testing, validation, and governance for safe, scalable AI deployments within Digital Core. /ppYou will work with quality...
- pAccenture the Netherlands is seeking a Senior ML QA Engineer to lead quality and performance evaluation across conversational AI and multi-step workflows. You will work with quality engineers, software engineers, AI specialists, and business consultants to embed testing...
- pObsidian seeks PhD-level scientists based in the Netherlands to help make advanced AI models safer. You’ll apply your scientific expertise to write prompts in Dutch and evaluate the accuracy and usefulness of model replies while ensuring responsible handling of sensitive...
- ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Dutch Type: Contract Compensation...Werken op afstandMet contractZomerbaan
- ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Dutch Type: Contract Compensation...Werken op afstandMet contractZomerbaan
- ...wat er van potentiële kandidaten wordt verwacht.brpAs a Frontier Engineer, you'll help build AI that moves the world forward. Working at the frontier... ...coding agents, applying human judgment to validate quality, safety and accuracy. /pbr/lilipCreate scalable RAG...
Wilt u meer vacatures ontvangen?
Abonneer u om vacatures voor Frontier AI Safety Evaluator te ontvangen. Solliciteer als eerste!

