Base Career helps you apply smarter for this job.
Key skills for this role
Design, implement, and continuously refine system prompts, character behavior definitions, and guardrails for our AI character platform.
Build evaluation frameworks to measure character fidelity, response quality, and safety — establishing quantitative standards for what "in character" actually means across interaction types.
Architect and implement RAG systems to ground character responses in canonical knowledge, including embedding strategies, vector retrieval, and relevance evaluation.
Partner with creative and brand teams to translate character attributes — personality, voice, canon, tone, and boundaries — into precise, testable AI behavior and system requirements.
Lead prompt versioning, regression testing, and model evaluation practices to ensure character behavior remains consistent across provider updates and model changes.
Integrate with multi-provider LLM APIs (OpenAI, Google Gemini, etc.) and make principled selection decisions as the landscape evolves.
Evaluate and adopt emerging techniques in prompt engineering, agentic design, and LLM evaluation that meaningfully advance the quality and reliability of character experiences.
3+ years of experience building and shipping production LLM-powered systems.
Deep expertise in prompt engineering — system prompt design, few-shot techniques, chain-of-thought, multi-turn conversation management, and structured output.
Experience designing and implementing evaluation frameworks for LLM outputs, including adversarial and security testing.
Hands-on experience with RAG architectures: embedding models, vector stores, chunking strategies, and retrieval evaluation.
Proficiency with major LLM APIs (OpenAI, Google Gemini, Anthropic, or equivalent) and experience making principled decisions between them.
Strong Python skills and comfort working in AWS environments.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
Montréal, CAN
Montréal, CAN
Renton, USA
Boston, USA
Bogota, USA
Renton, USA
London, GBR
London, GBR
Familiarity with how voice AI services work (TTS, STT) — enough to understand how spoken delivery shapes prompt and response design.
The ability to communicate AI system behavior clearly to creative, brand, and non-technical collaborators.
Familiarity with guardrails frameworks (NVIDIA NeMo Guardrails, Guardrails AI, or equivalent).
Conceptual understanding of fine-tuning and RLHF, even if not a primary area of focus.
Background in entertainment, gaming, or IP-driven product development.
The pay transparency range for this role is listed below. The hiring range will vary based on factors such as experience, skills, location and market conditions. Additionally, employees may be eligible for annual and long-term incentives as part of their overall compensation package.
Employees may be eligible for annual and long-term incentives as part of their overall compensation package, depending on role, location, and eligibility. Benefits and programs may include:
Health & Wellness
Time Off to Recharge
Financial Well-being
Life & Family Support
Volunteer and Community Initiatives
Learning & Development
Exclusive Perks
Please review our Applicant Privacy Notice to learn how we collect, use, and protect your personal information in connection with the application process.
Mid · 3+ years experience
Apply faster on company sites with our extension.