Base Career helps you apply smarter for this job.
Key skills for this role
You’ll be the safety net and enabler of speed at Vibecoderz. In a system where AI agents generate artifacts, mini-apps, and real-time interactions, quality isn’t just about correctness — it’s about trust.
As the founding QA Automation Engineer, you’ll design frameworks that validate not just code but also AI outputs. From traditional E2E flows to specialized agent evaluation pipelines, you’ll ensure that everything shipping to production meets our gold standard.
This role is about building autonomous, self-healing QA systems that scale with our multi-agent platform. You’ll work hand-in-hand with PM, FE, BE, AI, and DevOps to embed quality into every phase of the pipeline using Linear (execution), Notion (test cases/playbooks), and GitHub (CI/CD integration) as your tools of record.
You’ll be the safety net and enabler of speed at Vibecoderz. In a system where AI agents generate artifacts, mini-apps, and real-time interactions, quality isn’t just about correctness — it’s about trust.
As the founding QA Automation Engineer, you’ll design frameworks that validate not just code but also AI outputs. From traditional E2E flows to specialized agent evaluation pipelines, you’ll ensure that everything shipping to production meets our gold standard.
This role is about building autonomous, self-healing QA systems that scale with our multi-agent platform. You’ll work hand-in-hand with PM, FE, BE, AI, and DevOps to embed quality into every phase of the pipeline using Linear (execution), Notion (test cases/playbooks), and GitHub (CI/CD integration) as your tools of record.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
, IND
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Automated framework live for FE (Next.js) and BE (FastAPI).
AI TutorAgent “Text → Course” flow covered by golden set evals.
CI/CD pipeline enforces automated test runs before deploy.
85% test automation coverage across platform.
<1% escaped defects in production.
AI agent eval accuracy >90% on golden set outputs.
System proven under 100K concurrent users with <1% error rate.
Experience with LLM/AI output evaluation (evals).
Prior startup or founding engineer experience.
Knowledge of security testing (XSS, prompt injection defense).
Contributions to open-source QA frameworks.
Frontend: Playwright, Cypress
Backend: Pytest, Postman/Newman
AI Evals: Custom golden set framework, LangSmith/Langfuse for prompt monitoring
CI/CD: GitHub Actions integration
Observability: OpenTelemetry hooks into test reports
Infra: Cloud Run staging environments for test runs
Objective: Validate ability to design automated test frameworks that cover FE, BE, and AI outputs.
Frontend (FE) E2E Test Write Playwright/Cypress tests for: User onboarding (sign-up with Google). Chatting with TutorAgent mock API. Rendering a Quiz block with answer validation.
Write Playwright/Cypress tests for: User onboarding (sign-up with Google). Chatting with TutorAgent mock API. Rendering a Quiz block with answer validation.
User onboarding (sign-up with Google).
Chatting with TutorAgent mock API.
Rendering a Quiz block with answer validation.
Backend (BE) API Test Write Pytest suite for FastAPI service: /generate-course endpoint (input: “Teach React Hooks”). Assert valid JSON structure for output.
Write Pytest suite for FastAPI service: /generate-course endpoint (input: “Teach React Hooks”). Assert valid JSON structure for output.
/generate-course endpoint (input: “Teach React Hooks”).
Assert valid JSON structure for output.
AI Agent Eval Create a golden set with 5 input prompts and expected TutorAgent responses. Automate comparison with tolerance for natural language variation.
Create a golden set with 5 input prompts and expected TutorAgent responses.
Automate comparison with tolerance for natural language variation.
CI/CD Integration Configure GitHub Actions to run FE + BE + AI eval tests on each PR. Block merge if any critical tests fail.
Configure GitHub Actions to run FE + BE + AI eval tests on each PR.
Block merge if any critical tests fail.
GitHub repo with FE, BE, and AI test suites.
GitHub Actions workflow file with integrated tests.
Golden set JSON for TutorAgent eval.
README documenting design choices and edge cases.
Automation Framework Quality (30%)
AI Evals & Innovation (25%)
CI/CD Integration (20%)
Test Coverage & Reliability (15%)
Documentation & Clarity (10%)
Verified company details for this employer are not available yet.
USD 24000-32000 yearly / year
Full-time
Senior · 10+ years experience
Hybrid
Apply faster on company sites with our extension.