01Overview
Role at a Glance
Role: AI Quality Engineer (Playwright Automation)
Experience: 36 Years
Engagement: Full-Time | On-Site / Hybrid
Domain: AI-First Enterprise Procurement SaaS
Core Stack: Playwright TypeScript/JavaScript Python REST APIs SQL
Reporting To: Head of Quality Engineering
Team Context: Embedded in AI Product Squads
THE OPPORTUNITY
We are looking for a seasoned AI Quality Engineer with 36 years of hands-on experience to join our Quality Engineering team. This is not a conventional QA role. You will be embedded in our AI product squads, designing and executing intelligent test strategies that validate not just functionality, but the behaviour, accuracy, fairness, and reliability of AI-driven features on a live enterprise SaaS platform.
You won't just test software. You'll safeguard the intelligence powering enterprise procurement for the world's leading organizations.
KEY RESPONSIBILITIES
AI & Intelligent Feature Testing
Design and execute end-to-end test strategies for AI-driven procurement features including supplier recommendations, spend classification, contract risk analysis, and demand forecasting.
Develop robust test suites using Playwright to validate AI model outputs, UI behaviours, and data flows across the full application stack.
Define and implement test cases for non-deterministic AI features including boundary testing, edge case exploration, hallucination detection, and bias assessment.
Build evaluation frameworks to measure and track model accuracy, consistency, recall, and precision across releases.
Collaborate with AI/ML engineers and data scientists to validate training data quality, model versioning, and inference pipeline correctness.
Automation Framework Engineering
Architect, build, and scale a Playwright-based automation framework (TypeScript/JavaScript) supporting UI, API, and component-level testing.
Design Page Object Models (POM), fixtures, and reusable test utilities that enable rapid test authoring across multiple product areas.
Implement self-healing selectors and intelligent wait strategies to handle AI-driven dynamic content and async rendering.
Integrate test suites into CI/CD pipelines (GitHub Actions, Jenkins, or Azure DevOps) to enable shift-left, continuous quality enforcement.
Maintain test result dashboards, flakiness tracking, and failure triage workflows to ensure signal-to-noise discipline in the automation suite.
API & Data Validation
Validate REST and GraphQL API contracts through Playwright's API testing capabilities, ensuring AI service integrations behave correctly under all conditions.
Write SQL queries to verify data integrity across procurement entities POs, invoices, supplier records, contracts pre- and post-AI processing.
Design data-driven test scenarios leveraging realistic procurement datasets to stress-test AI features at scale.
Perform response schema validation, latency benchmarking, and error boundary testing on AI inference APIs.
Performance, Security & Reliability
Lead performance testing initiatives to ensure AI-powered features meet SLA requirements under realistic enterprise load conditions.
Collaborate with security teams to perform penetration and vulnerability testing on AI API endpoints and data pipelines.
Design and execute chaos and resilience tests to validate system behaviour during AI model degradation or outage scenarios.
Monitor production AI model drift and work with MLOps teams to define quality gates for model redeployment decisions.
Quality Leadership & Culture
Champion a quality-first engineering culture by embedding testing discipline into sprint ceremonies, design reviews, and definition of done.
Mentor junior QA engineers on Playwright, automation best practices, and AI-specific testing methodologies.
Produce comprehensive test plans, risk assessments, and quality reports for product leadership and enterprise stakeholders.
Continuously evaluate and introduce next-generation testing tools, AI-assisted test generation platforms, and quality intelligence solutions.
KEY SKILLS
Playwright & Automation (Must-Have)
36 years of hands-on Playwright experience with TypeScript or JavaScript portfolio or GitHub evidence strongly preferred.
Deep expertise in Playwright features: network interception, storage state, fixtures, parallel execution, trace viewer, and visual comparison.
Strong understanding of test design patterns: POM, screenplay pattern, data-driven, and BDD with Cucumber or Gherkin.
Experience building and maintaining large-scale automation suites (500+ tests) with low flakiness and high maintainability.
Proficiency in CI/CD integration configuring Playwright in GitHub Actions, Jenkins, Azure Pipelines, or equivalent.
AI & ML Quality Expertise (Must-Have)
Demonstrated experience testing AI/ML-powered features NLP, recommendation engines, classification models, or generative AI outputs.
Understanding of AI quality concepts: hallucination, model drift, data bias, fairness metrics, confidence .