Skip to content
PS
Available · Q3 2026
Portfolio · MMXXVI

PANKAJ PRATIMSARMAH

Senior Quality Engineer · Test Automation · AI / ML QA

Issue 01 · Volume 01

A working definition

I build the test automation that ships trustworthy software — across web, mobile, API, and AI. The work sits at the intersection of test engineering, product quality, and modern delivery — end-to-end QA strategies built from scratch, frameworks that survive contact with real teams, and a culture where quality is a product feature.

A working reel · 2020 — 2026

NoahBlockchain · CurrentGrowthDayEdtech · SaaSPrudentialInsurance · 14k staffGreyOrangeRobotics · LogisticsWells FagoBankingPSiDEOEvent technologiesCGIOil & GasLykaSocial MediaUpwork core teamMarketplace · Top ratedAI / ML ProductLLM Apps · AgentsPlaywrightTypeScriptCypressJavaScriptWebdriverIOMobile + WebAppiumAndroid · iOSPython · PytestAPI + LoadCI · GitHub ActionsParallel ShardsNoahBlockchain · CurrentGrowthDayEdtech · SaaSPrudentialInsurance · 14k staffGreyOrangeRobotics · LogisticsWells FagoBankingPSiDEOEvent technologiesCGIOil & GasLykaSocial MediaUpwork core teamMarketplace · Top ratedAI / ML ProductLLM Apps · AgentsPlaywrightTypeScriptCypressJavaScriptWebdriverIOMobile + WebAppiumAndroid · iOSPython · PytestAPI + LoadCI · GitHub ActionsParallel Shards

About · 01

A master of trustworthy software — fifteen years on.

SDET at noah.com working on blockchain-integrated products. Previously QA across fintech, marketplace, insurance, and edtech — through direct engagements and freelance platforms.

I'm Pankaj — a QA engineer and SDET based in Dubai. For the last 15+ years I've been the person teams call when they want their software to actually work in production, not just on staging.

My work sits at the intersection of test engineering, product quality, and modern delivery. I design end-to-end QA strategies from scratch, build the automation frameworks that hold them up, and ship alongside the engineers — Playwright and TypeScript when I can choose, Cypress, Selenium, WebdriverIO, or Appium when the stack calls for it.

Lately I've been deep in AI/ML QA: testing LLM applications, agent workflows, and the brittle seams where AI features meet real users. I'm equally comfortable with a GraphQL query, a JMeter load profile, or a flaky mobile test on BrowserStack.

I care about clean execution: small, fast feedback loops, frameworks that survive contact with real teams, and a culture where quality is a product feature, not an afterthought.

  • 15+ years in QA & test automation
  • SDET at noah.com — blockchain-integrated fintech
  • Built frameworks that cut manual testing by 60%
  • Polyglot: TypeScript, Python, Java, JavaScript

Expertise · 02

Four practices, one discipline.

I work where testing meets product. The work splits naturally into four overlapping practices — the seams between them are where the value lives.

Practice
Description
Stack
I

Test Automation

Read more

End-to-end frameworks in Playwright, Cypress, Selenium, and WebdriverIO. Page-object discipline, fixtures that survive, CI wired for parallel runs.

Playwright · Cypress · WDIO · Selenium

II

AI / ML QA

Read more

Strategy and tooling for LLM applications, agent workflows, and prompt quality. Where AI features meet real users — the brittle seams.

Dify · Ollama · Cursor · Eval frameworks

III

API & Performance

Read more

Contract testing, schema validation, load profiles, and the kind of slow-leak bugs that show up the day after launch.

Pytest · SuperTest · Locust · JMeter

IV

Mobile & Web

Read more

Cross-browser, cross-device, real-device on BrowserStack and Bitrise. The boring work that decides whether a release ships.

Appium · React Native · Flutter · Cypress

Table 01 · PracticesCareer

Career · 03

From manual QA to SDET — fifteen years of shipping.

The arc below is the same one most senior QA leads walk — but with the AI/ML inflection added in the last few years. Each role taught a different part of the craft.

A career · 2010 — Present

Originally a manual tester, I began writing automation at a small agency where the only QA tool that worked was the one I built myself. That early experience shaped how I work today: the framework is the product, not an afterthought. The career since has been a long, deliberate compounding — from Selenium + Java in 2013, to Cypress + JS in 2019, to Playwright + TypeScript in 2023, to the AI/ML work that dominates my current focus.

  1. 2025 — Present

    Full-time

    SDET (Software Development Engineer in Test)

    noah.com · Remote

    Owning end-to-end test automation for a fintech platform with blockchain-integrated features and AI-driven workflows.

    • Migrated legacy WebdriverIO test suites to Playwright using Cursor AI for intelligent code refactoring — 40% reduction in test execution time.
    • Built E2E coverage for blockchain-integrated features: smart contract interactions, wallet-based user flows, Ethereum testnet integrations.
    PlaywrightTypeScriptCursor AIGraphQL+8
  2. 2024 — 2024

    Contract

    Senior Quality Assurance Engineer

    Contract Role · Remote

    Embedded QA for an AI/ML product team building LLM applications, agents, and orchestration workflows.

    • Developed QA strategy and automated tests for LLM applications: chatbots, AI workflows, autonomous agents.
    • Performed manual QA on Dify (open-source LLM application orchestration) across multiple release cycles.
    PythonPytestLocustAI/ML+4
  3. 2022 — 2024

    Contract

    Senior QA Engineer

    Upwork core team · Remote

    Domain QA specialist across multiple product launches, leading both manual and automation workstreams.

    • Owned end-to-end QA for new product launches and feature rollouts.
    • Led and scaled both manual and automation testing teams.
    RubySeleniumSelenium GridGraphQL+4
  4. 2022 — 2024

    Contract

    QA Automation Engineer

    Grey Orange Inc. · Remote

    Central QA contact for multiple clients, building test automation for robotics-platform-adjacent products.

    • Created a test automation framework combining SuperTest for APIs and Appium Flutter Driver for mobile native apps.
    • Built a mobile UI automation framework with Appium + Java, integrated into CI.
    JavaJavaScriptFlutterAppium+8
  5. 2022 — 2023

    Contract

    QA Lead

    Chef Booking Platform · Remote

    End-to-end QA for a marketplace product — manual testing, no-code UI automation, and PR-level defect prevention.

    • Owned end-to-end manual testing of a marketplace platform across web and mobile.
    • Built and maintained no-code UI automated tests to widen coverage without inflating the automation codebase.
    Node.jsReactSwagger
  6. 2022 — 2022

    Contract

    QA Engineer

    Prudential · Remote

    QA on a 360-degree feedback tool revamp for a 14,000-employee Asian insurance provider.

    • Detailed test case writing, E2E manual testing, and requirement review.
    • Assisted the product owner with design reviews and acceptance criteria shaping.
    PostmanMySQLTrelloFigma+1
  7. 2020 — 2022

    Contract

    QA Engineer

    GrowthDay · Remote

    Owned QA across web, mobile (Android & iOS), and back-end — built a 500+ case API automation suite.

    • Owned end-to-end QA for web, mobile (Android & iOS), and back-end apps.
    • Developed an API test framework with SuperTest — 500+ test cases automated.
    CypressJavaScriptSuperTestReact+8

Toptal · 04

Top 3% talent, vetted by Toptal — and available for senior SDET engagements.

One of the world's most selective talent networks. Book a senior QA engineer the same way you'd book a top-tier consultant — without the recruiting overhead.

TOP 3% TALENT

Vetted byHire me
Badge 01 · 01Toptal · Verified
Hire me via ToptalOr use my referral link

Why Toptal · A short note

I've joined the Toptal network for the same reason clients use Toptal to find me: the signal-to-noise ratio is the best in the industry. The screening is what makes the marketplace work, and the trial period is what makes it risk-free.

  • Vetted, not self-claimed

    Toptal screens every engineer through a multi-stage process. The Top 3% figure is theirs, not mine.

  • Engagements start in days

    No sourcing, no recruiting, no HR paperwork. You get a matched expert and a signed SOW inside a week.

  • No-risk trial period

    If the first two weeks don't click, Toptal refunds the engagement. I've never had to use it.

  • Direct contracting

    NDAs, IP assignment, and invoicing are handled by Toptal. We both focus on the work.

Prefer email? — direct line always opencontact@pankajsarmah.com
Network
Top 3% global
Engagement model
Hourly · Part-time · Full-time
Trial period
14 days, no-risk
Profile
QA · Test Automation · AI/ML

Selected Engagements · 05

From solo assignments to long-running embedded engagements.

A working list of the teams I've shipped with. Most engagements started as a contract and ended as a long-term relationship.

Client
Domain
noah.com
Fintech · Blockchain · Current
GrowthDay
Edtech · SaaS
Prudential
Insurance · 14k staff
GreyOrange
Robotics · Logistics
Upwork core team
Marketplace · Top rated
AI / ML Product Team
LLM Apps · Agents
Toptal Network
Top 3% QA Talent

Table 02 · Selected engagements · 2020 — 2026

Selected Work · 06

The repos are the receipts.

Every project below is on GitHub — built around a problem I actually wanted to solve. Sorted with the current focus on top.

No.

01

Automation Framework · 2024

Playwright TypeScript E2E Framework (Sauce Demo)

Featured

Production-grade end-to-end test framework for Sauce Demo using Playwright + TypeScript with the Page Object Model.

TypeScriptPlaywrightPOMGitHub Actions
Repo2024

No.

02

AI Tool · 2025

AI Test Data Generator (Ollama)

Featured

Generate realistic, scenario-aware test data locally using Ollama — no PII leakage, no cloud round-trips.

PythonOllamaLLM
Repo2025

No.

03

Automation Framework · 2024

Excel Online Automation Framework

Playwright + TypeScript framework purpose-built for automating Excel Online's Today() function and formula behaviour.

TypeScriptPlaywrightMicrosoft Graph
Repo2024

No.

04

Automation Framework · 2023

Python Selenium Framework (BookCart)

Python + Selenium framework for the BookCart reference app — clean structure, pytest reporting, and POM patterns.

PythonSeleniumPytest
Repo2023

No.

05

Test Utility · 2024

Shadow DOM: Selenium vs Cypress vs Playwright

Side-by-side comparison of how Selenium, Cypress, and Playwright handle Shadow DOM piercing — with runnable code.

PythonSeleniumCypressPlaywright
Repo2024

No.

06

Automation Framework · 2024

Playwright Luma Automation

Playwright automation suite for a Luma-style event flow.

TypeScriptPlaywright
Repo2024

No.

07

Automation Framework · 2023

Cypress Auto Starter

A clean Cypress JavaScript starter for greenfield test automation projects.

JavaScriptCypress
Repo2023

No.

08

Automation Framework · 2023

Python Selenium Starter

Minimal, no-frills Python + Selenium test automation starter.

PythonSelenium
Repo2023

No.

09

Mobile · 2023

WebdriverIO + Appium Android Starter

Appium tests for Android with WebdriverIO as the client — ready to point at your APK.

JavaScriptWebdriverIOAppiumAndroid
Repo2023

No.

10

Scraper · 2022

Simple Yellow Pages Scraper

Lightweight Python scraper that turns Yellow Pages search results into a clean Excel file.

PythonBeautifulSoupopenpyxl
Repo2022
Table 03 · 10 selected reposBrowse all

Stack · 07

A stack tuned for shipping, not résumés.

Tools I reach for under deadline pressure — and the ones I keep around because they pay off on the long arc.

Test Automation

01 / 07

Frameworks I reach for, in order of how often I open them.

  • Playwright
  • Cypress
  • Selenium
  • WebdriverIO
  • Appium
  • Cucumber
  • Pytest
  • SuperTest
  • Mocha
  • Chai
  • JUnit

Languages

02 / 07

Comfortable shipping production code — not just tests.

  • TypeScript
  • JavaScript
  • Python
  • Java
  • Ruby

Web & Mobile

03 / 07

Where the products I test actually live.

  • React
  • React Native
  • Node.js
  • Flutter
  • HTML/CSS
  • REST
  • GraphQL

API & Performance

04 / 07

Validating contracts and load behaviour before users find the seams.

  • Postman
  • Swagger / OpenAPI
  • Locust
  • Apache JMeter
  • k6 (basics)

DevOps & Cloud

05 / 07

Pipelines and infra I integrate with, not just touch.

  • GitHub Actions
  • Docker
  • Kubernetes (basics)
  • AWS (CLI, DynamoDB, AppSync, Step Functions)
  • GCP (basics)
  • BrowserStack
  • Bitrise

AI / ML QA

06 / 07

Where my current focus sits — testing LLM apps and AI workflows.

  • LLM App Testing
  • Dify
  • Agent Workflow Testing
  • Prompt Quality
  • Cursor AI
  • Test Data Generation (AI)

Data & Tooling

07 / 07

The glue that holds a test stack together.

  • PostgreSQL
  • MySQL
  • MongoDB
  • Kibana
  • Datadog
  • Jira
  • Notion
  • Figma
  • LaunchDarkly

Approach · 08

Six principles, fifteen years of compounding.

The rules I work by. They're not a manifesto — they're a pattern that survived contact with dozens of teams and products.

01Principle

The framework is the product

I build the test system, not the test cases. A good framework grows with the team; a good test case is just today's expectation captured. Invest in the system.

02Principle

Quality is a product feature

The cost of a defect in production is not a line item — it's a trust withdrawal. I treat testability as a first-class design constraint, not a downstream concern.

03Principle

Boring stacks, sharp edges

I prefer the tools I can debug at 2am. Playwright, TypeScript, Python, GitHub Actions, Postgres. The novelty budget goes to the AI work, not the rest.

04Principle

Small loops beat big releases

CI on every PR, nightly regression, weekly exploratory. The compounding cost of slow feedback is the silent killer of quality programs.

05Principle

Embedded, not outsourced

I work inside the team. Same Slack, same standups, same off-site. The QA lead who shows up is the one whose fixes get merged.

06Principle

Test the seams, not the surface

The interesting bugs are at the contract boundaries — third-party APIs, AI outputs, async queues. That's where the most leverage is, and where I focus.

Lens · 09

What's on the
desk right now.

A snapshot of the things I'm reading, building, and working on this season. Updated roughly monthly.

  • ReadingDesigning Data-Intensive Applications · Kleppmann01
  • BuildingAI test-data generator with Ollama (open source)02
  • WritingA long-form essay on LLM evaluation pipelines03
  • TeachingMentoring two junior SDETs through Toptal04
  • WatchingShadow DOM automation across 3 frameworks05
  • ListeningStrange arrangements, badly mixed coffee06

Colophon · 10

The bits that don't fit in a CV.

A small editorial aside — the human parts of the work that don't show up in a job description.

Origin

Assam, India · 1988

Home

Dubai, UAE · 2018 —

Languages

English · Hindi · Assamese

Coffee

Strong, no sugar

Family

Married · two kids · Eva & Neev

Side project

AI test-data generator · Ollama · Open source

Dispatches · 11

Notes from the test trenches.

I write occasionally — about test engineering, AI / LLM QA, product thinking, and the long arc of a QA career. No roundups.

Index · 3 essaysAll dispatches

Contact · 12

Engagement ps conversations.

Open to senior SDET, QA lead, and test-automation consulting engagements. Drop a line — happy to chat about scope, timelines, and the actual shape of the work.

Send a brief

Form · 01

One follow-up. No newsletter.