Skip to main content

Automation architecture · Engineering consulting

Reliable automation.
From strategy to code.

I’m Dhiraj, an automation consultant with over a decade of experience across web, mobile, APIs, and enterprise workflows. I design test architectures, diagnose difficult failures, and bring the same discipline to AI agents.

Dhiraj Das

Dhiraj DasTest architecture → Agentic AI reliability

Selected work

Built to solve
real engineering problems.

From flaky browser tests and repeated login setup to visual regressions and AI-agent failures. These tools show how I investigate a problem, define its boundaries, and build a solution you can inspect.

Behind the Code

My work sits at the intersection of test automation and agentic AI reliability. After years of building test frameworks, stabilizing brittle browser flows, debugging CI failures, and turning ambiguous defects into reproducible evidence, I apply the same engineering discipline to AI-agent validation, run observability, and reliability tooling.

Agent testing is not just prompt evaluation. Reliable agentic systems need observability, replayable evidence, failure taxonomies, browser/runtime signals, safe redaction, and postmortems that explain risk. That is the bridge I am building through Agent Blackbox and related reliability tooling.

"Reliable agents need the same discipline that made reliable automation possible.|

Read full background
Dhiraj Das
🔬 Research Foundation

Finding Structure in Complex Systems

I co-developed the Triangle-Density based Clustering Technique (TDCT) to turn noisy spatial data into explainable structure. That same systems thinking now shapes how I design automation and diagnose unreliable AI agent behavior.

💡 See how this research connects to agent reliability →

Published research later expanded into a book on clustering concepts and techniques.

🎮 Interactive Game

Master Locator Strategies

Gamified learning experience to master XPath and CSS selectors. Solve puzzles, level up, and sharpen your automation skills.

📖 Engineering Playbook

Automation That Survives Reality

A practical guide to browser internals, eliminating flaky tests, maintainable architecture, and delivery backed by evidence. These are the engineering habits that now shape how I build reliable AI agent systems.

🌌 Agent Platform · 5.x Alpha

Starlight Protocol

A general-purpose agent platform that turns goals into inspectable outcomes. Coordinate domain agents, bound mission execution, and verify results across code, data, APIs, and more.

Why the shift is happening

Routine automation is becoming a commodity.

AI agents increasingly generate and execute the deterministic scripts that once required specialist effort. The durable engineering problem is proving that autonomous work is correct.

Why my background wins

Automation discipline is the advantage.

Ten years of assertions, fixtures, failure isolation, CI diagnostics, and evidence capture transfer directly to agents that call tools, edit files, and make decisions.

Where I am building

Agentic AI reliability is the next layer.

I build local-first run capture, validation harnesses, replay, redaction, and failure postmortems so agentic systems can be trusted under production pressure.

Automation Roots, Agentic AI Direction

Ten years of automation work shaped the habits I now bring to agents: observe the run, control the inputs, isolate the failure, and prove the fix.

10+
Years
Experience
~35%
Avg
Efficiency Boost
High
Impact
Cost Savings
7+
Years
Global Exp

Career Timeline

Automation Consultant

Present

Consulting on automation strategy and quality systems while independently building reliability tooling for AI-assisted engineering: agent run capture, validation workflows, and failure postmortems.

Extending automation discipline into observable, repeatable, evidence-backed agent workflows

  • Designing Python-first tools for agent diagnostics and postmortems as an independent focus.
  • Applying CI, browser automation, and failure-triage patterns to AI-agent runs.
  • Defining guardrails for reliable local-first and AI-assisted workflows.
Hover for details

Senior Automation Developer

7 Years

Built and stabilized large web, API, and mobile automation programs across high-pressure delivery environments.

Reduced flaky failures by 70% across 200+ test suites

  • Managed complex web, API, and mobile automation suites.
  • Worked with diverse tools like UFT, AutoIt, and QF Test.
  • Delivered critical automation solutions for major corporate clients.
  • Gained deep Airline Domain Expertise.
Hover for details

Junior Automation Developer

3 Years

Built the foundations: reliable Selenium suites, CI integration, maintainable test design, and close QA-engineering collaboration.

Compressed manual regression from 2 weeks to 3 days

  • Developed test scripts using Selenium WebDriver and Java.
  • Learned CI/CD integration and version control best practices.
  • Collaborated with QA teams on manual-to-automation transition.
  • Built foundational expertise in test design patterns.
Hover for details

Tech Arsenal

Pick the reliability gap: opaque agent runs, flaky CI, Cloudflare walls, login overhead, visual drift, or GenAI UIs. I build tools for the places normal automation and naive AI workflows break.

Recent Insights

Latest notes on automation, agentic AI, reliability, and engineering

Your Test Suite Does Not Need Another Dashboard: Local Test Pulse for Omarchy

Local Test Pulse reads the report your test runner already produced and turns it into a trustworthy Omarchy bar signal. ...

Read Article

A Security Indicator Must Refuse to Lie: Building Omaudit Status for Omarchy

Omaudit already scans plugin capabilities. Omaudit Status adds a bounded desktop signal, then treats every green or ambe...

Read Article

Stop Teaching Browser Agents the Same Workflow Twice

When a browser path already worked, rediscovery is waste. Flow2Skill treats a Playwright demonstration as compiler input...

Read Article