I've spent 15+ years making commercial enterprise software trustworthy, first as a System Verification Tester at IBM, then as a Senior Software Engineer at HCLTech building test frameworks, and release quality for commercially shipped enterprise products.
Now I bring that same discipline to AI systems: agents that call real tools, follow business rules and are tested like production software, not demos.
- π Building agentic AI apps with LangGraph, Claude API and MCP
- π§ͺ Testing LLM output the way I test APIs: assertions, edge cases, zero-escape goals
- π± Learning: RAG patterns, LLM evaluation, cloud-native deployment
- π¬ Ask me about: test strategy, CI/CD speed-ups, root-causing hard defects
|
Multi-agent ticket resolution A 5-node LangGraph
|
AI "Resolve with AI" for support desks Agents click Resolve with AI and Claude calls 4 MCP tools (
|
|
Ask your database in plain English A full-stack app that turns plain English into executable SQL via the Claude API, with generated SQL checked for correctness and result accuracy.
|
Spec-driven CI/CD, end to end Every pipeline stage is owned by a declarative spec file. The code must conform to the spec, not the other way around.
|
π§° More: MCP Document Tools server
A Python package that exposes document conversion and processing tools through an MCP server, so AI assistants can call them directly. Tools are typed Python functions with Pydantic Field descriptions for clean tool schemas.
| π€ AI / LLM |
|
| π» Languages |
|
| π§ͺ Testing |
|
| π CI/CD & DevOps |
|
| π APIs & Web |
|
| ποΈ Data |
|
| π€ Practice |
|
"If it isn't tested, it isn't done, and that includes the AI."
| Principle | In practice |
|---|---|
| π― Rules in structure, not hope | Agent guardrails enforced by graph topology and hard caps, not only by prompts |
| π§ͺ Test the output, not just the code | Assert generated SQL and LLM answers for correctness, like API contracts |
| β‘ Fast feedback wins | Parallel pipelines and auto-ticketing so defects surface in minutes, not days |
| π Find the real root cause | Binary-search isolation + log correlation before any fix ships |
