Engineering Manager at Iterate.ai · I lead the 12-member team building Generate, an enterprise AI platform
Generate runs as multi-tenant SaaS and also installs air-gapped inside regulated industries, as one system in both places. We build it alongside NetApp, AMD, Intel, IBM, and HP. It has won AI Product of the Year from both Pinnacle and TMC, and put Iterate.ai on the CRN AI 100.
- RAG and multi-agent orchestration on LangGraph, with 94% extraction accuracy on documents that break naive RAG
- Multi-cloud, on-prem, and edge deployment, including quantized models running fully local on an Intel AI PC
- Event-driven backend on Kafka, FastAPI, and Kubernetes, holding 99.8% uptime at 50,000+ requests/second
- Tenant isolation, with permission-filtered vector search and sandboxed code execution
- Engineering leadership: hiring, mentorship, and roadmap for a team grown from 6 to 12
More on the platform work
- Deployment targets are AWS and IBM Cloud, on-prem racks, and quantized models (GGUF, INT8/FP16) on Intel, AMD, and NVIDIA.
- The backend is observed through OpenTelemetry and Prometheus/Grafana.
- SSO identities resolve to POSIX UID/GID, so vector search is permission-filtered before it runs. Model-generated code executes under gVisor or Kata.
- The team includes 2 tech leads and 2 project managers, and I have promoted engineers into senior roles.
| Project | What it is | Traction |
|---|---|---|
| Stepgate | MCP server that shows an agent one step at a time and moves on only when mechanical checks pass | |
| MedBERT | Biomedical language model for named entity recognition | |
| NERP | Python framework for transformer-based named entity recognition |
- MASc in Electrical & Computer Engineering, McMaster University, and BSc (Hons.) in Computer Science & Engineering, University of Moratuwa
- 254+ citations across NLP, biomedical NER, and low-resource languages (Tamil and Sinhala)
- Co-inventor on 4 filed US patents in document extraction and multi-agent AI workflows
- Edge AI work shipped to 10,000+ Intel AI PCs and was demoed at the Intel Vision 2024 keynote
Happy to talk about applied AI, RAG at scale, edge and air-gapped deployment, or building engineering teams.



