About Me
I spent ten years learning why systems fail. I now do that for AI.
I'm Manvi Sharma — an applied AI engineer focused on making agentic and LLM systems reliable enough to ship inside regulated enterprises. For nearly a decade I built and validated mission-critical systems at Workday, AWS, IBM, and British Telecom, where the job was proving that complex software behaves correctly before it reaches production. I now bring that same reliability discipline to AI: the model is rarely the hard part — earning enough trust that a real organization will run the system is.
I've shipped across regulated domains — industrial & chemical manufacturing, enterprise HR & financials, healthcare, and telecommunications — each with its own auditors, failure costs, and definition of "safe." I hold a Master of Business Analytics (MBAN) from UBC Sauder, where I also served as a Graduate Teaching Assistant, and I collaborate on research applying AI to surgical decision-making.
My Journey
Professional Background
Research Collaboration
Personal Interests & Hobbies
Fun Facts About Me
Technical Focus Areas
Agentic AI & LLM Systems
- · Multi-agent orchestration: Plan-and-Execute, supervisor/worker patterns, tool use & function calling
- · RAG & retrieval: hybrid search (Azure AI Search), vector stores, chunking, grounding & citation
- · LLM engineering: LangChain/LangGraph, structured outputs, prompt & context engineering
- · Fine-tuning & NLP: transformers, BioBERT, PEFT/LoRA, named-entity recognition
- · ML/DL foundation: PyTorch, TensorFlow, CNNs, LSTMs, ensembles, computer vision, time-series anomaly detection
Reliability, Evaluation & Platform
- · Evaluation: eval harness design, LLM-as-judge, golden datasets, regression suites, offline & online eval
- · Observability & guardrails: tracing, drift monitoring, attribution/hallucination checks, human-in-the-loop
- · Data & pipelines: Databricks, Spark, ETL & crawling, feature pipelines, SQL, Pandas
- · Cloud & MLOps: Azure, AWS, GCP · Docker, CI/CD, model versioning, FastAPI, Streamlit
- · Languages: Python, SQL, JavaScript/TypeScript, Java
Education
Certifications
- · NLP, Data Science for Business – DataCamp
- · Intermediate & Introduction to Python – DataCamp
- · Intro to GenAI, LLM, Responsible AI – Google Cloud
- · AI for Everyone – Andrew Ng (DeepLearning.AI)
Achievements & Recognition
Speaking
Available for talks and panels on AI reliability, evaluation, and shipping ML inside regulated organizations. Prior speaking includes:
- · QA Global Summit 2022
- · Globant QUANTANXT 2021
AI Solution Architecture & Assurance
Alongside my full-time focus, I take on independent consulting as an AI solution architect — designing and de-risking AI systems for teams that need them to actually work in production.
Work spans solution architecture, evaluation harness design, agent reliability reviews, and readiness assessments for AI moving from pilot to production.
Available for consulting