Available for AI/ML & Agentic AI roles

Yuvaraj
Sriramoju

I build AI systems that reason, retrieve, and act.

AI/ML Engineer with 5+ years shipping production AI across financial services and healthcare — multi-agent orchestration, RAG, LLM systems, and end-to-end MLOps.

Press ⌘ K to navigate this portfolio like a command center.

5+ yrs
production AI/ML
4
domains shipped
2
cloud AI certs
MS
in Artificial Intelligence
scroll
01 profile

Designing AI that ships — reliable, explainable, in production.

I'm an AI/ML engineer focused on Agentic AI and the systems that make large models useful in the real world. Over the last five years I've designed and scaled production AI across financial services and healthcare, from multi-agent orchestration to enterprise RAG pipelines.

My work sits where research meets operations: LangChain, LangGraph, and the Model Context Protocol for orchestration; vector databases and knowledge graphs for retrieval; and FastAPI, Kubernetes, and MLflow for the MLOps that keeps it all running and auditable.

I care about the parts people skip — governance, interpretability, and compliance — so the systems I ship hold up under GDPR, HIPAA, and FINRA scrutiny, not just in a demo.

Current focus

  • Multi-agent orchestration & tool calling
  • Model Context Protocol (MCP) integrations
  • Enterprise RAG with hybrid retrieval & reranking
  • Knowledge-graph-grounded reasoning (Neo4j)
  • LLM observability & drift monitoring
  • Responsible & explainable AI
02 interactive AI lab

See how I think about production AI systems.

Run a browser-only RAG simulation, inspect the pipeline stages, and explore case studies grounded in my production experience. This is an interactive portfolio demo—not a hosted language model.

rag_pipeline.py · local simulation
Intent router
Hybrid retrieval
Reranker
Generation
Guardrails
Run the pipeline to generate a grounded architecture response.

No prompt is sent to a server. The interaction runs entirely in your browser.

02.1 selected systems

Production case studies.

Financial services · Agentic AI

Underwriting & policy copilot

−45%manual review

Specialized agents orchestrated over enterprise APIs, policy documents, vector stores, and compliance checks.

LangGraphMCPRAGNeo4j
Enterprise RAG · Quality

Grounded document intelligence

+28%answer relevance

Hybrid retrieval, reranking, vector memory, and evaluation loops designed to reduce hallucinations.

FAISSRerankingEvaluationGuardrails
Healthcare · Predictive ML

Clinical risk intelligence

99.5%service uptime

Readmission, length-of-stay, clinical NLP, and multimodal CV pipelines deployed with compliance controls.

PyTorchSageMakerKubernetesSHAP
LLMOps · Reliability

Observability & model operations

−35%downtime

Telemetry, drift monitoring, automated retraining, and production controls for model and LLM services.

PrometheusGrafanaMLflowKubeflow
03 experience

Where I've built.

Sep 2024 — PresentDallas, TX

Generative AI Engineer

Corebridge Financial
  • Architected a production Agentic AI platform with MCP, LangGraph & LangChain — orchestrating specialized agents over enterprise APIs and vector DBs to automate underwriting, policy analysis, and compliance.
  • Built and fine-tuned GPT, BERT, T5 and diffusion models for financial document summarization — +32% accuracy, -45% manual review time.
  • Designed enterprise RAG pipelines with hybrid retrieval, reranking & vector memory — +28% answer relevance with fewer hallucinations.
  • Integrated Neo4j knowledge graphs with RAG for relationship-aware reasoning — +30% policy analysis accuracy.
  • Stood up LLM telemetry with Prometheus & Grafana, cutting downtime by 35%; automated monitoring & retraining via MLflow/Kubeflow.
Oct 2023 — Aug 2024Dallas, TX

Machine Learning Engineer

Tenet Health Care
  • Built models predicting readmission risk & length of stay from EHR and claims data — +28% risk stratification accuracy.
  • Developed a multimodal CV pipeline (CLIP, BLIP, PyTorch) for X-ray anomaly detection & radiology triage — +19% diagnostic accuracy.
  • Designed clinical NLP with BERT to extract diagnoses, meds & symptoms from physician notes — 87% F1.
  • Shipped end-to-end MLOps on SageMaker & Kubernetes with CI/CD and monitoring — 99.5% uptime.
  • Ensured HIPAA/PHI compliance with de-identification and SHAP/LIME explainability.
Mar 2022 — Jul 2023Bengaluru, India

Data Scientist

Kotak Mahindra Bank
  • Deployed credit scoring & loan-risk models (XGBoost, LightGBM) — +22% prediction accuracy, faster approvals.
  • Cut fraud false positives by 15% via supervised learning and feature-importance analysis.
  • Improved liquidity forecasting +25% with LSTM & Prophet time-series models.
  • Served risk & fraud models as REST API microservices for real-time decisioning; governed with SHAP/LIME under RBI & GDPR.
Aug 2020 — Feb 2022Bengaluru, India

Data Scientist

SG Analytics
  • Built end-to-end ETL on SQL, Python & Snowflake across finance/healthcare data — -32% refresh latency.
  • Shipped forecasting & churn models (ARIMA, regression) — +21% predictive accuracy.
  • Automated validation frameworks (Pandas/NumPy) — -45% manual QA effort.
  • Delivered Power BI & Tableau dashboards, cutting reporting turnaround from 3 days to under 1.
04 stack

Tools I reach for.

All capabilities

Agentic & Generative AI

LangChainLangGraphMCP Multi-Agent SystemsTool CallingRAG Prompt EngineeringFine-TuningGPT / BERT / T5 Diffusion Models

Machine Learning

Deep LearningNeural NetworksTime Series ClassificationClusteringFeature Engineering XGBoostLightGBMModel Optimization

Vision & NLP

CLIPBLIPOpenCVCNNs ViTsNERQuestion Answering Transformers

MLOps & Serving

MLflowKubeflowDockerKubernetes CI/CDFastAPIPrometheusGrafana

Data & Vector Stores

PostgreSQLMongoDBSnowflakeNeo4j FAISSPineconeChromaDBWeaviate PySparkAirflow

Cloud & Languages

Azure MLAWS SageMakerVertex AI PythonSQLRJavaBash
05 certifications
AWS

ML Engineer — Specialty

Amazon Web Services

AZ

Azure AI Engineer Associate

Microsoft

06 contact

Let's build something intelligent.

Open to AI/ML and Agentic AI engineering roles. The fastest way to reach me is email — I usually reply within a day.

Uses your email app; no form data is stored.
YS

Yuvaraj Sriramoju

AI/ML Engineer · Agentic AI · Generative AI · MLOps
Dallas, TX
5+ years
4 companies
2 certs
Connect on LinkedIn

A styled preview of my profile — the button opens my real LinkedIn in a new tab.

View full profile →