Skip to content
View sailikhithk's full-sized avatar
🏠
Working from home
🏠
Working from home

Block or report sailikhithk

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
sailikhithk/README.md

Sai Likhith Kanuparthi

Typing SVG

I build the reliability layer for production AI in regulated industries. The model is the easy part. The hard part is the harness: routing across 30+ models (FacadeDriver), catching regressions before users do (23+ agent versions, 1,690 ground-truth samples), detecting synthetic media (SAI, 6 signals), and surviving in environments where failures have real consequences. Shipped in pharma (21 CFR Part 11), finance (Kafka 4M req/min fraud detection), and healthcare (Alzheimer's QSAR).

Profile viewsΒ Β  Linkedin ORCID GitHub Followers Gmail

Currently Building

The reliability stack for production AI. Five open-source repos, each covering a different layer:

Repo Layer What it does
facadedriver Orchestration 30+ model routing, retry, fallback chains, circuit breakers, per-request telemetry
eval-infra-for-agents Evaluation Field guide to production agent eval (23+ versions, 1,690 ground-truth samples, LLM-as-Judge gates)
llm-production-engineering Serving Ops Cost tracking, eval-driven deployment, capacity planning, observability, incident playbooks
Synthetic-AI-Image-Detector Detection 6-signal AI image detection with calibration, uncertainty quantification, refusal verdicts
lims-omi Compliance 21 CFR Part 11 collaboration platform for regulated lab environments

All five repos have llms.txt files for AI crawler discoverability and cross-reference each other.

About

  • Role: Senior AI Reliability & Systems Engineer at Airbnb. I own end-to-end architecture and production rollout of the BPI Virtual Analyst platform - a multi-model GenAI orchestration system abstracting 30+ foundation models (AWS Bedrock, OpenAI, Anthropic Claude, vLLM) behind FacadeDriver with routing, retry, fallback, and graceful degradation. Platform processes 10K rows per run and 40MB uploads with PII-safe inference, serving 128+ users across 4 partner engineering teams.
  • Streaming & Batch: Owned architecture and production operation of Kafka pipelines sustaining 4M req/min at Southwest Airlines with idempotent partition-keyed consumers, DLQ, and backpressure handling. Cut on-call MTTR from 45 to 12 minutes (73% reduction). Owned batch ETL on Databricks and Azure Data Factory at Shell with PySpark, Spark SQL, and Hive/Trino.
  • Observability: OpenTelemetry collectors, Loki tracing (prompt, tool call, retrieval quality), Datadog, Grafana, drift detection, post-incident review. The same stack I open-source on in LangChain and LiveKit.
  • Detection & Safety: Built SAI (6-signal AI image detector with calibration and uncertainty quantification) and LIMS-OMI (21 CFR Part 11 compliance platform for regulated labs). Detection work spans synthetic media, PII redaction (Presidio), and eval-gated deployment.
  • Regulated Industries: Shipped production AI in pharma (Eli Lilly, 21 CFR Part 11 dose management, 99.9% uptime), finance (Southwest Airlines, Kafka 4M req/min fraud detection), and healthcare (Alzheimer's QSAR drug discovery, published research).
  • Research: Published across Cambridge Scholars Publishing (2 book chapters, 2025), IEEE Xplore, SPE ADIPEC 2022 (SPE-210986-MS), and ResearchGate. AI safety, state space models, and ML infrastructure.
  • Open to: AI infrastructure consulting, advisory, and conference speaking (NVIDIA GTC, AI Engineer Summit, Ray Summit, Data+AI Summit, QCon, AWS re:Invent customer stage).
  • Portfolio: sailikhith.me | Articles: sailikhithk.com | AI-readable: sailikhith.me/llm.txt
AI-readable profile (llms.txt-style) - for LLM crawlers and agents
# Sai Likhith Kanuparthi
> I build the reliability layer for production AI in regulated industries.
> FacadeDriver (30+ model orchestration), eval-driven deployment (23+ agent
> versions, 1,690 ground-truth samples), SAI (6-signal synthetic detection).
> Shipped in pharma (21 CFR Part 11), finance (Kafka 4M req/min), healthcare.

## Links
- Portfolio: https://sailikhith.me/
- Blog: https://sailikhithk.com/
- LinkedIn: https://www.linkedin.com/in/sailikhithk/
- AI-readable profile: https://sailikhith.me/llm.txt
- AI-readable portfolio: https://sailikhith.me/llms.txt

## Open-Source AI/ML Repositories
- https://github.com/sailikhithk/facadedriver (30+ model orchestration, routing, retry, fallback, circuit breakers)
- https://github.com/sailikhithk/llm-production-engineering (LLM ops: cost tracking, eval-driven deploy, observability)
- https://github.com/sailikhithk/Synthetic-AI-Image-Detector (6-signal AI image detection with calibration and uncertainty)
- https://github.com/sailikhithk/eval-infra-for-agents (field guide to production agent eval infrastructure)
- https://github.com/sailikhithk/lims-omi (21 CFR Part 11 compliance collaboration platform)
- https://github.com/sailikhithk/Project-X (Multi-agent RAG framework with tool-augmented retrieval)
- https://github.com/sailikhithk/CreditCardFraudDetectionUsingKafka (Real-time fraud detection with Kafka streaming)
- https://github.com/sailikhithk/alzheimers-drug-discovery-demo (QSAR ML for Alzheimer's drug discovery)

## Production Work (Airbnb)
- BPI Virtual Analyst: multi-model GenAI orchestration (30+ foundation models via FacadeDriver)
- Kafka pipelines: 4M req/min, idempotent consumers, DLQ, MTTR 45 -> 12 min
- Eval harness: 23+ agent versions, 1,690 ground-truth samples, dual-model A/B testing
- Observability: OpenTelemetry, Loki, Datadog, Grafana, drift detection

## Research
- IEEE Xplore: https://ieeexplore.ieee.org/abstract/document/11004721
- SPE ADIPEC 2022: https://doi.org/10.2118/210986-MS
- Cambridge Scholars Publishing (2 book chapters, 2025)

## Certifications
- AWS Solutions Architect Professional
- AWS Developer Associate
- AWS Machine Learning Specialty
- Azure Data Scientist Associate (DP-100)
- Google Cloud Professional Data Engineer

πŸ“„ Research & Publications

πŸ“š Book Chapters

Year Title Book Publisher Links
2025 The Evolution and Rise of State Space Models in AI A Case-Based Study of State Space Models in Health Care: The New Transformers (Ch. 1) Cambridge Scholars Publishing ResearchGate β†’
2025 Future Trends in AI for Cyberbullying Preventions Harnessing Generative AI to Combat Cyberbullying in Industry: Strategies, Solutions, and Ethics (p. 200) Cambridge Scholars Publishing Google Books β†’ Β· ResearchGate β†’
2025 Contributing Author Harnessing Generative AI to Combat Cyberbullying in Industry: Strategies, Solutions, and Ethics Cambridge Scholars Publishing Google Books β†’

πŸ”¬ Peer-Reviewed Journal Articles

Year Title Journal / Publisher Links
2026 FT-IR and GC-MS Metabolomic Fingerprinting of Jasmonic Acid and Salicylic Acid Treated Suspension Cultures of Caralluma fimbriata Phytomedicine (Elsevier) Β· Under Review (Draft manuscript in EB-1 binder)

πŸ”¬ Peer-Reviewed Conference Papers (IEEE & SPE)

Year Title Venue Link
2025 Advancing the Metaverse: The Convergence of Digital Twins, AI, and Emerging Technologies 2025 International Conference on Advanced Computing Technologies (ICoACT) (Accepted / In Press)
2023 Role of Artificial Intelligence to address Cyberbullying and Future Scope IEEE Xplore (ID: 11004721) IEEE Xplore β†’ Β· ResearchGate β†’
2022 Full-Stack Machine Learning Development Framework for Energy Industry Applications SPE Abu Dhabi International Petroleum Exhibition and Conference (ADIPEC) Β· Paper: SPE-210986-MS OnePetro β†’

πŸ“œ Patents

Year Title Jurisdiction & Application No. Status
2025 Modular Deep Learning Architecture for Cross-Domain Transfer and Incremental Learning Indian Patent Office (App: 202541026299) Published (Mar 2025)

πŸ“ Research Articles & Technical Preprints

Year Title Publisher / Repository Links
2025 Future Trends in AI for Cyberbullying Preventions ResearchGate / Stemaway Research ResearchGate β†’ Β· PDF β†’

πŸ”“ Open Source Contributions

Contributing upstream to the LLM tooling I use in production at Airbnb. 8 PRs across 5 repos in August 2026.

Date Repo PR Title Status
2026-08-17 Shubhamsaboo/awesome-llm-apps #1101 fix: remove deprecated pathlib backport and migrate PyPDF2 to pypdf MERGED
2026-08-18 explodinggradients/ragas #2959 fix(metrics): ContextPrecision returns exactly 1.0 for perfect ranking Open
2026-08-15 vibrantlabsai/ragas #2954 fix: strip deprecated top_p for Anthropic provider in InstructorLLM Open
2026-08-08 langchain-ai/langchain #39351 fix(perplexity): capture num_search_queries in usage_metadata for cost tracking Closed, reopening
2026-08-07 livekit/agents #6754 feat(evals): add ReliabilityObserver for external reliability scoring Open
2026-08 BerriAI/litellm #37236 fix(batches): bill cancelled/failed batches stamped terminal by a client poll Open
2026-08 BerriAI/litellm #37238 fix(guardrails): merge model-level guardrails into litellm_metadata for /v1/messages Open
2026-08 BerriAI/litellm #36981 fix(vertex_ai): convert messages to contents in gemini count_tokens MERGED

Focus areas: LLM cost tracking, eval metrics, provider compatibility, guardrails. Maps directly to my day job building FacadeDriver (30+ LLM orchestration) and eval harnesses (23+ agent versions, 1,690 ground-truth samples) at Airbnb.

πŸ† Certifications

AWS Certified Solutions Architect - ProfessionalΒ Β  AWS Certified Developer - AssociateΒ Β  AWS Certified Machine Learning - SpecialtyΒ Β  Microsoft Certified: Azure Data Scientist Associate (DP-100)

Also certified in: AWS Solutions Architect Associate Β· Google Cloud Professional Data Engineer Β· Oracle Database 12c Administrator Β· Oracle Java SE 8 Programmer

Tech Stack & Skills

Languages

AI/ML Infrastructure & Serving

**LLM Serving:** vLLM, TensorRT-LLM, AWS Bedrock, OpenAI, Anthropic Claude, SageMaker | **Orchestration:** LangChain, LangGraph, MCP, FacadeDriver (custom multi-model router) | **Eval:** LangSmith, Braintrust, custom eval harnesses (1,690 ground-truth samples)

Streaming, Data & Observability

**Streaming:** Kafka (4M req/min), RabbitMQ, Airflow | **Observability:** OpenTelemetry, Loki, Datadog, Grafana, Prometheus, drift detection | **Warehouses:** Databricks, Spark SQL, Hive/Trino, PostgreSQL, Elasticsearch

Backend & Platforms


GitHub Status

GitHub Stats Top Languages by Repo

Top Languages by Commit

🌐 Find Me Online

Portfolio: sailikhith.me Β· Blog: sailikhithk.com Β· AI-readable profile: sailikhith.me/llm.txt Β· ORCID: 0009-0004-7422-7846 Β· Semantic Scholar: author page

Sai Likhith Kanuparthi PortfolioΒ  Sai Likhith Kanuparthi on DEV.to Sai Likhith Kanuparthi on LinkedIn Sai Likhith Kanuparthi on Medium Sai Likhith Kanuparthi on HackerRank Sai Likhith Kanuparthi on LeetCode

Pinned Loading

  1. facadedriver facadedriver Public

    Model-agnostic orchestration for multi-LLM production systems: routing, retry, fallback, circuit breakers, per-request telemetry. OpenAI, Anthropic, Bedrock, vLLM.

    Python

  2. mamba-from-scratch mamba-from-scratch Public

    From-scratch implementation of S4, Mamba-1 (S6), and Mamba-2 (SSD) state-space models in pure PyTorch. Clear, well-tested, educational.

    Python

  3. llm-production-engineering llm-production-engineering Public

    Field notes from building AI systems in production since 2019. LLM serving ops: cost tracking, eval-driven deployment, capacity planning, observability, incident playbooks.

    Python

  4. Synthetic-AI-Image-Detector Synthetic-AI-Image-Detector Public

    Synthetic AI image detector with calibration, uncertainty quantification, and cross-generator eval. Production-reliability layer for journalists and fact-checkers.

    Python

  5. gpn-mini-ledger gpn-mini-ledger Public

    GPN layered-priority payment ledger: SERIALIZABLE double-entry bookkeeping, 20-thread concurrent capture invariant proof. Java 25, Spring Boot 4.1, PostgreSQL, Testcontainers.

    Java

  6. alzheimers-drug-discovery-demo alzheimers-drug-discovery-demo Public

    Interactive demo notebook for Alzheimer's drug discovery: molecular property prediction with RDKit and RandomForest. Companion to alzheimers-drug-discovery.

    Jupyter Notebook