Lead Engineer · GenAI

Priyanshu Shekhar Sinha

AI/ML Engineer architecting enterprise RAG at Elastiq AI · Google Cloud Professional ML Engineer

Summary

AI/ML Engineer with 7+ years shipping production-grade systems. Currently leading the design and development of an enterprise RAG platform at Elastiq AI — multi-source ingestion, hybrid retrieval, cross-encoder re-ranking, agentic reasoning, and document-level access control synchronized across cloud providers.

Experience

Lead Engineer — GenAI
Elastiq AI · Bangalore, India
Jan 2025 — Present

Lead the architecture and development of Discover, an enterprise RAG platform spanning ingestion, retrieval, ranking, agentic reasoning, and access control across cloud providers.

  • Designed and shipped a ReAct-style reasoning agent for natural-language analytics — surfaces key insights as the final answer, generates large-tabular insight payloads, and harvests delegated-agent and web references into inline citations; added an orchestrator / sub-agent role toggle to the agent canvas.
  • Built a stateless Streamable-HTTP MCP server exposing retrieval as tools (ask, search, aggregate_jira/confluence), secured with OAuth 2.1 resource-server auth (Keycloak JWT) and Kubernetes liveness probes.
  • Engineered a staged connector-ingestion pipeline (fetch → Docling parse → contextualise → embed) with windowed processing and per-knowledge-base tuning presets; delivered Jira, Confluence, S3 and SVN connectors behind a unified experience.
  • Researched and implemented Citus DB for distributed, enterprise-scale Postgres workloads and migrated the backend to a production gunicorn multi-worker server.
  • Onboarded the first enterprise client for the Text2SQL solution; mentor a cross-functional team on data pipelines, LLM fine-tuning, and deployment strategy.
Senior Software Engineer
Elastiq AI · Bangalore, India
Sep 2024 — Jan 2025

Foundational AI engineering on Discover — Elastiq's flagship product — covering fine-tuning, retrieval, and agent design.

  • Fine-tuned LLaMA 3.1 70B via advanced QLoRA and LoRA techniques, reaching 85% target accuracy.
  • Built the foundational Text2SQL framework from scratch and shipped AI agents for natural-language-to-SQL plus unstructured file processing.
  • Integrated multiple Graph DB backends offering Classic RAG, Cypher-based RAG, and GraphRAG — extending Discover with structural retrieval.
  • Engineered custom datasets and context-aware chunking algorithms to lift unstructured-document RAG performance.
Senior Data Scientist
Acuity Knowledge Partners · Bangalore, India
May 2023 — Aug 2024

Research, analytics, and technology partner to the financial-services sector — asset managers, investment banks, private equity, and hedge funds.

  • Developed a versatile analysis dashboard for an investment firm enabling KPI monitoring and experimentation — 87% reduction in client analysis time.
  • Built a web dashboard integrated with an ESG data engine for a top-tier investment firm, shaping their investment-strategy decisions.
  • Shipped an end-to-end RAG system for parsing financial reports.
Associate, Data Scientist
TheMathCompany · Bangalore, India
Dec 2021 — Apr 2023
  • Led a 5-person team shipping a flexible A/B testing framework for a leading pharmaceutical client, measuring promotion success against revenue lift.
  • Co-led a 7-person team on a seasonality-aware sales-anomaly pipeline that drove customer outreach delivering +$125K revenue in Q1.
Data Analyst
Recruitment Smart Technologies · Ahmedabad, India
May 2021 — Dec 2021
  • Cut new-client integration from 10 days to 3 hours (97%); operational reports from 5 days to 30 minutes; opportunity reports from 1 day to 30 minutes.
  • Reduced advanced-analytics report generation from 7 days to 15 minutes (99%) and automated Power BI dashboard refreshes.
Data Scientist
Sumyag Data Sciences · Bangalore, India
Aug 2019 — Apr 2021
  • Designed a multi-phase pipeline enriching extracted document data points — generated 500+ features via NLP, data wrangling, and text mining.
  • Co-developed a matrix-multiplication ensemble on the NumPy stack across 10+ sources for entity classification, plus custom Bayesian models for word embeddings.

Skills

Languages
PythonSQLCypher
LLM & GenAI
LLaMA 3.1 70BQLoRALoRASFTRLHF Classic RAGGraphRAGCypher-RAGText2SQLMCP ReAct AgentsLangChainTransformersOllama
Search & Retrieval
OpenSearchBM25kNNHybrid Search Cross-encoder Re-rankingChromaDBEmbeddingsACL Sync
Multi-modal
TextPDFAudioVideoSpeaker DiarizationOCR
Machine Learning
PyTorchTensorFlowscikit-learnMLflow PandasPolarsNumPyPySparkMLOps
Cloud
Google CloudAWS · S3 / EKS / ECS / SageMakerGCSAzure Blob
Databases
PostgreSQLCitus DBSnowflakeDatabricksGraph DBChromaDB
Tooling
DockerApache AirflowFlask / GunicornStreamlitChainLitPlotlyPower BIGit

Education & Credentials

Education
PG Program — Data Science Engineering
Great Lakes Institute of Management · 2019
B.E. — Electrical & Electronics
Sir M. Visvesvaraya Institute of Technology · 2014–2018
Certifications
Google Cloud · Active
NLP Specialization
DeepLearning.AI · Coursera · 2021
Deep Learning Specialization
DeepLearning.AI · Coursera · 2018
Spark & Python for Big Data (PySpark)
Coursera
Awards
  • Star Employee of the Month — Elastiq AI · Mar 2026
  • Star Performer of the Month — Recruitment Smart · 2021
  • Arctic Code Vault Contributor — GitHub Archive Program
  • Pull Shark — GitHub Achievements