|
I build AI systems that go from research to production — not just notebooks. At Publicis Sapient, I designed Bodhi-Atomize, a multimodal AI pipeline that decomposes ad creatives (images, video, GIFs) into structured JSON using Gemini 2.5 Pro, YOLO, and PaddleOCR. What used to take hours now runs in ~1.5 minutes — achieving 84.5% consistency and 89% correctness evaluated with DeepEval. Previously at Lincode Vision Labs, I shipped CV models to production — RF-DETR at 1.8x faster inference than YOLOv8, and defect detection from 55% → 82% mAP through synthetic data and multi-stage training. My research: FedFV-CV — a federated learning framework for finger vein biometrics achieving 1.21% EER and 98.48% TAR, outperforming standard benchmarks across 122,600 images. |
# shivang.yml
role: AI Engineer
company: Publicis Sapient
location: Bengaluru, India
education:
degree: B.Tech CSE
school: IIIT SriCity
gpa: 8.09
building:
- Multimodal LLM pipelines
- Structured output systems
- Production ML on Kubernetes
exploring:
- Agentic AI & LangGraph
- Federated Learning
- ML System Design |
|
Multimodal AI system decomposing image, video & GIF ad creatives into structured JSON for competitive marketing intelligence.
|
FedFV-CV |
slackAgent |
RAG Deployment |

