Director, Platform Engineering — Mobius Networks
12+ years leading cross-functional engineering teams — up to 14 engineers across 4 squads — building cloud-native platforms, distributed systems, and enterprise AI infrastructure. Built and scaled a 12-member Platform Engineering team from the ground up.
Leading distributed platform and cloud-native engineering teams for over a decade.
Engineering Manager and Platform Architect with 12+ years leading cross-functional engineering teams and building large-scale distributed platforms and cloud-native systems. Experienced leading organizations of up to 14 engineers across 4 Agile squads, partnering with Product, Architecture, SRE, Operations, and Customer Success to turn business objectives into scalable platform capabilities.
Built and scaled a 12-member Platform Engineering team from the ground up, mentoring engineers, growing technical leaders, and supporting promotions into lead roles. My work spans an enterprise DevSecOps automation platform, a sub-millisecond Cloud Data Stream handling 1M+ msg/sec, a distributed DaaS platform across 150+ microservices at 99.5% availability, a unified enterprise AI gateway serving 50+ foundation models, and confidential computing infrastructure for regulated workloads.
Built & scaled a 12-member platform team from scratch · Led up to 14 engineers across 4 squads · Mentored engineers into lead-role promotions
Distributed Systems · Event-Driven Architecture · Microservices · API Design · System Reliability · Workflow Orchestration
150+ microservices · 99.5% platform availability · 3x AI delivery throughput · Zero security incidents
Mobius Networks, Gaian Solutions
Planon Software
Tech Mahindra
IFS Solutions
Building distributed platforms and leading engineering teams at scale.
Own architecture, engineering roadmap, and team leadership for multi-tenant distributed platforms and AI/ML infrastructure in regulated enterprise environments, leading cross-functional squads through delivery, quarterly planning, and production readiness.
Built and scaled a 12-member Platform Engineering team from the ground up across Platform, DevOps, and Infrastructure disciplines. Led 3 Engineering Leads and 14 engineers across 4 Agile squads on RunRun, driving quarterly planning, architecture reviews, and incident management. Mentored engineers into promotions to lead roles.
Designed multi-tenant distributed platform with event-driven communication patterns enabling asynchronous processing, improved resilience, and container-based runtime on Kubernetes across 150+ microservices and 12+ databases.
Led a team of 6 engineers delivering a production-grade Enterprise AI Gateway supporting 50+ foundation models, improving GPU utilization 30–50% and cutting inference costs 25–40%. Led a further team of 4 engineers on a Generative AI platform (Image/Video/3D Studio) serving 1000+ enterprise marketplaces on NVIDIA H100 clusters, cutting infra costs 35–40% and lifting throughput 3x.
Implemented GDPR/PII governance, mTLS, IAM compliance programmes, and confidential compute (CVM/TEE) infrastructure achieving full audit readiness for regulated clients.
Designed enterprise integration services connecting cloud platforms, enterprise applications, and operational systems for facility and asset management.
Designed microservices and real-time integration APIs enabling cross-system asset and facility data synchronisation between Azure, SAP, and enterprise platforms.
Delivered Java-based REST services and integration components for enterprise platforms, with focus on performance optimisation.
Built Java REST services and integration components for enterprise platforms. Optimised SQL queries and execution plans, improving enterprise reporting performance by 40%.
Developed ERP integrations using Java, SQL, and service APIs across enterprise systems.
Developed and maintained ERP integration components using Java, SQL, and service APIs for enterprise business systems across industries.
Leading engineering teams building production LLM and agentic infrastructure at enterprise scale.
Architected and led engineering delivery of a production-grade Enterprise AI Gateway. Defined platform architecture, technical roadmap, and governance standards for intelligent model routing, session memory, multi-agent orchestration, batching, caching, and fine-tuning.
Served self-hosted models (gpt-oss-20b, gpt-oss-120b, Llama 3.1) on NVIDIA H100 via vLLM, Ollama, and LoRAX, with ONNX models served through Triton Inference Server. Autoscaled GPU pools to zero with KEDA between traffic bursts.
Applied Anthropic/OpenAI-style prompt batching and prompt caching to cut redundant token spend across repeated context. Used Cache-Augmented Generation (preloading static context directly into the model's KV cache) as a lower-latency alternative to Retrieval-Augmented Generation for latency-sensitive request paths.
Improved GPU utilization by 30–50%, reduced inference costs by 25–40%, and enabled rapid onboarding of enterprise AI applications while ensuring scalability, security, and production reliability.
Led architecture, platform strategy, and engineering execution for an enterprise Generative AI platform powering Image Studio, Video Studio, and 3D Studio, delivered across 1000+ enterprise marketplaces on NVIDIA H100 Kubernetes clusters.
Established engineering standards, sprint planning, design reviews, and production readiness processes while improving engineering velocity across the team.
Optimized GPU utilization through intelligent scheduling and scale-to-zero capabilities, reducing infrastructure costs by 35–40%, increasing GPU utilization by 40%, and improving throughput by 3x.
Enterprise-grade platforms, shipped and running in production.
Multi-tenant DaaS integrating 12+ database technologies across transactional, analytical, time-series, geospatial, and streaming workloads, processing 10TB+ of data daily.
Led a 6-engineer team delivering a unified gateway for multi-provider LLM and on-premise AI inference, supporting 50+ foundation models with intelligent routing and GPU infrastructure management.
Owned architecture and delivery for a cloud-native self-service infrastructure provisioning platform. Led 3 Engineering Leads and 14 engineers across 4 Agile squads using Camunda + Terraform.
Owned strategy and delivery for an enterprise ML Platform with a 15-person cross-functional org (3 Engineering Leads, 6 ML Engineers, 3 Backend, 2 UI, 1 UX), building reusable MLOps/LLMOps capabilities from data pipelines to serving.
Ultra-low-latency microservice communication platform using RSocket TCP
Confidential VM and Trusted Execution Environment (TEE) platform for regulated enterprise workloads
AI platform ideas in active development
Unified wrapper around self-hosted and hosted LLM backends behind one API.
Wrapper unifying image-generation models behind a single interface.
Text-to-video generation wrapping the WAN model family, with multiple model support (TI2V-5B, T2V-A14B) and customizable output resolution. Ships as a lightweight Docker image that downloads models on first use, exposed through a RESTful API with GPU-accelerated inference.
Wrapper around the Hunyuan3D model for text/image-to-3D-asset generation.
12+ years across platform architecture, AI infrastructure, and engineering leadership.
Role match analysis, open to remote and global opportunities.
Personalized learning paths, message drafts, and job tracking.
Beyond code — learning, exploring, and staying grounded.
Google Cloud
Certification demonstrating expertise in generative AI strategy, implementation, and leadership on Google Cloud infrastructure.
The Linux Foundation / CNCF
Advanced Kubernetes application development, deployment strategies, and cloud-native patterns.
SnapLogic
Integration platform development: pipelines, connectors, and enterprise data flow design on SnapLogic's iPaaS.
IIM Ahmedabad
Executive leadership development and quantitative foundations, investing in the people-management side of the Engineering Manager track.
Coursera, University of California San Diego
Mental tools to master tough subjects and develop effective learning techniques applicable to complex engineering domains.
"Success is not just about individual merit but opportunity, timing, and cultural background."
"When you want something, all the universe conspires in helping you to achieve it."
"All people at root are time optimists. We always think there's enough time to do things with other people."
Promoting the adoption of advanced AI models in enterprise environments. Exploring the potential of Grok4 for real-world applications, from natural language processing to automated decision-making systems.
Critical analysis of emerging technologies, focusing on practical implementation challenges, scalability concerns, and real-world impact. Advocating for responsible technology adoption.
Sacred lakes, colorful markets, and spiritual vibes
Mountain serenity and colonial charm
Ancient ruins and architectural marvels
Beaches, culture and vibrant nights
Yoga capital with serene Ganges views
Spiritual heritage and history
City of pearls and modern tech
Practicing mindfulness and breathing techniques to maintain balance in the fast-paced tech world. The Art of Living philosophy teaches us that happiness is our nature, and stress is just a temporary state that can be managed through proper breathing and meditation.
"When you are fully in the present moment, you drop all your worries and anxieties."
Running isn't just physical exercise, it's mental therapy. It clears the mind, improves focus, and provides the mental resilience needed for complex problem-solving in software architecture.
Enhanced focus for debugging complex systems
Natural stress relief from high-pressure deadlines
Many breakthrough solutions come during runs
Just like in software development, consistency beats sporadic bursts of activity. A regular running routine has taught me the value of small, daily improvements that compound over time, a principle I apply to both code and career growth.
Personal projects and utilities on GitHub.
Cordova plugin for Zebra RFD8500 and TSL RFID scanners.
Web app embedding Autodesk Forge Viewer for 3D/BIM model visualization.
Backend service powering career/job-search analysis and insights.
Real-time message rooms backed by Redis Pub/Sub and Kafka.
Open to Director / Engineering Manager, Platform & AI roles.
I'm always interested in challenging projects and opportunities to work with innovative teams.
gurpreetgandhi3@gmail.com
linkedin.com/in/gandhigurpreet
github.com/grupret
Engineering Manager Journey · AI-Powered
Optional: enable Power Mode for deeper, open-ended Q&A
Quick questions work instantly with no key. Get a key at console.anthropic.com for freeform mode. Stored locally only.
Ready to focus, pick a session above