Engineering 0-to-1. Teams 1-to-100.
I grew Gamemano’s engineering team from scratch to over 100 people, and built Veda Studios to 50. I recently shipped Mudda Kendra solo—from first commit to both app stores—and I still write production code every week.
A bit about who I am and how I work.
Yogesh Bhandari
Co-founder, Cloud Cheers
Currently talking with founders building in cloud and AI
Over the last five years, I built engineering departments from scratch across three sister companies: 55 people at AITC, 50 at Veda Studios, and 100+ at Gamemano. I established the cloud foundations, hiring standards, and microservices architecture at each venture, handing the team over once it was stable and self-sufficient.
I don't believe in technology leaders who only live in spreadsheets and roadmaps. While a big part of my job is hiring, sprint cadence, and architecture planning, I still write backend code, design PostgreSQL schemas, and debug deployment pipelines. Staying close to the code keeps architectural decisions grounded in reality.
Today I'm co-founder at Cloud Cheers, helping companies modernize cloud infrastructure and put AI into production. The work I'm proudest of is Mudda Kendra, a legal practice platform I took solo from first commit to live on both app stores.
0-to-1 Product Execution
Taking a product from initial customer discovery and first commit to a live app in the store.
Org Scaling (0 to 100+)
Hiring standards, sprint rituals, and engineering culture that scales sustainably without operational chaos.
Backend & Cloud Systems
Distributed event streams (Kafka/Redis), clean database design, and automated CI/CD release rails on AWS.
Case Studies
Production systems and the engineering behind them.
A look at real problems I've solved: cutting cloud spend, handling traffic surges without downtime, and taking early prototypes into production.
Mudda Kendra: Legal Practice Management Platform
Designed, architected, and built a full-stack legal practice management SaaS from 0 to 1 as a solo engineer. Features cross-platform Flutter applications on iOS and Android with an offline-first Hive cache, backed by a FastAPI and PostgreSQL cloud infrastructure that schedules automated court hearing alerts and manages secure client document vaults.
Legal Retrieval and Semantic Search Engine
Engineered a specialized document retrieval and semantic search engine for law firms. The platform processes high-volume court filings and scanned briefs via asynchronous OCR, uses hybrid retrieval combining Pinecone dense vector indexing with OpenSearch lexical search, and enforces cross-encoder reranking with strict page-level citation validation to eliminate hallucinations.
Enterprise HRMS AI Assistant with Model Context Protocol (MCP)
Architected and built an enterprise AI assistant embedded into a Human Resource Management System using the Model Context Protocol (MCP). Implemented multi-layered role-based access control (RBAC), dynamic tool schema filtering based on user JWTs, PostgreSQL Row-Level Security, and immutable audit logging across payroll, leave, and hiring workflows.
High-Scale Real-Time Messaging and Encrypted Chat Infrastructure
Engineered a distributed real-time messaging platform supporting high-concurrency instant chat, channels, media delivery, and WebRTC calling. Built stateful connection gateways in Go and Node.js, backed by Apache Kafka for queueing, Redis Pub/Sub for inter-node routing, and Signal Protocol Double Ratchet End-to-End Encryption (E2EE).
E-Commerce Platform with Real-Time Tracking
Architected a high-concurrency retail platform featuring distributed checkout, real-time inventory locking, and live shipment tracking. The system normalizes third-party carrier webhook streams, delivers live status updates via WebSockets, and synchronizes real-time order context directly into Zendesk for customer support operations.
What I've learned building software, backends, and teams.
Notes on things I've built, broken, and fixed in production: practical tradeoffs and engineering decisions that only become clear under real traffic.
Cut Scope, Never Quality: How Early-Stage Startups Can Ship Fast Without Burning Customer Trust
When a founder says 'just ship it by Friday, we'll fix the bugs later,' what should engineering do? Why shipping buggy software destroys customer trust, how to cut scope instead of quality, and the pragmatic infrastructure that lets startups move fast without collapsing.
The CTO's Guide to AI Vendor Risk: Evaluating LLM Providers for Enterprise Use
A practical, battle-tested framework for evaluating LLM providers on cost, security, compliance, performance, and architecture patterns that keep you flexible.
Detailed Guide to Building MCP Server in Production: Complete Technical Deep Dive
A comprehensive technical guide to designing, building, testing, deploying, and operating MCP (Model Context Protocol) servers in production environments. Covers architecture design, security hardening, performance optimization, observability, disaster recovery, and real-world patterns for integrating AI capabilities with existing business systems. Includes complete code examples, deployment strategies, and lessons from production deployments.
Achieving 99.95% Uptime: Building Self-Healing Infrastructure for 200+ Microservices
A complete technical guide to architecting, deploying, and operating 200+ microservices with 99.95% uptime (4.4 hours downtime per year). Covers reliability engineering principles, multi-region architecture, observability at scale, self-healing automation, chaos engineering, and incident response. Includes detailed code examples, diagrams, and a proven roadmap from 98.2% to 99.95% uptime.
Cloud Cost Optimization at Scale: A $2.8M Reverse-Engineering Case Study
A detailed case study on how a high-growth SaaS company reverse-engineered their $5.4M annual cloud spend, identified inefficiencies across compute, storage, and networking, and achieved a 52% cost reduction ($2.8M in annual savings) through systematic optimization, intelligent right-sizing, and architectural redesign. Includes step-by-step technical implementation, code snippets, and a replicable FinOps framework.
Documented engineering leadership and delivery.
A chronological record of founding CTO leadership, organizational scaling from day zero to 100+ people, and production systems architecture across enterprise and high-growth products.
Team Scale
Scaled engineering, QA, design, and DevOps from founding team to 100+ across global operations.
Operating Cost
Optimized infrastructure budgets and vendor contracts across multi-million dollar allocations.
System Availability (Gamemano)
Maintained fault-tolerant AWS and Kubernetes infrastructure and automated CI/CD for Gamemano operations.
Project Kickoff
Accelerated kickoff times for new client projects by 70% at AITC International using reusable cloud infrastructure and automated CI/CD.
Cloud Cheers
Boutique AI and Cloud consultancy. Leading technical solution architecture, LLM agentic Proof-of-Concepts, and multi-cloud blueprints.
Gamemano Pvt Ltd
Built the entire technology organization from concept to 100+ multidisciplinary professionals across Backend, Frontend, QA, Game Dev, and DevOps.
Veda Studios
Built the engineering department from the ground up to 50+ engineers, launched 3 major commercial software products, and executed legacy-to-AWS microservices migration.
AITC International
Established the company's first software development division from scratch, hiring and directing a 55-person team delivering web applications and games.
Numeric Mind (Nimble Clinical Research)
Led 5 engineers building clinical research data analytics and visualization platforms adhering to strict CDISC regulatory standards.
From first commit to 100+ engineers. What are you building?
Most of my conversations are with founders and engineering leaders working through architecture decisions, system bottlenecks, or team growth.