Role- Lead Cloud Infrastructure & DevOps Engineer
Experience- 10+ Years
Location- Pune, Hyderabad, Bangalore
Required Skills & Experience
Must Have
- 8-15 years of technology experience, with recent hands-on responsibility for cloud infrastructure, DevOps, or platform automation.
- Strong hands-on experience with Google Cloud infrastructure supporting enterprise GKE and data platforms.
- Deep understanding of Shared VPC, routing, firewalls, Cloud NAT, Private Service Connect, Cloud DNS, load balancing, Cloud IAM, Cloud Storage, Cloud KMS, and Secret Manager.
- Advanced Terraform experience covering reusable GCP modules, GKE provisioning, remote state, testing, policy validation, drift detection, and controlled promotion.
- Strong experience designing CI/CD pipelines using Cloud Build, GitHub Actions, GitLab CI, Jenkins, or equivalent enterprise tooling.
- Experience with Artifact Registry, OCI images, artifact management, and CI/CD or GitOps enablement; sufficient Helm knowledge to support release pipelines.
- Working knowledge of GKE dependencies, including node infrastructure, Workload Identity Federation, load balancing, DNS, ingress, storage classes, and observability.
- Experience implementing Cloud IAM, Secret Manager integration, vulnerability scanning, policy-as-code checks, and infrastructure security controls.
- Experience supporting multi-environment infrastructure releases, cluster upgrades, rollback, observability, incident troubleshooting, backup, and recovery.
- Strong technical leadership, documentation, mentoring, stakeholder communication, and problem-solving skills.
Good to Have
- Experience enabling infrastructure for Spark, Kafka, Flink, Airflow, Iceberg, Trino, or related data platform services.
- Exposure to Vault, External Secrets Operator, cert-manager, Kyverno, Open Policy Agent/Gatekeeper, or cloud-native policy services.
- Experience with Cloud Monitoring, Cloud Logging, Prometheus, Grafana, Loki, OpenTelemetry, or centralized enterprise observability platforms.
- Experience supporting hybrid, sovereign, restricted, or regulated GCP environments.
- Understanding of FinOps, infrastructure cost optimization, capacity planning, disaster recovery, and platform SRE practices.