About the Company
We are a leading technology company dedicated to providing innovative solutions and services. Our mission is to empower businesses through cutting-edge technology while fostering a culture of collaboration, diversity, and continuous improvement.
About the Role
The DevOps Engineer will be responsible for managing and maintaining Kubernetes clusters, developing Infrastructure as Code, and ensuring the reliability of cloud-native infrastructure.
Responsibilities
- Manage, patch, upgrade, and maintain Kubernetes clusters in production environments across on-premises and public cloud platforms.
- Develop, maintain, and enhance Infrastructure as Code (IaC) using Terraform and automation tools.
- Build and maintain CI/CD pipelines for infrastructure deployments and GitOps workflows.
- Troubleshoot and resolve production incidents related to Kubernetes infrastructure and cloud platforms.
- Deploy and manage networking and storage solutions for Kubernetes workloads.
- Implement automation using Python, Go, and Ansible to improve platform operations.
- Monitor platform health using Prometheus, Grafana, and other observability tools.
- Improve monitoring, alerting, and incident response to reduce Mean Time to Resolution (MTTR).
- Collaborate with development, networking, security, and operations teams to deliver reliable cloud-native infrastructure.
- Follow Infrastructure as Code, GitOps, and DevOps best practices throughout the software delivery lifecycle.
- Leverage AI-enabled tools and automation to improve infrastructure management and operational efficiency.
Qualifications
3-8 years of experience in this role.
Required Skills
- 3+ years of hands-on experience managing production Kubernetes environments.
- Strong Linux administration experience (5+ years preferred).
- Hands-on experience with Terraform for Infrastructure as Code.
- Experience with Helm for Kubernetes application deployments.
- Experience with GitOps tools such as ArgoCD or FluxCD.
- Strong knowledge of Kubernetes networking, including:
- Ingress & Egress
- Load Balancers
- DNS
- TLS/SSL
- Firewalls
- Experience with Python or Go (Golang) for automation and scripting.
- Experience building and managing CI/CD pipelines.
- Hands-on experience with monitoring tools such as Prometheus and Grafana.
- Strong troubleshooting skills in production environments.
Preferred Skills
- Experience with AWS, Azure, or Google Cloud Platform.
- Experience with Ansible automation.
- Knowledge of Service Mesh technologies (Istio, Linkerd).
- Experience developing Kubernetes Custom Operators.
- Familiarity with ServiceNow.
- Exposure to AI-driven infrastructure automation and cloud operations.
Preferred Certifications
- Certified Kubernetes Administrator (CKA)
- Red Hat Certified Engineer (RHCE)
- Linux Foundation Certified System Administrator (LFCS)
- AWS, Azure, or Google Cloud Certifications
Pay range and compensation package
Competitive salary based on experience and qualifications.
Equal Opportunity Statement
We are an equal opportunity employer and are committed to fostering a diverse and inclusive workplace. We encourage applications from all qualified individuals regardless of race, gender, age, sexual orientation, disability, or any other characteristic protected by law.