About Awan Infotech
Awan Infotech is a key provider of IT services across the entire cloud life cycle, spanning cloud architecture, implementation, integration, and management, helping customers progress confidently through their cloud journey. The company delivers end-to-end cloud platform solutions on Amazon Web Services (AWS) and Azure, and is an AWS Consulting Partner. Awan Infotech provides secure, end-to-end cloud solutions focused on scalability, reliability, and performance, backed by 24/7 global support, and is recognized as a cloud services partner specializing in Site Reliability Engineering, Data Services, and DevSecOps. Founded in 2011 and headquartered in Chennai, Tamil Nadu, the company has over 14 years of experience delivering measurable results and building lasting partnerships with clients worldwide.
Job Description – DevOps Engineer
Experience: 3–4 Years
Work Mode: Work from Office (No remote/WFH)
Shift: 24×7 Rotational Support (including night shifts, weekends, and public holidays as per roster)
About the Role
We are looking for a skilled DevOps Engineer to join our infrastructure and operations team. The ideal candidate will have hands-on experience managing cloud infrastructure across Azure and AWS, container orchestration using AKS and EKS, and strong automation/scripting skills in Python. This role also involves Cloud Operations support, ensuring the stability, performance, and availability of cloud-hosted applications in production. The position requires participation in a 24/7 rotational support model to ensure high availability and reliability of production environments.
Key Responsibilities
- Design, deploy, and manage infrastructure on Azure and AWS cloud platforms.
- Manage and troubleshoot Kubernetes clusters on AKS (Azure Kubernetes Service) and EKS (Elastic Kubernetes Service).
- Provide Cloud Operations support for cloud-hosted applications, ensuring high availability, performance, and reliability of production systems.
- Monitor cloud-hosted application health, infrastructure performance, and system availability; proactively identify, troubleshoot, and resolve issues.
- Build and maintain CI/CD pipelines to support application deployment and release automation.
- Develop automation scripts and tools using Python to streamline infrastructure and operational tasks.
- Provide 24/7 rotational on-call/shift support, including incident response, root cause analysis, and escalation management for both infrastructure and application-level issues.
- Implement and maintain Infrastructure as Code (IaC) using tools such as Terraform, ARM templates, or CloudFormation.
- Manage configuration, patching, and security compliance across cloud environments.
- Perform routine health checks, capacity monitoring, and performance tuning of cloud-hosted applications and underlying infrastructure.
- Collaborate with development, QA, and infrastructure teams to support release cycles, deployments, and production stability.
- Maintain documentation for infrastructure, cloud operations processes, runbooks, and incident reports.
- Ensure adherence to security best practices, backup policies, and disaster recovery procedures.
- Support change management and deployment activities for cloud-hosted applications with minimal downtime.
Required Skills & Qualifications
- 3–4 years of relevant experience in a DevOps/Cloud Engineering/Cloud Operations role.
- Strong hands-on experience with Microsoft Azure and AWS cloud platforms.
- Proven experience managing containerized workloads on AKS and EKS.
- Hands-on experience supporting and operating cloud-hosted applications in production environments.
- Solid understanding of Kubernetes concepts — pods, deployments, services, ingress, namespaces, etc.
- Proficiency in Python for scripting and automation.
- Experience with CI/CD tools (e.g., Jenkins, Azure DevOps, GitHub Actions, GitLab CI).
- Familiarity with Infrastructure as Code tools (Terraform, ARM Templates, CloudFormation).
- Working knowledge of monitoring/logging tools (e.g., CloudWatch, Azure Monitor, Prometheus, Grafana, ELK).
- Understanding of networking concepts, security groups, IAM/RBAC, and cloud security fundamentals.
- Experience with version control systems (Git).
- Willingness and flexibility to work in rotational shifts, including nights, weekends, and holidays.
- Strong troubleshooting, analytical, and problem-solving skills, especially during live incidents.
- Good communication skills to coordinate with cross-functional teams during incidents.
Good to Have
- Certifications such as AWS Certified DevOps Engineer, Microsoft Certified: Azure DevOps Engineer Expert, or CKA (Certified Kubernetes Administrator).
- Experience with containerization tools like Docker.
- Exposure to configuration management tools (Ansible, Chef, Puppet).
- Familiarity with ITIL processes and incident/change management practices.
- Experience with Cloud Operations Command Center or NOC-style support environments.
Work Mode & Shift Details
- Location: Work from Office only (No remote/hybrid option)
- Support Model: 24×7 rotational shifts (Morning/General/Evening/Night rotation as per team schedule)
- Candidates must be flexible and comfortable working in a shift-based, on-call, cloud operations support environment.