Role Description
Sr. DevOps Engineer (SRE / Platform Engineering Focus)
Location: India (Preferably near UST BFF team)
Positions: 2
Experience: 7–12+ years
Role Overview
We are looking for highly hands-on DevOps / SRE Engineers with strong Platform Engineering experience supporting application and API hosting platforms in AWS and Kubernetes environments.
Candidates must be capable of operating, troubleshooting, and enhancing production platforms in a multi-vendor ecosystem
This is not a people-management role. Strong recent hands-on technical expertise is mandatory.
Strong communication skills and the ability to drive discussions with Tier-2 teams, stakeholders, and cross-functional engineering groups are required.
Key Responsibilities
- Maintain and enhance CI/CD pipelines (GitLab preferred).
- Manage AWS infrastructure and Kubernetes workloads.
- Perform production incident triage, debugging, and root cause analysis using logs, metrics, and Splunk.
- Enhance Terraform-based infrastructure and automation.
- Support repository restructuring and platform modernization initiatives.
- Drive stability, reliability, observability, and performance improvements.
- Collaborate with GCC team and UST BFF leads.
- Support platforms that host and enable application and API services.
Mandatory Skills
- AWS (strong hands-on experience).
- Kubernetes administration, deployments, troubleshooting, and operational support.
- CI/CD (GitLab preferred).
- Terraform / Infrastructure as Code.
- Splunk and observability tooling.
- Python and Shell scripting.
- Strong production support, troubleshooting, and incident management skills.
- Experience supporting platforms that host applications and APIs.
- Hands-on experience supporting production platforms hosting microservices, APIs, and enterprise applications at scale.
- Helm Charts and Kubernetes deployment automation.
- Configuration Management.
- Secrets Management.
- Governance, platform guardrails, and service testability practices.
- Strong communication and stakeholder collaboration skills.
Platform Engineering Expertise
- Kubernetes platform operations and troubleshooting.
- AWS infrastructure supporting production services.
- CI/CD pipelines for both application and infrastructure deployments.
- Observability, monitoring, ing, and operational excellence.
- Reliability engineering and production support best practices.
Good-to-Have Skills
- Grafana dashboard creation.
- Exposure to Conduktor.
- Basic understanding of Java / TypeScript services.
- Awareness of AI and automation use cases.
Experience Expectations
- Senior-level engineers with strong hands-on expertise.
- Demonstrated experience in production environments.
- Strong incident management, debugging, and root cause analysis skills.
- Must be able to explain technical implementations in detail.
- Must be able to clearly articulate troubleshooting approaches, production incidents handled, RCA findings, and remediation actions implemented.
- Hands-on engineering experience is required; leadership-only profiles will not be considered.
API Platform Support
- Knowledge of APIs and how to support them is essential.
- This is not an API development role; however, the platform supports and enables API development.
- Candidates should understand API deployments, monitoring, troubleshooting, security considerations, and production support.
Work Model
- Initial PST overlap required during KT phase.
- Daily call at 8:30 AM PST.
- Deployment support around 8:00 PM PST.
- Post-KT: Standard India shifts.
- No 24x7 support required.
Additional Expectations
- Ability to work independently with ownership.
- Strong collaboration skills in a multi-team setup.
- Prefer co-location with the UST BFF team.
- Prior experience working in distributed team models is a plus.
- Strong verbal communication skills and ability to drive technical discussions.
Skills
devops,site reliability engineering,aws cloudfront,kubernetes,splunk,python,docker,