DevOps and Platform Engineer
We are seeking a skilled DevOps Specialist with 3 + years of hands-on experience in managing production systems, automation, and cloud-native infrastructure. The ideal candidate will have a strong foundation in operations, along with deep expertise in Kubernetes and container orchestration, to ensure scalability, reliability, and performance of system
Key Responsibilities
- Design, build, and maintain scalable and highly available infrastructure
- Deploy, manage, and optimize Kubernetes clusters (on-prem or cloud platforms such as AWS, Azure, or GCP)
- Automate infrastructure provisioning using tools like Terraform, CloudFormation, or ARM templates
- Implement and manage CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI, Azure DevOps)
- Monitor system performance and troubleshooting issues across the stack
- Ensure system reliability through logging, monitoring, and alerting tools (Prometheus, Grafana, ELK, etc.)
- Collaborate with development teams to streamline deployments and improve system design
- Maintain security best practices across infrastructure and applications
- Manage containerized environments using Docker and Kubernetes
- Support incident management, root cause analysis, and system optimization initiatives
Experience and Qualifications
- 3–5 years of experience in DevOps, Site Reliability Engineering (SRE), Platform Engineering, or System Engineering roles.
- Hands-on experience managing production environments with a focus on availability, scalability, and operational excellence.
- Experience deploying and managing Kubernetes-based container platforms in cloud or on-premises environments.
- Experience working with at least one major cloud platform (Azure, AWS, or GCP).
- Hands-on experience implementing Infrastructure as Code (IaC) using tools such as Terraform, ARM, CloudFormation, or Ansible.
- Experience designing and maintaining CI/CD pipelines using tools such as Azure DevOps, Jenkins, GitHub Actions, or GitLab CI.
- Experience with Linux system administration, containerization technologies (Docker), and scripting using Bash, Python, or Shell.
- Experience working in Agile/DevOps teams and collaborating with development, infrastructure, and security teams.
- Experience supporting production incidents, troubleshooting complex technical issues, and performing root cause analysis.