DevOps and Platform Engineer

We are seeking a skilled DevOps Specialist with 3 + years of hands-on experience in managing production systems, automation, and cloud-native infrastructure. The ideal candidate will have a strong foundation in operations, along with deep expertise in Kubernetes and container orchestration, to ensure scalability, reliability, and performance of system

Key Responsibilities

  • Design, build, and maintain scalable and highly available infrastructure
  • Deploy, manage, and optimize Kubernetes clusters (on-prem or cloud platforms such as AWS, Azure, or GCP)
  • Automate infrastructure provisioning using tools like Terraform, CloudFormation, or ARM templates
  • Implement and manage CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI, Azure DevOps)
  • Monitor system performance and troubleshooting issues across the stack
  • Ensure system reliability through logging, monitoring, and alerting tools (Prometheus, Grafana, ELK, etc.)
  • Collaborate with development teams to streamline deployments and improve system design
  • Maintain security best practices across infrastructure and applications
  • Manage containerized environments using Docker and Kubernetes
  • Support incident management, root cause analysis, and system optimization initiatives

Experience and Qualifications

  • 3–5 years of experience in DevOps, Site Reliability Engineering (SRE), Platform Engineering, or System Engineering roles.
  • Hands-on experience managing production environments with a focus on availability, scalability, and operational excellence.
  • Experience deploying and managing Kubernetes-based container platforms in cloud or on-premises environments.
  • Experience working with at least one major cloud platform (Azure, AWS, or GCP).
  • Hands-on experience implementing Infrastructure as Code (IaC) using tools such as Terraform, ARM, CloudFormation, or Ansible.
  • Experience designing and maintaining CI/CD pipelines using tools such as Azure DevOps, Jenkins, GitHub Actions, or GitLab CI.
  • Experience with Linux system administration, containerization technologies (Docker), and scripting using Bash, Python, or Shell.
  • Experience working in Agile/DevOps teams and collaborating with development, infrastructure, and security teams.
  • Experience supporting production incidents, troubleshooting complex technical issues, and performing root cause analysis.