Job Description & Scope
Job description Job Summary We are seeking a motivated and detail-oriented Site Reliability Engineer (SRE) with 1 to 3 years of experience to join our growing team. The ideal candidate will have a strong foundation in Linux systems administration, AWS cloud infrastructure, and a basic understanding of DevOps tools and practices. This role involves maintaining system reliability, automating operational tasks, and ensuring high availability and performance of services. Role & responsibilities Monitor and maintain availability, performance, and reliability of production systems Manage and troubleshoot Linux-based servers in a production environment Deploy, manage, and optimize services on AWS (EC2, S3, IAM, CloudWatch, RDS, etc.) Respond to system outages and ensure timely resolution Collaborate with development and QA teams for deployment and support activities Participate in on-call rotations as needed Preferred candidate profile 1 to 3 years of hands-on experience as a System Administrator, DevOps Engineer, or SRE Strong expertise in Linux system administration (CentOS, Ubuntu, etc.) Good working knowledge of AWS cloud services EC2, S3, IAM, CloudWatch, RDS, etc. Basic experience with DevOps tools like Git, Jenkins, Docker, Ansible, or Terraform Familiarity with monitoring and alerting tools like Prometheus, Grafana, or CloudWatch Role: IT Operations Management Industry Type: IT Services & Consulting Department: IT & Information Security Employment Type: Full Time, Permanent Role Category: IT Infrastructure Services Education UG: Any Graduate Key Skills Skills highlighted with ‘‘ are preferred keyskills Linux AdministrationRedhat LinuxAws Cloud JenkinsTerraformDockerDevops ToolsAnsibleKubernetes