AutoRABIT Holding Inc.

Senior Site Reliability Engineer

Hyderabad, India - Full Time

AutoRABIT Profile

AutoRABIT is the leader in DevSecOps for SaaS platforms such as Salesforce. Its unique metadata-aware capability makes Release Management, Version Control, and Backup & Recovery complete, reliable, and effective. AutoRABIT’s highly scalable framework covers the entire DevSecOps cycle, which makes it the favourite platform for companies, especially large ones who require enterprise strength and robustness in their deployment environment. AutoRABIT increases the productivity and confidence of developers which makes it a critical tool for development teams, especially large ones with complex applications. AutoRABIT has institutional funding and is well positioned for growth. Headquartered in the CA, USA and with customers worldwide, AutoRABIT is a place for bringing your creativity to the most demanding SaaS marketplace.

Job Role

Responsible for reliability, scalability, security, and automation of cloud-native platforms on AWS. Works across ECS, EKS, EC2, databases, and CI/CD systems with strong ownership of observability, incident management, and compliance.

Roles & Responsibilities

  • Define and manage SLIs, SLOs, SLAs; enforce error budgets
  • Own incident response, on-call, RCA, and postmortems
  • Build and maintain observability using Elasticsearch and Kibana (logs, metrics, alerts)
  • Automate infrastructure using Terraform
  • Operate workloads on Amazon Web Services (ECS, EKS, EC2, Lambda, SSM)
  • Manage CI/CD using CodeBuild, CodeDeploy, CodePipeline
  • Implement security controls and hardening aligned with CIS Level 2
  • Integrate and manage endpoint/workload security using Trend Micro
  • Optimize cost, performance, and capacity
  • Drive self-healing, auto-remediation, and runbook automation
  • Support databases: DynamoDB, PostgreSQL, MySQL
  • Leverage Amazon Bedrock for automation/operational intelligence use cases
  • Responsibility to adhere to set internal controls.

Desired Skills and Knowledge

  • Strong experience with Kubernetes (EKS) and Docker
  • Deep AWS experience (ECS, EKS, EC2, Lambda, IAM, VPC, SSM)
  • Hands-on with Terraform and CI/CD pipelines
  • Expertise in ELK stack (log ingestion, parsing, alerting, dashboards)
  • OS hardening and administration (Ubuntu, Amazon Linux)
  • Networking fundamentals (DNS, TCP/IP, load balancing)
  • Scripting (Python/Bash)
  • Database operations and performance tuning
  • Experience with CIS benchmarks, SOC2/ISO controls
  • Multi-region/high availability design
  • AIOps or AI-assisted observability
Education
  • Bachelor’s in computers or any related field.

Location: Hyderabad, Hybrid - 3 Days from Office

Experience: 

  • 6+ years in SRE/DevOps/Platform Engineering
  • 3+ years in AWS-based production environments

Compensation: 24 - 28 LPA
Website: www.autorabit.com

Apply: Senior Site Reliability Engineer
* Required fields
First name*
Last name*
Email address*
Location
Phone number*
Resume*

Attach resume as .pdf, .doc, .docx, .odt, .txt, or .rtf (limit 5MB) or paste resume

Paste your resume here or attach resume file

Total years of experience*
How many years of experience do you have in cloud infrastructure and architecture ?*
How many years of hands-on experience do you have with Amazon Web Services (ECS, EKS, EC2, Lambda) ?*
Do you have production experience using Terraform for infrastructure provisioning ?*
How many years of experience do you have managing Kubernetes (EKS) clusters in production ?*
Do you have hands-on experience with ELK stack (Elasticsearch, Kibana) for observability and alerting ?*
Have you implemented CI/CD pipelines using AWS CodeBuild, CodeDeploy, or CodePipeline ?*
Rate your proficiency in Linux system administration (Ubuntu/Amazon Linux) and networking (DNS, VPC, VPN).*
Do you have experience implementing CIS Level 2 security hardening and working with tools like Trend Micro ?*
Have you worked with AWS Lambda (Python/Boto3) for automation or auto-remediation ?*
Have you defined/managed SLOs, handled incidents, and performed root cause analysis (RCA) in production environments ?*
Current Location*
Current Company*
Current CTC*
Expected CTC*
Notice Period*
Are you willing to relocate to Hyderabad and work in Hybrid mode(We are working in Hybrid mode - 3 days a week from office)?*
Are you willing to work in rotational shifts and rotational week-offs(5 days working and shift allowance for UK & US shift) ?*
Human Check*