Site Reliability Engineer, SRE, DevOps Engineer
  • Quantum World Technologies Inc.
2 Hours Ago
NA
C2C
Scottsdale-AZ
8-10 Years
Required Skills: EC2, S3, RDS, Lambda, VPC, CloudWatch, Terraform, AWS IAM, SRE practices, SLIs, SLOs, error budgets, AWS certifications
Job Description
We are looking for a skilled Site Reliability Engineer (SRE) / DevOps Engineer with strong expertise in AWS, Terraform, and IAM. The ideal candidate will be responsible for building, automating, and maintaining scalable, secure, and highly available cloud infrastructure. You will work closely with development and operations teams to improve system reliability, performance, and deployment efficiency.
 
Key Responsibilities:
  • Design, build, and manage scalable infrastructure on AWS Cloud
  • Implement Infrastructure as Code (IaC) using Terraform/Kubernetes
  • Configure and manage AWS IAM roles, policies, and permissions ensuring secure access control
  • Automate CI/CD pipelines for application deployment and infrastructure provisioning
  • Monitor system performance, availability, and reliability using observability tools
  • Troubleshoot production issues and implement root cause analysis (RCA)
  • Ensure system security, compliance, and best practices in cloud environments
  • Collaborate with development teams to improve application performance and resilience
  • Manage configuration, release, and deployment processes
  • Implement disaster recovery (DR) and backup strategies
 
Required Skills & Qualifications:
  • Strong hands-on experience with AWS services (EC2, S3, RDS, Lambda, VPC, CloudWatch, etc.)
  • Proficiency in Terraform for infrastructure automation
  • Deep understanding of AWS IAM (roles, policies, federation, least privilege access)
  • Experience with CI/CD tools (Jenkins, GitHub Actions, GitLab CI, etc.)
  • Knowledge of containerization tools like Docker and orchestration using Kubernetes (preferred)
  • Familiarity with monitoring tools like Prometheus, Grafana, or Datadog
  • Strong scripting skills (Python, Bash, or similar)
  • Understanding of networking concepts (DNS, load balancing, firewalls, etc.)
  • Experience in incident management and production support
 
Preferred Qualifications:
  • Experience with SRE practices (SLIs, SLOs, error budgets)
  • AWS certifications (e.g., AWS Solutions Architect, DevOps Engineer)
  • Experience with security best practices and compliance frameworks
  • Knowledge of multi-cloud or hybrid environments

Jobseeker

Looking For Job?
Search Jobs

Recruiter

Are You Recruiting?
Search Candidates