You are viewing a preview of this job. Log in or register to view more details about this job.

Site Reliability Engineer (SRE)

Job Title- Site Reliability Engineer (SRE)

Salary Range- 65k to 70k

Location- Pheonix, AZ

 

Entry-Level Site Reliability Engineer (SRE)
Role Overview

We are looking for an Entry-Level Site Reliability Engineer to help build, operate, and improve reliable, scalable, and secure production systems. This role is ideal for an engineer with strong technical fundamentals, curiosity about how systems work, and an interest in automation, cloud infrastructure, observability, and production engineering.

You will work alongside experienced SRE, DevOps, infrastructure, and application engineering teams to monitor services, troubleshoot issues, automate operational tasks, and continuously improve system reliability.

Required Qualifications
Bachelor's or Master's degree in Computer Science, Information Technology, Software Engineering, or a related technical field.
Up to 2 years of relevant professional experience, internship experience, academic experience, or equivalent hands-on project experience.
Basic understanding of Linux/Unix systems and command-line troubleshooting.
Familiarity with at least one scripting or programming language such as Python, Bash, Go, or Java.
Basic understanding of networking concepts such as DNS, HTTP/HTTPS, TCP/IP, ports, and load balancing.
Familiarity with Git and software development workflows.
Understanding of fundamental cloud and infrastructure concepts.
Strong problem-solving and troubleshooting skills.
Ability to learn new technologies and work collaboratively with engineering teams.
Preferred Qualifications

Experience or exposure to one or more of the following is helpful but not required:

Core Java
AWS, Azure, or Google Cloud
Docker and Kubernetes
Terraform or other Infrastructure-as-Code tools
Jenkins, GitHub Actions, GitLab CI, or similar CI/CD platforms
Prometheus, Grafana, Datadog, Splunk, ELK, or similar observability platforms
SQL and basic database concepts
REST APIs and distributed systems
Key Responsibilities
Monitor production systems, applications, infrastructure, and service health.
Assist with troubleshooting incidents, performance issues, and service failures.
Participate in incident response and root-cause analysis with guidance from senior engineers.
Build scripts and automation to reduce repetitive operational work.
Create and maintain dashboards, alerts, operational documentation, and runbooks.
Support application deployments and CI/CD pipelines.
Help identify reliability risks and opportunities for automation.
Participate in on-call support after appropriate training and onboarding.
Collaborate with software engineers to improve application reliability, observability, and operational readiness.
Nice-to-Have Skills
Linux Administration
Cloud Computing Fundamentals
Scripting and Automation
Containerization Technologies
Monitoring and Observability Concepts
Infrastructure as Code (IaC)
CI/CD Fundamentals
Networking Basics
Troubleshooting and Incident Management
Version Control Systems (Git)
Work Environment

This role provides an opportunity to gain hands-on experience with modern cloud platforms, production infrastructure, automation frameworks, and reliability engineering practices while working closely with experienced engineering teams in a collaborative environment.