Description:
We are seeking a motivated and detail-oriented Associate Site Reliability Engineer (ASRE) to join our growing engineering team. In this role, you will work closely with developers, security professionals, and operations teams to ensure the reliability, scalability, and security of our systems and services. This is an excellent opportunity for recent graduates or early-career professionals to gain hands-on experience in a fast-paced, cloud-native environment.
The Site Reliability Engineer is a starting level position on the Site Reliability Engineer/DevOps Career Path. SRE works as part of an agile development team, but their focus is building and operating the infrastructure on which the applications run. Key areas for the role include following:
- Building and Managing infrastructure
- Operations
- Development automation
- Security
- Infrastructure Architecture and High availability
- Troubleshooting & Support
- CI/CD Pipeline Management
- Infrastructure as Code (IaC)
- Monitoring & Incident Response
- Security Integration (DevSecOps)
Key Activities
- Experience of developing, managing, and debugging high-volume, low-latency, high performance applications which are highly available
- Experience in the deployment of APIs, object-oriented design, agile development e.g. Scrum etc
- Experience designing and developing unit, component and functional tests and test infrastructure
- Ability to contribute to all phases of the product lifecycle
- Ability to multi-task
- Ability to work independently and must be self-driven
- Contribute to software life cycle strategies
- Regularly attending team meetings to discuss projects, brainstorm ideas, and put forward solutions to any issues
- A Site Reliability Engineer is not expected to work autonomously, as such their key accountability is to work with their mentor to learn and demonstrate constant progress towards the next level of the career ladder
System administration, security and networking
The Site Reliability Engineer is expected to have a good understanding of system administration (Linux or Windows) and networking.
- Essential cli commands
- Operation of Running Systems
- User and Group Management
- Knowledge of networking concepts (DNS, TCP/IP, and Firewalls)
- Experience understanding virtualization technology
- Good grasp of fundamental Security concepts
Automation and deployment technologies
- Good understanding of "infrastructure as code" principles
- Knowledge of a scripting language such Bash, PowerShell etc
- Be able to create a build and deployment pipeline using an automation server such as Jenkins or Bamboo
- Understanding of container technologies such as Docker or Podman
- Some experience with container orchestration technologies such as ECS or Kubernetes would be beneficial
Cloud technologies
- Understanding of how to design available, cost-efficient, fault-tolerant, and scalable distributed systems on AWS or GCP or Azure
- Understanding of computer, networking, storage, and database cloud services
- An understanding of the basic architectural principles of building on the public cloud
- An understanding of network technologies as they relate to the public cloud
Qualifications
- Bachelor’s degree in computer science, Information Technology, or a related field
- Basic understanding of Linux/Unix systems and networking concepts
- Familiarity with at least one programming/scripting language (e.g., Python, Bash, Go)
- Exposure to version control systems like Git
- Awareness of cloud platforms (AWS, Azure, or GCP) and containerization (Docker, Kubernetes)