SyoS LLC logo

AWS Site Reliability Engineer

SyoS LLC

RemotemidPosted 2h ago

Job description

Company Description

SyoS LLC is focused on delivering tailored technology solutions rather than “one size fits all” implementations. The company brings industry experts from multiple verticals to each engagement, ensuring that solutions align with clients’ specific business and technical needs. SyoS LLC emphasizes high-quality service, proactive support, and close collaboration throughout the entire project lifecycle. Team members are encouraged to apply their expertise, share knowledge, and continuously improve the solutions provided to clients.

Role Description

The AWS Site Reliability Engineer is a full-time remote role responsible for ensuring the reliability, performance, and scalability of cloud-based systems running on AWS. Day-to-day tasks include designing and maintaining infrastructure as code, monitoring system health, implementing automation for deployment and operations, and responding to incidents to restore service quickly. The role involves collaborating with software development teams to improve application resiliency, optimizing resource usage and costs, and enhancing observability through logging, metrics, and alerting. The engineer will also participate in on-call rotations, conduct root cause analysis, and contribute to continuous improvement of operational practices and security controls.

Qualifications

  • Candidates should possess strong Site Reliability Engineering skills, including reliability design, incident response, observability, and performance optimization.

  • Candidates should possess solid Troubleshooting skills, including diagnosing complex production issues, conducting root cause analysis, and implementing corrective and preventive actions.

  • Candidates should possess Software Development skills, such as scripting or programming (e.g., Python, Go, or similar) and experience with CI/CD pipelines and automation frameworks.

  • Candidates should possess System Administration skills, including managing Linux-based environments, user and permission management, and basic networking concepts.

  • Candidates should possess Infrastructure skills, including hands-on experience with AWS services (e.g., EC2, RDS, Lambda, VPC, CloudWatch) and infrastructure-as-code tools (e.g., Terraform, CloudFormation).

  • Experience with containerization and orchestration (e.g., Docker, Kubernetes or ECS), configuration management tools, and security best practices in cloud environments.

  • Strong communication and collaboration skills, with the ability to work effectively in a distributed remote team and partner with cross-functional stakeholders.

  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience in site reliability or cloud operations.