Snr. Site Reliability Engineer (Remote)
Core
Build highly scalable, cloud-native, and resilient applications and infrastructure in AWS to ensure platform reliability, security, and efficiency.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Cloud-native distributed systems and infrastructure on AWS
Domain
Cybersecurity / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
AWS cloud architecture, Infrastructure as Code (Terraform), CI/CD pipeline design, observability and monitoring, distributed systems debugging, Python scripting
Preferred skills
AWS Professional Certification, GCP/Azure experience, secure coding practices, public company experience
Technologies
AWS (ECS, Lambda, DynamoDB, etc.), Terraform, GitLab, DataDog, Docker, Python, Ruby, Rust
Responsibilities
Design globally distributed systems, maintain deployment strategies, act as escalation point for production incidents, define solutions for complex technical problems, improve observability patterns
Seniority
Senior, hands-on IC