SRE Devops at Diverse Lynx in Sunnyvale CA, Austin TX
- Company: Diverse Lynx
- Location: Sunnyvale CA, Austin TX
- Posted: Sep 18, 2026
- Type: Contract
- Salary: $50/hr. to $65/hr.
- Experience: 4+ years
Overview
Job Title: SRE DevOps Location: Sunnyvale CA, Austin TX Pay Rate: $50/hr. to $65/hr. Position Type: Contract • We are looking for a skilled Site Reliability Engineer (SRE) to design, build, and maintain highly available, scalable, secure, and reliable cloud infrastructure and applications.
Job description
- Job Title: SRE DevOps
- Location: Sunnyvale CA, Austin TX
- Pay Rate: $50/hr. to $65/hr.
- Position Type: Contract
Responsibilities
- • We are looking for a skilled Site Reliability Engineer (SRE) to design, build, and maintain highly available, scalable, secure, and reliable cloud infrastructure and applications.
- • The ideal candidate will have strong hands-on experience with AWS, Kubernetes, Python, Linux, and cloud-native technologies.
- • You will work closely with Development, DevOps, Security, and Operations teams to improve system reliability, automation, observability, and operational efficiency.
- • Key Responsibilities Design, implement, and maintain highly available and scalable infrastructure on AWS. Deploy, manage, and troubleshoot containerized applications using Kubernetes.
- • Develop automation tools, scripts, and operational utilities using Python. Build and maintain CI/CD pipelines for reliable and repeatable application deployments.
- • Monitor system health, availability, performance, and capacity.
- • Implement observability using metrics, logs, traces, dashboards, and alerting.
- • Participate in incident response, troubleshooting, root-cause analysis, and post-incident reviews. Define and improve SLIs, SLOs, and SLAs. Automate repetitive operational tasks and reduce manual intervention.
- • Perform Kubernetes troubleshooting, including pods, deployments, services, ingress, networking, storage, and resource management.
- • Eligible to workimize AWS infrastructure for performance, reliability, security, and cost.
- • Implement infrastructure as code using tools such as Terraform or CloudFormation.
- • Establish and maintain backup, disaster recovery, and business continuity mechanisms.
- • Work with development teams to improve application reliability and production readiness.
- • Participate in an on-call rotation and respond to production incidents when required.
- • Continuously identify opportunities to improve system resilience and operational processes.
Requirements
- strong hands-on experience with AWS, Kubernetes, Python, Linux, and cloud-native technologies
Skills
Required
- AWS
- Kubernetes
- Python
- Linux
- cloud-native technologies
- CI/CD pipelines
- Terraform or CloudFormation
- incident response and troubleshooting
- on-call rotation