Site Reliability Engineer (Sre) - Evening Shift
By Brightspot At , Chicago $100,000 - $115,000 a year
Automate manual tasks and build tools for system monitoring, deployment, and configuration management.
2+ years of relevant experience in Cloud Operations
Proven troubleshooting and problem-solving skills in a cloud-based application environment
Outstanding communication skills with the ability to work in a client-facing role
Monitor the availability, performance, and reliability of our systems and applications during the evening shift.
Investigate and resolve incidents, troubleshooting any issues that arise and ensuring prompt resolution to minimize downtime.
Cdn Site Reliability Engineer (L5) - Open Connect
By Netflix At , Remote
Knowledge of and proven experience with CDNs and HTTP cache/proxy technologies
Service Reliability/Operational experience running large scale, high performance systems & internet services with focus on security and reliability
Expert-level knowledge of Unix or Linux system administration at scale. We happen to use FreeBSD
Knowledge of networking concepts and application protocols, especially TCP/IP, BGP, HTTP/S and DNS
Experience with distributed analytic processing technologies (Hive, Presto/Trino, Spark SQL, etc)
Some experience with container and container orchestration technologies (Docker, Kubernetes)
Sr. Site Reliability Engineer
By eHealth At , Remote $113,500 - $141,900 a year
A security certification and/or knowledge of DevSecOps would be a plus
5+ years of experience as System engineer or SRE engineer (DevOps culture)
Strong Linux skills and excellent skills in one major programming language (Python, Java would be great.)
Hands-on experience implementing and maintaining Container stack with all the security and compliance consideration.
Experience managing Hybrid infrastructure and configuration using tools like Terraform, Ansible and Puppet.
Understanding of CI/CD and experience with Jenkins, Pipeline as code
Site Reliability Engineer, Netflix Technology
By Netflix At , Remote
Experience with incident management and response
Improve our incident management lifecycle to identify, mitigate, and learn from reliability risks
Reads signals in aggregate to develop deeper insights into the quality of experience for our users to help inform business decisions
Experience with complex sociotechnical systems and their successful operations at scale
Experience conducting blame-aware incident reviews
Strong analytical and problem-solving skills
Site Reliability Engineer (Sre)
By Luxoft At , Remote
5+ years of experience with administrating Linux and at least 2 years in supporting production environments;
Fluent developer skills in any popular programming language (C++ / Python / Java / Go. Java is preferred);
Experience with designing large-scale distributed solutions accompanied with it's capacity planning;
Experience with monitoring and alerting tools like Grafana, Datadog, Prometheus etc;
Strong knowledge of virtualization and containerization principles including orchestration tools;
Experience with relational and NoSQL DBMS
Backend Engineer (Site-Reliability) Jobs
By Terraform Labs At , Remote
In-depth knowledge of database management systems, including relational databases (e.g., MySQL, PostgreSQL) and NoSQL databases (e.g., Cassandra, MongoDB, Redis).
Collaborate with cross-functional teams to understand requirements and translate them into technical designs and implementation plans.
An interest in DeFi, or background in finance / Fintech
3+ years of professional work experience
Proven experience as a Backend Software Engineer, with a focus on site reliability and DevOps.
Experience with containerization and orchestration technologies such as Docker and Kubernetes.
Site Reliability Engineer Ii
By Exact Sciences Corporation At , Remote $82,000 - $130,000 a year
Support and comply with the company’s Quality Management System policies and procedures.
3+ years of experience in systems engineering
3+ years of work and/or formal classroom experience with modern application design and cloud environments
3+ years of work and/or formal classroom experience working with software development and operations teams
1+ years of experience developing highly available systems architecture using modern technologies.
AWS Solutions Architect, AWS SysOps Administrator, or AWS Developer certification.
Site Reliability Engineer - Kubernetes
By Avantage Entertainment At , Remote $115,000 - $130,000 a year
Strong detail orientation, time management skills, dependability, and flexibility required (our team spans at least 12 time zones).
Support our DevOps team with management of application deployments using GitOps tooling in the Kubernetes environment.
Proactively researches new capabilities and trends and reports findings to senior leadership.
Bachelor's degree in computer science or equivalent occupational experience.
Experience in an AWS or other cloud environment.
In-depth experience in Kubernetes (Red Hat OpenShift preferred).
Principal Site Reliability Engineer
By GoDaddy At , Remote $168,000 - $252,000 a year
Process improvement, management, and development experience.
Translate core architecture and business requirements into technical cloud infrastructure solutions that consist of platform, network, software, cloud automation, security, etc.
3+ years of experience in complex distributed networking, system performance tuning, and monitoring.
Experience with CI/CD development using Kubernetes, Docker, etc.
Experience in virtualization technologies such as KVM, and OpenStack.
Experience with back-end services, highly distributed and scalable services, and deployment automation.
Site Reliability Engineer *Sre*
By Synchronoss Technologies At , Remote
Proven ability to deliver a superior operations support experience working directly with corporate clients’ technology teams and associated change management.
Experience in monitoring tools such as Prometheous, Thanos and Grafana
Experience with Terraform and Ansible.
Experience with Cloud platforms such as AWS
Excellent verbal, written and analytical skills, with the ability to tailor communication to the intended audience.
Experience working with ticketing systems.
Site Reliability Engineer * Sre*
By Synchronoss Technologies At , Remote
Experience with Configuration Management Automation tools (chef or puppet).
Deploy and manage Kubernetes (EKS) based docker applications in AWS/OCI.
Solid experience in building a solution on AWS or Oracle Cloud or other public cloud services using Terraform.
Knowledge in Infrastructure monitoring tools (ELK stack, Prometheus, Grafana, or similar)
Knowledge of AWS/OCI best practices. Very keen to learn new technologies, Flexible to work on new platforms/environments and models like Agile/Scrum.
Excellent written and verbal skills.
Sre - Site Reliability Engineer (Ambra Team)
By Intelerad At , Remote
Experience with Systems Lifecycle Management Products (Foreman, Katello, RedHat Satellite)
Demonstrated knowledge of configuration management tools like Puppet, Chef and Ansible
Own system designs, documentation, platform management, and capacity planning for Enterprise Imaging Systems in your area of responsibility
University or college education in science, technology, engineering, or equivalent industry experience
Build software and systems to manage platform infrastructure and applications
Excellent verbal and written communication skills and ability to communicate technical subjects to a broad range of stakeholders
Site Reliability/Devops Engineer
By Axoni At , Remote
Experience with automation and configuration management tools (Terraform, Ansible, Salt, Chef, Puppet)
Experience troubleshooting issues on a remote distributed system
Manage and configure all pre-production, production, and client facing infrastructure
Coordinate with the Applications team to satisfy all non-functional project requirements (security, performance, scalability, and resiliency)
Experience with at least one of the following scripting languages: Bash and/or Python
Experience with Docker (Docker compose, yamls, etc)
Aws Site Reliability Engineer
By Derivative Path At , Remote
Excellent communication, organizational and time-management skills
Work closely with architects, software engineers, quality engineers, product owners, and management to design scalable, robust systems using cloud architecture
Participate in system design consulting, platform management, and capacity planning
Proficient with AWS certification preferred
Prior experience within the Capital Markets, Financial Services, and IT & Services
Design and implement fully automated CI/CD Pipelines using industry tools
Software Engineer, Site Reliability
By Packback Inc At , Remote $108,000 - $140,000 a year
2+ years of devops experience using Docker and Kubernetes
Experience with CI/CD pipelines, containerization, and orchestration
Experience reviewing code to both give and receive constructive feedback.
Experience with helm and terraform
Experience working on highly scalable cloud infrastructures
Startup or small company experience
Senior Site Reliability Engineer
By Lumin Digital At , Remote $170,000 - $200,000 a year
Expert-level knowledge of at least one configuration management system (Chef, Ansible, Puppet, etc.).
Exceptional full stack and environment troubleshooting skills.
Exceptional written and verbal communication skills.
Experience with a microservice architecture running in containers (Docker or other containerization technology).
Experience with Terraform and Kubernetes
2+ years of experience as a software engineer. C#, Angular, JavaScript preferred.
Aws Site Reliability Engineer
By Zeektek At United States
Help set up and manage our AWS EKS environment.
Help set up and manage our GitLab CI/CD pipeline.
Can engage and manage the heterogenous CI/CD and deployment environments of the teams we collaborate with
Site Reliability Engineer, DevOps manager
1.5+ years experience in SRE/DevOps or equivalent role
Work with other teams to assist in deploying our microservices and code into their environments (on prem and AWS)
Staff Site Reliability Engineer, Multi-Cloud
By Okta At ,
Extensive experience with configuration management tools like Chef, Ansible, or Puppet and infrastructure-as-code tools such as Terraform
Experience with multi-cloud infrastructure is desired
Proficiency in distributed systems design, with a comprehensive understanding of failure modes, benefits, and potential drawbacks
In-depth knowledge of various types of data stores, including both SQL and NoSQL
Core contributor driving Okta’s multi-cloud initiatives
Design, build, and operate Okta's global production infrastructure
Site Reliability Engineer Jobs
By Adobe At , Lehi, 84043 $92,100 - $161,000 a year

What you need to succeed:

An understanding of SRE standard methodologies:

Site Reliability Engineer, Product - Usds
By TikTok At , Los Angeles $119,000 - $289,000 a year
Gain a solid understanding of the various components and services that power the TikTok experience
Maintain services to meet service-level-agreements (SLAs) and service-level-objectives (SLOs) by measuring and monitoring availability, performance, and overall system health
Scale systems sustainability through mechanisms such as automation; evolve systems reliability, efficiency, and velocity by pushing for changes
Provide user support, incident responses and postmortems
In this role, you will:
Our time off and leave plans are:
Site Reliability Engineer Jobs
By Fisker Inc At , Manhattan Beach $60,900 - $169,650 a year
Experience with artifact management (Artifactory, Nexus)
Experience with strict security requirements and implementation
Design, provision, deploy, and manage Kubernetes clusters and resources
Bachelor’s degree in computer science or related technical field or equivalent experience
5+ years of SRE / DevOps Engineer experience
Experience with cloud infrastructure (AWS, GCP, Azure)
Site Reliability Engineer Jobs
By Zscaler At , San Jose
Strong Centos/UNIX skills, FreeBSD specific experience is a plus.
5 -7 years experience in a SaaS/ Cloud/Distributed environment growing at a rapid scale.
Minimum 3+ years of scripting experience in Python is required.
Hands-on experience with infrastructure as code and automation tools (Ansible, Chef, Puppet, Terraform).
Basic Networking skills (TCP/IP, DNS, LACP, CARP) for testing and troubleshooting are required.
Competitive salary and benefits, including equity
Site Reliability Engineer (Sre)
By Agama Solutions At , San Jose
5+ years of US experience as in a SRE role
Good communication (and listening) skills.
Some experience administering Linux “web” servers, at scale.
Working knowledge of DNS, HTTP, TLS, web security.
Experience with networking troubleshooting using tools such as TCP Dump.
Well versed in *nix Operating Systems (we use CentOS and Ubuntu LTS).
Site Reliability Engineer Jobs
By Ascendion At , Alpharetta
Knowledge of the cloud and managed services such as MS Flex Server or AWS RDS.
Strong experience as a database administrator.
Strong experience in PostgreSQL and/or MySQL.
Automation skill in Bash, Golang, Python a plus.
Knowledge of IaC and CI/CD tools such as Terraform and GitHub Actions a plus.
Experience in query optimization and performance improvement.
Site Reliability Engineer Jobs
By eBay At , San Jose, 95125, Ca $168,400 - $262,900 a year
Develop automation systems for implementing eBay Traffic management
Manage eBay’s traffic infrastructure including SLB, CDN, etc.
Solid programming experience in languages like Golang, Java, C/C++
Experience with Kubernetes, docker is a must
Experience working with public cloud is a plus
Experience in software load balancer(IPVS, Envoy, Istio, Cilium etc) is a plus
Site Reliability Engineer, Systems
By Anthropic At , San Francisco, Ca
Automate operations and infrastructure management
Have significant experience with Kubernetes and cloud-native infrastructure
Have strong communication skills to work with a range of technical and non-technical colleagues
Python and Linux SysAdmin skills
Significant experience with Kubernetes architecture and administration
Strong Linux skills and cloud infrastructure expertise
Site Reliability Engineer (L4/5) - Core
By Netflix At , Los Gatos, Ca
Experience in risk management and/or analysis
Improve our incident management lifecycle to identify, mitigate, and learn from reliability risks
Read signals and metrics to develop deeper insights into our customers’ quality of experience to help inform business decisions
Strong writing and presentation skills
Development experience with Java, JavaScript/Node.js, Python, Go
Knowledge of cloud platforms (i.e. AWS, GCP, etc.) and microservices architecture
Sr. Site Reliability Engineer
By CCC At , Chicago, Il
Experience preparing and presenting operational artifacts to senior management
Gain and disseminate knowledge of our complex applications
2+ years experience working with the Azure tech stack in a production capacity
5+ years operational experience working with Microsoft technologies
Comfort and experience with Ops environment growing at a rapid scale.
Knowledge of Virtualization, Cloud Infrastructure and APIs
Site Reliability Engineer Jobs
By Nike At Beaverton, OR, United States

This overview explains our hiring process for corporate roles. Note there may be different hiring steps involved for non-corporate roles

Site Reliability Engineer (Sre) - $700,000
By Thurn Partners At New York, NY, United States
3+ years' experience in a similar software engineering or site reliability engineering position
Experience with SQL database operations
Experience with Kafka, CICD pipelines and virtualisation a bonus
Beautiful office space with generous overall benefits package
Proficiency with either Python or Golang
Extremely competitive compensation including performance bonuses
Site Reliability Engineer Jobs
By Spotify At Greater Chicago Area, United States
• 4+ years of IT experience needed
• Experience working in a Linux environment
• Good knowledge of Unix
• Basic experience in writing SQL queries
• Good verbal communicative skills
• Ability to manage priorities and deadlines
Site Reliability Engineer (.Net Engineer)
By Suzy At United States
Exposure to a Configuration Management System (Puppet, Chef, Salt, etc)
Optimize: Observe and improve performance, reduce cost, and improve the experience for millions of users
3+ years of experience in Software Engineering, Site Reliability Engineering, or a Development focused DevOps role.
Experience with Kubernetes and Cloud systems
Experience with the development and operation of high-traffic backend systems
Troubleshooting skills that span applications, networking (TCP/IP), and systems
Site Reliability Engineer - Usds
By TikTok At Seattle, WA, United States

Responsibilities TikTok is the leading destination for short-form mobile video. Our mission is to inspire creativity and bring joy. TikTok has global offices including Los Angeles, New York, ...

Site Reliability Engineer Jobs
By Therapy Brands At Birmingham, AL, United States
2+ years of experience programming or scripting. C# or Python is preferred.
1+ years of experience with cloud environments: AWS and Azure
1+ years of experience with SQL: writing basic select and update statements
Primary Responsibilities Of This Position
Familiarity with networking fundamentals: TCP/IP, DNS resolution
Familiarity with tools including or similar to: Grafana, InfluxDB, OpenTelemetry
Site Reliability Engineer Jobs
By Xforia Global Talent Solutions At United States
Support system design consulting, platform management, and capacity planning
Excellent communication skills and a high degree of technical leadership skills.
As Site Reliability Engineer you will:
Support the production environment by monitoring availability and the system health.
Improve reliability, quality, and time-to-release of the changes.
Provide primary operational support and engineering for multiple large-scale distributed software applications.