Maximum of 25 job preferences reached.
Top SRE Engineer Jobs in Denver & Boulder, CO
Cloud • Security • Software • Cybersecurity
Design, build, and operate scalable infrastructure and CI/CD/IaC systems. Implement observability (monitoring, logging, alerting), automate reliability improvements, mentor engineers, collaborate on incident response, and participate in on-call rotations to maintain Akamai Cloud services.
Top Skills:
AlertingAnsibleBashChefCi/CdGithub ActionsGitlab Ci/CdGoInfrastructure As CodeJenkinsLoggingMonitoringPuppetPythonSaltstackTelemetryTerraform
Cloud • Security • Software • Cybersecurity
Design, develop, test, and operate scalable infrastructure and services for Akamai Cloud. Implement and manage Infrastructure-as-Code (Terraform and similar tools), CI/CD, and observability. Automate reliability improvements, mentor engineers, collaborate on incident response and root-cause remediation, and participate in on-call rotations.
Top Skills:
Alerting)AnsibleChefCi/CdInfrastructure As CodeLinuxLoggingObservability (MonitoringPuppetSaltstackTerraform
Artificial Intelligence • Information Technology • Consulting
Build and operate Nebius's network infrastructure: define SLIs/SLOs, improve site and inter-site reliability, lead incident response and postmortems, develop observability and alerting, automate change workflows, and collaborate with network and platform teams to embed operability.
Top Skills:
Ci/CdContainer PlatformsGoInfrastructure As CodeLinuxPython
Software
The role involves managing compute infrastructure for decentralized applications, requiring critical thinking, documentation skills, and experience in Kubernetes and blockchain management.
Top Skills:
BlockchainGitopsInfrastructure-As-CodeKubernetesProgramming Languages
Artificial Intelligence • Information Technology • Software • Database
As a Site Reliability Engineer, you will design, implement, and maintain scalable infrastructure, ensure system reliability, automate processes, and collaborate with engineering teams.
Top Skills:
DockerElk StackGoGrafanaJavaKubernetesNode.jsPrometheusPulumiPythonRubyTerraform
Reposted 20 Days AgoSaved
Other • Social Impact
As a Senior Site Reliability Engineer, you will design, develop, and maintain reliable infrastructure for Wikimedia's API services, ensuring performance and availability while driving reliability engineering practices and improving developer experience.
Top Skills:
AnsibleArgocdAWSAzureGCPGitlabGoKubernetesOpentelemetryPrometheusPythonTerraform
Angel or VC Firm • Blockchain • Fintech • Cryptocurrency
Apply to join Galaxy Ventures' invite-only Talent Network for DevOps, SRE, QA, and Security professionals. Upon acceptance, your profile may be discreetly shared with portfolio companies for relevant roles, and you'll receive invitations to exclusive networking events. Participation is confidential and non-binding.
Healthtech • Software
Design, automate, and maintain scalable infrastructure and SRE tooling. Manage Kubernetes clusters, CI/CD, monitoring, and incident response. Improve processes, reduce toil via automation, and collaborate with engineering and data teams to support domestic and international workloads.
Top Skills:
AWSAzureContainerdDnsDockerFirewallsGCPGoGrpcHelmKubernetesLinuxLoad BalancingPrometheusPythonRoutingShell ScriptingTcp/IpUdp
Information Technology • Cybersecurity • Defense • Automation
Design, build, and maintain secure, highly available Azure cloud-native platforms and CI/CD pipelines for classified mission systems. Implement Infrastructure as Code, automation, monitoring, and security controls while supporting hybrid Windows environments and platform reliability.
Top Skills:
Arm TemplatesAzureAzure ComputeAzure DevopsAzure IdentityAzure Kubernetes ServiceAzure MonitorAzure NetworkingAzure StorageBicepC#Ci/CdDevsecopsDockerGitInfrastructure As CodeKubernetesLog AnalyticsPowershellPythonTerraformWindows Server
Artificial Intelligence • Insurance • Software • Automation
The Staff Site Reliability Engineer will build and scale infrastructure for Assured's platform, automate delivery, enhance observability, and lead mentoring initiatives.
Top Skills:
AWSKubernetesPostgresTerraform
Reposted 22 Days AgoSaved
Aerospace • Information Technology • Professional Services • Security • Software
Maintain and improve reliability, scalability, and performance of enterprise infrastructure across global sites. Implement automation and infrastructure-as-code, build monitoring and observability, perform RCA and incident response, support patching and RMF changes, integrate new capabilities, and maintain operational documentation and ITIL/ITSM processes to ensure mission-ready, high-availability environments.
Top Skills:
AnsibleElkNagiosPowershellPythonScomSolarwindsSplunkTerraform
Cloud • Security
Build and operate the production platform (Kubernetes, AWS, IaC, CI/CD, observability), automate self-service deployment, embed security and secrets management, run and modernize on-call, drive cost efficiency, mentor teammates, and maintain runbooks and post-incident reviews.
Top Skills:
AWSBashCi/CdClaudeGitGrafanaKubernetesLinuxPrometheusPythonSaltTerraform
New
Track Smarter, Apply Better.
Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.
Use For Free
Aerospace • Defense
Own and operate Loft's cloud and hybrid network infrastructure: design architectures, secure connectivity, manage network-as-code (IaC), define SLOs and observability, automate toil, and participate in SRE incident response and platform reliability.
Top Skills:
ArgocdCi/CdDnsDockerFirewallingFluxcdGCPGitGitopsGrafanaHybrid ConnectivityK8SKubernetesPeeringSdnSite-To-Site RoutingSoftware-Defined NetworkingTerraformVpcVpn
Healthtech • Social Impact • Software
Own the operational lifecycle of cloud-native data infrastructure: design and automate reliable deployments, observability, incident response, SLIs/SLOs, autoscaling and IaC, and improve platform efficiency and data freshness across GKE and Cloud Run.
Top Skills:
BashBigQueryCloud BuildCloud MonitoringCloud RunDatadogDockerGCPGithub ActionsGkeGoGrafanaJIRAKubernetesPrometheusPulumiPythonSentrySlackSnykSonarqubeTerraform
Artificial Intelligence • Other • Sales • Software
The role involves designing and advancing infrastructure for the engineering team, ensuring the reliability of Kubernetes clusters, automating operations, and building machine learning infrastructure.
Top Skills:
ArgoAWSAzureCloudFormationFluxGithub ActionsGoGCPKubernetesPostgresPythonTerraform
Software • Consulting
Lead 24x7 application support for external web applications: manage incidents, perform RCA, implement preventative fixes, build monitoring/alerts, expand Splunk functionality, create dashboards, collaborate with development and platform teams, and participate in on-call rotation.
Top Skills:
ApmAppdynamicsAWSDatadogGrafanaKubernetesLinuxMulesoftOpenshiftOpentelemetryPostmanPythonRumSeleniumServicenowShell ScriptingSplunk (Spl)Splunk CloudSplunk Observability CloudSplunk Synthetics
Software
Own and improve platform performance, reliability, and deployment automation. Manage cloud infrastructure, implement IaC, monitor systems with observability tools, provide operational support for distributed applications, and integrate production learnings into development workflows.
Top Skills:
Aiops ToolingAws Elastic ContainersAws RdsAws S3Claude CodeClaude CoworkDatadogHarness EngineeringInfrastructure As CodeKubernetesLlmsPrompt EngineeringRigorSplunk
Information Technology • Professional Services
Provide senior SRE expertise to improve reliability, scalability, performance, and resilience of a cloud-hosted geospatial platform. Design monitoring/observability, automate deployments, support incident response, optimize capacity and performance, and collaborate across DevSecOps, Kubernetes, database, and support teams in a mission-focused DoD environment.
Top Skills:
Alerting ToolsArcgis EnterpriseAw S Cloud OneAWSCi/CdContainerized SystemsEsriInfrastructure-As-CodeKubernetesLinuxLogging ToolsMonitoring ToolsRmfScriptingStig
Legal Tech • Software
Design and improve observability (monitoring, logging, tracing, SLIs/SLOs), build automation and CI/CD, lead incident response and reliability improvements, mentor SREs, run on-call, and apply AI/ML to operational signals to forecast and reduce risks.
Top Skills:
AWSBashCi/CdDistributed TracingGoInfrastructure As CodeKubernetesLoggingMonitoringPythonSlisSlos
Cloud • Security • Software • Cybersecurity
Lead reliability for a serverless AI inference platform: own observability and SLO/SLI frameworks, build automation and tooling, manage incidents and on-call, define deployment safety (canaries, rollbacks), influence architecture with product teams, and mentor other SREs.
Top Skills:
AutoscalingCi/CdContainer OrchestrationContainerizationGoGpu WorkloadsInfrastructure-As-CodeKubernetesModel ServingPythonResource Scheduling
Database • Analytics
This role involves ensuring the reliability and performance of ClickHouse's cloud infrastructure, collaborating with engineering teams, incident management, and driving continuous improvement in service availability.
Top Skills:
AnsibleAWSAzureClickhouseDocker SwarmGoGoogle Cloud PlatformKubernetesPuppetPythonTerraform
Artificial Intelligence • Marketing Tech • Mobile • Software
Lead design and implementation of scalable, reliable platform systems; define SLIs/SLOs and observability; drive cross-team strategic initiatives; mentor engineers; own production standards, incident management, and cost/operational optimization to improve platform reliability and scalability.
Top Skills:
GoJavaPythonTypescript
Information Technology
Design, build, and operate a reliable, scalable developer platform and cloud infrastructure (GCP/GKE). Lead SRE practices: SLO/SLI, observability, incident response, on-call, automation, security-by-default, Terraform IaC, CI/CD, capacity planning, mentoring, and platform enablement across teams.
Top Skills:
Ci/CdConfluenceDatadogDockerGCPGitGithub ActionsGkeGoGrafanaGsm (Gcp Secrets Manager)HclHelmHoneycombJIRAKubernetesNew RelicNode.jsOpentelemetryPrometheusPythonSslTerraform
Fintech • Real Estate • Software
Lead reliability and observability efforts across the org: design and maintain Kubernetes and AWS infrastructure, build CI/CD pipelines, drive IaC standards (Terraform/Crossplane), partner with 16+ teams to roll out tools and processes, participate in on-call rotation and incident response, and use AI tools to accelerate work.
Top Skills:
Ai ToolsArgoAurora PostgresAWSCi/Cd PipelinesCrossplaneDatadogDocumentdb (Mongo)EcsEksGithub ActionsHelmKubernetesMongoDBPostgresRdsTerraform
Hardware • Information Technology • Other • Software • Analytics
Design, deploy, and operate enterprise-grade, multi-cloud infrastructure and CI/CD pipelines; manage Kubernetes/Docker environments; integrate enterprise APIs and identity frameworks; implement observability and cost-optimization; collaborate with product, data, and security teams to drive AI-enabled automation and operational ROI.
Top Skills:
Agentic AiAWSAzureCi/CdDockerFlux CdGCPGithub ActionsGrafanaHelmInfrastructure As Code (Iac)KubernetesModel Context Protocol (Mcp)New RelicOidcOktaPrometheusSalesforceSAMLWorkday
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Denver & Boulder, CO Companies Hiring SRE Engineers
See AllPopular Denver & Boulder, CO Engineering Job Searches
Engineering Jobs in Denver & Boulder, CO
Software Engineer Jobs in Denver & Boulder, CO
Android Developer Jobs in Denver & Boulder, CO
C# Jobs in Denver & Boulder, CO
C++ Jobs in Denver & Boulder, CO
DevOps Jobs in Denver & Boulder, CO
Front End Developer Jobs in Denver & Boulder, CO
Golang Jobs in Denver & Boulder, CO
Hardware Engineer Jobs in Denver & Boulder, CO
iOS Developer Jobs in Denver & Boulder, CO
Java Developer Jobs in Denver & Boulder, CO
Javascript Jobs in Denver & Boulder, CO
Linux Jobs in Denver & Boulder, CO
Engineering Manager Jobs in Denver & Boulder, CO
.NET Developer Jobs in Denver & Boulder, CO
PHP Developer Jobs in Denver & Boulder, CO
Python Jobs in Denver & Boulder, CO
QA Jobs in Denver & Boulder, CO
Ruby Jobs in Denver & Boulder, CO
Salesforce Developer Jobs in Denver & Boulder, CO
Scala Jobs in Denver & Boulder, CO
Associate Software Engineer Jobs in Denver & Boulder, CO
Automation Engineer Jobs in Denver & Boulder, CO
Backend Engineer Jobs in Denver & Boulder, CO
Cloud Engineer Jobs in Denver & Boulder, CO
Controls Engineer Jobs in Denver & Boulder, CO
CTO Jobs in Denver & Boulder, CO
Design Engineer Jobs in Denver & Boulder, CO
DevOps Engineer Jobs in Denver & Boulder, CO
Director of Engineering Jobs in Denver & Boulder, CO
Electrical Engineering Jobs in Denver & Boulder, CO
Embedded Software Engineer Jobs in Denver & Boulder, CO
Full-Stack Engineer Jobs in Denver & Boulder, CO
Infrastructure Engineer Jobs in Denver & Boulder, CO
Manufacturing Engineer Jobs in Denver & Boulder, CO
Mechanical Design Engineer Jobs in Denver & Boulder, CO
Mechanical Engineering Jobs in Denver & Boulder, CO
Network Engineer Jobs in Denver & Boulder, CO
Platform Engineer Jobs in Denver & Boulder, CO
Principal Engineer Jobs in Denver & Boulder, CO
Principal Software Engineer Jobs in Denver & Boulder, CO
Process Engineer Jobs in Denver & Boulder, CO
Project Engineer Jobs in Denver & Boulder, CO
QA Engineer Jobs in Denver & Boulder, CO
Robotics Engineer Jobs in Denver & Boulder, CO
Security Engineer Jobs in Denver & Boulder, CO
Software Architect Jobs in Denver & Boulder, CO
Solutions Architect Jobs in Denver & Boulder, CO
Solutions Engineer Jobs in Denver & Boulder, CO
SRE Engineer Jobs in Denver & Boulder, CO
Staff Software Engineer Jobs in Denver & Boulder, CO
Systems Engineer Jobs in Denver & Boulder, CO
All Filters
Total selected ()
No Results
No Results
































