Outpost Logo

Outpost

Site Reliability Engineer (Contract)

Posted One Month Ago
Remote
Hiring Remotely in USA
75K-90K Annually
Mid level
Remote
Hiring Remotely in USA
75K-90K Annually
Mid level
Own and improve uptime for backend services, APIs, workers, and ML pipelines; build monitoring, alerting, auto-remediation, optimize GCP infrastructure and Postgres performance, run on-call rotation and blameless postmortems.
The summary above was generated by AI

About Us:

Outpost is building the backbone of freight. We’re reinventing how supply chain infrastructure works in America with carrier agnostic truck terminals. As a vertically integrated real estate, operations, and technology company, we acquire and operate mission-critical real estate across the country to serve the largest logistics providers in the world. Backed by $1B from Greenpoint Partners, we’re scaling and building the most valuable logistics network in the country.

We thrive on accountability, integrity, and a shared drive to raise the bar. If you’re excited to reshape the industry alongside a high-performance team with a championship mindset that executes relentlessly, welcome aboard.

Role Summary:

Our platform combines AI-powered gate automation, computer vision, and operational software to help logistics operators run smarter, faster facilities. We're a small, high-conviction team shipping real software that ends up in real yards, at real gates, moving real freight; if our system goes down, trucks stop moving and a customer's yard stops running. What we build is mission-critical to the people who depend on it.

We’re scaling fast, with load expected to 10X over the next 18 months, and reliability is now core to whether customers trust us to run their gates. We need an SRE to own uptime and incident response as the system grows, and to help the team get proactive about issues instead of reactive.

Project Details:

  • Own reliability targets across our backend/API, worker services, applications and CV pipeline; MTD, MTM, MTR, and follow-through on root causes.

  • Level up our monitoring and alerting, and build out auto-remediation, so on-call load scales with automation, not headcount.

  • Partner with our agentic engineering work to build agents that triage alerts and handle routine remediation.

  • Harden and optimize our GCP infrastructure (Cloud Run, Cloud SQL, GCS) for cost and performance as load scales.

  • Own database scale and performance; connection pooling, query optimization and indexing, read replicas, and capacity planning, so Postgres doesn't become the bottleneck as data volume grows.

  • Improve the reliability of our ML training and monitoring infrastructure, in partnership with the CV/ML team.

  • Run blameless postmortems and drive fixes for root causes, not just symptoms.

  • Participate in on-call rotation.

Qualifications:

  • 4+ years in an SRE, infrastructure, or backend engineering role with production on-call ownership.

  • Deep experience with a major cloud provider (GCP preferred); compute, managed databases, object storage, networking.

  • Experience building monitoring/alerting/observability stacks (Grafana, Prometheus, Zabbix, Datadog, or similar).

  • Strong scripting/automation skills (Python, Bash, or similar).

  • Comfortable with containerized workloads (Docker) and CI/CD pipelines.

  • Track record of reducing incident volume or improving reliability metrics — not just responding to incidents.

  • Strong communication skills, comfortable working with both technical and non-technical stakeholders, know when and how to escalate urgency, and build strong working relationships across teams.

  • Strong communication skills in English — you write clearly and engage well async.

Preferred Qualifications:

  • Experience with ML/data infrastructure — training pipelines, model monitoring, feature stores.

  • Experience building or integrating AI agents for operational automation (alert triage, auto-remediation).

  • Infrastructure-as-code experience (Terraform or similar).

  • PostgreSQL performance tuning at scale.

  • Background supporting physical/IoT systems (edge devices, cameras, on-site hardware).

  • Experience with bare-metal infrastructure in colocation environments, hardware monitoring, redundancy, and failover/high-availability configuration.

Our Stack:

TypeScript / Node.js · Next.js · React · Apollo Server · Express · PostgreSQL · GCP · Docker · Python (ML/data workloads). We’re pragmatic — the right tool matters more than the familiar one.

Location:

Remote — Latin America preferred

Working Arrangements:

This role is a full-time contract position. You’ll work closely with our core engineering team — embedded in our sprints, standups, and Slack channels — but employment is managed through the agency. We’ve built this model successfully with engineers in Latin America, and it’s been a great fit for both sides.

Outpost is an Equal Opportunity Employer and Prohibits Discrimination of Any Kind.

Similar Jobs

17 Days Ago
Remote
U.S.
Senior level
Senior level
Agency • Cloud • Professional Services • Software
Improve AWS production infrastructure reliability, observability, performance, and operational maturity. Build Terraform infrastructure, enhance CI/CD, automate operational work, manage incident response and on-call operations, lead postmortems, improve application resilience, support capacity planning and database reliability, and collaborate on security hardening and compliance. Mentor engineers and promote reliability practices across the organization.
Top Skills: AWSCi/CdCircleCIDatadogGithub ActionsGitlab CiLinuxNew RelicPostgresRubyRuby On RailsSlisSlosTerraform
21 Minutes Ago
Remote
USA
180K-270K Annually
Senior level
180K-270K Annually
Senior level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Leads Shield AI’s Enterprise Data delivery organization, managing data engineering, analytics engineering, domain enablement, and governance teams. Translates enterprise data strategy into sequenced roadmaps, manages priorities, capacity, dependencies, risks, and stakeholder expectations, and ensures production-ready, governed data products. Partners with architecture, platform engineering, security, and business leaders on technical tradeoffs, data quality, lineage, reliability, and investment priorities. Also owns hiring, coaching, performance development, delivery metrics, and continuous improvement of the data operating model.
Top Skills: DatabricksDatabricks SqlDatabricks WorkflowsDelta LakeLakehouse ArchitectureUnity Catalog
22 Minutes Ago
Remote or Hybrid
United States
120K-180K Annually
Mid level
120K-180K Annually
Mid level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Owns Shield AI’s website user experience and content journeys across key audience personas. Maintains and publishes WordPress content, builds digital experiences, improves navigation and conversion paths, and partners with agencies, designers, product marketers, communications, recruiting, and business stakeholders. Monitors web performance and user behavior while ensuring accessibility, SEO, brand, quality, and performance standards. Uses AI-enabled workflows and manages multiple web updates, launches, and stakeholder requests.
Top Skills: CSSHTMLWordpress

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account