Site Reliability Engineer

2 days, 3 hours ago
Full-time
Senior
DevOps and Infrastructure
Raydar

Raydar

Raydar is a company that provides high-impact and results-driven services in partnership with clients like Metawork. Their team's enthusiasm and dedication set them apart, delivering tangible results and fostering long-term partnerships.

Professional Services

Description

  • Own and improve CI/CD pipelines, including caching, architecture, self-service, and deployment workflows.
  • Build developer tools and platform capabilities that reduce operational toil and accelerate delivery.
  • Extend the agent harness with continuous integration, sandboxes, guardrails, and autonomous-agent validation.
  • Operate Kubernetes-based cloud infrastructure while improving cost management, reliability, and observability.
  • Develop monitoring, alerting, and incident-response practices across logs, metrics, and traces.
  • Participate in an on-call rotation for core infrastructure, with rare non-business-hours pages.

Requirements

  • 6+ years of experience in SRE, platform engineering, DevEx, or similar software development enablement roles.
  • End-to-end ownership experience with CI/CD platforms, including caching, system design, and developer self-service.
  • Hands-on experience with TypeScript or Node.js, Python, Terraform, and Kubernetes; familiarity with Helm.
  • Practical experience with observability, including logging, metrics, tracing, monitoring, alerting, and incident management.
  • Understanding of LLMs and experience using AI development tools with sound judgment.
  • Experience working on fully remote teams.
  • Ability to pass a software-engineer-oriented technical screen.
  • Experience building platforms for autonomous AI agents or agent harnesses preferred.
  • Based in Canada or the United States and able to work U.S. time zones; new visa sponsorship is unavailable.

Benefits

  • Base salary of $160,000–$210,000 USD plus competitive equity; Canadian compensation bands are lower.
  • Medical, dental, and vision coverage.
  • Flexible vacation.
  • Monthly wellness stipend.
  • Technology and learning stipend.
  • Fully remote work across Canada or the United States.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

OMS Application Engineer

Warner Music Group is seeking an OMS Application Engineer to maintain, upgrade, and support the technical systems powering its global commerce ecosystem, with a focus on reliability, scalability, and long-term stability.

AWS CI/CD GitHub Actions Java Microservices
2 hours, 17 minutes ago

[Job-32080] Senior SRE / Cloud Engineer (Pessoa Engenheira de Plataforma), Brazil

CI&T 5K-10K Internet Software & Services

A CI&T busca uma pessoa Engenheira de Plataforma Sênior para projetar e operar a infraestrutura OCI que sustenta agentes de IA, garantindo alta disponibilidade e baixa latência.

Generative AI Kafka Kubernetes Terraform
2 hours, 17 minutes ago

Sr. Sustaining and Forward Deployed Engineer

Abacus Insights 51-250 Insurance

Abacus Insights is seeking a Senior Site Reliability Engineer – Forward Deployed to operate and improve its AWS- and Databricks-based healthcare data platform while resolving complex production issues and supporting customers.

Apache Spark AWS CI/CD Databricks Kubernetes Python Snowflake
1 day, 3 hours ago

Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)

Coforge 10K-50K IT Services

Coforge is hiring a remote Senior Site Reliability Engineer to improve application observability, reliability, and production readiness across Azure-based systems.

Azure Bash CI/CD Databricks dbt GitHub Actions New Relic PowerShell SQL
4 days, 1 hour ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers