Flip App

Flip App

Flip is the employee app reshaping workplace communication by empowering every employee with a digital workspace for effective communication and workflow management.

Internet Software & Services
51-250
Founded 2018

Description

  • Expand and optimize cloud infrastructure on Azure and Kubernetes to support global growth.
  • Design and implement zero-downtime deployments, rollback mechanisms, and disaster-recovery strategies.
  • Evolve the observability stack using Loki, Grafana, Tempo, and Mimir and help define and optimize SLOs.
  • Design, develop, and optimize infrastructure as code with Pulumi in Go to reduce toil and enable self-service.
  • Promote CI/CD best practices, incident management, post-mortems, and developer experience across engineering.
  • Collaborate with the squad and engineering leadership on platform direction, including scalability, cost optimization, security, and compliance.
  • Support platform reliability through on-call participation and operational ownership.

Requirements

  • 1–3 years of hands-on experience as an SRE, Platform Engineer, DevOps Engineer, Infrastructure Engineer, Cloud Engineer, or Backend Engineer with an infrastructure focus.
  • Experience operating and scaling cloud infrastructure on Azure, GCP, or AWS.
  • Deep knowledge of Kubernetes and container orchestration in production environments.
  • Hands-on experience with observability tools such as Prometheus, Mimir, Loki, or ELK, including SLOs and error budgets.
  • Solid software development skills in Go, Python, or Kotlin; Go is preferred.
  • Experience with infrastructure as code tools such as Pulumi, OpenTofu, or Terraform.
  • Experience with configuration management tools such as Ansible or Chef.
  • Collaborative mindset, strong communication skills, and business-fluent English.
  • Willingness to participate in on-call rotations.
  • Preferred experience building and operating high-throughput, highly available systems in production.
  • Preferred experience with Azure Kubernetes Service (AKS).
  • Preferred experience with Kubernetes Gateway API and Envoy Gateway.
  • Preferred familiarity with GitOps workflows and CI/CD pipeline design.
  • Preferred knowledge of service mesh technologies such as Linkerd or Istio.
  • Preferred experience with Kubernetes Operators such as Strimzi or CNPG.
  • Preferred experience operating highly available PostgreSQL.

Benefits

  • Remote-first work with flexibility to work from home.
  • Occasional in-person collaboration in the Berlin or Stuttgart offices with advance notice.
  • Covered E-Gym-Wellpass membership and job bike leasing.
  • Relaxed working atmosphere with motivated and committed colleagues.
  • Regular team events and culture days.
  • Opportunity to shape the company and grow with a fast-growing tech organization.
  • Ability to work abroad within the European Union, subject to discussion.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Site Reliability Engineer

VantageScore 11-50 Banks

Site Reliability Engineer at a growing engineering team, focused on DevSecOps for maintaining the reliability, security, and compliance of cloud infrastructure, APIs, and software supply chains.

Agile AWS AWS CDK Bash CI/CD CloudFormation CodePipeline Datadog DevSecOps Docker EC2 GitHub Actions Grafana HashiCorp Vault Kong Kubernetes Microservices Python REST API Scrum Terraform
22 hours, 56 minutes ago

Application Site Reliability Engineer (SRE)

CXM Direct 51-250 Capital Markets

Application Site Reliability Engineer at a trading technology company, responsible for keeping .NET/C# Windows-based trading and back-office services highly reliable, observable, and resilient.

AWS Bash C# CI/CD Docker Grafana Kubernetes Microservices .NET OpenTelemetry PowerShell Prometheus Python Terraform Windows Server
23 hours, 11 minutes ago

Customer Reliability Engineer

iPiD 11-50 Internet Software & Services

iPiD is hiring a Customer Reliability Engineer to own production reliability, customer deployments, and operational excellence for its global KYP verification platform.

Ansible CI/CD GitOps Helm Kubernetes Linux Microservices Terraform
23 hours, 26 minutes ago

Site Reliability Engineer

CSC Generation 251-1K Internet Software & Services

Backcountry is hiring a Site Reliability Engineer in Costa Rica to keep its ecommerce platform reliable, scalable, and observable across a multi-cloud environment.

Ansible Argo CD AWS AWS CDK Bash CI/CD Docker GCP GitOps Grafana Helm Kubernetes Linux Node.js OpenSearch Prometheus Python Terraform TypeScript
1 day, 22 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers