Teikametrics

Teikametrics

Teikametrics is a top Marketplace Optimization Agency leveraging AI to empower ecommerce businesses on Amazon and Walmart for strategic growth.

Media
251-1K
Founded 2015
$65M raised

Description

  • Build and maintain cloud infrastructure that hosts Teikametrics applications and platforms.
  • Develop internal DevOps tools and best practices to improve software development and deployment efficiency.
  • Design, deploy, and manage automation tools that improve predictability and reduce operational cost.
  • Manage deployment infrastructure and CI/CD automation, including GitHub workflows and related tooling.
  • Work with product development teams in a DevOps model to support releases and infrastructure decisions.
  • Create reusable and immutable infrastructure using Terraform.
  • Improve observability through monitoring, logging, and alerting systems.
  • Investigate application issues, advise teams on deployment and infrastructure choices, and support root cause analysis.
  • Participate in on-call rotations and post-mortems to support overall platform reliability.
  • Address security and scaling challenges across AWS and third-party infrastructure.

Requirements

  • 5+ years of professional experience.
  • Ability to take full ownership of significant system components and their reliability/performance.
  • Experience managing infrastructure lifecycle from design through deployment, maintenance, and optimization.
  • Familiarity with industry standards and DevOps best practices.
  • Experience supporting a platform in on-call rotations.
  • Ability to operate with minimal supervision.
  • Experience with deployment workflows and pipelines such as CircleCI, Argo Workflows, or similar tools.
  • Hands-on experience with AWS and Kubernetes (EKS, GKE, or similar container orchestration platforms).
  • Experience with infrastructure-as-code tools such as Terraform.
  • Proficiency in at least one scripting/programming language such as Bash or Python.
  • Hands-on experience with authentication and authorization technologies.
  • Experience with containerization technologies such as Docker.
  • Experience with monitoring tools such as Datadog, Opensearch, or Sentry.
  • Preferred: experience using AI agents and writing project-specific standard guidelines.
  • Preferred: experience operating data pipelines with Databricks or Kafka.
  • Preferred: experience with Java or JavaScript.
  • Preferred: experience managing infrastructure costs and budgets.
  • Preferred: experience with databases such as AWS RDS/Postgres.
  • Must be able to work US hours, 9 AM–6 PM EST.
  • Bengaluru, India-based candidates are preferred.

Benefits

  • Company equity.
  • Remote work flexibility, including work from home or from office with flexible remote options.
  • Broadband reimbursement.
  • Group medical insurance coverage of INR 7,50,000 per annum for a family.
  • Crèche benefit.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its document workflow platform reliable, resilient, and available for customers.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
21 hours, 3 minutes ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its document workflow platform reliable through incident management, observability, production support, and resilience work across services.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
21 hours, 3 minutes ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to keep its document workflow platform highly available and resilient while supporting production operations and reliability improvements.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
1 day, 20 hours ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its production document workflow platform reliable, resilient, and low-downtime for customers.

AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
1 day, 20 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers