MLabs

MLabs

MLabs is a Haskell, Rust, Blockchain, and AI consultancy specializing in mission-critical software development, cross-team collaboration, and cutting-edge value delivery for fintech, blockchain, and information technology sectors.

Internet Software & Services
11-50
Founded 2018

Description

  • Build and maintain infrastructure for concurrent AI trading agents, including cron schedules, state files, and trailing stop processes.
  • Deploy and manage agent environments with workspace persistence, isolated sessions, and MCP server connectivity.
  • Design and operate CI/CD pipelines to ship trading skills and plugins without interrupting live trading activity.
  • Implement zero-downtime deployment strategies such as blue/green and canary releases.
  • Build alerting and monitoring across metrics, logs, and traces to detect failures, corruption, and regressions before financial loss occurs.
  • Operate and scale core platform infrastructure across Kubernetes, Redis, Postgres, ClickHouse, and Kafka.
  • Maintain blockchain node infrastructure and stable connectivity to exchange APIs and on-chain transaction systems.
  • Lead incident response, on-call practices, debugging, mitigation, and post-mortems to improve reliability.

Requirements

  • Extensive experience in DevOps, SRE, or Infrastructure Engineering, preferably in a startup environment.
  • Proven experience deploying, scaling, and debugging production workloads in AWS EKS Kubernetes.
  • Proficiency with infrastructure as code tools such as Terraform, Ansible, or equivalent frameworks.
  • Hands-on experience with Docker and Helm for packaging and deploying production services.
  • Experience operating production-grade systems such as Redis, Postgres/RDS, ClickHouse, and Kafka.
  • Strong experience with observability tools such as Prometheus, Grafana, Datadog, Loki, or OpenTelemetry.
  • Ability to debug across multiple languages, including Python, Node.js, and Go.
  • Understanding of real-time systems where latency and reliability have direct financial consequences.
  • Familiarity with blockchain node infrastructure, exchange APIs, wallet operations, and on-chain monitoring.
  • Experience managing secrets, access controls, and production hardening in sensitive environments.
  • Experience defining SLOs and building mature on-call practices.
  • Experience with OpenClaw agent deployments and workspace templates (preferred).
  • Familiarity with Model Context Protocol (MCP) server deployment and auth management (preferred).
  • Direct experience with Hyperliquid or other DEX protocols (preferred).
  • Background in fintech, market data infrastructure, or high-frequency trading systems (preferred).

Benefits

  • Competitive compensation of $120K - $150K.
  • Remote role with US-based coverage aligned to GMT timezones.
  • High-autonomy environment with significant technical ownership.
  • Opportunity to build infrastructure for autonomous AI agents.
  • Commitment to equality, accessibility, and reasonable accommodations during hiring.
  • Privacy and secure handling of applicant data during recruitment.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Manager Site Reliability Engineer

GoDaddy 5001-10000 Technology, Information and Internet

GoDaddy is hiring a Senior Manager of Site Reliability Engineering in India to lead its SRE Center of Excellence and build a global, platform-focused operations capability that improves reliability, developer productivity, and scale.

AWS CI/CD Datadog GitHub Actions Go Grafana Kafka Kubernetes Microservices PostgreSQL Prometheus Python Terraform
8 hours, 27 minutes ago

Configuration / Deployment Specialist (.NET)

AI that benefits humanity 11-50 Internet Software & Services

SquareDev (part of QnR Group) is hiring a remote Configuration / Deployment Specialist to support an EU public sector client in Poland with enterprise application build, integration, configuration, and deployment work.

Active Directory Agile Azure Azure Pipelines CI/CD Git .NET
8 hours, 27 minutes ago

Vice President, Global Production Operations & Reliability

Everbridge 1K-5K Internet Software & Services

Everbridge is hiring a Vice President, Global Production Operations & Reliability to lead the company’s global production operations for its cloud-native SaaS platform and drive reliability, scalability, security, and operational excellence.

AWS CI/CD Kubernetes
1 day, 7 hours ago

DevOps Engineer - SRE Observability

Lingaro 5K-10K IT Services

An infrastructure-focused role at Lingaro responsible for monitoring, automating, and designing cloud systems within an Azure-based environment.

Azure Azure Pipelines CI/CD Docker GitHub GitHub Actions Grafana Kubernetes MySQL PostgreSQL Prometheus SQL Terraform
1 day, 7 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers