Point Wild

Point Wild

Point Wild specializes in providing comprehensive online security solutions, including consumer VPN and antivirus services, while also supporting businesses with tools to enhance customer retention and address data breaches.

Internet Software & Services

Description

  • Architect and manage scalable ML infrastructure on GCP using Vertex AI, GKE, GCS, Cloud Run, and GPU/TPU compute.
  • Own the end-to-end deployment and serving lifecycle for machine learning models.
  • Build high-throughput, low-latency inference services using containers and tools such as Triton, vLLM, and MLflow.
  • Develop automated, reproducible pipelines for model training, testing, evaluation, and deployment.
  • Implement monitoring for system health and ML metrics, including drift, accuracy, and data distribution shifts.
  • Provide scalable training environments, optimized runtimes, and standardized deployment templates for AI teams.
  • Integrate ML pipelines with feature stores, dataset versioning, and batch or streaming data workflows.
  • Transition AI prototypes and notebooks into secure, resilient, auto-scaling production microservices.

Requirements

  • 5+ years of experience designing, deploying, and maintaining production ML workloads in cloud environments.
  • Deep practical experience with GCP, including Vertex AI, Cloud Storage, GKE, Cloud Run, IAM, and VPC configuration.
  • Expertise with Docker, Kubernetes/GKE, and serving tools such as Triton, vLLM, and MLflow.
  • Experience with Airflow, Vertex AI Pipelines, GitHub Actions, and ArgoCD.
  • Experience managing cloud infrastructure with Terraform.
  • Proficiency in Python and SQL for automation, API development, scripting, and data manipulation.
  • Hands-on experience with logging, telemetry, and ML observability tools such as Grafana, Prometheus, and GCP Cloud Monitoring.
  • Experience with large-scale LLM or deep learning workloads (preferred).
  • GCP Professional Machine Learning Engineer or Cloud Architect certification (preferred).
  • Familiarity with feature stores such as Feast or Vertex AI Feature Store (preferred).

Benefits

  • Opportunity to solve real customer problems in cybersecurity.
  • High individual visibility and impact in a scrappy, nimble organization.
  • Opportunities to learn new technologies, products, and markets as the company grows.
  • Inclusive workplace committed to equal opportunity and freedom from discrimination and harassment.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Manager, AI/ML Engineering

Lyra Health 1K-5K Health Care Providers & Services

Lyra Health is seeking an engineering leader to execute its machine learning roadmap by scaling and mentoring an AI/ML engineering team that delivers reliable, production-grade systems for mental health care.

AWS HIPAA Java Kotlin Kubernetes Machine Learning Microservices MLOps Neural Networks Prototyping Transformers
2 days, 3 hours ago

Senior/Staff Machine Learning Engineer (Model Dev)

Artera 51-250 Construction & Engineering

Artera is seeking an experienced machine learning engineer to lead end-to-end AI biomarker development for cancer care, from clinical problem definition and model validation through regulatory submission and production deployment.

AWS Deep Learning Kubernetes Machine Learning PyTorch TensorFlow
2 days, 3 hours ago

[Job-31696] Tech Lead – AI Orchestrator | Android, Brazil

CI&T 5K-10K Internet Software & Services

A CI&T busca uma pessoa Tech Lead – AI Orchestrator | Android para liderar a construção e modernização de jornadas de pagamentos em um SuperApp, integrando fluxos via BFF e orquestrando agentes de IA ao longo do SDLC.

Android Espresso Generative AI Jetpack Compose JUnit Kotlin REST API XML
3 days, 3 hours ago

MLOps Senior [Risk Team]

Banco Plata, S.A., Institución de Banca Múltiple. 1001-5000 Banking / financial services

As a Senior MLOps Engineer on the Risk team, you will build and operate infrastructure that industrializes the machine-learning lifecycle, from automated training to real-time production inference and feature serving.

CI/CD Docker Kubeflow MLflow MLOps Neural Networks Python SQL
4 days, 3 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers