Nebius

Nebius

Nebius enables B2B companies to build local hyperscaling cloud platforms with cost-effective GPUs, InfiniBand network, and 50% less compute cost. They offer managed Kubernetes and a launch-ready business model for innovative cloud solutions.

Internet Software & Services
51-250

Description

  • Develop and improve advanced fine-tuning methods for large language models, including LoRA-based and full-parameter approaches.
  • Identify and reduce LLM inference bottlenecks to improve production speed and efficiency.
  • Build model training and evaluation pipelines in JAX for speculative decoding and scaling experiments.
  • Experiment with different model architectures, including dense and MoE and auto-regressive and parallel approaches.
  • Investigate low-precision training and inference methods such as FP8 and NVFP4/MXFP4 for supervised fine-tuning and reinforcement learning.
  • Work on large-scale AI systems that train and deploy foundation models across text, vision, audio, and multimodal use cases.
  • Collaborate on resource allocation decisions by deriving scaling laws and evaluating hardware and performance tradeoffs.
  • Contribute to the development of reliable, high-performance infrastructure for training and deploying models at massive scale.

Requirements

  • Strong theoretical understanding of machine learning and reinforcement learning.
  • Deep expertise in modern deep learning for language processing and generation.
  • Experience training large models across multiple computational nodes.
  • Understanding of large neural network performance topics such as sharding strategies, custom kernels, and hardware features.
  • Strong software engineering skills, with Python as the primary language.
  • Deep experience with modern deep learning frameworks, especially JAX.
  • Experience with contemporary software engineering practices, including CI/CD, version control, and unit testing.
  • Strong communication and leadership abilities.
  • Previous experience working with language models or similar NLP technologies is preferred.
  • Familiarity with LLM concepts such as MHA, RoPE, ZeRO/FSDP, Flash Attention, and quantization is preferred.
  • Experience building and delivering products in a dynamic startup-like environment is preferred.
  • Experience developing large distributed systems or high-load web services is preferred.
  • Open-source projects that demonstrate engineering ability are preferred.
  • Excellent English writing, articulation, and communication skills are preferred.

Benefits

  • Competitive salary and comprehensive benefits package.
  • Opportunities for professional growth within Nebius.
  • Flexible working arrangements.
  • A dynamic and collaborative work environment that values initiative and innovation.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

AI DevOps Engineer

Zscaler 1K-5K Internet Software & Services

Zscaler is hiring a remote AI DevOps Engineer, with a hybrid preference in San Jose, to build and operate multi-cloud platform automation across AWS, GCP, and Azure while integrating AI-assisted workflows.

Argo CD AWS Azure Bash CI/CD GitOps Kubeflow Kubernetes MLflow MLOps Python Secrets Management Terraform Vertex AI
9 hours, 10 minutes ago

Senior Machine Learning Engineer, Digital Twin Platform

instacart.careers 1K-5K Internet Software & Services

Instacart is seeking a Senior Machine Learning Engineer to build and deploy real-time inventory and shelf-stocking intelligence for its Digital Twin Platform, supporting grocery products used by millions of customers.

AWS Azure Computer Vision GCP Machine Learning Python PyTorch Scikit-learn TensorFlow
10 hours, 10 minutes ago

Applied AI Engineer - India

dscout 51-250 Professional Services

Dscout is hiring an Applied AI Engineer to build, evaluate, and improve reliable LLM-powered agents and product features for enterprise UX research customers.

Machine Learning
10 hours, 10 minutes ago

Product & Technical Co-founder (CPTO) - AI Compliance Platform

FutureSight 11-50 Internet Software & Services

FutureSight is seeking a technical co-founder and CPTO to build and lead an explainable AI RegTech platform that automates compliance review for regulated financial firms and produces audit-ready evidence.

Encryption
10 hours, 25 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers