ML Engineer
Undisclosed company
Mid-levelFull-timeAI & ML
Confirmed open at the employer less than an hour ago · Posted 21 hours ago
Requirements
machine learninghugging facetransformerspython
Job description
we are looking for a ML Engineer.
Own the model plane end to end - vLLM inference, fine-tuning (GRPO/RL and LoRA), and the evaluation-and-promote pipeline that takes an owned model from checkpoint to production serving. We train and serve our own models on GPUs in our own cluster.
Own the model plane end to end - vLLM inference, fine-tuning (GRPO/RL and LoRA), and the evaluation-and-promote pipeline that takes an owned model from checkpoint to production serving. We train and serve our own models on GPUs in our own cluster.
Requirements:
2+ years of ML engineering or applied ML research
Deep familiarity with HuggingFace transformers
Experience running inference servers (vLLM, TGI, or Triton)
Python and CUDA fundamentals
Understanding of quantization (AWQ, GPTQ, GGUF)
Bonus: GRPO/RLHF/DPO training, GPU workloads on Kubernetes
2+ years of ML engineering or applied ML research
Deep familiarity with HuggingFace transformers
Experience running inference servers (vLLM, TGI, or Triton)
Python and CUDA fundamentals
Understanding of quantization (AWQ, GPTQ, GGUF)
Bonus: GRPO/RLHF/DPO training, GPU workloads on Kubernetes
This position is open to all candidates.
Listing published by its original source and linked back to it. JobsWarm does not receive payment from employers.