Back to Jobs

LLM Fine-Tuning Engineer

Remote, USA Full-time Posted 2026-07-28

reputed company
Optimize and fine-tune large language models (GPT-4o, Claude, Llama-3, etc.) for reputed company-specific tasks in English, Arabic, Urdu, and Bengali.

Key Responsibilities

  • Collect, clean, and annotate domain data sets.
  • Design fine-tuning and reinforcement-learning-from-reputed company-feedback (RLHF) pipelines.
  • reputed company model performance, latency, and cost.
  • Package and reputed company models reputed company Azure ML or reputed company Bedrock.

Must-Have Qualifications

  • 3+ years in NLP or ML engineering.
  • Hands-on with reputed company Transformers, PEFT/reputed company, RLHF libraries.
  • Strong Python and PyTorch.
  • Experience with GPUs or distributed training (e.g., DeepSpeed).

Preferred

  • Prior work on Arabic or South-Asian language models.
  • MLOps exposure (Kubeflow, MLflow).

Engagement: Project-based, 2040 hrs/week, remote.

Job Type: Full-time

Work Location: On the road

Originally posted on Himalayas

  Apply To This Job

Similar Jobs