The role involves post-training of LLMs, model alignment, server operation for checkpoint routing, and building evaluation pipelines.
Job Responsibilities:
1. Advanced post-training of large language models (e.g. SFT, RLHF/RLAIF, continual pretraining).
2. Aligning models for reliable JSON-schema function calls and external tool usage.
3. Design, deploy, and operate Model Context Protocol (MCP) servers that handle checkpoint routing, manage context windows, and enforce safety gates.
4. Experience in distributed training and inference with DeepSpeed/FSDP, LoRA/QLoRA, mixed precision, and performance tuning on vLLM or Triton clusters.
5. Build offline and live eval pipelines for alignment, factuality, grounding, and hallucinations.
Qualifications
1. Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Machine Learning, or a related field.
2. 3+ years of experience in developing and optimizing large language models.
3. Proven track record in implementing advanced post-training techniques (SFT, RLHF, RLAIF, continual pretraining).
4. Hands-on experience with distributed training frameworks (DeepSpeed, FSDP) and optimization techniques (LoRA, QLoRA, mixed precision).
5. Familiarity with model alignment, JSON-schema function calls, and external tool integration.
6. Experience in building and maintaining evaluation pipelines for model performance assessment.
7. Proficiency in Python and relevant machine learning frameworks (e.g., PyTorch, TensorFlow).
8. Strong understanding of distributed systems and high-performance computing.
9. Experience with model deployment and inference optimization on vLLM or Triton clusters.
10. Knowledge of JSON-schema and API development.
Top Skills
Deepspeed
Fsdp
Lora
Python
PyTorch
Qlora
TensorFlow
Triton
Vllm
Similar Jobs
Cloud • Fintech • Information Technology • Machine Learning • Software
The Principal Reward Partner develops and manages Xero's reward programs for global Revenue and Marketing teams, ensuring alignment with business goals and talent retention.
Top Skills:
Compensation Policy DesignMarket AnalysisProgram ManagementRelationship Management
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
The Strategic Program Manager will lead transformational programs in Customer Success, ensuring effective planning, collaboration, and operational improvements to enhance customer experience.
Top Skills:
AgileProject ManagementSaaSScrum
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
The role involves leading the Mobile Core Team and driving core platform evolution while managing mobile infrastructure engineers to enhance mobile application development.
Top Skills:
Ci/CdGraphQLJavaKotlinObjective-CReact NativeReduxSwift
What you need to know about the Dublin Tech Scene
From Bono and Oscar Wilde to today's tech leaders, Dublin has always attracted trailblazers, with more than 70,000 people working in the city's expanding digital sector. Continuing its legacy of drawing pioneers, the city is advancing rapidly. Ireland is now ranked as one of the top tech clusters in the region and the number one destination for digital companies, with the highest hiring intention of any region across all sectors.