H

Senior Python Engineer - LLM Fine-Tuning and Benchmarking

hired • India
Remote
Apply
AI Summary

We are hiring a Senior Python Engineer to enhance large language models for a foundational LLM company. The role involves designing and maintaining Python code for model training, conducting evaluations, and leading supervised fine-tuning efforts. Candidates must have strong Python proficiency and experience with SFT and RLHF.

Key Highlights
Work with a global AI leader to advance large language models
Design Python code for training, optimization, and fine-tuning
Conduct model evaluations and benchmarking across diverse domains
Lead supervised fine-tuning and RLHF with human feedback
Key Responsibilities
Design, develop, and maintain efficient, high-quality Python code to train and optimize AI models
Conduct evaluations to benchmark model performance and analyze results for continuous improvement
Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria
Lead efforts in supervised fine-tuning, including creating and maintaining high-quality, task-specific datasets
Collaborate with researchers and annotators to execute reinforcement learning with human feedback and refine reward models
Technical Skills Required
Python Supervised Fine-Tuning (SFT) Reinforcement Learning with Human Feedback (RLHF)
Benefits & Perks
Remote work

Job Description


  • Role: Python Engineer (Remote)
  • Location: Remote (Work from Anywhere)
  • Job Type: Full-Time
  • Payout: Competitive, based on experience


Role Overview:

We are hiring for one of our clients, seeking a Senior Python Developer to assist a foundational LLM company in enhancing their large language models. The goal is to provide high-quality proprietary data for fine-tuning and benchmarking model performance.


Key Responsibilities:

• Design, develop, and maintain efficient, high-quality Python code to train and optimize AI models.

• Conduct evaluations to benchmark model performance and analyze results for continuous improvement.

• Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria.

• Lead efforts in supervised fine-tuning, including creating and maintaining high-quality, task-specific datasets.

• Collaborate with researchers and annotators to execute reinforcement learning with human feedback and refine reward models.


Required Skills & Qualifications:

• Proficiency in Python and related frameworks/libraries for AI/ML tasks.

• Experience with supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF).

• Strong understanding of evaluation strategies and benchmarking processes for AI models.

• Ability to design and implement Python code for data generation and model optimization.

• Familiarity with AI model response evaluation and ranking methodologies.


More About the Opportunity:

This role offers a unique opportunity to work with a global leader in artificial intelligence, contributing to the advancement of large language models. Candidates will collaborate with top researchers and engineers in the field.


Equal Opportunity Employer:

We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications.


Apply Now!


Similar Jobs

Explore other opportunities that match your interests

Visa Sponsorship Relocation Remote
Job Type Contract
Experience Level Not Applicable

Jobgether

India

Engineering Team Lead

Programming
•
15h ago
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Mid-Senior level

Jobgether

India

Senior Software Engineer

Programming
•
19h ago
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Mid-Senior level

epg group

India

Subscribe our newsletter

New Things Will Always Update Regularly