For Employers

Vetto

Post-Training Research Scientist (LLMs) — Experimental Track

Posted 4 days ago
Worldwide
5-10 years experience
Apply Now

Please mention DailyRemote when applying

?
Resume Match Score

See how much of this job your resume covers, and what’s missing.

Want a recruiter to go through it line by line?

Get professional review

Create a cover letter for this job

Upload your resume and we draft a letter for this exact role, tailored to what it asks for.

  • Tailored to this role
  • Based on your resume
  • Fully editable
AI Summary

Design and execute post-training experiments for LLMs, including SFT and preference-based methods. Translate raw expert data into training-ready datasets and iterate on reward signals and evaluation pipelines.

About us

Vetto is a global talent platform connecting top-tier professionals to high-impact AI projects around the world. Our mission is to build trust, quality, and long-term value in the AI ecosystem - for both exceptional talents and companies operating at the frontier of technology. 

About the role

This role sits at the heart of Vetto’s mission: using high-quality human data to build AI systems that make the world better. You’ll take raw expert signals and turn it into tangible model improvement, experimenting rapidly and carving new paths in post-training. With full autonomy and no production constraints, you’ll have the freedom to try unconventional ideas and see their impact quickly.

Key Responsibilities

  • Design and run post-training experiments on frontier and open-weight LLMs (SFT, preference-based methods, rubric-driven training)
  • Translate raw annotation artifacts (multi-step solutions, evaluations, adversarial prompts) into training-ready datasets.
  • Prototype new reward signals beyond pairwise preferences (rubrics, constraints, structured critics).
  • Analyze failure modes; propose data-centric fixes (sampling, curriculum, counterfactuals).
  • Build lightweight training/eval pipelines; iterate quickly.
  • Produce short internal memos: what worked, what didn’t, why.

About you

We’re looking for a researcher who thrives with autonomy, is hands-on, and brings a strong execution mindset and startup mentality. You are opinionated about data quality, pragmatic about tradeoffs, and comfortable moving quickly with incomplete information. You have strong experimental instincts — you can design, run, and interpret messy experiments and extract meaningful insights from them.

Minimum Qualification

  • PhD (or equivalent experience) in ML/AI, applied math, stats, or adjacent.
  • Hands-on experience with LLM post-training (at least one of SFT/DPO/RLHF/RLVR).
  • Solid Python + PyTorch/JAX; comfortable with training infra basics.
  • Fluent English

Preferred Qualification 

  • Worked with rubric-based evaluation or tool-augmented tasks.
  • Experience mixing synthetic and human data.
  • Familiarity with failure analysis and dataset audits.

Work Model

We operate remote-first. We focus on outcomes, not where the work is done. To support flexibility and personal choice, we maintain offices in select locations as an optional resource for the team.

Location: Flexible (EU-friendly time zones preferred)

Type: Full-time or long-term contract

Equal Employment Opportunity

Vetto is proud to be an equal opportunity employer and values diversity at our company. We do not discriminate on the basis of race, color, religion, national origin, sex, sexual orientation, gender identity, age, disability, veteran status, or any other protected characteristic.

Type: Full-time or long-term contract

Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Others jobs →

AI Response Evaluator

Freelance France, Japan, Mexico +3 more $10K – $20K Others

Proposal Lead - Government Business Solutions [543]

Full Time United States $111K - $153K per year Others

Director, Software Engineering Manager

Full Time United States $107K - $159K per year Others

Implementation Specialist - KitCheck - TRAVEL REQUIRED

Full Time United States $60000 - $75000 per year Others

Social Operations Lead - NO THIRD PARTIES

Freelance United States Others

Retail Catalog Optimization Specialist

Full Time United States $76000 - $82000 per year Others
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Featuring 218,987+ Jobs in Research Scientist

Answer easy questions

Answer easy questions

218,987+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified