Applying through our link may earn us a small commission — at no extra cost to you.
About this role
Senior Machine Learning Expert (AI Training)
About the Role
What if your deep understanding of machine learning could directly influence how the next generation of AI systems reason, plan, and make decisions? We're looking for a Senior Machine Learning Expert to help build the training data that teaches large language models to think more reliably — step by step, in real-world scenarios.
This is a fully remote, flexible contract role built for senior ML practitioners who want meaningful, high-impact work on their own schedule.
- Organization: Alignerr
- Type: Hourly Contract
- Location: Remote
- Commitment: 10–40 hours/week
What You'll Do
- Author complex, high-fidelity reasoning traces that capture how an LLM should plan, use tools, and make decisions across sophisticated technical tasks
- Design and document structured step-by-step thought processes that help AI navigate intricate, real-world problem-solving scenarios
- Review and mentor the quality of reasoning traces — ensuring logical rigor, clarity, and optimal decision-making documentation
- Develop data strategies that improve how LLMs handle ambiguity, multi-step reasoning, and tool use
- Apply senior-level architectural insight to ensure the training data you produce drives more reliable and trustworthy model behavior
Who You Are
- Experienced in machine learning, AI research, or a closely related technical discipline
- Skilled at decomposing complex problems into clear, structured, logical steps
- Deeply familiar with how large language models are trained, evaluated, and improved
- Comfortable thinking rigorously about model behavior, edge cases, and failure modes
- Detail-oriented with a consistent, methodical approach to quality
- Able to work independently and asynchronously without hand-holding
Nice to Have
- Prior experience in data annotation, data quality assurance, or AI evaluation systems
- Top-tier Kaggle competition results (Grandmaster or Master level), demonstrating mastery of model performance and feature engineering
- Familiarity with agentic AI frameworks, tool-use architectures, or chain-of-thought prompting methodologies
- Background in NLP, reinforcement learning, or AI safety research
Why Join Us
- Work directly on cutting-edge AI projects alongside leading research labs and teams pushing the frontier of what LLMs can do
- Fully remote and async — work when and where it suits you, with no fixed hours
- Freelance autonomy with the structure of impactful, task-based work that genuinely matters
- Gain deep exposure to how frontier AI models are trained and evaluated at the highest level
- Potential for ongoing work and contract extension as new projects launch
About Alignerr
A legitimate, well-funded platform (built by Labelbox) with strong rates for experts. The real catch is availability — per-approved-task pay and quiet stretches between projects.
Trust score 7.6/10Read our Alignerr review →
Next steps
AITrainerGigs aggregates this listing from Alignerr.