
Engineering Manager
Overview
Turing, a San Francisco-based research accelerator, supports frontier AI labs and enterprises with data pipelines and training systems. This Delivery Leader role sits in the LLM Training business, where you will manage teams of 20+ software engineers and data scientists executing Supervised Fine Tuning (SFT) and Reinforcement Learning from Human Feedback (RLHF) tasks. The work focuses on delivering datasets that hit strict quality, throughput, and cost targets. The contract runs 10 months and is open to candidates in India and LATAM countries, with a required 4-hour daily overlap with PST.
What You'll Do6
- 1Direct teams of 20+ developers working in Python, JavaScript, and Java to execute technical delivery tasks.
- 2Partner with researcher clients to align on expectations and keep them satisfied with completed datasets.
- 3Run rigorous review checkpoints that protect dataset quality before release.
- 4Move projects from kickoff to stable delivery state while owning the outcome end to end.
- 5Implement quality improvements for Python, JavaScript, and Java code produced by the team.
- 6Identify and resolve issues related to question clarity, methodology, result communication, and bugs across all team output.
Requirements5
- 1At least 8 years of professional software engineering experience, including 3+ years in an engineering management position.
- 2A track record of managing large technical teams in a delivery-focused environment, with working proficiency in Python or Java/JavaScript (both are ideal).
- 3Comfortable doing hands-on technical work to spot and resolve quality issues in code and datasets.
- 4Strong leadership and people management skills that keep large, mixed teams of engineers and data scientists motivated.
- 5Excellent communication and stakeholder management skills for working with external clients and internal partners.
Who Should Apply
The right candidate has 8+ years of engineering experience, at least 3 years managing engineers, and a history of running teams of 20+ in delivery-heavy environments. They are comfortable opening the codebase themselves to fix quality issues in Python, Java, or JavaScript, and they know how to communicate with researcher clients. This role is less suitable for hands-off managers who prefer to delegate all technical review, or for engineers who want to stay individual contributors. Common rejection reasons include showing no evidence of leading teams at that scale, failing a hands-on technical review, or struggling to explain delivery issues to stakeholders.
Location
Required Skills
Application Tip
In your application, describe a specific delivery project where you managed 20+ engineers, and include metrics for throughput, quality, or cost improvements. Be ready to walk through your own code review process in Python, Java, or JavaScript during that project.
See NearSkill jobs more often in your search
How your application is processed
1Application received
Your resume and details are logged the moment you apply.
2ATS + eligibility screening
We check your profile against the role’s skills, seniority, and requirements.
3Employer sees qualified profiles only
Only candidates who clear screening move forward.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedEngineering & Software Domain Expert
Frontier AI models improve only when someone with real engineering experience checks their work. This role puts you inside a leading AI lab's GenAI team, where you will review engineering knowledge tasks, write instruction specs and golden solutions, and build benchmarks that measure model progress. You will work in the client's own tools, on-site in the Bay Area several days each week, and your output directly shapes how the model reasons about software and systems. Employment is W-2 through Cincinnatus LLC, with client-issued accounts and equipment.

Turing
VerifiedPython Machine Learning Engineer
Turing pairs frontier AI labs with data, training pipelines, and specialized researchers, and helps enterprises move AI from prototype to production systems that deliver measurable business results. This remote contract role focuses on machine learning solution delivery using Python, with ownership across data pipelines, model design, deployment, and monitoring. The engineer sets technical direction, mentors peers, and keeps ML initiatives aligned with business priorities. Hands-on experience in Kaggle competitions or ML benchmarks is a strong signal for this position.

Micro1
VerifiedForward Deployed Engineer
This role sits at the intersection of applied AI, ML infrastructure, and partner-facing product development. You'll work directly with leading AI labs and enterprises to transform ambiguous research questions into production-grade systems. As a Member of Technical Staff, you'll own everything from data curation and LLM agent workflows to deployment and partner success — all while operating in a fast-moving, remote-first environment.

Turing
VerifiedSenior Software Engineer – LLM Evaluation (US/Canada/WEU based)
Turing, a San Francisco-based research accelerator, is hiring a contract software engineer to evaluate AI-generated code and build datasets for large language models. You will curate code examples and write precise corrections in Python, JavaScript/ReactJS, C/C++, Rust, plus Java and Go. The role includes scoring model outputs for efficiency, scalability, and reliability, designing verification mechanisms, and partnering with researchers to strengthen enterprise coding solutions. This remote engagement runs one month, offers 10 to 40 flexible hours per week, and accepts candidates in the US, Canada, and Western Europe.

Micro1
VerifiedSenior Software Engineer
This is a contract role for a Senior Software Engineer to help train the next generation of AI systems. You'll build and maintain scalable backend applications, design robust RESTful APIs, and work across AWS, GCP, and Azure to power high-performance distributed architectures. Your contributions will directly influence how AI models learn, reason, and perform using real-world data.

Turing
VerifiedSenior Software Engineer – LLM Evaluation
In this contract role, you will build and refine training datasets that help large language models improve their coding skills. You will write, correct, and evaluate code in Python, JavaScript (ReactJS), C/C++, Java, Rust, and Go, collaborating with researchers and cross-functional teams. Your daily work includes assessing AI-generated code for efficiency, scalability, and reliability, plus building verification agents that catch error patterns. The one-month engagement runs 10-40 hours per week with partial PST overlap.

