Micro1Verified Source
RemoteHotAI & Machine LearningSoftware Engineering

Member of Technical Staff, Coding Research | $8-$9/hr Remote

$8–$9/hr
Posted June 8, 2026
full-time

Overview

This role sits at the intersection of AI research, software engineering, and model evaluation. You'll design the benchmarks, methodologies, and data systems that define how next-generation coding models are measured and improved. It's a chance to directly shape the capabilities of frontier coding agents in a remote-first, research-driven environment.

What You'll Do8

  • 1Own the design and maintenance of evaluation frameworks for coding agents, including benchmark specs, scoring rubrics, and quality standards.
  • 2Lead end-to-end research initiatives that measure and enhance coding model performance across a wide range of software engineering tasks.
  • 3Build high-quality datasets, golden examples, and evaluation protocols to enable reliable assessment of cutting-edge coding systems.
  • 4Analyze model behavior and failure modes to identify systematic weaknesses and translate findings into actionable improvements for training and evaluation.
  • 5Develop tooling and infrastructure that support large-scale experimentation, data generation, review workflows, and evaluation pipelines.
  • 6Establish best practices for coding-agent assessment — ensuring methodological rigor, reproducibility, and measurement quality.
  • 7Collaborate closely with researchers, engineers, and applied AI teams to design experiments and evaluate emerging model capabilities.
  • 8Contribute to technical reports, benchmark studies, and client-facing research that communicate model performance and insights.

Requirements8

  • 1Strong software engineering background with expertise in Python, C++, or similar languages.
  • 23+ years of experience in software engineering, machine learning, AI research, evaluation, or related technical fields.
  • 3Proven experience designing, reviewing, or validating technical assessments, benchmarks, coding tasks, or evaluation methodologies.
  • 4Familiarity with large language models, coding agents, reinforcement learning, model evaluation, or related AI systems.
  • 5Ability to build tooling, automate workflows, and improve technical processes through systematic experimentation.
  • 6Strong analytical skills to investigate model behavior and extract insights from complex technical systems.
  • 7Excellent written and verbal communication, including the ability to explain technical findings to varied audiences.
  • 8Comfortable working in fast-moving research environments with significant ambiguity and evolving priorities.

Who Should Apply

We're looking for someone who thrives at the intersection of AI research and software engineering — a builder who cares deeply about how coding models are evaluated and improved. You should have a track record of designing rigorous benchmarks or datasets, a curiosity about model behavior, and the ability to drive ambiguous technical projects from idea to execution. Ideal candidates are comfortable wearing both research and engineering hats and are excited by the challenge of measuring frontier AI capabilities.

Salary Insight

The base salary for this full-time position ranges from $140,000 to $180,000 USD annually, plus eligibility for equity compensation and performance-based bonuses. micro1 also offers a comprehensive benefits package including up to 100% reimbursement for health-insurance premiums, paid time off, a 401(K) plan with company match, and additional remote-first perks.

Required Skills

LLMsCoding EvaluationAI EvaluationML Systems

Application Tip

When applying, emphasize any hands-on work you've done designing evaluation benchmarks or building coding agent test suites. Concrete examples of how your frameworks improved model performance or caught failure modes will make your application stand out.

Share:

Similar open positions

Explore active roles that match your skills and interests.

Micro1

26d agoRemote
Hot

Member of Technical Staff, Research Engineering | $7-$8/hr Remote

This role places you at the cutting edge of Reinforcement Learning, where you'll build novel training environments, scalable pipelines, and robust evaluation frameworks that push the boundaries of AI. You'll bridge the gap between research and production, turning experimental concepts into high-performance systems that power real-world applications. As a key member of the research engineering team, you'll own the full lifecycle of RL system development, from environment design to model fine-tuning.

$7–$8/hr
Reinforcement LearningML-Oriented Data DesignRL Environments+1 more

Micro1

1mo agoRemote

Coding Expert (Multi-Language) | $15-$40/hr Remote

This is a remote contract role where your deep coding expertise helps train the next generation of AI systems. You'll evaluate real-world coding challenges across multiple languages, ensuring they meet high standards for accuracy and efficiency. No prior AI experience is needed—just strong software engineering skills and a passion for clean, maintainable code.

$15–$40/hr
PythonJavaScriptTypescript+8 more

Micro1

1mo agoRemote

Member of Technical Staff, Medical Research | $7-$8/hr Remote

Join a team focused on pushing the boundaries of AI in healthcare by designing evaluation frameworks and benchmarks for medical reasoning systems. As a Member of Technical Staff, you'll help shape how AI supports clinical decisions, evidence synthesis, and biomedical research. This remote role blends applied research with rigorous quality standards to build trustworthy healthcare AI.

$7–$8/hr
Medical/Health ResearchEnterprise AIBenchmarking+1 more

Micro1

1mo agoRemote

Member of Technical Staff, Legal Research | $7-$8/hr Remote

This role sits at the crossroads of advanced legal reasoning and cutting-edge artificial intelligence. As a Member of Technical Staff, you'll design rigorous evaluation frameworks to measure and improve how AI systems handle complex legal tasks — from statutory interpretation to contract analysis. You'll collaborate with researchers and engineers to push the boundaries of enterprise AI, shaping standards that define what reliable legal AI looks like.

$7–$8/hr
Legal ResearchEnterprise AIBenchmarking+1 more

Micro1

15d agoRemote
Hot

Member of Technical Staff, Finance Research | $6-$8/hr Remote

This role sits at the cutting edge where artificial intelligence meets finance. As a Member of Technical Staff on the Finance Research team, you'll design and own the evaluation frameworks that measure how well AI systems handle complex financial tasks—from reasoning and decision-making to full workflow automation. You'll work closely with researchers and engineers to push the boundaries of what AI can do in enterprise finance, turning cutting-edge research into tangible improvements.

$6–$8/hr
Financial ResearchEnterprise AIBenchmarking+1 more

Micro1

1mo agoRemote

AI Evaluation Specialist | $20-$35/hr Remote

As an AI Evaluation Specialist, you'll help train next-generation AI systems by designing and executing hands-on evaluation tasks. Your insights will directly shape how models learn, reason, and perform on practical computer-based workflows. This is a fully remote contract role where meticulous observation and clear documentation are key.

$20–$35/hr
· 50 openings
rubric-based evaluationstructured observation and reportinghigh attention to detail+2 more