Mercor
MercorVerified listing
Remote

Agent Engineer | $100-$500 One-Time Remote

500 fixed
Remote
Posted August 13, 2026
task-based

Overview

Mercor pays LLM agent engineers $100-$500 for a 30-minute conversation about their production systems. The engagement is task-based and remote, with no coding test or take-home assignment. Conversation topics include internal monoagents wired to company data, shared company memory, reusable skills and playbooks, and the tool and MCP surfaces agents call. The focus is on how engineers measure reliability, evaluate changes, and track adoption and cost.

What You'll Do6

  • 1Describe a production agent you shipped and supported when it failed.
  • 2Explain how you determined whether a change improved or degraded agent behavior.
  • 3Discuss your experience with scheduled jobs or long-running agent workflows in cloud sandboxes.
  • 4Detail how your organization built an internal assistant and which user groups adopted it.
  • 5Share specific tradeoffs you made around tool access, memory, or MCP integrations.
  • 6Walk through how you monitored agent actions and associated costs.

Requirements6

  • 1Shipped an LLM agent that real users depended on and took responsibility for when it broke.
  • 2Built or used methods to evaluate agent performance after changes.
  • 3Operated agents that run beyond a single request, such as scheduled jobs or long-running work.
  • 4Observed internal assistant adoption patterns, including non-adoption.
  • 5Familiarity with internal monoagents, shared company memory, reusable skills and playbooks, or MCP surfaces.
  • 6Ability to articulate reliability and evaluation tradeoffs in real systems.

Who Should Apply

The ideal candidate has at least one production agent story with real users and real breakage. This role suits engineers who can explain how they evaluated agent quality, ran long-lived workflows, and handled adoption resistance. Less suitable for this role are people with only demo or prototype experience. Candidates often score low when they describe systems without specific failure modes, cost data, or user feedback. Another frequent miss is failing to name the actual tools, memory layers, or MCP endpoints involved.

Salary Insight

Mercor pays $100 to $500 for the completed 30-minute conversation, with no additional compensation. The amount varies by the depth of experience you demonstrate, and it sits near the range for paid expert interviews in the AI engineering field. Treat it as an interview honorarium, not an employment rate.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

llm agentsproduction systemsobservabilitydata integrationplaybookscost analysisincident responsecollaboration

Application Tip

Prepare one specific war story about an agent failure you debugged in production. Include the evaluation metric you used to confirm the fix, the cost or latency impact, and how many users were affected. Mention the exact MCP tools or company memory systems involved.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Turing

Turing

21d agoRemotecontract

Senior Software Engineer – LLM Evaluation

In this contract role, you will build and refine training datasets that help large language models improve their coding skills. You will write, correct, and evaluate code in Python, JavaScript (ReactJS), C/C++, Java, Rust, and Go, collaborating with researchers and cross-functional teams. Your daily work includes assessing AI-generated code for efficiency, scalability, and reliability, plus building verification agents that catch error patterns. The one-month engagement runs 10-40 hours per week with partial PST overlap.

Competitive salary
PythonJavaScriptReact+14 more
Turing

Turing

10d agoRemotecontract
Hot

LLM Trainer - Agent Function call

Work on a project with a foundational LLM company on a 6-week contractor assignment that produces high-quality proprietary data. You will design multi-turn conversations between a user and a smart assistant, playing both sides and simulating function-calling tools like calendar, email, maps, and drive. These dialogues help fine-tune models and benchmark their performance against competitors. The role demands strong technical reasoning, API fluency, and consistent adherence to an internal formatting playbook.

Competitive salary
PythonJavaJavaScript+7 more
Turing

Turing

12d agoRemotefull-time

Senior LLM Engineer

A remote role based in India centers on designing and building Generative AI and LLM systems with Python and Langchain. The engineer will create RAG pipelines, prompt techniques, and agent-based workflows that run in production. The position expects 7-12 years of experience and close work with engineering teams, business SMEs, and data teams to shape the LLM roadmap. Strong SQL and cloud familiarity across AWS, Azure, or GCP support the day-to-day work.

Competitive salary
PythonLangchainSQL+13 more
Mercor

Mercor

30d agoRemotetask-based

Software Engineer — Agentic Search Systems

We're looking for engineers with hands-on experience shipping production search systems — especially agentic ones in the age of LLMs and AI agents. This isn't a typical coding screen: it's a 25-minute conversational interview focused on how you think about search quality, evaluation, and real-world tradeoffs. Standout chats lead to a paid 30-minute follow-up conversation with our team.

150 fixed
· 10 openings
LLMAI AgentsAgentic Search+4 more
Micro1

Micro1

30d agoRemoteContract
Hot

Forward Deployed Engineer

This role sits at the intersection of applied AI, ML infrastructure, and partner-facing product development. You'll work directly with leading AI labs and enterprises to transform ambiguous research questions into production-grade systems. As a Member of Technical Staff, you'll own everything from data curation and LLM agent workflows to deployment and partner success — all while operating in a fast-moving, remote-first environment.

Competitive salary
PythonLLM SystemsML Infrastructure+1 more
Micro1

Micro1

1mo agoRemotecontract
Hot

Ai Ml Engineer for Internal Platforms Remote

Remote role building and refining an AI recruiting agent with production-ready features. You’ll push newer capabilities, test cutting edge LLMs and agent frameworks, and deploy AI-powered services that improve reliability and usefulness. Expect collaboration with product and engineering to ship tangible enhancements. You will work on backend Python services and keep pace with evolving AI tools and practices. Bold technologies: Python, LLMs, LangChain and AWS.</br>

160–300/hr
PythonLlmsAWS+2 more