Mercor
MercorVerified listing
Remote

Kubernetes Task Auditor | $70-$90/hr Remote

70–90/hr
Remote · Remote — United States
Posted September 1, 2026
hourly
3 openings

Overview

At Mercor, you will assess Kubernetes tasks that train a frontier AI lab's models. The work involves grading cluster-operation scenarios, checking manifest correctness, and judging whether failure-mode troubleshooting is realistic and complete. Each review ends with rubric-based written feedback that the AI team uses to improve model performance. This remote hourly role is open to candidates across the United States.

What You'll Do6

  • 1Review Kubernetes task prompts and their expected solutions for technical accuracy and alignment with production practices.
  • 2Trace cluster failure scenarios like CrashLoopBackOff, OOMKilled, and eviction to check if the stated root causes and fixes are correct.
  • 3Audit Helm charts and raw manifests for errors in RBAC, storage, ingress, resource limits, and security context.
  • 4Write rubric-based feedback that explains exactly what a task got right or wrong and what would make it production-ready.
  • 5Flag ambiguous or misleading tasks that could bias model evaluation.
  • 6Use Go, Python, or TypeScript to inspect or run task code when needed.

Requirements6

  • 13+ years of hands-on production Kubernetes experience with EKS, GKE, AKS, or self-managed clusters.
  • 2Strong grasp of cluster internals including CNI, DNS, ingress, persistent volumes, RBAC, and scheduling/eviction behavior.
  • 3Proven ability to debug live cluster incidents such as CrashLoopBackOff, OOMKilled, and resource exhaustion.
  • 4Comfort authoring and reviewing Helm charts and YAML manifests.
  • 5Working proficiency in Go, Python, or TypeScript.
  • 6Preferred: CKA/CKAD certification, Istio, HPA, Prometheus/Grafana, and prior SRE or platform-engineering work.

Who Should Apply

This role is for a production-tested SRE or platform engineer who can spot a wrong manifest or an unrealistic failure path at a glance. The work rewards people who enjoy writing precise, useful feedback and can apply the same rubric across dozens of tasks. Candidates who have only used Kubernetes through tutorials or managed services will find the depth required here uncomfortable. Rejections often happen when a candidate cannot describe a specific production incident they debugged, or when their Kubernetes knowledge is wide but thin on internals like RBAC, CNI, or persistent storage.

Salary Insight

The hourly rate is $70-$90, which sits above typical remote Kubernetes contract work and aligns with senior SRE or platform engineering experience.

Location

Typeremote
LocationRemote — United States
Eligible countriesUnited States
This is a remote position

Required Skills

kuberneteseksgkeakscnidnsingresspvpvcrbachelmgopythontypescriptckackadistiohpaprometheusgrafanasreplatform-engineeringtask-grading

Application Tip

In your application, include a short incident postmortem that shows how you diagnosed a production failure, including the exact kubectl commands and manifest changes you used.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

3d agoRemotehourly

SWE-Bench Task Auditor

This role puts you inside the evaluation pipeline for a frontier AI lab's models. You will audit SWE-Bench style repository tasks, checking reference patches, test harnesses, and Docker isolation for correctness and reproducibility. Your written, rubric-based feedback shapes which tasks get used for training and evaluation. The work is remote and hourly, paying $70-$90 per hour, and it demands strong open-source credentials plus fluency in Python and at least one of Java, Go, TypeScript, or C++.

70–90/hr
· 3 openings
PythonJavaGo+10 more
Mercor

Mercor

3d agoRemotehourly

AWS Serverless & Infrastructure-as-Code Task Auditor

This contract role puts you inside the evaluation loop for training data used by a frontier AI lab. You will audit AWS serverless and infrastructure-as-code tasks for architectural correctness, cross-service integration, and IaC fidelity. Expect to work through Lambda, Step Functions, DynamoDB, EventBridge, and related services, then deliver written feedback guided by a set rubric. The position is fully remote at an hourly rate.

70–90/hr
· 3 openings
AWSLambdaAPI Gateway+14 more
Mercor

Mercor

3d agoRemotehourly

ML Challenge Task Auditor

Contract reviewers with at least three years of hands-on applied ML experience will audit challenge tasks used to train and evaluate models for a frontier AI lab. The work focuses on experiment design, model-selection reasoning, and evaluation methodology, with feedback delivered through a defined rubric. Each review checks for data-quality problems such as leakage, metric gaming, and weak train/test/CV hygiene. This role is not an LLM-application-building or MLOps position; it demands reproducible critique of ML claims using frameworks like PyTorch, TensorFlow, scikit-learn, and XGBoost.

70–90/hr
· 3 openings
PyTorchTensorFlowScikit Learn+10 more
Mercor

Mercor

3d agoRemotehourly

AI Developer Trace Task Auditor

Every coding trace you receive comes from AI-assisted developer sessions. You will assess each trace for code correctness, workflow soundness, and reasoning, then write clear feedback that follows a defined rubric. A frontier AI lab uses these evaluations to train and refine its models, so your judgments shape model behavior. The role is remote within the United States and pays $70 to $90 per hour.

70–90/hr
· 3 openings
CursorGitHub CopilotClaude Code+9 more
Mercor

Mercor

9d agoRemotehourly

GPU Kernel Expert

This remote contract role puts you inside the evaluation loop for GPU/accelerator kernel tasks generated for a frontier AI lab. You will review assignments built on CUDA, Triton, NKI, and Pallas, checking whether they are numerically sound, correctly scoped, and safe to run. Your written, rubric-based feedback helps decide which tasks are used to train and evaluate the lab's models.

70–90/hr
· 3 openings
CudaTritonNki+10 more
Micro1

Micro1

1mo agoRemotefull-time

DevOps Engineer

This DevOps Engineer role is your opportunity to apply your cloud infrastructure expertise to directly influence how next-generation AI systems learn and perform. You'll design and automate scalable environments using AWS and Kubernetes, ensuring reliability and efficiency for model training workflows. No prior AI experience is needed—your proven DevOps skills are what matter most.

20–70/hr
DevOpsKubernetesAWS