Government Careers
  • AI Code Evaluator - Fully Remote | Upto $90/hr (1)

  • Visa Hunt
  • all cities, Alaska 1 United States View Map

Summary

About the jobMercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .Position: AI Developer Trace Task AuditorType: ContractCompensation: $70–$90/hourLocation: RemoteRole ResponsibilitiesEvaluate the quality and correctness of AI-assisted software-development traces to enhance model training and evaluation.Assess end-to-end coding sessions produced with AI-assisted developer tools for correctness, workflow soundness, and reasoning.Provide clear, rubric-based written feedback to improve AI model outputs .Review complex coding trajectories for alignment with best practices and correctness.Collaborate with AI research teams to ensure consistency and relevance in training data.Work independently and asynchronously to meet deadlines while enhancing AI model performance .QualificationsMust-Have3+ years professional software development.Hands-on experience with AI-assisted coding tools and agentic/spec-driven workflows ( Cursor , GitHub Copilot , Claude Code , or similar).Strong code-reading and debugging skills across full-stack or backend systems.Ability to evaluate multi-step coding trajectories for correctness and best practice.PreferredExperience with Kiro or Amazon CodeCatalyst .Prior work evaluating or grading AI-generated code .Contributions to developer tooling .Resources & SupportFor details about the interview process and platform information, please check:For any help or support, reach out to:#J-18808-Ljbffr

Job Description

About the jobMercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .Position: AI Developer Trace Task AuditorType: ContractCompensation: $70–$90/hourLocation: RemoteRole ResponsibilitiesEvaluate the quality and correctness of AI-assisted software-development traces to enhance model training and evaluation.Assess end-to-end coding sessions produced with AI-assisted developer tools for correctness, workflow soundness, and reasoning.Provide clear, rubric-based written feedback to improve AI model outputs .Review complex coding trajectories for alignment with best practices and correctness.Collaborate with AI research teams to ensure consistency and relevance in training data.Work independently and asynchronously to meet deadlines while enhancing AI model performance .QualificationsMust-Have3+ years professional software development.Hands-on experience with AI-assisted coding tools and agentic/spec-driven workflows ( Cursor , GitHub Copilot , Claude Code , or similar).Strong code-reading and debugging skills across full-stack or backend systems.Ability to evaluate multi-step coding trajectories for correctness and best practice.PreferredExperience with Kiro or Amazon CodeCatalyst .Prior work evaluating or grading AI-generated code .Contributions to developer tooling .Resources & SupportFor details about the interview process and platform information, please check:For any help or support, reach out to:#J-18808-Ljbffr

Government Careers

Government Careers

Government jobs offer stability, competitive benefits, and the chance to make a meaningful impact on your community and country.

Whether you’re starting your career or seeking new opportunities, these roles provide pathways for growth, security, and service.

Explore positions across a wide range of fields and take the first step toward a rewarding future in public service.

Show more

MORE JOBS