AI Engineer / Architect

Pave Talent
Fully remote, United States onlyContract$312,000–$624,000Posted Aug 31, 2026

About the role

𝗧𝗵𝗶𝘀 𝗶𝘀 𝗻𝗼𝘁 𝗮 𝗰𝗼𝗱𝗶𝗻𝗴 𝗷𝗼𝗯. 𝗜𝘁 𝗶𝘀 𝗮 𝗷𝘂𝗱𝗴𝗺𝗲𝗻𝘁 𝗷𝗼𝗯. 𝗬𝗼𝘂 𝘄𝗶𝗹𝗹 𝗻𝗼𝘁 𝘀𝗵𝗶𝗽 𝗳𝗲𝗮𝘁𝘂𝗿𝗲𝘀 𝗵𝗲𝗿𝗲.

You will read what a coding agent did across an entire session on a real production codebase, work out exactly where its solution falls short, and write up why. To do that you need years of hands-on engineering behind you. The deliverable is your analysis, not your commits.

𝗧𝗛𝗘 𝗖𝗟𝗜𝗘𝗡𝗧

Pave Talent is hiring on behalf of a frontier artificial intelligence (AI) research lab. This is a contract engagement through their staffing partner.

𝗧𝗛𝗘 𝗪𝗢𝗥𝗞

  • Review a coding agent's full session, not the final diff. What it investigated, what it verified, what it assumed, what it left undone
  • Locate where model-generated code breaks and explain precisely why, in writing a researcher can act on
  • Build the evaluation problems these models get tested against
  • Reproduce builds and failures locally in containers
  • Work directly with researchers on questions they are investigating right now

𝗧𝗛𝗘 𝗦𝗖𝗛𝗘𝗗𝗨𝗟𝗘, 𝗔𝗡𝗗 𝗪𝗛𝗬 𝗣𝗘𝗢𝗣𝗟𝗘 𝗧𝗔𝗞𝗘 𝗧𝗛𝗜𝗦

30 to 40 hours a week, and you choose when. Weekday evenings. Weekends. A standard Monday to Friday week if you prefer. Some people on this project keep their full-time job and run this alongside it. Others do it as a straight 40-hour week. Both work.

𝗪𝗛𝗔𝗧 𝗪𝗘 𝗔𝗥𝗘 𝗟𝗢𝗢𝗞𝗜𝗡𝗚 𝗙𝗢𝗥

  • 6+ years writing production code, hands-on. Not managing people who write it
  • Heavy code review history. You are the person on your team who catches what continuous integration (CI) and the author both missed
  • Comfort in unfamiliar repos and languages outside your daily stack. The codebases rotate weekly
  • Working fluency with Docker or equivalent, reproducing builds and failures locally
  • You write clearly. If you cannot explain why a solution is wrong, the analysis has no value
  • United States work authorization, no sponsorship

𝗗𝗢 𝗡𝗢𝗧 𝗔𝗣𝗣𝗟𝗬 𝗜𝗙 𝗔𝗡𝗬 𝗢𝗙 𝗧𝗛𝗘𝗦𝗘 𝗔𝗥𝗘 𝗧𝗥𝗨𝗘

We would rather say this plainly than waste your time or ours.

  • You have not written production code in the last two years
  • You are not willing to sit a proctored coding assessment with one attempt
  • You are not willing to verify your identity on video before submission
  • You need someone else to interview or work on your behalf
  • You are outside the United States

𝗛𝗢𝗪 𝗪𝗘 𝗦𝗖𝗥𝗘𝗘𝗡

Three steps, in this order, and no exceptions for anyone.

  • Technical video screening with us
  • Identity verification. You record a brief video holding a government photo ID. You may mask everything except your photo and name
  • A CodeSignal Industry Coding Assessment. One attempt, roughly 90 minutes to 2 hours, minimum score 500. We send prep material in advance

After onboarding there are two further assessments the lab built, over about two weeks, that determine whether you continue on the project.

𝗧𝗘𝗥𝗠𝗦

𝗣𝗮𝘆: $150 per hour, W-2. No corp-to-corp, no 1099

𝗛𝗼𝘂𝗿𝘀: 30 to 40 per week, scheduled by you

𝗟𝗲𝗻𝗴𝘁𝗵: 6-month contract. Conversion to a permanent role is unlikely, and you should hear that now

𝗟𝗼𝗰𝗮𝘁𝗶𝗼𝗻: Fully remote, United States only

Apply through LinkedIn. Answer the screening questions honestly. Every answer gets verified.

𝗣𝗮𝘃𝗲 𝗧𝗮𝗹𝗲𝗻𝘁 | 𝗛𝗶𝗿𝗶𝗻𝗴 𝗥𝗲𝗶𝗺𝗮𝗴𝗶𝗻𝗲𝗱

Responsibilities

  • Review a coding agent's full session
  • Locate where model-generated code breaks and explain why
  • Build evaluation problems for model testing
  • Reproduce builds and failures locally in containers
  • Work directly with researchers on current questions

Qualifications

  • 6+ years writing production code
  • Heavy code review history
  • Comfort in unfamiliar repos and languages
  • Working fluency with Docker or equivalent
  • Ability to write clear explanations

Benefits

  • Flexible scheduling
  • Remote work

Skills mentioned

AI AgentsModel EvaluationDebuggingGitDockerContainerizationCI/CDSoftware TestingSystem DesignIntegration Testing

About Pave Talent

Pave Talent helps high-growth companies hire the people who set them apart. The best people are rarely looking. They're already doing great work somewhere else, often for your competitors. Finding them, earning their trust, and bringing them onto your team takes more than a job post and a database. It takes a real process, genuine outreach, and human judgment. That's all we do. We're a boutique, founder-led firm that runs modern AI tooling at every step, sourcing, matching, outreach, and screening, so we move faster than firms many times our size. But the technology is the floor, not the finish. When everyone has the same tools, the difference is still the people: the ones we find for you, and the ones who decide. Your edge isn't the AI. It's the people who run it. Since 2018 we've made 1,000+ placements across autonomous vehicles, technology and software, manufacturing, life sciences, insurance, and healthcare, for everyone from 50-person startups to the Fortune 500. Over 90% of the people we place are still there and thriving, because we hire for fit, not just a keyword match. We also believe in transparency. We tell you exactly where a search stands, what it will cost, and what to expect, the same straight talk our candidates get. Most firms hide their process. We're happy to show you ours. If you're building a team that has to be exceptional, let's talk.

Staffing and Recruiting11-50 employeesSan Diego, CA