Remote job
Go Engineer, AI Code Reviewer
Job details
About this role
Role overview
This contract position centers on evaluating how AI coding agents behave during real-world developer interactions. Rather than writing production code, the role focuses on whether the model demonstrates sound engineering judgment, useful reasoning, and a trustworthy tone. The work is part-time and well suited to senior engineers who want to influence how AI-assisted development tools are judged and improved.
Responsibilities
- Review end-to-end AI coding interactions and assess whether responses are coherent, useful, and reflect strong engineering judgment. - Judge the quality of explanations and reasoning, not just whether generated code compiles or passes tests. - Distinguish between tiers of response quality and articulate what separates weaker outputs from strong ones. - Provide direct, opinionated written feedback on what works, what does not, and what feels misleading or off-target. - Help define the standard for "good" when interacting with AI-first developer tools such as Cursor-class IDEs. - Make rigorous subjective calls about tone, helpfulness, and whether an interaction would build or erode a developer's trust in the model.
Requirements
- Staff or Principal-level engineering background, or equivalent industry experience. - Strong hands-on proficiency with TypeScript/JavaScript or Python. - Direct experience using modern AI coding agents such as OpenAI Codex, Claude Code, and Cursor. - Deep familiarity with AI-assisted development workflows and how developers actually use these tools. - Ability to evaluate generated code without needing to execute or line-by-line review every snippet. - Comfort providing candid, opinionated feedback and holding a high bar for engineering quality.
Nice to have
- Prior exposure to prompt design or evaluation pipelines. - Experience mentoring senior engineers or defining engineering standards across a team.
Benefits and work setup
- Contract engagement at $100–$200/hour. - Approximately 10–20 hours per week, supporting a flexible side commitment alongside other work. - Engagement runs through early May, with possible extension depending on fit and need. - Remote-friendly format, given the evaluative and asynchronous nature of the work. - Hiring process includes a take-home evaluation exercise followed by a single behavioral interview.