About the Role
This a Full Remote job, the offer is available from: New York (USA)
We are sharing a specialised part-time consulting opportunity for experienced Software Engineers with strong code-review, debugging, and AI-assisted development experience across backend or full-stack systems.
This role focuses on evaluating end-to-end software development sessions produced with AI-assisted coding tools. Selected engineers will review code, development workflows, debugging decisions, and multi-step implementation trajectories for correctness, technical quality, and sound engineering practice, then provide clear rubric-based feedback.
Key Responsibilities
Software Development Review
Evaluate complete software development sessions from initial requirements through implementation
Assess whether resulting code is functionally and technically correct
Review implementation choices for maintainability, reliability, and engineering quality
Identify bugs, incomplete solutions, or problematic technical assumptions
Apply practical software engineering judgement across realistic development scenarios
Code Review & Debugging
Review backend or full-stack code for correctness and implementation quality
Identify logical errors, regressions, edge cases, and debugging gaps
Assess whether debugging approaches efficiently identify and resolve underlying problems
Evaluate proposed fixes for completeness and technical soundness
Distinguish substantive engineering issues from minor stylistic differences
AI-Assisted Development Workflows
Evaluate coding sessions performed with AI-assisted developer tools
Review how developers interact with coding assistants throughout implementation and debugging
Assess agentic and specification-driven development workflows
Identify ineffective prompting, unnecessary iteration, or poor use of available development context
Evaluate whether AI-generated suggestions are appropriately verified before implementation
Development Trace Evaluation
Review multi-step coding trajectories rather than isolated code snippets
Assess the sequence of investigation, implementation, testing, and refinement
Determine whether development decisions follow a coherent and effective workflow
Identify points where a stronger engineering approach should have been taken
Evaluate both final outcomes and the process used to reach them
Technical Reasoning & Best Practices
Assess engineering decisions against established software development practices
Evaluate architecture, implementation strategy, testing, and debugging choices
Review whether assumptions are appropriately validated during development
Identify unnecessary complexity or technically weak approaches
Apply professional judgement to ambiguous or imperfect coding scenarios
Rubric-Based Evaluation
Assess development traces against structured project criteria
Provide clear written explanations supporting evaluation decisions
Identify specific evidence within the development session for each judgement
Apply scoring and review standards consistently across assignments
Distinguish correct but unconventional approaches from genuinely flawed implementations
Developer Tools & Engineering Workflows
Work with development sessions involving tools such as Cursor, GitHub Copilot, Claude Code, or comparable AI-assisted coding environments
Evaluate workflows involving repositories, specifications, code changes, debugging, and testing
Assess how developers use tooling to investigate and modify existing systems
Review interactions between automated assistance and human engineering judgement
Contribute practical insight into effective developer-tool workflows
Ideal Profile
3+ years of professional software development experience
Strong background in backend or full-stack software engineering
Hands-on experience with AI-assisted coding tools such as Cursor, GitHub Copilot, Claude Code, or similar platforms
Experience with agentic or specification-driven development workflows
Strong code-reading and debugging skills
Ability to evaluate multi-step software development trajectories for correctness and best practice
Strong written communication and ability to provide precise technical feedback
Experience with Kiro or Amazon CodeCatalyst is preferred
Previous experience evaluating or grading AI-generated code is advantageous
Contributions to developer tooling are advantageous
Engagement Details
Part-time independent contractor engagement
Fully remote within the United States
Flexible scheduling based on project requirements
Compensation: $60–$80/hour
Work focuses on software development trace evaluation, code review, debugging, AI-assisted development workflows, and technical quality assessment
Projects may be extended, shortened, or concluded based on project needs and performance
Work must be completed without using confidential or proprietary information belon