Summary: The Software Engineering & Systems Design Expert role involves evaluating LLM-generated responses to software engineering queries for accuracy and clarity. Responsibilities include conducting fact-checking, executing code for validation, and annotating model responses. Candidates should possess a strong background in software engineering and programming languages, with experience in LLMs and open-source contributions. Attention to detail and the ability to assess complex technical reasoning are essential for this position.
Key Responsibilities:
- Evaluate LLM-generated responses to coding and software engineering queries for accuracy, reasoning, clarity, and completeness
- Conduct fact-checking using trusted public sources and authoritative references
- Conduct accuracy testing by executing code and validating outputs using appropriate tools
- Annotate model responses by identifying strengths, areas of improvement, and factual or conceptual inaccuracies
- Assess code quality, readability, algorithmic soundness, and explanation quality
- Ensure model responses align with expected conversational behavior and system guidelines
- Apply consistent evaluation standards by following clear taxonomies, benchmarks, and detailed evaluation guidelines
Key Skills:
- BS, MS, or PhD in Computer Science or a closely related field
- Significant real-world experience in software engineering or related technical roles
- Expert in at least one relevant programming language (e.g., Python, Java, C++, JavaScript, Go, Rust)
- Able to solve HackerRank or LeetCode Medium and Hard–level problems independently
- Experience contributing to well-known open-source projects, including merged pull requests
- Significant experience using LLMs while coding and understanding their strengths and failure modes
- Strong attention to detail and comfortable evaluating complex technical reasoning, identifying subtle bugs or logical flaws
Salary (Rate): £80.00/hr
City: undetermined
Country: United Kingdom
Working Arrangements: undetermined
IR35 Status: undetermined
Seniority Level: undetermined
Industry: IT
Software Engineering & Systems Design Expert [$45-$80/hr]
Role Responsibilities
- Evaluate LLM-generated responses to coding and software engineering queries for accuracy, reasoning, clarity, and completeness
- Conduct fact-checking using trusted public sources and authoritative references
- Conduct accuracy testing by executing code and validating outputs using appropriate tools
- Annotate model responses by identifying strengths, areas of improvement, and factual or conceptual inaccuracies
- Assess code quality, readability, algorithmic soundness, and explanation quality
- Ensure model responses align with expected conversational behavior and system guidelines
- Apply consistent evaluation standards by following clear taxonomies, benchmarks, and detailed evaluation guidelines
Good Candidature
- BS, MS, or PhD in Computer Science or a closely related field
- Significant real-world experience in software engineering or related technical roles
- Expert in at least one relevant programming language (e.g., Python, Java, C++, JavaScript, Go, Rust)
- Able to solve HackerRank or LeetCode Medium and Hard–level problems independently
- Have experience contributing to well-known open-source projects, including merged pull requests
- Significant experience using LLMs while coding and understand their strengths and failure modes
- Strong attention to detail and are comfortable evaluating complex technical reasoning, identifying subtle bugs or logical flaws
Nice-to-Have Specialties
- Prior experience with RLHF, model evaluation, or data annotation work
- Track record in competitive programming
- Experience reviewing code in production environments
- Familiarity with multiple programming paradigms or ecosystems
- Experience explaining complex technical concepts to non-expert audiences