Role Overview
As a Software Engineering evaluator, you will create cutting-edge datasets for training, benchmarking, and advancing large language models, collaborating closely with researchers. This includes curating code examples, providing precise solutions, and making corrections primarily in Java, as well as in Python, JavaScript, ReactJS, C/C++, Rust, and Go.
Responsibilities
- Working on AI model training initiatives by curating code examples, building solutions, and correcting code primarily in Java, Python, JavaScript, ReactJS, C/C++, Rust, and Go.
- Evaluate and refine AI-generated code to ensure that it is efficient, scalable, and reliable.
- Collaborate with cross-functional teams to enhance AI-driven coding solutions against industry performance benchmarks.
- Build agents that can verify the quality of the code and identify error patterns.
- Hypothesize on steps in the software engineering cycle and evaluate model capabilities on them.
- Design verification mechanisms that can automatically verify a solution to a software engineering task.
Requirements
- Several years of software engineering experience (3 years or more).
- Strong expertise in building full-stack applications and deploying scalable, production-grade software using modern languages and tools.
- Deep understanding of software architecture, design, development, debugging, and code quality/review assessment.
- Excellent oral and written communication skills for clear, structured evaluation rationales.
Skills
- Java
- Python
- JavaScript
- ReactJS
- Software Architecture