Software Engineer - AI Code Evaluation & Benchmarking (Remote)
Remote · Full Time
Be the first to apply
- Experience
- 3+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 6 గంటలు క్రితం
- Work mode
- Work from home
- Education
- Bachelor's or master's degree in computer science or related technical field
- Eligibility
- Candidates with a bachelor’s or master’s degree in computer science or related technical fields and at least 3 years of professional software engineering experience are eligible to apply. US candidates only.
- Resume
- Required to apply
Job description
Role Summary
We are currently seeking a Software Engineer specializing in AI code evaluation and benchmarking to join our client’s team on a full-time remote basis. This position involves scrutinizing AI-generated software code to verify its accuracy, performance, and reliability, and benchmarking it against real-world software development challenges.
Key Responsibilities
- Examine AI-produced code for correctness, maintainability, efficiency, and compliance with specifications.
- Analyze engineering tasks and ensure the solutions meet the desired outcomes.
- Debug, replicate software issues, and validate resolutions in multiple programming environments.
- Evaluate explanations, reasoning, and implementation methods generated by AI models for technical precision.
- Develop and update evaluation datasets, benchmarking tools, and grading standards for coding assignments.
Required Skills and Qualifications
- Bachelor’s or master’s degree in computer science, software engineering, or similar technical discipline.
- Minimum of three years professional experience in software engineering.
- Strong command of programming languages including but not limited to Python, Java, C/C++, Go, Swift, Objective-C, PHP, or SQL.
- Sound knowledge of data structures, algorithms, software design principles, and debugging techniques.
- Experience in code reviews with an ability to identify edge cases, failure modes, and challenges AI faces in coding tasks.
Additional Information
This opportunity enables participation with a global leader in AI technology, supporting the evolution of large language models through detailed human feedback and assessments. This role is suited for engineers who possess excellent problem-solving capabilities and a passion for high-quality software evaluation.
Equal Employment
We are committed to hiring candidates based solely on skills and qualifications regardless of background or prior employment, ensuring fairness and diversity in our recruitment process.