We're working with a leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems, based in San Francisco.
The Role
- Curate code examples, build solutions, and correct code for AI model training
- Evaluate and refine AI-generated code to ensure efficiency, scalability, and reliability
- Collaborate with cross-functional teams to enhance AI-driven coding solutions
- Build agents that can verify the quality of code and identify error patterns
- Hypothesize on steps in the software engineering cycle and evaluate model capabilities
- Design verification mechanisms that can automatically verify solutions to software engineering tasks
What You'll Need
- 3+ years of software engineering experience
- Strong expertise in building full-stack applications and deploying scalable, production-grade software
- Deep understanding of software architecture, design, development, debugging, and code quality/review assessment
- Excellent oral and written communication skills for clear, structured evaluation rationales
- Ability to work with Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go
What's On Offer
- Opportunity to work on cutting-edge datasets for training, benchmarking, and advancing large language models
- Flexible commitment, minimum 10 hrs/week up to 40 hrs/week (partial PST overlap required)
- Competitive hourly rate
- Fully remote work with flexible hours
Apply via Haystack today!