A cutting-edge AI company is looking for experienced engineers to help shape the future of autonomous agents. This organization partners with leading AI research groups to train Large Language Models (LLMs) to function as proactive, multi-step agents, focusing on complex, real-world architectural workflows.
The Role
- Develop objective, verifiable criteria to evaluate system performance and ensure outputs meet strict functional requirements.
- Review system logs and "trajectories" to refactor code, improve execution paths, and reach a "Golden Path" of perfect reliability.
- Test systems for vulnerabilities, including improper data exposure, unauthorized access escalations, and edge-case failures.
- Contribute your expertise by training generative AI models in a freelance, fully remote, and flexible hourly role.
What You'll Need
- 2+ years of experience in backend engineering, AI automation, or complex systems integration.
- Proven ability to build and maintain production-grade software with modular separation.
- Strong command of at least two major languages (e.g., Python, JavaScript, Go, or Java) and experience with SQL databases.
- Practical experience building for live, non-mocked environments and handling multi-turn system interactions.
- Outstanding attention to detail and the ability to provide clear, high-density technical feedback on complex system behaviors.
- Expertise building multi-stage coordination tasks where data acquisition leads to reasoned output.
What's On Offer
- Opportunity to work on groundbreaking AI technology.
- Fully remote position with flexible hours.
- Competitive hourly rates up to USD $50, with additional incentives.
- Be part of a dynamic team shaping the future of generative AI.
Apply via Haystack today!