Job Description
Responsibilities
- Write clever and tricky prompts designed to stump the AI and discover its weak points.
- Review the AI’s answers and annotate (label) them for quality, accuracy, safety, and usefulness.
- Create test cases that go beyond normal questions to find hidden problems in the AI.
- Write clear notes and reports explaining why the AI succeeded or failed.
- Work with the AI development team to share your findings and improve the models.
- Help build better training data by combining creative testing with careful evaluation.
Requirements
- 3+ years of experience in AI evaluation, testing, or data annotation.
- Strong experience in creating prompts that stump LLMs (this is the most important skill).
- Good attention to detail and consistency in reviewing AI responses.
- Excellent communication skills — you must be able to explain your thoughts clearly in writing.
- Good understanding of how AI models work and where they usually go wrong.
- Ability to be both creative (for stumping) and organized (for annotation work).
- Basic knowledge of Python or annotation tools is a plus.
Are you interested in this position?
Apply by clicking on the “Apply Now” button below!
#GraphicDesignJobsOnline
#WebDesignRemoteJobs
#FreelanceGraphicDesigner
#WorkFromHomeDesignJobs
#OnlineWebDesignWork
#RemoteDesignOpportunities
#HireGraphicDesigners
#DigitalDesignCareers
# Dynamicbrand guru