Responsibilities
- Review and verify the technical accuracy of AI-generated descriptions related to model architectures, loss functions, and backpropagation processes.
- Examine machine learning code and Jupyter notebooks for correctness, efficiency, and adherence to best practices, including training pipelines and data transformation logic.
- Deliver precise human feedback to improve reinforcement learning from human feedback (RLHF) systems, ensuring outputs align with safety, clarity, and user intent.
- Evaluate how artificial intelligence models process multi-step reasoning tasks, identifying logical inconsistencies or breakdowns in chain-of-thought workflows.
- Perform structured comparisons of model outputs using defined technical criteria and performance indicators to assess quality and reliability.
Benefits
- Competitive compensation for project-based work
- Flexible scheduling to accommodate different time zones and personal availability
- Fully remote work environment with no location restrictions
Compensation
Competitive pay rates
Work Arrangement
Remote (Worldwide) — Glasgow, UK
Other
- Applicants must be available to complete paid assignments requiring up to one hour of focused effort, though most tasks take less time.
- After passing the qualification assessment, onboarding is completed within 15 minutes.