Responsibilities
- Evaluate LLM Architecture Logic: review AI-generated explanations of model architectures, loss functions, and backpropagation for technical accuracy.
- Audit Code & Notebooks: validate ML-specific code (e.g., training loops, data preprocessing scripts, or model evaluations) for efficiency and correctness.
- Refine RLHF Frameworks: provide the high-quality human feedback necessary to align models with human intent, safety, and helpfulness.
- Analyze Model Reasoning: critically assess how an AI model navigates complex chain-of-thought (CoT) prompts and identify where the reasoning breaks down.
- Benchmark Performance: conduct comparative testing between different model outputs based on specific technical taxonomies and performance metrics.
Benefits
- Competitive pay rates
- Flexible hours
- Ability to work from home
Additional Information
- Must be prepared to complete paid tasks that require one hour of uninterrupted work
- Flexible hours
- Ability to work from home