Responsibilities
- Develop standardized CI/CD pipelines for training, evaluating, and deploying models using tools like Langfuse, GitHub Actions, and experiment tracking platforms.
- Automate processes for model version control, approval routing, and compliance validation across development, staging, and production environments.
- Design and implement a flexible, future-proof AI infrastructure stack including vector databases, feature stores, model registries, and monitoring solutions.
- Collaborate with engineering and data science teams to integrate AI models and autonomous agents into live, real-time systems and workflows.
- Evaluate and adopt cutting-edge AI frameworks and libraries such as LangChain, LlamaIndex, vLLM, MLflow, and BentoML to enhance platform capabilities.
- Lead initiatives to ensure AI system reliability, governance, and security while supporting rapid experimentation and high availability.
- Improve the performance, scalability, and efficiency of AI and machine learning models in production.
- Maintain high standards of data integrity, consistency, and reliability to support accurate model training and inference.
- Deploy systems enabling both offline and online evaluation of large language models and agents, including regression testing, cost tracking, and human-in-the-loop validation.
- Empower research teams with self-service sandboxes, dashboards, and reproducible environments to accelerate development cycles.
Work Arrangement
Hybrid — Los Angeles, San Francisco, New York, Washington D.C., London, Singapore
Learn about TRM Speed in this position
- Engineers resolve critical issues within minutes to hours, using rapid-response virtual war rooms and delivering fixes with full transparency to stakeholders within 48 hours.
- Teams proactively navigate procedural obstacles, build trust with decision-makers, and identify alternative paths to keep projects advancing in complex settings.
- Real-time documentation and knowledge sharing ensure all team members—onsite and remote—are aligned on plans, blockers, and resolutions, reducing delays and improving delivery speed.
About TRM's Engineering Levels
- Engineer: Contributes to defining project milestones and independently executes small-scale technical decisions with balanced tradeoffs. Mentors junior engineers and improves operational efficiency through code quality and knowledge sharing.
- Senior Engineer: Designs and documents system-level improvements or features from scratch for OKRs or projects. Delivers reusable, efficient systems, mentors team members, and enhances cross-team collaboration through clear documentation.
- Staff Engineer: Leads the planning and execution of multi-team OKRs or projects. Works with stakeholders to define technical roadmaps and team vision. Serves as a mentor and role model across engineering, ensuring system quality through rigorous testing and monitoring.
Leadership Principles
- Impact-Oriented Trailblazer: Prioritizes customer needs and acts with speed, focus, and adaptability. Treats every initiative as an experiment—test, deploy, measure, and iterate quickly.
- Master Craftsperson: Demonstrates deep commitment to technical excellence. Balances velocity with high standards, owns outcomes fully, and continuously improves through deliberate practice.
- Inspiring Colleague: Adds clarity and energy to interactions. Practices humility, direct communication, and a unified team mindset, strengthening the group through constructive feedback.
Life at TRM
- The organization operates at high speed and expects high ownership, with emphasis on clarity, execution, and measurable impact.
- Individuals who succeed here are driven by complex challenges, experimentation, and ongoing feedback.
- Work that might take months elsewhere is delivered in days due to streamlined processes and autonomy.
- Our mission intersects AI, national security, and crime prevention, creating high-stakes, rapidly evolving problems.
- Challenges are technically demanding, with real-world consequences and fast-changing conditions.
- The pace and intensity reflect the critical nature of the work we do.
- Expect frequent shifts in priorities and objectives as we test, learn, and adapt.
- Work often involves navigating ambiguous situations and making progress without complete information.
- Individuals are expected to take full ownership of their responsibilities and outcomes.
- Close collaboration across teams and functions is a standard practice.
- Communication is frequent, direct, and highly engaged.
- Creative and unconventional thinking is encouraged to solve difficult problems.
- The environment rewards urgency, flexibility, and results-driven behavior.
- This operating model may not suit everyone; those prioritizing predictability or steady workloads should assess fit carefully.
- We seek team members who thrive in this environment, not just endure it.
- Many find the work deeply meaningful and professionally fulfilling.
- If you're inspired by impactful missions, ambitious goals, and collaborative, purpose-driven colleagues, you're likely to excel here.
AI Fluency at TRM
- AI proficiency is a fundamental requirement. We believe AI transforms how top performers work, and expect team members to use AI to transform their work, not just automate tasks.
- AI fluency means ranking in the top 10% of professionals in your field in using AI to accelerate routine workflows.
- AI fluency means ranking in the top 10% of professionals in your field in using AI to structure and solve complex problems.
- AI fluency means ranking in the top 10% of professionals in your field in using AI to enhance the quality of outputs.
- AI fluency means ranking in the top 10% of professionals in your field in using AI to increase speed and leverage.
- Candidates will be assessed on practical AI fluency during the interview process.
Join our Mission
- We value craftsmanship and seek individuals who want their work to have significance, who experiment with speed and rigor, and who take pride in contributing to a safer world for billions.
- If you're aligned with our mission but don't meet every requirement, we still encourage you to apply—we value learning agility, judgment, and drive.
Privacy Notice
- By applying, you consent to the processing of your personal data as described. Information provided (e.g., resume, work history, contact details) is used solely for evaluating your candidacy for current and future roles.
- Due to extended hiring cycles, applicant data may be retained for up to 36 months from the application date.
- After this period, data is deleted unless legal requirements necessitate longer retention.
- Applicants in the EEA, UK, or jurisdictions with data protection laws have the right to access, correct, or request deletion of their personal data before the retention period ends.
- To exercise these rights, contact privacy@trmlabs.com.
Recruitment agencies
- We do not accept unsolicited resumes from recruitment agencies. Please do not send unsolicited candidate submissions.
- We are not liable for fees related to unsolicited resumes and will not pay third parties without a formal agreement in place.
Other
- AI fluency is a baseline expectation. We believe AI fundamentally changes how high performers operate and expect all team members to use AI to accelerate and reimagine their work, not just automate surface tasks.
- Use of AI tools—including notetakers, interview assistants, or real-time coaching platforms like Otter.ai, Fireflies, Fathom, Cluey, or similar—is prohibited during interviews without prior authorization.
- We provide reasonable accommodations for applicants with disabilities; requests can be submitted via a dedicated form.