Responsibilities
- Identify and fix cross-cutting technical issues that impact system performance and reliability
- Enhance platform-wide observability, monitoring, and debugging capabilities
- Examine AI agent behavior and boost output quality via structured evaluation and iterative improvements
- Refactor and improve legacy or complex code to increase scalability and maintainability
- Spot architectural bottlenecks and implement system-wide enhancements
- Work with product and customer teams to grasp real-world failures and convert them into technical solutions
- Boost performance and reliability across infrastructure and services
- Raise technical standards and promote cleaner code practices throughout the system
- Act proactively by identifying and solving problems before they are formally defined
Requirements
- At least 8 years of professional software engineering experience
- Strong experience with Python (preferred), especially in backend or systems environments
- Proven experience working with complex or distributed systems
- Demonstrated ability to diagnose and resolve system-wide performance issues
- Experience working with LLM-based systems or AI agents
- Experience evaluating, tuning, or improving AI model outputs
- Strong understanding of software architecture, scalability, and maintainability principles
- Experience improving code quality across large or evolving codebases
- Solid computer science fundamentals (algorithms, data structures, system design)
- Experience working independently with high ownership and minimal oversight
- Strong problem-solving skills and ability to context-switch across infrastructure, backend, and AI-related challenges
- Excellent communication skills and ability to translate ambiguous problems into structured solutions
Nice to Have
- Strong observability background (logging, tracing, monitoring, performance metrics)
- Experience in platform engineering or infrastructure-focused roles
- Familiarity with cloud environments (GCP preferred, AWS/Azure acceptable)
- Experience with containerization (Docker) and CI/CD pipelines
- Experience working in early-stage startup environments
- Experience improving system performance across services
Benefits
- Educational resources
- Flexible schedule and Work From Anywhere
- Referral Program
- Supportive and chill atmosphere
Work Arrangement
Remote (Worldwide)
Other
We are accepting applications from LATAM countries