Remote (Country)

Braze is hiring a Senior Site Reliability Engineer (SRE)

About the Role

Braze is hiring a Senior Site Reliability Engineer (SRE) to help change how customers control their observability data. In this role, you will unlock the value of all observability data, contribute to envisioning, creating, deploying, testing, and shipping products, and be involved from initial conception through production.

What You'll Do

  • Engage with teams and improve service delivery and reliability across their entire lifecycle.
  • Measure and monitor all production systems with an eye towards availability, latency and overall system health.
  • Seek out the cause of errors and instability in our production cloud services and drive teams towards better operational excellence.
  • Engage with product and platform teams to improve and evolve systems by lobbying for changes that improve reliability, resilience, and observability.
  • Help identify and drive down toil with creative innovation and automation.
  • Participate in stand-by, on-call, or off-hours duties.

What We're Looking For

  • Extensive experience with enterprise scale continuous delivery environments.
  • Development with JavaScript/Node.js/TypeScript in a Linux/Mac environment.
  • Experience with sustainable incident response in a blameless environment.
  • Experience with Configuration Management Tools like Terraform (preferred) or Puppet, Chef, Ansible.
  • Knowledge of cloud platforms (prefer AWS) and container + orchestration technologies.
  • Experience with APM and Observability and related tools such as, New Relic, Splunk, CloudWatch, Prometheus, Grafana/Kibana, Sentry etc.
  • Background in Linux Systems Engineering.
  • Experience with Incident response related tools for instance, PagerDuty, FireHydrant, Blameless etc.
  • Comfortable with a high level of autonomy and working with a distributed team.
  • Knowledge of Cloud and application security.
  • Strong knowledge of cloud design patterns for scale, data management, resiliency, etc.
  • A love for high quality and a knack for testing.
  • Opinions about dashboards, metrics, and SLO’s.

Nice to Have

  • Terraform experience.
  • AWS knowledge.

Technical Stack

  • Languages: JavaScript, Node.js, TypeScript
  • OS: Linux, Mac
  • Infrastructure as Code: Terraform, Puppet, Chef, Ansible
  • Cloud & Containers: AWS, Container technologies, Orchestration technologies
  • Observability: New Relic, Splunk, CloudWatch, Prometheus, Grafana, Kibana, Sentry
  • Incident Response: PagerDuty, FireHydrant, Blameless

Team & Environment

You will be part of the engineering organization, joining a team of technical engineers.

Benefits & Compensation

  • Compensation range: $180,000 - $240,000
  • Health, dental, vision, short-term disability, and life insurance.
  • Paid holidays and paid time off.
  • A fertility treatment benefit.
  • 401(k).
  • Equity.
  • Eligibility for a discretionary company-wide bonus.

Work Mode

This is a remote position. Candidates must be located within the United States.

All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, or any other applicable legally protected characteristics in the location in which the candidate is applying.

Required Skills
JavaScriptNode.jsTypeScriptLinuxTerraformAWSPuppetChefAnsibleSite Reliability EngineeringInfrastructure as CodeCloud ArchitectureAutomationMonitoringIncident Management
Got hired remotely?

Get paid like a professional

Remote clients expect company invoices, not personal PayPal requests. Glopay forms an EU partnership that makes you look legitimate while you stay independent.

Professional invoices with EU company details
Compliance handled automatically
Withdraw to any bank account
Income reports for easy tax filing
Create free account
Free signup • 5 min setup
About company
B

Braze is the leading customer engagement platform that empowers brands to Be Absolutely Engaging.™ Braze allows any marketer to collect and take action on any amount of data from any source, so they can creatively engage with customers in real time, across channels from one platform.

Visit website
Job Details
Category infrastructure
Posted 5 months ago