Senior Systems Engineer - Site Reliability Engineering
CurrentThe responsibilities of this role include maintaining and improving critical infrastructure in all environments including Dev/QA/Perf/Production. We focus on creating baselines and trends based on percentile data or historical data of an application to then make informed decisions on the health of our applications. We support our Command Center and On-Call admins by creating visuals of their application's data to help navigate an issue and perform deep dive analysis using various APM tools like Dynatrace, DataDog, or Splunk. Our mission is to responsibly drive improvements in the reliability and performance of our systems as well as root cause analysis for major incidents.