Sr. Observability Engineer
CurrentAs a Senior Observability Engineer, I am responsible for implementing, and maintaining the observability infrastructure to ensure optimal performance, availability, and reliability of critical systems. My role involves collaborating with development, operations, and security teams to establish best practices for monitoring, logging, and alerting across the organization's infrastructure.Key Responsibilities: Infrastructure Management: Manage and maintain observability platforms such as Cisco Appdynamics, Grafana, or SIEM tools, ensuring they are optimized for performance and scalability. Data Analysis and Visualization: Create and maintain dashboards and reports that provide real-time insights into system performance, resource utilization, and potential bottlenecks. Incident Response: Lead the troubleshooting and resolution of production issues by leveraging observability data to quickly identify and address root causes. Collaboration: Work closely with cross-functional teams to define observability requirements and ensure that applications and services are built with observability in mind. Automation: Automate the collection and analysis of metrics, logs, and traces to enable proactive monitoring and alerting. Continuous Improvement: Stay up to date with the latest observability technologies and practices, continuously improving observability strategies and implementations.