In this role, you will be a vital instrument in collecting, analyzing, and interpreting data to provide valuable insights into system behavior, performance trends, and user experiences. You will design and implement solutions to monitor, trace, and provide insights into complex, multi-technology environments, with a focus on system reliability, security, and performance.
Responsibilities
Developing and maintain observability solutions using tools like Datadog, Splunk, New Relic, AWS CloudWatch, and Azure Application Insights.
Ensuring that observability tools and configurations are managed through Infrastructure as Code for consistency and reliability.
Working closely with development, DevOps, and operations teams to ensure seamless integration of observability solutions.
Documenting observability configurations, procedures, and best practices.
Providing guidance and mentorship to junior engineers.
Requirements
Have experience in application observability and performance management.
Show experience managing managers and technical teams with either a data, infrastructure or platform focus
Be proficient with observability tools such as Datadog, Splunk, New Relic, AWS CloudWatch, and Azure Application Insights.
Have good expertise in AWS and Azure cloud environments.
Have hands-on experience with infrastructure as code (IaC) using CloudFormation and Terraform.
Have an engineering, Computer Science, or equivalent experience.
20 Initiatives to Boost Employee EngagementAre you struggling with improving employee engagement at work? This article covers everything from better communication to building a strong workplace culture.
30 Common Interview Mistakes to AvoidThis piece examines 30 of the most common mistakes applicants make at interviews, so you know how to better avoid them.