Observability Engineer
Observability Engineer
Irving, TX
Position Overview
We are seeking a hands-on Application Performance Monitoring & Automation Engineer to support application monitoring, observability, automation, reporting, and operational performance initiatives.
This is an individual contributor role requiring strong experience with AppDynamics and Splunk, along with exposure to automation, scripting, Python, Prometheus, and business intelligence tools. The ideal candidate is comfortable troubleshooting application performance issues, analyzing monitoring data, developing automation solutions, and working across technical teams in a large, geographically dispersed environment.
Key Responsibilities
-
Configure, maintain, and utilize AppDynamics for application performance monitoring, troubleshooting, and performance analysis.
-
Use Splunk to analyze logs, identify trends, investigate issues, and support operational monitoring.
-
Develop and support automation solutions that improve monitoring, reporting, and operational processes.
-
Create scripts using Python or other scripting languages to automate repetitive technical tasks.
-
Work with automation tools and frameworks such as Ansible, Selenium, or similar technologies.
-
Utilize Prometheus and related monitoring technologies to track system and application performance.
-
Develop dashboards, reports, and data visualizations using Power BI and/or Tableau.
-
Support ServiceNow (SNOW) processes, including incident, request, or operational workflow activities.
-
Assist with monitoring and analysis involving mainframe technologies and environments.
-
Identify opportunities to improve existing processes through scripting and automation.
-
Work with technical teams to troubleshoot application performance and operational issues.
-
Apply automation principles to increase efficiency and reduce manual processes.
-
Leverage or evaluate AI-enabled technologies where applicable to monitoring, automation, analytics, or operational processes.
-
Organize and manage multiple priorities while meeting deadlines in a fast-paced environment.
-
Collaborate effectively with teams and stakeholders across geographically dispersed locations.
Required Qualifications
-
4–6 years of AppDynamics Application Performance Monitoring experience.
-
4–6 years of Splunk experience.
-
1–2+ years of automation experience.
-
1–2+ years of Python or comparable scripting experience.
-
1–2+ years of Prometheus experience.
-
1–2+ years of Power BI and/or Tableau experience.
-
1–2+ years of ServiceNow (SNOW) experience.
-
Understanding of automation principles and tools such as Ansible, Selenium, or similar platforms.
-
Experience with scripting and developing technical automation solutions.
-
Experience with Microsoft Office products, including Excel, PowerPoint, SharePoint, and Word.
-
Experience or familiarity with mainframe technology.
-
Exposure to Artificial Intelligence (AI) technologies or AI-enabled solutions.
-
Strong troubleshooting, analytical, and problem-solving skills.
-
Ability to prioritize work, meet deadlines, and perform effectively under pressure.
-
Ability to organize and manage multiple priorities in a dynamic and complex environment.
-
Strong communication and collaboration skills with the ability to work effectively across geographically dispersed teams.
Core Technical Skills
Primary: AppDynamics, Splunk
Monitoring/Observability: Prometheus, Application Performance Monitoring
Automation/Scripting: Python, Ansible, Selenium, scripting/automation tools
Reporting/Analytics: Power BI, Tableau
IT Operations: ServiceNow (SNOW), Mainframe Technology
Additional: AI exposure, Microsoft Office, SharePoint
Required Skills
Required Languages
🇬🇧 English