Principal Observability Engineer
FreedomPay- Location
- US
- Workplace
- Hybrid
- Employment
- Full Time
- Salary
- —
Posted 1mo ago
The Senior Observability Engineer is responsible for defining, leading, and advancing enterprise observability strategy, architecture, and implementation across applications, platforms, infrastructure, and operational services. This role requires deep hands-on expertise with Dynatrace SaaS and recent Dynatrace platform capabilities and innovations, as well as strong experience with OpenTelemetry, cloud-native observability, AIOps, and agentic operations.
This individual will serve as a principal-level technical leader, applying strategic thinking, systems architecture expertise, and sound decision-making to establish modern observability standards and scalable telemetry practices. Working in a fast-paced, highly collaborative technology environment, this role partners closely with engineering, site reliability, platform, security, infrastructure, operations, and architecture teams to guide the organization beyond traditional monitoring toward more intelligent and adaptive operational models.
Essential Duties & Responsibilities
- Define and evolve our enterprise observability vision, standards, principles, and roadmap, using strategic thinking and sound judgment to align technical direction with business needs.
- Lead the implementation, optimization, and adoption of Dynatrace SaaS and its latest platform capabilities across the enterprise.
- Establish scalable telemetry architecture and OpenTelemetry standards that improve consistency, interoperability, and long-term flexibility.
- Design observability solutions for Kubernetes, containers, microservices, distributed applications, and public cloud environments.
- Apply AIOps and agentic operations capabilities to strengthen detection, event correlation, diagnosis, automation, operational response, and continuous improvement.
- Develop and maintain service health models, SLOs, dashboards, alerting strategies, telemetry governance, and observability best practices that support operational excellence and service reliability.
- Collaborate across engineering, site reliability, platform, security, infrastructure, operations, and architecture functions to expand observability adoption and maturity.
- Guide the transition from legacy monitoring practices to modern, adaptive, and outcome-focused observability models, including initiatives involving production systems and enterprise platform transformation.
- Mentor engineers, influence technical direction, solve complex problems, and help shape enterprise architecture and engineering standards.
Required Qualifications
- Master’s degree from an accredited college or university in Computer Science, Information Systems, Engineering, or a related technical field.
- 8+ years of experience in observability, monitoring, site reliability engineering, platform engineering, infrastructure engineering, or related technical disciplines.
- Deep hands-on experience with Dynatrace SaaS and recent Dynatrace platform capabilities and innovations.
- Strong experience with OpenTelemetry and telemetry instrumentation, collection, and architecture practices.
- Experience designing and implementing observability solutions for cloud-native infrastructure, distributed systems, and modern application environments.
- Hands-on experience implementing AIOps and/or agentic operations capabilities.
- Strong understanding of metrics, logs, traces, event correlation, service health, alerting models, and operational intelligence.
- Demonstrated ability to apply lessons from traditional monitoring approaches to modern observability strategy and technical design.
- Proven ability to lead technical initiatives, influence architectural decisions, and drive adoption across multiple teams and stakeholders.
- Strong verbal and written communication skills, with the ability to collaborate across functions and influence technical and non-technical stakeholders.
Preferred Qualifications
- Experience leading enterprise-scale observability transformation initiatives.
- Experience with AWS, Azure, and/or Google Cloud Platform.
- Strong knowledge of site reliability engineering principles, including SLIs, SLOs, incident response, and operational resilience.
- Experience with automation, orchestration, and remediation workflows.
- Familiarity with additional observability platforms, frameworks, or ecosystems beyond Dynatrace.
- Experience with scripting or programming languages such as Python, Go, Bash, or JavaScript.
Skills
- Dynatrace
- OpenTelemetry
- Kubernetes
- AIOps
- AWS
- Azure
- Google Cloud Platform
- Python
- Go
- Bash
- JavaScript
More jobs at FreedomPay
All 11Sr. Android Software Engineer
FreedomPay · Philadelphia, PA/ Las Vegas, NV · 7d ago
Platform Operations Analyst
FreedomPay · Philadelphia, Pennsylvania · 9d ago
Associate Site Reliability Engineer
FreedomPay · US · 17d ago
Director, Global Brand Design
FreedomPay · Philadelphia, Pennsylvania · 1mo ago
Platform Operations Engineer
FreedomPay · Philadelphia, Pennsylvania · 1mo ago
Similar roles
Site Reliability Engineer (FedRAMP / Security)
Coralogix · New York, NY, United States · USD 170,000–350,000/yr · today
Senior Staff Product Manager, Conversational AI
ServiceNow · Santa Clara, California, United States · USD 190,900–334,100/yr · today
Senior Business Engineer - Ads
Reddit · Remote · USD 180,200–252,300/yr · today
Machine Learning Engineer
Reddit · US · USD 185,800–303,400/yr · today
Backend Software Engineer
ClearlyRated · Portland, OR, United States · USD 90,000–120,000/yr · today
Firmware Engineer
Anduril Industries · Costa Mesa, California, United States · USD 166,000–220,000/yr · today