J
Principal Software Engineer- Telemetry/Observability
Analytics
Business Intelligence
Cloud
Cloud Infrastructure
Cloud Operations
Cloud Platform
Cloud Platforms
Cloud Technology
Cybersecurity Tools
Dashboards
Data Analysis
Data Analytics
Data Governance
Data Observability
Data Platform
Data Processing
Data Security
Data Visualization
DevOps
DevSecOps
Engineering
Fleet Management
Grafana
Information Technology (IT)
Infrastructure As Code
Log Management
Monitoring
Observability
Opentelemetry
Operations Management
Platform Engineering
Programming Languages
Prometheus
Reporting and Analytics
Splunk
Telematics
Telemetry
Visual Design
Job Description
JPMorganChase seeks a Principal Software Engineer focused on observability and telemetry platforms within the Core Foundational Platforms group. Based in Seattle, WA onsite, this role designs end-to-end telemetry pipelines, builds platform services, and guides observability strategy across VPC and cloud networking environments, with a salary range of USD 204,250 to 285,000 per year and a requirement of 10+ years of software engineering experience.
Responsibilities
- Designs scalable software frameworks and platform services for observability within VPC environments, applying robust software design patterns.
- Delivers secure, production-grade code and conducts reviews and debugging across telemetry collection, enrichment, storage, and visualization components.
- Leads end-to-end telemetry pipelines covering metrics, logs, traces, and events, establishing standards, schemas, and instrumentation patterns for platform and application teams.
- Develops and evolves CI/CD pipelines to enable reliable delivery, automated testing, policy as code, and safe progressive rollouts for observability services.
- Architects solutions for software-defined networking and cloud networking environments, including troubleshooting across layered network paths and dependencies.
- Builds and maintains monitoring, alerting, and SLO/SLI practices using tools like Splunk, Grafana, and Prometheus, including dashboards, alerts, and runbooks.
- Advises cross-functional teams on telemetry strategy, platform integrations, and readiness for operations.
- Acts as the function's subject matter expert for observability and telemetry across the VPC platform, guiding standards and best practices.
- Creates durable, reusable software frameworks used across teams, such as SDKs, libraries, templates, reference architectures, and golden paths.
- Informs senior stakeholders by communicating tradeoffs, risk, and roadmaps for platform reliability and resiliency.
- Champions diversity, opportunity, inclusion, and respect within the firm.
Requirements
- 10+ years of software engineering experience, including building and operating production platform services.
- Hands-on involvement in system design, application development, testing, and operational stability, ideally for platform, infrastructure, or developer productivity products.
- Expert Python development experience, building services, automation, tooling, and integrations for telemetry and operations workflows.
- Strong Linux systems and networking expertise (TCP/IP, DNS, TLS, routing, NAT, packet capture), with the ability to diagnose cross-system issues in distributed environments.
- Practical experience with CI/CD and modern delivery practices (progressive delivery, automated quality gates, artifact/version management).
- Advanced knowledge of software development processes with in-depth expertise in cloud platforms, platform engineering, reliability engineering, networking, or observability.
- Experience with observability tooling such as Splunk, Grafana, or Prometheus, including instrumentation, queries, dashboards, alerts, and operationalization.
- Experience designing or operating telemetry data pipelines (collection agents, exporters, scraping, indexing, retention, cardinality management, performance and cost controls).
- Proven ability to apply expertise and new methods to solve complex technology problems across multiple disciplines.
- Working knowledge of software-defined networking concepts and environments, including integrations and operational considerations for VPC-like network platforms.
Technologies
- Python
- Splunk
- Grafana
- Prometheus
- OpenTelemetry
- Terraform
Benefits
- Salary: USD 204,250 - 285,000 per year, base determined by role, experience, skill set and location.
- Commission-based pay
- Discretionary incentive compensation in cash and/or forfeitable equity
- Comprehensive health care coverage
- On-site health and wellness centers
- Retirement savings plan
- Backup childcare support
- Tuition reimbursement
- Mental health support
- Financial coaching
Similar Jobs
J
J
J
J
J
J