DeveloperJobs.io
← Back to all jobs

Job Description

Oracle invites you to join the Seattle on-site team shaping Oracle Cloud Infrastructure's observability platform. This role centers on metrics, logs, traces, and telemetry data, with ownership spanning the design, development, operation, and optimization of scalable, low-latency pipelines and services that deliver timely insights for cloud services.

As a Software Developer, you will lead the design, development, and operation of cloud-scale observability platforms that handle telemetry data across metrics, logs, and traces. The position offers opportunities to mentor engineers, influence technical strategy, and collaborate across a hyperscale cloud environment to advance next generation capabilities.

You will work closely with product management, architects, SREs, and engineering teams to improve instrumentation, telemetry quality, and operational visibility across cloud services, while establishing and monitoring health, scalability, performance, and cost-efficiency metrics for observability platforms.

Responsibilities

  • Lead the design, development, and operation of cloud-scale observability platforms supporting metrics, logs, traces, and related telemetry data.
  • Architect and implement scalable, resilient, and cost-efficient telemetry collection, ingestion, processing, storage, and query systems.
  • Drive the evolution of end-to-end observability pipelines, from instrumentation and data collection through real-time analytics and long-term retention.
  • Design and optimize distributed systems capable of ingesting and processing massive telemetry data with stringent latency and availability requirements.
  • Develop scalable storage and indexing solutions for high-cardinality metrics, large-scale log analytics, and distributed tracing workloads.
  • Build and enhance query, search, and retrieval services that deliver fast, reliable, and intuitive access to observability data.
  • Collaborate with product management, architects, SREs, and engineering teams to define and deliver next-generation observability capabilities.
  • Identify and resolve performance bottlenecks across the observability stack, including ingestion, storage, indexing, aggregation, and query execution.
  • Design systems with a strong focus on reliability, fault tolerance, scalability, security, and operational excellence.
  • Drive technical strategy and architectural decisions for observability services operating at hyperscale cloud environments.
  • Mentor senior and junior engineers, provide technical leadership, and foster engineering best practices across the organization.
  • Partner with service teams to improve instrumentation, telemetry quality, and operational visibility across cloud services.
  • Establish and monitor key service health, scalability, performance, and cost-efficiency metrics for observability platforms.
  • Lead troubleshooting and root-cause analysis efforts for complex distributed systems and large-scale production environments.
  • Stay current with emerging trends, technologies, and best practices in observability, distributed systems, data processing, and cloud-native architectures.

Similar Jobs