Senior Engineer

Roku · Bengaluru

  • Experience15–19 yrs
  • SalaryNot disclosed
  • Work modeonsite
  • Levelexecutive
  • Posted2 Sept 2026

About Roku

Roku is hiring in Bengaluru in media advertising. This role looks for around 15+ years of experience.

Skills

  • Golang
  • Python
  • Shell
  • Prometheus
  • Grafana
  • Loki
  • Tempo
  • ELK
  • OpenSearch
  • ClickHouse
  • OpenTelemetry
  • OpenMetrics
  • Kubernetes
  • Istio
  • Envoy
  • Terraform
  • AWS
  • GCP
  • distributed systems
  • cloud infrastructure
  • service mesh
  • multicloud
  • TSDB
  • Parquet
  • distributed processing
  • infrastructure security

The role

A platform engineer at a media and advertising product company architects observability platforms and cloud infrastructure with distributed systems, Kubernetes, and service mesh technologies. The role evolves telemetry pipelines, storage architectures, and multicloud reliability while improving developer experience and platform security.

Full job description

We are building a next-generation observability and cloud platform that is high-performance, cost-efficient, secure, and scalable across multi-region, multi-cloud clusters. You will lead the architecture and evolution of Roku's observability and cloud infrastructure stack. This includes metrics, logs, traces, telemetry pipelines, service mesh, developer experience, and reliability of systems that power thousands of services and millions of devices. You will drive a vision where developers gain deep visibility with minimal overhead, onboarding is seamless, and insights are available in real time. Your work will directly help Roku scale efficiently while maintaining reliability, cost control, and performance.

Responsibilities:

Architect and lead Roku's observability platform across metrics, logs, and traces; evolve data pipelines and storage layers optimized for high throughput, performance, and cost at Roku scale (TSDBs, Parquet, and distributed processing).

Extend and harden open-source observability systems; overhaul core components (e. g., storage layers, query paths) to improve performance, reliability, and usability at scale.

Implement features such as preaggregation, down-sampling, and sampling to reduce load and accelerate queries across the platform.

Collaborate across platform, SRE, and product teams to migrate hundreds of workloads to our common platform and augment and automate CI/CD flows and onboarding.

Integrate security into infrastructure and platform services; ensure robust multitenant, multicluster, and multicloud designs.

Contribute improvements back to open source and CNCF-aligned projects; shape standards adoption (OpenTelemetry, OpenMetrics) across the company.

Mentor engineers; establish best practices for reliability, efficiency, and cost management across service mesh and observability domains.

Requirements:

15+ years in software engineering with a track record of architecting distributed systems or platforms at scale.

Strong hands-on experience in Golang and one scripting language (e. g., Python or Shell).

Experience operating observability at pb-scale ingestion and hundreds of millions of series.

Expertise in observability platforms and tooling (Prometheus, Grafana, Loki, Tempo, ELK/OpenSearch, ClickHouse) and standards (OpenTelemetry, OpenMetrics).

Deep experience building systems of scale and operating cloud infrastructure with Kubernetes; strong proficiency with service mesh technologies (Istio/Envoy), infrastructure as code (Terraform) and experience in multicloud (AWS, GCP).

Demonstrated ability to evolve storage and query architectures for cost, scale, and latency (e. g., TSDB, Parquet, distributed processing).

Proven experience integrating security as part of infrastructure and platform development.

Exceptional crossfunctional communication; effective collaboration with both technical and nontechnical stakeholders.

Culture fit: independent thinker, pragmatic problem solver, low-ego collaborator who moves fast and focuses on company success.

Experience integrating AI tools to improve processes and reduce toil.

Open-source contributions to CNCF projects are preferred but not mandatory.