NVIDIA is looking for a skilled engineer to design and scale observability platforms in Santa Clara, California. This role entails building backend services for telemetry ingestion and collaborating with various teams to ensure platform health. Candidates should have a relevant Bachelor’s degree and over 5 years of experience in distributed systems, ideally with programming skills in Python, Go, or Java. The salary range for this position is competitive and includes equity and benefits.
Qualifications
5+ years of experience building backend or distributed systems in production environments.
Hands-on experience with modern observability architectures.
Strong understanding of distributed systems and fault-tolerant design.
Responsibilities
Design and scale observability platforms for metrics, logs, and traces.
Build high-performance backend services for telemetry ingestion.
Collaborate with platform engineering and infrastructure teams.
Skills
Python
Go
Java
Kubernetes
Kafka
Spark
Flink
OpenTelemetry
Prometheus
Observability
Education
Bachelor’s degree in Computer Science, Computer Engineering, or related field
Job description
NVIDIA is looking for a skilled engineer to design and scale observability platforms in Santa Clara, California. This role entails building backend services for telemetry ingestion and collaborating with various teams to ensure platform health. Candidates should have a relevant Bachelor’s degree and over 5 years of experience in distributed systems, ideally with programming skills in Python, Go, or Java. The salary range for this position is competitive and includes equity and benefits.