Cloudera Logo

Cloudera

Principal Engineer - Observability Telemetry Client Infrastructure

Reposted Yesterday
Be an Early Applicant
In-Office
Austin, TX
Expert/Leader
In-Office
Austin, TX
Expert/Leader
The Principal Engineer will architect and implement an observability telemetry framework, enabling seamless telemetry data integration for multi-tenant environments, focusing on high performance and scalability.
The summary above was generated by AI

Business Area:

Engineering

Seniority Level:

Director

Job Description: 

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we're the preferred data partner for the top companies in almost every industry.  Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world’s largest enterprises.

Cloudera is seeking a Principal Engineer to serve as the primary architect and visionary for our Observability Telemetry client interactions framework as we build a multi-tenant, high-throughput telemetry fabric to support the world’s largest data estates. In this role, you will lead the evolution of Cloudera’s Observability product by designing and building a "self-service" Open Telemetry-based ecosystem that allows internal clients to integrate telemetry data seamlessly for multiple downstream consumers. 

You will architect the self-service interfaces that allow thousands of distributed components to emit high-cardinality logs, metrics, and traces that are automatically correlated, context-aware, and ready for downstream analysis in ClickHouse and other massive-scale engines. You will be Cloudera’s voice in the CNCF/OTel community, influencing the direction of open-source observability to meet the needs of hybrid-cloud data platforms.

As a technical leader, you will be responsible for defining the semantic conventions to enable log events, metrics, spans, and traces from diverse, multi-language clients to be correlated into a unified, actionable view for customers, Support and Cloudera product engineering. This data is the foundation for troubleshooting, forecasting, workload analysis, financial governance, and other administrative functions. You will also work closely with open source products for integration and can influence OTel integration directions in the open source ecosystem for these components

This position is a high-visibility role requiring a blend of deep systems architecture, hands-on implementation, and cross-organizational influence to ensure our telemetry infrastructure scales with the world’s most complex data workloads.

This role is not eligible for immigration sponsorship.

As an Observability Telemetry Principal Engineer you will:

  • Architect and drive the implementation of automated "on-ramps" for observability clients that handles the complexity of multi-cloud, hybrid environments without sacrificing performance, ensuring teams can integrate their services with minimal friction.

  • Establish and enforce the semantic conventions needed to ensure telemetry data carries the appropriate context for easy correlation across the entire Cloudera stack.

  • Develop and support high-performance interfaces and SDKs for clients across various languages (Java, Go, Python, etc.) to contribute high-fidelity signals.

  • Build the logic to stitch together disparate signals into a unified trace, enabling deep-dive workload analysis and financial governance across massive distributed systems.

  • Work alongside engineering teams to turn architectural blueprints into production reality, conducting deep-dive code reviews and resolving complex systemic bottlenecks.

  • Serve as the "go-to" expert for observability, resolving technical disagreements and making high-stakes decisions on the future of our telemetry platform.

We are excited if you have: (Required Experience)

  • 10+ years of experience (or equivalent advanced degree + experience) designing and maintaining large-scale distributed systems and observability platforms.

  • A proven track record of designing and shipping complex, critical features that serve as foundational infrastructure for other engineering teams.

  • Deep, hands-on experience with the OpenTelemetry Collector architecture, custom processors, and the challenges of high-cardinality data.

  • Experience with high-volume OLAP engines (e.g., ClickHouse, StarRocks) and an understanding of how to structure telemetry data for sub-second queries at large scale.

  • Excellent communication and collaboration skills and the ability to build relationships across the company to drive adoption of new standards and remove technical roadblocks.

  • The ability to map business requirements to technical roadmaps, ensuring our observability tools support Cloudera’s long-term strategic goals.

  • Experience coaching senior and staff-level engineers, acting as a "force multiplier" for a technical organization.

  • Bsc/Msc in related field or equivalent experience

You might have:

  • Significant contributions to major observability or data projects (e.g., CNCF or Apache projects).  Bonus points if you’re already a CNCF OTel maintainer.

  • Deep experience with Kubernetes-native observability and managing telemetry at scale in hybrid-cloud environments.

  • Experience representing technical initiatives at industry conferences or internal company-wide summits.

  • Experience using machine learning or advanced analytics to derive "AIOps" insights from raw telemetry data.

Why this role matters:

At Cloudera, our customers manage some of the largest and most complex data estates in the world. Without robust, correlated observability, managing these environments is an impossible task. This role is the linchpin of our visibility strategy.

By building a self-service, standardized telemetry framework, you are not just helping one team; you are empowering every developer at Cloudera and every one of our customers to understand their data life cycle. Your work ensures that when a performance bottleneck occurs or a system fails, the path to resolution is visible, traceable, and immediate. You are building the "nervous system" of the Cloudera Data Platform.

What you can expect from us:

  • Generous PTO Policy 

  • Support work life balance with Unplugged Days

  • Flexible WFH Policy 

  • Mental & Physical Wellness programs 

  • Phone and Internet Reimbursement program 

  • Access to Continued Career Development 

  • Comprehensive Benefits and Competitive Packages 

  • Paid Volunteer Time

  • Employee Resource Groups

EEO/VEVRAA

#LI-CP1

#LI-HYBRID

HQ

Cloudera Santa Clara, California, USA Office

Santa Clara, CA, United States

Cloudera San Jose, California, USA Office

6220 America Center Dr, 5th Floor, San Jose, California, United States, 95002

Similar Jobs at Cloudera

An Hour Ago
In-Office
Senior level
Senior level
Artificial Intelligence • Cloud • Software • Big Data Analytics
Embed with strategic enterprise customers to prototype, build, and productionize full-stack generative AI and agentic applications; advise on AI strategy, codify repeatable solution patterns, mentor teams, and feed product feedback to improve the AI platform.
Top Skills: Agentic WorkflowsApache IcebergApache NifiSparkAWSAzureClouderaFine-TuningFoundation ModelsGCPInference OptimizationLlmopsMlopsModel ServingNvidia GpusRetrieval-Augmented Generation (Rag)Semantic Search
Yesterday
In-Office or Remote
3 Locations
165K-230K Annually
Senior level
165K-230K Annually
Senior level
Artificial Intelligence • Cloud • Software • Big Data Analytics
The Staff Software Engineer will architect and build scalable solutions for Cloudera's Data Platform, contribute to Apache Spark, enhance engineering processes, and work with large-scale distributed systems.
Top Skills: Apache IcebergApache ParquetSparkJavaPythonScalaSQL
13 Days Ago
In-Office
Senior level
Senior level
Artificial Intelligence • Cloud • Software • Big Data Analytics
Design, build, and deliver scalable enterprise AI inference services and model registry capabilities. Enable generative AI applications using foundation models, RAG, and vector databases; collaborate with frontend, data scientists, product, and UX teams to drive production-ready AI platform features on Kubernetes-based microservices.
Top Skills: AWSAzureC#CSSFoundation ModelsGCPGoGrpcHiveHTMLHuggingfaceJavaKnativeKserveKubeflowKubernetesMilvusMlflowNimNode.jsNvidia Ai FrameworksPineconePrompt EngineeringPythonReactRetrieval-Augmented Generation (Rag)SparkSQLTensorFlow

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account