Cloudera Logo

Cloudera

Principal Engineer - Observability Telemetry Client Infrastructure

Reposted Yesterday
Be an Early Applicant
In-Office
Austin, TX
Expert/Leader
In-Office
Austin, TX
Expert/Leader
The Principal Engineer will architect and implement an observability telemetry framework, enabling seamless telemetry data integration for multi-tenant environments, focusing on high performance and scalability.
The summary above was generated by AI

Business Area:

Engineering

Seniority Level:

Director

Job Description: 

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we're the preferred data partner for the top companies in almost every industry.  Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world’s largest enterprises.

Cloudera is seeking a Principal Engineer to serve as the primary architect and visionary for our Observability Telemetry client interactions framework as we build a multi-tenant, high-throughput telemetry fabric to support the world’s largest data estates. In this role, you will lead the evolution of Cloudera’s Observability product by designing and building a "self-service" Open Telemetry-based ecosystem that allows internal clients to integrate telemetry data seamlessly for multiple downstream consumers. 

You will architect the self-service interfaces that allow thousands of distributed components to emit high-cardinality logs, metrics, and traces that are automatically correlated, context-aware, and ready for downstream analysis in ClickHouse and other massive-scale engines. You will be Cloudera’s voice in the CNCF/OTel community, influencing the direction of open-source observability to meet the needs of hybrid-cloud data platforms.

As a technical leader, you will be responsible for defining the semantic conventions to enable log events, metrics, spans, and traces from diverse, multi-language clients to be correlated into a unified, actionable view for customers, Support and Cloudera product engineering. This data is the foundation for troubleshooting, forecasting, workload analysis, financial governance, and other administrative functions. You will also work closely with open source products for integration and can influence OTel integration directions in the open source ecosystem for these components

This position is a high-visibility role requiring a blend of deep systems architecture, hands-on implementation, and cross-organizational influence to ensure our telemetry infrastructure scales with the world’s most complex data workloads.

This role is not eligible for immigration sponsorship.

As an Observability Telemetry Principal Engineer you will:

  • Architect and drive the implementation of automated "on-ramps" for observability clients that handles the complexity of multi-cloud, hybrid environments without sacrificing performance, ensuring teams can integrate their services with minimal friction.

  • Establish and enforce the semantic conventions needed to ensure telemetry data carries the appropriate context for easy correlation across the entire Cloudera stack.

  • Develop and support high-performance interfaces and SDKs for clients across various languages (Java, Go, Python, etc.) to contribute high-fidelity signals.

  • Build the logic to stitch together disparate signals into a unified trace, enabling deep-dive workload analysis and financial governance across massive distributed systems.

  • Work alongside engineering teams to turn architectural blueprints into production reality, conducting deep-dive code reviews and resolving complex systemic bottlenecks.

  • Serve as the "go-to" expert for observability, resolving technical disagreements and making high-stakes decisions on the future of our telemetry platform.

We are excited if you have: (Required Experience)

  • 10+ years of experience (or equivalent advanced degree + experience) designing and maintaining large-scale distributed systems and observability platforms.

  • A proven track record of designing and shipping complex, critical features that serve as foundational infrastructure for other engineering teams.

  • Deep, hands-on experience with the OpenTelemetry Collector architecture, custom processors, and the challenges of high-cardinality data.

  • Experience with high-volume OLAP engines (e.g., ClickHouse, StarRocks) and an understanding of how to structure telemetry data for sub-second queries at large scale.

  • Excellent communication and collaboration skills and the ability to build relationships across the company to drive adoption of new standards and remove technical roadblocks.

  • The ability to map business requirements to technical roadmaps, ensuring our observability tools support Cloudera’s long-term strategic goals.

  • Experience coaching senior and staff-level engineers, acting as a "force multiplier" for a technical organization.

  • Bsc/Msc in related field or equivalent experience

You might have:

  • Significant contributions to major observability or data projects (e.g., CNCF or Apache projects).  Bonus points if you’re already a CNCF OTel maintainer.

  • Deep experience with Kubernetes-native observability and managing telemetry at scale in hybrid-cloud environments.

  • Experience representing technical initiatives at industry conferences or internal company-wide summits.

  • Experience using machine learning or advanced analytics to derive "AIOps" insights from raw telemetry data.

Why this role matters:

At Cloudera, our customers manage some of the largest and most complex data estates in the world. Without robust, correlated observability, managing these environments is an impossible task. This role is the linchpin of our visibility strategy.

By building a self-service, standardized telemetry framework, you are not just helping one team; you are empowering every developer at Cloudera and every one of our customers to understand their data life cycle. Your work ensures that when a performance bottleneck occurs or a system fails, the path to resolution is visible, traceable, and immediate. You are building the "nervous system" of the Cloudera Data Platform.

What you can expect from us:

  • Generous PTO Policy 

  • Support work life balance with Unplugged Days

  • Flexible WFH Policy 

  • Mental & Physical Wellness programs 

  • Phone and Internet Reimbursement program 

  • Access to Continued Career Development 

  • Comprehensive Benefits and Competitive Packages 

  • Paid Volunteer Time

  • Employee Resource Groups

EEO/VEVRAA

#LI-CP1

#LI-HYBRID

Cloudera New York, New York, USA Office

151 W 26th St, 10th Floor, Suite 1002 , New York, United States, 10001

Similar Jobs at Cloudera

2 Hours Ago
In-Office
Senior level
Senior level
Artificial Intelligence • Cloud • Software • Big Data Analytics
Embed with strategic enterprise customers to prototype, build, and productionize full-stack generative AI and agentic applications; advise on AI strategy, codify repeatable solution patterns, mentor teams, and feed product feedback to improve the AI platform.
Top Skills: Agentic WorkflowsApache IcebergApache NifiSparkAWSAzureClouderaFine-TuningFoundation ModelsGCPInference OptimizationLlmopsMlopsModel ServingNvidia GpusRetrieval-Augmented Generation (Rag)Semantic Search
Yesterday
In-Office or Remote
3 Locations
165K-230K Annually
Senior level
165K-230K Annually
Senior level
Artificial Intelligence • Cloud • Software • Big Data Analytics
The Staff Software Engineer will architect and build scalable solutions for Cloudera's Data Platform, contribute to Apache Spark, enhance engineering processes, and work with large-scale distributed systems.
Top Skills: Apache IcebergApache ParquetSparkJavaPythonScalaSQL
13 Days Ago
In-Office
Senior level
Senior level
Artificial Intelligence • Cloud • Software • Big Data Analytics
Design, build, and deliver scalable enterprise AI inference services and model registry capabilities. Enable generative AI applications using foundation models, RAG, and vector databases; collaborate with frontend, data scientists, product, and UX teams to drive production-ready AI platform features on Kubernetes-based microservices.
Top Skills: AWSAzureC#CSSFoundation ModelsGCPGoGrpcHiveHTMLHuggingfaceJavaKnativeKserveKubeflowKubernetesMilvusMlflowNimNode.jsNvidia Ai FrameworksPineconePrompt EngineeringPythonReactRetrieval-Augmented Generation (Rag)SparkSQLTensorFlow

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account