Ness Digital Engineering Logo

Ness Digital Engineering

Data Engineer

Posted 2 Days Ago
Remote
Hiring Remotely in United States
Entry level
Remote
Hiring Remotely in United States
Entry level
Build and maintain Databricks data pipelines across bronze, silver, and gold layers. Ingest APIs, logs, billing exports, and reference data; implement attribution logic, governance, data quality monitoring, and cost optimization. Manage Unity Catalog permissions, lineage, refresh schedules, incremental processing, and CI/CD workflows while supporting multi-cloud storage and high-volume caller-identity data.
The summary above was generated by AI

Key responsibilities 
• Build ingestion into the bronze layer for assigned sources: gateway and observability logs, productivity 
tool admin APIs, AI-enabled SaaS usage, hyperscaler billing exports and reference data. Land raw and 
untransformed, on a scheduled refresh, replayable if the downstream design changes. 
• Work to the shared bronze landing contract so each tool is ingested once and serves both this program 
and the parallel productivity initiative, rather than being integrated twice. 
• Build the silver layer: typed, deduplicated and conformed to the canonical dimensions, refreshed 
independently of any downstream publication schedule. 
• Build gold marts carrying attribution method, attribution level, cost basis and provisional status alongside 
cost and usage. 
• Implement the attribution and allocation logic designed by the analysts, including precedence resolution 
and ratio-based splitting of shared endpoint cost. 
• Work within Unity Catalog governance — shared bronze and silver, separate gold marts with a recorded 
owner per dataset — including permissions, lineage and cataloging. 
• Implement data quality rules and monitoring: completeness, freshness and tag-coverage checks with 
alerting, so pipeline problems surface before they reach a divisional invoice. 
• Manage the volume impact of enabling caller-identity data in the cost and usage report, which multiplies 
row counts by the number of calling identities per model. 
• Work to the per-source cadence — daily where controls and anomaly detection depend on it, monthly 
where they do not — within the team's existing CI/CD and promotion practices. 
Essential skills and experience 
• Advanced Databricks engineering: Delta Lake, medallion architecture, Databricks Workflows, Auto 
Loader and incremental ingestion patterns. 
• Unity Catalog to a governance standard — catalogs, schemas, permissions, lineage — not merely as a 
place tables happen to live. 
• Strong Python and PySpark, and strong SQL. Notebook-based development. 
• Ingestion from REST APIs including pagination, throttling, incremental watermarks and credential 
handling, plus cloud object storage across AWS, Azure and GCP. 
• Performance and cost optimization of Spark workloads: partitioning, clustering, file sizing and cluster 
configuration. 
Tokenomics Program - Contract Role Descriptions  |  Page 7 
• CI/CD for Databricks — asset bundles or equivalent — and Git-based development workflow. 
• Able to work to an existing catalog structure and coding standard rather than introducing a parallel 
approach. 

HQ

Ness Digital Engineering Teaneck, New Jersey, USA Office

Teaneck, NJ, United States

Ness Digital Engineering New York, New York, USA Office

1 World Trade Center, Suite 83-E,, New York, New York, United States, 1007

Similar Jobs

Yesterday
Remote
United States
121K-164K Annually
Junior
121K-164K Annually
Junior
Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Build and operate production data pipelines and dimensional models using Spark, SparkSQL, and cloud lakehouse technologies. Own pipelines from requirements through deployment, monitoring, and iteration; improve data quality, lineage, reliability, and cost efficiency. Partner with data scientists, analysts, product managers, and engineers to support datamarts, KPIs, reporting, and analysis. Participate in business-hours on-call rotations and improve runbooks and alerting.
Top Skills: AirflowC++DatabricksJavaKafkaKinesisMonte CarloPythonScalaSparkSparksqlSQLStructured Streaming
Yesterday
Remote or Hybrid
OH, USA
Junior
Junior
Financial Services
Develop and maintain secure, scalable data pipelines and production code using AWS, PySpark, Python, and ETL technologies. Extract and transform data, implement quality checks, optimize data processing workflows, support cloud data modernization, and ensure reliable data availability. Collaborate with cross-functional teams, troubleshoot technical issues, automate recurring remediation, and communicate with technical and non-technical stakeholders. The role also involves software testing, operational stability, continuous delivery, and potentially mentoring other engineers.
Top Skills: Ab InitioAirflowAWSCi/CdData LakesDatabricksETLInformaticaJavaPysparkPythonSnowflakeSQL
13 Days Ago
In-Office or Remote
New York City, NY, USA
124K-207K Annually
Senior level
124K-207K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Build and operate production data pipelines supporting analytics, AI, and agentic workflows. Responsibilities include implementing canonical data models, maintaining Databricks or Snowflake platforms, monitoring reliability, responding to incidents, validating healthcare data mappings, improving performance and cost, and documenting architecture. Requires strong SQL and Python skills, cloud data platform experience, ETL/ELT orchestration expertise, and healthcare or pharmaceutical data experience.
Top Skills: Ai/Ml WorkflowsDatabricksEltETLHedisOmopPythonSnowflakeSQL

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account