Maximum of 25 job preferences reached.
Top Remote Data Engineer Jobs in NYC, NY
Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Build and operate production data pipelines and dimensional models using Spark, SparkSQL, and cloud lakehouse technologies. Own pipelines from requirements through deployment, monitoring, and iteration; improve data quality, lineage, reliability, and cost efficiency. Partner with data scientists, analysts, product managers, and engineers to support datamarts, KPIs, reporting, and analysis. Participate in business-hours on-call rotations and improve runbooks and alerting.
Top Skills:
AirflowC++DatabricksJavaKafkaKinesisMonte CarloPythonScalaSparkSparksqlSQLStructured Streaming
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Build and operate production data pipelines supporting analytics, AI, and agentic workflows. Responsibilities include implementing canonical data models, maintaining Databricks or Snowflake platforms, monitoring reliability, responding to incidents, validating healthcare data mappings, improving performance and cost, and documenting architecture. Requires strong SQL and Python skills, cloud data platform experience, ETL/ELT orchestration expertise, and healthcare or pharmaceutical data experience.
Top Skills:
Ai/Ml WorkflowsDatabricksEltETLHedisOmopPythonSnowflakeSQL
AdTech • Automotive • Big Data • Consumer Web
Administer and enhance Edmunds’ Databricks data platform and AWS infrastructure. Build and maintain ETL pipelines, infrastructure-as-code tooling using Terraform or CDK, and operational dashboards, alerts, and reports. Collaborate with business, engineering, analytics, security, and infrastructure teams to support data platform users and AI solutions. Evaluate new technologies, troubleshoot platform issues, and improve operational, cost, and security visibility.
Top Skills:
SparkAWSAws CdkDatabricksInfrastructure As Code (Iac)PythonScalaSQLTerraform
Artificial Intelligence • Cloud • Payments • Software • Business Intelligence • Generative AI • Automation
Define and govern enterprise-scale data architecture across batch, streaming, warehouse, lakehouse, transactional, and AI use cases. Establish standards for data quality, lineage, access, cataloging, governance, observability, and SLAs. Architect AI-enabled workflows, resolve complex architecture issues, influence roadmaps, and mentor engineers through hands-on technical leadership. The role requires 15+ years of software, data engineering, or architecture experience and expertise in large-scale data platforms and modeling.
Top Skills:
AIBatch ProcessingBigQueryData CatalogsData WarehousesDbtFeature StoresGCPLakehousesOlapOltpStreaming ArchitecturesVector Stores
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Owns the design, development, operation, and governance of data pipelines and integration patterns supporting Medical Affairs AI products. Builds APIs and ETL/ELT solutions connecting enterprise systems to analytics platforms, RAG pipelines, vector databases, and GraphRAG applications. Ensures data quality, privacy, lineage, compliance, and protection of sensitive information. Partners with architecture, engineering, product, and business stakeholders to deliver reusable, production-grade data integrations.
Top Skills:
Ai/MlApi GatewaysCi/CdEtl/EltEvent-Driven IntegrationGraph DatabasesGraphQLGraphragLow-Code/No-Code ToolsMiddlewareRagRestSalesforce Life Sciences/Marketing CloudSnowflakeSQLStreaming IntegrationVector DatabasesVeeva Crm
Information Technology • Productivity • Software • Infrastructure as a Service (IaaS)
Design and scale lakehouse architecture, streaming and batch pipelines, and reliable data platforms using Kafka, Spark, Airflow, Iceberg, Databricks, and related technologies. Build Medallion-layer data systems, manage open table formats, optimize distributed queries, monitor platform reliability, resolve data quality issues, and collaborate with analysts, data scientists, and product teams.
Top Skills:
Apache AirflowApache HudiApache IcebergApache KafkaSparkDatabricksDelta LakePythonSQLStarburstTrino
Information Technology • Professional Services • Consulting
Build and maintain large-scale data processing pipelines using Apache Spark (PySpark/Scala) and Python. Design ETL workflows, implement distributed data processing, use version control (Git), and work with cloud platforms (GCP/AWS/Azure).
Top Skills:
SparkAWSAzureETLGCPGitPysparkPythonScala
Marketing Tech • Energy
Build and maintain production data pipelines and foundational datasets for Market Intelligence. Automate survey ingestion, cleaning, joining, scheduling, monitoring, and reporting-layer integration. Develop repeatable queries, improve data quality practices, document assumptions, support reproducible outputs, and translate research needs into technical solutions. The role requires close collaboration with technical and nontechnical stakeholders and ownership of reliable data infrastructure.
Top Skills:
Cloud Data WarehousingETLPythonSQL
Information Technology • Consulting
Build and operate platform infrastructure for product analytics, data processing, and ML systems. Responsibilities include event tracking pipelines, Databricks Bronze-to-Gold lakehouse layers, schema evolution, data contracts, CI validation, backfills, monitoring, feature computation, model pipelines, lineage, PHI policy enforcement, and self-service developer tooling. The role focuses on making data and ML changes safe, observable, reproducible, and accessible to engineers and subject-matter experts.
Top Skills:
AirflowBigQueryDagsterDatabricksDbtDltKafkaPythonSnowflakeSpark Structured StreamingSQLTerraform
Software
Own and evolve WorkOS’s internal data platform, including ingestion, orchestration, Snowflake, dbt models, data governance, reverse ETL, semantic views, and AI-enabled data workflows. Ensure pipeline reliability, freshness, security, and scalability while partnering with Product, Finance, RevOps, GTM, Engineering, and Security. Build monitoring, runbooks, CI/CD gates, access controls, masking policies, and infrastructure automation across AWS and Kubernetes.
Top Skills:
AirflowAWSCi/CdDagsterDbtIamKubernetesPostgresPrefectPythonRbacSalesforceSlackSnowflakeSQLTerraform
Fintech • Software • Analytics • Financial Services
Design, build, and maintain scalable data pipelines, data lakes, and databases; ingest and map customer financial datasets; monitor pipeline reliability; translate business requirements into data flows and analytical insights to ensure high-quality, usable data.
Top Skills:
APIsBigQueryBigtableData LakesDatabricksDbtFivetranGithub ActionsMariadbMongoDBMySQLNoSQLPostgresRedshiftSnowflakeSQL
Consulting
Build and maintain data pipelines, governance documentation, and data architectures supporting Department of State systems. Collect, clean, transform, integrate, validate, and analyze data from multiple sources using Palantir Foundry and relational databases. Collaborate with nontechnical stakeholders to gather requirements, identify data-quality issues, develop reporting tools and visualizations, and recommend improvements supporting data-driven decisions.
Top Skills:
ArchibusData PipelinesKahuaPalantir FoundryPower BIPythonRelational DatabasesSQLTableau
New
Track Smarter, Apply Better.
Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.
Use For Free
Artificial Intelligence • Information Technology • Software • Consulting
Build and optimize scalable ETL pipelines, monitoring and alerting systems, and reliable data solutions for legal and eDiscovery infrastructure. Responsibilities include query and workflow optimization, cross-functional requirements gathering, system documentation, deployment planning, and compliance with data privacy and security standards. The role also contributes to internal consulting, thought leadership, and practice development.
Top Skills:
Apache AirflowETLMySQLPostgresPrestoPythonSQL
Artificial Intelligence • Information Technology • Software • Consulting
Design, build, and optimize complex ETL pipelines, data workflows, monitoring, and alerting systems supporting legal and eDiscovery infrastructure. Lead end-to-end data engineering initiatives, improve query and pipeline performance, document deployment strategies, and ensure privacy, security, and compliance. Collaborate with attorneys, data scientists, engineers, and other stakeholders in regulated environments while contributing to internal consulting and thought leadership initiatives.
Top Skills:
Ai/MlApache AirflowETLMySQLPostgresPrestoPythonSQL
Healthtech
Designs and builds cloud-based, data-centric applications supporting clinical and operational healthcare processes. Develops data pipelines, transformations, enrichment processes, provisioning layers, and user interfaces. Collaborates with Product, Platform, and Architecture teams using modern software development practices, source control, documentation, and regular delivery methods.
Top Skills:
Big DataCloud ComputingData PipelinesData ScienceData TransformationsSource Control
Consulting
Designs, builds, tests, deploys, and maintains secure ETL/ELT pipelines for public health data. Integrates Power Platform, Azure services, APIs, SQL Server, CDC/NIOSH systems, and approved file-transfer platforms. Develops batch, event-driven, and near-real-time data workflows supporting intake, validation, reporting, analytics, notifications, and results delivery while meeting FISMA Moderate, federal privacy, Agile, and DevSecOps requirements.
Top Skills:
AgileAPIsDataverseDevsecopsEltETLAzureMicrosoft Power PlatformPower AppsPower AutomatePower PagesPythonSQLSQL Server
Biotech • Agriculture
Designs and maintains scalable batch and streaming data pipelines and datasets in Databricks using Python, Spark, and SQL. Supports SQL Server, SSIS, SSRS, Azure, data warehousing, reporting, analytics, and AI/ML workloads. Responsibilities include optimizing performance and reliability, implementing data quality controls, integrating enterprise data sources, modernizing platforms, and partnering with business and IT teams to deliver scalable data solutions.
Top Skills:
SparkDatabricksDelta LakeAzureMicrosoft FabricPythonSQLSQL ServerSsisSsrsT-Sql
Fintech • News + Entertainment • Software • Database • Financial Services
Lead the architecture and development of scalable AWS-based data ingestion, transformation, and orchestration pipelines. Build reliable data infrastructure using Python, SQL, Airflow, Lambda, ECS, SQS, and Terraform. Establish data modeling, quality, lineage, monitoring, testing, and observability practices while partnering with analysts, scientists, and backend engineers. Mentor senior engineers, guide technical decisions, and provide hands-on leadership for complex data platform initiatives.
Top Skills:
Amazon EcsAmazon KinesisAmazon RedshiftAmazon S3Amazon SqsApache AirflowApache FlinkAWSAws GlueAws LambdaBeautifulsoupCi/CdDatabricksDockerGreat ExpectationsKafkaMonte CarloMwaaPythonScrapySnowflakeSQLTerraform
Agriculture
Designs, builds, and maintains scalable batch and streaming data pipelines in Databricks using Python, Spark, SQL, and Delta Lake. Supports SQL Server, SSIS, and SSRS environments while improving performance, reliability, data quality, and scalability. Integrates enterprise data sources, applies warehousing best practices, and partners with business and IT teams to deliver reporting, analytics, and AI/ML data solutions.
Top Skills:
SparkDatabricksDelta LakeAzureMicrosoft FabricPythonSQLSQL ServerSsisSsrsT-Sql
Artificial Intelligence • Cloud • Software
Lead the architecture and development of Vercel’s next-generation data platform, supporting batch and real-time integrations, analytics, data warehousing, and AI/ML workloads. Design scalable systems using Kafka, ClickHouse, Tinybird, and Snowflake; establish data governance and security standards; guide architectural decisions and roadmaps; collaborate with engineering, product, security, compliance, and leadership teams; write production code; and mentor engineers.
Top Skills:
AWSAzureBig Data FrameworksClickhouseConfluent PlatformData GovernanceData WarehousingETLGCPKafkaKafka StreamsSnowflakeTinybird
Travel
Build and maintain scalable ETL/ELT pipelines, integrate travel and customer data from internal and external systems, develop cloud data warehouse models, implement data quality monitoring, support real-time streaming, and ensure data privacy and regulatory compliance. Collaborate with analysts, data scientists, and product teams to deliver reliable datasets for personalization, forecasting, pricing, route optimization, and customer experience initiatives.
Top Skills:
AirflowAWSAzureBigQueryCcpaDbtGCPGdprIdmcJavaKafkaPythonRedshiftScalaSnowflakeSQL
Healthtech
Design, build, and maintain scalable data pipelines, data models, and analytics solutions using Microsoft Fabric and Azure. Develop Lakehouse and Warehouse architectures, Fabric Notebooks with PySpark and Spark SQL, ETL/ELT workflows, and Power BI semantic models. Optimize performance, ensure data quality, security, governance, and regulatory compliance, and collaborate with analysts, data scientists, and business stakeholders to support business intelligence and advanced analytics.
Top Skills:
Azure Data LakeAzure Data ServicesAzure Synapse AnalyticsData FactoryDataflows Gen2Delta LakeDirect LakeFabric LakehouseFabric PipelinesFabric WarehouseMicrosoft FabricOnelakePower BIPysparkPythonScalaSpark SqlSQL
Software
Build and maintain Databricks data pipelines across bronze, silver, and gold layers. Ingest APIs, logs, billing exports, and reference data; implement attribution logic, governance, data quality monitoring, and cost optimization. Manage Unity Catalog permissions, lineage, refresh schedules, incremental processing, and CI/CD workflows while supporting multi-cloud storage and high-volume caller-identity data.
Top Skills:
SparkAsset BundlesAuto LoaderAWSAzureCi/CdDatabricksDatabricks WorkflowsDelta LakeGCPGitPysparkPythonRest ApisSQLUnity Catalog
Consumer Web • Digital Media • eCommerce • News + Entertainment
Build and maintain ETL/ELT pipelines, ingest and transform structured and unstructured data, integrate APIs and external sources, improve data quality and monitoring, support dashboards and ML pipelines, optimize warehouse and cloud performance, and contribute to data governance and documentation.
Top Skills:
APIsBigQueryEltETLGCPPostgresPythonSnowflakeSQL
Big Data • Cloud • Productivity • Software • Database • Analytics • Automation
Build and maintain Databricks-based data platforms, including ingestion, transformation, storage, governance, data modeling, and serving pipelines. Establish medallion architecture standards, canonical data models, quality controls, lineage, schema evolution, and reliable batch or incremental processing. Improve pipeline observability, scalability, idempotency, and recoverability while moving curated data to systems such as ClickHouse. Collaborate across application and analytics teams to create durable, governed production datasets.
Top Skills:
Amazon AuroraAmazon RdsApache AirflowSparkBigQueryCdcClickhouseCloud Object StorageDatabricksDelta LakeIamOpenmetadataPostgresSnowflakeUnity Catalog
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Companies in NYC, NY Hiring Remote Data Engineers
See AllPopular NYC Remote Job Searches
Remote Jobs in NYC
Remote .NET Developer Jobs in NYC
Remote Account Executive (AE) Jobs in NYC
Remote Account Manager (AM) Jobs in NYC
Remote Analysis Reporting Jobs in NYC
Remote Analytics Jobs in NYC
Remote Android Developer Jobs in NYC
Remote Business Intelligence Jobs in NYC
Remote C# Jobs in NYC
Remote C++ Jobs in NYC
Remote Content Jobs in NYC
Remote Customer Success Jobs in NYC
Remote Cyber Security Jobs in NYC
Remote Data & Analytics Jobs in NYC
Remote Data Engineer Jobs in NYC
Remote Data Management Jobs in NYC
Remote Data Science Jobs in NYC
Remote UX Designer Jobs in NYC
Remote DevOps Jobs in NYC
Remote Engineering Manager Jobs in NYC
Remote Finance Jobs in NYC
Remote Front End Developer Jobs in NYC
Remote Golang Jobs in NYC
Remote Hardware Engineer Jobs in NYC
Remote HR Jobs in NYC
Remote Internships in NYC
Remote iOS Developer Jobs in NYC
Remote IT Jobs in NYC
Remote Java Developer Jobs in NYC
Remote Javascript Jobs in NYC
Remote Legal Jobs in NYC
Remote Linux Jobs in NYC
Remote Machine Learning Jobs in NYC
Remote Marketing Jobs in NYC
Remote Office Manager Jobs in NYC
Remote Operations Jobs in NYC
Remote Operations Manager Jobs in NYC
Remote PHP Developer Jobs in NYC
Remote Product Manager Jobs in NYC
Remote Project Management Jobs in NYC
Remote Python Jobs in NYC
Remote QA Engineer Jobs in NYC
Remote Ruby Jobs in NYC
Remote Sales Development Representative Jobs in NYC
Remote Sales Engineer Jobs in NYC
Remote Sales Jobs in NYC
Remote Sales Leadership Jobs in NYC
Remote Sales Operations Jobs in NYC
Remote Salesforce Developer Jobs in NYC
Remote Scala Jobs in NYC
Remote Software Engineer Jobs in NYC
Remote Tech Support Jobs in NYC
All Filters
Total selected ()
No Results
No Results



.png)















.jpg)









