Redhorse Logo

Redhorse

Senior Data Engineer

Posted 9 Days Ago
Remote
Hiring Remotely in USA
190K-210K Annually
Senior level
Remote
Hiring Remotely in USA
190K-210K Annually
Senior level
Support aircraft sustainment and logistics initiatives through data collection, ETL pipeline development, exploratory and statistical analysis, predictive modeling, knowledge graph capabilities, and production model integration. Monitor data quality and model performance while collaborating with engineering teams and government stakeholders in an Agile environment. The role requires NLP, Databricks, PySpark, SQL, predictive modeling, data governance, visualization, and an active Secret clearance.
The summary above was generated by AI
About the Organization
Now is a great time to join Redhorse Corporation. We are a solution-driven company delivering data insights and technology solutions to customers with missions critical to U.S. national interests. We’re looking for thoughtful, skilled professionals who thrive as trusted partners building technology-agnostic solutions and want to apply their talents supporting customers with difficult and important mission sets.

About the Role

Redhorse Corporation is seeking a Data Scientist to support Condition-Based Maintenance Plus (CBM+) initiatives to enhance aircraft sustainment and optimize maintenance and logistics efficiency. The Data Scientist will help provide quality data and develop predictive maintenance and logistics forecasting models to support decision-making tools for program-wide initiatives serving U.S. Air Force (USAF) stakeholders.

This role involves working within cross-functional teams in an Agile development environment. We operate in a fast-paced, evolving domain, and are seeking creative, motivated, and talented individuals who are eager to learn, grow, and deliver effective, high-impact solutions.

Key Responsibilities

  • Support CBM+ program initiatives from inception to deployment, ensuring alignment with overarching business and mission objectives.
  • Oversee end-to-end data collection and processing, including the implementation, sustainment, and optimization of ETL pipelines.
  • Perform Exploratory Data Analysis (EDA), statistical analysis, and data visualization to identify trends, correlations, and actionable insights that drive product development.
  • Build and maintain knowledge graph capabilities.
  • Collaborate with multi-functional engineering teams to integrate trained models into production applications and APIs.
  • Continuously monitor data quality and model performance to detect and mitigate drift and degradation.

Required Qualifications

  • Active U.S. Government Secret Security Clearance (U.S. Citizenship required; applicants without an active Secret Clearance cannot be considered).
  • Bachelor’s degree in a STEM (Science, Technology, Engineering, or Mathematics) field or proven equivalent professional experience.
  • NLP Experience: 2+ years of experience incorporating applied Natural Language Processing (NLP), data labeling, and entity or keyword extraction.
  • Analytics & Data Engineering: Proven proficiency with Databricks, PySpark, SQL, data governance frameworks, and technical documentation standards.
  • Statistical Foundations: Deep understanding and practical application of statistical distributions for data assessment, data analysis, and predictive modeling.
  • Collaboration & Self-Direction: Demonstrated self-starter capabilities with strong interpersonal skills to foster positive stakeholder relationships and execute tasks independently.
  • Communication: Excellent verbal and written communication skills, with the ability to translate complex technical findings for non-technical and executive audiences.
  • Tools: Experience with project management and productivity platforms including Jira, Confluence, Lucidchart, and Microsoft Office Suite (Word, Excel, PowerPoint). Experience with data end products such as Qlik Sense, Streamlit and other visualization or dashboard tools.
  • Travel: Ability to travel, minimal, as needed to engage with government customers and project stakeholders.
  • Growth: Possess a desire to learn and advance skills in the aviation data space.

Preferred Qualifications

  • Master’s degree in a STEM field, with a preference for Data Science, Data Analytics, or Computer Science.
  • Advanced expertise in applying artificial intelligence and machine learning (AI/ML) to predictive maintenance and defense logistics challenges.
  • Direct experience developing sensor-based failure prediction models and equipment health indicators.
  • Practical experience conducting Monte Carlo simulations and calculating statistical confidence intervals.
  • Demonstrated expertise in end-to-end data pipeline engineering including customer-facing UI/UX integration.

The salary range provided for this position represents the anticipated base salary for successful candidates. Actual compensation will be determined based on a variety of factors, including relevant experience, education, certifications, skills, security clearance level, geographic location, market conditions, and internal equity. In addition to base salary, eligible employees may participate in Redhorse's comprehensive benefits programs and may be eligible for performance-based or other incentive compensation, where applicable.
 
Redhorse Corporation is an equal opportunity employer. All qualified applicants will receive consideration for employment and will not be discriminated against on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, disability, or any other protected class.
 
If you are a qualified individual with a disability or a disabled veteran, you may request a reasonable accommodation if you are unable or limited in your ability to access job openings or apply for a job on this site as a result of your disability. You can request reasonable accommodations by contacting Talent Acquisition at [email protected]
 
Redhorse Corporation shall, in its discretion, modify or adjust the position to meet Redhorse’s changing needs. This job description is not a contract and may be adjusted as deemed appropriate in Redhorse’s sole discretion.

Similar Jobs

6 Days Ago
In-Office or Remote
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Senior Data Engineer responsible for Epic and EHR integrations supporting value-based care, risk and quality workflows, member attribution, roster management, and care-gap tools. The role designs and executes EHR development tasks, coordinates with clinical, data, operations, and development teams, documents business and data flows, evaluates AI and automation opportunities, and develops reporting to identify data-quality issues and prevent outages.
Top Skills: Ai ToolsCaboodleClarityEhr IntegrationsEpicHealthy PlanetSQL
10 Days Ago
Remote or Hybrid
United States
165K-235K Annually
Senior level
165K-235K Annually
Senior level
Big Data • Cloud • Productivity • Software • Database • Analytics • Automation
Build and maintain Databricks-based data platforms, including ingestion, transformation, storage, governance, data modeling, and serving pipelines. Establish medallion architecture standards, canonical data models, quality controls, lineage, schema evolution, and reliable batch or incremental processing. Improve pipeline observability, scalability, idempotency, and recoverability while moving curated data to systems such as ClickHouse. Collaborate across application and analytics teams to create durable, governed production datasets.
Top Skills: Amazon AuroraAmazon RdsApache AirflowSparkBigQueryCdcClickhouseCloud Object StorageDatabricksDelta LakeIamOpenmetadataPostgresSnowflakeUnity Catalog
16 Days Ago
Remote or Hybrid
17 Locations
110K-200K Annually
Senior level
110K-200K Annually
Senior level
Information Technology • Productivity • Software • Infrastructure as a Service (IaaS)
Design and scale lakehouse architecture, streaming and batch pipelines, and reliable data platforms using Kafka, Spark, Airflow, Iceberg, Databricks, and related technologies. Build Medallion-layer data systems, manage open table formats, optimize distributed queries, monitor platform reliability, resolve data quality issues, and collaborate with analysts, data scientists, and product teams.
Top Skills: Apache AirflowApache HudiApache IcebergApache KafkaSparkDatabricksDelta LakePythonSQLStarburstTrino

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account