Photon Logo

Photon

SPARK Data Reconciliation Engineer- NJ

Posted 2 Days Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in United States
Senior level
In-Office or Remote
Hiring Remotely in United States
Senior level
Design, implement, and maintain PySpark applications to automate large-scale financial data reconciliations. Build transformation and matching algorithms, integrate with rules engines, analyze data gaps, and collaborate with analysts and architects to ensure data quality and system resilience.
The summary above was generated by AI

Job Title: PySpark Data Reconciliation Engineer

Summary:

We're seeking a skilled PySpark Data Reconciliation Engineer to join our team and drive the development of robust data reconciliation solutions within our financial systems. You will be responsible for designing, implementing, and maintaining PySpark-based applications to perform complex data reconciliations, identify and resolve discrepancies, and automate data matching processes. The ideal candidate possesses strong PySpark development skills, experience with data reconciliation techniques, and the ability to integrate with diverse data sources and rules engines.

Key Responsibilities:

Data Reconciliation Development:

  • Design, develop, and test PySpark-based applications to automate data reconciliation processes across various financial data sources, including relational databases, NoSQL databases, batch files, and real-time data streams.
  • Implement efficient data transformation, matching algorithms (deterministic and heuristic) using PySpark and relevant big data frameworks.
  • Develop robust error handling and exception management mechanisms to ensure data integrity and system resilience within Spark jobs.

Data Analysis and Matching:

  • Collaborate with business analysts and data architects to understand data requirements and matching criteria.
  • Analyze and interpret data structures, formats, and relationships to implement effective data matching algorithms using PySpark.
  • Work with distributed datasets in Spark, ensuring optimal performance for large-scale data reconciliation.

Rules Engine Integration:

  • Integrate PySpark applications with rules engines (e.g., Drools) or equivalent to implement and execute complex data matching rules.
  • Develop PySpark code to interact with the rules engine, manage rule execution, and handle rule-based decision-making.

Problem Solving and Gap Analysis:

  • Collaborate with cross-functional teams to identify and analyze data gaps and inconsistencies between systems.
  • Design and develop PySpark-based solutions to address data integration challenges and ensure data quality.
  • Contribute to the development of data governance and quality frameworks within the organization.

Qualifications and Skills:

  • Bachelor's degree in Computer Science or a related field.
  • 5+ years of hands-on experience in big data development, preferably with exposure to data-intensive applications.
  • Strong understanding of data reconciliation principles, techniques, and best practices.
  • Proficiency in PySpark, Apache Spark, and related big data technologies for data processing and integration.
  • Experience with rules engine integration and development 
  • Strong analytical and problem-solving skills, with the ability to translate business requirements into technical solutions.
  • Excellent communication and collaboration skills to work effectively with business analysts, data architects, and other team members.
  • Familiarity with data streaming platforms (e.g., Kafka, Kinesis) and big data technologies (e.g., Hadoop, Hive, HBase) is a plus.

Photon New York, New York, USA Office

New York, United States

Similar Jobs

24 Minutes Ago
In-Office or Remote
New York, NY, USA
140K-185K Annually
Senior level
140K-185K Annually
Senior level
Blockchain • Fintech • Payments • Financial Services • Cryptocurrency • Web3
Lead full-cycle recruiting for Finance and core business roles, partner with hiring leaders and cross-functional teams, maintain TA systems, and build/test agentic AI bots and Skills using Claude Code/OpenAI Codex to automate recruiting workflows and improve process quality, governance, and candidate experience.
Top Skills: Claude CodeGemGoogle WorkspaceLinkedin RecruitermacOSModernloopOpenai CodexSlackWorkday
2 Hours Ago
Remote or Hybrid
East Hanover, NJ, USA
143K-235K Annually
Senior level
143K-235K Annually
Senior level
Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Lead SnackFutures Ventures investments in early and growth-stage companies, manage deal sourcing, screening, negotiation, and portfolio oversight, coordinate cross-functional teams, contribute to investment strategy, and support exits or handovers to Corporate Development.
2 Hours Ago
Remote or Hybrid
CA, USA
37K-73K Hourly
Mid level
37K-73K Hourly
Mid level
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Provide white-glove technical support and onboarding for high-value resellers: troubleshoot integrations, manage escalations, run onboarding and training, track and drive issue resolution with engineering, and collaborate cross-functionally to improve processes and product experience for enterprise sellers.
Top Skills: APIsGoogle MeetsJIRASdksThird-Party Integrations

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account