Photon Logo

Photon

SPARK Data Reconciliation Engineer- NJ

Posted Yesterday
Be an Early Applicant
In-Office or Remote
Hiring Remotely in United States
Senior level
In-Office or Remote
Hiring Remotely in United States
Senior level
Design, implement, and maintain PySpark applications to automate large-scale financial data reconciliations. Build transformation and matching algorithms, integrate with rules engines, analyze data gaps, and collaborate with analysts and architects to ensure data quality and system resilience.
The summary above was generated by AI

Job Title: PySpark Data Reconciliation Engineer

Summary:

We're seeking a skilled PySpark Data Reconciliation Engineer to join our team and drive the development of robust data reconciliation solutions within our financial systems. You will be responsible for designing, implementing, and maintaining PySpark-based applications to perform complex data reconciliations, identify and resolve discrepancies, and automate data matching processes. The ideal candidate possesses strong PySpark development skills, experience with data reconciliation techniques, and the ability to integrate with diverse data sources and rules engines.

Key Responsibilities:

Data Reconciliation Development:

  • Design, develop, and test PySpark-based applications to automate data reconciliation processes across various financial data sources, including relational databases, NoSQL databases, batch files, and real-time data streams.
  • Implement efficient data transformation, matching algorithms (deterministic and heuristic) using PySpark and relevant big data frameworks.
  • Develop robust error handling and exception management mechanisms to ensure data integrity and system resilience within Spark jobs.

Data Analysis and Matching:

  • Collaborate with business analysts and data architects to understand data requirements and matching criteria.
  • Analyze and interpret data structures, formats, and relationships to implement effective data matching algorithms using PySpark.
  • Work with distributed datasets in Spark, ensuring optimal performance for large-scale data reconciliation.

Rules Engine Integration:

  • Integrate PySpark applications with rules engines (e.g., Drools) or equivalent to implement and execute complex data matching rules.
  • Develop PySpark code to interact with the rules engine, manage rule execution, and handle rule-based decision-making.

Problem Solving and Gap Analysis:

  • Collaborate with cross-functional teams to identify and analyze data gaps and inconsistencies between systems.
  • Design and develop PySpark-based solutions to address data integration challenges and ensure data quality.
  • Contribute to the development of data governance and quality frameworks within the organization.

Qualifications and Skills:

  • Bachelor's degree in Computer Science or a related field.
  • 5+ years of hands-on experience in big data development, preferably with exposure to data-intensive applications.
  • Strong understanding of data reconciliation principles, techniques, and best practices.
  • Proficiency in PySpark, Apache Spark, and related big data technologies for data processing and integration.
  • Experience with rules engine integration and development 
  • Strong analytical and problem-solving skills, with the ability to translate business requirements into technical solutions.
  • Excellent communication and collaboration skills to work effectively with business analysts, data architects, and other team members.
  • Familiarity with data streaming platforms (e.g., Kafka, Kinesis) and big data technologies (e.g., Hadoop, Hive, HBase) is a plus.

Photon San Francisco, California, USA Office

San Francisco, United States

Similar Jobs

40 Minutes Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
196K-245K Annually
Senior level
196K-245K Annually
Senior level
Cloud • Information Technology • Security • Software • Cybersecurity
Lead the enterprise data platform strategy and architecture, driving Snowflake/dbt-based platform evolution, self-service Data Mesh and medallion models. Build AI-ready pipelines, RAG systems, and observability/cost frameworks while managing a central data team, supporting federated BI, and executing hands-on technical work.
Top Skills: AiopsAutomated TestingCi/CdCortexData MeshDbtGitGraph RagMatillion Data Productivity Cloud (Matillion Dpc)Medallion ArchitecturePythonRetrieval-Augmented Generation (Rag)SnowflakeSnowparkSQLStreamlit
43 Minutes Ago
Remote or Hybrid
121K-205K Annually
Senior level
121K-205K Annually
Senior level
Aerospace • Hardware • Information Technology • Security • Software • Cybersecurity • Defense
Design and develop FPGA systems using VHDL and vendor toolchains. Create testbenches and perform simulation with QuestaSim, perform lab verification and debugging, manage configurations with Git, provide architectural input, and mentor/lead a team of engineers on communications-related FPGA projects.
Top Skills: GitIntel QuartusLinuxMentor Graphics QuestasimVhdlWindowsXilinx Vivado
44 Minutes Ago
Remote or Hybrid
133K-226K Annually
Senior level
133K-226K Annually
Senior level
Aerospace • Hardware • Information Technology • Security • Software • Cybersecurity • Defense
Lead and deliver enterprise-scale IT projects (infrastructure modernization, cloud migration, cybersecurity, compliance) across the project lifecycle. Manage schedules, EVM, budgets, vendors, resources, risks, stakeholder engagement, change management, and benefits realization. Drive agile execution and executive communications.
Top Skills: AgileCloud MigrationEarned Value Management (Evm)ExcelMicrosoft ProjectPowerPointServicenow SpmWord

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account