Photon Logo

Photon

SPARK Data Reconciliation Engineer- NJ

Reposted 21 Days Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in United States
Senior level
In-Office or Remote
Hiring Remotely in United States
Senior level
Design, implement, and maintain PySpark applications to automate large-scale financial data reconciliations. Build transformation and matching algorithms, integrate with rules engines, analyze data gaps, and collaborate with analysts and architects to ensure data quality and system resilience.
The summary above was generated by AI

Job Title: PySpark Data Reconciliation Engineer

Summary:

We're seeking a skilled PySpark Data Reconciliation Engineer to join our team and drive the development of robust data reconciliation solutions within our financial systems. You will be responsible for designing, implementing, and maintaining PySpark-based applications to perform complex data reconciliations, identify and resolve discrepancies, and automate data matching processes. The ideal candidate possesses strong PySpark development skills, experience with data reconciliation techniques, and the ability to integrate with diverse data sources and rules engines.

Key Responsibilities:

Data Reconciliation Development:

  • Design, develop, and test PySpark-based applications to automate data reconciliation processes across various financial data sources, including relational databases, NoSQL databases, batch files, and real-time data streams.
  • Implement efficient data transformation, matching algorithms (deterministic and heuristic) using PySpark and relevant big data frameworks.
  • Develop robust error handling and exception management mechanisms to ensure data integrity and system resilience within Spark jobs.

Data Analysis and Matching:

  • Collaborate with business analysts and data architects to understand data requirements and matching criteria.
  • Analyze and interpret data structures, formats, and relationships to implement effective data matching algorithms using PySpark.
  • Work with distributed datasets in Spark, ensuring optimal performance for large-scale data reconciliation.

Rules Engine Integration:

  • Integrate PySpark applications with rules engines (e.g., Drools) or equivalent to implement and execute complex data matching rules.
  • Develop PySpark code to interact with the rules engine, manage rule execution, and handle rule-based decision-making.

Problem Solving and Gap Analysis:

  • Collaborate with cross-functional teams to identify and analyze data gaps and inconsistencies between systems.
  • Design and develop PySpark-based solutions to address data integration challenges and ensure data quality.
  • Contribute to the development of data governance and quality frameworks within the organization.

Qualifications and Skills:

  • Bachelor's degree in Computer Science or a related field.
  • 5+ years of hands-on experience in big data development, preferably with exposure to data-intensive applications.
  • Strong understanding of data reconciliation principles, techniques, and best practices.
  • Proficiency in PySpark, Apache Spark, and related big data technologies for data processing and integration.
  • Experience with rules engine integration and development 
  • Strong analytical and problem-solving skills, with the ability to translate business requirements into technical solutions.
  • Excellent communication and collaboration skills to work effectively with business analysts, data architects, and other team members.
  • Familiarity with data streaming platforms (e.g., Kafka, Kinesis) and big data technologies (e.g., Hadoop, Hive, HBase) is a plus.

Similar Jobs

10 Minutes Ago
In-Office or Remote
Senior level
Senior level
Artificial Intelligence • Big Data • Cloud • Machine Learning • Software • Business Intelligence • Data Privacy
The role involves designing cloud infrastructure, managing production Kubernetes clusters, optimizing CI/CD pipelines, enhancing developer experience, and ensuring reliable AI workloads. Candidates should have extensive experience in infrastructure and distributed systems engineering with strong coding skills and cloud expertise.
Top Skills: AWSAzureDatadogDockerElkGCPGoGrafanaJavaKubernetesPrometheusPythonTerraform
12 Minutes Ago
Easy Apply
In-Office or Remote
IN, USA
Easy Apply
Junior
Junior
Artificial Intelligence • Healthtech • Software • Telehealth
Own end-to-end recruiting for Registered Dietitians: vet candidates, manage outreach and communications, set up offers, track hiring metrics, and improve recruiting processes across teams to increase funnel conversion and candidate experience.
17 Minutes Ago
Remote
United States
20-20 Annually
Entry level
20-20 Annually
Entry level
Healthtech • Information Technology • Social Impact • Software • App development
Perform telephonic outreach to individuals with serious mental illness (SMI) to connect them with local firsthand Guides, schedule appointments, build trust using lived-experience approaches, document encounters, participate in team case reviews, and support long-term engagement in recovery-focused services. Complete paid training and ongoing professional development.
Top Skills: AppsEmailMessaging ServicesSmartphones

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account