EXL Logo

EXL

Data Engineer

Reposted 7 Days Ago
Remote or Hybrid
Hiring Remotely in United States
Mid level
Remote or Hybrid
Hiring Remotely in United States
Mid level
Design, build, and maintain scalable data pipelines and ETL/ELT frameworks using Python and PySpark. Develop reusable components, microservices, and cloud-based data solutions on AWS and Snowflake. Ensure code quality via reviews, CI/CD, logging, and best practices while collaborating with stakeholders to solve data problems.
The summary above was generated by AI

We are looking for a strong hands-on Data Engineer/Developer with solid coding and software development experience. The ideal candidate should have a development-first mindset, be comfortable writing production-quality code, building scalable data applications, and working with modern cloud and data platforms.
The candidate should have strong expertise in Python, PySpark, AWS, Snowflake, APIs, Microservices, and software engineering best practices.

For more information on benefits and what we offer please visit us at https://www.exlservice.com/us-careers-and-benefits

Base Compensation Range: 57,000 - 90,000

The posted range is the hiring range for this role — a subset of the broader range available to employees over time — and reflects base salary across our national hiring scale. Final offers are based on several factors, including the candidate's skills and experience, internal pay equity, work location, market conditions for the role, and the specific scope and responsibilities of the position. The top of the range is reserved for candidates who notably exceed the requirements; the lower end applies to those with less experience or fewer preferred qualifications. For positions based in higher-cost zones (e.g., California, New York, New Jersey), actual compensation may exceed the posted range; your recruiter will share specifics during the process.

Responsibilities

Key Responsibilities Application & Data Development

  • Design, develop, and maintain scalable data pipelines, applications, and utilities using Python and PySpark.
  • Build robust ETL/ELT frameworks and data processing solutions for large-scale datasets.
  • Develop reusable components, libraries, and automation solutions following coding standards and best practices.
  • Write clean, maintainable, and efficient code with proper error handling and logging.
  • Participate in code reviews and ensure adherence to development standards.

Must-Have Skills

  • Strong hands-on development experience in Python.
  • Experience with PySpark and large-scale data processing.
  • Strong understanding of software engineering principles, coding standards, design patterns, and best practices.
  • Experience with AWS Cloud services such as Glue, Lambda, S3, EC2, API Gateway, ECS, etc.
  • Hands-on experience with Snowflake or other cloud data platforms.
  • Solid understanding of relational and non-relational databases.
  • Experience developing Microservices.
  • Knowledge of CI/CD pipelines, Git, version control, and DevOps practices.
  • Strong SQL and data modeling skills.
  • Excellent analytical, debugging, and problem-solving skills.
  • Strong communication and stakeholder management skills
Qualifications

Bachelor's or Master's degree in Computer Science, Information Technology, Engineering, or a related field.

Similar Jobs

4 Days Ago
Remote or Hybrid
113K-193K Annually
Senior level
113K-193K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design and operate scalable batch and streaming data platforms supporting machine learning and generative AI. Build pipelines for structured, unstructured, OCR, document, and image data; develop RAG, semantic search, and LLM-powered solutions; and establish data quality, observability, governance, orchestration, and deployment practices. Partner with stakeholders, lead platform scalability and cost optimization, mentor engineers, and translate ambiguous needs into production-ready technical roadmaps while securely handling sensitive data.
Top Skills: AirflowAmazon KinesisSparkAWSAzureAzure Event HubsChart.JsDatabricksDeequDelta LakeDockerGithub ActionsGCPGreat ExpectationsJavaKafkaKubernetesLlmsMlopsPlotlyPysparkPythonRagScalaSeabornSnowflakeSQLTerraform
5 Days Ago
Remote
United States
Mid level
Mid level
Artificial Intelligence • Information Technology • Professional Services • Software • Analytics • Generative AI • Big Data Analytics
Configure and maintain Adobe Experience Platform and Real-Time CDP data pipelines, XDM schemas, datasets, identity resolution, Profile enablement, segmentation readiness, and destination activation. Validate data quality, troubleshoot ingestion and identity issues, manage sandboxes, and implement privacy, consent, governance, and access controls. Partner with architects, data engineers, consultants, analysts, data scientists, and marketing teams to support reporting, personalization, and machine learning readiness.
Top Skills: Adobe AnalyticsAdobe Experience PlatformAdobe I/O RuntimeAdobe Journey OptimizerAdobe Real-Time CdpAdobe Source ConnectorsAdobe TargetAPIsAWSAzureBatch IngestionBigQueryCcpaData PrepEltETLGCPGdprJavaScriptPythonQuery ServiceRedshiftSalesforce CdpSegmentSnowflakeSQLStreaming IngestionWeb SdkXdm
3 Days Ago
In-Office or Remote
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Lead cloud data modernization by designing and building cloud-native ETL/ELT pipelines, data warehouses, and lakehouse architectures using Azure, Snowflake, and Databricks. Ensure data quality, security, governance, CI/CD, and healthcare interoperability while mentoring engineers and collaborating with stakeholders to deliver scalable, HIPAA-compliant analytics platforms.
Top Skills: AdlsApache AirflowAzureAzure Data Factory (Adf)Azure Data Lake Storage Gen2 (Adls Gen2)Azure OpenaiCi/CdCptDatabricksDatabricks GenieDatabricks Mosaic AiDockerEltETLFhirGitGithub ActionsHcpcsHl7Icd-10KafkaKubernetesLlmsLoincPysparkPythonRag PipelinesSnowflakeSnowflake CortexSparkSQL ServerSsisTerraformVector StoresVisioX12 Edi (837/835/834)

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account