SMASH Logo

SMASH

Data Engineer (P-175)

Posted 11 Days Ago
Remote
Hiring Remotely in Northeast Brazos, TX, USA
Senior level
Remote
Hiring Remotely in Northeast Brazos, TX, USA
Senior level
Design, build, and maintain scalable production data pipelines and cloud-based solutions in Microsoft Azure. Develop Python, SQL, Spark, Spark SQL, and PySpark workflows for data ingestion, transformation, integration, analytics, reporting, and machine learning. Monitor, troubleshoot, validate, and optimize data pipelines and large-scale processing workloads while collaborating with technical teams and stakeholders. Document architectures and follow software engineering best practices for testing, version control, and deployment.
The summary above was generated by AI
SMASH, Who we are?

We believe in long-lasting relationships with our talent. We invest time getting to know them and understanding what they seek as their professional next step.

We aim to find the perfect match. As agents, we pair our talent with our US clients, not only by their technical skills but as a cultural fit. Our core competency is to find the right talent fast.

We purposefully move away from the “contractor” or “outsourcing” type of relationship. Our clients don’t want contractors or “just a service.” Neither does our talent.

This position and offers the opportunity to work with a US-based company. To be eligible for this role, you must have US Citizenship or valid US work authorization.

Role summary

We are looking for an experienced Data Engineer with a strong background in designing, building, and maintaining scalable data pipelines and cloud-based data solutions.

The ideal candidate will bring hands-on expertise with SQL, Python, Spark, Spark SQL, PySpark, and Microsoft Azure, with a strong emphasis on coding and end-to-end pipeline development. Experience with Microsoft Fabric and within the Pharmaceutical, Life Sciences, or Insurance industries will be highly valued.

Responsibilities
  • Design, build, and maintain automated data pipelines that move and transform data across systems.
  • Develop scalable data workflows to support analytics, reporting, and machine learning use cases.
  • Build data ingestion, transformation, processing, and integration solutions within Microsoft Azure.
  • Develop and maintain production-quality code using Python and SQL.
  • Use Apache Spark, Spark SQL, and PySpark to process and transform large-scale datasets.
  • Design data transformations that convert raw data into reliable and usable formats for downstream consumers.
  • Develop efficient data integration processes across multiple data sources and destinations.
  • Monitor and troubleshoot data pipelines to ensure reliability, accuracy, and performance.
  • Identify and resolve data quality, pipeline, and processing issues.
  • Optimize data workflows and code for performance, scalability, and maintainability.
  • Collaborate with Data Analysts, Data Scientists, engineering teams, and business stakeholders to understand data requirements.
  • Support the implementation and continuous improvement of cloud-based data engineering solutions.
  • Document pipeline architecture, transformations, dependencies, and technical processes.
  • Follow software engineering best practices for coding, testing, version control, and deployment.
Requirements – Must-haves
  • 5–6+ years of professional Data Engineering experience.
  • Proven hands-on experience designing, building, and maintaining production data pipelines.
  • Strong coding and software development capabilities.
  • Strong hands-on Python experience.
  • Advanced SQL skills.
  • Hands-on experience with Apache Spark.
  • Strong experience with Spark SQL.
  • Strong experience developing data solutions using PySpark.
  • Experience building data solutions within Microsoft Azure.
  • Experience developing automated workflows for data ingestion, transformation, and delivery.
  • Experience processing and transforming large and complex datasets.
  • Strong understanding of data integration, ETL/ELT, and data pipeline architecture.
  • Ability to troubleshoot and optimize data pipelines and processing workloads.
  • Strong understanding of data quality and validation practices.
  • Strong analytical and problem-solving skills.
  • Ability to work independently while collaborating effectively with cross-functional technical teams.
Nice-to-haves (optional)
  • Hands-on Microsoft Fabric experience – strongly preferred.
  • Experience building data pipelines or engineering solutions using Microsoft Fabric.
  • Pharmaceutical industry experience.
  • Life Sciences industry experience.
  • Insurance industry experience.
  • Experience with enterprise-scale cloud data platforms and distributed data processing.
  • Experience supporting data solutions used for analytics, BI, reporting, or machine learning.

Similar Jobs

One Month Ago
Remote
United States
156K-180K Annually
Senior level
156K-180K Annually
Senior level
Big Data • Real Estate • Software
The Senior Data Engineer will optimize data ingestion pipelines, manage PostgreSQL databases, design ETL processes, and ensure system reliability in a remote-first, collaborative environment.
Top Skills: AWSAws AuroraCi/CdCloudFormationPostgresRubyRuby On RailsTerraform
24 Days Ago
Remote
United States
Senior level
Senior level
Healthtech
The Senior Data Engineer will design and maintain data pipelines, write production-grade Python and SQL code, and support AI initiatives while collaborating across teams.
Top Skills: AirflowSparkAuroraAws GlueCloudFormationCloudwatchIamLookerMySQLPythonRedshiftSQLTableau
One Month Ago
Remote
USA
160K-195K Annually
Senior level
160K-195K Annually
Senior level
Artificial Intelligence • Machine Learning • Software • Analytics
The Senior Data Engineer will design and build scalable data platforms, handle automation for data operations, and ensure reliability and performance of data pipelines while collaborating with various teams.
Top Skills: AirflowSparkAws EmrAws LambdaAws S3KafkaKinesisPythonScalaSQL

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account