Hexaware Technologies Logo

Hexaware Technologies

Big Data Engineer

Posted 9 Days Ago
Remote
Hiring Remotely in United States
Senior level
Remote
Hiring Remotely in United States
Senior level
Build and maintain ETL and data pipelines using Python, PySpark, and AWS services. Orchestrate workflows, develop event-driven integrations, optimize data storage and queries, process API and JSON data, and implement data quality monitoring. Support CI/CD, production operations, cloud migration, data lake and lakehouse initiatives, and near-real-time systems while collaborating with stakeholders to deliver technical solutions.
The summary above was generated by AI

Build and maintain ETL pipelines using Python and PySpark on AWS Glue and related platforms. 

Orchestrate workflows using AWS Step Functions and Lambda. - Implement messaging and event-driven integrations using SNS and SQS. Design and optimize storage and querying solutions in Amazon Redshift, RDS, Oracle and S3-based architectures. 

Write efficient SQL for transformations, validation, and reporting. - Integrate data from APIs and process structured and semi-structured JSON data. - Implement data quality checks, monitoring, and operational support processes. 

Participate in CI/CD and version control practices for deployment and release management. 

Collaborate with cross-functional teams to translate business requirements into technical solutions. 

8 years of software development experience across the appropriate platform. 

Strong hands-on experience with Python, PySpark, API’s and SQL. 

Experience with ETL/data pipeline development and Orchestration using Step functions / AirFlow .

 Working knowledge of AWS services including Glue, Lambda, Step Functions, Redshift, S3, SNS, and SQS. 

Experience with Athena, EMR, Kinesis, DynamoDB, or RDS. - Good Knowledge on CloudWatch, logging, and production support. 

Understanding of data warehousing, data lakes, Lake House and query optimization. 

Experience with GitLab/Terraform or similar and CI/CD workflows. 

Good understanding of using AI tools like Github Copilot or similar for code productivity .

Exposure to enterprise data lake or cloud migration initiatives. 

Have an eye to solving complex problems, great communication with stakeholders .

Have a good understanding of performance engineering of code pipelines and near real time systems - Good understanding on Agents and MCP".

A Bachelor’s Degree in Information Technology, Computer Science, Business Administration, or a related field is required.

Similar Jobs

16 Days Ago
Easy Apply
Remote or Hybrid
14 Locations
Easy Apply
120K-150K Annually
Mid level
120K-150K Annually
Mid level
Automotive • Big Data • Insurance • Software • Transportation
Designs, builds, and maintains scalable ETL/ELT pipelines and finance data models using Snowflake, AWS, dbt Core, Python, and SQL. Develops Medallion architecture data marts, integrates ERP systems, and partners with stakeholders on financial workflows. Implements data governance, RBAC, automated quality testing, observability, documentation, CI/CD, and secure deployments. Monitors Snowflake costs and optimizes queries while supporting Oracle EBS integrations and the migration to Workday.
Top Skills: AWSCi/CdDbt CoreGitOracle EbsPythonSnowflakeSQLWorkday
2 Days Ago
Remote
United States
Entry level
Entry level
HR Tech • Information Technology • Professional Services • Consulting
Design, build, and maintain scalable big data pipelines, architectures, distributed databases, and storage systems. Use Python and frameworks such as Hadoop, Spark, or Flink for data integration, transformation, and processing. Ensure system performance, availability, security, governance, and data quality while collaborating with cross-functional teams and documenting technical solutions.
Top Skills: AWSAzureData WarehousingETLFlinkGCPHadoopNosql DatabasesPythonRelational DatabasesSpark
15 Days Ago
In-Office or Remote
243K-500K Annually
Expert/Leader
243K-500K Annually
Expert/Leader
Social Media
Lead Pinterest’s technical strategy and roadmap for scalable big data and AI infrastructure. Build frameworks for compute, job, resource, scheduling, and shuffle management across petabyte-scale datasets. Partner with internal customers, provide company-wide technical leadership, and improve data processing reliability, speed, and efficiency.
Top Skills: AWSFlinkGoJavaKubernetesPythonPyTorchRayScalaSparkTensorFlow

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account