WISEcode Logo

WISEcode

Software Developer & Data Engineer

Posted Yesterday
Remote
Hiring Remotely in United States
Mid level
Remote
Hiring Remotely in United States
Mid level
Build and operate WISEcode’s food data lakehouse using Python, SQL, DuckLake, S3 Parquet, Postgres, and Prefect. Responsibilities include source ingestion, entity resolution, pipeline consolidation, data enrichment, downstream connectors, data quality reporting, configuration discipline, and architectural documentation. The role requires designing and shipping production lakehouse or warehouse systems, managing scalable API-driven pipelines, and collaborating closely with engineers in a fully remote environment.
The summary above was generated by AI
Build the food intelligence layer at WISEcode.Most people have no real way to know what's in their food or what it does to them. The information is scattered, inconsistent, and written to sell rather than to inform. WISEcode exists to change that: to give people a source they can actually trust, and the freedom to decide for themselves once they have it.Getting there takes more than a better app. It takes a system underneath that turns fragmented food data into answers that hold up: consistent, explainable, and the same whether they reach a parent scanning a label, a brand seeking verification, or an institution making decisions at scale. That system is what we are building, and it is the reason the work compounds instead of resetting with every new product.

The Role

    We are hiring a Software Developer & Engineer to work as a senior individual contributor, serving as the second engineer on The Pantry, our new data lakehouse (DuckLake on S3 Parquet, Postgres catalog, Prefect orchestration, Python throughout). You will work alongside our Data team, combining hands-on implementation with platform and data architecture design in an early-stage environment where the shape of the work is still yours to define.

    You bring 3+ years of experience in a similar role, along with real comfort with ambiguity and pace. What you get in return is a problem that matters and the room to shape how it gets solved.

Key Responsibilities

    • Lakehouse Architecture & Implementation: Own end-to-end implementation and collaborate on data/platform architecture alongside our data team.
    • Source Ingestion: Build and establish patterns for a variety of sources including APIs, text from public sources, and external databases.
    • Identity & Entity Resolution: Merge the many records we collect for each product, across retailers and dates, into one trustworthy published record. Decide which source wins on each disagreement, and keep every original record so each published value can be traced to its source.
    • Pipeline Consolidation & Enrichment: Consolidate existing food enrichment pipelines down to a single pipeline in collaboration with our engineering team.
    • Downstream Connectors: Build robust connectors to feed collapsed food data to WISE Intel and the consumer app.
    • Operational Discipline & Configuration: Maintain strict fail-fast config discipline (magic values in config files; early loud failures for missing settings).
    • Data Quality & Reporting: Establish ingest capabilities and quality reporting for Data Assembly readiness.
    • Technical Documentation: Write plain, short, dated decision records prior to code implementation.

Technical Skills

  • Languages: Native proficiency in Python and SQL.
  • Storage & Query Engines: DuckDB, DuckLake, S3 Parquet.
  • Databases: Postgres, DynamoDB (state management).
  • Orchestration & Frameworks: Prefect, Kubernetes/Fargate job runners; familiarity with dbt, Dagster, or Airflow.
  • AWS Infrastructure: S3, ECS, Lambda, EventBridge, SNS, Terraform, IAM.
  • Reporting & Data Modeling: Lakehouse, analytic/dimension design; Quarto reporting; Medallion architecture (bronze/clean), star schema (food sighting, ingredient, nutrient, provenance).
  • CI/CD: github actions

Minimum Qualifications

    • 3+ years of experience in a similar role 
    • Proven track record of designing and shipping a lakehouse or data warehouse from scratch, and operating it for at least one year in production.
    • Demonstrated experience treating data identity and entity resolution as first-class architectural questions.
    • Experience running batch data pipelines that invoke paid cloud/third-party APIs at scale with strict cost efficiency and budget controls.
    • Experience pairing daily with other engineers in a fast-paced environment.

Preferred Qualifications

    • Experience making architectural trade-offs between lightweight/embedded query engines (e.g., DuckDB) versus distributed systems (Spark, Databricks, Snowflake).
    • Familiarity with data acquisition methods (e.g., web scraping, OCR API integration).
    • Note: Food or nutrition domain knowledge, deep Spark/Databricks/Snowflake experience, and people management experience are explicitly NOT required for this role.

Compensation

    Competitive base, bonus, and meaningful equity. Details shared in the hiring process.

Benefits

    • Location: Fully remote 

    • Unlimited PTO & Paid Holidays: Unlimited time off and 10 paid holidays

    • Financial Wellness: 401(k) with a 4% company match, immediately vested

    • Health Benefits: Comprehensive medical, dental, vision, life, and additional ancillary coverage. 

      • 100% coverage for employees, with 80% covered for dependents 

      • 100% employer paid STD, LTD, Identity Theft, and Life insurance 

      • FSA and HSA plans available, for our HSA:

        • $200/mo contributions for Employee plans

        • $400/mo contributions for Employee + Dependent plans

Why WISEcode

    We are building the layer that food decisions will run on. That is quiet, compounding work: broadening coverage, sharpening precision, and making every output repeatable and defensible. We value transparency, rigor, and respect for the person making the decision, who deserves clear information and the freedom to reach their own conclusion. Come in early enough and you help shape how this system works instead of operating inside one someone else already designed.

     

Research shows that some groups hesitate to apply unless they meet every qualification. If you’re excited about this role but don’t check every box, we encourage you to apply. At WISEcode, we value diverse experiences, transferable skills, and the unique strengths each person brings.

WISEcode is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. 

Similar Jobs

6 Days Ago
Remote or Hybrid
OH, USA
Senior level
Senior level
Financial Services
Build and operate scalable Databricks-on-AWS data pipelines using PySpark, Delta Lake, and lakehouse patterns. Optimize performance, implement data quality, monitoring, alerting, and automated remediation, and deliver curated datasets for BI and analytics partners. Collaborate with stakeholders on architecture and design while applying secure software engineering, CI/CD, agile, and operational stability practices. The role also uses AI-assisted development tools and supports workforce data analytics.
Top Skills: AlteryxAmazon AthenaAmazon EmrAmazon S3Apache AirflowApache IcebergSparkAutosysAWSAws CloudwatchAws GlueAws LambdaBitbucketClaudeDatabricksDatabricks WorkflowsDelta LakeDelta Live TablesGitGithub CopilotJavaJenkinsOracleParquetPysparkPythonScalaSigmaSpinnakerSQLTableau
15 Days Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
232K-348K Annually
Senior level
232K-348K Annually
Senior level
Artificial Intelligence • Cloud • Software
Lead the architecture and development of Vercel’s next-generation data platform, supporting batch and real-time integrations, analytics, data warehousing, and AI/ML workloads. Design scalable systems using Kafka, ClickHouse, Tinybird, and Snowflake; establish data governance and security standards; guide architectural decisions and roadmaps; collaborate with engineering, product, security, compliance, and leadership teams; write production code; and mentor engineers.
Top Skills: AWSAzureBig Data FrameworksClickhouseConfluent PlatformData GovernanceData WarehousingETLGCPKafkaKafka StreamsSnowflakeTinybird
21 Days Ago
Easy Apply
In-Office or Remote
IN, USA
Easy Apply
165K-221K Annually
Senior level
165K-221K Annually
Senior level
Healthtech • Information Technology • Mobile • Productivity • Software • Analytics • Telehealth
Design and build maintainable data pipelines, ETL processes, and data architecture standards. Collaborate with product managers, analysts, and machine learning engineers to translate requirements into automated, scalable solutions. Apply software engineering best practices, asynchronous systems, APIs, streaming, testing, monitoring, and Python and SQL expertise. Help prioritize data initiatives based on business impact and communicate complex technical concepts to technical and non-technical stakeholders.
Top Skills: APIsArtificial IntelligenceETLGitMachine LearningMultithreadingPythonSQLStreaming

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account