
- B2B
- Early StageStartup in initial stages
Senior Software Engineer - Data
- $140k – $180k • No equity
- |Remote ()
- |4 years of exp
- |Full Time
Remote only
Not Available
About the job
Senior Software Engineer - Data
Location: Fully Remote (US-based)
Reports to: Lead Data Engineer
The Opportunity
Most startups burn cash to find a market, at Symmetric, we've spent the last few years winning it. We've been profitable from Year 1, remain employee-owned and doubled in 2025 with line of sight to repeat that growth in 2026.
Our product is the data engine behind the world's most sophisticated healthcare systems — 17 of the Gartner Top 25 supply chains already rely on Symmetric. We provide cutting-edge solutions that streamline procurement, item master data management, and improve efficiency across healthcare systems.
We have a fully remote team that's making a big impact on the healthcare supply chain by working closely with enterprises (health systems, hospitals, consulting firms, etc.), regulatory agencies and industry groups.
Role Overview
We are seeking a Senior Data Engineer to design, build, and optimize the data pipelines and platforms that power Symmetric's core product. You will work with large, complex, and often messy healthcare supply chain datasets — combining classical ETL/ELT with LLM-powered enrichment and ML-driven data quality to transform raw, unstructured source data into reliable, high-quality assets that drive our application, analytics, and customer-facing solutions.
This is a high-impact individual contributor role for someone who thrives on solving hard data problems at the intersection of traditional data engineering and applied AI, building scalable infrastructure, and collaborating closely with engineering, product, and data operations teams.
Key Responsibilities
Data Pipeline Development & Architecture
- Design, build, and maintain scalable ETL/ELT pipelines that ingest, transform, and deliver healthcare supply chain data from both structured and unstructured sources.
- Architect data solutions on AWS, selecting the right tools and patterns for performance, reliability, and cost efficiency.
- Develop and enforce data quality frameworks including validation, reconciliation, and monitoring across all pipeline stages.
LLM & ML Data Processing
- Build and operate LLM-powered pipelines for entity extraction, classification, normalization, and matching of unstructured healthcare supply chain data (e.g., product descriptions, catalogs, invoices, and contracts).
- Design prompting, evaluation, and tuning workflows that turn LLMs into reliable production components — with guardrails, structured outputs, caching, retries, and cost controls.
- Integrate ML models into the data platform — training pipelines, feature stores, batch and online inference, and monitoring for drift, accuracy, and data quality.
- Partner with domain experts to build labeled datasets, evaluation benchmarks, and human-in-the-loop feedback systems that continuously improve model performance.
Collaboration & Leadership
- Partner with Product Management, Application Engineering, and Data Operations to understand requirements and deliver data solutions aligned with the product roadmap.
- Prototype and iterate on tools for data review and validation used by domain experts.
- Mentor junior engineers and contribute to team best practices around code quality, testing, and documentation.
- Participate in code reviews, architectural discussions, and technical planning sessions.
Qualifications
Required
- 4+ years of experience as a software or data engineer, with at least 3+ years focused on data engineering.
- Strong proficiency in Python and SQL for data pipeline development and analysis.
- Hands-on experience architecting and operating data solutions on AWS (S3, Glue, Athena, or equivalent).
- Experience designing and maintaining ETL/ELT pipelines at scale, including orchestration tools (Airflow, Dagster, or similar).
- Solid understanding of data modeling, schema design, and data warehouse best practices.
- DevOps experience: CI/CD, infrastructure-as-code, containerization, and operationalizing production data workloads.
- Hands-on experience integrating LLMs or ML models into production data pipelines — including prompt design, evaluation, structured output handling, and managing cost, latency, and reliability at scale.
- Strong communication skills with the ability to collaborate effectively across technical and non-technical teams.
Preferred
- Experience in healthcare IT, supply chain technology, or data platforms serving enterprise customers.
- Experience with high-performance Rust code for data processing.
- Familiarity with data quality, governance, and master data management challenges.
- Experience with NLP, information extraction, entity resolution, record linkage, or fine-tuning/embedding-based retrieval (RAG) systems.
- Experience with web scraping or large-scale data ingestion from heterogeneous sources.
- Familiarity with MLOps and LLMOps tooling — model/prompt versioning, offline and online evaluation, vector databases, and observability for AI systems.
What Success Looks Like
- 30 days: Solid understanding of our data architecture, pipeline codebase, and key datasets; built relationships with cross-functional partners in Engineering, Product, and Data Operations.
- 60 days: Independently delivering pipeline improvements and new data integrations; contributing to code reviews and architectural discussions.
- 90 days: Shipped a meaningful pipeline or platform improvement that measurably increases data quality, reliability, or throughput; established as a go-to technical resource on the data team.
- 1 year: Driven significant evolution of our data platform; mentored teammates; played a key role in shaping the technical strategy for data engineering at Symmetric.
What We Offer
- Competitive salary with performance-based bonuses
- Comprehensive benefits: 100% covered health, dental, vision, 401K matching, and flex PTO
- Meaningful work with real-world impact on the healthcare supply chain
- Growth potential at a profitable, high-growth, founder-led startup
- Autonomy and ownership — we trust our team members to be the experts in their craft
- US-based remote with scheduled team in-person meetings
Compensation
Base: $140,000–$180,000/year
Bonus: 10–20% of base
About the company

Symmetric Health Solutions
- B2B
- Early StageStartup in initial stages
Employees joined from
Perks
Founders
Similar Jobs


