Avatar for TrialSpark
TrialSpark
Actively Hiring
Our mission is to bring new treatments to patients faster and more efficiently
  • B2B
  • Scale Stage
    Rapidly increasing operations
  • Top Investors
    This company has received a significant amount of investment from top investors
  • +4

Engineering Manager, Infrastructure

Posted: 2 weeks ago• Recruiter recently active
Job Location
Visa Sponsorship

Not Available

RelocationNot Allowed
Hiring contact
Linhao Zhang
CTO
New York City
image

About the job

About the Position

As the Engineering Manager for Infrastructure at Formation Bio, you will lead the team responsible for the reliability, security, and operational foundation of our engineering organization. You will set technical direction, develop a high-performing team, and create the systems and operating practices that allow engineers to ship quickly and safely. You're accountable for the pace, reliability, and quality of the systems your team builds and operates, and for creating the conditions in which engineers can grow, perform at a high level, and sustain that performance over time.

Formation Bio is an AI-native engineering organization, and Infrastructure has a central role in defining what that means operationally. You and your team will develop the systems, access patterns, guardrails, delivery workflows, and observability needed for engineers and agents to build and operate production software safely. We are looking for a leader with a credible point of view on how infrastructure and software delivery should evolve as AI-assisted and agent-authored code becomes a larger part of our systems. You do not need to have solved every aspect of the agentic SDLC already, but you should be actively exploring the space and capable of turning emerging practices into reliable engineering systems.

You will partner closely with engineering, data science, tech ops, security, and domain experts across the drug development organization, representing Infrastructure and SRE in cross-functional planning and promoting strong reliability practices throughout the company.

This position is ideal for a manager who leads from the front: someone who is passionate about developing engineers, sets an urgent, high bar for the team's pace and quality, and is restless about finding the next thing that's slowing the org down before it becomes a blocker. Your experience with engineering management, technical leadership, and infrastructure operations will be instrumental in driving Formation Bio's mission to bring new treatments to patients faster and more efficiently.

Responsibilities

  • Hire, coach, and develop a team of infrastructure and site reliability engineers, including setting goals, delivering feedback, and managing performance.
  • Set the technical direction and roadmap for infrastructure and SRE, prioritizing work across reliability, security, cost, and developer experience in partnership with engineering leadership. Build operating practices and paved paths that allow the team to meet the pace required by the broader organization without compromising reliability or security.
  • Partner with engineers and data scientists to develop and deploy secure, observable, and reliable software across containerized applications, internal tools, ML pipelines, and AI model training workloads, including software developed with AI-assisted and agentic workflows. Ensure your team designs solutions that fill gaps and sees them through from concept to implementation.
  • Oversee the research, development, and maintenance of infrastructure services spanning multiple cloud providers and operating environments, including compute, load balancers, databases, and secrets management across development, staging, and production, as well as infrastructure that supports validated and GxP-relevant workloads.
  • Ensure the team creates, reviews, maintains, and optimizes IaC, container images, and CI/CD pipelines, and holds a high bar for engineering quality.
  • Review and provide feedback on requirements and design documents as well as operating and maintenance procedures, and ensure documentation and knowledge sharing practices stay strong across the team.
  • Drive knowledge sharing, training, and mentoring on infrastructure and SRE fundamentals throughout the organization.
  • Own the team's support rotation coverage including staffing, escalation paths, and incident retrospectives.
  • Manage team budgeting, resourcing, and reporting for infrastructure initiatives.

About You

  • 7+ years of total experience across hands-on infrastructure/SRE work and people management, including 3+ years directly managing Site Reliability, Infrastructure, or DevOps engineers.
  • A track record of hiring, developing, and retaining strong engineers, with the ability to give direct, actionable feedback and set clear expectations.
  • Strong operational and reliability judgment, including experience overseeing advanced diagnostics, incident response, and root cause analyses. Digital forensics is a plus, but not required.
  • A strong bias for automation and a low tolerance for repetitive manual work, paired with the judgment to know the rare cases when clickops is the right call and to coach your team accordingly.
  • A credible point of view on AI-native engineering and what an agentic SDLC requires operationally, from access and guardrails for agents to how CI/CD and incident response change when agents are writing and shipping code. Able to set a high bar for how AI is used responsibly to accelerate infrastructure development, investigate incidents, perform diagnostics, and make operational improvements.
  • Exceptional collaboration and communication skills to effectively engage and work alongside stakeholders from diverse disciplines and varying levels of technical expertise, including engineering leadership and non-technical partners.
  • Experience managing or coordinating technical projects and programs, including resource budgeting, estimates, tracking, and reporting.
  • Working knowledge of managing workloads in AWS and Snowflake. Experience with Azure, GCP, and/or Vercel is a plus.
  • Familiarity with managing multiple shared COTS and FOSS software applications in a multi-tenant environment.
  • Working knowledge of Docker, GitHub, Kubernetes, Python, Terraform/OpenTofu, and virtual networking sufficient to guide and evaluate your team's technical decisions.
  • Experience supporting regulated or validated workloads is helpful, but not required. We value strong infrastructure leadership, operational judgment, and the ability to learn our regulatory context more than prior pharmaceutical or biotech experience.

Total Compensation Range: $185,500 - $232,000

About the company

TrialSpark company logo

TrialSpark

Actively Hiring
Our mission is to bring new treatments to patients faster and more efficiently51-200 Employees
  • B2B
  • Scale Stage
    Rapidly increasing operations
  • Top Investors
    This company has received a significant amount of investment from top investors
  • Valuation $1B+
    This company has a valuation of $1B or more
  • 4.8
    Highly rated
    TrialSpark is highly rated on Glassdoor, with 4.8 out of 5 stars
  • 4.7
    Work / Life Balance
    Employees rate TrialSpark 4.7/5 on Glassdoor for work / life balance
  • 4.8
    Strong Leadership
    Employees rate TrialSpark 4.8/5 on Glassdoor for faith in leadership
Learn more about TrialSpark image

Perks

Parental leave
We believe your personal life enriches your professional life. We provide 12 weeks of parental leave (plus a flexible return to work) so you can start or grow your family. Plus, you’ll get a cute TrialSpark onesie for your new addition.
Flexible PTO
We have a flexible PTO policy (so you don’t have to cut corners on your dream trip). If you’re under the weather, we actually want you to rest with take-it-as-you-need-it sick leave.
Dog Friendly Office
We...love dogs. Please bring yours to work.
Team Bonding
We like to get outside the office from time to time. Count on frequent team outings that aren't your standard happy hour (think rock climbing, food tours). Plus, you can put your personal values into practice by joining our resource groups.
Annual Retreat
We set aside time each year to go OOO. Spend time at our annual offsite connecting with your colleagues from other offices. Whether you're into kayaking or karaoke, our workplace experience team plans a weekend with activities for everyone.

Founders

Linhao Zhang
CTO
New York City
image
Benjamine Liu
Founder
New York City
image
Kit Dobyns
Founder
New York City
image
View the team image

Similar Jobs

Kinetic Trials company logo
Kinetic Trials
Accelerate development of life changing therapies with the first agentic OS for clinical trials
Phare Health company logo
Phare Health
A lighthouse for healthcare claims to increase transparency and ease of processing clinical billing
TrialSpark company logo
TrialSpark
Our mission is to bring new treatments to patients faster and more efficiently
Brex company logo
Brex
The AI-powered spend platform
Govio.ai company logo
Govio.ai
Building tech-first solutions to help small businesses find & win government contracts
Klaviyo company logo
Klaviyo
Klaviyo is the AI-first CRM built for B2C brands