Avatar for Influxion
Influxion
Actively Hiring
Automating reliability and performance engineering for production AI agents

Founding Infrastructure Engineer

  • $120k – $180k • 0.5% – 2.0%
  • |
  • |5 years of exp
  • |Full Time
Posted: 1 month ago• Recruiter recently active
Job Location
Remote Work Policy

In office

Visa Sponsorship

Not Available

RelocationNot Allowed
Skills
Python
Cloud Computing
Redis
Security
Performance Monitoring
TypeScript
Docker / Docker Compose / Kubernetes
postgres
AWS/GCP/Azure
Observability

About the job

About Influxion

Influxion is building an AI engineer: an autonomous teammate that continuously monitors, debugs, and improves AI agents in production.

AI agents are performing increasingly complex work, creating a growing operational burden for teams. Understanding what an agent did, investigating failures, uncovering unexpected behavior, and improving quality requires reasoning across thousands of steps, tool calls, model decisions, traces, evaluations, logs, and code. As agents grow more capable and organizations rapidly increase their deployment, it's becoming impossible for humans to manage using traditional tools and techniques.

Our thesis is that AI must manage AI. Influxion combines state-of-the-art AI with systems and performance engineering to deploy teams of specialized agents across the production AI stack. These agents work together to process and reason over massive volumes of semantic, qualitative data, automating complex AI engineering work at scale.

Influxion is a venture-backed startup based in San Francisco at an eight-figure pre-seed valuation, with core technology and a working product that now needs to scale.

About the role

We are now looking for an early Founding Engineer focused on infrastructure and platform to help take the product to production and scale.

This person will have broad ownership across cloud architecture, deployment, reliability, security, data infrastructure, DevOps, and the foundations required to run complex agentic systems safely at enterprise scale.

This is an opportunity to join at an early stage, work directly with the founders, and help shape both the product and the company that will become essential as AI agents proliferate.

Please reach out if this sounds like you—or someone exceptional you know.

Possible Week 0 Projects

  • Productionize authentication and core database infrastructure
  • Establish initial cloud deployments with observability, security, backup, and reliability mechanisms
  • Build infrastructure for tracking AI model and agent usage and costs
  • Establish initial deployment, testing, and release practices
  • Begin SOC 2 readiness scoping and control implementation

First 30 days

  • Improve agent deployment reliability, tenant isolation, and failure recovery
  • Take ownership of deployment, testing, release, and upgrade workflows
  • Strengthen observability, monitoring, backup, and incident-response practices
  • Develop new product and infrastructure integrations
  • Design deployment architectures for initial enterprise environments
  • Help complete SOC 2 Type 1 readiness and begin the Type 2 effort

First 90 days

  • Partner directly with early customers to deploy and operate enterprise pilots
  • Own the reliability, security, scalability, and operational health of the production platform
  • Establish a repeatable architecture for cloud and customer-hosted deployments
  • Automate provisioning, upgrades, migrations, rollback, backup, and recovery
  • Introduce platform safeguards for tenant isolation, agent resource usage, costs, and failure handling
  • Establish the operational standards, tooling, and platform roadmap needed to support additional customers and engineers
  • Operationalize SOC 2 controls and evidence collection as part of normal engineering workflows

We're Looking for someone who

  • Has prior startup experience owning design and implementation of production cloud systems
  • Has strong infrastructure and backend engineering experience, with the ability to contribute across the stack
  • Is independent, moves fast, and is adaptable to changing priorities and information
  • Thrives in ambiguous, fast-changing environments
  • Is constantly thinking, evaluating, prioritizing, and is willing to push back and offer alternatives
  • Tolerates grunt work necessary to improve the product and customer experience
  • Has experience with and wants to work directly with customers/users
  • Is familiar with SOC2 and other security frameworks
  • Enjoy operating with high ownership and limited process

Technical Experience

  • Cloud infrastructure, Infrastructure as Code, deployment automation, and DevOps
  • Authentication and authorization systems
  • Postgres database architecture and operations
  • Redis and distributed systems
  • Python and TypeScript
  • React and full-stack product development
  • Observability/monitoring, security, reliability, backup and recovery
  • Incident response and production operations
  • Test infrastructure and release automation
  • Strong AI-assisted coding skills - using it as a productivity multiplier, not a substitute (know its limitations and when not to use it)
  • Experience deploying (not necessarily developing) AI agents

You do not need to have worked with every technology in our stack. We care more about strong fundamentals, ownership, learning speed, and demonstrated experience operating production systems.

Interview Process

  • 30 min call with each founder
  • Technical screening
  • 3–5 day work trial
  • Offer

About the company

Influxion company logo

Influxion

Actively Hiring
Automating reliability and performance engineering for production AI agents1-10 Employees
Learn more about Influxion image

Founders

David Kim
Founder
image
View the team image

Similar Jobs

Kick Health company logo
Kick Health
The Online Performance Medicine Clinic for Energizing Sleep and Confident Presentations
Crema Social company logo
Crema Social
Fast Growting International Dating through a Social Meal Experiment