
Founding Infrastructure Engineer
- $120k – $180k • 0.5% – 2.0%
- |
- |5 years of exp
- |Full Time
In office
Not Available
About the job
About Influxion
Influxion is building an AI engineer: an autonomous teammate that continuously monitors, debugs, and improves AI agents in production.
AI agents are performing increasingly complex work, creating a growing operational burden for teams. Understanding what an agent did, investigating failures, uncovering unexpected behavior, and improving quality requires reasoning across thousands of steps, tool calls, model decisions, traces, evaluations, logs, and code. As agents grow more capable and organizations rapidly increase their deployment, it's becoming impossible for humans to manage using traditional tools and techniques.
Our thesis is that AI must manage AI. Influxion combines state-of-the-art AI with systems and performance engineering to deploy teams of specialized agents across the production AI stack. These agents work together to process and reason over massive volumes of semantic, qualitative data, automating complex AI engineering work at scale.
Influxion is a venture-backed startup based in San Francisco at an eight-figure pre-seed valuation, with core technology and a working product that now needs to scale.
About the role
We are now looking for an early Founding Engineer focused on infrastructure and platform to help take the product to production and scale.
This person will have broad ownership across cloud architecture, deployment, reliability, security, data infrastructure, DevOps, and the foundations required to run complex agentic systems safely at enterprise scale.
This is an opportunity to join at an early stage, work directly with the founders, and help shape both the product and the company that will become essential as AI agents proliferate.
Please reach out if this sounds like you—or someone exceptional you know.
Possible Week 0 Projects
- Productionize authentication and core database infrastructure
- Establish initial cloud deployments with observability, security, backup, and reliability mechanisms
- Build infrastructure for tracking AI model and agent usage and costs
- Establish initial deployment, testing, and release practices
- Begin SOC 2 readiness scoping and control implementation
First 30 days
- Improve agent deployment reliability, tenant isolation, and failure recovery
- Take ownership of deployment, testing, release, and upgrade workflows
- Strengthen observability, monitoring, backup, and incident-response practices
- Develop new product and infrastructure integrations
- Design deployment architectures for initial enterprise environments
- Help complete SOC 2 Type 1 readiness and begin the Type 2 effort
First 90 days
- Partner directly with early customers to deploy and operate enterprise pilots
- Own the reliability, security, scalability, and operational health of the production platform
- Establish a repeatable architecture for cloud and customer-hosted deployments
- Automate provisioning, upgrades, migrations, rollback, backup, and recovery
- Introduce platform safeguards for tenant isolation, agent resource usage, costs, and failure handling
- Establish the operational standards, tooling, and platform roadmap needed to support additional customers and engineers
- Operationalize SOC 2 controls and evidence collection as part of normal engineering workflows
We're Looking for someone who
- Has prior startup experience owning design and implementation of production cloud systems
- Has strong infrastructure and backend engineering experience, with the ability to contribute across the stack
- Is independent, moves fast, and is adaptable to changing priorities and information
- Thrives in ambiguous, fast-changing environments
- Is constantly thinking, evaluating, prioritizing, and is willing to push back and offer alternatives
- Tolerates grunt work necessary to improve the product and customer experience
- Has experience with and wants to work directly with customers/users
- Is familiar with SOC2 and other security frameworks
- Enjoy operating with high ownership and limited process
Technical Experience
- Cloud infrastructure, Infrastructure as Code, deployment automation, and DevOps
- Authentication and authorization systems
- Postgres database architecture and operations
- Redis and distributed systems
- Python and TypeScript
- React and full-stack product development
- Observability/monitoring, security, reliability, backup and recovery
- Incident response and production operations
- Test infrastructure and release automation
- Strong AI-assisted coding skills - using it as a productivity multiplier, not a substitute (know its limitations and when not to use it)
- Experience deploying (not necessarily developing) AI agents
You do not need to have worked with every technology in our stack. We care more about strong fundamentals, ownership, learning speed, and demonstrated experience operating production systems.
Interview Process
- 30 min call with each founder
- Technical screening
- 3–5 day work trial
- Offer
About the company

Influxion
Similar Jobs









