Avatar for Jumeau Capital
Jumeau Capital
Actively Hiring
Building developer infrastructure in the API economy. Real product. Real revenue. Lean team

Infrastructure engineer

Posted: today• Recruiter recently active
Job Location
Remote Work Policy

In office

Visa Sponsorship

Not Available

RelocationAllowed
Skills
Python
Infrastructure
Linux
Bash
DevOps
AWS/EC2/ELB/S3/DynamoDB
Networking & TCP/IP
Monitoring
AWS Cloud Services
Cloud Infrastructure
Kubernetes
Terraform
Go
SRE
Cloud FinOps

About the job

Roles and Responsibilities

Production operations — Own daily infrastructure health across AWS, EKS, routing, and the proxy fleet. Review failed workloads, resource utilisation, capacity, and service dependencies. Maintain operational runbooks and follow issues through to verified resolution.
• Platform troubleshooting and incident response — Lead triage and service restoration during your on-call coverage and assigned incidents. Trace failures across DNS, TLS, load balancing, nginx, proxies, Kubernetes, databases, caches, and streams. Apply safe mitigations or rollbacks, engage service owners for application fixes, and coordinate recovery updates and post-mortems.
AWS, Kubernetes, and Terraform — Operate EKS, Aurora, DynamoDB, MSK, Kinesis, Elasticache, S3, IAM, and supporting AWS services across regions. Maintain reviewed Terraform changes, networking, ingress, autoscaling, resource limits, and configuration drift controls. Test capacity assumptions before traffic exposes them.
• Observability and reliability — Own Sentry / Coralogix / PagerDuty and the monitoring conventions service teams use. Agree service-level objectives with engineering, close visibility gaps, and make alerts actionable. Service teams own application instrumentation; you ensure infrastructure and critical dependencies are observable.
Backup, recovery, and maintenance — Own infrastructure backup policies and coordinate database and service recovery with their owners. Test restores and failover against agreed recovery-time and data-loss targets. Plan patching, cluster upgrades, certificate renewal, and secrets rotation with rollback procedures.
CI/CD and production changes — Own and improve Jenkins and GitHub Actions pipelines for frontend, backend, and data workloads. Maintain deployment checks, rollback paths, and post-change health verification. Diagnose failed releases with service teams and keep routine deployments independent of manual intervention.
• Security and access — Maintain least-privilege IAM, secrets management, network controls, and production access discipline. Partner with the architect and service owners on security requirements, remediate infrastructure findings, and keep evidence of changes and controls.
FinOps and operational automation — Own cloud and infrastructure vendor cost visibility, allocation, forecasts, and budget tracking with Finance. Maintain tagging and cost dashboards, investigate spend anomalies, and report monthly on spend, variance, cost per API call, and realised savings. Drive rightsizing, autoscaling, storage lifecycle, and vendor optimisation; recommend Savings Plans or reservations against measured demand and agreed approvals. Automate recurring operations and validate AI-assisted changes before production use

What You Bring

• 5–7 years of hands-on DevOps, SRE, or infrastructure engineering experience, with direct responsibility for customer-facing production systems.
• Strong Linux and networking troubleshooting: processes, memory, disk, DNS, TCP/IP, TLS, HTTP, load balancing, and reverse proxies. You can follow a failed request beyond the first unhealthy dashboard.
• Production AWS and Kubernetes experience across compute, networking, storage, IAM, deployments, ingress, autoscaling, and workload troubleshooting.
• Terraform in production: reusable modules, state management, drift detection, and reviewed infrastructure changes. Strong scripting in Python, Bash, or Go.
• Solid scripting in Python, Bash, or Go.
• Hands-on monitoring, logs, metrics, traces, alerting, and service-level objectives. Experience building reliable CI/CD pipelines and rollback procedures.
• Practical backup, restore, and disaster-recovery experience. Comfortable diagnosing database connections, cache failures, queue backlogs, and dependency bottlenecks with service owners.
• Hands-on FinOps: AWS cost analysis, tagging and allocation, budgets, forecasting, anomaly investigation, and optimisation. You can explain rightsizing and commitment trade-offs, distinguish realised savings from estimates, and work with Finance on spend decisions.
• Bonus: experience operating high-volume API platforms, nginx, Kafka, ClickHouse, Snowflake, or Airflow; familiarity with Node.js services and AI-assisted operational tooling.
• You stay calm when the platform is noisy. You restore service, find the cause, and make the next incident less likely

About the company

Jumeau Capital company logo

Jumeau Capital

Actively Hiring
Building developer infrastructure in the API economy. Real product. Real revenue. Lean team1-10 Employees
Company Size
1-10
Company Industries
Developer APIs
Company Industries
E-Commerce Platforms
Learn more about Jumeau Capital image

Founders

John Sutton
Founder
image
View the team image

Similar Jobs

Aikenist company logo
Aikenist
Welcome to the Future of HealthCare!
OneAssure company logo
OneAssure
Insurance Distribution Platform & Technology Service Provider
EaseOps company logo
EaseOps
Salesforce for hospital operations—automating compliance, workflows, and patient services for India