In office
Not Available
About the job
Housing.Cloud builds software that colleges and universities rely on to run their student housing operations. Our platform supports high-stakes workflows like roommate matching, housing assignments, and room selection—often under intense, time-bound usage where thousands of students are interacting with the system simultaneously.
We are looking for a Senior Platform Engineer to take ownership of the performance, reliability, and operational maturity of our platform. This is a high-leverage role focused on strengthening how our system behaves under real-world conditions—especially during peak usage events.
You will work across infrastructure, application architecture, and release processes to ensure our platform is fast, resilient, and predictable at scale. This includes improving how we test, validate, and ship code, as well as how we simulate and prepare for high-concurrency scenarios.
Our stack is TypeScript end-to-end, including Next.js, React, tRPC, and Postgres.
This role is ideal for an engineer who enjoys diagnosing complex system behavior, building robust systems, and creating the operational discipline that allows teams to ship with confidence.
Responsibilities
Own Platform Performance & Reliability
- Identify, diagnose, and resolve system performance bottlenecks across the stack
- Design and execute load testing and stress testing strategies to simulate real-world usage patterns
- Improve system behavior under high concurrency and peak traffic conditions
- Establish performance benchmarks and ensure the system consistently meets them
Strengthen Release & Testing Processes
- Own and improve the code release lifecycle, from pre-release validation through production rollout
- Define and enforce standards for testing, QA, and release readiness
- Build processes that ensure code is stable, predictable, and production-ready before deployment
- Partner with engineering to raise the bar on code quality and reliability
Build Scalable Platform Infrastructure
- Improve system architecture to better support high-volume, time-sensitive workflows
- Optimize database performance (Postgres), query efficiency, and data access patterns
- Strengthen background jobs, queues, and asynchronous workflows
- Enhance observability (logging, monitoring, alerting) to surface issues early and clearly
Operationalize Engineering Excellence
- Introduce and refine practices around system reliability, performance monitoring, and incident response
- Build tooling and internal frameworks that help engineers ship more reliable
- Drive a culture of measured, testable improvements to system performance and stability
Collaborate Across Engineering
- Work closely with product and engineering teams to align system design with real-world usage patterns
- Provide technical leadership on performance, scalability, and reliability decisions
- Help the team move faster by reducing uncertainty around system behavior
Qualifications
Core Experience
- Strong experience building and operating production systems at scale
- Deep understanding of performance optimization, system bottlenecks, and reliability engineering
- Experience designing and running load testing / stress testing frameworks
- Strong backend and systems experience, ideally in TypeScript / Node.js environments
- Experience with Postgres performance tuning and query optimization
Platform & Systems Thinking
- Experience with distributed systems, asynchronous workflows, queues, and event-driven architectures
- Strong understanding of high-concurrency system design
- Familiarity with observability tools (monitoring, logging, tracing)
Release & Operational Discipline
- Experience owning or improving CI/CD pipelines, release processes, and testing strategies
- Strong opinions on what “production-ready” means—and how to enforce it
Execution
- Ability to work across layers (infrastructure, backend, application) to solve real problems
- Comfortable operating in a fast-moving environment with evolving requirements
- Strong ownership mindset—you see problems through to resolution
Preferred Qualifications
- Experience with Next.js, React, and tRPC in production systems
- Experience in high-traffic, event-driven platforms (marketplaces, ticketing, logistics, etc.)
- Startup or small-team experience where you’ve owned systems end-to-end
- Comfort using modern AI tools (Cursor, Claude, etc.) as part of your workflow
Benefits
- Competitive base salary.
- Opportunity to earn equity options based on performance.
- 20 Days PTO.
- Health, dental, and vision coverage.
- Collaborative and inclusive company culture.
- The chance to make a real impact by helping institutions improve their housing services through cutting-edge technology.
Housing.Cloud is an equal opportunity employer committed to fostering diversity and inclusion in the workplace. We evaluate qualified applicants without regard to race, color, religion, sex, sexual orientation, disability, veteran status, or other protected characteristics. We are dedicated to providing reasonable accommodations to applicants and employees with disabilities to ensure they can fully participate in our hiring process and perform their job duties. Our commitment extends to creating an accessible and supportive environment where everyone can thrive.
Applicants must be eligible to work in the United States and must reside within the United States to be considered for this position.
About the company
Similar Jobs









