Summary
β¨ AIβGenerated
A fast-growing technology organization is seeking a Software Development Team Lead to own the cloud infrastructure platform supporting its products. The role is split approximately 50/50 between hands-on engineering and people leadership, covering platform reliability, security, cost efficiency, production operations, and team development.
Highlights
Lead an infrastructure platform team while remaining hands-on, with a balanced focus on technical contribution and people leadership. Build secure, reliable, and cost-efficient cloud foundations in an autonomous environment.
Description
About Us
Spare is a fast-growing, successful startup.
We thrive on innovation, rapid execution, and delivering top-tier products that make a real impact in the transit industry.
We are now seeking a Software Development Team Lead to join our Infrastructure Platform Team and help build and operate the cloud platform that every Spare product runs on.
About This Role
As leader of the team responsible for Spare's core infrastructure platform, you will play a critical role in designing and building the reliable, secure, and cost-efficient foundation that powers our customers' on-demand transit systems.
Working closely with others on the team, you will own the correct and efficient operation of the platform across a variety of real-world production scenarios.
In this role, you'll split your time 50/50 between hands-on technical contribution and people leadership, working in an autonomous environment where you'll solve interesting technical challenges while growing and mentoring a high-performing team.
Currently there are four software developers reporting to this role.
Given the nature of our business, this role requires someone who can balance technical expertise with strong product sensibilities, creating solutions that are both technically sound and accessible to end users.
The role includes some travel as part of the job responsibilities β specifically, up to four customer site visits per year to gain firsthand insights, plus participation in our biannual software development hackathons in Vancouver.
Key Responsibilities
Own the design and development of core infrastructure platform capabilities from inception to launchBuild and evolve tooling, automation, and platform services that make every engineering team at Spare faster and saferArchitect and implement high-performance, scalable distributed systems on GCP and KubernetesDrive improvements in cluster reliability, application resilience, and internal access securityOperate and maintain Spare's Redis and PostgreSQL databases β availability, performance tuning, scaling, upgrades, backups, and disaster recoveryDrive Spare toward AI SRE practices β embed AI agents into incident detection, alert triage, and operational workflows to make reliability work faster and more proactiveManage and continuously improve the SRE on-call rotation β healthy schedules, clear escalation paths, and blameless post-mortems that turn incidents into systemic improvementsDrive FinOps practices across the organization β own cloud spend visibility, right-sizing, and cost optimization initiatives that deliver measurable savingsUse AI agentic tooling daily to accelerate your own and your team's engineering output, and coach the team to do the sameActively mentor software developers of all levels and uplift team capacityCollaborate cross-functionally with product managers, designers, and other software developersEnsure 99.99% uptime and maintain exceptional system performanceParticipate in team agile rituals and help improve software development processes
Who You Are
A highly productive software developer with a proven track record of delivering high-quality code in complex environmentsHighly proficient with modern AI development tools and agentic workflows β you use AI to move faster and think better, not as a crutchA passionate mentor and technical leader who enjoys helping others growPassionate about distributed systems, cloud infrastructure, and platform engineeringCost-conscious by default β you treat cloud spend as an engineering problem and know how to deliver savings without sacrificing reliabilityAdept at balancing reliability and security rigor with practical developer-experience constraints
Requirements
7+ years of software development experience, with at least 2+ years in a people leadership roleExpert in backend technologies with strong distributed systems experienceDemonstrated proficiency with AI-assisted and agentic development workflows (AI coding agents, automation of engineering and operational tasks)Experience operating systems at scale with a strong reliability and uptime mindsetExperience running or managing an SRE on-call rotation, including incident response and post-mortem cultureDeep experience with cloud infrastructure (GCP) and container orchestration (Kubernetes)Experience driving cloud cost optimization and FinOps initiatives β right-sizing, spend visibility, and measurable cost savingsExperience with infrastructure-as-code and configuration management tooling (Terraform)Understanding of security best practices, especially around internal access control and container securityDemonstrated success in managing a team of software developers, with a focus on team and individual performanceDemonstrated ability to mentor other developers and provide technical leadershipStrong problem-solving, debugging, and system design skillsExcellent communication and collaboration skills
It Will Be Considered a Plus (nice-to-have)
Experience in the transit industry or another real-time, safety-critical domainExperience building internal developer platforms and golden-path toolingExperience applying AI/LLM tooling to SRE or infrastructure operations (AIOps, automated incident triage)Experience with CI/CD systems at scale, including test sharding and build performanceExperience operating and maintaining PostgreSQL and Redis in production β performance tuning, indexing, replication, backup and restore
Don't meet every single requirement?
Studies have shown that women and people of colour are less likely to apply to jobs unless they meet every single qualification in the job posting.
At Spare, we are committed to creating a diverse and inclusive environment so we strongly encourage you to apply even if you don't believe you meet every single qualification outlined.
We also do our best to respond to all applications we receive.
About The Infrastructure Platform Team
The Infrastructure Platform Team's job is to build and run the foundation that Spare's on-demand transportation platform depends on β where circumstances change in real time and downtime is not an option.
Primarily this involves keeping our Kubernetes clusters reliable and secure, hardening our application resilience, driving cloud cost efficiency, and making sure our distributed systems can handle real-time updates efficiently.
Why Join Us?
Work on challenging technical problems with real-world impact in the transportation industryA fast-paced, high-impact role in a rapidly growing startupThe opportunity to take ownership of core systems and drive innovationA dynamic, collaborative, and supportive team cultureCompetitive salary and equity options
If you are ready to tackle complex routing and optimization challenges and thrive in a high-performance environment, we'd love to hear from you!
Compensation Range: CA$150K - CA$260K