management
Posted Jun 23Director of Engineering, Infrastructure
at Klaviyo
Boston, United StatesOn-site
Responsibilities
- Lead the definition of platform primitives such as compute runtimes, storage options, and service networking, ensuring they are scalable, secure, and aligned with Klaviyo's standards for operational excellence.
- Create and disseminate golden paths and decision trees that simplify the technological choices for R&D teams, enhancing consistency and self-sufficiency across engineering efforts.
- Drive initiatives that enhance the reliability of production systems, focusing on incident prevention, transparent response protocols, and proactive capacity planning.
- Coordinate with product teams to identify and eliminate infrastructure bottlenecks, aiding in improving the time-to-market for new services and increasing developer satisfaction.
- Establish frameworks for cost-effective infrastructure management, balancing financial discipline with flexibility and efficiency to maximize value delivery.
- Mentor and develop high-performing teams, fostering a culture of inclusivity and ownership, while setting clear, impactful goals that align with business priorities.
- Collaborate with cross-functional partners to manage platform investments, clarify ownership, and safely implement infrastructure changes that drive strategic outcomes.
- Track and report critical performance metrics, such as system reliability, developer productivity, and infrastructure costs, enabling data-driven decision-making and accountability.
- Champion operational readiness by establishing robust SLAs and SLIs, ensuring all infrastructure components meet defined performance thresholds conforming to Klaviyo's quality standards.
- Facilitate a culture of continuous learning and experimentation with AI tools, deploying enhancements that intelligently streamline engineering workflows.
- Lead a disciplined approach to incident management and postmortems, establishing a blameless culture of learning and innovation to minimize future disruptions. Who You Are
- Proven track record of optimizing both cost-to-serve and reliability metrics in a scaling SaaS company, driving significant impact on the bottom line.
Requirements
- Optimize the use of AI to enhance infrastructure management and development processes, pioneering innovative workflows that keep Klaviyo at the forefront of technological advancement.
- experience in infrastructure, SRE, platform engineering, or security engineering, with at least 5 years managing managers and senior ICs, demonstrating strong leadership and team-building skills.
- You possess a deep understanding of SRE principles, with proven expertise in designing effective SLOs and SLIs, managing incidents, and ensuring capacity planning and operational continuity. You have
- experience with LiveSite (the enterprise web platform) and understand the infrastructure requirements, integration patterns, and operational demands that come with supporting enterprise-grade web properties at scale.
- You have a proven track record of standing up or strengthening LiveSite culture inside engineering orgs.
- You excel at simplifying complexity and creating logical clarity, using your strong communication skills to drive organizational changes across product, data, and security domains.
- experience with AI makes you AI-curious, and you're eager to leverage it to drive smarter, more efficient infrastructure operations.
- You possess a keen ability to balance strategic thinking with tactical execution, ensuring infrastructure solutions not only meet current needs but are future-proof and scalable.
- You are an org builder, you know how to scale a team thoughtfully, develop high-potential engineers into managers, and create the conditions for leaders to emerge and grow.
- experience partnering closely with security engineering teams to build platforms that are secure by design and operationally hardened.
- experience partnering closely with security engineering teams to build platforms that are secure by design and operationally hardened. Nice to Have Previous
- experience in transforming internal platforms into productized services, complete with SLAs and developer experience-focused roadmaps.
- Background in architecting data-centric and event-driven systems at a large scale, particularly within a high-growth environment.
- Familiarity with the challenges and opportunities characteristic of a high-growth SaaS milieu, with a focus on leveraging those to drive platform innovation and efficiency.
Benefits
- Our salary range reflects the cost of labor across various U.S. geographic markets.
- The range displayed below reflects the minimum and maximum target salaries for the position across all our US locations.
- The base salary offered for this position is determined by several factors, including the applicant’s job-related skills, relevant experience, education or training, and work location.
- In addition to base salary, our total compensation package may include participation in the company’s annual cash bonus plan, variable compensation (OTE) for sales and customer success roles, equity, sign-on payments, and a comprehensive range of health, welfare, and wellbeing
- Your recruiter can provide more details about the specific salary/OTE range for your preferred location during the hiring process.
- Base Pay Range For US Locations: $244,000 — $366,000 USD
Contact
- Want to learn more about life at Klaviyo? Visit klaviyo.com/careers to see how we empower creators to own their own destiny.
Additional details
- At Klaviyo, we value the unique backgrounds, experiences and perspectives each Klaviyo (we call ourselves Klaviyos) brings to our workplace each and every day.
- We believe everyone deserves a fair shot at success and appreciate the experiences each person brings beyond the traditional job requirements.
- If you’re a close but not exact match with the description, we hope you’ll still consider applying.
- Klaviyo’s mission is to empower businesses to independently drive their growth, and the Engineering Department's contribution to this mission is crucial.
- As Director of Production Infrastructure, you'll helm the creation and management of a high-performance platform designed to support the rapid innovation demanded by our R&D teams.
- This role is all about defining the pillars of our infrastructure, compute, storage, networking, observability, and setting a robust set of principles that guide their use.
- In this position, you'll be entrusted with the responsibility of developing and maintaining platform primitives that empower our engineering teams to bring ideas to life seamlessly.
- Collaborating with industry leaders across engineering, security, and finance, your decisions will shape the infrastructure blueprint that underpins our scalable, secure, and cost-effective operations.
- As a leader, your mission is to foster a culture of ownership, innovation, and productivity while steering teams toward achieving critical reliability and performance metrics.
- Your role will span across defining clear service contracts, instituting capacity plans, and honing our developer enablement strategy to reduce friction and enhance developer velocity.