jobloom

JobLoom finds jobs directly from company career sites before many job boards, then routes you into detailed role pages like this one.

management

Posted Jun 23

Director of Engineering, Infrastructure

at Klaviyo

Boston, United StatesOn-site

Responsibilities

  • Lead the definition of platform primitives such as compute runtimes, storage options, and service networking, ensuring they are scalable, secure, and aligned with Klaviyo's standards for operational excellence.
  • Create and disseminate golden paths and decision trees that simplify the technological choices for R&D teams, enhancing consistency and self-sufficiency across engineering efforts.
  • Drive initiatives that enhance the reliability of production systems, focusing on incident prevention, transparent response protocols, and proactive capacity planning.
  • Coordinate with product teams to identify and eliminate infrastructure bottlenecks, aiding in improving the time-to-market for new services and increasing developer satisfaction.
  • Establish frameworks for cost-effective infrastructure management, balancing financial discipline with flexibility and efficiency to maximize value delivery.
  • Mentor and develop high-performing teams, fostering a culture of inclusivity and ownership, while setting clear, impactful goals that align with business priorities.
  • Collaborate with cross-functional partners to manage platform investments, clarify ownership, and safely implement infrastructure changes that drive strategic outcomes.
  • Track and report critical performance metrics, such as system reliability, developer productivity, and infrastructure costs, enabling data-driven decision-making and accountability.
  • Champion operational readiness by establishing robust SLAs and SLIs, ensuring all infrastructure components meet defined performance thresholds conforming to Klaviyo's quality standards.
  • Facilitate a culture of continuous learning and experimentation with AI tools, deploying enhancements that intelligently streamline engineering workflows.
  • Lead a disciplined approach to incident management and postmortems, establishing a blameless culture of learning and innovation to minimize future disruptions. Who You Are
  • Proven track record of optimizing both cost-to-serve and reliability metrics in a scaling SaaS company, driving significant impact on the bottom line.

Requirements

  • Optimize the use of AI to enhance infrastructure management and development processes, pioneering innovative workflows that keep Klaviyo at the forefront of technological advancement.
  • experience in infrastructure, SRE, platform engineering, or security engineering, with at least 5 years managing managers and senior ICs, demonstrating strong leadership and team-building skills.
  • You possess a deep understanding of SRE principles, with proven expertise in designing effective SLOs and SLIs, managing incidents, and ensuring capacity planning and operational continuity. You have
  • experience with LiveSite (the enterprise web platform) and understand the infrastructure requirements, integration patterns, and operational demands that come with supporting enterprise-grade web properties at scale.
  • You have a proven track record of standing up or strengthening LiveSite culture inside engineering orgs.
  • You excel at simplifying complexity and creating logical clarity, using your strong communication skills to drive organizational changes across product, data, and security domains.
  • experience with AI makes you AI-curious, and you're eager to leverage it to drive smarter, more efficient infrastructure operations.
  • You possess a keen ability to balance strategic thinking with tactical execution, ensuring infrastructure solutions not only meet current needs but are future-proof and scalable.
  • You are an org builder, you know how to scale a team thoughtfully, develop high-potential engineers into managers, and create the conditions for leaders to emerge and grow.
  • experience partnering closely with security engineering teams to build platforms that are secure by design and operationally hardened.
  • experience partnering closely with security engineering teams to build platforms that are secure by design and operationally hardened. Nice to Have Previous
  • experience in transforming internal platforms into productized services, complete with SLAs and developer experience-focused roadmaps.
  • Background in architecting data-centric and event-driven systems at a large scale, particularly within a high-growth environment.
  • Familiarity with the challenges and opportunities characteristic of a high-growth SaaS milieu, with a focus on leveraging those to drive platform innovation and efficiency.

Benefits

  • Our salary range reflects the cost of labor across various U.S. geographic markets.
  • The range displayed below reflects the minimum and maximum target salaries for the position across all our US locations.
  • The base salary offered for this position is determined by several factors, including the applicant’s job-related skills, relevant experience, education or training, and work location.
  • In addition to base salary, our total compensation package may include participation in the company’s annual cash bonus plan, variable compensation (OTE) for sales and customer success roles, equity, sign-on payments, and a comprehensive range of health, welfare, and wellbeing
  • Your recruiter can provide more details about the specific salary/OTE range for your preferred location during the hiring process.
  • Base Pay Range For US Locations: $244,000 — $366,000 USD

Contact

  • Want to learn more about life at Klaviyo? Visit klaviyo.com/careers to see how we empower creators to own their own destiny.

Additional details

  • At Klaviyo, we value the unique backgrounds, experiences and perspectives each Klaviyo (we call ourselves Klaviyos) brings to our workplace each and every day.
  • We believe everyone deserves a fair shot at success and appreciate the experiences each person brings beyond the traditional job requirements.
  • If you’re a close but not exact match with the description, we hope you’ll still consider applying.
  • Klaviyo’s mission is to empower businesses to independently drive their growth, and the Engineering Department's contribution to this mission is crucial.
  • As Director of Production Infrastructure, you'll helm the creation and management of a high-performance platform designed to support the rapid innovation demanded by our R&D teams.
  • This role is all about defining the pillars of our infrastructure, compute, storage, networking, observability, and setting a robust set of principles that guide their use.
  • In this position, you'll be entrusted with the responsibility of developing and maintaining platform primitives that empower our engineering teams to bring ideas to life seamlessly.
  • Collaborating with industry leaders across engineering, security, and finance, your decisions will shape the infrastructure blueprint that underpins our scalable, secure, and cost-effective operations.
  • As a leader, your mission is to foster a culture of ownership, innovation, and productivity while steering teams toward achieving critical reliability and performance metrics.
  • Your role will span across defining clear service contracts, instituting capacity plans, and honing our developer enablement strategy to reduce friction and enhance developer velocity.

Find more real-time jobs on JobLoom.