Director, Infrastructure & Site Reliability Engineering

Other Jobs To Apply

No other job posts for this day.

Meet the Moment with Alteryx

 

We're living through a once-in-a-generation shift in how work gets done. Data, automation, and AI are quickly becoming the center of every business decision - and Alteryx is leading the transformation.

 

You'll be working on the challenges that sit at the heart of modern business. No matter your role, the work you do will help organizations move faster, see more clearly, and tackle questions that used to feel impossible.

 

If you're ready to meet the moment with innovation, curiosity, and excellence, there's a place for you here.

Alteryx is searching for a Director, Infrastructure & Site Reliability Engineering. This position is remote-friendly.

Position Overview:

We are seeking an experienced engineering leader to lead Alteryx’s Infrastructure, Site Reliability Engineering (SRE), Observability, and Performance Engineering organizations. In this role, you will define the technical vision and execution strategy for the platforms and operational capabilities that power our cloud services, enabling engineering teams to build, deploy, and operate reliable, secure, and highly scalable products.

You will lead multiple engineering teams responsible for cloud infrastructure, reliability engineering, observability, performance optimization, and operational excellence. This leader will partner closely with Product Engineering, Security, Compliance, and Customer Operations to ensure our platform meets the highest standards for availability, scalability, security, and customer experience.

Primary Responsibilities:

  • Define and execute the strategy for Alteryx’s centralized Infrastructure, Site Reliability Engineering (SRE), Observability, and Performance Engineering organizations.

  • Lead the design, operation, and continuous evolution of cloud infrastructure across AWS and GCP, ensuring scalability, reliability, security, and cost efficiency.

  • Drive Infrastructure-as-Code adoption and governance through Terraform, establishing consistent platform standards, automation, and operational best practices.

  • Own the company’s observability strategy by building and operating enterprise-grade telemetry platforms using Datadog and related technologies, enabling actionable insights into system health, performance, and customer experience.

  • Partner with Security, Compliance, and Engineering teams to meet regulatory and customer requirements, including HIPAA, FedRAMP, SOC 2, and other compliance frameworks.

  • Establish and continuously improve incident management practices, including operational readiness, on-call excellence, postmortem culture, root cause analysis, and measurable reliability improvements.

  • Develop proactive reliability programs including capacity planning, resiliency testing, disaster recovery, performance benchmarking, and operational risk management.

  • Define reliability engineering frameworks that enable product teams to own service health through Service Level Objectives (SLOs), Service Level Indicators (SLIs), error budgets, performance objectives, and operational accountability.

  • Lead the evolution of centralized platform capabilities that simplify how engineering teams build, deploy, monitor, and operate services at scale.

  • Partner with engineering leadership to improve developer productivity through platform automation, self-service infrastructure, deployment tooling, and operational best practices.

  • Build, mentor, and develop high-performing engineering managers and technical leaders while fostering a culture of operational excellence, customer focus, accountability, continuous learning, and innovation.

Qualifications:

  • 10+ years of software engineering, infrastructure, or platform engineering experience, with 5+ years leading multiple engineering teams or managers.

  • Proven experience leading Infrastructure, SRE, Platform Engineering, or Cloud Operations organizations supporting large-scale SaaS products.

  • Deep expertise operating production environments on AWS and/or GCP.

  • Strong experience with Infrastructure-as-Code technologies such as Terraform.

  • Experience building and operating modern observability platforms using Datadog, OpenTelemetry, Prometheus, Grafana, or similar technologies.

  • Demonstrated success implementing SRE practices including SLOs, SLIs, error budgets, incident management, operational reviews, and reliability engineering programs.

  • Experience supporting regulated environments and working with compliance frameworks such as HIPAA, FedRAMP, SOC 2, ISO 27001, or similar.

  • Strong understanding of distributed systems, cloud networking, Kubernetes, container orchestration, CI/CD pipelines, and production operations.

  • Proven ability to influence technical strategy and drive alignment across engineering, security, product, and executive stakeholders.

  • Excellent communication skills with the ability to translate technical strategy into business outcomes.

  • Passion for building high-performing teams and developing engineering leaders.

Valued Skills:

  • Experience leading platform transformations for enterprise SaaS organizations.

  • Familiarity with software performance engineering, load testing, and large-scale distributed systems optimization.

  • Experience supporting data-intensive cloud services

Compensation:

Alteryx is committed to fair, equitable, and transparent compensation. Final compensation is determined by several factors, including but not limited to relevant work experience, education, certifications, skills, and geographic location.

The salary range for this role in the United States is $181,900 - $239,610.

Bonus payouts are based on individual and company performance.

In addition to base pay and bonus eligibility, this role includes clear forms of additional compensation, such as:

  • A monthly Connectivity Plus stipend of $150 to support remote work-related expenses

  • An annual $200 home office reimbursement

Alteryx offers a comprehensive benefits package designed to support your health, financial security, and overall well-being, including:

  • Medical, dental, and vision coverage

  • 401(k) with company match

  • Paid parental leave, caregiver leave, and flexible time off

  • Mental health support and wellness reimbursement

  • Career development and education assistance

Interested? Learn more and apply today at alteryx.com/careers!

#LI-EM1

#LI-REMOTE

Find yourself checking a lot of these boxes but doubting whether you should apply? At Alteryx, we support a growth mindset for our associates through all stages of their careers. If you meet some of the requirements and you share our values, we encourage you to apply. As part of our ongoing commitment to a diverse, equitable, and inclusive workplace, we’re invested in building teams with a wide variety of backgrounds, identities, and experiences.

Benefits & Perks:

Alteryx has amazing benefits for all Associates which can be viewed here.

For roles in San Francisco and Los Angeles: Pursuant to the San Francisco Fair Chance Ordinance and the Los Angeles Fair Chance Initiative for Hiring, Alteryx will consider for employment qualified applicants with arrest and conviction records.

This position involves access to software/technology that is subject to U.S. export controls. Any job offer made will be contingent upon the applicant’s capacity to serve in compliance with U.S. export controls.

Back to blog

Common Interview Questions And Answers

1. HOW DO YOU PLAN YOUR DAY?

This is what this question poses: When do you focus and start working seriously? What are the hours you work optimally? Are you a night owl? A morning bird? Remote teams can be made up of people working on different shifts and around the world, so you won't necessarily be stuck in the 9-5 schedule if it's not for you...

2. HOW DO YOU USE THE DIFFERENT COMMUNICATION TOOLS IN DIFFERENT SITUATIONS?

When you're working on a remote team, there's no way to chat in the hallway between meetings or catch up on the latest project during an office carpool. Therefore, virtual communication will be absolutely essential to get your work done...

3. WHAT IS "WORKING REMOTE" REALLY FOR YOU?

Many people want to work remotely because of the flexibility it allows. You can work anywhere and at any time of the day...

4. WHAT DO YOU NEED IN YOUR PHYSICAL WORKSPACE TO SUCCEED IN YOUR WORK?

With this question, companies are looking to see what equipment they may need to provide you with and to verify how aware you are of what remote working could mean for you physically and logistically...

5. HOW DO YOU PROCESS INFORMATION?

Several years ago, I was working in a team to plan a big event. My supervisor made us all work as a team before the big day. One of our activities has been to find out how each of us processes information...

6. HOW DO YOU MANAGE THE CALENDAR AND THE PROGRAM? WHICH APPLICATIONS / SYSTEM DO YOU USE?

Or you may receive even more specific questions, such as: What's on your calendar? Do you plan blocks of time to do certain types of work? Do you have an open calendar that everyone can see?...

7. HOW DO YOU ORGANIZE FILES, LINKS, AND TABS ON YOUR COMPUTER?

Just like your schedule, how you track files and other information is very important. After all, everything is digital!...

8. HOW TO PRIORITIZE WORK?

The day I watched Marie Forleo's film separating the important from the urgent, my life changed. Not all remote jobs start fast, but most of them are...

9. HOW DO YOU PREPARE FOR A MEETING AND PREPARE A MEETING? WHAT DO YOU SEE HAPPENING DURING THE MEETING?

Just as communication is essential when working remotely, so is organization. Because you won't have those opportunities in the elevator or a casual conversation in the lunchroom, you should take advantage of the little time you have in a video or phone conference...

10. HOW DO YOU USE TECHNOLOGY ON A DAILY BASIS, IN YOUR WORK AND FOR YOUR PLEASURE?

This is a great question because it shows your comfort level with technology, which is very important for a remote worker because you will be working with technology over time...