Director of Production Engineering
פורסם אתמול · 58 מועמדים
- 10+ years of engineering experience
- 4+ years in a Director or senior engineering leadership role, managing managers and team leads in a high-growth SaaS environment
- Experience owning both a central DevOps/platform-engineering function and a 24/7 operations or NOC function, including incident management and escalation protocols
- Public cloud expertise (primarily AWS)
- Container orchestration at scale (such as Kubernetes)
- Workflow automation platforms and engineering ownership tooling
- Exposure to SecOps and compliance programs (such as SOC 2, ISO 27001)
- Feature flagging and large-scale platform migration projects
- Modern analytical data platforms (such as Snowflake, ClickHouse, BigQuery)
- Autonomous agent frameworks and agentic AI tooling
חולץ מתיאור המשרה · מתעדכן אוטומטית
תיאור המשרה המלא
המשרה המקורית · נשמר לעיוןAbout the Company:
We are a leading global technology company that leverages advanced AI and machine learning to deliver intelligent, data-driven solutions.
The Role:
As the Director of Production Engineering, you will be a key leader within our Infrastructure Engineering Organization. This is a leadership role with a hands-on technical scope and significant organizational impact. You will own the reliability, velocity, and cost efficiency of our engineering infrastructure.
You will partner closely with R&D leadership to ensure our production systems are resilient, developer-friendly, and continuously improving, while directly managing multiple distributed engineering groups focused on platform engineering, automation, and cloud operations.
Responsibilities:
• Own the platform engineering roadmap and internal developer platform, including cluster operations, environment provisioning, and establishing architecture standards for reliability, observability, and deployment tooling.
• Lead major infrastructure modernization, upgrade, and migration programs, partnering closely with the Security team on compliance requirements, secrets management, and overall cloud security posture.
• Own the health, uptime, and SLAs of production and pre-production environments, overseeing 24/7 monitoring capabilities, incident response, severity classification, and post-incident reviews.
• Champion the adoption of AI and automation tooling across engineering and operations, overseeing the rollout of internal AI gateways and driving AI-assisted workflows and intelligent alerting.
• Focus on optimizing the developer experience and reducing toil by stabilizing CI/CD pipelines, improving local and lab environments, and tracking productivity metrics to accelerate deployment frequency.
• Build, mentor, and grow a high-performing distributed organization of engineers and managers, setting direction, managing quarterly planning, and defining OKRs aligned with company-wide priorities.
• Define and execute a FinOps strategy to manage compute, storage, and networking costs, embedding cost-awareness into the development lifecycle and driving saving initiatives.
Requirements:
• 10+ years of engineering experience, with 4+ years in a Director or senior engineering leadership role. Proven track record of managing managers and team leads in a high-growth SaaS environment.
• Experience owning both a central DevOps/platform-engineering function and a 24/7 operations or NOC function, including incident management and escalation protocols.
• Deep technical background in DevOps and infrastructure engineering, with strong public cloud expertise (primarily AWS).
• Comprehensive experience with container orchestration at scale (such as Kubernetes) and proficiency with Infrastructure-as-Code (such as Terraform or Pulumi).
• Hands-on familiarity with CI/CD systems and build reliability at scale, alongside experience with data platform operations and relational databases.
• Fluent in FinOps principles and cloud cost governance. Solid understanding of Linux, networking, and security fundamentals, with a demonstrated ability to integrate AI/GenAI tooling into engineering workflows.
Advantages:
• Familiarity with workflow automation platforms and engineering ownership tooling.
• Background in or exposure to SecOps and compliance programs (such as SOC 2, ISO 27001, etc.).
• Experience with feature flagging and large-scale platform migration projects.
• Familiarity with modern analytical data platforms (such as Snowflake, ClickHouse, BigQuery).
• Exposure to autonomous agent frameworks and agentic AI tooling.
שאלות על המשרה
- המשרה לא ציינה שכר. אנחנו מציגים שכר רק כשהמעסיק מפרסם אותו.
- 10+ years of engineering experience, 4+ years in a Director or senior engineering leadership role, managing managers and team leads in a high-growth SaaS environment, Experience owning both a central DevOps/platform-engineering function and a 24/7 operations or NOC function, including incident management and escalation protocols, Public cloud expertise (primarily AWS), Container orchestration at scale (such as Kubernetes)