Notifications

Loading notifications...
Horizon3.ai Cover

Senior Engineering Manager, Cloud Platform

Horizon3.ai
Worldwide Full Time USD 260,000 - 280,000 19 days ago

About the Job

We are a fusion of former U.S. Special Operations cyber operators, startup engineers & operators, and formerly frustrated cybersecurity practitioners. We're committed to helping solve our common security problems: ineffective security tools and false positives, resulting in alert fatigue, blind spots, "checkbox” sec...

We are looking for an engineering leader to lead Horizon3’s infrastructure platform team. Our platform team is building capabilities for product development teams to be able to rapidly launch new product features, by providing self-service environments and infrastructure. This leader will also be chartered to establ...

Lead software engineering teams providing infrastructure-as-code to manage cloud infrastructure. Provide high quality IaC components and frameworks to support application development teams to leverage and extend to self-service their infrastructure provisioning.

Key Responsibilities

You will:
Establish governance and mechanisms for application development teams to self-service infrastructure provisioning, while providing for best practices and controls.
Provide documentation, training, and support to ensure feature dev teams are leveraging self-service capabilities.
Hire experienced site reliability staff, and a line manager to grow and oversee the SRE team.
Professionalize incident management. Define and document incident processes and practices for your SRE team and for the application feature teams. Make tool and vendor decisions to support processes.
Drive incident professionalism across the engineering organization through training and process adoption.
Establish design-before-build discipline. Facilitate lightweight design documents, architectural decision records, and working group reviews. Outline operational lifecycles, “Day 2” concerns, and developer experience as part of infrastructure architecture decisions.
Use design reviews, code reviews, and blameless retrospectives to drive a culture of quality and excellence in engineering.
Balance providing developer support while also executing on a roadmap of infrastructure engineering initiatives. Establish intake, allocate resources, provide visibility into backlogs to stakeholders, and manage prioritization against capacity.
Directly manage a growing team of infrastructure engineers. Hire and develop line managers and staff / principal engineers. Ensure a strong bench of technical and leadership talent in your group. As a Manager, you will be responsible for:
Recruiting and onboarding talented individuals to support our organizational goals
Mentoring, coaching, equipping, and developing your team
Recognizing and retaining high performers
Leading horizontally with peer Management & Senior Leaders
Please note this job description is not designed to cover or contain a comprehensive listing of activities, duties or responsibilities that are required of the employee. Duties, responsibilities, and activities may change at any time with or without notice.
In any materials you submit, you may redact or remove age-identifying information such as age, date of birth, or dates of school attendance or graduation. You will not be penalized for redacting or removing this information.

Required Skills & Abilities

Demonstrated experience leading teams operating SaaS service infrastructure.
Deep hands-on experience deploying and operating production infrastructure on public cloud platforms (AWS strongly preferred; Azure and GCP familiarity a plus).
Strong command of Infrastructure as Code, including Terraform; experience with Crossplane and GitOps patterns strongly preferred.
Experience managing production Kubernetes environments at scale.
Solid understanding of security best practices including zero trust architecture, secrets management, identity and access management, and software supply chain security.
Experience building and operating self-service infrastructure platforms that enable application development teams, while balancing self-service and developer productivity with maintainability and security.
Experience leading or building SRE functions, including incident management processes, on-call programs, SLO/SLA definition, and operational runbooks.
Deep hands on experience with observability: application performance management, logs and traces, and golden signals and service-specific metrics.

Apply now

Please let Horizon3.ai know you found this job on Job Vista. This helps us grow!