Jobtailor Deutschlandweit vor 3 Tagen

Principal Engineer, CSRE Provisioning

Jetzt bewerben bei Jobtailor Sichere Bewerbung über StudySmarter
Aus der Stellenanzeige

Die ganze Ausschreibung von Jobtailor

Das ist der Job

• Set technical direction and a multi-quarter roadmap for platform lifecycle work across the CSRE hosted estate • Lead the hardest and highest-risk platform initiatives, including platform-side execution • Define and evolve Provisioning’s supported operating model and standards • Act as a technical escalation point for production incidents and drive durable structural fixes • Represent Provisioning in architecture and design reviews with Platform Engineering, application teams, and senior leadership • Build and evolve reliability, disaster recovery, and observability practices across the estate • Lead service acceptance for complex or high-risk new platforms • Identify infrastructure and processes that can be retired before advocating for new ones • Participate in the team’s on-call rotation and escalates complex incidents • Mentor and develop senior and staff-level engineers through pairing, design review, and career coaching • Contribute to sprint planning and PI-cycle delivery on the PROV Jira board • Feed operational insights back into CSRE Platform tooling and Provisioning standards Requirements • Extensive practical experience operating and evolving production infrastructure at scale • Experience with complex, high-risk migrations and decommissioning • Deep expertise designing and troubleshooting distributed systems • Expert understanding of SRE principles, including SLIs, SLOs, and error budgets • Extensive experience with on-premises data center infrastructure and cloud-native environments • Experience with governance, cost trade-offs, and vendor evaluation at scale • Track record improving observability and alerting across many platforms • Deep experience with infrastructure-as-code and automation • Excellent software engineering fundamentals • Extensive experience with production readiness, DR validation, and controlled failure testing • Expert-level incident analysis skills • Practical experience directing where LLM and AI-assisted tooling helps engineering work • Excellent written and verbal communication • Ability to participate in an on-call rotation • Ability to mentor senior and staff-level engineers Core Competencies Demonstrates extensive experience in operating and evolving production infrastructure at scale, with a strong focus on incident analysis, observability, and disaster recovery practices.

Darum lohnt es sich

Capable of mentoring engineers and leading high-risk platform initiatives while ensuring adherence to SRE principles.

Highest-signal resume keywords • Production Infrastructure Management • Distributed Systems Design • SRE Principles Expertise • Infrastructure-as-Code • Incident Analysis Hard Skills • Production Readiness • Disaster Recovery Validation • Error Budgets • Automation • Observability Improvement • Software Engineering Fundamentals • High-Risk Migrations • Decommissioning • Governance • Cost Trade-offs Soft Skills • Excellent Written Communication • Excellent Verbal Communication • Mentoring Industry Keywords • Platform Lifecycle • Service Acceptance • Technical Direction • Production Incidents • Operational Insights Tools & Technologies • Cloud-Native Environments • On-Premises Data Center Infrastructure • Jira

Bereit?

Bewerbung wird direkt an Jobtailor übergeben — kein Konto nötig.

Jetzt bewerben
Weitere Stellen bei diesem Arbeitgeber

Jobtailor hat 1 weitere offene Stelle:

Ähnliche Stellen

Wenn dir dieser Job gefällt, schau dir auch an:

Weiter stöbern:

Kostenfrei starten