GitHub is the world’s leading platform for agentic software development, and they are seeking a Staff Software Engineer for their Enterprise Cloud Platform team. This role involves defining and driving the technical strategy for GitHub's regional infrastructure, leading architecture decisions, and mentoring engineers to enhance operational excellence.
Responsibilities:
- Lead architecture and execution for large platform initiatives across regional deployments
- Define and drive the technical strategy for provisioning, capacity management, and service onboarding at scale
- Drive end-to-end improvements to provisioning, automation, and service onboarding
- Set technical direction across multiple repositories and partner teams
- Shape the platform's reliability, scalability, and operational model as it grows
- Identify and address architectural gaps, cross-service dependencies, and systemic risks before they become incidents
- Create durable patterns, frameworks, and documentation that enable other engineers and teams to operate independently
- Mentor engineers and raise the bar for system design, incident response, and cross-team execution
- Represent the team's technical perspective in leadership discussions and cross-org planning
Requirements:
- 9+ years' experience in software engineering, computer science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, Go, Ruby, Rust, Python, JavaScript, C, C++, C#, Java
- OR associate's degree in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field AND 8+ years' experience in software engineering, computer science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, Go, Ruby, Rust, Python, JavaScript, C, C++, C#, Java
- OR bachelor's degree in Computer Science or related field AND 7+ years' experience in software engineering, computer science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, Go, Ruby, Rust, Python, JavaScript, C, C++, C#, Java
- OR master's degree in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field AND 5+ years' experience in software engineering, computer science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, Go, Ruby, Rust, Python, JavaScript, C, C++, C#, Java
- OR doctorate in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field AND 3+ years' experience in software engineering, computer science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, Go, Ruby, Rust, Python, JavaScript, C, C++, C#, Java
- OR equivalent experience
- 3+ years Experience with infrastructure-as-code, provisioning automation, or platform orchestration (e.g., Terraform, Bicep, Pulumi)
- 2+ years of demonstrated ability to effectively leverage AI-assisted development tools (e.g., GitHub Copilot, AI code review) to improve productivity, code quality, and engineering workflows
- 5+ years of hands-on On-call experience, driving post incident retrospectives and driving systemic improvement and repairs through
- Experience with multi-region or data-residency-aware architectures at scale
- Experience with capacity planning, quota management, or cloud resource lifecycle across multiple environments
- Experience leading incident response and driving systemic reliability improvements
- Experience mentoring senior engineers or leading technical working groups
- Track record of setting technical direction that influenced how multiple teams build or operate, creating reusable systems, patterns, or platforms adopted beyond your immediate team