M1 Finance has created a personal wealth-building platform made for the modern era, uniting personal perspective and automated ease. They are seeking a Senior DevOps & Infrastructure Engineer to help design and evolve their AWS infrastructure, ensure network reliability, and operate Kubernetes clusters.
Responsibilities:
- Help design and evolve our AWS foundation — account structure, network architecture (VPC, DNS, routing), and IAM — and shape the Terraform conventions the infra team builds on
- Harden our network's reliability, including the hybrid and on-prem connectivity behind systems that can't go down
- Operate and evolve our Kubernetes (EKS) clusters and the in-cluster platform services that run on them — ingress, service mesh, autoscaling, and policy
- Build the observability layer — metrics, logs, traces, and alerting — so teams can see how their systems behave and catch issues before customers do
- Expand our self-service platform and CI/CD + GitOps automation, so routine changes are safe and repeatable instead of manual and risky
- Keep security, secrets management, and least-privilege access sane as the platform grows
- Share in on-call, and bring an SRE mindset to keeping reliability, cost, and toil all trending in the right direction
Requirements:
- Senior-level experience is the one thing we really need
- You've spent several years (typically 5+) building and running production infrastructure or platforms, and you've earned real judgment from owning systems when they broke — not just when they behaved
- You've got hands-on depth in several of these, and pick up the rest fast:
- Cloud infrastructure (we run on AWS)
- Kubernetes in production (we use EKS)
- Infrastructure as code (we use Terraform)
- CI/CD and GitOps-style automated delivery (we use CircleCI and Flux)
- Observability — metrics, logs, traces, and alerting
- Secrets management and secure, least-privilege access
- Networking and Linux fundamentals — TCP/IP, DNS, routing and subnetting, firewalls and security groups, load balancing, and TLS
- General database administration and systems administration experience
- Experience with Kafka or another event-streaming platform
- Experience in regulated environments (e.g., ISO 27001 / SOC 2 certified) or working directly with auditors
- Windows experience — administration, PowerShell, and SQL Server
- Experience with agentic AI tooling and securing agentic workflows