Senior Infrastructure Engineer, AWS
Capacity
- Location
- Remote (Remote (United States))
- Compensation
- $125k - $150k/yr
- Employment
- Full-time
- Level
- Senior Level
About the Role
Capacity is an AI-powered support automation platform helping clients reduce costs and improve efficiency. This role involves building and evolving the AWS cloud platform, managing Kubernetes clusters, and leading infrastructure initiatives as a senior hands-on engineer.
Skills
Benefits
- Health insurance
- 401(k)
- Short term disability
- Life insurance
Perks
- Unlimited vacation
- Equity
- Remote work
Full job details
Why this job is exciting
The role:
You'll be the senior hands-on engineer for the cloud foundation everything else at Capacity runs on. That means building and evolving our AWS platform, operating the Kubernetes clusters our services run on, authoring the infrastructure as code that provisions and governs our environments, designing the network architecture, and keeping cloud costs under control as we scale across acquired platforms.
You'll report to the Manager, Infrastructure and set the technical bar for how infrastructure is defined, reviewed, and versioned. You own the "what exists" layer; you partner with DevOps at the boundary where your IaC gets executed, and with SRE where your platform is observed and kept reliable. As the senior IC on the team, you'll also mentor other infrastructure engineers and lead the design work on our larger platform initiatives.
Responsibilities:
- Build and evolve Capacity’s cloud platform on AWS, including compute, storage, database (Aurora / MariaDB on RDS), and networking architecture across a multi-account AWS Organization
- Provision, operate, and scale Capacity’s Kubernetes platform (EKS): cluster infrastructure, node groups, ingress (AWS ALB), Linkerd service mesh, Helm-based platform components, and autoscaling (KEDA, cluster-autoscaler)
- Author and own infrastructure as code using Terraform and Terragrunt; set and maintain standards for how infrastructure is defined, reviewed, and versioned
- Design and manage network architecture: VPCs, segmentation, DNS, AWS VPN, and certificate infrastructure
- Define and enforce environment provisioning standards across dev, staging, and production
- Lead cloud cost optimization and capacity planning; evaluate architectural tradeoffs with cost in mind
- Drive design and execution on major platform initiatives (e.g. AWS account consolidation across acquired platforms, network redesign, database migrations, cross-platform standardization)
- Mentor infrastructure engineers through code review, pairing, and design guidance
- Partner with DevOps at the IaC execution boundary (you write the infrastructure, their pipelines run it)
- Partner with Information Security on cloud security posture, network segmentation, and compliance work supporting SOC 2, ISO 27001, HIPAA, and PCI obligations
- Author and maintain infrastructure documentation, architecture diagrams, and provisioning standards
Requirements:
- 5+ years in cloud infrastructure or platform engineering
- Deep hands-on experience with AWS, including networking, compute, storage, and IAM
- Hands-on experience operating Kubernetes in production (EKS preferred), including cluster networking, ingress, Helm, and autoscaling (KEDA, cluster-autoscaler)
- Strong Terraform experience; comfortable owning an IaC codebase at scale
- Experience with GitOps deployment tooling (ArgoCD preferred) and CI/CD in GitHub Actions
- Solid understanding of network architecture: VPCs, segmentation, DNS, CDN, certificates
- Experience with observability tooling and concepts, including metrics, dashboards, monitoring, and alerting (Datadog preferred)
- Experience with cloud cost management and capacity planning
- AWS certifications (e.g. Solutions Architect, SysOps Administrator) are a big plus
- Experience supporting compliance programs (SOC 2, ISO 27001, HIPAA, or PCI) is a plus
- A track record of leading infrastructure design work and mentoring other engineers
- Strong cross-functional communication; comfortable working across DevOps, Security, and product engineering teams
- Experience in a SaaS engineering environment with containerized workloads
You are motivated by:
- Hustle: You inspire others to work as hard as you. You will find a way, no matter how hard the task is.
- Ownership: You have an owner/builder mentality. You care about what you deliver and own your mistakes.
- Proactivity: You don’t wait for someone to tell you what to do or what problems to solve. You are always looking for ways to learn and improve.
- Excellence: You set a high bar and surpass expectations. You hit your goals and ask for more.
- Humility: You are not above any task in the organization and are willing to drop what you’re doing to help a teammate.
What you can expect from us
The team:
Capacity team members enjoy the opportunity and benefits of working at an artificial intelligence startup, but with leaders who’ve worked at places like Apple, Ebay, Visa, Answers.com, Oracle, Boeing, and many more world-class companies. The culture at Capacity encourages innovation, independent problem solving, and collaboration as we continue to mature our product in the ever-changing world of AI.
We provide:
- Employer-paid health insurance (for you and your eligible dependents)
- Company Profit Interest Units eligibility
- Unlimited vacation policy
- 401(k) with a company match
- Short term disability insurance
- Group life AD&D insurance
- A supportive, diverse workplace where we prioritize respect for each other and our clients
- A fun and collaborative team culture
Salary range:
- The expected base salary for this role is between $125,000 and $150,000; actual salary will be commensurate with a candidate's experience, skill and location.
Still unsure?
At Capacity we value more than just hard skills. Our goal is to build a holistic and diverse team. If you aren’t sure if you qualify, just apply! We will carefully consider your application and are always grateful for any time and effort invested in Capacity.
But wait, there’s more!
At Capacity we believe in more than just building amazing products and helping our customers. Although we are a remote workforce, we remember the neighborhood where we started. We still strive to elevate our community by furthering access to education and careers in the tech space. Our affiliated nonprofit, Create A Loop, brings rigorous computer science courses to underserved communities with little to no access to formal computer science education. There are many opportunities for our Capacity team members to serve and educate our Create A Loop students throughout the year.