The fab2 team can build anything. We design all the hardware and software needed to make chips and we make the fabs, tools, and components ourselves.
We're looking for exceptional, hands-on people who can iterate to the limits of what's possible.
fab2 was founded by Sam Zeloof and Jim Keller. Sam is best known for making chips in his garage, and Jim has been a leader in the semiconductor industry for the past 40 years.
We are seeking a Infrastructure and Site Reliability Intern for the summer term. The internship begins in May or June, with a preferred commitment of 4 to 8 months. You will design, build, deploy, and manage the on-prem backend infrastructure that powers our small, fast semiconductor fab.
This is a broad role encompassing all aspects of backend infrastructure and services.
Our philosophy towards infra is minimal, understandable, on-site, and close to the hardware.
What you'd do
- Build and evolve our infrastructure-as-code tooling to manage complex, multi-environment deployments
- Help manage and maintain our fleet of on-premises servers, ensuring high availability and optimal resource utilization
- Monitor and maintain system performance and health across diverse production workloads
- Design, implement, and refine alerting systems to catch issues before they impact users
- Build automation to reduce toil and improve operational efficiency
- Implement secure secret management
- Debug production issues across the stack
- Work closely with engineering teams to understand their infrastructure needs and deliver robust, scalable solutions
- Experience with configuration management or fleet management at scale
- Background in monitoring and observability tools (Prometheus, Grafana, ELK stack, Datadog)
What they want
- Pursuing a BS in Computer Science, Computer Engineering, or demonstrated exceptional skill in software engineering
- Strong programming skills in Rust, Go, or other systems-oriented languages
- Solid understanding of Linux systems, networking fundamentals, and distributed systems concepts
- Experience building or maintaining non-trivial infrastructure, automation, or tooling projects (personal, academic, or professional)
- Interest in one or more of the following areas: systems reliability, observability, infrastructure automation, performance optimization