Site Reliability Engineer (Network)
San Francisco, CA - USA
Job Summary
As a Site Reliability Engineer with strong networking skills in our Cloud Infrastructure (SRE) team you help the team own the networks that keep Loft running: cloud networking VPN and site-to-site connectivity segmentation and routing across our offices test equipment and ground stations. You keep them reliable observable and secure as we scale. You coordinate with IT on the network components that interface with the physical infrastructure used by employees. You coordinate with the Test-Infrastructure team to ensure our test equipment is securely accessible.
Youll work the way an SRE team works: defining reliability in measurable terms (SLOs for connectivity and ground-station availability) instrumenting the systems that prove it managing the network as code and automating away toil so on-call load doesnt grow with the fleet.
Beyond the network youll help the SRE team across its broader scope by participating in all the daily work just like a regular SRE: platform reliability and expansion incident response planning and user support.
- Participate in the design and implement Lofts network architecture; cloud networking VPN/site-to-site segmentation and routing across offices ground stations and test equipment
- Tighten the network security posture: segmentation access and connectivity hardening
- Manage the network as code: IaC provisioning CI/CD and GitOps
- Define SLOs for connectivity and ground-station availability and own the observability (metrics logs tracing) that measures them; Grafana-centric stacks a plus
- Investigate and resolve reliability issues with root-cause analysis and durable fixes; reduce operational toil through automation
- Contribute to the SRE teams broader reliability and incident-response work fostering the SatDevOps culture that power Lofts DNA
- 45 years in network engineering with deep hands-on networking: routing VPN/site-to-site segmentation DNS firewalling
- Hands-on Software-Defined Networking: comfortable managing networks programmatically / as code rather than appliance-by-appliance (everything we run is IaC)
- Strong public-cloud networking experience ideally GCP (VPCs peering hybrid connectivity)
- Comfortable operating on a Kuburnetes platform (k8s Docker) and occasionally writing code
- Infrastructure-as-Code (Terraform or similar) and a CI/CD-driven GitOps workflow
- An SRE mindset: SLOs observability and reducing toil through automation
- Ability to debug connectivity issues in a complex hybrid network environment
- Degree in Computer Science or a related field or equivalent practical experience
- Network certifications (CCNP or similar)
- Hands-on experience with GitOps frameworks (ArgoCD FluxCD)
- Interest or experience in FinOps and cost-optimized architectures
- Familiarity with security practices: vulnerability scanning threat detection risk mitigation
- Understanding of orchestration in resource-constrained environments like space systems
- 100% company-paid medical dental and vision insurance option for employees and dependents
- Flexible Spending (FSA) and Health Savings (HSA) Accounts offered with an employer contribution to the HSA
- 100% employer paid Life AD&D Short-Term and Long-Term Disability insurance
- Flexible Time Off policy for vacation and sick leave and 12 paid holidays
- 401(k) plan and equity options
- Daily catered lunches and snacks in office
- International exposure to our team in France
- Fully paid parental leave; 14 weeks for birthing parent and 10 weeks for non-birthing parent
- Carrot Fertility provides comprehensive inclusive fertility healthcare and family-forming benefits with financial support
- Off-sites and many social events and celebrations
- Relocation assistance when applicable
Required Experience:
IC
About Company
Loft Orbital builds software and hardware products to fly and operate any payload on a standard satellite bus. See how we're making space simple.