Senior Network Engineer (On Site)
Campbell, OH - USA
Job Summary
About the Role
Mirantis is seeking a Senior Network Engineer to join our IT team in a hands-on role spanning two areas: the networking that underpins our internal AI labs and the broader company network infrastructure. This is the second dedicated networking role on the team working alongside our existing network engineer in Europe to scale a function that has grown well beyond what one person can carry. The role is based in Silicon Valley with regular on-site presence expected at our local datacenter/colocation facilities.
As Mirantis invests in AI compute the lab environments behind it have become critical infrastructure that IT operates as a shared platform engineering groups across the company consume them as tenants. Our job is to make that possible: high-throughput fabric for GPU and compute-intensive workloads clean segmentation between tenants and reliable connectivity back to corporate and cloud environments. This is a hardware- networking- and architecture-focused mandate building and running the underlying infrastructure not the AI/ML work that runs on top of parallel our day-to-day company network offices datacenter/colo presence cloud interconnects and remote access requires ongoing engineering ownership rather than best-effort maintenance.
In this role you will design build and operate the network and physical infrastructure for the AI labs and evolve it as demand grows while sharing responsibility for the corporate and datacenter network. You will work close to the hardware where it matters partner with the IT Security and engineering teams who depend on these environments and help establish the standards and documentation that let this function scale as a proper team.
Key Responsibilities
AI Labs Network Architecture & Operations: Design build and operate the network fabric for our internal AI labs high-throughput switching for GPU and compute-intensive workloads with the segmentation needed to run them as a shared platform for engineering groups across the company. Own capacity planning and expansion as demand grows.
Datacenter & Colocation Infrastructure: Own the physical and logical infrastructure at our Silicon Valley datacenter/colocation facility network fabric rack and cabling standards cross-connects power/port planning and hands-on structured work with regular on-site presence for buildout changes and troubleshooting. Operate the network at our other datacenter locations as well (including in Europe) where physical work is handled through contractors or remote hands rather than direct access.
Server & Hardware Support: Provide hands-on support for server and datacenter hardware racking cabling provisioning firmware/BMC (iDRAC/iLO) management and physical troubleshooting. Perform this work directly on-site where practical or coordinate it through third-party contractors and remote-hands services as needed.
Corporate & Campus Networking: Share ownership of the day-to-day company network offices LAN/WLAN remote access and connectivity between on-prem/colo corporate and cloud environments (site-to-site VPN cloud interconnects such as AWS and SD-WAN where appropriate) alongside our existing network engineer in Europe. Bring these environments up to a consistent well-engineered standard rather than best-effort maintenance ensuring routing redundancy and performance meet the needs of both the AI labs and corporate users.
Network Security & Segmentation: Partner with Corporate/Product Security to operationalize network security baselines firewall policy ACLs network access control least-privilege segmentation and secure management planes across all environments.
Monitoring Observability & Capacity: Implement and maintain network and infrastructure monitoring alerting and traffic visibility across labs datacenter corporate and cloud environments. Use it to catch issues early plan capacity and inform architecture decisions.
Standards Documentation & Automation: Establish network and infrastructure standards runbooks and clear architecture documentation as this function scales into a proper team. Automate configuration and repetitive operations (IaC / network automation) to reduce manual work and drift.
Incident Response & On-Call: Participate in an on-call rotation to respond to network and infrastructure incidents across labs and corporate environments. Triage and resolve production issues perform root cause analysis and contribute to post-incident reviews to prevent recurrence.
Qualifications :
5 years of experience in a Network Engineering role with networking as the core of your background.
Strong hands-on experience designing and operating switching and routing in datacenter and/or colocation environments spine/leaf or top-of-rack fabrics VLANs VXLAN/EVPN BGP/OSPF and structured cabling.
Solid experience with firewalls ACLs VPNs (site-to-site and remote access) and network segmentation. Vendor experience with the likes of Cisco Arista Juniper Palo Alto or Fortinet.
Comfortable working physically on-site in a datacenter racking cabling cross-connects and troubleshooting at the hardware level and coordinating remote-hands or contractor work at sites without direct access.
Working knowledge of server hardware provisioning firmware/BMC management (iDRAC/iLO) and physical troubleshooting with the willingness to own the compute layer alongside the network.
Experience with cloud and hybrid connectivity (AWS interconnects VPN SD-WAN) and how on-prem colo and cloud environments fit together.
Familiarity with network monitoring and observability (e.g. SNMP-based tooling NetFlow Prometheus/Grafana or similar) and using it for capacity planning.
Scripting/automation and network automation or Infrastructure as Code (Python Bash Ansible Terraform or similar) hands-on experience is a strong plus; alternatively a demonstrated ability and strong desire to pick these up quickly.
Prior experience with on-call rotations and incident management.
Good understanding of network and infrastructure security fundamentals least-privilege access secure management planes and segmentation.
High-throughput and GPU/AI cluster networking (RDMA/RoCE InfiniBand high-speed Ethernet) existing experience is a strong plus; alternatively a strong desire and ability to ramp up quickly in this area.Experience with SOC 2 ISO 27001 or similar compliance frameworks is a plus.
Strong communication skills and the ability to document standards runbooks and architectural decisions clearly.
Additional Information :
What does Mirantis offer you
- Work with an established Silicon Valley leader in the cloud infrastructure industry;
- Work with exceptionally passionate talented and engaging colleagues helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
- Be a part of cutting-edge open-source innovation;
- Thrive in the high-energy environment of a young company where openness collaboration risk-taking and continuous growth are valued;
- Professional development and training;
- Attend conferences and working groups;
- Company outings happy hours hackathons and tech talks;
- Receive a competitive compensation package with a strong benefits plan.
We are a Leader for Container Management in G2 (#2 after AWS)!
Remote Work :
No
Employment Type :
Full-time
About Company
Mirantis is an open cloud company that helps organizations achieve digital self determination by giving them complete control over their strategic infrastructure. The company combines intelligent automation and cloud-native expertise for managing and operating virtual machines, contai ... View more