Solutions Architect
Boston, United States · Remote · Full-time
- Posted 3h ago
- From Mirantis’s careers page
- Location
- Boston, United States
- Work mode
- Remote
- Type
- Full-time
- Level
- Senior
- Experience
- 8+ years
- Department
- Engineering
Opens the listing on jobs.smartrecruiters.com
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.
About the role
Company Description:
Mirantis, an IREN company, is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment, on-premises, in the cloud, at the edge, or in sovereign data centers. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.
Job Description:
We are seeking a Senior Solutions Architect with deep networking expertise to join the Voyager team in the Mirantis Office of the CTO. You will own the network solutioning of k0rdent AI, our platform for building and operating GPU clouds and AI factories, from metal to model.
You will design how large GPU clusters are interconnected, how tenants are isolated, and how workloads reach the network from containers, virtual machines and bare metal nodes. You will turn that design into published reference architectures, working code and proofs of concept that run on real hardware with our customers and partners.
As part of the Office of the CTO, research is a core part of the job. You will explore designs and technologies before customers ask for them, prototype ideas that may not ship, and challenge established practice when a better approach exists.
This role combines research with hands-on, outward-facing work. You will split your time between exploring new designs, writing architecture, building and validating it in the lab, and explaining it to engineers, executives and conference audiences.
Key Responsibilities:
- Design and publish network reference architectures and solution designs for k0rdent AI, from a single rack to multi-thousand-GPU clusters.
- Define compute (east-west) fabrics on InfiniBand and on RoCEv2 Ethernet, including rail-optimised and fat-tree / Clos designs.
- Define front-end, storage and management networks, and how they connect to customer data centers, public clouds and hybrid environments.
- Specify multi-tenant isolation across the stack: InfiniBand partitions (PKeys), VRFs and EVPN-VXLAN on Ethernet, and Kubernetes-level network policy.
- Document scale limits and trade-offs (cost, performance, operability, vendor lock-in) with defensible reasoning.
- Define how Linux hosts expose NICs, DPUs and SuperNICs to workloads, including PCI passthrough, SR-IOV, IOMMU groups, NUMA and GPU-NIC affinity.
- Design Kubernetes networking for AI workloads, including primary and secondary CNIs, Multus, SR-IOV and RDMA device plugins, and Dynamic Resource Allocation (DRA).
- Design networking for virtual machines on KubeVirt, including passthrough and SR-IOV for GPU and RDMA traffic.
- Track and evaluate emerging AI networking technologies and standards, and prototype alternative designs in the lab.
- Build and run proofs of concept with customers and partners, and report results against agreed success criteria.
- Write automation and tooling (Python, Go, Bash, Ansible, Helm, Kubernetes manifests, Terraform) that makes reference architectures reproducible.
- Run validation and benchmarking of fabrics and host configurations (e.g. NCCL tests, perftest, ib_write_bw).
- Act as the network subject matter expert in customer discovery, design reviews and architecture workshops.
- Work with hardware and networking partners such as NVIDIA, server OEMs, and switch vendors on joint designs and validations.
- Present at industry events, webinars and partner summits, and write technical content.
Qualifications:
- Bachelor's degree in Computer Science, Electrical Engineering, Telecommunications or a related field, or equivalent practical experience.
- 8+ years in network engineering or network architecture, with at least 3 years in data center, HPC or cloud infrastructure networking.
- Customer-facing experience as a solutions architect, pre-sales engineer, consultant or technical lead.
- Strong knowledge of data center network design (spine-leaf, Clos, rail-optimised GPU fabrics), InfiniBand, and RoCEv2 (PFC, ECN, DCQCN).
- Experience with routing and overlays (BGP, EVPN-VXLAN, VRFs), hybrid and multi-cloud connectivity, and Linux networking.
- Experience with Kubernetes networking (CNI plugins, Multus, SR-IOV and RDMA device plugins, network policy) and KubeVirt.
- Programming or scripting in at least one language (Python or Go preferred), and comfortable with Git, CI and infrastructure-as-code.
- Excellent written and spoken English, and ability to present to audiences from network engineers to customer executives.
Nice to Have:
- Hands-on experience with NVIDIA networking (Quantum InfiniBand, Spectrum-X Ethernet, ConnectX SuperNICs, BlueField DPUs, UFM).
- Experience with GPU cloud, neocloud or HPC operators.
- Bare-metal provisioning and lifecycle experience (Metal3, Ironic, Redfish, PXE / iPXE, Cluster API).
- Network automation, SDN controllers, and switch operating systems such as Cumulus Linux or SONiC.
- Contributions to open-source networking or Kubernetes projects, or participation in standards bodies.
Location, Travel and Working Model:
- This role is remote, open to candidates based in Europe.
- Travel of up to 25% for customer engagements, partner meetings, lab work and industry events, primarily within the EU and US.
- Compensation and benefits are set according to the local market and employment arrangement.
Additional Information:
- Build the observability foundation for the AI cloud era, working directly with leading GPU cloud operators, NeoClouds, sovereign clouds, and AI-first enterprises.
- Collaborate with a world-class, distributed team committed to openness and technical excellence.
- Shape the product narrative and influence go-to-market success.
Skills they ask for
Pick one to see other roles that ask for it.
About Mirantis
Build AI infrastructure your wayMirantis provides tools for building and operating AI-ready infrastructure across environments, including its k0rdent AI platform.
See all 10 roles at MirantisMore roles at Mirantis
See all 10- Network EngineerCampbell · Senior · On-siteInformation Technology · Senior · On-siteCampbell, United States2d
- Product ManagerAustin · RemoteProduct Management · RemoteAustin, United States2w
- Senior AI Infrastructure & Platform Operations EngineerSenior · RemoteInformation Technology · Senior · Remote2w
- Solutions ArchitectAustin · SeniorSoftware Development · SeniorAustin, United States2w
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.