Principal Front-End Network Engineer at Nscale | San Francisco, CA, USA | Rezi

Principal Front-End Network Engineer at Nscale

Principal Front-End Network Engineer

Nscale · San Francisco, CA, USA

1 weeks ago

Principal Front-End Network Engineer

Nscale · San Francisco, CA, USA

14 days ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

The Network Operations and Engineering teams at Nscale operate demanding networking environments supporting large-scale AI GPU clusters. This role provides technical leadership for Nscale’s front-end networking domain, focusing on reliability, scalability, and evolution of high-performance Ethernet front-end networks. You will act as a senior technical authority, influencing architecture, standards, and operational practices while addressing complex network challenges.

Responsibilities

  • Owning the technical direction and operational strategy for Nscale’s front-end AI infrastructure networks at the highest level.
  • Designing, reviewing, and evolving large-scale Ethernet leaf-spine / Clos fabric architectures (including Arista and Nokia platforms) to support future growth, inference workloads, and storage requirements.
  • Acting as the senior-most escalation point for complex front-end network incidents, guiding investigations and fixes.
  • Driving cross-team initiatives to improve fabric reliability, performance predictability, observability, and operational maturity.
  • Defining standards for hardware configuration, routing, congestion management, firmware lifecycle management, automation, and change safety across Nvidia Cumulus, Arista EOS and Nokia platforms.
  • Partnering with SRE, Compute Platform, Storage, and Network Architecture teams to influence end-to-end system design.
  • Mentoring senior and principal-level network engineers, raising the bar for operational rigor and technical excellence.
  • Driving measurable improvements in uptime, latency consistency, capacity efficiency, and incident reduction for front-end services.

Requirements

  • 12+ years of experience in network engineering, with deep focus on hyperscale data centre, cloud, or AI infrastructure networking.
  • Expert-level operational and architectural experience with large-scale Ethernet data centre fabrics (leaf-spine / Clos topologies).
  • Strong hands-on expertise with Nvidia (Cumulus), Arista(EOS / Etherlink) and/or Nokia(7220 IXR, 7250 IXR, 7750 SR series) platforms in production environments at scale.
  • Deep understanding of modern data centre routing and control planes (BGP, OSPF, ECMP, EVPN-VXLAN).
  • Proven experience with long-haul circuits, DCI, and optical transport (dark fiber, carrier Ethernet, coherent optics, ZR/ZR+).
  • Strong background in storage networking over Ethernet and shared storage connectivity at hyperscale.
  • Demonstrated ability to debug and resolve complex cross-layer issues spanning hardware, optics, routing, and application layers.
  • Proven ability to lead complex technical initiatives across teams and influence strategy without direct authority.
  • A systems-level mindset, balancing performance, reliability, scalability, and operational cost at the highest level.

Skills

  • Nvidia (Cumulus)
  • Arista (EOS / Etherlink)
  • Nokia (7220 IXR, 7250 IXR, 7750 SR series)
  • BGP
  • OSPF
  • ECMP
  • EVPN-VXLAN
  • Python
  • Ansible

Experience Level

  • Senior Principal

Salary/Compensations

  • The range below reflects the base salary for the position. Actual compensation may vary based on job-related factors such as skill set, experience, education, and location. In addition to base salary, this role may be eligible for bonus, equity, and/or commission programs.

Benefits

  • Highly competitive package (base + equity) with reviews every 12 months.
  • Dynamic progression plan tailored to your ambitions.
  • Medical
  • Dental
  • Vision
  • Flexible paid time off
  • Parental leave
  • Retirement plan participation

About the Company

  • Nscale is the GPU cloud engineered for AI, providing cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers.
  • Nscale enables AI-focused companies to achieve superior results by reducing the complexity of AI development.
  • Our GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.
  • We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency.
  • As an Nscaler, you’ll build trust through openness and transparency, where everyone is inspired to do their best work.
  • If you join our team, you’ll be contributing to building the technology that powers the future.

Equal Opportunity

  • We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds.
  • If there’s anything we can do to accommodate your specific situation, please let us know.