Principal Network Engineer
This is the employer's own posting, not a copy on a job board.
What we know
Is it still open?
Confirmed still open
Last checked 1d ago — checked against the employer's own applicant tracking system, which is the company answering directly.
We re-read the employer's own applicant tracking system and the posting was still there. That is the company answering directly.
How old is it?
Posted 45d ago
The date the source published, not the day we noticed it (2026-07-31). Last seen at its source 1h ago.
Is it remote?
Remote
That is the location the employer filed this posting under. Quoted as written — we do not re-word the source's own location.
Who may apply?
United States
No source stated where this role may be worked. This is read from the ad's own words.
What the ad says
…authorization to work in the United States, as required by law…
Pay not stated
Similar roles pay $196.0k–255.7k/yr
Middle 50% of 551 listings that do state pay — Engineering · Lead · United States · USD/year. This employer has published no salary; this is what comparable listings we hold disclose, never converted between currencies or periods. How this is calculated.
Skills named in the ad
Recognised terms only, from a fixed vocabulary — this is what CV matching compares against.
Carried by 1 source
-
ashby employer's own board first seen 11d ago · last seen 1h ago
- location not stated
The listing
About TensorWave
Our mission is simple: deliver seamless, secure, reliable, and resilient AI compute at scale. We've built a versatile cloud platform that eliminates infrastructure barriers, empowering builders to focus on innovation instead of fighting their stack. Because breakthrough AI should move at the speed of ideas, not infrastructure.
About the Role
We’re seeking a Principal Network Engineer (L7) focused on owning and evolving large-scale, RoCEv2 data center networks powering next generation AI and ML infrastructure.
You’ll work closely with our network architect and infrastructure leadership to define how the network is designed, implemented, and operated at scale, keeping over 8,000 GPUs burring today and scaling to cluster sizes reaching over 100,000 GPUs. You will be responsible for the architectural decisions that determine performance, reliability, and operational sanity at scale.
You’ll remain hands-on with high-speed optics, switching, routing, and congestion management in production clusters, while also setting the standards, patterns, and tooling other engineers build and operate against.
What You’ll Do
As a Principal Engineer, you own end-to-end network architecture, make high-impact design decisions, and set technical direction across teams, with clear examples of systems you’ve defined and scaled
Define, evolve, and standardize large-scale RoCEv2 data center networks supporting AI and ML clusters from thousands to 100,000+ GPUs
Set and validate congestion management strategy across RDMA fabrics, including PFC, ECN, and DCQCN, based on real production behavior
Establish automation, validation, and observability patterns that prevent misconfiguration and eliminate manual operational work
Act as the technical escalation point for complex failures, scaling limits, and architectural tradeoffs in always-on, multi-tenant environments
Essential Skills & Qualifications
Bachelor’s degree in Computer Science, Electrical Engineering, or a related technical field, or equivalent practical experience
Deep experience designing and operating RDMA and RoCEv2 networks in large-scale production environments supporting AI or HPC workloads
Expert-level knowledge of switching hardware and their NOS, such as Arista, Juniper, and custom solutions using SONiC, including high-speed Ethernet fabrics
Proven hands-on experience with congestion management and performance tuning using PFC, ECN, and DCQCN
Strong experience with high-speed optics and cabling including 400G, 800G, and AEC, AOC, DAC, and structured cabling at scale
Strong automation mindset, with experience using Python, Ansible, Terraform, Git, and production observability tooling
We’re looking for engineers who operate comfortably at scale, make hard calls with incomplete data, and take responsibility for systems that must work under sustained load. The solutions that work on a handful of devices will not work at Exascale.
What We Offer
Stock Options
100% paid Medical, Dental, and Vision insurance for Employees
Company Health Savings Account Contributions
100% paid Short Term and Long Term Disability Insurance for Employees
Life and Voluntary Supplemental Insurance Options
Other Insurance Options, such as Pet & Legal Insurance
Various Supplementary Health Benefits, such as discounted Virtual Healthcare Appointments and Serious Illness Support
Flexible Spending Account
401(k)
Employee Assistance Program
Flexible PTO
Paid Holidays
Parental Leave
Other In-Office Perks
Equal Employment Opportunity
TensorWave is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. We do not discriminate on the basis of any protected status under applicable law.
Reasonable Accommodations
TensorWave provides reasonable accommodations in accordance with applicable laws. If you require accommodation during the hiring process, please contact accomodations@tensorwave.com.
Employment Eligibility
All offers of employment are contingent upon verification of identity and authorization to work in the United States, as required by law.
Background Checks
Where permitted by law, employment may be contingent upon the successful completion of a job-related background check.
Data Privacy Notice
By submitting an application, you acknowledge that TensorWave may collect, use, and retain your personal information for recruiting and employment-related purposes in accordance with applicable data privacy laws.