TensorWave Logo

TensorWave

Senior Network Engineer

Reposted 19 Days Ago
Remote
Hiring Remotely in USA
Senior level
Remote
Hiring Remotely in USA
Senior level
This role involves developing and managing a networking infrastructure for AI cloud services, integrating new technologies, ensuring network reliability, and collaborating with IT and AI teams.
The summary above was generated by AI

Our mission at Tensorwave Cloud is to build seamless, secure, reliable, and resilient AI infrastructure at scale, eliminating barriers and challenging the status quo to empower builders and support AI innovation.

About the role

We are seeking a Senior Network Engineer focused on implementing and operating large-scale, Arista-based RoCEv2 data center networks powering next generation AI and ML infrastructure.

You’ll work hand-in-hand with our network architect to design the infrastructure that keeps over 8,000 GPUs burring, and play a critical role in the implementation and maintenance of our next generation systems, with cluster sizes reaching over 100,000 GPUs.

You’ll work hands-on with high-speed optics, switching, and routing in production clusters and implement modern automation and tooling critical to how the network is deployed, validated, and operated.

Responsibilities

  • Design, deploy, and operate large-scale RoCEv2 data center networks supporting AI and ML clusters from thousands to 100,000+ GPUs

  • Own congestion management and performance tuning across RDMA fabrics, including PFC, ECN, and DCQCN, in production environments

  • Implement and maintain automation, validation, and observability tooling using Python, Ansible, Terraform, and modern DevOps workflows

  • Ensure high availability and reliability across multi-tenant environments by leading operational excellence, incident response, and continuous improvement

Required Experience

  • Bachelor of Science in Computer Science, Computer Engineering, or a related technical field, or equivalent practical experience

  • Deep experience with RDMA and RoCEv2 in large-scale production data centers supporting AI or HPC workloads

  • Strong Arista expertise, including EOS, hardware platforms, and operating high-speed Ethernet fabrics

  • Proven knowledge of congestion management and performance tuning using PFC, ECN, and DCQCN

  • Hands-on experience with high-speed optics and cabling including 400G, 800G, and AEC, AOC, DAC, and structured cabling in dense environments

  • Automation and operations mindset, with experience using Python, Ansible, Terraform, Git, and observability tooling in always-on production systems

What We Bring

  • Mission driven company

  • Competitive Salary

  • Stock Options

  • 100% paid Medical, Dental, and Vision insurance

  • Flexible PTO

  • Paid Holidays

  • 401(k)

  • Parental Leave

  • Flexible Spending Account

  • Short Term Disability Insurance

  • Life and Voluntary Supplemental Insurance

  • Mental Health Benefits through Spring Health

We’re looking for resilient, adaptable people to join our team, people who believe in the mission and think at massive scale. The solutions that worked on a handful of devices will not work at Exascale. Be prepared to be pushed daily, to learn a lot, and literally build the future.

Tensorwave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, national origin, or veteran status.

Top Skills

Amd Gpu
Bgp
Ethernet Protocols
Nvidia Gpu
Rocev2

Similar Jobs

8 Days Ago
Remote
United States
203K-274K Annually
Expert/Leader
203K-274K Annually
Expert/Leader
Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Responsible for managing Dropbox's cloud networking infrastructure and ensuring reliability, scalability, and security across multi-cloud environments while collaborating with various teams to drive improvements in automation and observability.
Top Skills: AnsibleAWSAzureDatadogGCPOciPythonTerraform
5 Hours Ago
Remote
US
119K-200K Annually
Senior level
119K-200K Annually
Senior level
Big Data • Machine Learning • Software • Analytics
The Sr. Network Engineer designs, implements, and operates cloud and hybrid network infrastructure, emphasizing automation and customer-centric solutions while supporting 24x7 operations.
Top Skills: AnsibleArista TechnologiesAWSAzureBashChefCi/CdGCPGitPowershellPuppetPythonTerraformWireshark
7 Days Ago
Remote
US
111K-198K Annually
Senior level
111K-198K Annually
Senior level
Cloud • Fintech • Insurance • Software
The Cloud Senior Network Engineer designs, builds, and supports networking in cloud environments, focusing on configuration automation and system reliability, with an emphasis on cross-team collaboration.
Top Skills: AWSAzureAzure DevopsBashBicepDirect ConnectExpressrouteF5GCPHyper-VNginxPowershellPythonTerraformVMwareVpns

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account