# Senior Network Engineer

Hiring organization: [Together AI](https://career.thegoodapps.co/organizations/together-ai)

Canonical page: https://career.thegoodapps.co/jobs/efe3a69a-740b-4da8-aefb-a15ea12c8d1f

Listed on Together AI's own careers site. Applications go to them directly.

- Employment type: full time
- Seniority: Senior
- Location: San Francisco, CA
- Salary: 190000 – 280000 USD per year

## Summary

Together AI is hiring a Senior Network Engineer to design and operate global infrastructure for their AI compute platforms, working across multiple data centers and vendor networks. This role suits experienced network engineers comfortable troubleshooting complex, large-scale systems and collaborating across teams to solve problems that span networking, infrastructure, and applications.

_Our summary, not Together AI's wording._

## Skills named

Amazon Web Services (AWS), Ansible, Border Gateway Protocol (BGP), Cisco, Git, Juniper, Kubernetes, Linux, Microsoft Azure, Nmap, NVIDIA Triton Inference Server, OSPF, Python, VLAN, Wireshark

## Required

- 8+ years designing, building, and supporting large-scale production data center, cloud, service-provider, or HPC networks
- Deep understanding of TCP/IP and experience with BGP, OSPF, VXLAN, EVPN, ECMP, and QoS
- Experience designing and supporting multi-tenant network environments using VRFs, VLANs, overlays, and policy-based segmentation
- Hands-on experience deploying and troubleshooting networks from Arista, Cisco, Juniper, or NVIDIA
- Strong troubleshooting with Wireshark, tcpdump, MTR, curl, nmap, and Linux networking utilities
- Ability to diagnose connectivity, latency, packet loss, routing, and performance issues across network, host, and application layers
- Experience developing or maintaining network automation using Python, Ansible, or similar tools
- Experience with Git-based software development lifecycle including branching, code review, CI/CD, and deployment
- Working knowledge of Kubernetes networking including pods, services, CNIs, and connectivity troubleshooting
- Foundational knowledge of RDMA networking and RoCE or InfiniBand
- Experience with cloud networking in AWS, GCP, or Azure
- Strong Linux administration and troubleshooting skills

## Nice to have

- Hands-on experience deploying or operating RoCE or InfiniBand fabrics
- Experience supporting GPU clusters, HPC environments, distributed storage, or high-bandwidth latency-sensitive workloads
- Understanding of AI training and inference traffic patterns and network demands
- Experience operating networks spanning thousands of devices across multiple data centers and geographic regions
- Familiarity with AI-assisted engineering tools and ability to validate and safely deploy AI-generated automation

Apply on Together AI's site: https://job-boards.greenhouse.io/togetherai/jobs/5180977007
