The Senior Network Engineer will operate, maintain, and automate large-scale data center networks supporting AI and GPU workloads. Responsibilities include participating in on-call rotations, troubleshooting connectivity issues, and building automation tools to streamline network operations.
TensorWave
9 Remote Job Openings at TensorWave
Lead the design, technical validation, and optimization of advanced thermal management systems for hyperscale data centers. Design cooling infrastructure and perform CFD modeling to ensure efficient thermal containment for AI compute hardware.
You will own the end-to-end zero-touch provisioning and automation platform for GPU network fabrics while leading a team of engineers and SREs. Your role involves building intent-based configuration generation, streaming telemetry pipelines, and ensuring software engineering rigor across all network operations.
You will own the front-end network architecture for large-scale AI and GPU-accelerated infrastructure, including DCI, edge, and control-plane networks. You will also lead hands-on deployment, validation, and troubleshooting while collaborating with cross-functional teams to deliver scalable network solutions.
The engineer will perform detailed design reviews and technical validation of high-power electrical distribution systems for AI compute infrastructure. They will also provide subject matter expertise during construction, commissioning, and operation to ensure system reliability and safety compliance.
The Director of Construction will oversee the global data center construction portfolio, managing multi-billion-dollar capital expenditure budgets and delivery pipelines. They are responsible for developing construction strategy, managing vendor relationships, and ensuring all projects meet technical and quality standards for AI compute operations.
The Senior Data Center Design Manager leads the planning, design, and coordination of mission-critical data center projects from site selection through construction. This role acts as the primary liaison between multidisciplinary teams and external consultants to ensure projects meet technical standards, budget, and schedule requirements.
Serve as the primary technical escalation point for complex infrastructure issues and bridge the gap between customer needs and engineering development. Drive root cause analysis for P1 incidents and translate recurring technical blockers into actionable product roadmap improvements.
Architect and develop high-performance customer and internal dashboards for real-time GPU telemetry using TypeScript and React. Establish frontend engineering standards and collaborate with backend engineers and designers to ship scalable, accessible interfaces.