39 network performance engineer jobs at 35 companies in San Rafael, CA
1w
Save
Mark Applied
Hide
1w
AI/HPC Network Performance Engineer
Menlo Park, California, United States
$184k-$257k/yrOnsiteFull Time
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
8+ YOEBachelor's or equivalent,8+ years in system or network performance engineering for large-scale distributed/HPC environments; experience with datacenter networks, network automation, and coding in Python,C++,Go.
Gimlet Labs: Infrastructure for scaling and deploying high-performance agentic AI workloads.
Design, deploy, and operate production network infrastructure for high-performance AI compute; strong data center networking and automation experience.
Hewlett Packard EnterpriseNYSE: HPE: Providing global edge-to-cloud infrastructure and IT solutions for businesses.
12+ YOEBachelor's or Master's in CS/ECE required, 12+ years networking/systems product development, expert C programming, data structures, networking protocols knowledge, debugging and performance optimization skills.
C, C++, DevOps, Microservices, Full Stack Development
RakutenTokyo Stock Exchange: 4755: Provides global e-commerce, financial services, and telecommunications technology.
5+ YOEDesign, implement, and manage enterprise network infrastructure across LAN/WAN, data center, and cloud with emphasis on security, performance, and automation.
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
5+ YOE5+ years in large-scale enterprise/SRE networking; expert in high-performance networking, cloud connectivity, IaC, and network automation; bachelor’s degree; professional certifications.
Cresset: Independent multi-family office and private investment firm.
6+ YOE6+ years IT infrastructure experience, hands-on Cisco Meraki, strong TCP/IP and networking knowledge, Microsoft 365/Entra/Intune and Windows Server experience, troubleshooting and documentation skills, ability to perform off-hours and travel as needed.
Cisco Meraki, Meraki Dashboard, Microsoft 365, Entra ID, Azure, Intune, Windows Server, PowerShell, Python, APIs
OpenAI: Develops artificial intelligence models and generative AI software services.
Design, build, and operate networking systems for large-scale AI training; improve performance and reliability; develop automation and observability tooling.
12+ YOEBachelor's or higher in CS (or equivalent),12+ years engineering experience with senior technical leadership, strong C/C++ on Unix, networking and security knowledge, mentoring and performance optimization experience.
Dedalus Labs: Infrastructure for building and deploying AI agent applications.
Fluency in Rust/Go/C/C++; strong software engineering fundamentals and systems knowledge (OS, networking, distributed systems); performance-oriented debugging and reliability focus.
Eridan: Building power-efficient 5G transceivers using gallium nitride technology.
5+ YOE5+ years telecom/network/systems engineering (3+ considered with RAN experience). BS in EE/CS/Telecom or equivalent, fluency with Linux, RAN architecture knowledge, debugging and networking skills, ability to travel and perform lab/field testing.
Vast.ai: Decentralized marketplace for GPU cloud computing resources.
Expertise in systems and GPU engineering, GPU architectures, neural network performance, C++/CUDA/Python proficiency, and strong research background with publications preferred.
Woodinville or Palo Alto or New York or New Jersey
OnsiteContract
Redapt: Provides integrated data center infrastructure and cloud computing solutions.
5+ YOE5+ years infrastructure engineering experience with enterprise Microsoft Hyper-V, PowerShell automation, Windows Server administration, storage/networking performance analysis, infrastructure automation (DSC/SCCM/Ansible), and a technical bachelor's degree.
Microsoft Hyper-V, PowerShell, WMI, C#, .NET 8, Windows Server, PowerShell DSC, SCCM, Ansible, CI/CD, elbencho, VMFleet, diskspd, fio, Qualys, Microsoft System Center (VMM, SCOM, SCCM), VMware ESXi, KVM, Windows Server Failover Clustering (WSFC), Storage Spaces Direct (S2D), SET
E2B: Open-source cloud infrastructure for running autonomous AI agents.
5+ YOE5+ years building distributed systems; deep Linux/kernel and VM hypervisor expertise; systems programming in Go/Rust/C/C++; production orchestration (Kubernetes/Nomad); networking and performance optimization experience.
Luma AI: Develops multimodal AI for video generation and creative production.
5+ YOE5+ years SRE or infrastructure experience, deep Linux and low-level performance debugging, Terraform, Airflow, Ray, AWS or OCI, high-performance networking (InfiniBand/RDMA/RoCE), security/compliance familiarity.
5+ YOEBachelor's degree or equivalent; 5+ years software development; 3+ years testing/launching software; 3+ years infrastructure/distributed systems experience.