116 monitoring engineer jobs at 57 companies in Prunedale, CA

1mo
Save
Mark Applied
Hide
Software Development Engineer, Network Monitoring & Alerts - San Jose
San Jose, California, United States
$213k-$450k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
5+ YOEBachelor's degree in CS/EE or above,5+ years in network monitoring/alert systems,proficiency in Java/C++/Python,deep TCP/IP knowledge,EMR? no,ability to develop monitoring,alerting and network-diagnostics solutions.
Java, C++, Python, Telemetry, SNMP, Syslog, sFlow, Erspan, INT, TCP/IP
4d
Save
Mark Applied
Hide
Security Engineer Cybersecurity Monitoring - Secret
Alexandria or Seaside
$94k-$127k/yr OnsiteFull Time
VMD Corp
VMD Corp: Provides airport security screening and federal IT services.
5+ YOEBachelor's degree and five years of experience required. U.S. citizenship and active Secret clearance required; cybersecurity monitoring, security scanning, compliance, troubleshooting, and security engineering expertise needed.
SIPRNet
2mo
Save
Mark Applied
Hide
Software Engineer, Fleet Monitoring
Mountain View or San Francisco
$175k-$215k/yr HybridFull Time
Waymo
Waymo: Autonomous driving technology for ride-hailing and logistics.
5+ YOE5 years full-stack experience with Angular/React/TypeScript and backend languages.
Angular, React, TypeScript, HTML, CSS, C++, Java, Kotlin, Python, Go
2w
Save
Mark Applied
Hide
System Development Engineer, AWS Manufacturing Infrastructure Services
Cupertino, California, United States
$149k-$201k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
3+ YOE3+ years software and systems engineering experience; proficiency in modern languages, Linux, networking, automation, enterprise-scale infrastructure, and monitoring; strong troubleshooting and architecture skills.
C++, C#, Java, Python, Golang, PowerShell, Ruby, Perl, PHP, Bash, Shell, EC2, VPC, IAM, Outposts, CloudWatch, S3, Prometheus, Grafana, PXE, IPMI, BMC
5d
Save
Mark Applied
Hide
ASIC Design and Integration Engineer – Silicon Security
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Design secure ASIC subsystems and hardware defenses for SoC security, including cryptographic engines and side-channel attack monitors.
3w
Save
Mark Applied
Hide
Senior Platform Engineer
Austin or Sunnyvale
$158k-$205k/yr OnsiteFull Time
GFiber: An Alphabet providing Google Fiber and Webpass internet services to homes and businesses.
5+ YOE5+ years public cloud and infrastructure experience with GCP, Terraform, Linux, Kubernetes, CI/CD; 2+ years coding with Python or Java; experience with Netcracker/Alepo/Axiros and monitoring stacks.
Netcracker, Alepo, Axiros, GCP, Terraform, Linux, Kubernetes, VPC, Cloud Run, Ansible, SaltStack, Python, Java, GCP Monitoring, Prometheus, Datadog, Dynatrace, ELK, Gemini, Vertex AI, ADK
5d
Save
Mark Applied
Hide
ZfG Operations Engineer
San Jose, California, United States
$99k-$229k/yr HybridFull Time
Zoom
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
Proven experience in site reliability, DevOps, or infrastructure operations; proficiency with Python, Go, or Bash; knowledge of distributed systems, monitoring/observability, incident management, and automation; experience supporting production systems and on-call rotations.
Python, Go, Bash, BrightHire
1mo
Save
Mark Applied
Hide
Sr. Facilities Engineer
Milpitas, California, United States
$122k-$207k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management for semiconductor manufacturing.
8+ YOEManage cleanroom facilities operations and preventative maintenance, coordinate tool installations, ensure EHS compliance, troubleshoot MEP systems, monitor BMS, and support after-hours responses. Requires 8+ years relevant experience with 5+ years in semiconductor/cleanroom.
BMS, SPC, SEMI, ECN, SOP, RO, ID, 6S
2mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Santa Clara, California, United States
$152k-$245k/yr OnsiteFull Time
Palo Alto Networks
Palo Alto NetworksNASDAQ: PANW: Provides enterprise-grade network, cloud, and endpoint security software.
BS or MS in CS or related field; expertise in configuration management (Ansible, Terraform, Kubernetes); Python and/or Go; Kubernetes with autoscaling; production engineering/DevOps/SRE experience; public cloud (GCP/AWS); Linux networking; CI/CD with GitLab/GitHub; distributed systems; strong communication; ownership and monitoring as code.
Kubernetes, Docker, GCP, AWS, Ansible, Terraform, Vault, GitLab, Spinnaker, Pub/Sub, Bigtable, Memorystore, BigQuery, RabbitMQ, Kafka, MySQL, Python, Go, Shell scripting, Golang
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale, California, United States
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Manufactures medical devices, diagnostics, and nutritional health products.
Ensure reliability, scalability, and performance of a medical-device remote monitoring platform; expertise in cloud (Azure), Kubernetes, observability, automation, and incident management; bachelor's in a technical discipline.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK/EFK, Datadog, Linux
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale or Sylmar
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
Senior SRE with strong distributed systems, cloud (Azure), Kubernetes, observability, automation, incident management, and cross-functional communication skills for a medical device remote monitoring platform.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK, EFK, Datadog, Linux
1mo
Save
Mark Applied
Hide
Sr Infrastructure Capacity Engineer
Seattle or San Jose or Virginia or Texas or Massachusetts or United States
$162k-$242k/yr RemoteFull Time
F5
F5NASDAQ: FFIV: Provides application delivery networking and multi-cloud security solutions.
8+ YOE8+ years in infrastructure capacity planning or related discipline; hyperscale/SaaS experience; deep compute architecture and forecasting skills; proficiency with monitoring tools and Python/SQL/Excel; strong executive communication.
Prometheus, Grafana, Datadog, Python, SQL, Microsoft Excel, Microsoft PowerPoint
1mo
Save
Mark Applied
Hide
Software Engineer 6 - Creative Studio
Los Gatos or New York
$499k-$820k/yr OnsiteFull Time
Netflix
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
9+ YOE9+ years building highly available, scalable backend services; expertise in microservices, data modeling, cloud, API design (REST/gRPC), async processing, monitoring, and mentoring; experience with Generative AI/ML integrations preferred.
Java, Go, Python, REST, gRPC, Generative AI, machine learning platforms
2mo
Save
Mark Applied
Hide
Principal Hardware Diagnostics Engineer
Austin or Milpitas
OnsiteFull Time
Graphcore
Graphcore: Produces specialized processors for artificial intelligence workloads.
Bachelor's/Master's/PhD in CS/CE or related; strong Python/C++/C#; Linux; experience with diagnostics and monitoring systems; distributed systems; collaboration with hardware and firmware teams.
Python, C++, C#, Linux, Diagnostics tools
1mo
Save
Mark Applied
Hide
Senior Systems Reliability Engineer II
Mountain View, California, United States
HybridFull Time
ThoughtSpot
ThoughtSpot: AI-powered analytics platform for enterprise business intelligence.
Experience troubleshooting Linux systems and cloud platforms, hands-on with monitoring tools, on-call/incident management experience, scripting in Python/Go/Bash/Java, B.S. in CS or equivalent preferred.
Grafana, Prometheus, Datadog, Splunk, VMware, AWS, Azure, GCP, Python, Go, Bash, Java, Spotter, SpotterViz
1mo
Save
Mark Applied
Hide
Senior Server Operations & Maintenance Engineer
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
3+ YOEManage server lifecycle, design operation/maintenance solutions, implement hardware monitoring and troubleshooting, collaborate with R&D, and support online quality and standardization.
BMC, IPMI, SMBIOS, Redfish, SEL, BMC OneKeyLog, Linux, Shell, Python, PHP, Perl, Lua, AIOps, PCIe, NVMe, x86, ARM
3w
Save
Mark Applied
Hide
Senior Storage Production Engineer - DGX Cloud
Santa Clara or United States
$176k-$334k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEBS or equivalent and 8+ years experience in distributed/high-performance storage, expertise with storage protocols and automation, programming and monitoring skills, and strong communication and troubleshooting abilities.
Kubernetes, OpenStack, C/C++, Java, Python, Go, NodeJS, Bash, Ansible, Chef, Puppet, Terraform, InfluxDB, Prometheus, Grafana, Elastic Stack, Git, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - HPC
Santa Clara or Austin or Durham
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years building/supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes); IaC and CI/CD proficiency; coding in Python/Go/Perl/Ruby; monitoring, capacity planning, and incident response skills.
Slurm, LSF, Kubernetes, Infrastructure as Code (IaC), CI/CD, AWS, GCP, OCI, Python, Go, Perl, Ruby
6d
Save
Mark Applied
Hide
Staff Data Center Solutions Engineer
San Jose, California, United States
$165k-$200k/yr FieldFull Time
Supermicro
SupermicroNASDAQ: SMCI: Designs and manufactures high-performance server and storage solutions.
12+ YOE12+ years datacenter infrastructure experience, strong GPU cluster and NeoCloud knowledge, BMS/infrastructure monitoring familiarity, understanding of cooling and power systems, excellent communication and leadership skills.
NeoCloud, Agentic AI, NVIDIA, CUDA, Hyperview, Grafana, PagerDuty, ServiceNow CSM, InfiniBand, RoCEv2
1mo
Save
Mark Applied
Hide
Senior AWS AgentCore Platform Engineer/Architect IRC292142
San Jose or Reading
$130k-$140k/yr HybridFull Time
GlobalLogic
GlobalLogic: Digital product engineering and software development services provider.
10+ YOE10+ years experience building and hardening AWS platform components for AI agents; expertise in observability, cost tracking, monitoring, IAM/ABAC, Terraform, Python, and security.
AWS, AWS Bedrock, AWS CloudTrail, AWS CloudWatch, AWS IAM, MCP, Python, Strands Agents, Terraform, Dynatrace, AWS X-Ray, LangFuse, LiteLLM, Microsoft Teams, Confluence, AWS Budgets, Cedar