Northland PowerToronto Stock Exchange: NPI: Global developer and operator of sustainable energy infrastructure.
Post-secondary degree in electrical engineering; valid G-class driver's license; advanced Excel and PowerPoint; knowledge of reliability engineering in power systems.
Microsoft Excel, PowerPoint, Relays and protection systems
iManage: Intelligent document and email management software for professionals.
Experience in reliability engineering with automation, cloud platforms, observability, and on-call responsibility; strong collaboration and architectural skills.
5+ YOE5+ years IT experience (or equivalent), proficiency in Python/Golang, Linux administration, infrastructure-as-code, networking and identity/auth systems, Agile experience, and strong problem-solving and collaboration skills.
Python, Golang, Linux, Kubernetes, ECS, CI/CD, Docker, Istio, Envoy, Consul, Mesos/Marathon, Amazon Web Services, EC2, RDS, Dynamo DB, Route53, Elastic Load Balancers, AMIs, IAM Roles, Ops Works, Cloud Formation
Domino Data Lab: Enterprise MLOps platform for developing and managing AI models.
Deep SRE/platform engineering experience, hands-on operational ownership, fluency with Kubernetes and Linux, strong software engineering skills in Python or Go, observability and SLO/SLI expertise, mentoring and cross-team influence.
SimCorp: Provides integrated software solutions for investment and asset managers.
3+ YOE3+ years in Site Reliability, DevOps, or Cloud Engineering; Azure expertise; IaC with Bicep/ARM/Terraform; monitoring/logging tools; IdP onboarding and security; Kubernetes/Docker; ITIL familiarity.
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
15+ YOE15+ years in SRE/infrastructure/platform engineering with large-scale fleets; experience in capacity management, orchestration, observability, SLOs/SLIs, incident response, and cross-team architecture.
5+ YOE5+ years SRE/DevOps experience, Bachelor’s in CS/Engineering or equivalent, strong AWS and IaC (Terraform/CDK/CloudFormation) skills, CI/CD and container expertise, Python/Bash scripting, monitoring and SRE practices, on-call experience.
Boson AI: Develops generative AI and large-scale audio foundation models.
4+ YOE4+ years in SRE or related production-operations; deep expertise in networking, cluster scheduling, storage, GPU systems or AI infra; strong Linux and scripting; production availability, performance, security, and automation focus.
Reliability & Process Engineer; Production/Manufacturing; Mechanical Engineer, Electrical; On Site Amherst, NY
Amherst, New York, United States
$92k-$142k/yrOnsiteFull Time
Saint-GobainEuronext Paris: SGO: Global leader in light and sustainable construction and materials.
5+ YOEBachelor's in engineering required; 5–7 years manufacturing/maintenance or reliability engineering experience; experience leading maintenance teams, troubleshooting industrial equipment, SCADA/MES familiarity, and Microsoft Office proficiency.
SCADA, MES, Ignition, Microsoft Excel, Microsoft Word, Microsoft PowerPoint
Dominion Dynamics: Developing autonomous defense platforms and Arctic sensing technology.
Senior-level experience running Linux on embedded/edge devices, DDIL design, ARM/Jetson/embedded toolchains, networking fundamentals, secure boot/attestation and provisioning pipelines; ability to define reliability and observability for constrained field devices.
Linux, SELinux, AuraNet SDK, ROS, TPM, Jetson, ARM
5+ YOEExpert SRE experience with observability, incident management, Terraform, Azure, automated testing, and 5+ years systems/application experience; strong troubleshooting and communication skills.
ScotiabankToronto Stock Exchange: BNS: Provides global personal, commercial, and investment banking services.
3+ YOE3+ years experience in ETL platforms and application support, Unix shell scripting, Java, and SQL. Experience with observability tools, incident management, automation, IaC, and on-call rotations. Undergraduate degree in CS or equivalent required.
Electronic ArtsNASDAQ: EA: Develops and publishes video games and interactive entertainment software.
7+ YOE7+ years experience with cloud, containers, virtualization, Linux, automation and distributed systems; strong scripting/programming in Python, Golang or Java; experience with Terraform, Helm, Chef, Puppet, Packer and Kubernetes.
Staff Site Reliability Engineer - Confluent Incident Management & Reliability
Markham or Toronto
$134k-$248k/yrRemoteFull Time
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
10+ YOE10+ years SRE/incident management experience, cloud experience (AWS, GCP, or Azure), deep incident tooling knowledge (Rootly, PagerDuty), Kubernetes and observability expertise, strong communication and coaching skills.
Reliability & Process Engineer; Production/Manufacturing; Mechanical Engineer, Electrical; On Site Amherst, NY
Amherst, New York, United States
$92k-$142k/yrOnsiteFull Time
Saint-GobainEuronext Paris: SGO: Global leader in light and sustainable construction solutions.
5+ YOEBachelor's degree in engineering, 5+ years manufacturing/maintenance or reliability engineering experience, leadership of maintenance teams, project/CAPEX management, SCADA/MES (Ignition) familiarity, and proficiency with Microsoft Office.
SCADA, MES, Ignition, Microsoft Excel, Microsoft Word, Microsoft PowerPoint