This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

Beacon AI
Posted 2w ago

Software Engineer, Cloud Infrastructure (Multiple Seniority Levels)

Beacon AI
San Carlos, California, United States
$135k-$260k/yrHybridFull Time
Responsibilities
  • designing infrastructure
  • operating services
  • building pipelines
Requirements
  • Experience building and operating AWS cloud and LLM infrastructure
  • CI/CD for infra and ML pipelines
  • Data pipelines
  • Security and observability for production systems
Technical tools mentioned
AWS CDKTerraformGitHub ActionsCodeBuildCodePipelineVPCPrivateLinkVPC endpointsIAMKMSSecrets ManagerCloudWatchOpenTelemetryAWS BedrockSageMakerECSEKSLambdaBatchLangChainStep FunctionsAirflowOpenSearch ServerlessAurora PostgreSQLpgvectorPineconeS3GlacierAthenaGlueLake FormationRedisEventBridgeSQSTriton Inference ServerTensorRT-LLMPython

Job description

About Beacon AI

We’re a fast-moving team of aviators, engineers, and operators building an AI platform to make flying safer, more efficient, and more capable. Backed by top investors, we’ve secured a dozen Department of Defense contracts and partnered with major airlines to deliver mission-critical systems. We operate without silos or heavy processes. Small, focused teams own what they build, ship quickly, and learn fast, pushing the boundaries of how humans and AI work together in aviation.

Role Overview

We are seeking skilled Cloud and ML Infrastructure Engineers to lead the buildout of our AWS foundation and our LLM platform. You will design, implement, and operate services that are scalable, reliable, and secure.

The broad scope means focus areas in LLM/ML Infra and IoT infra are strong bonus points. For ML Infra, build the stack that powers retrieval-augmented generation and application workflows built with frameworks like LangChain. Experience with IoT AWS services is a plus.

You will work closely with other engineers and product management. The ideal candidate is hands-on, comfortable with ambiguity, and excited to build from first principles.

Key Responsibilities

  • Cloud Infrastructure Setup and Maintenance

    • Design, provision, and maintain AWS infrastructure using IaC tools such as AWS CDK or Terraform.

    • Build CI/CD and testing for apps, infra, and ML pipelines using GitHub Actions, CodeBuild, and CodePipeline.

    • Operate secure networking with VPCs, PrivateLink, and VPC endpoints. Manage IAM, KMS, Secrets Manager, and audit logging.

  • LLM Platform and Runtime

    • Stand up and operate model endpoints using AWS Bedrock and/or SageMaker; evaluate when to use ECS/EKS, Lambda, or Batch for inference jobs.

    • Build and maintain application services that call LLMs through clean APIs, with streaming, batching, and backoff strategies.

    • Implement prompt and tool execution flows with LangChain or similar, including agent tools and function calling.

  • RAG Data Systems and Vector Search

    • Design chunking and embedding pipelines for documents, time series, and multimedia. Orchestrate with Step Functions or Airflow.

    • Operate vector search using OpenSearch Serverless, Aurora PostgreSQL with pgvector, or Pinecone. Tune recall, latency, and cost.

    • Build and maintain knowledge bases and data syncs from S3, Aurora, DynamoDB, and external sources.

  • Evaluation, Observability, and Cost Governance

    • Create offline and online eval harnesses for prompts, retrievers, and chains. Track quality, latency, and regression risk.

    • Instrument model and app telemetry with CloudWatch and OpenTelemetry. Build token usage and cost dashboards with budgets and alerts.

    • Add guardrails, rate limits, fallbacks, and provider routing for resilience.

  • Safety, Privacy, and Compliance

    • Implement PII detection and redaction, access controls, content filters, and human-in-the-loop review where needed.

    • Use Bedrock Guardrails or policy services to enforce safety standards. Maintain audit trails for regulated environments.

  • Data Pipeline Construction

    • Build ingestion and processing pipelines for structured, unstructured, and multimedia data. Ensure integrity, lineage, and cataloging with Glue and Lake Formation.

    • Optimize bulk data movement and storage in S3, Glacier, and tiered storage. Use Athena for ad-hoc analysis.

  • IoT Deployment Management

    • Manage infrastructure that deploys to and communicates with edge devices. Support secure messaging, identity, and over-the-air updates.

  • Analytics and Application Support

    • Partner with product and application teams to integrate retrieval services, embeddings, and LLM chains into user-facing features.

    • Provide expert troubleshooting for cloud and ML services with an emphasis on uptime and performance.

  • Performance Optimization

    • Tune retrieval quality, context window use, and caching with Redis or Bedrock Knowledge Bases.

    • Optimize inference with model selection, quantization where applicable, GPU/CPU instance choices, and autoscaling strategies.

What Will Make You Successful

  • End-to-End Ownership: Drives work from design through production, including on-call and continuous improvement.

  • LLM Systems Experience: Shipped or operated LLM-powered applications in production. Familiar with RAG design, prompt versioning, and chain orchestration using LangChain or similar.

  • AWS Depth: Strong with core AWS services such as VPC, IAM, KMS, CloudWatch, S3, ECS/EKS, Lambda, Step Functions, Bedrock, and SageMaker.

  • Data Engineering Skills: Comfortable building ingestion and transformation pipelines in Python. Familiar with Glue, Athena, and event-driven patterns using EventBridge and SQS.

  • Security Mindset: Applies least privilege, secrets management, network isolation, and compliance practices appropriate to sensitive data.

  • Evaluation and Metrics: Uses quantitative evals, A/B testing, and live metrics to guide improvements.

  • Clear Communication: Explains tradeoffs and aligns partners across product, security, and application engineering.

Bonus Points

  • 4+ years working with serverless or container platforms on AWS.

  • Experience with vector databases, OpenSearch, or pgvector at scale.

  • Hands-on with Bedrock Guardrails, Knowledge Bases, or custom policy engines.

  • Familiarity with GPU workloads, Triton Inference Server, or TensorRT-LLM.

  • Experience with big data tools for large-scale processing and search.

  • Background in aviation data or other safety-critical domains.

  • DevOps or DevSecOps experience automating CI/CD for ML and app services.

Work Location
This is a hybrid role based in San Carlos, CA, with 3+ days per week onsite and the option to work remotely on remaining days.

Perks & Benefits (Full-Time Employees)

  • Healthcare: 100%* of employee medical premiums covered; 25% for dependents

  • Time Off: 3 weeks PTO plus 13+ paid company holidays

  • Stipends: Monthly phone and wellness benefits

  • 401(k): Offered (no current employer match, but we are committed to enhancing this benefit in the future).

Due to U.S. export control regulations, we can only hire U.S. Persons (U.S. citizens, Green Card holders, lawful permanent residents, or individuals granted asylum or refugee status). We are unable to provide visa sponsorship or support visa transfers. All work must be performed in the United States.


Beacon AI is an equal opportunity employer and does not discriminate based on race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected characteristic. We prohibit harassment or discrimination of any kind in the workplace and comply with all applicable federal, state, and local employment laws.

About Beacon AI

Developing an AI-powered-pilot for safer flight operations.

Year founded
2021
Employees
21
Organization type
Private
Latest investment
Raised $15.00M Series A (2024) — led by Costanoa Ventures
Headquarters
US

Similar jobs

Software Engineer roles near San Carlos, California
6h
Save
Mark Applied
Hide
Sr. / Staff Software Engineer, Infrastructure (Autonomy)
San Jose or Mountain View
$170k-$339k/yr OnsiteFull Time
DiDi Autonomous Driving
DiDi Autonomous Driving: Develops Level 4 autonomous driving technology for shared mobility.
6+ YOEBachelor’s degree or higher in a technical field; 6–10+ years developing complex real-time or Linux infrastructure systems in C++; expertise in concurrency, IPC, compilers, build systems, debugging, and architecture.
C++, Linux, ccatch, ezsim, Clang, GCC, LLVM, Bazel, CMake, QNX, Linux RT, NVIDIA Orin, eBPF, gperf
7h
Save
Mark Applied
Hide
Senior Software Engineer, Google Health
Mountain View, California, United States
$174k-$252k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
5+ YOEBachelor's degree or equivalent practical experience; 5+ years with general-purpose programming languages and data structures and algorithms; experience with AI agents, software testing, and scalable systems.
Java, C, C++, Python, Go
7h
Save
Mark Applied
Hide
Senior Software Engineer (Semiconductor Process and Device Simulation)
San Francisco or Santa Clara or Hsinchu or Sapporo or Tokyo or Seoul
$90k-$162k/yr HybridFull Time
Siemens
SiemensXETRA: SIE: Manufactures industrial automation, infrastructure, and energy technology systems.
5+ YOEPhD or MS with 5+ years of research or industry experience in computer science, applied mathematics, or engineering; expertise in scientific computing, high-performance computing, and C/C++, CUDA, and Python.
C, C++, CUDA, Python, Code Copilot, AI/ML, Message Passing Interface (MPI)
11h
Save
Mark Applied
Hide
Staff Software Engineer, Data Engineering
London or San Francisco
HybridFull Time
Ripple
Ripple: Provides blockchain solutions for global payments and liquidity.
10+ YOE10+ years of data engineering experience; mastery of Databricks, Delta Live Tables, Unity Catalog, Delta Lake, and Spark; expertise in SQL, Python, and AWS; experience with AI data engineering tooling.
Databricks, Delta Live Tables, Unity Catalog, Delta Lake, Spark, SQL, Python, AWS, AI, LLMs
12h
Save
Mark Applied
Hide
Senior Software Engineer (Semiconductor Process and Device Simulation)
San Francisco or Hsinchu or Hokkaido or Tokyo or Seoul
$90k-$162k/yr HybridFull Time
Siemens Healthineers
Siemens HealthineersXetra: SHL: Global provider of medical technology and diagnostic imaging solutions.
5+ YOEPhD, or MS with 5+ years of research or industry experience, in computer science, applied mathematics, or engineering. Requires scientific computing, HPC, C/C++, CUDA, Python, and collaborative problem-solving expertise.
C, C++, CUDA, Python, Code Copilot, Message Passing Interface (MPI), AI, ML, Calibre, EDA, CAD, FEM, FD, FVM, BEM, Monte Carlo
14h
Save
Mark Applied
Hide
Senior Software Engineer, Developer Experience
Santa Clara, California, United States
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOEBachelor's or master's degree in computer science or related field; 5+ years building or supporting large software projects; Python, Java or Go; databases, CI/CD tools, AI, APIs and data analysis.
Python, Java, Go, SQL, NoSQL, MySQL, MongoDB, Elasticsearch, Jenkins, GitLab CI, Packer, Terraform, Artifactory, Ansible, Chef, Make, Maven, Ant, RESTful APIs, Large Language Models (LLMs), Machine Learning (ML), MCP
14h
Save
Mark Applied
Hide
Software Engineer - FPGA-Accelerated RTL Simulation
Berkeley, California, United States
$179k-$219k/yr OnsiteFull Time
SiFive
SiFive: Designs and licenses high-performance RISC-V processor intellectual property.
7+ YOEBS, MS, or PhD in CS, CE, or EE; 7+ years building complete modern C++ applications; strong C++ concurrency and multithreaded application experience; strong oral and written communication.
C++, C++20, Chisel, Xilinx FPGAs, MLIR, CIRCT, LLVM, RISC-V, FireSim
14h
Save
Mark Applied
Hide
Staff Software Engineer - Video Performance - (Bay area only)
San Francisco, California, United States
$251k-$329k/yr OnsiteFull Time
Canva
Canva: Cloud-based visual communication and graphic design software platform.
Strong C++ proficiency; systems performance optimization, CPU/GPU architecture, SIMD, graphics APIs, multimedia codecs, profiling, diagnostics, telemetry, and cross-team technical collaboration experience.
C++, Rust, GLSL, HLSL, Metal, Vulkan, WebGPU, OpenGL, Perf, Instruments, Chrome DevTools, Systrace, CMake, Web, iOS, Android, H.264, H.265, VP9, AV1, SIMD
This job has expired