Beacon AI
Posted 2w ago

Software Engineer, Artificial Intelligence/LLM (Multiple Seniority Levels)

Beacon AI
San Carlos, California, United States
$135k-$260k/yrHybridFull Time
Responsibilities
  • building features
  • integrating models
  • monitoring metrics
Requirements
  • Experience building LLM-powered features
  • RAG and tool-calling
  • Production services in Python or TypeScript
  • Vector search and embeddings
  • Evals/metrics, and safety/compliance for a regulated domain
Technical tools mentioned
LangChainPythonTypeScriptAWS BedrockOpenAIAnthropicOpenSearchpgvectorPineconeWeaviateS3AuroraDynamoDBTritonTensorRT-LLM

Job description

About Beacon AI

We’re a fast-moving team of aviators, engineers, and operators building an AI platform to make flying safer, more efficient, and more capable. Backed by top investors, we’ve secured a dozen Department of Defense contracts and partnered with major airlines to deliver mission-critical systems. We operate without silos or heavy processes. Small, focused teams own what they build, ship quickly, and learn fast, pushing the boundaries of how humans and AI work together in aviation.

You will ship LLM-powered product features end-to-end. That means designing retrieval and tool-calling flows, writing the services that run them, building evals and guardrails, and watching cost, latency, and quality in production. You’ll partner with the ML/infra teammates on embeddings, indexing, and model hosting, and with the product teammates on user experience and outcomes. We move fast, and we care about reliability in a safety-critical domain.

We’re hiring across levels. Senior engineers own features and services. Staff engineers own systems, standards, and cross-team technical direction.

What you’ll do

Build user-facing LLM features

  • Design and implement retrieval-augmented generation and tool-calling flows using frameworks like LangChain or equivalent primitives, where simpler is better.

  • Deliver robust JSON and schema-bound outputs with validation, retries, and fallbacks.

  • Add function calling to integrate with internal tools, search, routing, and data services.

Own the service layer

  • Ship APIs and workers in Python or TypeScript with clear contracts, streaming, and backoff.

  • Add caching, request shaping, prompt templates, and context packing to control latency and cost.

  • Integrate with AWS Bedrock, OpenAI, Anthropic, or self-hosted endpoints as needed.

Retrieval and data prep

  • Collaborate with infrastructure teammates to develop chunking, embeddings, and indexing capabilities for documents, time series, and multimedia.

  • Choose and tune vector backends such as OpenSearch, pgvector, or Pinecone.

  • Keep knowledge bases fresh with data syncs from S3, Aurora, DynamoDB, and external sources.

Evaluation and quality

  • Create offline evals and golden sets for prompts, retrievers, and tools.

  • Stand up online metrics for task success, hallucination rate, retrieval precision/recall, p95 latency, and cost per request.

  • Run A/B tests and prompt/version rollouts with guardrails and canaries.

Safety, privacy, and compliance

  • Implement content and policy checks, PII detection and redaction, access controls, and auditing.

  • Design human-in-the-loop paths for sensitive actions.

  • Handle aviation data with care and follow internal security standards.

Operate what you build

  • Add tracing, logs, and dashboards for model calls, token usage, errors, and saturation.

  • Debug tricky failures across retrieval, prompts, tools, and providers.

What will make you successful

  • Shipped LLM apps: You’ve put LLM features in front of users and improved them with data.

  • Strong builder: Comfortable writing production code, tests, and docs. You keep things simple and observable.

  • RAG and tools depth: You understand embeddings, chunking, vector search tradeoffs, and function calling.

  • Quality mindset: You design evals, define success metrics, and iterate based on evidence.

  • Cost and latency aware: You track p95, hit SLAs, and reduce cost without hurting quality.

  • Clear communicator: You explain tradeoffs and align partners across product, infra, and security.

Nice to have

  • Experience with Bedrock, OpenSearch Serverless, pgvector, Pinecone, or Weaviate.

  • Prompt versioning, guardrails, and provider routing in production.

  • Multimodal work with time series or video.

  • Familiarity with GPU inference, Triton, or TensorRT-LLM.

  • Aviation or other safety-critical domain exposure.

  • DevOps basics for CI/CD, IaC, and secure secrets handling.

Example problems you might tackle in month one

  • Transform an internal knowledge base into a low-latency RAG service, complete with explicit schemas and evaluations.

  • Add tool-calling to automate a repetitive cockpit or ops workflow with guardrails and audit trails.

  • Reduce the cost per request through improved chunking, caching, and prompt refactoring, while maintaining task success rates.

Work Location
This is a hybrid role based in San Carlos, CA, with 3+ days per week onsite and the option to work remotely on remaining days.

Perks & Benefits (Full-Time Employees)

  • Healthcare: 100%* of employee medical premiums covered; 25% for dependents

  • Time Off: 3 weeks PTO plus 13+ paid company holidays

  • Stipends: Monthly phone and wellness benefits

  • 401(k): Offered (no current employer match, but we are committed to enhancing this benefit in the future).

Due to U.S. export control regulations, we can only hire U.S. Persons (U.S. citizens, Green Card holders, lawful permanent residents, or individuals granted asylum or refugee status). We are unable to provide visa sponsorship or support visa transfers. All work must be performed in the United States.


Beacon AI is an equal opportunity employer and does not discriminate based on race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected characteristic. We prohibit harassment or discrimination of any kind in the workplace and comply with all applicable federal, state, and local employment laws.

About Beacon AI

Developing an AI-powered-pilot for safer flight operations.

Year founded
2021
Employees
21
Organization type
Private
Latest investment
Raised $15.00M Series A (2024) — led by Costanoa Ventures
Headquarters
US

Similar jobs

Software Engineer roles near San Carlos, California
2h
Save
Mark Applied
Hide
Senior Software Engineer, Apple Services Engineering
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Build distributed, large-scale data processing systems, frameworks, and platforms using big data technologies while collaborating with Apple TV and Video teams.
3h
Save
Mark Applied
Hide
Senior Software Engineer, Hyperscale Build Environments and Tools
Santa Clara, California, United States
$180k-$270k/yr OnsiteFull Time
Pure Storage
Pure StorageNYSE: PSTG: Provides all-flash enterprise data storage and management solutions.
5+ YOE5+ years of software engineering experience in infrastructure, developer productivity, build systems, or platform engineering, with C/C++, Linux, Docker, CI, and modern build-system expertise.
C, C++, Make, CMake, Linux, Docker
3h
Save
Mark Applied
Hide
Software Engineer, Linux Kernel / Android
San Jose or Los Angeles or Bellevue
$180k-$240k/yr OnsiteFull Time
Rivet Industries
Rivet Industries: Building integrated task systems for frontline industrial and defense operators.
3+ YOERequires 3+ years developing Android/Linux system software, strong Linux kernel and C/C++ skills, AOSP, HALs, drivers, hardware bring-up, embedded builds, debugging, security mechanisms, and U.S. Person status.
Android, Linux, Linux kernel, AOSP, HALs, C, C++, USB, MIPI, I2C, UART, GPIO, PCIe, U-Boot, Android Bootloader, Yocto, Buildroot, Bazel, Soong, OTA, TPM, AR/XR, NPU, DSP
3h
Save
Mark Applied
Hide
Principal Staff Software Engineer, Systems Infrastructure
Mountain View, California, United States
$226k-$369k/yr HybridFull Time
LinkedInNASDAQ: MSFT: Professional social network for career development and job recruitment.
10+ YOEBA/BS or equivalent practical experience, 10+ years in software or reliability engineering, 5+ years in technical leadership, distributed systems expertise, and experience defining reliability standards across teams.
Java, Go, C++, Python, LLM, SLO, SLI
5h
Save
Mark Applied
Hide
Exceptional Software Engineer
Redwood City, California, United States
$180k-$400k/yr OnsiteFull Time
Dyna Robotics
Dyna Robotics: Develops general-purpose robots powered by proprietary embodied AI models.
Exceptional software engineering ability, high tolerance for ambiguity, resilience in changing environments, and clear communication. Robotics experience is welcome but not required.
6h
Save
Mark Applied
Hide
Staff+ Software Engineer, Product Sandboxing
San Francisco or New York City or Seattle or California
$405k-$485k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
8+ YOERequires 8+ years building scalable distributed systems, strong service-oriented architecture, networking, and systems design expertise, proficiency in Python, Go, or Rust, and experience with cloud infrastructure and Kubernetes.
Python, Go, Rust, GCP, AWS, Azure, Kubernetes
7h
Save
Mark Applied
Hide
Staff Software Engineer, Middle Office
San Francisco, California, United States
$240k-$300k/yr OnsiteFull Time
Carta
Carta: Software platform for cap table management and fund administration.
10+ YOEExpertise in distributed systems and 10+ years of professional software engineering experience recommended, with high-level technical leadership and experience guiding architecture across Python/Django, React, Postgres, JVM languages, gRPC, and AWS.
Python, Django, React, Postgres, gRPC, AWS
8h
Save
Mark Applied
Hide
Staff Software Engineer, MetalDev
New York City or Sunnyvale
$207k-$275k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
8+ YOE8+ years in software engineering focused on infrastructure, cloud engineering, and distributed systems; expert Go, REST/gRPC APIs, Kubernetes, observability, CI/CD, GPU fleets, technical leadership, and incident response.
Go, REST, gRPC, Kubernetes, Prometheus, Grafana, PromQL, CI/CD, Kafka, ClickHouse, CRDB, DMTF, RedFish