📋 External Recruiting Agencies

GinasTechJobs (also known as Next Step Systems) is a specialized IT recruiting and staffing agency that posts job openings on behalf of its various client companies.

This company was flagged and excluded from default search results. Proceed with caution.

Ginas Tech Jobs
Posted 1mo ago

Principal Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home

Ginas Tech Jobs
San Francisco, California, United States
$170k-$200k/yrRemoteFull Time
Responsibilities
  • architecting systems
  • training models
  • deploying models
Requirements
  • Deep learning and transformer expertise
  • Hands-on production ML experience
  • Distributed training and inference
  • GPU optimization
  • Strong software engineering and production deployment skills
Technical tools mentioned
PyTorchJAXDeepSpeedFSDPMegatronZeRORayvLLMTensorRT-LLMFasterTransformerApache ArrowSparkPPODPOORPO

Job description

Job Description:

Principal Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home

 

As a Principal Machine Learning Engineer, you are a deep technical authority responsible for designing and evolving the most critical ML systems in the company.  The Principal Machine Learning Engineer will operate across training, inference, evaluation, and infrastructure, solving the hardest architectural and performance problems.  While Technical Leads may own execution at the team level, you set the technical standard and shape how ML systems are built across the organization.  This is a hands-on, high-impact role focused on depth.  This position is 100% Remote.

 

Principal Machine Learning Engineer Responsibilities:

 

- Architect and build large-scale ML systems spanning data, training, evaluation, inference, and deployment.

- Design reproducible, high-performance training pipelines across GPU infrastructure.

- Architect inference systems that balance latency, throughput, cost, and reliability at scale.

- Design and maintain data systems for high-quality synthetic and real-world training data.

- Implement evaluation pipelines covering performance, robustness, safety, and bias, in partnership with research leadership.

- Own production deployment, including GPU optimization, memory efficiency, latency reduction, and scaling policies.

- Collaborate closely with application engineering to integrate ML systems cleanly into backend, mobile, and desktop products.

- Make pragmatic trade-offs and ship improvements quickly, learning from real usage.

- Work under real production constraints: latency, cost, reliability, and safety

 

Principal Machine Learning Engineer Outcomes:

 

- ML systems (training, inference, evaluation) are reliable, scalable, and meet defined performance targets.

- Models deployed to production achieve measurable quality improvements and meet user-impact goals.

- Production issues are proactively monitored, debugged, and resolved with clear root-cause analysis.

- Team and cross-functional collaborators benefit from clear guidance, best practices, and scalable ML solutions.

- Research-to-production cycles are efficient, safe, and continuously improve the product experience.

Qualifications:

Principal Machine Learning Engineer Qualifications:

 

- Strong background in deep learning and transformer-based architectures.

- Artificial Intelligence (AI) experience required.

- Hands-on experience training, fine-tuning, or deploying large-scale ML models in production.

- Proficiency with at least one modern ML framework (e.g. PyTorch, JAX), and ability to learn others quickly.

- Experience with distributed training and inference frameworks (e.g. DeepSpeed, FSDP, Megatron, ZeRO, Ray).

- Strong software engineering fundamentals; you write robust, maintainable, production-grade systems.

- Experience with GPU optimization, including memory efficiency, quantization, and mixed precision.

- Comfort owning ambiguous, zero-to-one ML systems end-to-end.

- A bias toward shipping, learning fast, and improving systems through iteration.

- Experience with LLM inference frameworks such as vLLM, TensorRT-LLM, or FasterTransformer.

- Contributions to open-source ML or systems libraries.

- Background in scientific computing, compilers, or GPU kernels.

- Experience with RLHF pipelines (PPO, DPO, ORPO).

- Experience training or deploying multimodal or diffusion models.

- Experience with large-scale data processing (Apache Arrow, Spark, Ray).

 

Benefits include medical insurance, Dental, Vision, Savings Plan Options, PTO, etc.

 

Keywords:  San Francisco CA Jobs, Principal Machine Learning Engineer, Apache Arrow, DeepSpeed, DPO, FasterTransformer, FSDP, GPU Kernels, JAX, LLM, Machine Learning, Megatron, ML, ORPO, PPO, Principal Machine Learning Engineer, Pytorch, RLHF Pipelines, Spark, TensorRT-LLM, Virtual Large Language Model, vLLM, Work From Home, ZeRO Ray, California Recruiters, IT Jobs, California Recruiting

 

Looking to hire a Principal Machine Learning Engineer in San Francisco, CA or in other cities?  Our IT recruiting agencies and staffing companies can help.

 

We help companies that are looking to hire Principal Machine Learning Engineers for jobs in San Francisco, California and in other cities too.  Please contact our IT recruiting agencies and IT staffing companies today!

Additional Information:

Please check out all of our jobs at www.ginastechjobs.com.

Similar jobs

Machine Learning Engineer roles near San Francisco, California
19h
Save
Mark Applied
Hide
Founding Engineer - Machine Learning
Mountain View, California, United States
$220k-$300k/yr OnsiteFull Time
Clera
Clera: AI talent agent matching professionals with high-growth startup roles
3+ YOERequires 3–10 years of ML engineering, applied science, or research engineering experience; Python and PyTorch, TensorFlow, or JAX; distributed systems, cloud ML infrastructure, MLOps, and large-scale data experience.
Python, PyTorch, TensorFlow, JAX, AWS, GCP, Azure, Weights & Biases, MLflow
1d
Save
Mark Applied
Hide
Machine Learning Engineer
San Francisco, California, United States
$170k-$300k/yr HybridFull Time
Clay
Clay: Platform for automated lead enrichment and sales workflows.
5+ YOE5+ years in machine learning engineering or ML-heavy software engineering, production ML features, strong coding, LLMs or classical ML, data-intensive systems, and product-oriented problem solving.
LLMs, Snowflake, dbt, Dagster
1d
Save
Mark Applied
Hide
Machine Learning Engineer - Satellite Capacity Optimization & Planning
Carlsbad or San Jose or San Francisco or New York City
$141k-$222k/yr OnsiteFull Time
Viasat
ViasatNASDAQ: VSAT: Provider of global satellite-based connectivity and secure communication solutions.
7+ YOERequires 7+ years in ML or optimization, optimization techniques, production cloud ML systems, ML frameworks, geospatial visualization, containerized development, SQL, RESTful APIs, and up to 10% travel.
AWS, GCP, TensorFlow, PyTorch, SQL, RESTful APIs, Airflow, Athena, BigQuery, ECS, Batch
2d
Save
Mark Applied
Hide
Staff Machine Learning Engineer, Generative AI Modeling and Inference
Los Angeles or Seattle or Palo Alto or New York City or Bellevue
$229k-$343k/yr OnsiteFull Time
Snap
SnapNYSE: SNAP: Provides visual messaging software and augmented reality wearable devices.
8+ YOEBachelor's degree or equivalent experience and 8+ years of post-bachelor's machine learning experience, with expertise in generative modeling, computer vision, deep learning, and ML frameworks.
TensorFlow, PyTorch, JAX, MLX, scikit-learn
2d
Save
Mark Applied
Hide
(USA) Staff, Machine Learning Engineer
Sunnyvale, California, United States
$169k-$338k/yr OnsiteFull Time
Walmart
WalmartNYSE: WMT: Operates a chain of hypermarkets, discount stores, and grocery stores.
4+ YOEBachelor's degree and 4 years, or 6 years of experience, in software engineering, machine learning engineering, AI systems, or a related field. Requires ML, MLOps, cloud, and programming expertise.
TensorFlow, PyTorch, Scikit-learn, Python, SQL, AWS, GCP, Azure, Kubernetes, CI/CD, MLOps, Web Content Accessibility Guidelines (WCAG) 2.2 AA
2d
Save
Mark Applied
Hide
Machine Learning Engineer
Chicago or Washington or Tallahassee or Sarasota or Hartford or Nashville or Costa Mesa or San Jose or Atlanta or Boston or Burlington or Cleveland or Columbus or Dallas or Denver or Fort Lauderdale or Fort Wayne or Grand Rapids or Indianapolis or Knoxville or Lexington or Livingston or Plano or Louisville or Los Angeles or Miami or The Woodlands or New York City or Oakbrook Terrace or Sacramento or San Francisco or South Bend or Springfield or Tampa or Houston or Austin or Charlotte
$62k-$100k/yr OnsiteFull Time
Crowe
Crowe: Global professional services firm providing audit, tax, and consulting.
Expert Python and Git skills; ability to develop, test, debug, and document production AI solutions, navigate ambiguity, communicate technical details, and support ethical and security reviews.
Python, Git, ChatGPT Enterprise, Microsoft Copilot, Azure AI Foundry
2d
Save
Mark Applied
Hide
Lead Machine Learning Engineering, (Hybrid)
Seattle or San Jose
$198k-$250k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
8+ YOESTEM degree and 8+ years relevant experience, or equivalent advanced-degree pathways; 3+ years scaling ML datasets; 5+ years programming and ML framework experience.
Python, C++, Go, PyTorch, TensorFlow, Spark, Ray, Beam, Large Language Models (LLMs)
2d
Save
Mark Applied
Hide
Staff Machine Learning Engineer, Generative AI Modeling and Inference
Los Angeles or Seattle or Palo Alto or New York City or Bellevue
$195k-$343k/yr OnsiteFull Time
Snap
SnapNYSE: SNAP: Develops social media applications and augmented reality technology.
8+ YOEBachelor's degree or equivalent experience and 8+ years of post-bachelor's ML experience, or advanced degree with equivalent experience. Requires computer vision or generative modeling and ML framework experience.
TensorFlow, PyTorch, JAX, MLX, scikit-learn, GPU, CPU, NPU, RSUs