2 model serving engineer jobs at 2 companies in Commerce, TX

PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
Python, PyTorch, vLLM, SGLang, TensorRT, LLMs
1mo
Save
Mark Applied
Hide
Machine Learning & Data Platform Engineer
Richardson, Texas, United States
$107k-$183k/yr OnsiteFull Time
RealPage
RealPage: Software and data analytics for the real estate industry.
5+ YOE5+ years Python development on production ML systems; strong NLP/transformer experience (Hugging Face, TensorFlow); proficiency with pandas, NumPy, SQLAlchemy, PostgreSQL; ETL pipeline and model serving experience; technical leadership.
Python, Pandas, TensorFlow, Hugging Face Transformers, PostgreSQL, SQLAlchemy, Wallaroo, Paramiko, SFTP, openpyxl, xlrd, REST APIs, OAuth2/JWT, SageMaker, Vertex AI, Data Management Gateway (DMG)
2mo
Save
Mark Applied
Hide
AI/ML Engineer
Plano, Texas, United States
$65-$75/hr OnsiteFull Time
CCS INC
CCS INC: Provider of customer experience solutions and contact center technologies.
6+ YOE6+ years cloud architecture; 3+ years GenAI/LLM on AWS; strong Python/AWS; experience with vector DBs, prompt design, model serving infra; IaC tools; GenAI libraries.
Python, AWS, Lambda, ECS, EKS, S3, SageMaker, Docker, Kubernetes, Terraform, CloudFormation, GitOps, CI/CD, OpenAI APIs, LangChain, LlamaIndex, HuggingFace, Pinecone, Milvus, Weaviate, Qdrant, NVIDIA Triton, Ray Serve