AI Inference Engineer Jobs in California

241 jobs (page 1)

Categories

All Categories

Engineering (98)

Software/IT (35)

AI Inference Engineer

quadric.io, Inc (Burlingame, CA)

…GPNPU executes both NN graph code and conventional C++ DSP and control code. Role: The AI Inference Engineer in Quadric is the key bridge between the world ... of AI /LLM models and Quadric unique platforms. The AI Inference Engineer at Quadric will [1] port AI models to Quadric platform; [2] optimize the… more

quadric.io, Inc (11/25/25)
- Related Jobs
Senior Software Engineer , AI…

NVIDIA (CA)

…distributed model management systems, including Rust-based runtime components, for large-scale AI inference workloads. + Implement inference scheduling ... people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An...We are now looking for a Senior System Software Engineer to work on user facing tools for Dynamo… more

NVIDIA (11/29/25)
- Related Jobs
Senior Software Engineer , AI…

NVIDIA (Santa Clara, CA)

…seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency. You'll ... architect and implement high-performance inference stacks, optimize GPU kernels and compilers, drive industry...teams to push the frontier of accelerated computing for AI . What you'll be doing: + Contribute features to… more

NVIDIA (01/10/26)
- Related Jobs
Senior Technical Marketing Engineer…

NVIDIA (Santa Clara, CA)

…and ensure a consistent, high-impact go-to-market strategy. This role will focus on AI inference at scale, ensuring that customers and partners understand how ... scale. We are looking for a Senior Technical Marketing Engineer to join our growing accelerated computing product team....+ Develop Technical Positioning & Messaging - Translate NVIDIA's AI inference and accelerated computing technologies into… more

NVIDIA (11/06/25)
- Related Jobs
Forward Deployed Engineer , AI…

Red Hat (Sacramento, CA)

…logic (affinity/tolerations) for **GPU workloads** and troubleshoot complex **CNI** failures. + ** AI Inference Proficiency** : You understand how a LLM forward ... developer to join our team as a **Forward Deployed Engineer ** . In this role, you will not just...software; you will be the bridge between our cutting-edge inference platform (LLM-D (https://llm-d. ai /) , and vLLM… more

Red Hat (01/08/26)
- Related Jobs
Lead AI Engineer (FM Hosting, LLM…

Capital One (San Francisco, CA)

Lead AI Engineer (FM Hosting, LLM Inference ) **Overview** At Capital One, we are creating responsible and reliable AI systems, changing banking for good. ... AI and ML algorithms or technologies (eg LLM Inference , Similarity Search and VectorDBs, Guardrails, Memory) using Python,...regularly worked. Cambridge, MA: $193,400 - $220,700 for Lead AI Engineer McLean, VA: $193,400 - $220,700… more

Capital One (11/04/25)
- Related Jobs
Software Development Engineer AI…

Amazon (Cupertino, CA)

…and Trainium machine learning accelerators, designed to deliver high-performance, low-cost inference at scale. The Neuron Serving team develops infrastructure to ... and efficiently on AWS silicon. We are seeking a Software Development Engineer to lead and architect our next-generation model serving infrastructure, with a… more

Amazon (12/21/25)
- Related Jobs
Senior Software Development Engineer…

Amazon (Cupertino, CA)

…of applied scientists, system engineers, and product managers to deliver state-of-the-art inference capabilities for Generative AI applications. Your work will ... ML frameworks like PyTorch and JAX enabling unparalleled ML inference and training performance. The Inference Enablement...expertise to push the boundaries of what's possible in AI acceleration. As part of the broader Neuron organization,… more

Amazon (01/06/26)
- Related Jobs
Senior Software Development Engineer…

Amazon (Cupertino, CA)

…of applied scientists, system engineers, and product managers to deliver state-of-the-art inference capabilities for Generative AI applications. Your work will ... ML frameworks like PyTorch and JAX enabling unparalleled ML inference and training performance. The Inference Enablement...expertise to push the boundaries of what's possible in AI acceleration. As part of the broader Neuron organization,… more

Amazon (12/10/25)
- Related Jobs
Lead Engineer , Inference Platform

MongoDB (Palo Alto, CA)

We're looking for a Lead Engineer , Inference Platform to join our team building the inference platform for embedding models that power semantic search, ... retrieval, and AI -native features across MongoDB Atlas. This role is part...Atlas and optimized for developer experience. As a Lead Engineer , Inference Platform, you'll be hands-on with… more

MongoDB (12/27/25)
- Related Jobs

"Alerted.org

Advanced Search

Recent Searches

Recent Jobs

Account Login

Sign Up

Forgot your password?