AI Inference Engineer Jobs

417 jobs (page 1)

Categories

All Categories

Engineering (159)

Software/IT (76)

Management (8)

Entry Level (5)

AI Inference Engineer

quadric.io, Inc (Burlingame, CA)

…GPNPU executes both NN graph code and conventional C++ DSP and control code. Role: The AI Inference Engineer in Quadric is the key bridge between the world ... of AI /LLM models and Quadric unique platforms. The AI Inference Engineer at Quadric will [1] port AI models to Quadric platform; [2] optimize the… more

quadric.io, Inc (11/25/25)
- Related Jobs
Senior Engineer - AI…

Bank of America (Addison, TX)

Senior Engineer - AI Inference Addison, Texas;Plano, Texas; Newark, Delaware; Charlotte, North Carolina; Kennesaw, Georgia **To proceed with your application, ... must be at least 18 years of age.** Acknowledge (https://ghr.wd1.myworkdayjobs.com/Lateral-US/job/Addison/Senior- Engineer - AI - Inference \_25029879) **Job Description:** At Bank… more

Bank of America (12/22/25)
- Related Jobs
Software Engineer - AI…

Argonne National Laboratory (Lemont, IL)

…and AI . In this position, the candidate can expect to explore and engineer solutions for AI inference integrated within scientific workflows, via ... an opening for a Software Engineer working in the space of enabling AI for science, specifically targeting scalable inference leveraging HPC systems and … more

Argonne National Laboratory (01/06/26)
- Related Jobs
Senior Software Engineer , AI…

NVIDIA (CA)

…distributed model management systems, including Rust-based runtime components, for large-scale AI inference workloads. + Implement inference scheduling ... people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An...We are now looking for a Senior System Software Engineer to work on user facing tools for Dynamo… more

NVIDIA (11/29/25)
- Related Jobs
Senior Software Engineer , AI…

NVIDIA (Santa Clara, CA)

…seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency. You'll ... architect and implement high-performance inference stacks, optimize GPU kernels and compilers, drive industry...teams to push the frontier of accelerated computing for AI . What you'll be doing: + Contribute features to… more

NVIDIA (01/10/26)
- Related Jobs
Senior Technical Marketing Engineer…

NVIDIA (Santa Clara, CA)

…and ensure a consistent, high-impact go-to-market strategy. This role will focus on AI inference at scale, ensuring that customers and partners understand how ... scale. We are looking for a Senior Technical Marketing Engineer to join our growing accelerated computing product team....+ Develop Technical Positioning & Messaging - Translate NVIDIA's AI inference and accelerated computing technologies into… more

NVIDIA (11/06/25)
- Related Jobs
Forward Deployed Engineer , AI…

Red Hat (Sacramento, CA)

…logic (affinity/tolerations) for **GPU workloads** and troubleshoot complex **CNI** failures. + ** AI Inference Proficiency** : You understand how a LLM forward ... developer to join our team as a **Forward Deployed Engineer ** . In this role, you will not just...software; you will be the bridge between our cutting-edge inference platform (LLM-D (https://llm-d. ai /) , and vLLM… more

Red Hat (01/08/26)
- Related Jobs
Senior Principal MLOps Engineer , AI…

Red Hat (Raleigh, NC)

…this role, your primary responsibility will be to build and release the Red Hat AI Inference runtimes, continuously improve the processes and tooling used by the ... open-source LLMs and vLLM to every enterprise. Red Hat Inference team accelerates AI for the enterprise...on Github. We are seeking an experienced ML Ops engineer to work closely with our product and research… more

Red Hat (01/07/26)
- Related Jobs
Principal Machine Learning Engineer…

Red Hat (Boston, MA)

…bring the power of open-source LLMs and vLLM to every enterprise. Red Hat Inference team accelerates AI for the enterprise and brings operational simplicity to ... ai -leads-typescript-to-1/#the-top-open-source-projects-by-contributors)** on Github. As a Principal Machine Learning Engineer focused on vLLM, you will be at the… more

Red Hat (12/15/25)
- Related Jobs
Lead AI Engineer (FM Hosting, LLM…

Capital One (San Francisco, CA)

Lead AI Engineer (FM Hosting, LLM Inference ) **Overview** At Capital One, we are creating responsible and reliable AI systems, changing banking for good. ... AI and ML algorithms or technologies (eg LLM Inference , Similarity Search and VectorDBs, Guardrails, Memory) using Python,...regularly worked. Cambridge, MA: $193,400 - $220,700 for Lead AI Engineer McLean, VA: $193,400 - $220,700… more

Capital One (11/04/25)
- Related Jobs

"Alerted.org

Advanced Search

Recent Searches

Recent Jobs

Account Login

Sign Up

Forgot your password?