- quadric.io, Inc (Burlingame, CA)
- …GPNPU executes both NN graph code and conventional C++ DSP and control code. Role: The AI Inference Engineer in Quadric is the key bridge between the world ... of AI /LLM models and Quadric unique platforms. The AI Inference Engineer at Quadric will [1] port AI models to Quadric platform; [2] optimize the… more
- Bank of America (Addison, TX)
- Senior Engineer - AI Inference Addison, Texas;Plano, Texas; Newark, Delaware; Charlotte, North Carolina; Kennesaw, Georgia **To proceed with your application, ... must be at least 18 years of age.** Acknowledge (https://ghr.wd1.myworkdayjobs.com/Lateral-US/job/Addison/Senior- Engineer - AI - Inference \_25029879) **Job Description:** At Bank… more
- Argonne National Laboratory (Lemont, IL)
- …and AI . In this position, the candidate can expect to explore and engineer solutions for AI inference integrated within scientific workflows, via ... an opening for a Software Engineer working in the space of enabling AI for science, specifically targeting scalable inference leveraging HPC systems and … more
- NVIDIA (CA)
- …distributed model management systems, including Rust-based runtime components, for large-scale AI inference workloads. + Implement inference scheduling ... people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An...We are now looking for a Senior System Software Engineer to work on user facing tools for Dynamo… more
- NVIDIA (Santa Clara, CA)
- …seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency. You'll ... architect and implement high-performance inference stacks, optimize GPU kernels and compilers, drive industry...teams to push the frontier of accelerated computing for AI . What you'll be doing: + Contribute features to… more
- NVIDIA (Santa Clara, CA)
- …and ensure a consistent, high-impact go-to-market strategy. This role will focus on AI inference at scale, ensuring that customers and partners understand how ... scale. We are looking for a Senior Technical Marketing Engineer to join our growing accelerated computing product team....+ Develop Technical Positioning & Messaging - Translate NVIDIA's AI inference and accelerated computing technologies into… more
- Red Hat (Sacramento, CA)
- …logic (affinity/tolerations) for **GPU workloads** and troubleshoot complex **CNI** failures. + ** AI Inference Proficiency** : You understand how a LLM forward ... developer to join our team as a **Forward Deployed Engineer ** . In this role, you will not just...software; you will be the bridge between our cutting-edge inference platform (LLM-D (https://llm-d. ai /) , and vLLM… more
- Red Hat (Raleigh, NC)
- …this role, your primary responsibility will be to build and release the Red Hat AI Inference runtimes, continuously improve the processes and tooling used by the ... open-source LLMs and vLLM to every enterprise. Red Hat Inference team accelerates AI for the enterprise...on Github. We are seeking an experienced ML Ops engineer to work closely with our product and research… more
- Red Hat (Boston, MA)
- …bring the power of open-source LLMs and vLLM to every enterprise. Red Hat Inference team accelerates AI for the enterprise and brings operational simplicity to ... ai -leads-typescript-to-1/#the-top-open-source-projects-by-contributors)** on Github. As a Principal Machine Learning Engineer focused on vLLM, you will be at the… more
- Capital One (San Francisco, CA)
- Lead AI Engineer (FM Hosting, LLM Inference ) **Overview** At Capital One, we are creating responsible and reliable AI systems, changing banking for good. ... AI and ML algorithms or technologies (eg LLM Inference , Similarity Search and VectorDBs, Guardrails, Memory) using Python,...regularly worked. Cambridge, MA: $193,400 - $220,700 for Lead AI Engineer McLean, VA: $193,400 - $220,700… more