Company:Qualcomm Technologies, Inc. Job Area:Engineering Group, Engineering Group Graphics Software Engineering General Summary: The Qualcomm GPU Software team is looking for talented software engineers interested in developing high performance kernels and runtimes for accelerating GPGPU operations such as
Company:Qualcomm Technologies, Inc. Job Area:Engineering Group, Engineering Group Graphics Software Engineering General Summary: As a leading technology innovator, Qualcomm pushes the boundaries of whats possible to enable next-generation experiences and drives digital transformation to help create
Were now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance ceiling? Our LLM Inference Performance Analysis and Optimization team builds the answer from
About the role: CoreWeave is the top-rated AI-cloud for high-performance GPU infrastructure across AI/ML, visual effects, rendering, and real-time inference. Our stack is engineered for speed, scale, and cost-efficiency—an unmatched alternative to traditional hyperscalers. At CoreWeave, infrastructure
CoreWeave is seeking a Senior Engineer for its Benchmarking & Performance team to own kernel-level optimization for LLM inference and end-to-end model serving, focusing on CUDA kernels and throughput/latency improvements. You will lead kernel design reviews, mentor engineers, and
SIG Susquehanna in Bala Cynwyd, Pennsylvania, is seeking a GPU Performance Engineer to develop optimized CUDA kernels for low-latency inference. The ideal candidate has strong CUDA programming skills and expertise in GPU architecture. Responsibilities include optimizing kernel performance, analyzing models
NVIDIA is seeking a Sr. Inference Engineer to push GPU kernel optimization for LLM inference. The role focuses on silicon-measured kernel benchmarking, model-level performance projection, and agentic optimization systems that improve kernels at the assembly level. You will collaborate with
Baseten is seeking an Engineering Manager to lead our GPU Kernel Engineering team, directing CUDA kernel work and GPU architecture optimization within Basetens inference stack. This player-coach role combines hands-on coding with building processes, culture, and a roadmap for a
NVIDIA Corporation in Santa Clara, CA seeks a Senior Architect to model and simulate next-generation GPU/SOC architectures. You will develop modelling tools and environments, enabling teams to build hardware models efficiently with SystemC, C++, and advanced modelling
A leading tech firm is seeking a talented Senior Staff Software Engineer to design and develop software for high-density Data Center Compute racks. This remote role requires expertise in GPU programming and LINUX driver development, with a
Susquehanna International Group, LLP is seeking a GPU Performance Engineer to craft highly optimized CUDA kernels for low-latency inference. You will collaborate with researchers to translate mathematical models into production-ready GPU code. The role emphasizes deep GPU hardware knowledge, memory
Pantera Capital in Palo Alto is seeking a talented engineer to revolutionize AI supercomputing infrastructure. You’ll be tasked with designing, building, and optimizing massive GPU clusters that drive AI training and inference. The role demands hands-on expertise
DigitalOcean is looking for a Staff Software Engineer with a passion for Linux kernel development. This role involves optimizing and maintaining Linux kernels to ensure performance and security across their cloud infrastructure. The ideal candidate should have a
Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to
SummaryApple’s Compute Frameworks team in GPU, Graphics and Displays org provides a suite of high-performance data parallel algorithms for developers inside and outside of Apple for iOS, macOS and Apple TV. Our efforts are currently focused in
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to
Magic’s mission is to build safe AGI that accelerates humanity’s progress on the world’s most important problems. We believe the most promising path to safe AGI lies in automating research and code generation to improve models
About the Role As a Systems Research Engineer specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. Working closely with the modeling and algorithm team, you will
Company:Qualcomm Technologies, Inc. Job Area:Engineering Group, Engineering Group GPU ASICS Engineering General Summary: Architects, designs, implements, verifies, and optimizes performance and power of GPU cores. Responsible for verification of Graphics IP , and performing pre- and post-silicon verification