CoreWeave is The Essential Cloud for AI. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs,
Company:Qualcomm Technologies, Inc. Job Area:Engineering Group, Engineering Group Software Engineering General Summary: Qualcomm is seeking a Sr. Staff / Principal-level Software Engineer to provide system-level technical leadership for next-generation ARM server platforms. This role spans Linux
Minimum qualifications: Bachelors degree or equivalent practical experience. 5 years of experience in Linux Kernel development using the programming languages C/C++. 3 years of experience testing, maintaining, or launching software products, and 1 year of experience with software
DescriptionThe Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon’s custom machine learning accelerators, Inferentia and Trainium. DescriptionThe Annapurna Labs
Yoh, a Day & Zimmermann company, is seeking an ESXi CPU Server Architect / SW Engineer to join the ESXi-x86 kernel team focused on x86 CPUs and servers. You will work on large-scale system and hardware enablement
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs We are building AWS Neuron, an SDK that accelerates deep learning and GenAI workloads on Amazon’s custom machine learning accelerators, Inferentia and Trainium. As part of the Acceleration
The Senior Linux Kernel Camera/ISP Driver Engineer will design, develop, and optimize Linux kernel-level drivers for camera and image signal processing systems in embedded platforms. The role involves deep low-level development, hardware bring-up, and performance tuning of multimedia
Responsibilities Design and implement high-performance compute kernels for AI primitives such as GEMM, attention, normalization, and convolution. Optimize for throughput, latency, and memory hierarchy across heterogeneous compute units (SIMD, matrix engines, DMA). Collaborate with compiler and runtime
Veriipro is seeking a Senior Linux Kernel Camera/ISP Driver Engineer based in Palo Alto, California, to design and optimize Linux kernel drivers for camera systems. This role emphasizes deep development in multimedia subsystems, focusing on high-performance and power-efficient
Description The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon’s custom machine learning accelerators, Inferentia and Trainium. Description The
A cutting-edge AI technology firm in California is seeking an entry-level programmer to design and implement high-performance compute kernels for AI primitives. Responsibilities include optimizing for throughput and memory hierarchy, collaborating with teams, and writing reusable code
ABOUT xAI xAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization
The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon’s custom machine learning accelerators, Inferentia and Trainium. The Acceleration Kernel Library
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the servers that use them. As the Software Development Engineer for the Neuron Runtime Team, you will be
The application window is expected to close on: 08/14/2026 Job posting may be removed earlier if the position is filled or if a sufficient number of applications are received. This is a hybrid position requiring 2-3
Reality Labs (RL) focuses on delivering Metas vision through Virtual Reality (VR), Augmented Reality (AR) and Wearable AI Devices. The compute performance and power efficiency requirements of our AI devices require custom silicon. Reality Labs Silicon
SummaryApple’s ML Frameworks (MetalLM) team in GPU, Graphics and Machine Learning works on enabling Apple Intelligence through high-performance, distributed inference of GenAI applications (such as LLMs) on Private Cloud Compute. You will get to work on
SummaryAt Apple, our Platform Architecture group is responsible for connecting our hardware and software into one unified system. You’ll collaborate with engineers across Apple to design how all of our technologies work in unison, drive development
About Clockwork Systems Clockwork.io – Software Driven Fabrics to increase GPU cluster utilization Clockwork Systems was founded by Stanford researchers and veteran systems engineers who share a vision for redefining the foundations of distributed computing. As
At d-Matrix, we are focused on unleashing the potential of generative AI to power the transformation of technology. We are at the forefront of software and hardware innovation, pushing the boundaries of what is possible. Our