Susquehanna International Group is seeking a GPU Performance Engineer in Bala Cynwyd, Pennsylvania. In this role, you will design and optimize CUDA kernels for low-latency inference workloads, working closely with quantitative researchers. This position requires strong expertise
SIG Susquehanna in Bala Cynwyd, Pennsylvania, is seeking a GPU Performance Engineer to develop optimized CUDA kernels for low-latency inference. The ideal candidate has strong CUDA programming skills and expertise in GPU architecture. Responsibilities include optimizing kernel performance, analyzing
AVEVA is creating software trusted by over 90% of leading industrial companies. Salary Range:$124,200.00 - $236,900.00 This pay range represents the minimum and maximum compensation that the position offers, and final compensation can vary within the
Overview We are looking for a GPU Performance Engineer to build highly optimized CUDA kernels for low-latency inference. This role is focused on workloads where off-the-shelf runtimes and vendor libraries do not fully exploit the structure of
Founder Vice President, AI Inference Software About the Company Confidential AI systems company Industry Information Technology and Services Type Privately Held, VC-backed About the Role The Company is seeking a senior software leader to take on
• Design, develop, and deploy production models, services, and pipelines that are reliable, scalable, and maintainable • Partner with data science, product, data engineering, and platform teams to translate business problems into technical solutions • Build
Vice President of AI Inference, Software About the Company Confidential venture-backed AI systems company building high-performance inference software and future hardware. Industry Computer Software Type Privately Held, VC-backed About the Role The Company is seeking a
Overview We are looking for a GPU Performance Engineer to build highly optimized CUDA kernels for low-latency inference. This role focuses on workloads where off-the-shelf runtimes and vendor libraries do not fully exploit the structure of the
Overview We are looking for a GPU Performance Engineer to build highly optimized CUDA kernels for low‑latency inference. This role is focused on workloads where off‑the‑shelf runtimes and vendor libraries do not fully exploit the structure of
Overview We are looking for a Machine Learning Engineer focused on low-latency inference optimization to help build, tune, and productionize high-performance model serving systems. This role sits at the intersection of machine learning, systems engineering, and
Make sure they complete AI Torc Test ! Job Responsibilities We are looking for a visionary AI Engineer to lead the integration of advanced artificial intelligence into our flagship products. You will be at the forefront