ABOUT BASETEN Baseten powers mission-critical inference for the worlds most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambdas mission is to make compute as ubiquitous as
ABOUT BASETEN Baseten powers mission-critical inference for the worlds most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies
We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of whats
About the role The Model Shaping team at Together AI works on products and research focused on tailoring open foundation models to downstream applications. We build services that enable machine learning developers to choose the best
About Black Forest Labs Were the team behind Latent Diffusion, Stable Diffusion, and FLUX—foundational technologies that changed how the world creates images and video. We’re creating the generative models that power how people make images and
The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence. As a Machine Learning and System Optimization Engineer, you will orchestrate and allocate overall system
About Us Alembic is the pioneering Causal AI platform. We help the worlds largest enterprises move past correlation to prove what actually drives business outcomes — the question marketing and growth teams have never been able
About Us: AI needs a new infrastructure layer. Were building it at Modal. Every era of computing brought new workloads that previous infrastructure couldnt support: mainframes, databases, and the cloud. Each time, the company that rebuilt
About Us: AI needs a new infrastructure layer. Were building it at Modal. Every era of computing brought new workloads that previous infrastructure couldnt support: mainframes, databases, and the cloud. Each time, the company that rebuilt
Who We Are Physical Intelligence is bringing general-purpose AI into the physical world. We are a group of engineers, scientists, roboticists, and company builders developing foundation models and learning algorithms to power the robots of today
ABOUT BASETEN Baseten powers mission-critical inference for the worlds most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies
Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack —
About Us Eon is building the infrastructure for large-scale connectomics data collection, reconstruction, and brain simulation. Our mission is to enable the safe and scalable development of brain emulation technology, beginning with digital twins of model
ABOUT BASETEN Baseten powers mission-critical inference for the worlds most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies
Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack —
Introducing Moonlake, AI for creating world simulations. Scope of Work Training efficiency Dataloaders, fusion, activation remat, gradient checkpointing. FSDP/ZeRO/tensor+pipeline parallel; NCCL tuning. GPU + kernel performance Nsight profiling, Triton/CUDA kernels, fused ops. Flash-attention–style speedups, sequence packing, KV-cache tricks. Inference
About Fluidstack We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had
About Reducto Reducto is the agentic document platform for leading AI teams who demand enterprise performance at scale. We provide a comprehensive toolkit for working with documents the way a human would, combining custom in-house and
The Video ML Foundations team optimizes video ranking infrastructure to balance latency, cost, and freshness for an optimal user experience. We drive diverse initiatives, including co-designing ranking models and systems, accelerating model training, optimizing GPU inference, and