Loading memory…
Loading memory…
Luminal optimizes AI models to accelerate and simplify model deployment using a search-based compiler. The AI stack needs to be rethought from the ground up to achieve this. As demand for inference grows, teams will need to run models across a wider range of hardware, not just the platforms with the most mature software support. Luminal makes AI workloads faster, more portable, and easier to deploy by automatically optimizing models for the ideal hardware. Our mission is to make state-of-the-art AI production-ready on any compute platform. We have already closed multiple contracts with non-Nvidia hardware platforms. Luminal is backed by Y Combinator and Felicis as well as top-tier angels such as Paul Graham (founder of Y Combinator), Guillermo Rauch (founder of Vercel) and many others. Design and build core compiler infrastructure in Rust Build search-based optimization systems for discovering faster kernels Develop backend code generation for multiple targets (NVIDIA, Trainium, AMD, TPU, etc.) Implement compiler passes for fusion, scheduling, memory planning, and kernel selection Profile real models, identify bottlenecks, and improve latency and throughput Help shape Luminal’s eng