Staff Software Engineer - GenAI Performance and Kernel
Databricks · San Francisco, California
About this role
P-1285
About This Role
As a staff software engineer for GenAI Performance and Kernel, you will own the design, implementation, optimization, and correctness of the high-performance GPU kernels powering our GenAI inference stack. You will lead development of highly-tuned, low-level compute paths, manage trade-offs between hardware efficiency and generality, and mentor others in kernel-level performance engineering. You will work closely with ML researchers, systems engineers, and product teams to push the state-of-the-art in inference performance at scale.
What You Will Do
Lead the design, implementation, benchmarking, and maintenance of core compute kernels (e.g. attention, MLP, softmax, layernorm, memory management) optimized for various hardware backends (GPU, accelerators)
Drive the performance roadmap for kernel-level improvements: vectorization, tensorization, tiling, fusion, mixed precision, sparsity, quantization, memory reuse, scheduling, auto-tuning, etc.…
Summary from Databricks's official Greenhouse career feed — read the full description on the original posting ↗
The full job description lives on Databricks's official careers page. AI Stack Jobs links you straight to it — we never sit between you and the employer.
More AI/ML roles at Databricks
- 2026-08-17 Remote - California; Remote - Colorado; Remote - Oregon; Remote - Washington H-1B SponsorRemote
- 2026-08-17 Bengaluru, India H-1B Sponsor
- 2026-08-14 United States H-1B Sponsor
- 2026-08-14 Seattle, Washington H-1B Sponsor
- 2026-08-14 San Francisco, California H-1B Sponsor
- 2026-08-14 Bengaluru, India H-1B Sponsor