RSS feedhuggingface.co
Hugging Face Releases 207 WebGPU Kernels for Faster Browser AI
Summary
Hugging Face has released @huggingface/kernels, a JavaScript library for loading 207 optimized WebGPU kernels from the Hub for browser-based local AI inference. Each kernel includes versioned contracts, WGSL shaders, correctness tests, benchmarks, and documentation. Fleet adds browser-based testing that can contribute private, consented evidence from real GPUs. On an Apple M4, the kernels were 2.57 times faster by geometric mean and 1.90 times faster at the median than ONNX Runtime Web across 809 matching cases.