ZML's free inference server lets enterprises drop Nvidia's hardware stranglehold

ZML's free inference server lets enterprises drop Nvidia's hardware stranglehold

French startup ZML released LLMD on July 8, 2026, a free software layer that runs AI models on Nvidia, AMD, Google TPU, Apple Metal, and Intel Arc chips interchangeably. The move addresses enterprise and cloud providers' growing need to diversify hardware suppliers and avoid vendor lock-in. Turing Award winner Yann LeCun endorsed the approach. LLMD competes directly with commercial engines vLLM and SGLang.

Published

Read at another depth