Hot French startup ZML releases free product to speed inference across lots of AI chips
↗ZML launches a free multi-chip LLM inference server (ZML/LLMD) to run models across Nvidia, AMD, Google TPU, Apple Metal, and Intel Arc, aiming to reduce vendor lock-in and lower inference costs.
Jul 8, 20261%