French AI startup ZML, based in Paris, has released ZML/LLMD, a free inference server designed to run open-source LLMs at high speed across a wide range of AI chips. Offered for now as an alpha technical preview, it supports not only NVIDIA but also AMD, Google TPU, Apple Metal and Intel Arc/OneAPI accelerators, TechCrunch reported.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.