Liquid AI's Open d1 models run on NVIDIA hardware with llama.cpp support
Original titleBoth Open d1 models run across NVIDIA DGX, RTX, and Jetson, with day-one llama.cpp support to run anywhere.
AISummary
Liquid AI's Open d1 models run across NVIDIA DGX, RTX, and Jetson hardware, with day-one llama.cpp support for deployment anywhere. Measured one request at a time, the d1-3B model's single-question latency is 8 ms on an NVIDIA RTX 4090, 16 ms on Jetson AGX Thor, 26 ms on Jetson AGX Orin 64 GB, and 50 ms on Jetson Orin Nano.
Source: Liquid AI · x.comPublished · added here