Myrtle.ai, a recognized leader in accelerating machine learning inference, today released support for its VOLLO inference accelerator on the NT400D1x series of SmartNICs from Napatech. VOLLO achieves industry-leading ML inference compute latencies, which can be less than one microsecond. This new release enables those who need the very lowest latencies possible to run inference next to the network in a SmartNIC. A wide range of models may be run on VOLLO, including LSTM, CNN, MLP, as well as Random Forests and Gradient Boosting decision trees.
This has been developed to meet the needs of a wide range of applications including financial trading, wireless telecommunications, cyber security, network management and others, where running ML inference at the lowest possible latency confers advantages in security, safety, profit, efficiency and cost.