Specifications

AI Performance

40 TOPS (INT4) / 20 TOPS (INT8)

Chipset

Hailo-10H — 2nd-gen neural core

Form Factor

M.2 Key M (2242, 2280)

Interface

PCIe Gen3 x4 lanes

On-Module Memory

4GB or 8GB LPDDR4/4X, 4266 MT/s

Power

2.5W typical

Operating Temp

-40°C to 85°C (industrial grade)

Host Architectures

x86, ARM

OS Support

Linux, Windows, Android

Frameworks

TensorFlow, PyTorch, ONNX, Keras, TFLite

Overview

The Hailo-10H M.2 is the industry's first edge AI accelerator purpose-built for on-device generative AI, delivering 40 TOPS INT4 at just 2.5W typical power. Featuring Hailo's second-generation neural core architecture with up to 8GB on-module LPDDR4, it enables local inference of large language models (LLMs), vision-language models (VLMs), and real-time computer vision — without cloud dependency. The standard M.2 2242/2280 form factor with PCIe Gen3 x4 interface allows drop-in integration into existing edge PCs, industrial gateways, and embedded systems, backed by Hailo's Model Zoo and full TensorFlow/PyTorch/ONNX framework support.

Key Benefits

Generative AI at the Edge: First-to-market M.2 accelerator capable of running LLMs and VLMs locally — 8GB on-module RAM enables models previously requiring cloud or datacenter GPUs. 2.5W Power Envelope: Best-in-class power efficiency — 16 TOPS per watt — enables fanless, battery-powered, and thermally constrained deployments. Industrial Grade: -40°C to 85°C operating range with M.2 socket reliability for automotive, outdoor, and factory-floor AI at the edge. Drop-In Upgrade: Standard M.2 Key M slot with standard OS drivers — upgrade any x86 or ARM edge system with 40 TOPS in minutes.

Request a Quote — HAILO 10H M2

QS Compute — NVIDIA Goldendisk Partner. Global B2B supply, 15-day sample lead time.

Get Your Quote →