Software

Edge AI Software

Runtimes, orchestration, fleet management and MLOps for the edge

Edge AI software covers inference runtimes, container orchestration for constrained sites, device fleet management, model deployment pipelines and observability.

Solutions in this category

Representative platforms with the details engineering and procurement teams shortlist on.

Small cluster of edge servers on a rack shelf
SUSE (Rancher)

K3s

Apache 2.0 lightweight CNCF-conformant Kubernetes distribution shipped as a single binary under 100 MB.

Type
Lightweight Kubernetes distribution
Licence
Apache 2.0
Packaging
Single binary under 100 MB
Deployment
Edge server, industrial PC or Arm device cluster
Orchestrating inference containers at sites
Staged model rollouts
Air-gapped edge clusters
Dashboard showing a fleet of connected edge devices on a dark interface
balena

balenaCloud

Container-based fleet platform with over-the-air updates and support for more than 80 device types.

Type
Device fleet management and container deployment platform
Operations
Over-the-air application and OS updates, remote diagnostics
Integration
Full API and SDK
Deployment
Distributed edge estate
Model rollout and rollback
Remote diagnostics
Fleet observability

Specifications are summarised from publicly published manufacturer material and are provided for orientation only. Always confirm current figures directly with the manufacturer before purchasing or designing in.

Companies working in this category

Siemens

Industrial AI company

Industrial automation and software group whose SIMATIC IPC industrial PCs and Industrial Edge platform bring application and AI workloads alongside controllers and plant data.

Munich, Germany

Roboflow

Software provider

Provides computer vision tooling including the open-source Inference package and inference server, covering model loading, pre- and post-processing and workflow execution on edge devices.

Des Moines, Iowa, United States

Edge Impulse

Software provider

Edge MLOps platform, now part of Qualcomm, for collecting data, training and deploying models to microcontrollers, NPUs, CPUs and GPUs across a broad partner hardware ecosystem.

San Jose, California, United States

balena

Software provider

Container-based platform for deploying and managing fleets of Linux edge devices, with over-the-air updates, an API and SDK, and support for a large catalogue of device types.

Remote / Seattle, United States

SUSE (Rancher)

Software provider

Originator of K3s, the lightweight CNCF-conformant Kubernetes distribution now a CNCF project, widely used to orchestrate containerised AI workloads on edge and Arm hardware.

Nuremberg, Germany

Ollama

Software provider

Maintains the MIT-licensed Ollama runtime for pulling and serving open-weight language and multimodal models locally on workstations, servers and capable edge systems.

Palo Alto, California, United States

Microsoft

Software provider

Originated ONNX Runtime, the MIT-licensed cross-platform inference accelerator, and publishes the Phi family of small open-weight models aimed at low-latency local inference.

Redmond, Washington, United States

Edge AI Software: common questions

What is edge MLOps?
The practice of packaging, deploying, monitoring and updating models across distributed devices, including staged rollouts and rollback when a model regresses in the field.