Industry

Edge AI in Healthcare

Imaging assistance, patient monitoring and workflow intelligence delivered without moving patient data off site.

Common workloads

  • Medical imaging
  • Patient monitoring
  • Clinical workflow

Companies active in Healthcare

NVIDIA

Chip manufacturer

Develops the Jetson family of edge system-on-modules, the DGX Spark desktop AI system and the JetPack, TensorRT and DeepStream software stacks used for edge inference, robotics and vision AI.

Santa Clara, California, United States

AMD

Chip manufacturer

Produces Ryzen AI processors that combine Zen 5 CPU cores, RDNA 3.5 graphics and the XDNA 2 NPU, with large unified memory configurations used for local generative AI workloads.

Santa Clara, California, United States

ADLINK Technology

Hardware manufacturer

Builds embedded computing modules and the DLAP family of deep learning acceleration platforms, including fanless Jetson-based inference systems with Power over Ethernet camera inputs.

Taoyuan, Taiwan

Dell Technologies

Cloud & infrastructure provider

Supplies the PowerEdge XR range of short-depth, edge-focused servers for industrial automation, video analytics and inference at retail, telecom and factory sites.

Round Rock, Texas, United States

Lenovo

Hardware manufacturer

Manufactures ThinkEdge edge servers and ThinkStation workstations, including multi-GPU platforms for local model development and EPYC-based edge systems for distributed sites.

Beijing, China / Morrisville, United States

Ollama

Software provider

Maintains the MIT-licensed Ollama runtime for pulling and serving open-weight language and multimodal models locally on workstations, servers and capable edge systems.

Palo Alto, California, United States

ggml.ai (llama.cpp)

Software provider

Maintainers of the ggml tensor library and llama.cpp, an MIT-licensed C and C++ inference engine that runs open-weight language models on CPUs and on CUDA, Metal and Vulkan backends.

Sofia, Bulgaria

Microsoft

Software provider

Originated ONNX Runtime, the MIT-licensed cross-platform inference accelerator, and publishes the Phi family of small open-weight models aimed at low-latency local inference.

Redmond, Washington, United States

Mistral AI

Software provider

Publishes open-weight language and multimodal models, including Apache 2.0 licensed Mistral Small builds designed for latency-sensitive and self-hosted deployments.

Paris, France

Hugging Face

Software provider

Hosts open-weight models and publishes the SmolVLM family of compact vision-language models designed explicitly for on-device inference under the Apache 2.0 licence.

New York, United States