Software provider

Microsoft

Originated ONNX Runtime, the MIT-licensed cross-platform inference accelerator, and publishes the Phi family of small open-weight models aimed at low-latency local inference.

Redmond, Washington, United StatesProfile reviewed editorially

Categories

Industries served

Solutions from Microsoft

Terminal showing an inference session starting
Microsoft

ONNX Runtime

MIT-licensed cross-platform inference accelerator with hardware-specific execution providers.

Type
Cross-platform inference and training accelerator
Licence
MIT
Model sources
PyTorch, TensorFlow and Keras, TensorFlow Lite, scikit-learn and others via ONNX
Deployment
Edge device, client PC or server
Cross-vendor edge deployment
AI PC applications
Server and device parity
Laptop running a small language model locally
Microsoft

Microsoft Phi-4

MIT-licensed 14-billion-parameter small language model built for low-latency reasoning tasks.

Parameters
14 billion
Licence
MIT
Design goal
Low-latency scenarios with strong reasoning quality
Deployment
Workstation, edge server or capable AI PC
Local reasoning assistants
Private summarisation
Structured extraction

Company information is summarised from publicly available material. If you represent Microsoft and would like to update or expand this profile, please get in touch.