Instructions to use onnx-community/Phi-4-mini-instruct-ONNX-GQA with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers.js
How to use onnx-community/Phi-4-mini-instruct-ONNX-GQA with Transformers.js:
// npm i @huggingface/transformers import { pipeline } from '@huggingface/transformers'; // Allocate pipeline const pipe = await pipeline('text-generation', 'onnx-community/Phi-4-mini-instruct-ONNX-GQA');
https://huggingface.co/microsoft/Phi-4-mini-instruct with ONNX weights to be compatible with Transformers.js.
Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using 🤗 Optimum and structuring your repo like this one (with ONNX weights located in a subfolder named onnx).
- Downloads last month
- 52
Model tree for onnx-community/Phi-4-mini-instruct-ONNX-GQA
Base model
microsoft/Phi-4-mini-instruct