https://huggingface.co/meta-llama/Llama-3.2-3B-Instruct with ONNX weights to be compatible with Transformers.js.

Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using πŸ€— Optimum and structuring your repo like this one (with ONNX weights located in a subfolder named onnx).

Downloads last month
141
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for onnx-community/Llama-3.2-3B-Instruct-ONNX

Quantized
(404)
this model

Space using onnx-community/Llama-3.2-3B-Instruct-ONNX 1