https://huggingface.co/facebook/MobileLLM-350M with ONNX weights to be compatible with Transformers.js.
Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using 🤗 Optimum and structuring your repo like this one (with ONNX weights located in a subfolder named onnx
).
- Downloads last month
- 0
Inference API (serverless) does not yet support model repos that contain custom code.
Model tree for onnx-community/MobileLLM-350M
Base model
facebook/MobileLLM-350M