https://huggingface.co./premai-io/prem-1B-chat with ONNX weights to be compatible with Transformers.js.

Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using 🤗 Optimum and structuring your repo like this one (with ONNX weights located in a subfolder named onnx).

Downloads last month
15
Inference Examples
Inference API (serverless) does not yet support transformers.js models for this pipeline type.

Collection including ucalyptus/prem-1B-chat-onnx-q4