Loom / Guides / Hugging Face

Cloud · Hugging Face

Hugging Face

Use open models through Hugging Face Inference Providers.

Overview

Hugging Face’s router sends your request to one of several inference companies. The model name can end in :fastest or :cheapest to let it choose.

Get a key

  1. 1
    Sign in at huggingface.co.
  2. 2
    Open Settings → Access Tokens and choose Create new token.
  3. 3
    Pick Fine-grained and turn on Make calls to Inference Providers. A plain “read” token is not enough.
  4. 4
    Create it and copy the token (it starts with hf_).

Cost. Free accounts get a small monthly credit; beyond that you add billing (or use PRO). Each inference company may bill differently.

Add it to Loom

  1. 1
    Open Loom, then Providers in the menu, and choose Add provider.
  2. 2
    Pick Hugging Face (the search box finds it).
  3. 3
    Fill in the fields in the table below.
  4. 4
    Choose Connect. Loom checks the key and says how many models it found. If it cannot check, it says so, and you can Save anyway and try a message.
  5. 5
    Open the model chip below the message box and pick a model.
Field in LoomWhat to putNeeded
API keyPaste the key you copied. Create a token at huggingface.co/settings/tokens with the “Inference Providers” permission.Yes
Endpointhttps://router.huggingface.co/v1Pre-filled
NameAnything you like. It is shown in the model pickerOptional

At a glance

Appears in Loom asHugging Face
Sign-inYour token, sent as Authorization: Bearer
ModelsLoom fetches the list from your account
Tools (connectors, search)Depends on the model
PhotosDepends on the model
PDFsNo

If something goes wrong

What you seeWhat to do
403, “does not have sufficient permissions to call Inference Providers”The token lacks the permission above. Make a new fine-grained token with it ticked.
401 on chat, although the model list loadsThe list works with any token. Check your remaining credits and that the model supports chat.

Official documentation