Name		Name	Last commit message	Last commit date
parent directory ..
mlserver_huggingface		mlserver_huggingface
tests		tests
LICENSE		LICENSE
README.md		README.md
setup.py		setup.py

README.md

HuggingFace runtime for MLServer

This package provides a MLServer runtime compatible with HuggingFace Transformers.

Usage

You can install the runtime, alongside mlserver, as:

pip install mlserver mlserver-huggingface

For further information on how to use MLServer with HuggingFace, you can check out this worked out example.

Settings

The HuggingFace runtime exposes a couple extra parameters which can be used to customise how the runtime behaves. These settings can be added under the parameters.extra section of your model-settings.json file, e.g.

---
emphasize-lines: 5-8
---
{
  "name": "qa",
  "implementation": "mlserver_huggingface.HuggingFaceRuntime",
  "parameters": {
    "extra": {
      "task": "question-answering",
      "optimum_model": true
    }
  }
}

These settings can also be injected through environment variables prefixed with `MLSERVER_MODEL_HUGGINGFACE_`, e.g.

```bash
MLSERVER_MODEL_HUGGINGFACE_TASK="question-answering"
MLSERVER_MODEL_HUGGINGFACE_OPTIMUM_MODEL=true
```

Reference

You can find the full reference of the accepted extra settings for the HuggingFace runtime below:


.. autopydantic_settings:: mlserver_huggingface.settings.HuggingFaceSettings

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

huggingface

huggingface

README.md

HuggingFace runtime for MLServer

Usage

Settings

Reference

Files

huggingface

Directory actions

More options

Directory actions

More options

Latest commit

History

huggingface

Folders and files

parent directory

README.md

HuggingFace runtime for MLServer

Usage

Settings

Reference