Quick Start

This guide shows you how to download your first model and start chatting with it.

Switch to the Models Page

  1. Open Herdsman and click Models in the top-left corner to switch to the Models page.

    Models page entry

Filter and Download a Model

  1. Click Expand Filters on the right to open the filter panel.

    Expand filters

  2. In the filter checkboxes, select < 10B to narrow the list to small-parameter models.

  3. Select the Gemma 4 E2B model, then click Download Model.

    Select Gemma 4 E2B

  4. On first download, you will be prompted to install required components. Click Install to continue.

  5. Once components are installed, click Continue Downloading Model.

  6. The model begins downloading and installs automatically when finished.

    Model files are large. We recommend changing the model storage path to a non-system drive first. See Installing Herdsman → Configuring the Model Storage Path.

    Download progress

Launch the Model

  1. After installation completes, a green Launch Now button appears. Click Launch Now.

    Launch Now button

  2. Configure the launch settings on the Launch Model page:

    • Drag the Context Size slider to the middle position, or pull it to maximum.
    • Toggle Enable Thinking to off.

    Launch settings

    Enable Thinking toggle

  3. Click Launch.

    Launch button

  4. The model is now running.

    Model running

Using the Model

Once a model is running in Herdsman, you can use it directly inside Herdsman or expose it as a local inference backend for FlowyAIPC.

Use the Model Inside Herdsman

  1. Click Apps in the top-left corner to switch to the Apps view.

    Apps view

  2. Type your question in the chat box and click Start Chat.

    Chat input

    Chat interface

    Chat reply

Use the Local Model from FlowyAIPC

Make sure you have already completed Download Model and Launch Model. FlowyAIPC automatically connects to any running Herdsman model.

  1. Open FlowyAIPC and click the model selector in the top-left, then scroll to the bottom of the list.

    FlowyAIPC model selector

  2. Under Local, select Gemma 4-E2B-IT.

    Select local model

  3. You can now chat with your local model from FlowyAIPC.