Afterschool Agent curriculum
Private hub · Course 2

Course 2 — Install LM Studio and run a local model

Before this course, finish Course 1 (Hermes installed) and Course 0 (concepts). You will need about 3 GB of free disk space and 30 minutes.

What this course does

By the end of this course, you will have:

After this course, every chat in Hermes Desktop runs locally. No internet needed.


Lesson 1 — Why local?

Before installing anything, two minutes on why this matters.

When you finish this course, every chat you have with Hermes happens on your laptop. The model file sits on your hard drive. Your prompt never leaves the machine. The reply is generated on the machine. Nothing travels.

This is the same as writing in a notebook. The words stay where you wrote them.

Cloud models (the default in Hermes when you finish Course 1) can do the same thing your local model does, but they send your messages to a server somewhere. Most cloud providers log your conversation. Some train future versions on what you typed.

Local-first isn't paranoia — it's hygiene.


Lesson 2 — Download LM Studio

What you need

Steps

  1. Open your browser
  2. Go to https://lmstudio.ai
  3. Click Download for [Mac / Windows]
  4. The download is about 350 MB
  5. Save the installer to your Downloads folder

Check


Lesson 3 — Install on Mac

  1. Open your Downloads folder
  2. Double-click LM-Studio-Installer.dmg
  3. Drag the LM Studio icon onto the Applications shortcut
  4. Wait for the copy (about 30 seconds)
  5. Open LM Studio from your Applications folder
  6. If macOS asks for permission, click Open

Check


Lesson 3 — Install on Windows

  1. Open your Downloads folder
  2. Double-click LM-Studio-Setup.exe
  3. Click Install when prompted
  4. Wait for the install (about 1 minute)
  5. Click Finish
  6. LM Studio opens automatically

Check


Lesson 4 — Find the model

The app has a catalogue of free, open-source models. For this course, we're using LFM2-700M from Liquid AI.

Why this model

Steps

  1. In LM Studio, look at the left sidebar. Click the magnifying glass icon (Search)
  2. In the search bar, type: LiquidAI/LFM2-700M
  3. Look for the model card with the Liquid AI logo
  4. You'll see several versions. Pick the one labelled GGUF with quantisation Q4_K_M
  5. Click Download
  6. The download is about 1.4 GB. Wait for it to finish (5–15 minutes depending on your internet)

Check


Lesson 5 — Test the model

Before connecting Hermes, confirm the model works on its own.

  1. Click the chat bubble icon (left sidebar)
  2. At the top, select LFM2-700M from the model dropdown
  3. In the message box, type: "Tell me a joke about teenagers"
  4. Press Enter
  5. Wait for the reply (3–8 seconds)
  6. Try: "What's one thing I should do today to be healthier?"

What you're seeing

The model file on your hard drive is reading your prompt, generating tokens (one at a time), and printing them as a reply. Every word is being generated on your laptop. To confirm, you can turn off your Wi-Fi right now and the chat will still work.

Check


Lesson 6 — Start the local server

LM Studio can also run as a small HTTP server. That's how Hermes talks to it.

  1. Click the developer icon in the left sidebar (looks like )
  2. Click Start Server
  3. You'll see a URL appear. It will be something like http://localhost:1234/v1
  4. Leave this window open. The server stays running in the background.

Check

Don't close LM Studio

The server only runs while LM Studio is open. If you close LM Studio, the local server stops and Hermes will fall back to a cloud model. Keep LM Studio running in the background while you use Hermes.


Lesson 7 — Connect Hermes to the local model

Steps

  1. Open the Hermes Desktop app (if not already open)
  2. In the top-right corner, click your profile icon (or the menu button)
  3. Click Settings
  4. Find the section Local Model or Provider
  5. Choose LM Studio as the local provider
  6. Paste the URL: http://localhost:1234/v1
  7. Set the model name: lfm2-700m
  8. Click Save
  9. Back in the main chat, type: "Hello. Are you running locally?"
  10. The agent should reply. Look at the bottom of the reply — it should say something like "via LM Studio" or "local model"

Check


Lesson 8 — Test it offline

The real test of local-first is that it works without internet.

  1. Turn off your Wi-Fi (or unplug your ethernet)
  2. Wait 5 seconds
  3. Send a message to Hermes: "Are you still there?"
  4. The agent should reply normally
  5. When done, turn Wi-Fi back on

What just happened

The agent ran on your laptop. No internet. The model, the tools, the memory — all local. The only thing that requires internet is the cloud fallback, which you can also disable if you want everything 100% local.

Check


Lesson 9 — Adjust the model (optional)

If the model is too slow or too dumb, you can swap it. Try a bigger one.

Speed up

If responses feel slow (10+ seconds), your laptop is struggling. Try a smaller model or close other apps.

Make it smarter

If responses feel too short or basic, try a bigger model:

In LM Studio, search for the new model, download it, and select it in the chat dropdown. Then in Hermes, update the model name in settings.

When you outgrow the local model

That's what Course 3 is for — it adds a cloud backup that kicks in only for hard questions.

Check


Lesson 10 — Make it a habit

Every time you start your laptop for the next week:

  1. Open LM Studio (it stays in the menu bar / tray)
  2. Confirm the local server is running (Developer tab)
  3. Open Hermes
  4. Chat normally — the agent is using your local model

If you forget to start LM Studio, Hermes will use the cloud model instead. That's fine for now — the goal is to use the local model most of the time, not all of the time.


You're done with Course 2

When you finish this course, you have:

Next course: Course 3 — Connect Hermes to OpenCode Go. That adds a cloud backup for the hard questions — fast, smart, but only when you need it.


Troubleshooting

LM Studio installer is blocked by Windows SmartScreen

The model download is failing

The agent is slow after connecting to LM Studio

The agent says "I can't connect to the local server"

The model replies don't make sense

I want to use Ollama instead of LM Studio

Hermes is using the cloud model even though LM Studio is running

I need to talk to a person