PTENES
MODULE 2.5

🔌 Connect the local model to the Hermes Agent

You already have the brain (qwen3-coder-64k), and now comes the body: the Hermes Agent. In this module, you install/update Hermes, point it to your local model, and prove that everything runs 100% on your machine — with no calls to the cloud.

6
Topics
~25
Minutes
Practical
Level
Hands-on
Type
1

🌐 What the Hermes Agent is

O Hermes Agent and the "AI OS" we saw in Track 1: the home where memory, skills, connections, and agents live. It's open-source, with MIT license, maintained by Nous Research. Under the hood, Hermes needs a model to do the reasoning — and this is where your qwen3-coder-64k comes in.

Nous Research’s Hermes Agent website, showing that it’s open source under the MIT license, with the motto THE AGENT GROWS YOU
Frame from the official site: notice open-source / MIT license and in the signature Nous Research. Open-source + MIT means you can freely read, run, and modify the code — it fits the course’s emphasis on ownership.

New here? "Open-source" means the code is open for anyone to see and use. The "MIT License" is one of the most permissive licenses: you can use, copy, and even modify it commercially, as long as you keep the attribution notice. It’s the opposite of a closed black box.

Key concepts

Hermes Agent

The AI OS that connects memory, skills, and models.

Open source MIT

Open-source code and a permissive license.

Nous Research

The organization behind Hermes.

Needs a model

Hermes reasons using the model you point it to.

2

⬇️ Install / update Hermes

With Hermes already on the system, the command that ensures you have the latest version is hermes update. In the video, after it, the dashboard appears "HERMES IS READY" with the button Launch Hermes — a sign that the agent is ready to open.

🎯 Objective

Update Hermes to the latest version and confirm it’s ready.

hermes update

How to verify: when it’s done, the panel appears HERMES IS READY with Launch Hermes. If you want a quick status check at any time, run hermes status.

Terminal running hermes update and the HERMES IS READY panel with the Launch Hermes button
Video frame: the terminal runs hermes update and the result is the dashboard HERMES IS READY. That "READY" is your green light to move on to connecting the model in the next topic.

Key concepts

hermes update

Update Hermes to the latest version.

HERMES IS READY

Panel that confirms it’s ready.

Launch Hermes

Button that opens the agent.

hermes status

Quick check of the current state.

3

🔗 Select the local model

This part is in the Hermes interface, not in the terminal. In the model selector, you choose a local model instead of a cloud one — in our case, the qwen3-coder-64k that we created in module 2.4. Once selected, the active model appears in the bottom-right corner of the screen.

🖱️ Steps in the interface (faithfully described from the video)

1

Open the model selector

In Hermes, open the model list. Local models (from Ollama) appear alongside cloud models.

2

Choosing qwen3-coder-64k

Select the local model you created. It’s what gives the agent 64k of context.

3

Confirm in the bottom-right corner

The active model appears in the bottom-right corner. If it shows the name of your local model, the connection is set up.

Honesty: the exact menu names may change between Hermes versions. What matters is the flow—open the selector, choose a model local, and see its name in the lower-right corner. There is no "CLI flag" for this: it's an action in the interface.

Key concepts

Model selector

Where you choose local or cloud in the UI.

Local model

The Ollama one (qwen3-coder-64k), not the cloud one.

Bottom-right corner

Where the active model is displayed.

UI action

Connect and click, don’t type a command.

4

📏 The 64k requirement

Now it’s clear why we did the work for module 2.4. Hermes needs a large window to fit the agent system—instructions, memory, tool descriptions, and conversation history. If you point it to a model with a short context window, the agent freezes or forgets things along the way.

the local model’s 64k context window agent instructions tool descriptions memory / task context conversation history Hermes Agent uses the local model to reason short context = the window fills up and the agent "forgets"—that's why 64k

Everything the agent needs to remember (instructions, tools, memory, and conversation) shares the same window. With 64k, there's room to spare; with a short context, the window fills up and the agent loses track — exactly the problem qwen3-coder-64k solves.

💡 Practical tip

If the agent starts ignoring instructions or "loses its way" midway through a task, suspect the context. Check that the selected model is the 64k one (not the chat one) by looking at the bottom-right corner.

Key concepts

Agent system

Instructions + tools that take up context.

Full window

A short context fills up and the agent forgets.

That's why 2.4

The 64k model exists to power Hermes.

Wrong model = failure

Selecting the chat one locks up the agent.

5

🩺 Troubleshooting when something doesn't add up

If the connection seems off, Hermes gives you two health-check commands. The hermes status shows the overall status; the hermes doctor does a deeper check and points out common issues. These are the first places to look before changing anything else.

🎯 Objective

Check Hermes’s status and run a diagnostic when the model connection doesn’t seem right.

hermes status
hermes doctor

How to verify: hermes status summarizes the current state; hermes doctor lists checks and flags anything out of place. If it points something out (e.g., Ollama is unreachable), fix that first and run it again.

✓ Before asking for help

  • ✓Is Ollama running? (ollama list responds)
  • ✓The 64k model appears in ollama list?
  • ✓Does the bottom-right corner show the local model?
  • ✓hermes doctor no red alerts?

✗ Common pitfalls

  • ✗Ollama is closed: Hermes can’t find the model.
  • ✗You selected the chat model, not the 64k one.
  • ✗RAM at its limit: the model won’t even load.
  • ✗Forgot the hermes update and is running an old version.

Key concepts

hermes status

Summary of the agent's state.

hermes doctor

In-depth check; points out problems.

Ollama running

Prerequisite: the service needs to be up.

Resolve and repeat

Fix the alert and run the doctor again.

6

✅ Test the connection (100% local)

The final test is simple: with the local model selected, send a “hi” to the agent. If it responds, you have an agent running entirely on your machine. The privacy proof comes next: disconnect from the internet and send another message—if it keeps responding, nothing was being sent to the cloud.

1

Send a "hi"

With qwen3-coder-64k active, write a simple message and watch the response arrive.

2

Disconnect from the internet

Turn off Wi-Fi (or unplug the cable). This is the “Vault mode” test we’ll see in Track 3.

3

Send another message

If the agent still responds without a network connection, that proves everything runs locally and nothing leaked.

🎉 What you just built

A complete agent—Hermes + Ollama + qwen3-coder-64k—reasoning 100% on your machine, for free and without internet. This is the heart of the course. In Track 3, you’ll use this foundation in real projects.

Key concepts

“Hi” test

The fastest way to validate the connection.

Offline test

No network, and it still responds = proof it runs locally.

Bridge to Vault

Disconnecting is the foundation of air-gapped mode.

Full agent

Hermes + Ollama + 64k model together.

Optional self-check: What’s the best proof that the agent runs 100% locally?

🎯 Module summary

✓
Hermes = open AI OS — open source, MIT-licensed, from Nous Research; it needs a model underneath.
✓
hermes update — updates and takes you to the “HERMES IS READY / Launch Hermes” panel.
✓
Select the local model — in the UI, choose qwen3-coder-64k; it appears in the lower-right corner.
✓
Diagnostics and testing — hermes status/doctor for problems; send an offline “hi” to prove it’s 100% local.

Next module:

2.6 — Desktop app, terminal, and Telegram