PTENES
TRACK 1

🌱 Fundamentals

Start from scratch, without any jargon: what exactly is a "Jarvis," what is the LLM that thinks behind it, and the exact moment when AI stops just answering and starts ACTING for you. Three modules that unlock everything else in the course.

1.1 · Chatbot only responds 1.2 · LLM the brain that thinks in text context = memory 1.3 · Agent thinks and ACTS with tools web calendar email

Read from left to right—the trail map: you start from chatbot that only responds (1.1), understands the LLM like the brain and the context as your working memory (1.2), and reaches agent that thinks and ACTS using tools real (1.3).

3
Modules
18
Topics
~2h30
Duration
Beginner
Level
Path progress0%
0 of 18 topics

Learning path map

Detailed content

1.1~50 min

🤖 What is a “Jarvis” — from fiction to your home

The Iron Man assistant fantasy has become something you can build. Here, you’ll understand the concept without jargon: why a chatbot isn’t Jarvis, what an “AI operating system” is, and what you can (and still can’t) do today.

0 of 6 · 0%
What it is:

Iron Man’s Jarvis is the fantasy of an assistant that talks, remembers everything, and DOES tasks. In this course, “Jarvis” = a personal AI assistant that you build yourself — and today this is real, not fiction.

Why learn:

Having the right picture in your head avoids disappointment and overpromising: you’ll aim for what’s possible to build now.

Key concepts:

Personal AI assistant: chat, remember, and act, from fantasy to something you can build.

What it is:

A chatbot like ChatGPT responds. A Jarvis acts in the world: it sends an email, reads your calendar, runs a task on its own. The difference isn't how smart the text is — it's whether it makes something happen.

Why learn:

It's the distinction that organizes the entire course; without it, everything else becomes just "a better chat."

Key concepts:

Responding vs. acting, taking action in the world, "the leap" that defines a Jarvis.

What it is:

Just as Windows organizes programs, memory, and peripherals, a AI operating system organizes the model, memory, tools, and channels so the assistant can "live." It's the home that brings all the pieces together.

Why learn:

This is exactly what you’ll build in the following tracks; having the metaphor ready makes everything easier.

Key concepts:

AI ONLY, model + memory + tools + channels, orchestration.

What it is:

Three things came together: good-enough models, capable personal hardware, and open standards (such as the MCP, which lets tools connect to any agent). That’s why you can build today what was fiction yesterday.

Why learn:

Getting in early on a wave that's still forming is where the advantage lies—you learn before it becomes obvious.

Key concepts:

Capable models, personal hardware, open standards (MCP), timing.

What it is:

CAN do: text, search, calendar, drafts, automations. CAN'T (yet) do: long tasks without supervision—and it hallucinates (makes things up confidently). Calibrating this gets you halfway there.

Why learn:

Knowing the limits helps you avoid the two classic mistakes: trusting too much or giving up too soon.

Key concepts:

Serious cases, hallucination, human supervision, realistic expectations.

What it is:

Fundamentals (here) → Overview (what already exists) → Anatomy (the 6 layers) → Build → Mobile → Children. Each track builds on the previous one; the order was carefully planned.

Why learn:

Seeing the whole path gives you context so each part makes sense at the right time.

Key concepts:

6 tracks, progression T1→T6, from concept to building.

View Full
1.2~50 min

🧠 The brain of the thing: what an LLM is (without jargon)

The engine that thinks behind every Jarvis, demystified: what an LLM is, how it “thinks” by predicting the next word, what context and tokens are, and the choice between running it in the cloud or on your computer.

0 of 6 · 0%
What it is:

A LLM ("Large Language Model") is a program trained on A LOT of text that learned to predict the next word. It’s what runs behind ChatGPT and Claude — the brain that "thinks in text".

Why learn:

It's the central piece of any Jarvis; understanding what it is demystifies everything else.

Key concepts:

LLM, next-word prediction, trained on text, “brain.”

What it is:

It doesn't query a database or “search”: it statistical forecast learned during training. That's why it's creative AND sometimes gets things wrong with complete confidence — what we call hallucination.

Why learn:

Knowing it's a prediction explains both the magic and the mistakes — and teaches you to check answers.

Key concepts:

Statistical prediction, not a database, hallucination.

What it is:

A context window and it’s everything the model "is seeing right now": your conversation + instructions + files. It’s like the RAM of the computer — limited, and it clears when you close the conversation.

Why learn:

Explains why it "forgets" — and lays the groundwork for the persistent memory in Track 3.

Key concepts:

Context window, working memory, RAM analogy, forgotten when closed.

What it is:

The model reads and writes to tokens — word pieces (a common word usually takes about one token; long words become several). The cloud charges by token, and the context window is measured in tokens.

Why learn:

And it’s the unit used to measure cost and context; it shows up in every practical decision.

Key concepts:

Tokens, tokenization, token ≠ word, cost per token.

What it is:

Cloud (Claude, GPT): powerful, you pay per use. Local (via Ollama): free after downloading, private, but requires hardware. Both serve the same purpose: being the brain.

Why learn:

It’s the trade-off that will come up throughout the course—privacy/cost on one side, power/convenience on the other.

Key concepts:

Cloud vs. local, Ollama, pay-per-use vs. hardware, privacy.

What it is:

On its own, the LLM hallucinates, doesn’t know things that happened after its training, and forgets what falls outside the context window. The solution: give it tools (search, read) and memory — exactly what Track 3 teaches.

Why learn:

Every limitation has a concrete solution; knowing what it is helps you move past frustration.

Key concepts:

LLM limits, training cutoff, tools + memory as the solution.

View Full
1.3~50 min

⚙️ From chatbot to agent: when AI starts taking action

The course’s defining moment: when the LLM stops just responding and starts to ACT. Here you’ll understand the agentic loop, what tools are, and the “LLM as an operating system” metaphor driving the entire industry.

0 of 6 · 0%
What it is:

When you give an LLM tools and let it use them in a loop, it becomes an agent: an LLM that can ACT, not just respond. That’s the leap that turns chat into an assistant.

Why learn:

It’s at the center of everything: it’s the definition of "agent" that carries the course through to the end.

Key concepts:

Agent, LLM + tools in a loop, answering vs. doing.

What it is:

O agentic loop and the cycle: receives the request → thinks → calls a tool → reads the result → calls more tools if needed → responds. There's a iteration limit for safety, so it doesn't run forever.

Why learn:

It’s the mechanism that makes an agent “work on its own”; understanding it demystifies autonomy.

Key concepts:

Agentic loop, think-act-read-repeat, iteration limit.

What it is:

Tools (or “tools”) are the agent’s hands and eyes: searching the web, running code, reading a file, sending a message. Without tools, the model “responds like a stranger” — relying only on what it knows by heart.

Why learn:

They connect the brain to the real world; all of Track 3 revolves around them.

Key concepts:

Tool, hands and eyes, action in the world, "reply as a stranger".

What it is:

Andrej Karpathy (2023) proposed: the LLM is the kernel (core) of a new operating system. Context = RAM, tools = peripherals, agents = processes. The industry adopted this mental model.

Why learn:

It's the metaphor that connects "AI OS" to what you already know about computers—and underpins the course.

Key concepts:

LLM-OS, kernel, context=RAM, tools=peripherals, agents=processes.

What it is:

The idea became reality: AIOS (a real “kernel” for agents), OpenAI Operator (an agent that navigates screens) and Google Project Jarvis (an agent that operates Chrome for you). Proof that the mental model works.

Why learn:

Shows that this isn't isolated hype — major labs are building in the same direction as you.

Key concepts:

AIOS, Operator, Project Jarvis, from metaphor to product.

What it is:

In practice, you stop "asking" and start "delegating": instead of requesting a text, you hand over a task. It's the direct bridge to Anatomy (Track 3), where you assemble the 6 layers that make this possible.

Why learn:

Wraps up the Fundamentals with the change in mindset that unlocks the rest of the course.

Key concepts:

Asking vs. delegating, a shift in approach, a bridge to the Anatomy.

View Full