Skip to main content
🤖 From Chat To Agents
What modern AI can do for work, code, and real life
Victor Szoltysek · 🏗 🤖 🎤 · May 13, 2026
...and why GitHub Copilot is only a thin slice of the experience
More 201 AI talks: Toronto Java User Group (June 2026) · Toronto DevOps Enterprise Meetup (June 2026)
🎯 The Case I’m Making
At work: GitHub Copilot is useful, but not the full modern AI experience.
We are leaving value on the table for tool integration, research, docs,
diagrams, reviews, and follow-through.
In your personal life: Agentic tools are not just for coding. I’ll show real
use cases from my own life where agents helped with training, planning,
writing, paperwork, research, and personal systems.
Call to Action: Make the leap from chat to agents. Try Codex or Claude
Code directly.
📈
Architecture
Leverage
A rare leverage point
There may not be a higher-
leverage skill right now than
learning how to use agents well.
🧠
Model
From Model to Agent
🧠 💬
Model Chat
From Model to Agent
🧠 💬 🤖
Model Chat Agent
From Model to Agent
The question is: how far along this ladder are we?
🧠 💬 🤖
Model Chat Agent
From Model to Agent
🧑✈
Assistant
GitHub Copilot is powerful assistant-layer AI, not the full agent experience.
Useful, but heavily constrained compared with modern agent tools.
Prompt
📥 🔮 📤
🧠
Model
Predict Text Answer
An LLM model is really just a function: text in, predicted text out.
A very sophisticated parrot: text in, predicted text out.
LLM
LIMITATIONS
🍓
Strawberry Test
How many r's are in
strawberry?
There are 2 r’s.
🍓
Ask an LLM how many r's
are in strawberry, or to do
precise counting and
symbol tricks, and the
illusion can breaks. Very
fancy autocomplete is still
autocomplete.
🤦
Strawberry Test
LLM can’t think
Prompt
📥 🔮 📤
💬
Chat
Predict Text Answer
+ Rolling History
The same model, wrapped in rolling history.
LLM
LIMITATIONS
🪪
Remembering Names
My name is Jo.
Sincerely,
[INSERT NAME]
(after a while)
A chat can carry history
forward, but that is not the
same as real memory.
Once enough messages
pile up, important context
gets buried or dropped.
🪪 🤦
LLMs forget
Remembering Names
Context Window
📥 🔮 📤
🤖 Agent
Predict Text Answer/Action
(relevant history,
fi
les, multimodal,
AGENTS.md, skills, plugins, etc. as
needed)
(tool use, loops, etc.)
The same model, wrapped in dynamic context and action.
LLM
LIMITATIONS
🧻
Passing Tests
Fix the failing tests.
All tests are now
passing.
Assertions weakened
Tests deleted / skipped
Edge cases removed
If you don’t de
f
ine success
clearly, the agent will
rede
f
ine it for you. Green
tests, broken behavior.
That’s the fastest path to AI
slop.
🧻 🤦
LLMs cheat
Passing Tests
🔎
Deep Research
(The
f
irst agent most people should try)
* or Claude Research / Advanced Research (paid version only)
ChatGPT
🔎
One of the clearest
ways to feel agentic
work inside ChatGPT. It
handles multi-step
research and synthesis
that plain chat struggles
to do well in one shot.
Deep Research
🔎
Deep Research
(Comparison Example)
Create a Vision deck for how our organizati
on should approach Azure networking over
the next 2-3 years.
Use current Microsoft Azure Well-
Architected guidance, Cloud Adoption Fra
mework material, and recent enterprise arc
hitecture patterns. Compare the main optio
ns, explain where each pattern
f
its, and call
out what decisions need to be made.
Output a slide-by-
slide outline with speaker notes, a decision
matrix, and a target-state diagram.
Check what people are saying after taking the
new AWS Gen AI Pro exam, especially what
was actually on it, what was less important,
and what surprised them.
Prioritize
f
irst-hand reports from places like
LinkedIn and Reddit. I do not want generic
training material, o
ff
icial study guides, or
vendor marketing summaries.
Output a study-focused summary with areas
to prioritize, areas that seem less important,
repeated themes across real test-takers, and
any notable disagreements or uncertainty.
🔎
Deep Research
(Real world signal analysis)
🤖
ChatGPT
Agent Mode
(Agents directly inside of ChatGPT)
A stronger agent mode
built right into ChatGPT.
Useful when you want
action and follow-
through, but do not need
a full Codex workspace.
ChatGPT
Agent Mode
🤖
ChatGPT
Agent Mode
🤖
(Example)
Take these meeting notes, a draft deck,
and a list of stakeholder concerns.
Rewrite the narrative, tighten the slides,
draft a follow-up email, and produce a
next-steps checklist.
🛠
Codex
(The full agent workspace)
* Or Claude Code / CoWork with Anthropic
🛠
Codex
OpenAI’s full agent
workspace for coding
and non-coding tasks,
with
f
iles, tools,
planning, and durable
context. This is where
the same model starts to
feel much more
powerful.
Help me build a native iOS todo app with a clean
SwiftUI interface, local task storage, and the basic
f
lows
people expect: create, edit, complete, delete, and
f
ilter
tasks.
Prefer a tight, production-like scope over too many
features. Start with a solid end-to-end app, sensible
structure, and clean code. Use modern iOS patterns,
keep the UX simple, and avoid overengineering.
Treat quality as part of the deliverable. Include unit
tests, static analysis, and a codebase structure that
feels maintainable.
Ask me any questions you need answered to get
building, and leave me with something tangible,
credible, and easy to demo as a real app rather than a
throwaway prototype.
🛠
Codex
(Coding Example)
Appendix
Extra Functionality
🖼
(More than text)
ChatGPT
Images
* Huge upgrade - April 21/2026 (ChatGPT Images 2.0) - Consistency, Intelligence, Proper Text
🖼
Three useful modes: image
input, image output, and
image generation from an
existing image. It can
understand what you upload,
create something new from
scratch, or use a photo as a
starting point for a new result,
like a home reno visualization.
ChatGPT
Images
🖼
ChatGPT
Images
(Architectural Diagram)
Create a professional Microsoft-style Azure
architecture diagram using o
ff
icial Azure icons
showing Codex CLI securely connecting via
Azure APIs into a hybrid Azure and on-
premises enterprise environment with Entra ID
authentication, Key Vault, private networking,
SIEM/security systems, and security artifact
generation. Use a clean enterprise white-and-
blue presentation style with labeled secure
data
f
lows.
🖼
ChatGPT
Images
(Architectural Diagram)
architecture diagram using o
ff
icial Azure icons
showing Codex CLI securely connecting via
Azure APIs into a hybrid Azure and on-
premises enterprise environment with Entra ID
authentication, Key Vault, private networking,
SIEM/security systems, and security artifact
generation. Use a clean enterprise white-and-
blue presentation style with labeled secure
data
f
lows.
🖼
ChatGPT
Images
(Reno Example)
Create a new image showing me how things
would look if I added a laundry nook under
the stairs with shelving, washer and dryer.
🖼
ChatGPT
Images
(Reno Example)
Create a new image showing me how things
would look if I added a laundry nook under
the stairs with shelving, washer and dryer.
🎤
ChatGPT Record
(Capturing audio context)
⏺ Record Button in ChatGPT App
( MacOS only)
🎤
A fast way to turn spoken
context into something
usable. Great for
recording ideas, debriefs,
or messy thinking, then
dropping the transcript
or markdown into
Projects as durable input.
ChatGPT Record
🔌
Plugins
(Agents connecting to the outside world)
* Connectors in Claude (with more variety
including iMessage/Apple Note)
🔌
A way to connect
agentic work
f
lows to
real tools like
Con
f
luence, MS
Teams, iMessage,
Outlook and let the
work reach outside the
chat.
Plugins
🔌
Plugins
(Calendar Example)
Review my calendar for this
week,
f
ind the best 90-
minute focus block for deep
work, and add it to my
calendar as “AI Workshop
Prep.”
🚘
ChatGPT
CarPlay
(Chat on the go)
* Also now available with Grok
🚘
A new voice-
f
irst
ChatGPT experience in
Apple CarPlay. Start a
new voice chat, continue
a recent chat, or
continue inside a Project
while you drive.
ChatGPT
CarPlay
🌳
Worktrees
(why parallel work does not
have to collide)
🌳
Local means working
directly in the current
directory and live state.
A worktree is a separate
working area, essentially
another branch
checkout, which is useful
for parallel isolated
tasks.
Worktrees
⏰
Automations
(The scheduled agent)
* Schedules in Claude
⏰
Automations
Automations take an
agentic work
f
low and
put it on a schedule. This
is useful for recurring
work like daily planning,
weekly reviews, and
recurring reminders.
⏰
Automations
(Update Example)
Review the latest entries in Active/Journal,
Active/Current Priorities.md, and Active/
Trackers, then create a short morning update
inbox item with: current state, today’s top
priorities, latest workout progression if there is
a new entry, one likely avoidance risk, and one
concrete next step.
Also include a high-signal news section for the
last 24 hours only: important AI updates from
o
ff
icial or primary sources, especially new
models, tools, or product changes from
OpenAI, the Codex app, Anthropic, and Claude.
Include Starship only if there has been a
meaningful update that materially a
ff
ects the
next launch.
Appendix
Context Engineering Fundamentals
🪟
Context
Window
(What the model can see
right now)
🪟
This is the working set
the model can actively
use.
Prompt, history,
f
iles, and
other context all have to
f
it inside it.
Context
Window
🫠
Context
Drift
(Why long threads get
weird)
🫠
As threads get longer,
important context
competes with noise.
That is when the model
starts drifting, forgetting,
and mixing things up.
Context
Drift
🧹
Context
Hygiene
(Keep the input clean)
🧹
Start a new thread when
the task changes. Write a
short summary before
switching threads.
Move durable guidance
into AGENTS.md, PRDs,
or focused docs.
Context
Hygiene
End
🚀
Call To
Action
Make the leap from chat to
agents
🚀
Call To
Action
Do not stop at Copilot.
Evaluate the full agentic
experience.
Do not stop at chat.
Try a work
f
low with context,
tools, and follow-through.
Do not wait for perfect.
Start with one real
architecture or personal
work
f
low.