AI Assistant vs Browser Automation: Where Each Fits in a Day
Maker's note: 기준일 2026-10-03. We're the AZET team. This post separates two things that get discussed as if they were one — AI assistants and browser automation — using our own tools as the worked examples. Status statements are current as of October 3, 2026.
"AI assistant" and "browser automation" often arrive in the same sentence, and the confusion is expensive: people expect judgment from a script and reliability from a model. They are different tools with different failure modes, and most ordinary days have room for both. Here is the distinction as we draw it — including where our own products actually stand, which is not the same place for each.
What browser automation actually is
Browser automation is a script driving a browser through steps a person wrote in advance. It is the oldest, least glamorous category here, and still one of the most useful. Our own version lives in the Azet app and CLI — a macOS tool, currently at version 0.3 — whose task language is explicit enough to read aloud:
open https://example.com and return the page title
click "Learn more"
type "hello" into "Query"
select "Seoul" from "Destination"
extract table "Results"
export final list to CSV
assert title "Example Domain"
Every step names its action and its target. Steps chain with "and" or "then," a run ends with assertions about the title, URL, or text of the final page, and nothing about the sequence is decided at runtime. That rigidity is the point: the same chore — check a page, fill a form, pull a table into a CSV — happens the same way every time, and its success is checkable rather than vibes.
What an AI assistant actually is
An assistant takes an open request and decides the steps itself. You don't hand it a choreography; you hand it a goal, and it works out what to look at, in what order, and when the goal is met. That is what azet, the assistant at the center of azet.io, is being built for — the input bar's one-line instruction reads "Ask anything. Get it done." — and it is why the assistant's hard problems are judgment problems: noticing when a page contradicts an assumption, knowing when to stop, and saying "I couldn't verify this" instead of smoothing over the gap.
As of this date, azet has not launched; it is in an early-access waitlist joined by email. We would rather say that plainly than blur it.
The middle ground, and the line we drew
Between scripted steps and free judgment sits the interesting territory: letting something choose the next action while keeping the choices bounded. We drew a hard line in the Azet app, and it is worth describing because most of the confusion lives exactly here: text-generating models are used only to produce text. A language model receives text and returns text; it never emits a click, a selector, a coordinate, or code that the browser then executes. When a task does need a decision — which button, which option — the app sends a bounded set of action and target candidates to a small classification model and executes only responses from an allowed set. Judgment is invited, but through a door with a guest list, not an open mic.
Side by side
| Browser automation | AI assistant | |
|---|---|---|
| Who decides the steps | You, in advance, step by step | The assistant, as the task unfolds |
| Best at | Repetitive, well-defined chores | Open questions and mixed errands |
| Verification | Assertions on title, URL, text | Harder — judgment needs review |
| Typical failure | Halts when the page changes | Can be confidently wrong |
| Cost shape | Predictable, often trivial | Model use per question |
Where each fits in a normal day
A concrete morning shows the split. A scripted check that a supplier's price page still loads, with today's table exported to a CSV before you're awake — that's automation. It was worth writing down once, and it's worth trusting precisely because it never improvises. "Three suppliers raised prices this quarter — which contract should I look at first?" — that's an assistant question: it needs reading, comparison, and a recommendation you can check. The assistant's value is being a single place to bring the second kind of question; the script's value is making the first kind disappear. Neither substitutes for the other, and a tool that claims both without distinction should be asked which one it actually is.
Where ours stand, honestly
Two different statuses, stated separately. The azet assistant: in waitlist, not launched, no date announced, no pricing published. The Azet app and CLI: real and working today — Chromium sessions per profile, a persistent JavaScript REPL, a usage ledger that accounts for model use in credits, with spending caps and step limits on every run — and also visibly unfinished: the desktop build is signed ad hoc rather than notarized, macOS still has first-run steps to walk through, and we do not present it as a finished product. We mention it here not as a pitch but because it is where our opinions about agentic browsing were earned.
FAQ
Is an AI assistant just browser automation with a chat box? No. Automation executes steps a person wrote in advance; an assistant decides its own steps toward a goal. The Azet app keeps the two separate even internally — text models produce text, and only a bounded classifier selects executable actions.
Which is better for repetitive web chores? Automation. Deterministic steps with final assertions are cheaper, more predictable, and easier to trust for work that never changes shape.
What is the azet assistant's status? Not launched. As of October 3, 2026, it is in an early-access waitlist joined by email at azet.io, with no announced launch date or pricing.
Can a language model click things on its own in the Azet app? No. Language-model output is treated as text only; browser actions come from explicit instructions or an allowed-response classifier, never from free-form model text.
One distinction, kept
If you remember one sentence: automation is a choreography, an assistant is a judgment you can question. The waitlist at https://azet.io is where the assistant side of that line is being built, and these notes will keep the two labeled separately as it grows.