The post has been translated automatically. Original language: Russian
ChatGPT itself is not an agent. There is often confusion about this.
They opened a chat and wrote: "explain what a tax deduction is." He replied with a text. That's it. It's just a conversation with the language model.
Useful? - yes. Smart? — Sometimes very much. But that doesn't make him an agent.
For me, the boundary is not based on the name of the interface, but on whether the system has the right and opportunity to act. Not to "respond beautifully", but to take a step in the "outside world": open the file, find the data, call the tool, create an email, go to the calendar, update the CRM, check the result and decide what to do next.
Here is the usual working difference.
You're asking for ChatGPT:
"Compose a letter to the client on this case."
He writes the text.
In a similar task, the agency system must first find the client's card, view the history of correspondence, check the status of the transaction, prepare a letter, perhaps create a task for the manager and stop before sending if, according to the rules, a person needs confirmation.
And then there are not beautiful words about AI, but boring engineering questions: what tools are available, what data can be read, what can be changed, where logs are needed, who is responsible if the model confidently misunderstood the task.
The chat itself is a shell. An ordinary AI chatbot that answers questions can live in it. Or maybe an agent system with tools, memory, constraints, and a cycle of actions.
In this article, I discussed in detail what an AI agent, a chatbot, and an AI-based chatbot are. Because without a clear separation of terms, it's easy to order an "agent" and just get a window that politely retells the knowledge base.
The phrase I would keep in mind is: the agent begins where the model ceases to be just an interlocutor and becomes part of the task execution process.
Up to this point, it can be a good ChatGPT scenario. But not yet an agent.
ChatGPT сам по себе — не агент. С этим часто бывает путаница.
Открыли чат, написали: «объясни, что такое налоговый вычет». Он ответил текстом. Всё. Это просто разговор с языковой моделью.
Полезный? — Да. Умный? — Иногда очень. Но агентом от этого он не становится.
Для меня граница проходит не по названию интерфейса, а по тому, есть ли у системы право и возможность действовать. Не «ответить красиво», а сделать шаг во "внешнем мире": открыть файл, найти данные, вызвать инструмент, создать письмо, сходить в календарь, обновить CRM, проверить результат и решить, что делать дальше.
Вот обычная рабочая разница.
Вы просите ChatGPT:
«Составь письмо клиенту по этому кейсу».
Он пишет текст.
Агентская система в похожей задаче должна сначала найти карточку клиента, посмотреть историю переписки, проверить статус сделки, подготовить письмо, возможно, создать задачу менеджеру и остановиться перед отправкой, если по правилам нужно подтверждение человека.
И тут уже появляются не красивые слова про ИИ, а скучные инженерные вопросы: какие инструменты доступны, какие данные можно читать, какие можно менять, где нужны логи, кто отвечает, если модель уверенно поняла задачу неправильно.
Сам по себе чат — это оболочка. В ней может жить обычный ИИ-чат-бот, который отвечает на вопросы. А может быть агентская система с инструментами, памятью, ограничениями и циклом действий.
В этой статье подробно разбирал, что такое ии-агент, чат-бот и чат-бот на базе ИИ. Потому что без четкого разделения терминов легко заказать «агента», а получить просто окно, которое вежливо пересказывает базу знаний.
Фраза, которую я бы держал в голове: агент начинается там, где модель перестает быть только собеседником и становится частью процесса выполнения задачи.
До этого момента это может быть хороший ChatGPT-сценарий. Но ещё не агент.