endueendue
← Releases
2026.10.05

Your agent comes to your site

One line of code opens an endue Live chat window inside your own site, and the endue.live page is yours to shape. One endue API key now calls models straight from the OpenAI and Anthropic SDKs or Claude Code, paid from your credits. Studio shows usage per agent. Artifacts are now Outputs and Abilities are now Resources. Dream tidies memory while an agent rests, and answers start sooner.

endue 2026.10.05 release key visual

Highlights

endue Live on your own site

Register your site, paste one line of code, and a chat window with your agent opens right on your own pages. You can shape the endue.live page too.

LLM API

Call models directly with an endue API key. It takes OpenAI and Anthropic requests as they are, so you only swap the base URL and the key, and each call comes out of your credits.

Usage in Studio

The Usage tab on the canvas shows the conversations one agent took over channels, Live and the API, and the credits each path used.

Outputs and Resources

Everything an agent makes sits in one Outputs list, and everything you give an agent sits in one Resources panel. You can also pick the format you want back, like slides or Excel.

Dream

Your agent tidies its memory while it rests. It merges duplicates and fixes stale facts, and you set its sleeping hours on a round clock.

Answers that start sooner

The checks before a run starts are fewer and run together. The first words of a new conversation arrive sooner.

endue Live on your own site

Studio, Embed on your site: two allowed sites, a preview of the floating button, and the one line of code to paste

The last release gave your agent an endue.live/@handle address. Anyone with the link could talk to it, but they had to come to endue.live to do it. Now the agent comes to your site.

In the Live panel in Studio, open Embed on your site, register your site’s address, and paste the one line of code it gives you just before </body>. A button appears in the corner of your site, and clicking it opens the same chat window as endue Live, right there. You can also place the window inside a page with no button. The button’s position, icon and text and the window’s theme are all settings, and a preview shows how it looks on a desktop and on a phone. The window only opens on sites you registered. Pasting the code anywhere else does nothing. Saved changes reach your site within a minute.

Studio, Customize page: fields for the intro and what the agent is good at, with a preview of the endue.live first screen

The endue.live page is yours to shape. In Customize page you set a one-line intro, a welcome message, questions visitors can tap to get started, links, an accent color and a background. You can write Korean and English separately, and the preview changes as you type.

Conversations on endue.live changed too. In the last release the agent showed its answer in one piece once it was done. Now the answer streams in as it is written, one line shows what the agent is doing, and visitors can stop it midway. In the Live settings you also choose whether the agent answers in the visitor’s language or always in the one you set.

LLM API

Settings, Account, LLM API: the base URL and a Python example

Until now, using a model on endue meant going through an agent. Some jobs need an agent with a prompt, memory and tools. In your own code, one model call is often all you want.

You can now call models directly with an endue API key. Requests use the OpenAI Chat Completions and Anthropic Messages formats as they are. If your code already uses the OpenAI or Anthropic SDK, set the base URL to https://platform.endue.ai/api/v1/llm and the key to your endue key. Claude Code connects with two environment variables. Settings › Account › LLM API has the address and examples for Python, Node.js, curl and Claude Code, ready to copy.

Calls are paid from your credits. Before the model is called, endue holds the most that call could cost. When it finishes, only what it used is charged and the rest goes back. If your balance is too low, the model is not called and you get a 402. The amount charged comes back in the response as usage.cost. The models you can call and their per-token prices are on the models page and at GET /llm/models. One credit is US$1.

The LLM API draws only on credits you bought or received in a promotion. Your plan’s allowance is for agents and is not used here. Use an account-wide key from Settings › Account › API keys. A key restricted to one agent cannot call models directly.

The Sources tab in Usage: daily tokens split into Agents, Agent API, endue Live and LLM API, with usage by source

What you spend shows up in the Sources tab in Usage. There are four sources: Agents, Agent API, endue Live and LLM API, each with the models it used. Request formats and error codes are in the LLM API docs.

Usage in Studio

Usage tab on the Studio canvas: conversations and credits on each path, the Analytics panel on the right, the daily chart below

Usage already showed what your whole account spent. There was no place to see what one agent spent and on which requests: how many credits visitor conversations take once Live is on, or which API key calls the most among the services you connected.

The Studio canvas now has a Usage tab. It is the same canvas you see in Setup, with numbers on every path a request comes in by. Channel connections, the Live address and API keys show their conversations, replies and credits, and the lines into the agent get thicker with use. Pick 7, 30 or 90 days.

Click a node and the Analytics panel opens on the right. A table lists conversations, replies, tokens and credits for six sources (App, API, Live, Channel, Routine and Space), and another breaks them down by API key and channel connection. The chart under the canvas stacks each day by source, and it can show just the path you clicked. Only the agent’s owner sees the Usage tab.

Artifacts are now Outputs

Outputs panel: slides, a sheet, docs, a page and apps in one list

The things your agents made lived in two places. Documents and files were artifacts, and anything built from several files was an app. To find something you first had to remember which kind it was, and the word artifact needed explaining every time.

Both are Outputs now. Open Outputs in the left rail and files and apps sit in one list, ordered by when they last changed, each card labeled doc, sheet, slides, page or app. One search covers both. The outputs list beside a conversation and the ⌘K palette use the same list. Old /artifacts and /apps links still open.

The format menu in the composer (Auto, Document, Slides, Excel, App) with a slide preview open on the right

You can also pick the format you want back. Choose Document, Slides, Excel or App from the format chip in the composer, send your request, and the agent makes it in that format. Slides are saved as a .pptx file and Excel as .xlsx, and both preview in the browser without a download. Leave it on Auto and the agent decides. The choice applies to that one request.

Abilities are now Resources

Resources panel: Skills, Connectors and Devices, with the list of connected connectors

Skills and connectors used to be grouped as Abilities, meaning what an agent can do. Then the computers, browsers and phones an agent works on moved into the same place, and the name stopped fitting. A device is hard to call an ability.

So the name is now Resources: everything you can give an agent, in one place. The Resources panel in the left rail has three areas, Skills, Connectors and Devices. Skills and connectors each split into a Discover tab and the ones you have. Devices groups Computer, Chrome, Android and Desktop by product, and what used to be called Computer and Chrome separately are all devices here. The agent card in Studio uses the same words, in three groups: Identity, Resources and Inventory.

Old /abilities addresses lead to the same place under Resources.

Dream

Trait panel in Studio: Dream switched on, with a round clock set to bedtime 01:30 and wake 07:00

An agent you use for a long time piles up memories. The same thing gets written twice, a fact that changed stays in its old form, and a commitment that is done stays open. Until now the cleanup was a routine attached to the Self-evolving trait.

That cleanup is now a trait of its own, called Dream. Turn it on and the agent tidies its memory while it rests. It merges overlapping memories, fixes stale facts, closes finished commitments and archives what is no longer useful. Then it adds a one-line summary and search terms to each memory so the next conversation finds the right one faster. Archived memories are kept under Archived in Memory, where you can restore them. Pinned memories are never changed or archived.

You set the sleeping hours on a round clock. Its two handles are bedtime and wake time. Cleanup starts at bedtime, and once wake time has passed it will not start that day. Pick the days below the clock. On a day with no changed memories and no new conversations since the last cleanup, Dream skips the run and uses nothing. After each cleanup a report of what it merged, fixed and archived is saved to Outputs.

Self-evolving now covers what the agent learns during a conversation. Teach it a report format or a sequence of steps and it saves a skill and follows it from the next conversation. It also keeps a list of the files it made so it can open the right one. Skills and files are only learned in conversations with the owner. The skills it learned are listed in the trait settings, newest first, and you can revert any of them.

Answers that start sooner

Most of the wait between sending a question and seeing the first words was being spent outside the model. Before a run started, the conversation, permissions, skills and connectors were checked one after another, and the run itself happened far from the data.

Checks that read the same thing twice are gone, and checks that do not depend on each other now run together. New conversations run close to the data. When an answer ends, the run reports done first and leaves cleanup such as naming the conversation for afterwards. Measured in our development environment, the stretch from receiving a request to calling the model went from about 6.5 seconds to between 1 and 2, and a short question gets its first words in about 3 seconds. Existing conversations keep running where they were, so new conversations see the biggest difference.

Effort in the composer now starts at the model’s default in a new conversation. It used to start at Medium every time, which made the model think longer than its default. Choosing None now turns reasoning off.

Everything else

  • Pinned agents are saved to your account. They are still pinned after you sign out or open endue on another device.
  • The Agents list no longer shows an Online label that meant nothing. In its place are counts of conversations that need you, are running or finished recently, and scheduled routines, with marks for Live, API and connected connectors.
  • The @handle in the top bar shows in full when there is room.
  • Agents that are Live get an ON AIR sign and a Continue on Live button on their home. The agent’s Live tab is now called Home.
  • Conversations with endue Live visitors stay out of your own conversation list and search and are kept apart.
  • Device names lost the endue prefix. They are Computer, Chrome, Android and Desktop, told apart by product and OS logos.
  • Fixed the Connected connectors table in the Resources panel breaking at tablet widths.
  • Two new use cases: publishing an agent with endue Live, and an on-call agent.

Security

API keys and tokens that appear in a conversation are masked by default. The prefix stays and the rest turns into dots, and copying the masked text still copies the real value.

How far memory reaches now depends on who the agent is talking to. endue Live visitors and callers who reach the agent through the API read only the agent’s shared memory. They cannot see the owner’s personal memories or save new ones. When someone in a channel says “remember this”, it is saved as that person’s memory only.

The LLM API accepts account-wide keys only. Someone holding a key restricted to one agent cannot use it to call models directly.

endue.live pages cannot be framed by other sites. Only the embedded window opens elsewhere, and only on sites the owner registered.

endue Live8
  • LiveEmbed on your site. Register your site, paste one line of code, and an endue Live window opens on your own pages.
  • LiveThe window can be a floating button or sit inside a page. Set the button’s position, icon and text and the window’s theme.
  • LiveThe embedded window only opens on sites you registered. endue.live pages cannot be framed by other sites.
  • LiveCustomize page. Set a one-line intro, a welcome message, starter questions, links, an accent color and a background.
  • LiveConversations on endue.live stream the answer as it is written, show progress in one line, and can be stopped midway.
  • LiveResponse language. Follow the visitor’s language or always answer in the one you set.
  • LiveAgents that are Live get an ON AIR sign and a Continue on Live button on their home. The Live tab is now Home.
  • LiveConversations with endue Live visitors stay out of your own list and search and are kept apart.
LLM API9
  • APIThe LLM API is open. The base URL is https://platform.endue.ai/api/v1/llm.
  • APIIt takes OpenAI Chat Completions and Anthropic Messages requests. Streaming comes back in each format too.
  • APIClaude Code connects with two lines: ANTHROPIC_BASE_URL and ANTHROPIC_API_KEY.
  • APIGET /llm/models lists the models you can call and their per-token prices. It needs no key.
  • BillingA call holds its maximum cost up front and charges only what it used. If the balance is too low, the model is not called and you get a 402.
  • BillingLLM API calls are paid from purchased and promotional credits only.
  • SettingsSettings › Account › LLM API has the base URL and examples for Python, Node.js, curl and Claude Code.
  • UsageUsage has a Sources tab that splits Agents, Agent API, endue Live and LLM API.
  • DocsLLM API docs are up in English and Korean.
Studio4
  • StudioThe canvas has two tabs, Setup and Usage. The Usage tab covers 7, 30 or 90 days.
  • StudioChannel, Live address and API key nodes show their conversations, replies and credits.
  • StudioAn Analytics panel with tables for the six sources and for each API key and channel.
  • StudioA daily chart under the canvas, stacked by source. Pick tokens, credits or replies.
Outputs and Resources8
  • OutputsArtifacts and apps are one Outputs list. Each card says what it is: doc, sheet, slides, page or app.
  • OutputsOld /artifacts and /apps addresses lead to /outputs.
  • OutputsA format chip in the composer: Auto, Document, Slides, Excel or App.
  • OutputsPowerPoint (.pptx) and Excel (.xlsx) files preview in the browser.
  • ResourcesAbilities are now Resources. The panel has three areas: Skills, Connectors and Devices.
  • ResourcesOld /abilities addresses lead to /resources.
  • ResourcesDevice names are shorter (Computer, Chrome, Android, Desktop) and carry product and OS logos.
  • ResourcesFixed the Connected connectors table breaking at tablet widths.
Traits and memory7
  • TraitsThe Dream trait. While the agent rests it merges, fixes and archives memories, then adds a summary and search terms.
  • TraitsThe Dream clock sets bedtime, wake time and days. Days with nothing new are skipped.
  • TraitsSelf-evolving saves procedures you teach as skills and keeps an index of the files it made.
  • TraitsA Learned skills list, with Revert.
  • MemoryMemory has an Archived group. Archived memories can be restored.
  • MemoryMemory search matches by word, so a different phrasing still finds it.
  • Memoryendue Live visitors and API calls cannot read the owner’s personal memories.
Chat and agents8
  • ChatFewer checks before a run starts, and they run in parallel, so the first words arrive sooner.
  • ChatWhen an answer ends, the run reports done first and names the conversation afterwards.
  • ChatEffort starts at the model default in a new conversation. None now turns reasoning off.
  • ChatAPI keys and tokens in a conversation are masked by default. Copying gives the real value.
  • AgentsPinned agents are saved to your account and survive sign-out and a change of device.
  • AgentsThe Agents list drops Online for counts of needs-you, running, recently finished and scheduled, plus Live, API and connector marks.
  • AgentsThe @handle in the top bar shows in full when there is room.
  • UsecasesTwo new use cases: publishing an agent with endue Live, and an on-call agent.