Skip to content

WebAgent ​

WebAgent is Qoni's web action layer. It lets agents perform search, extraction, browser operations, tracking, and controlled web tasks on the open web. After reading these docs, you can integrate WebAgent through the current Node SDK or REST API and stream execution results.

WebAgent is not a traditional crawler SDK. It is designed for the LLM agent execution workflow: developers provide an instruction, while WebAgent manages runtime resources, page state, retries, structured results, and the run lifecycle. The Console and SDKs are clients for the same API.

An agent that lives in your browser ​

Tell WebAgent what you want done, and it turns the request into a sequence of executable browser actions: open a page, log in, search, fill a form, click, read the result, and return it to your business system.

In enterprise production, about 90% of computer work happens in a browser. WebAgent lets AI take over that work: people no longer have to click, copy, and switch between pages one step at a time. Give it the goal, and the agent in your browser can carry the task through.

WebAgent is more than a click bot: it turns the browser into an execution surface that business systems can call:

  • Faster: reuse sessions, Profiles, and browser context to avoid repeated startup, login, and waiting.
  • More resilient: manage page state, retries, timeouts, and human confirmation together, so failures are visible and recoverable.
  • More controllable: isolate permissions and spend by project, then follow every action through SSE events.
  • Easier to integrate: create a run from one instruction while Console, OpenAPI, and raw HTTP share one contract.

Traditional browser vs. WebAgent ​

Comparison areaTraditional browserWebAgent
Primary roleDisplays pages and waits for user inputUnderstands a goal and executes it in the browser
How work happensThe user clicks, types, and switches pages step by stepTurns a request into executable actions and keeps going
Who does the workThe user completes every stepThe agent handles repetitive work; the user confirms key points
Task continuityState often has to be reconstructed after a page or session endsSessions, Profiles, and runs keep the context
ResultsResults stay on the pageStructured results and SSE events return to the business system

Use cases ​

WebAgent is for the work people do in a browser every day, but should not have to repeat by hand:

  • Enterprise marketing: collect channel and customer feedback, turn it into insights, and draft or publish content automatically.
  • Sales and customer success: log in to the CRM, summarize customer activity, update follow-ups, and trigger the next action.
  • Operations and commerce: collect prices, inventory, and promotions across sites, update back-office systems, and flag anomalies.
  • Research and intelligence: search public sources, cross-check evidence, and keep watch on competitors and industry changes.
  • Internal workflows: query systems, fill forms, download files, reconcile records, and write results back to the business system.

Toward the agentic web ​

Qoni's goal is to make Agents first-class citizens of the Web. WebAgent is not only a background click simulator, and not only a renderer for humans; it places Human, Agent, and Web in the same auditable collaboration layer, so an Agent can read, act, produce results, and render state back to people within an explicit permission boundary.

Toward the agentic webA layered model: a person delegates scoped, audited authority to the agent, which works on the web in hosted cloud browsers and hands the results back.HTMLWebAgentHumanDelegate to agentEvery step auditedResults to humanHosted cloud browsersNo API neededSign-ins carry overRuns tasks in parallel

When to use WebAgent ​

  • Your agent needs real-time web data instead of relying only on model training data or a fixed knowledge base.
  • You need search, extraction, browser actions, and long-running tasks behind auditable APIs.
  • You want the Console, SDKs, and backend services to share the same REST API contract.
  • You need a session / run model for state, event streaming, or steps that require user confirmation.

When not to use WebAgent ​

  • The task only needs your own backend APIs and does not need open web access.
  • You need large-scale offline crawling, warehouse synchronization, or search-index construction.
  • The target site's terms do not allow automated access and you do not have the required authorization.
  • You have not defined API keys, project scope, run budget, and failure handling.

Core capabilities ​

CapabilityDescription
DoAnything APIProvide a natural-language instruction. WebAgent chooses tools, runs steps, and returns a result.
Shaped APIsDedicated API contracts for artifact-shaped workflows such as DeepResearch, WebSearch, and Track.
Session / run modelA session owns runtime resources. A run represents one instruction or follow-up action.
Event streamSubscribe to run state, output chunks, errors, and user-confirmation requests through SSE.
SDK and raw HTTPNode.js / TypeScript and cURL docs use the same API semantics; Python uses raw HTTP.

Documentation entry points ​