Harness as a Service

Not a search API.
The agent harness for vertical SaaS teams.

Agents. Skills. Workflows. Runners. Scheduled tasks. Injectable prompts. A white-label embed SDK. Everything your vertical SaaS product needs to ship AI features—already built.

Embed It

Ship inside your product

Two ways in, same harness. Drop in our white-label UI and ship this afternoon, or build your own UI on the REST API — every agent, workflow, runner, tenant, and document is an endpoint. Both from Business. Either way it ships as part of your app, themed to your brand, authenticated to your users.

REST API — everything, not just embed

Anything the UI can do, the API can do: create tenants, attach skills, run agents, start workflows, schedule tasks, ingest documents, read results. Build your own front-end or wire agents into your backend with no UI at all.

Business & above, full parity with the UI.

  • Scoped vapi_ keys (read / write / admin)
  • Ephemeral tokens for the browser
  • Webhooks and WebSocket for results

Embed SDK

@vectrify-ai/ui drops the harness into your product as one component: <VectrifyEmbed /> for React, a vanilla JS loader for everything else.

Business & above.

  • npm install @vectrify-ai/ui
  • Ephemeral token auth — no long-lived keys in the browser
  • Context object carries injectable prompts

Full theming API

Override every CSS variable at the account, tenant, or runtime level. Light and dark mode, per-tenant themes.

Business & above.

  • CSS variable overrides
  • Light + dark mode
  • Per-tenant theme overrides
  • Runtime override via embed context

Injectable prompts

Named slots in your prompt templates take runtime values — user name, account tier, date, feature flags — with no redeploy.

  • Passed via API or embed SDK context
  • Per-session override
  • Audit-logged per invocation

Scoped API keys

Issue read, write, or admin keys per integration. Rotate without downtime. Every use is audited.

Make It Act

Agents that act, not chat

Configure an agent once — it answers, executes, delegates, and proposes actions your app can commit, without you writing orchestration code.

Configurable agents

Set reasoning depth, tone, context window, and persona per agent.

  • Subagents spawn, delegate, and aggregate results
  • Departments group agents and auto-route by topic or workspace

Propose / commit actions

Agents don't just answer — they propose an action and your app commits it. Live in the HelpDesk demo: auto-triage, copilot panel, trace drawer.

Skills

Packaged, versioned units of capability. Write one, attach it to any agent across any tenant; update once, every agent picks it up.

  • Custom skills: SKILL.md manifest + reference docs + runnable scripts, Python runtime
  • First-party library: extraction, summarization, structured output, web search, CRM lookups — all plans
  • Starter 10
  • Business 25
  • Platform/Scale 100+

Workflows

Chain agents, skills, and data operations into projects with named stages. An LLM routes between stages; every step streams to your UI over WebSocket.

Keep It Running

Work while you sleep

Runners and scheduled tasks give agents a place to live between requests — persistent, monitored, and driven on a clock, not just a chat turn.

Runners

The managed execution context for an agent: tool calls, retries, memory, session state, result persistence. Start, pause, resume, cancel.

  • Per-tenant run controls and limits
  • Streamed to WebSocket or webhook
  • Full execution log per run

Scheduled tasks

Point a cron expression at any agent and prompt. Vectrify fires it on schedule, stores the result, calls your webhook when done.

  • Per-tenant schedules, each customer's own cadence
  • Webhook on completion
  • Pause, skip, or run now

Local runner

A Go runner daemon handles work that has to happen on the customer's own machine — same lifecycle, running where the data lives.

Isolate & Meter Per Tenant

Built for multi-tenant resale

Every tenant is isolated by default and metered on pages and tenants — never seats, never runs — so you can build your own pricing on top.

Workspaces

Each workspace is an isolated knowledge namespace, isolated per tenant. Connect multiple sources; assign one or many workspaces to any agent. Pages pool for billing.

Connectors & ingestion

Pull data from where your customers already keep it: Google Drive, Confluence, SharePoint, Notion, Slack, GitHub, Jira (OAuth), or direct upload of PDF, DOCX, HTML, and MD. Parsing, chunking, embedding, and indexing run automatically, with live status.

  • Hybrid search — vector + BM25 — per workspace
  • Auto re-ingestion: watch a Drive folder; new, edited, and deleted files reflow automatically
  • Dedicated Qdrant namespace per tenant — no shared indexes

Per-tenant run controls

Cap daily or monthly runs per tenant from your dashboard. Build your own tier structure on top of Vectrify's.

BYOK per tenant

Bring your own model keys, for your account or for each of your tenants individually. Your customers can run on their own model accounts.

One-call provisioning

Provision a new, fully isolated tenant with a single API call. No infrastructure work on your end.

Six months of engineering. Available today.

14-day free trial. 1,000 pages. All features. No credit card required.