Harness as a Service
Not a search API.
The agent harness for vertical SaaS teams.
Agents. Skills. Workflows. Runners. Scheduled tasks. Injectable prompts. A white-label embed SDK. Everything your vertical SaaS product needs to ship AI features—already built.
Embed It
Ship inside your product
Two ways in, same harness. Drop in our white-label UI and ship this afternoon, or build your own UI on the REST API — every agent, workflow, runner, tenant, and document is an endpoint. Both from Business. Either way it ships as part of your app, themed to your brand, authenticated to your users.
REST API — everything, not just embed
Anything the UI can do, the API can do: create tenants, attach skills, run agents, start workflows, schedule tasks, ingest documents, read results. Build your own front-end or wire agents into your backend with no UI at all.
Business & above, full parity with the UI.
- Scoped
vapi_keys (read / write / admin) - Ephemeral tokens for the browser
- Webhooks and WebSocket for results
Embed SDK
@vectrify-ai/ui drops the harness into your product as one component: <VectrifyEmbed /> for React, a vanilla JS loader for everything else.
Business & above.
npm install @vectrify-ai/ui- Ephemeral token auth — no long-lived keys in the browser
- Context object carries injectable prompts
Full theming API
Override every CSS variable at the account, tenant, or runtime level. Light and dark mode, per-tenant themes.
Business & above.
- CSS variable overrides
- Light + dark mode
- Per-tenant theme overrides
- Runtime override via embed context
Injectable prompts
Named slots in your prompt templates take runtime values — user name, account tier, date, feature flags — with no redeploy.
- Passed via API or embed SDK context
- Per-session override
- Audit-logged per invocation
Scoped API keys
Issue read, write, or admin keys per integration. Rotate without downtime. Every use is audited.
Make It Act
Agents that act, not chat
Configure an agent once — it answers, executes, delegates, and proposes actions your app can commit, without you writing orchestration code.
Configurable agents
Set reasoning depth, tone, context window, and persona per agent.
- Subagents spawn, delegate, and aggregate results
- Departments group agents and auto-route by topic or workspace
Propose / commit actions
Agents don't just answer — they propose an action and your app commits it. Live in the HelpDesk demo: auto-triage, copilot panel, trace drawer.
Skills
Packaged, versioned units of capability. Write one, attach it to any agent across any tenant; update once, every agent picks it up.
- Custom skills:
SKILL.mdmanifest + reference docs + runnable scripts, Python runtime - First-party library: extraction, summarization, structured output, web search, CRM lookups — all plans
- Starter 10
- Business 25
- Platform/Scale 100+
Workflows
Chain agents, skills, and data operations into projects with named stages. An LLM routes between stages; every step streams to your UI over WebSocket.
Keep It Running
Work while you sleep
Runners and scheduled tasks give agents a place to live between requests — persistent, monitored, and driven on a clock, not just a chat turn.
Runners
The managed execution context for an agent: tool calls, retries, memory, session state, result persistence. Start, pause, resume, cancel.
- Per-tenant run controls and limits
- Streamed to WebSocket or webhook
- Full execution log per run
Scheduled tasks
Point a cron expression at any agent and prompt. Vectrify fires it on schedule, stores the result, calls your webhook when done.
- Per-tenant schedules, each customer's own cadence
- Webhook on completion
- Pause, skip, or run now
Local runner
A Go runner daemon handles work that has to happen on the customer's own machine — same lifecycle, running where the data lives.
Isolate & Meter Per Tenant
Built for multi-tenant resale
Every tenant is isolated by default and metered on pages and tenants — never seats, never runs — so you can build your own pricing on top.
Workspaces
Each workspace is an isolated knowledge namespace, isolated per tenant. Connect multiple sources; assign one or many workspaces to any agent. Pages pool for billing.
Connectors & ingestion
Pull data from where your customers already keep it: Google Drive, Confluence, SharePoint, Notion, Slack, GitHub, Jira (OAuth), or direct upload of PDF, DOCX, HTML, and MD. Parsing, chunking, embedding, and indexing run automatically, with live status.
- Hybrid search — vector + BM25 — per workspace
- Auto re-ingestion: watch a Drive folder; new, edited, and deleted files reflow automatically
- Dedicated Qdrant namespace per tenant — no shared indexes
Per-tenant run controls
Cap daily or monthly runs per tenant from your dashboard. Build your own tier structure on top of Vectrify's.
BYOK per tenant
Bring your own model keys, for your account or for each of your tenants individually. Your customers can run on their own model accounts.
One-call provisioning
Provision a new, fully isolated tenant with a single API call. No infrastructure work on your end.
Six months of engineering. Available today.
14-day free trial. 1,000 pages. All features. No credit card required.