OttBot Vision

Ask your videos a question. Get an instant answer.

Stop scrubbing through recordings looking for the one thing you need. OttBot Vision analyses every word spoken, everything visible on screen, and every piece of text that appears in your videos — and turns it all into a knowledge base you can search in plain language.

The problem

You already have the knowledge. It is stuck in the video.

A product demo. A customer walkthrough. A recorded training session. A discovery call. Between them, those videos contain everything a business knows about its product and its customers. The problem is that getting to that knowledge means watching the video again — scrubbing, note-taking, manually pulling quotes. If you have a library of videos, the problem multiplies. Nobody can keep track of what is in 50 recordings. The knowledge sits there, unused. OttBot Vision exists to change that.

How it works

Upload a video. Start asking it questions.

I

It hears everything

Every word spoken is transcribed with precise timestamps. Ask where the presenter mentions pricing and Vision tells you exactly where to find it — without you watching a second of footage.

II

It sees everything

OttBot Vision analyses what is on screen throughout the video — actions, objects, features being shown. A demo that shows a feature without narrating it is understood just as well as one that explains it verbally.

III

It reads everything

On-screen text — slides, forms, UI screens, captions, signage — is captured and indexed. If it appears in the video, it becomes searchable.

Speak your customer’s language

Serve Customers In Their Own Language

OttBot holds natural conversations in more than one language, so you can look after customers across borders without hiring a separate team for each one. Available today:

English (UK)English (US)FrenchGermanSpanishItalian

More languages are on the way. If you need one that isn’t here yet, just ask and we’ll add it.

Video analysis

Three simultaneous tracks, combined into one knowledge base

When you upload a video, OttBot Vision processes it in three parallel tracks at once. Everything is fused into a single, timestamped, searchable knowledge base — ready to ask questions of straight away.

  • Speech-to-text: every word spoken, precisely timestamped
  • Visual analysis: what is on screen at every moment throughout the video
  • On-screen text (OCR): slides, forms, UI screens, captions, and signage
Before and after

From buried recordings to searchable video knowledge.

OttBot Vision turns hours of demos, calls, walkthroughs, and training videos into timestamped answers your team can use instantly.

Before Manual review

Knowledge Stuck in Video

Teams know the answer is somewhere in the recording, but finding it means rewatching, scrubbing, note-taking, and relying on memory.

45m rewatch time8 recordings open0 clear timestamps
  • "Where did we explain pricing?"Someone has to scrub through the demo
  • "Which call mentioned that feature?"Answers depend on team memory
  • "Can we turn this into training?"Manual notes become another job
  • "What text was on screen?"Slides and UI details get missed

Useful knowledge stays locked inside recordings, even when the business has already captured it.

After OttBot Vision live

Searchable Video Intelligence

Vision analyses speech, visuals, and on-screen text, then turns each recording into a searchable knowledge source with timestamped answers.

Instant video answers3-track analysisExact timestamps
  • Ask in plain languageFind the answer without rewatching
  • Speech, visuals, and OCR combinedWhat is said, shown, and written is indexed
  • Timestamped evidenceJump straight to the exact moment
  • Reusable team knowledgeTurn recordings into training, support, and content

Every processed video becomes usable knowledge for sales, support, marketing, and training.

Advanced features

From individual videos to a whole knowledge library

Two features that turn a folder of recordings into something genuinely useful at scale — whether you have ten videos or a hundred.

IV

Collections

Combine a set of processed videos into one unified knowledge source. Ask a question and get an answer drawn from all the videos at once — great for building training guides or extracting consistent themes from a series of demo recordings.

V

Hubs

A Hub changes how Vision responds to the same video depending on who is asking. A Marketing Hub and a Support Hub can point at the same demo and give meaningfully different answers — without reprocessing anything.

Use cases

What businesses use OttBot Vision for

These are the most common ways Vision delivers value from the first week.

  • Sales teamsAsk demo recordings "does this cover X?" before a follow-up call — instead of rewatching 45 minutes of footage.
  • Training and onboardingTurn a library of training videos into a knowledge base new starters can interrogate directly, at their own pace.
  • Support teamsFind the answer to a customer query in a how-to video in seconds, with the exact timestamp included.
  • Agencies and content teamsExtract key angles, quotes, and themes from hours of client interview footage without a manual review day.
The ecosystem

Vision is where video becomes knowledge the whole ecosystem can use

OttBot Vision works on its own as a searchable video knowledge base. Paired with the rest of OttBot, the knowledge it extracts becomes the foundation for how the business talks to customers, designs automation, and builds its CRM.

Review a recorded sales call and update the contact's CRM record in the same screen via OttBot Data. Use what Vision extracts from customer recordings to design better chatbot flows in OttBot Build. The knowledge in a Vision Collection can become what the live chatbot in OttBot Connect answers from.

Works with the ecosystem

Vision powers the knowledge the rest of OttBot draws on

What Vision extracts from your videos does not have to stay in Vision. It feeds into how the business automates, communicates, and manages customer relationships.

The knowledge in recordings can inform chatbot logic in OttBot Build, feed into CRM records in OttBot Data, and train what the live chatbot in OttBot Connect answers from.

Who it is for

Built for anyone sitting on video they cannot use fast enough

IX

Founders and product teams

You know the product. You have the demos. Vision turns the knowledge inside those recordings into content, FAQs, and training guides — without a blank-page copywriting session.

X

Content teams and agencies

Client interview footage. Strategy session recordings. Campaign review calls. Extract what matters without anyone spending a day rewatching and summarising.

XI

Sales and support teams

Query demo recordings before calls. Find how-to answers in seconds. Give new starters a training library they can interrogate directly.

Questions

Frequently asked questions

Is OttBot Vision available now?

Yes. OttBot Vision is live today. Upload a video and once it has finished processing you can start asking it questions.

Does it just transcribe audio, or does it actually understand the video?

Both, plus more. OttBot Vision analyses speech, on-screen activity, and on-screen text in parallel — so a demo that shows a feature without narrating it is understood as well as one that explains everything verbally.

What is a Collection and when would I use one?

A Collection brings together a set of already-processed videos so you can ask questions across all of them at once, rather than one at a time. The most common uses are building a training guide from a video library, or extracting consistent themes from a series of customer interviews or demo recordings.

What is a Hub?

A Hub is a configured lens that changes how OttBot Vision responds to the same video depending on who is asking. A Marketing Hub and a Support Hub pointed at the same demo give meaningfully different answers — without reprocessing anything.

Can I use OttBot Vision without the rest of the OttBot ecosystem?

Yes. OttBot Vision works as a standalone video knowledge base. Pairing it with OttBot Data adds CRM contact lookup within Vision; pairing it with OttBot Build and OttBot Connect allows extracted knowledge to inform chatbot logic and live customer replies.

Your videos already contain the answers. Now you can actually find them.

Upload your demos, walkthroughs, training recordings, and calls. Ask them what they know. Turn what you find into content, training guides, and chatbot knowledge.