CogneraGet in touch
A team of cats at a round table, each working at their own screen — the machines are theirs, and the work stays in the room.
Local-first AI

An AI receptionist that never sends a call to someone else’s cloud.

Cognera answers your phone and your website chat from a machine in your building. It listens, understands, answers from your own documents, and hands off to a person when it should — and the recording, the transcript, and the answer never leave the premises.

Runs on
Your hardware
Call data egress
None
Usage cap
Unmetered
Models
Yours to keep

That is the whole point. Most businesses that would benefit most from an AI receptionist are the ones told they cannot use one.

01
The service

The voice agent.

Speech in, speech out, entirely on your own machine. Every stage of the round trip runs locally — there is no step where audio is uploaded to be processed.

A caller speaks. The audio is transcribed on your machine. The question is matched against your own indexed documents, answered by a local model, and spoken back in a natural voice. Nothing is sent out for processing at any point in that chain.

It knows what you have told it and admits what it has not. Asked something outside its material, it says so and offers to take a message rather than inventing a price, a policy, or an appointment.

Transcription
Local
Answering
Local + your documents
Speech synthesis
Local
Recordings leaving
None
Concurrent calls
Hardware-boundnot license-bound
Fig. 1
Fig. 1 — Live session view with running transcript. Screenshot to be placed.
A

Answers questions

Hours, services, pricing, policies, directions, availability — whatever your documents actually say.

B

Takes messages

Captures the caller’s details and the reason for the call when the answer is not something it should attempt.

C

Hands off

Escalates to a person on the conditions you set, rather than trapping the caller in a loop.

02
How it knows things

It answers from your documents.

The part people mean when they ask whether it “knows our business.” It does, and here is the actual mechanism — no magic, and nothing uploaded.

1

You add documents

The things that already describe your business — service descriptions, policies, price lists, FAQs, manuals. Markdown, plain text, PDF, or Word.

2

It builds an index

Each document is split into passages and converted into a numeric form by an embedding model running on your machine. The result is one index file on your own disk.

3

A question is matched

An incoming question is converted the same way and compared against every passage. The closest ones are retrieved — and only those clearing a confidence threshold.

4

The model answers

Those passages become the source material for the reply. If nothing clears the threshold, the honest answer is that it does not know.

Embedding
Localno document leaves the machine
Index location
Your disk
Accepted formats
.md · .txt · .pdf · .docx
Updating
Re-index after edits

Two settings decide how talkative it is: how many passages to retrieve, and how confident a match must be. Set them too tight and a well-stocked knowledge base will still answer with almost nothing — which is usually the real cause when retrieval appears to “stop working.”

1Documents.md · .txt · .pdf · .docx2Indexembedded locally, one file on disk3Matchclosest passages retrieved4Answerwritten from those passages onlyQUESTIONCONFIDENCE THRESHOLDALL STAGES ON YOUR MACHINE
Fig. 2 — Retrieval path. A question is matched against the index, and only passages clearing the confidence threshold become source material for the answer. Nothing in the chain leaves the machine.
03
Continuity

It remembers the conversation.

The most common complaint about automated answering is not that it sounds robotic. It is that it forgets — making the caller repeat a name, a number, or the reason they rang.

Live now
A

The original request

What the caller actually rang about is held for the whole call, so the answer at minute four still relates to the question at minute one.

B

Corrections as they happen

A caller who says “no, Tuesday” is not arguing with a system that keeps offering Monday. The correction sticks.

C

Details it was given

Name, number, email and the reason for the call are captured and kept, so a message arrives complete rather than as a fragment.

D

What it does not know

Anything outside the approved documents stays outside them. It says so and offers to take a message.

Planned

Currently in development, and listed here because it is fair to know what a system does not do yet. Memory today lasts for the duration of a call.

Recall across calls
Plannedrecognise a returning caller
Completed steps
Planneda durable log per engagement
Pending approval
Plannedheld for a human to confirm
Next permitted action
Plannedbounded by rules you set
04
Rationale

Why run it yourself.

Three reasons, in the order they usually start to matter.

i

Confidentiality

Some calls cannot go to a third-party API — not as a preference, but because of the industry you are in. Local execution is the only answer that survives that question.

ii

Cost shape

Metered AI turns usage into a variable cost that grows exactly when business is good. Owned hardware is a fixed cost that stops growing.

iii

Continuity

Hosted models get deprecated, repriced, and have their behaviour changed underneath you. A model file on your disk does none of those things.

05
Also available

Two more things the same machine can do.

Secondary to the agent, and priced as additions rather than as separate products. Both reuse the hardware and the model runtime already installed.

Website Studio

Generates a complete marketing site from a short description of the business, with a chat concierge grounded in that business’s own information — the same retrieval mechanism as the agent.

Media Studio

Image and video generation and upscaling on your own GPU. No queue, no credit balance, no per-image cost once the hardware is in place.

06
Engagement

What it costs.

Sold as a setup on your hardware, not as a per-seat subscription. Hardware is separate; a capable GPU you already own is used as-is.

Agent

Start here
$399

The voice agent, installed

  • Voice agent on your machine
  • Knowledge base indexed
  • Local inference, unmetered

Agent + Studio

$299

Added to an agent install

  • Website Studio
  • Grounded concierge
  • Media Studio available

À la carte

Custom

Individual capabilities

  • Security setup
  • Service activation
  • Add-ons as needed
07
Next step

Getting started.

The first conversation is about whether local execution is actually right for your situation. Sometimes it is not, and that is a useful thing to establish early.