Kanban StudiosKanban Studios
SELECTED WORK - 10

Common Ground - speak once, meet the room.

Common Ground is a real-time communication-adaptation engine for meetings, demos, onboarding, and presentations. It transcribes live speech (or reprocesses uploads), keeps the verbatim source visible, rewrites the same explanation for five distinct audience types, and produces a structured recap of decisions, action items, risks, and follow-ups - with a deterministic demo mode that runs with no API keys at all.

Common Ground landing page: Speak once. Meet people where they are - with transcript, audience rewrite, and clear recap tiles
5audience personas per message
4recap fields extracted live
0API keys needed for demo mode
1-clickreframe of the same source
THE PROBLEM

One explanation almost never lands the same way for an executive, a client, an engineer, and a new hire in the same room. Presenters re-explain on the fly, lose the thread, and leave the meeting without a clean record of what was actually decided.

Common Ground treats a single message as adaptable material: it keeps the original words intact, rewrites the framing per audience in real time, and distills the conversation into a decision-grade recap - so the same words meet every person where they are, and nothing gets lost after the call.

WHAT IT DOES

One message, adapted for every room.

01

Captures speech as it happens

Common Ground taps the browser's Web Speech API for live transcription, with a graceful non-blocking fallback when recognition is unavailable - and an upload path that reprocesses recorded audio through the same pipeline. The original transcript is never discarded; it stays visible next to every rewrite.

02

Rewrites one message for five rooms

The same source is re-framed on demand for Executive, Client, Engineer, New-hire, and Non-native-speaker audiences. Each mode is a distinct rhetorical target - altitude, jargon density, and sentence complexity all shift - while the underlying facts stay locked to the source so nothing is invented in translation.

03

Extracts a decision-grade recap

A structured extraction layer parses free-form conversation into Summary, Decision, Action, and Risk fields - and reports honestly when a field has no signal ("No decision detected yet") instead of hallucinating one. The recap is exportable the moment the meeting ends.

04

Runs fully offline for demos

Demo mode is deterministic and provider-agnostic: three preloaded scenarios and every audience mode run with zero external AI calls, so a live demo can never fail on a flaky network or a missing key. Live mode layers real speech recognition on top when the environment supports it.

Live demo workspace: input mode (type, speak, sample, upload), scenario presets, the verbatim source panel, and the audience selector that reframes the same message in one click.
Live demo workspace: input mode (type, speak, sample, upload), scenario presets, the verbatim source panel, and the audience selector that reframes the same message in one click.
THE ADAPTATION ENGINE

Source in. Audience-ready language out.

The hard part is not paraphrasing - it is re-framing the same facts across five audiences without drifting from what was actually said. Common Ground pins every rewrite to the verbatim source, so an engineer's version and an executive's version disagree on altitude and jargon but never on substance. Switching audience is a single click, and the source stays visible the whole time.

FROM MESSAGE TO RECAP

Capture, adapt, recap.

The whole flow is built to survive a live room: speech is captured as it happens, the source is preserved, the framing adapts on demand, and a structured recap is assembled in real time - deterministically enough to demo with the network unplugged.

  1. 01Capture speech live, or ingest an uploaded recording
  2. 02Preserve the verbatim source transcript, always visible
  3. 03Adapt the message per audience persona in one click
  4. 04Extract summary, decision, action, and risk in real time
  5. 05Export a shareable recap and persist the session locally
Recap view: summary, decision, action, and risk extracted from the current source and audience - with honest 'no signal yet' states - above the three-step product flow.
Recap view: summary, decision, action, and risk extracted from the current source and audience - with honest 'no signal yet' states - above the three-step product flow.
UNDER THE HOOD

Three layers, one transcript.

Capture layer

Browser speech recognition with capability detection and a graceful fallback, plus an upload-and-reprocess path - all feeding one normalized transcript model.

Adaptation layer

A per-audience rewrite pipeline that re-frames the same source without mutating its facts, driven by a provider abstraction that swaps between deterministic demo output and live AI.

Recap layer

Structured extraction into decision-grade fields with explicit no-signal handling, exportable output, and local session persistence for the history and detail views.

BUILT WITH
NEXT.JS 16 APP ROUTERREACT 19TYPESCRIPTTAILWIND CSS V4RADIX PRIMITIVESFRAMER MOTIONWEB SPEECH APISTRUCTURED EXTRACTIONPROVIDER ABSTRACTIONVITESTPLAYWRIGHTLOCAL PERSISTENCE
Chat on WhatsAppWhatsApp