Comparison

Meeting notes without a bot: how to record a meeting in the same room

Search for an "AI meeting assistant" and almost every tool that comes back does the same thing: a bot joins your Zoom, Meet or Teams call. So what happens when the meeting is around a table and there is no link to join?

The category quietly makes an assumption

Fireflies, tl;dv, Fathom, Read, Otter — every one of them starts its flow with a calendar invitation. The bot joins the link, records, produces a transcript. That is good design: the bot receives each participant's audio on a separate channel, so it never has to guess who is speaking. The channel already says who.

But a large share of meetings in Turkey still happen around a table: a client visit, a supplier negotiation, an interview, a class, a consulting session, a meeting with a lawyer. These meetings have no link. None of the category-leading tools is any use in that room.

Why recording one room is technically harder

The difference in one sentence: the bot has one audio channel per person, the phone on the table has a single mixed channel.

Three difficulties follow:

If a tool says "I do speaker separation", the question is under what conditions. A tool that reads the channel in an online meeting will not give you the same result on your recording at the table.

The real options for a face-to-face meeting

1. Record audio, paste it into an AI afterwards

Record with the phone's Voice Memos, hand the file to a transcription service, then hand the text to ChatGPT. It is free and it works — up to a point.

Where it jams: there is no speaker separation, so "who promised what" is lost. A two-hour recording runs into upload size limits. And most importantly, you have to remember to do those three steps after the meeting. That is why most recordings sit on the phone, never transcribed.

2. Tools with an on-device recording mode

Some tools, like Notta and Transkriptor, have a path other than the bot: recording directly from the browser or the mobile app. These can be used for a face-to-face meeting, and both support Turkish.

What to expect: after the recording ends the file goes to a server, is transcribed there, and the speaker separation is done there too, specific to that file. There is no screen being transcribed live and no speaker identity carried from one recording to the next. Plans are sold by the minute.

3. An app written for the meeting at the table

Ses Notu is that third path. You put the phone in the middle of the table and press record:

Speaker separation runs inside the phone: the CAM++ voice print model (WeSpeaker, Apache-2.0) on ONNX Runtime, a ~28 MB file shipped inside the app. A voice print is biometric data and never leaves the phone. How it was done.

ApproachWorks in one roomLiveSpeaker separationOutput
Bot-based tools
Fireflies, tl;dv, Fathom, Read
No — needs a linkYesFrom the channel, exactMeeting note
Voice Memos + ChatGPTYesNoNoneRaw text
On-device recording mode
Notta, Transkriptor
YesNoOn the server, per fileTranscript plus summary
Ses NotuYes — that is the designYesOn the device, persistent voice printSummary, decisions, action items

Before you record: tell the room

Bot-based tools have a side benefit — a participant appears on screen saying "recording", and everyone sees it. When you put a phone on the table there is no such visual warning.

In Turkey, under KVKK (the Turkish data protection law), an audio recording counts as personal data and a voice print as special-category personal data. Telling the room before you start the recording is required by the legal side and by the relationship alike. Ses Notu reminds you of this inside the app too, but the responsibility belongs to whoever is recording.

Ses Notu — takes the notes for the meeting at the table. No bot, no Zoom. iPhone and Apple Watch. ₺199.99/mo, first week free. Download on the App Store