メインコンテンツへスキップ

Audio Transcription & Meeting Notes

AI-Public can turn audio recordings into text and create meeting notes from the transcript. Transcription uses the provider from the central model catalog, such as OpenAI or European AI. When you start, choose whether the recording is for personal use, a meeting, or a lesson/presentation.

Start screen

On the transcription screen you can start a new recording or upload an existing audio file.

Providing audio

There are two ways to provide audio for transcription.

Record directly in AI-Public

録音中に一時停止を選ぶと、音声の収録を一時的に止められます。一時停止中の時間と音声は含まれません。同じ録音と文字起こしを続けるには再開を選び、会議が終了したときだけ停止を使用してください。

Click Start recording to begin. Before recording starts, a dialog opens with the recording settings.

Recording settings

When starting a recording you can set:

  1. Recording type:
    • Private recording: one person close to the microphone.
    • Meeting: multiple speakers in one room.
    • Lesson or presentation: one main speaker with possible interaction.
  2. Specialist vocabulary and keywords: add names, abbreviations, product names, or terms that are often recognized incorrectly.
  3. Language: AI-Public uses your account language to guide transcription.

Behavior per recording type

The selected type determines how the recording is processed:

  • Private recording with OpenAI uses realtime transcription. Text appears while you speak. Because this is meant for one person, speaker diarization is not applied and no interim audio files are stored.
  • Meeting and Lesson or presentation use file-based processing. AI-Public processes audio parts during the recording and also processes the complete final recording when you stop. This path is suitable for longer recordings, multiple speakers, and recovery after interruptions.
  • European AI/Mistral uses file-based processing. For private recordings, diarization is disabled so the transcript is not unnecessarily split into speakers.

Use an existing audio file

You can also upload an existing recording. Supported formats include MP3, WAV, M4A, and WebM. After upload, the file is processed with the same transcription approach as a recording of the same type.

Transcription and speakers

The transcript can contain time blocks and speaker labels. For conversations and meetings, the model tries to distinguish speakers. For private recordings this is disabled because the transcript is intended for one speaker. Sometimes labels such as Speaker A and Speaker B are used instead of real names. AI-Public post-processes explicit introductions in the text when possible.

Speaker recognition still depends on audio quality, overlapping speech, and the selected model. If names or terms are not recognized correctly, you can improve the transcript with AI.

Improve with AI

After processing, use Improve with AI for targeted corrections, such as renaming speakers, fixing a technical term, or applying a spelling correction consistently. Always check the result when the transcript is used for reporting or decisions.

Meeting notes

After the recording and transcription, open Meeting notes and choose Create meeting notes. The notes are based on the transcript and the active prompt.

Advanced settings

Manage prompts

AI-Public provides default prompts for general meeting notes and notes with speaker recognition. You can also add your own prompts, for example for a fixed meeting structure, action list, or public report. Custom prompts are stored in your account for future transcriptions.

Manage history

Via History you can search, load, rename, edit, or delete earlier transcriptions and play linked audio.

Use the transcription

You can copy the transcript, export it as PDF, use it in chat, or export meeting notes to PDF or Word.

Audio parts and storage

For private realtime transcriptions with OpenAI, AI-Public does not store interim audio files. For meetings, lessons/presentations, and European AI, AI-Public automatically processes audio parts for progress, reliability, and recovery. When you stop, the complete final recording is processed and has priority as the definitive basis. If one interim part fails, the recording can continue; check afterwards whether the final recording was processed correctly.

録音の再生と復元

完全な録音がある場合は、音声プレーヤーを1つ表示します。ない場合はリストから録音部分を選択します。モバイルアプリでは、端末内の録音を再アップロードしたり、完全に削除したりできます。

アップロード待ちです。成功後にローカルコピーが削除されます。アプリを削除すると端末内の録音も消えます。

標準処理:音声ファイルごとに最大 1 GiB。 録音と処理に十分な空き容量を確保してください。

Androidでは画面を開き直しても録音は停止しません。アプリを強制終了した後は保存済みの音声を復元し、同じ録音を再開できます。通話でマイクが使えない場合、録音は一時停止します。マイクが利用可能になったら再開してください。ヘッドセットが切断されると内蔵マイクへの切り替えを試みます。中断履歴は残り、内部の音声部分は一つの録音として表示されます。

Android アプリでは、録音ボタンからネイティブの録音画面が直接開きます。停止後、同じウィンドウで話者認識の有無を選び、必要に応じて専門用語を追加できます。キャンセルすると録音は端末内に保存されます。通知には録音時間と一時停止・再開・停止の操作が表示され、通知設定で許可されていればロック画面でも使えます。通知から停止した後は、アプリに戻って録音を処理してください。モバイルブラウザーでの録音も引き続き利用できます。

WhatsApp