convert.express

Transcribe Korean Audio or Video to Text

Upload any Korean recording — OpenAI Whisper detects the language automatically and returns an accurate transcript in seconds.

Drag a file here or browse

MP3, WAV, M4A, FLAC · MP4, MOV, MKV, WebM · up to 2 GB

Set language, translation, or vocabulary hints below before transcribing

OR

YouTube, Dropbox, Google Drive, or any direct MP3/MP4 link

Not applied on Gemini 3.5 Transcribe — the model can't combine vocabulary hints with word-level timestamps

MP3, WAV, M4A, FLAC · MP4, MOV, MKV · up to 2 GB · 10 minutes free every day

한국어 · convert.express →

About Korean transcription

Korean sits in its own language family, Koreanic, with no close relatives among the world's major tongues, and that isolation shapes everything about how transcription works. The writing system, Hangul, stacks consonants and vowels into syllable blocks rather than stringing letters in a line, which means the model has to commit to syllable boundaries as it reconstructs text. That is non-trivial work, but Hangul is a highly systematic script with consistent phoneme-to-grapheme mappings, so the reconstructed output lands in standard Hangul with high accuracy. One of the more interesting challenges is word spacing: spoken Korean carries no reliable acoustic cue for where one word ends and the next begins, so spacing has to be inferred from grammatical structure rather than heard directly. For a language where postpositions and verb endings attach fluidly to stems, that inference has real consequences for readability.

OpenAI Whisper, MAI-Transcribe 1.5 and Gemini 3.5 Transcribe all support Korean, detected automatically from the audio — no manual language selection is needed. Upload any Korean recording up to 2 GB — MP3, WAV, M4A, FLAC, MP4, MOV, and most common audio and video formats are accepted. Every account includes 10 free minutes of transcription per day; longer files are billed at €0.02 per minute with a €0.50 minimum per transaction. Transcripts download as plain text and are automatically deleted within 24 hours of upload.

High accuracy

Available transcription models

OpenAI WhisperMAI-Transcribe 1.5PreviewGemini 3.5 TranscribePreview

Compare models →

10 min

Free every day

€0.02

Per minute beyond quota

< 1 min

Typical turnaround

Tips for the best results

  • Use a quiet environment — background noise is the biggest source of transcription errors.
  • Speak at a natural pace; extremely fast speech reduces accuracy in any language.
  • MP3 or M4A files work great; uncompressed WAV is ideal for professional recordings.
  • Files up to 2 GB are supported — long recordings are automatically split and merged.

Frequently asked questions

How accurate is AI transcription for Korean?
Korean transcription on convert.express is rated high accuracy. A lot of that comes down to Hangul itself: the script encodes pronunciation systematically through syllable blocks, which keeps the gap between spoken sound and written output relatively tight. The main difficulty is word spacing, since spoken Korean gives no acoustic signal for word boundaries, and the model has to reconstruct them from grammatical context. For clean recordings of standard South Korean speech, the output is reliable and legible with minimal cleanup needed.
How long does Korean transcription take?
Most Korean recordings are ready in under a minute. Processing time scales with file length — a 10-minute recording typically returns results in 20–40 seconds, and a one-hour recording in around 4–6 minutes. Longer files are automatically split, transcribed in parallel, and merged, so you never need to cut your audio before uploading.
Can I translate Korean audio to English or other languages?
Yes, two ways. For a quick English-only result, enable "Translate to English" in the upload options above — OpenAI Whisper transcribes and translates Korean audio to English text in a single pass, at no extra cost. For any other target language, first transcribe normally, then use the Translate action on the finished transcript in your dashboard — it's powered by Claude AI and can translate the Korean transcript (or its summary and action list) into any of convert.express's other supported languages, not just English.
Which transcription model should I use for Korean?
OpenAI Whisper, MAI-Transcribe 1.5 and Gemini 3.5 Transcribe all support Korean. OpenAI Whisper is the established choice, with broad language coverage and single-pass translation to English. MAI-Transcribe 1.5 is in preview, typically returns results faster, and supports entity biasing for proper names and terminology. Gemini 3.5 Transcribe is in preview, has the widest language coverage, and labels speakers in-model. Try them on a short clip to see which output suits your recording.
What audio and video formats are supported?
MP3, WAV, M4A, FLAC, OGG, Opus, MP4, MOV, MKV, and most other common audio and video formats are accepted. Files up to 2 GB can be uploaded. The language is detected automatically from the audio — no manual language selection is needed.

← See all 64 supported languages

한국어 · convert.express

한국어는 세계 주요 언어들과 친족 관계가 없는 고립된 언어 계통, 즉 한국어족에 속합니다. 이러한 특성은 음성 인식 작동 방식 전반에 영향을 미칩니다. 한글은 자음과 모음을 한 줄로 늘어놓는 대신 음절 단위의 블록으로 묶는 문자 체계로, 모델이 텍스트를 복원할 때 음절 경계를 정확히 판단해야 합니다. 쉽지 않은 작업이지만, 한글은 음소와 문자 간의 대응이 일관된 매우 체계적인 문자이기 때문에 복원된 결과물은 높은 정확도로 표준 한글에 가깝게 출력됩니다. 특히 흥미로운 과제 중 하나는 띄어쓰기입니다. 구어 한국어에서는 단어와 단어 사이를 구분하는 뚜렷한 음향적 단서가 없어, 문법 구조를 바탕으로 띄어쓰기를 추론해야 합니다. 조사와 어미가 어간에 자연스럽게 붙는 언어 특성상, 이 추론은 가독성에 직접적인 영향을 미칩니다. OpenAI Whisper, MAI-Transcribe 1.5, Gemini 3.5 Transcribe 모두 한국어를 지원하며, 언어는 오디오에서 자동으로 감지되므로 별도로 선택할 필요가 없습니다. 최대 2 GB의 한국어 녹음 파일을 업로드할 수 있으며, MP3, WAV, M4A, FLAC, MP4, MOV 등 대부분의 일반적인 오디오 및 비디오 형식이 지원됩니다. 모든 계정에는 하루 10 free minutes의 무료 전사가 포함되며, 그 이상은 분당 €0.02, 거래당 최소 €0.50로 청구됩니다. 전사 결과물은 일반 텍스트로 다운로드할 수 있으며, 업로드 후 24시간 이내에 자동으로 삭제됩니다.

한국어 AI 전사의 정확도는 어느 정도인가요?
convert.express의 한국어 전사는 높은 정확도로 평가됩니다. 그 핵심에는 한글 자체의 특성이 있습니다. 한글은 음절 블록을 통해 발음을 체계적으로 표기하는 문자이기 때문에, 발화 음성과 텍스트 출력 간의 간극이 비교적 작습니다. 주된 어려움은 띄어쓰기입니다. 구어 한국어는 단어 경계를 알려주는 음향적 신호가 없어, 모델이 문법적 맥락을 바탕으로 이를 추론해야 합니다. 표준 한국어로 녹음된 깨끗한 음성의 경우, 출력 결과는 신뢰할 수 있으며 별도의 수정 작업이 거의 필요하지 않습니다.
한국어 전사는 얼마나 걸리나요?
대부분의 한국어 녹음은 1분 이내에 처리됩니다. 처리 시간은 파일 길이에 따라 달라지며, 10분짜리 녹음은 보통 20~40초, 1시간짜리 녹음은 약 4~6분 내에 결과가 반환됩니다. 긴 파일은 자동으로 분할되어 병렬로 전사된 후 합쳐지므로, 업로드 전에 오디오를 직접 나눌 필요가 없습니다.
한국어 오디오를 영어나 다른 언어로 번역할 수 있나요?
네, 두 가지 방법이 있습니다. 영어 결과만 빠르게 얻으려면 위 업로드 옵션에서 Translate to English를 활성화하세요. OpenAI Whisper가 한국어 오디오를 영어 텍스트로 한 번에 전사 및 번역하며, 추가 비용은 없습니다. 다른 언어로 번역하려면 먼저 일반 전사를 진행한 후, 대시보드에서 완성된 전사물에 번역 기능을 사용하세요. 이 기능은 Claude AI를 기반으로 하며, 한국어 전사물(또는 요약 및 액션 목록)을 영어뿐만 아니라 convert.express에서 지원하는 모든 언어로 번역할 수 있습니다.
한국어 전사에는 어떤 모델을 사용해야 하나요?
OpenAI Whisper, MAI-Transcribe 1.5, Gemini 3.5 Transcribe 모두 한국어를 지원합니다. OpenAI Whisper는 폭넓은 언어 지원과 영어 단일 패스 번역을 갖춘 검증된 선택지입니다. MAI-Transcribe 1.5는 현재 프리뷰 단계로, 일반적으로 더 빠른 결과를 제공하며 고유명사와 전문 용어에 대한 엔터티 바이어싱을 지원합니다. Gemini 3.5 Transcribe도 프리뷰 단계로, 가장 넓은 언어 지원 범위를 자랑하며 모델 내에서 화자를 구분해 표시합니다. 짧은 클립으로 각 모델을 테스트해보고 녹음에 가장 잘 맞는 출력을 선택하세요.
어떤 오디오 및 비디오 형식이 지원되나요?
MP3, WAV, M4A, FLAC, OGG, Opus, MP4, MOV, MKV 및 기타 대부분의 일반적인 오디오·비디오 형식이 지원됩니다. 최대 2 GB의 파일을 업로드할 수 있습니다. 언어는 오디오에서 자동으로 감지되므로 별도로 언어를 선택할 필요가 없습니다.