GUIDE · 사용법

PDF·TXT·MD 텍스트 음성 리더기 사용법

TxtSays에서 PDF·TXT·MD를 열고 TTS 음성으로 원하는 문장부터 듣는 방법과 지원 범위를 안내합니다. 로그인이나 파일 업로드 없이 바로 사용할 수 있으며, 로그인은 기기 간 설정 동기화를 원하는 경우에만 선택합니다.

1. 빠른 시작 / Quick start

  1. 파일 열기: 홈의 선택 영역을 누르거나 PDF·TXT·MD 파일을 끌어다 놓습니다.
  2. 시작 위치 선택: 변환된 문서에서 듣고 싶은 문장을 누릅니다.
  3. 재생 조절: 하단 플레이어에서 재생·일시정지, 앞뒤 문장, 속도·목소리를 조절합니다. 모바일에서는 볼륨·재생·포모도로가 첫 줄에 함께 표시됩니다.
  4. 위치 따라가기: 자동 스크롤을 켜면 현재 읽는 문장이 화면 안에 유지됩니다.

Choose or drop a PDF, TXT, or MD file, click the sentence where playback should begin, then use the bottom player to pause, move by sentence, or change speed. The Pomodoro timer defaults to 25 minutes of focus and a 5-minute break, and both times can be adjusted. Turn on auto-scroll to keep the current sentence in view.

2. 지원 파일과 한도 / Supported files and limits

  • PDF: 최대 50MB, 최대 2,000페이지
  • TXT: 최대 4MB, UTF-8·UTF-16·EUC-KR 인코딩 지원
  • MD: 최대 4MB, Markdown 원문을 일반 텍스트로 읽으며 TXT와 같은 인코딩 지원
  • 추출 텍스트: 최대 5,000,000자, 읽기 구간 최대 50,000개
  • 이미지로만 된 스캔 PDF는 글자가 없으므로 먼저 OCR 처리가 필요합니다.
  • 복잡한 표, 다단 편집, 각주가 많은 PDF는 원래의 시각적 순서와 읽기 순서가 다를 수 있습니다.

PDF files are limited to 50 MB and 2,000 pages; TXT and MD files are limited to 4 MB. MD source is read as plain text without rendering Markdown. Extracted text is capped at 5,000,000 characters and 50,000 playback sections. Image-only scans require OCR first, and complex columns or tables may not follow the visual reading order.

3. 음성과 읽기 위치 / Voices and saved position

문서의 글자를 분석해 한국어 또는 영어 읽기 언어를 자동으로 선택합니다. 기본 음성은 Supertonic 3 FP16의 F1이며, F1~F5와 M1~M5의 10개 스타일을 브라우저에서 직접 합성합니다. 첫 화면에서 동의하면 전체 모델·실행 파일 약 217MB를 백그라운드에서 디스크 캐시에 저장하고 동의 여부를 이 기기에 기억합니다. 다음 방문부터는 첫 화면에서 캐시를 확인하고, 필요한 다운로드와 음성 엔진 초기화를 문서를 고르는 동안 미리 진행합니다. 파일은 기다리지 않고 열 수 있으며 같은 모델 버전은 다음 방문에 재사용합니다. 저장과 문서 준비가 끝나면 첫 문장을 미리 만들고, 긴 문장은 완성된 음성 조각부터 이어 재생하면서 기기 추론 속도에 맞춰 다음 음성을 초 단위로 준비합니다. 플레이어의 ‘TTS / 목소리’ 설정에서 Supertonic 3, Google TTS, Microsoft TTS를 그룹별로 직접 선택할 수 있습니다. 브라우저가 모델을 저장할 수 없으면 현재 방문에서만 사용하는 상태를 따로 표시합니다.

읽을 때마다 현재 문장 번호가 이 브라우저에 저장됩니다. 로그인하지 않으면 현재 브라우저에만 남고, Google 로그인하면 계정별로 다시 해시한 문서 식별값과 읽기 위치, 화면 테마·언어만 동기화되어 같은 계정으로 로그인한 다른 브라우저에서 복원할 수 있습니다. 파일명, 문서 텍스트 및 음성은 동기화하지 않습니다. 서버 동기화 데이터를 없애려면 우측 상단 계정 메뉴에서 계정을 삭제하고, 로컬 저장값까지 없애려면 브라우저에서 txtsays.kr의 사이트 데이터를 삭제하세요.

포모도로 타이머는 기본 25분 집중과 5분 휴식을 자동으로 번갈아 진행하며 집중은 1~120분, 휴식은 1~60분으로 조정할 수 있습니다. 시간을 바꾸면 현재 타이머가 정지 상태로 초기화됩니다. 데스크톱에서는 왼쪽 카드에서, 작은 화면에서는 하단 재생바의 설정을 펼쳐 시간을 바꿀 수 있습니다. 집중과 휴식이 전환될 때는 타이머를 한 번 시작해 소리가 허용된 브라우저에서 짧은 알림음이 한 번 납니다. 알림음은 TTS 볼륨을 따르므로 볼륨이 0이면 들리지 않습니다. 타이머 전환은 TTS 재생과 독립적으로 동작하며 새로고침하면 기본 시간의 집중 단계로 초기화됩니다.

PDF 텍스트 조각, TXT, MD와 붙여넣기 원문에는 공백 보정이나 단어 복원 같은 문장 교정을 적용하지 않습니다. 음성 대본에서는 숫자와 숫자식만 읽는 말로 바꿉니다. 예를 들어 1,490.45는 한국어에서 천사백구십 점 사 오로, 영어에서 one thousand four hundred ninety point four five로 읽고, ×와 =도 각각 곱하기·이콜 또는 times·equals로 바꿉니다. 문장이 다음 PDF 페이지까지 이어지면 마침표·물음표·느낌표 같은 종결부호가 나올 때까지 한 문장으로 연결합니다. 종결부호 뒤에 공백·줄바꿈이 있거나 문서가 끝날 때만 문장을 나눕니다. 일반 줄바꿈·문단·페이지 경계에는 별도 쉼을 넣지 않고, 종결부호와 목차 항목, 독립된 대화문 사이에서만 쉽니다. 짝이 맞는 괄호·따옴표 안의 종결부호는 내부 내용으로 유지하며, 짝이 없거나 순서가 틀리면 종결부호 기준으로 다시 나눕니다.

TxtSays detects Korean or English and defaults to F1 from its ten local Supertonic 3 FP16 styles. With your consent, it saves about 217 MB of model and runtime files in the background and remembers that choice on this device. On later visits, TxtSays checks the cache and starts any required download and speech-engine initialization on the first screen while you choose a document. File opening remains available throughout, and the browser cache reuses that version on later visits. Once storage and the document are ready, TxtSays prepares its first sentence and then buffers upcoming audio according to this device's measured synthesis speed. Under “TTS / voice,” you can choose a Supertonic 3 style, Google TTS, or Microsoft TTS. When signed out, the current sentence stays only in this browser. Optional Google sign-in synchronizes only an account-specific hashed document identifier, reading position, display theme, and interface language. File names, document text, and audio are not synchronized. Delete the account from the upper-right account menu to remove server-side synchronization data. If model storage is unavailable, the interface identifies the local voice as temporary for that visit.

The Pomodoro timer defaults to 25 minutes of focus and a 5-minute break. Focus can be set from 1 to 120 minutes and breaks from 1 to 60 minutes. A short chime plays once when phases change after the timer has been started. The chime follows the TTS volume and is silent when the volume is set to 0. Phase changes run independently from TTS playback, and the timer returns to the default focus phase when the page reloads.

PDF text items, TXT, MD, and pasted text are not corrected or rewritten. Only numbers and numeric expressions in the hidden speech script are converted into spoken words, including grouped decimals such as 1,490.45 and symbols such as × and =. A sentence ends only when terminal punctuation is followed by whitespace or the end of the document.

4. 한·영 단어 뜻 찾기 / Korean–English word lookup

이 기능은 기본적으로 꺼져 있습니다. 문서 도구나 모바일 플레이어의 ‘한·영’ 버튼으로 켠 뒤 단어를 선택하거나 오른쪽 직접 찾기에 입력하세요. 앱은 먼저 첫 글자·초성별로 나눈 로컬 사전에서 정확히 일치하는 항목을 찾습니다. 영어 단어에는 일치하는 접두사·어근·접미사의 한글 뜻도 보조 정보로 표시하지만, 이를 정확한 단어 뜻으로 대신 사용하지는 않습니다. 영어→한국어는 한국어 위키낱말사전과 Kaikki/Wiktextract 가공 데이터, 한국어→영어는 국립국어원 한국어기초사전 가공 데이터를 사용합니다. 결과 아래에서 출처와 가공 상세를 확인할 수 있습니다.

정확한 사전 항목이 없을 때만 사용자가 기기 번역을 선택할 수 있습니다. 지원되는 Chrome은 언어팩을 먼저 내려받을 수 있으며, 브라우저 기능이 지원되지 않으면 사전 결과만 사용할 수 있습니다.

Word lookup is off by default. Enable the KO–EN control, then select a word or type it in the lookup field. Exact matches are searched in locally downloaded dictionary shards first. English words can also show conservative Korean guidance for matching prefixes, roots, and suffixes without replacing the exact meaning. English-to-Korean data is derived from Korean Wiktionary through Kaikki/Wiktextract, while Korean-to-English data is derived from the National Institute of Korean Language Basic Korean Dictionary. On-device translation is an explicit fallback only.

5. 파일이 처리되는 위치 / Where your file is processed

파일 선택, PDF 해석, TXT·MD 디코딩, 문장 분할은 브라우저 안에서 수행됩니다. 문서 파일이나 추출한 전체 텍스트를 TxtSays 서버에 업로드하지 않습니다. 음성 모델은 정적 파일로 내려받고 합성은 브라우저 안에서 수행합니다. 다만 사이트를 여는 일반 웹 요청, 정적 모델·사전 파일, 대체 음성·번역을 제공하는 브라우저 기능에는 각각의 네트워크 및 제공자 정책이 적용될 수 있습니다. 선택적으로 Google 로그인하면 계정 인증 정보와 화면 설정·읽기 위치만 Google/Firebase를 통해 처리됩니다. 문서 파일, 파일명, 추출 텍스트 및 음성은 로그인 데이터에 포함되지 않습니다. 세부 내용은 개인정보처리방침에서 확인하세요.

File selection, PDF parsing, TXT/MD decoding, and sentence planning happen in the browser. The document and extracted full text are not uploaded to TxtSays. Speech models are static downloads and synthesis runs in the browser. Normal page requests, model and dictionary assets, and fallback speech or translation features can still be subject to their respective network and provider policies. Optional Google sign-in uses Google and Firebase only for account authentication, display settings, and reading positions; document files, file names, extracted text, and audio are not included in account synchronization.

6. 문제 해결 / Troubleshooting

PDF에서 글자를 찾지 못한다고 나와요.
페이지가 사진으로 저장된 스캔 PDF일 가능성이 큽니다. OCR 기능이 있는 앱에서 문자 인식을 한 뒤 다시 열어주세요.
문장이 다른 순서로 읽혀요.
PDF 내부의 텍스트 순서가 화면 배치와 다를 수 있습니다. 다단 문서라면 TXT로 내보내 순서를 정리한 뒤 여는 방법이 가장 확실합니다.
소리가 나지 않아요.
첫 모델 다운로드와 합성이 끝날 때까지 기다린 뒤, 왼쪽 볼륨과 브라우저 탭·운영체제 음량을 확인하세요. WebGPU 또는 WebAssembly를 지원하는 최신 브라우저를 권장합니다.
이전 파일의 위치가 복원되지 않아요.
파일 내용이 달라졌거나 브라우저의 사이트 데이터가 삭제된 경우 새 문서로 인식합니다. 시크릿 모드의 저장값도 세션 종료 후 사라질 수 있습니다.
Can I start in the middle of a document?
Yes. Click any extracted sentence and playback starts there. The progress slider and previous/next controls provide finer movement.
Does TxtSays perform OCR?
No. It extracts an existing text layer. Run OCR on image-only scans before opening them in TxtSays.

도움이 더 필요한가요? / Need help?

오류 화면의 ‘진단 정보 복사’를 눌러 결과와 사용한 브라우저 이름을 wjdxogh4020@gmail.com으로 보내주세요. 진단 정보에는 파일명과 문서 내용이 포함되지 않습니다.