Cara mengambil teks dari foto atau tangkapan layar (screenshot) tanpa mengetik ulangHow to extract text from a photo or screenshot without retyping it
Perlu menyalin tulisan dari foto dokumen, papan pengumuman, atau tangkapan layar? Begini cara mengubah gambar menjadi teks yang bisa disalin, memilih antara mode standar dan mode AI, serta tips agar hasilnya akurat.
Need to copy text from a document photo, a notice board or a screenshot? Here is how to turn an image into text you can copy, choose between standard and AI mode, and get accurate results.
Di halaman ini
Mengetik ulang tulisan dari foto adalah pekerjaan yang membosankan dan rawan salah ketik: nomor resi dari tangkapan layar, kutipan dari halaman buku, pengumuman di papan sekolah, atau isi surat yang hanya tersedia dalam bentuk foto.
Teknologi pengenalan teks (OCR) bisa membaca tulisan pada gambar dan mengubahnya menjadi teks biasa yang bisa disalin, diedit, dan dicari.
Dua mode, dua kebutuhan
| Mode | Di mana diproses | Paling cocok untuk |
|---|---|---|
| Standar · di perangkat | Di browser Anda, gambar tidak diunggah | Teks cetak yang jelas: dokumen ketikan, buku, tangkapan layar |
| AI · lebih akurat | Dikirim lewat KASKOW ke layanan AI Google (Gemini) | Foto miring atau agak buram, tulisan tangan yang cukup rapi |
Mode standar mengenali teks Bahasa Indonesia dan Inggris. Mode AI hanya dipakai bila Anda sendiri memilihnya. KASKOW tidak menyimpan gambar maupun teks hasilnya. Untuk dokumen yang sangat rahasia, tetap gunakan mode standar.
Langkah mengambil teks dari gambar
- Buka Gambar ke Teks di KASKOW.
- Klik Pilih file, seret dan lepas gambar, atau tempel tangkapan layar langsung dengan Ctrl+V. Di HP, Anda bisa memotret dokumen dengan Ambil foto.
- Pada Cara membaca, pilih Standar · di perangkat atau AI · lebih akurat.
- Klik Ambil teks. Pada pemakaian pertama mode standar, data pengenal teks diunduh dulu (beberapa MB), lalu progres tiap gambar tampil.
- Baca hasilnya dan perbaiki huruf yang keliru.
- Klik Salin teks, atau Unduh untuk menyimpan file .txt.
Bila Anda memproses banyak gambar sekaligus, hasilnya berupa satu file .txt per gambar yang dikemas dalam satu file .zip.
Tips agar hasilnya akurat
- Cahaya terang dan merata. Bayangan tangan atau HP di atas kertas menurunkan akurasi.
- Kamera lurus di atas dokumen. Foto yang miring membuat baris tulisan sulit dikenali.
- Potong bagian yang tidak perlu. Gunakan Potong Gambar agar hanya area bertulisan yang dibaca.
- Tegakkan foto yang terbalik dengan Putar & Balik Gambar.
- Tangkapan layar lebih baik daripada memotret layar. Hasilnya jauh lebih tajam.
Periksa ulang hasilnya
Pengenalan teks bisa keliru pada karakter yang mirip, misalnya angka 0 dan huruf O, atau angka 1 dan huruf l. Untuk nomor rekening, nomor resi, NIK, atau nominal uang, selalu cocokkan dengan gambarnya.
Hasilnya berupa teks polos. Susunan tabel, kolom, huruf tebal, dan format tulisan lainnya tidak ikut tersimpan, jadi rapikan kembali setelah disalin. Untuk membersihkan baris kosong atau spasi berlebih, Anda bisa memakai Alat Teks.
Bagaimana dengan PDF?
- PDF hasil ketikan biasanya sudah menyimpan teks di dalamnya. Ambil langsung dengan PDF ke Teks. Panduannya ada di cara mengambil teks dari file PDF.
- PDF hasil pindai berisi gambar halaman. Ubah dulu halamannya menjadi gambar dengan PDF ke Gambar, lalu masukkan gambar-gambar itu ke Gambar ke Teks.
Pertanyaan yang sering muncul
Apakah bisa membaca tulisan tangan? Mode standar dirancang untuk teks cetak, sehingga tulisan tangan umumnya tidak terbaca dengan benar. Untuk tulisan tangan yang cukup rapi, pilih AI · lebih akurat, lalu periksa ulang hasilnya.
Kenapa pemakaian pertama terasa lama? Browser mengunduh data pengenal teks berukuran beberapa MB. Data ini cukup diunduh sekali. Bila memakai paket data seluler, Anda mungkin lebih nyaman memulainya saat terhubung ke Wi-Fi.
Bahasa apa yang didukung? Mode standar membaca teks Bahasa Indonesia dan Inggris.
Apakah gambar saya diunggah? Pada mode standar, tidak. Gambar dibaca sepenuhnya di browser. Pada mode AI, gambar dikirim lewat KASKOW ke layanan AI Google (Gemini) untuk dibaca, dan KASKOW tidak menyimpan gambar maupun teksnya.
Retyping text from a photo is tedious and prone to typos: a tracking number from a screenshot, a quote from a book page, a notice on a school board, or a letter that only exists as a photo.
Text recognition (OCR) can read the writing in an image and turn it into plain text you can copy, edit and search.
Two modes, two needs
| Mode | Where it's processed | Best for |
|---|---|---|
| Standard · on device | In your browser; the image is not uploaded | Clear printed text: typed documents, books, screenshots |
| AI · more accurate | Sent via KASKOW to Google's AI service (Gemini) | Tilted or slightly blurry photos, fairly neat handwriting |
Standard mode recognises Indonesian and English text. AI mode is only used when you choose it yourself. KASKOW stores neither the image nor the resulting text. For highly confidential documents, stick with standard mode.
Steps to extract text from an image
- Open Image to Text on KASKOW.
- Click Choose files, drag and drop the images, or paste a screenshot directly with Ctrl+V. On a phone you can photograph a document with Take photo.
- Under Recognition, choose Standard · on device or AI · more accurate.
- Click Extract text. The first time you use standard mode, the recognition data is downloaded first (a few MB), then progress for each image is shown.
- Read the result and correct any misread characters.
- Click Copy text, or Download to save a .txt file.
When you process many images at once, you get one .txt file per image, packed into a single .zip file.
Tips for accurate results
- Bright, even light. The shadow of your hand or phone over the paper lowers accuracy.
- Camera straight above the document. A tilted photo makes lines of text harder to recognise.
- Crop out what you don't need. Use Crop Image so only the text area is read.
- Straighten upside-down photos with Rotate & Flip Image.
- Screenshots beat photos of a screen. The result is much sharper.
Double-check the result
Text recognition can confuse similar characters, such as the digit 0 and the letter O, or the digit 1 and the letter l. For account numbers, tracking numbers, ID numbers or amounts of money, always compare with the image.
The output is plain text. Table layouts, columns, bold text and other formatting are not kept, so tidy it up after copying. To clean up empty lines or extra spaces, you can use Text Tools.
What about PDFs?
- Typed PDFs usually already contain their text. Extract it directly with PDF to Text. The guide is in how to extract text from a PDF.
- Scanned PDFs contain images of pages. Turn the pages into images first with PDF to Image, then put those images into Image to Text.
Frequently asked questions
Can it read handwriting? Standard mode is designed for printed text, so handwriting is generally not read correctly. For fairly neat handwriting, choose AI · more accurate, then double-check the result.
Why is the first use slow? The browser downloads text recognition data of a few MB. It only needs downloading once. If you're on mobile data, you may prefer to start while connected to Wi-Fi.
Which languages are supported? Standard mode reads Indonesian and English text.
Are my images uploaded? In standard mode, no. The image is read entirely in your browser. In AI mode, the image is sent via KASKOW to Google's AI service (Gemini) to be read, and KASKOW stores neither the image nor the text.