Kruti Dev to Unicode: convert Hindi typed in Kruti Dev

Kruti Dev 010 is the typewriter-layout font Hindi offices, courts and presses typed in for years, and Marathi and Sanskrit were typed in it too. A file set in it stores the keys the typist pressed, which is why its text pastes as Hkkjr and not भारत. The converter reads those keys through a Kruti Dev table and writes Unicode Devanagari.

Takes PDF, Word .docx or .doc, PageMaker .pmd files up to 25 MB · Files deleted within 4 hours · No sign-up

In the file

vlekurk ds fojqn~/k mBh ,d

Read as Unicode

असमानता के विरुद्ध उठी एक

A line of a Hindi PageMaker publication typed in Kruti Dev 010: the bytes the typist’s keys stored, and the same bytes read through the Kruti Dev table.
What do you have?
Back to Tools
🔤

Legacy PDF to DOCX

A PDF whose text was set in a legacy Indic font, recovered as editable Unicode Word — read from the font’s own table rather than by OCR.

How to use?

Leave it on detect if you are not sure. The ones marked “not yet” will tell you so rather than converting badly.

Files are stored encrypted
Deleted within 4 hours
SSL Encrypted

How Kruti Dev stores Hindi

Kruti Dev follows the Remington typewriter layout. Most consonants are typed as a half form and completed by the vertical stem on k, so भ is Hk. The short i is typed before the consonant it follows, which is how कविता becomes dfork, and the reph is typed after the letter it sits over, so कर्म is deZ. The converter reads the keys and puts each sign back where Unicode wants it.

The layout leaves no room for punctuation. The comma key draws ए and the full-stop key draws ण्, so a typist who wants a comma switches to another font for it, as the typist of the publication above did, and the danda is on A. Converted text follows the font, not the key: a comma typed inside a Kruti Dev run is ए.

A letter one position out

Drawn in the real font, 32 of Kruti Dev’s 33 consonants matched their Unicode shapes and झ did not. The key > already draws a whole झ, with its stem, and the table had read it as half a letter waiting for one. So a typist’s समझ came out as समझ् and समझा lost its ा. Written through the table and drawn in the font, six test words holding झ went from one right to six. It hid for a long time because the words it broke look like typing mistakes rather than a mapping fault.

A report that श and ष were swapped was checked the same way and was false: the font draws श and ष exactly where the table reads them. OCR confuses those two letters, which is why a check made by reading output back suggested a swap that is not there.

Measured

Fourteen real PDFs typed in Kruti Dev, from government tenders to a 103-page gazette notification, hold 19,885 words once converted, and 88.0% of them are in Tesseract’s Hindi or Devanagari word lists; fourteen DevLys PDFs hold 17,025 words at 95.2%. The words outside the lists include names, places and abbreviations that are right, so this is not an accuracy figure. One printed line of every document was checked against the conversion by eye, and all matched.

A Hindi PageMaker publication reads 122 of its 126 words exactly as its Unicode source. Eleven of its letters had been turned into curly quotes by Word or PageMaker, because श् and ष् sit on the quote keys; the file still says which key was pressed, so they are read back as the letters they were. The four words left are damage no key explains: three nuktas stored as ऽ, and quotation marks saved as question marks.

Where it still goes wrong

Kruti Dev 010 has a table, and Kruti Dev 714 is read as a variant of it: drawn side by side the two fonts put the same letters on the same keys except the digits, which 714 draws in Devanagari and 010 in Western form, and the codes for श् and ष् that Word’s curly quotes produce, which the two draw the other way round. A file is read as the face it names draws, and where a document was typed in a third build of 010 that swaps those two letters, the reading holding more real words is chosen. DevLys 010 is that swapped build under another name, and is converted; Chanakya has a table of its own, for PDFs. Dozens of codes in the font draw glyphs nobody has identified yet, mostly rare conjuncts: they come through as characters that are plainly not Devanagari, not as plausible wrong letters, so they are easy to spot. And a letter lost before the file reached you, such as a nukta saved as ऽ, cannot be recovered; the curly quotes above can be, and are.

Frequently asked questions

Does it work for Marathi and Sanskrit typed in Kruti Dev?

Yes. The font does not say which language it carries, and the table reads Hindi, Marathi and Sanskrit the same way. Marathi’s ळ is on the G key.

What about Kruti Dev 011, 016 or 714?

Files naming Kruti Dev 011 or 016 are read with the 010 table, on the assumption that they share its layout; only 010 has been measured. Kruti Dev 714 was drawn beside 010 code by code: the letters are the same, the digits are Devanagari where 010 draws Western ones, and the two quote codes that hold श् and ष् are the other way round. It is read as 010 with those differences.

Are DevLys and Chanakya supported?

Both are. DevLys 010 draws 501 of its 503 glyphs as Kruti Dev 010 does, and fourteen DevLys PDFs of recruitment notices were checked; it has its own page. Chanakya is converted for PDFs and has its own page: the newspaper layout, read from the real font and measured on twenty-six PDFs from at least seven publishers, including the variant that puts digits where others put half letters.

Why does my comma come out as ए?

Because in Kruti Dev it is ए. The comma key draws ए in that font, so a comma typed inside a Kruti Dev run was always ए on the page. Typists put commas in another font for this reason, and those convert as commas.

Can I convert Unicode Hindi into Kruti Dev?

Yes. DOCX to PageMaker-ready writes Kruti Dev 010 bytes into an RTF that PageMaker places already set in the font, with English kept in an ordinary face.

My Word file is in Kruti Dev. Can I get a Word file back?

Legacy DOCX to PDF reads it and writes a Unicode PDF. For an editable Unicode document, convert that PDF with PDF to Word. Word files have not been measured the way the PDFs above were, so check the result.

What happens to my file?

It is stored encrypted while it converts and removed by the cleanup that clears every conversion, so nothing is kept longer than 4 hours after upload. Files up to 25 MB, no account needed.

Related tools

Other legacy font pages