हमारे पास हाथ से बनाई गई हाइफ़नेशन (hyphenation) डेटा की एक बड़ी सूची है, जिसे एक निचले स्तर के PT-प्रोजेक्ट से PTXPrint में खींचा गया है, मार्ग के माध्यम से: > Fonts+Scripts > Hyphenation > Import List...
मुझे यहाँ एक फ़ाइल मिली है, जो उस इम्पोर्ट का परिणाम प्रतीत होती है:
"D:\[path]\PT9\[project]\shared\ptxprint\hyphen-[project].tex"
यह .tex फ़ाइल UTF8 में एन्कोडेड है और मुझे केवल U+002D के रूप में "सामान्य" हाइफ़न दिख रहे हैं। अब मैं एक regex का उपयोग करना चाहता हूँ ताकि वे हाइफ़नेशन विकल्प कम कर सकूँ जो "संभव हैं लेकिन कुरीतियाँ" हैं:
मैं PTXPrint बंद कर दूँगा, एक बैकअप-कॉपी बनाऊँगा, कुछ हाइफ़न हटाकर अपने परिवर्तन लागू करूँगा, फिर PTXPrint को फिर से चलाऊँगा, और "Print" अर्थात् "Make PDF" का प्रयास करूँगा। क्या यह उचित लगता है, या इस तरह से मैं PTXPrint-प्रोजेक्ट को नष्ट कर दूँगा? यदि दूसरा मामला है, तो हाइफ़नेशन को समायोजित करने का कोई बेहतर तरीका है?
नोट: मुझे regex में मदद की आवश्यकता नहीं है, धन्यवाद। यहाँ प्रश्न केवल PTXPrint के आंतरिक कार्यप्रणाली के बारे में है।
TL/DR:
हम PT-प्रोजेक्ट को कई संभावित प्रकाशनों के लिए "स्रोत" मानते हैं, जैसा कि PTXPrint ऑनलाइन सैंपल गैलरी में अच्छी तरह से प्रदर्शित किया गया है। इसलिए Paratext के अंदर हम हाइफ़नेशन मास्टर-सूची का सावधानीपूर्वक और हाथ से रखरखाव करते हैं, जहाँ हम सभी कानूनी विकल्पों को चिह्नित करते हैं, भले ही वे बहुत सुंदर न हों।
वर्तमान प्रकाशन के लिए हमारे पास एक निश्चित वयस्क पाठकों का लक्षित दर्शक वर्ग है और हमारे पास एक लेआउट, पेज-साइज़, फ़ॉन्ट-साइज़ है, जो हमें "थोड़ी अधिक स्वतंत्रता" और लचीलापन प्रदान करता है। इसलिए हम अनुमत लेकिन आदर्श न होने वाले कुछ हाइफ़नेशन-विकल्पों को छानकर हटा सकते हैं। उदाहरण के लिए, मैं सभी विकल्पों को हटाना चाहता हूँ, जो छपी हुई पंक्ति के अंत में केवल एक अक्षर को "अनाथ" (orphan) छोड़ देंगे, और इसी तरह अंत में अकेले खड़े अक्षरों के लिए भी।
We are having a substantial list of hand-crafted hyphenation data, pulled into PTXPrint from an underlying PT-project via > Fonts+Scripts > Hyphenation > Import List...
I have found a file here, that seems to be a result of that import:
"D:\[path]\PT9\[project]\shared\ptxprint\hyphen-[project].tex"
This .tex file is encoded in UTF8 and I see only "normal" hyphens in form of U+002D. Now I would like to use a regex to reduce hyphenation-options which are "possible but ugly":
I would shut down PTXPrint, make a backup-copy, apply my changes by removing certain hyphens, run PTXPrint again, and attempt to "Print" aka "Make PDF". Does this sound reasonable, or will I destroy the PTXPrint-project that way? If the latter, is there a better way to tune the hyphenation?
Note: I do not need help with the regex, thank you. The question here is only about the inner workings of PTXPrint.
TL/DR:
We consider the PT-project as the "source" for many possible publications, as demonstrated nicely on the PTXPrint online sample gallery. So inside Paratext we carefully manually maintain the hyphenation master-list, where we mark all legal options, even the not-very-pretty ones.
For the present publication we have a certain adult-readers target audience and we have a layout, page-size, font-size, which gives us "a little more freedom" and flexibility. So we can filter-out some of the allowed but not-ideal hyphenation-options. For example, I want to remove all options, which would leave only one character like a "orphan" on the end of a printed line and similar for trailing lonely characters.
English से मशीन-अनुवादित