मुझे टेक्स्ट संग्रह (Text Collection) में किसी संसाधन के टेक्स्ट प्रस्तुति के साथ एक समस्या का सामना करना पड़ रहा है। इसे स्वयं देखने के लिए, संसाधन ShuLatn, भाषा अरबी (shu-Latn) डाउनलोड करें और इसे टेक्स्ट संग्रह विंडो में जोड़ें। टेक्स्ट की सूची में नीले रंग के ShuLatn लिंक पर क्लिक करें, और दाईं ओर प्रकट होने वाला विस्तृत टेक्स्ट अब हाइफ़न के बिना है:

image777×298 84.5 KB
इनमें से सभी सामान्य U+002D हाइफ़न वर्ण हैं। आप देख सकते हैं कि हाइफ़न पूरी तरह हटा दिए गए नहीं हैं, क्योंकि उदाहरण के लिए al-makruubiin अभी भी एक पंक्ति में टूटा हुआ है। लेकिन हाइफ़न गायब है।
मैं भाषा गुणों (language properties) में देखता हूँ कि हाइफ़न को शब्द-विभाजक वर्ण (word-breaking character) के रूप में सूचीबद्ध किया गया है (शायद शब्द के भीतर विराम चिह्न के बजाय), लेकिन यदि यही एकमात्र समस्या होती, तो कम से कम जब कोई शब्द पंक्ति में टूटता है, तब वह हाइफ़न दिखाई देना चाहिए।
मेरे पास इस परियोजना (जो DBL में है) तक सीधा पहुंच नहीं है, लेकिन यदि समाधान के लिए परियोजना में कोई परिवर्तन आवश्यक है, तो मुझे विश्वास है कि हम उसे बदलवा सकते हैं।
I’m having a problem with text presentation of a resource in a Text Collection. To see this yourself, download the resource ShuLatn, language Arabic (shu-Latn), and add it to a Text Collection window. Click on the blue ShuLatn link in the list of texts, and the expanded text that appears to the right no longer has hyphens:

image777×298 84.5 KB
All of these are normal U+002D hyphen characters. You can see that it hasn’t just completely removed the hyphens, since for example al-makruubiin is still broken across a line. But the hyphen is missing.
I see in the language properties that hyphen is listed as a word-breaking character (instead of maybe punctuation inside a word), but if that were the only problem, then at least when a word is broken across a line that hyphen should appear.
I don’t have direct access to this project (which is in the DBL), but if the solution requires a change to the project, I believe we could get it changed.
English से मशीन-अनुवादित