लंबे समय से Wordlist को देखने के बाद, मुझे एक बड़ा लाल बॉक्स देखकर हैरानी हुई जिसमें असंगत एन्कोडिंग का उल्लेख किया गया था।
तो सवाल यह है कि क्या नॉर्मलाइज़ेशन चालू करना चाहिए या नहीं। निम्नलिखित सलाह पढ़ने के बाद:
हम "Off (no normalization)" चुनने की सलाह नहीं देते जब तक कि परियोजना भाषा की वर्तनी दीक्रिटिक्स का गैर-मानक तरीके से उपयोग करती है और दीक्रिटिक्स के होने की क्रम महत्वपूर्ण है।
मुझे आश्चर्य है कि दीक्रिटिक्स का गैर-मानक तरीके से उपयोग को क्या माना जाता है? उदाहरण के लिए, यह संभव है कि आप इनसे मिलें: ŋ́ या Ŋ́ या ŋ̀, ɛ̀, Ɛ̀ɛ̀, ɛ́ɛ́ और ɛ̀ɛ̀ ɔ̀, ɔ́। क्या इसे मानक उपयोग या गैर-मानक उपयोग माना जाता है?
इसे कैसे संभालना है, इस पर सलाह या सुझाव स्वागत योग्य हैं।
बार्ट।

Having not looked at the Wordlist for a long time, I was surprised to see a big red box mentioning inconsistent encoding.
So the question is whether or not to turn normalisation on. After reading the following advice :
We do NOT recommend selecting "Off (no normalization)" unless the orthography of the project language uses diacritics in a non-standard way and the order in which the diacritics occur is important.
I wonder what is considered using diacritics in a non-standard way? For example, it is possible to encounter : ŋ́ or Ŋ́ or ŋ̀, ɛ̀, Ɛ̀ɛ̀, ɛ́ɛ́ and ɛ̀ɛ̀ ɔ̀, ɔ́. Is this considered standard use or non-standard use?
Advice or tips on how to handle this are more than welcome.
Bart.

English से मशीन-अनुवादित