0
366 次瀏覽

As I am gradually getting more involved in language and dialect adaptations, I keep thinking about what might be the best way for doing this. Earlier this year I posted a question regarding the use of the PT interlinearizer feature vs. Adapt It. But now I’m wondering if there has ever been any discussion about using the interlinearizer vs. doing a Transliteration project (using Encoding Converter and the TEC kit Mapping Editor). Did anyone ever compare the two methods for adaptation work? … and if so, could you point me to a link where I might find a discussion on this topic with pros and cons, etc.

舊文章 - 以原文顯示
Paratext (252 點) 提出
已重新顯示 | 366 次瀏覽

4 個回答

0
最佳回答

Isn’t it a wonderful thing to have choices these days about which tools to use for adaptations! And yet at the same time, we end up staring at the supermarket shelf full of choices wondering which brand of cornflakes will taste the best, or cost the least, or be best for our health, etc. So, similarly, with choosing the right tool do do adaptation we need to ask a whole bunch of questions:

  1. Who is going to be doing the actual work (and what level of skill do they have)? This includes both computer skills [which are easy to build] and translation principles skills [which take a lot more time to build].
  2. Does the user understand both the source and target languages adequately? And will they be required to make on-the-spot decisions based on minimal understanding of the text? If so, are they capable of doing so? Will there be a second round of checking to ensure that the right choices were made?
  3. Are there going to be multiple people working on the adaptation, or will it all be handled by one person?
  4. Do you prefer a “slow, manually walking through the entire text” approach, or a “one off rapid global change” approach (which can be refined and re-run as many times as you want). Or perhaps a hybrid with global changes followed by manual tweaks?
  5. How much freedom do you really want the user to have in terms of adapting the text? How will you know that they have been consistent and thorough when revising the text with find-and-replace operations (some of which may have been more comprehensive than others)? And how will you know when they have gone too far and changed the meaning so much that it will require another round of comprehension testing and consultant-checking?
  6. How closely are the dialects and cultures related? Will the same idioms still work in the target?
  7. How predictable are the changes like to be? Will they primarily be phonological/orthographic (and mostly predictable) or are they mainly lexical (and unpredictable)?
  8. Are the languages closely related grammatically? Do they use the same/similar affixes? Is word order most likely to remain unchanged?
  9. Is the source text complete, final and authoritative, or is it still being worked on? And are you expecting the adaptation process to give feedback to the original text (when room for improvement is discovered)?
  10. Is a script conversion (transliteration) also needed as part of the adaptation?
  11. Is this a one-off job, or will there be other materials which will need to follow the same path later (in which case it may be worth a little extra effort to build a bridge)?
  12. How many target dialects are you planning to adapt into? Could you piggy-back off the work done in one to speed up the next?

And the list of questions could go on and on… It would be an interesting exercise to put these questions in a matrix with the various tools in columns and score the appropriateness of each method.

I think that there are a number of situations where a clever set-up using Transliteration method (as a one-off process) will get you 95% of the way, and then final editing or manual tweaking can polish off the job nicely. And so it is possible that you would do both (i.e. Source-A > automatically modified Target-Bv1 > manually tweaked/Interlinearized Target-Bv2).

Remember that with the Transliteration system you can use simple CC tables, TECkit maps, Python scripts, transliterators, and all kinds of powerful tools to do the initial transformation. You can also have a “main” converter followed by a “fall-back” converter which only acts on data that was unchanged by the main converter. You can daisy-chain a bunch of different converters together, and so on.

If using the Interlinearizer in Paratext, you can set it up so that the source and model text both point to the same text - which will auto-fill all the target words. That way you will only have to check/tweak a few words in each verse before approving and moving on. It becomes quite efficient (especially if all the predictable changes have already been dealt with through a previous automatic process), but still takes time to step through.

I hope these ideas are helpful as you stare at the shelf containing a dozen of varieties of corn flakes!

舊文章 - 以原文顯示
(3.2k 點) 提出
0

Interesting idea. You could in theory use a transliteration map to make consistent changes to one text to produce another. You could have some rules that look at context, but it could get very complicated.

Just to let you know you could look into doing rule-based adaptation using FLExTrans (https://github.com/rmlockwood/FLExTrans/wiki) It relies on FLEx dictionaries and linguistic rules to transform one language to another.

Ron

舊文章 - 以原文顯示
(336 點) 提出
已重新顯示
0

This is helpful - especially the idea of using FLEx. Thank you for your replies.

舊文章 - 以原文顯示
(252 點) 提出
0

If the changes are largely phonological as it sounds they are, then another option is to use a CC table or teckit to pre-process the data in Adapt it. Using cc tables or Flextrans would require that you know what all the rules are before starting. My opinion would be that if you are going to lots of different texts over along period of time would be to invest in making FlexTran work.

舊文章 - 以原文顯示
[Expert]
(2.9k 點) 提出

相關問題

0
4 個回答 393 次瀏覽
We need to do two dialect adaptations for our project. There will be a minor change in vocabulary but fair change in ... we might find a discussion with pros and cons about this?
anon233143 252 提出 已提問 5月 15, 2018
0
3 個回答 844 次瀏覽
我正在嘗試微調一個 .tec 編碼轉換器,以確保它對於從右到左(RTL)語言專案的從左到右(LTR)轉寫而言是安全的往返操作 (這使得能夠與原文進行比較,以確保拉丁字母側正字法上的調整沒有對文本引入任何更改 ) 這方面的主要障礙是 ... 用正則表達式來完成,但為了方便那些不熟悉 RegexPal 的人,我希望將其自動化 關於如何做到這一點,有什麼建議嗎?
DVM 175 提出 已提問 3月 17, 2021
0
1 個回答 176 次瀏覽
I have the main vernacular project in a special script and also a Romanized version using transliteration. When trying to do ... you want to do spell checking on the primary text?
anon650332 163 提出 已提問 12月 31, 2018
0
3 個回答 472 次瀏覽
I am trying to navigate my first transliteration project. Background information The team is starting typesetting, and has ... are supposed to work and what behavior is expected.
[Moderator]
dhigby
1.3k 提出
已提問 8月 3, 2018
0
1 個回答 210 次瀏覽
We have a user who is doing a transliteration from one script in a language to another script in the same language. ... has the reference number of PTXS-17276 Thank you, MSEAIT/LT
MSEAIT_LT 478 提出 已提問 8月 3, 2018
Welcome to Support Bible, where you can ask questions and receive answers from other members of the community.
Don’t you know that you yourselves are God’s temple and that God’s Spirit dwells in your midst?
1 Corinthians 3:16
3,049 個問題
6,007 個回答
5,672 則評論
2,029 位使用者