重新發佈自 @KimB
謝謝 anon044949。我們從 PT9 之前就開始使用基於詞幹(stems)的匹配(Match)。我的問題不在於如何在 PT 中進行逐字對照(interlinearizing);而是 PT 現在似乎正在與 FLEx 雙向共享其現有的形態學(morphology)和釋義(gloss)數據。這絕對不是我們想要的,因為我們一直刻意(且系統性地)在 PT 中簡化形態素(morpheme)切分,以避免讓 PT 變得笨重、加快處理速度,並提供翻譯顧問更容易理解的釋義。許多顧問不希望閱讀充斥著晦澀、類似語言學論文風格的形態素釋義文本。我們現在開始將舊的 Toolbox 文本輸入 FLEx 並進行語言學分析。
因此,我現在發現 FLEx 中出現了一堆 PT 風格的詞語形態學和釋義數據(僅限於 PT9 中新的 PT–FW 常規功能啟用後進行的近期分析)。我原本以為 FW–>PT 的共享是單向的。我猜將 PT 的分析結果發送到 FW 是合理的,但這一功能在 PT9 中新增時並未發出警告,提醒用戶注意可能出現的意外後果。我曾與許多其他像我們一樣進行 PT 逐字對照的人交流過——我們的目標是服務於翻譯顧問而非語言學家,所以我認為也會有其他對這個話題感興趣的人。
給 PT9 用戶的提示:如果您的 Paratext 專案未掛載 FLEx 資料庫,那麼目前這與您無關。
因此,如果有人能描述一下在 PT9 中,當關聯了 FLEx 資料庫時,逐字對照過程中具體發生了什麼,那將非常有帮助。例如,
在 FLEx 中定義的解析規則是否適用於 PT?
是否可以關閉 PT 向 FW 發送解析和釋義資訊的功能?
當 FLEx 中的形態素切分/釋義被編輯後,PT 的逐字對照會即時更新嗎?
當我在 PT 的逐字對照節中編輯某些詞語(但忽略我今天不感興趣的詞語),然後重新核准該節時,會發生什麼情況?在這種情況下,舊的形態素切分/釋義會被發送到 FW 嗎?
提前感謝任何見解。
Reposted from @KimB
Thank you, anon044949. We have been using Match based on stems since before PT9. My problem is not how to do interlinearizing in PT; it is that PT now seems to be sharing its existing morphology and gloss data with FLEx—bidirectionally. Which is definitely not what we want, because we have been deliberately (and systematically) underspecifying morpheme breaks in PT to avoid bogging PT down, make it faster to do, and provide glossing that is more understandable by translation consultants, many of whom do not want to wade through a text with cryptic, linguistic paper-type morpheme glosses. We are now starting to enter our old Toolbox texts into FLEx and analyze them for linguistics work.
So now I seem to have a bunch of PT-style word morphology and gloss data in FLEx (only recent analyses from PT9 since the new PT–FW routines became active.) I had been assuming that the FW–>PT sharing was in one direction. I guess it makes sense to send PT analyses to FW, but it was added to PT9 without a warning to watch for unforeseen consequences. I have interacted with many other people who also “do” PT interlinearization like we do—targeting translation consultants rather than linguists, so I’m thinking there will be others interested in this topic as well.
Note for PT9 users: If you do not have a FLEx database attached to your Paratext project this won’t apply to you at this time.
So it would be helpful if someone could give us a description of what exactly happens during interlinearization in PT9, with an associated FLEx database. For example,
Do parsing rules defined in FLEx apply in PT?
Is it possible to turn off PT sending parsing and glossing info to FW?
Is the PT interlinear updated on the fly when FLEx morpheme breaks/glosses have been edited?
What happens when I edit some words in a PT interlinear verse (but ignore words that I’m not interested in today), and then re-approve the verse? Are old morpheme breaks/glosses sent to FW in this case?
Thanks in advance for any insights.
機器翻譯自 English 顯示原文