转载自 @KimB
谢谢,anon044949。早在 PT9 之前,我们就一直在使用基于词干(stems)的匹配。我的问题不在于如何在 PT 中进行逐词对照(interlinearizing);而在于 PT 现在似乎正在与 FLEx 双向共享其现有的形态学(morphology)和释义(gloss)数据。这绝对不是我们想要的,因为我们一直有意(且系统地)在 PT 中简化词素(morpheme)切分,以避免拖慢 PT 的速度,使其操作更快捷,并提供让翻译顾问更容易理解的释义——其中许多顾问并不想阅读那些充满晦涩、类似语言学论文风格的词素释义。我们现在开始将旧的 Toolbox 文本导入 FLEx 并进行语言学分析。
因此,我现在发现 FLEx 中出现了一堆 PT 风格的词形形态学和释义数据(仅包括自新的 PT–FW 例程启用以来,从 PT9 进行的最近分析)。我一直以为 FW–>PT 的数据共享是单向的。将 PT 的分析发送到 FW 似乎合乎逻辑,但这一功能在 PT9 中被添加时并未发出警告,提醒用户注意可能出现的意外后果。我与许多其他像我们一样“做”PT 逐词对照的人交流过——我们的目标受众是翻译顾问而非语言学家,所以我想也会有其他人对这个话题感兴趣。
给 PT9 用户的提示:如果您的 Paratext 项目没有关联 FLEx 数据库,那么目前这与您无关。
因此,如果有人能描述一下在 PT9 中关联 FLEx 数据库时,逐词对照过程中具体发生了什么,那将非常有帮助。例如,
在 FLEx 中定义的解析规则是否适用于 PT?
是否可以关闭 PT 向 FW 发送解析和释义信息的功能?
当 FLEx 中的词素切分/释义被编辑后,PT 的逐词对照是否会实时更新?
当我编辑 PT 逐词对照节中的某些单词(但忽略我今天不感兴趣的单词),然后重新批准该节时,会发生什么?在这种情况下,旧的词素切分/释义会被发送到 FW 吗?
提前感谢任何见解。
Reposted from @KimB
Thank you, anon044949. We have been using Match based on stems since before PT9. My problem is not how to do interlinearizing in PT; it is that PT now seems to be sharing its existing morphology and gloss data with FLEx—bidirectionally. Which is definitely not what we want, because we have been deliberately (and systematically) underspecifying morpheme breaks in PT to avoid bogging PT down, make it faster to do, and provide glossing that is more understandable by translation consultants, many of whom do not want to wade through a text with cryptic, linguistic paper-type morpheme glosses. We are now starting to enter our old Toolbox texts into FLEx and analyze them for linguistics work.
So now I seem to have a bunch of PT-style word morphology and gloss data in FLEx (only recent analyses from PT9 since the new PT–FW routines became active.) I had been assuming that the FW–>PT sharing was in one direction. I guess it makes sense to send PT analyses to FW, but it was added to PT9 without a warning to watch for unforeseen consequences. I have interacted with many other people who also “do” PT interlinearization like we do—targeting translation consultants rather than linguists, so I’m thinking there will be others interested in this topic as well.
Note for PT9 users: If you do not have a FLEx database attached to your Paratext project this won’t apply to you at this time.
So it would be helpful if someone could give us a description of what exactly happens during interlinearization in PT9, with an associated FLEx database. For example,
Do parsing rules defined in FLEx apply in PT?
Is it possible to turn off PT sending parsing and glossing info to FW?
Is the PT interlinear updated on the fly when FLEx morpheme breaks/glosses have been edited?
What happens when I edit some words in a PT interlinear verse (but ignore words that I’m not interested in today), and then re-approve the verse? Are old morpheme breaks/glosses sent to FW in this case?
Thanks in advance for any insights.
机器翻译自 English