0
189 次瀏覽

[reposting from another list - TCOP]

I occasionally concatenate USFM files into a complete Bible text file in order to process them outside of Paratext. This consists of all of the files copied into one text file in sequential order.

One of the cumbersome manual steps in this process has been to separate them back into individual book files for long term storage.

Today, when I am importing a processed file back into Paratext, I haven’t separated the file properly: Each book file exists, but the text file Mark also contains John to Revelation instead of ending with Mark 16. Paratext recognized this as “multiple files contain information about the same book.”

When I try to import only the concatenated file which contains the books Mark-Rev, I get all these books in Paratext, with no apparent loss or added information.

This is foundationally different than the USFM I know, where 1 Bible book == 1 text file. Is this {Bible Book Aggregation|File concatenation} a feature? Does Paratext support importing concatenated files officially?

Taking it one step further: my post processing usually includes file concatenation. Aggregating is to a degree part of our local implementation of USFM. Is book aggregation actually part of the USFM specification? (and if not, could it be?)

I might owe someone a coffee if this is true … 1 bible = 1 text file would save so so much file manipulation time for me.

Thanks,

Michael Hart
Senior Publishing Services Specialist
Bible League International

舊文章 - 以原文顯示
Paratext (149 點) 提出 | 189 次瀏覽

2 個回答

+1
最佳回答

I believe that an attempt was made to many years ago to allow importing multiple USFM books from a single file. Most people don’t know about this and it is (I believe) it is rarely done so I can’t guarantee that you won’t find some way to break it.

Another thing that is probably mostly unknown is that you can import USX.

舊文章 - 以原文顯示
提出
0

After I posted this I did more research and discovered the linux/unix command csplit provides most of the functionality I need to restore a concatenated USFM file into it’s component pieces. That is, the linux command

$csplit -k /\\id\ / {1,100} *.sfm

breaks a joined file into SFM compatible files that still need naming properly.

Is there a windows equivalent to csplit? That is a known way to split a file by its contents from the command line?

Note that I"m working with American English files. In the past, I’ve had problem with Chinese characters being corrupted during the join step using the linux command line in some systems, even those that claimed POSIX compliance. If you try this at home and are working with unicode < Ux2100, I’d like to hear how it works.

舊文章 - 以原文顯示
(149 點) 提出

相關問題

0
3 個回答 287 次瀏覽
I see that Paratext 8.0.100.1 comes with USFM v. 2.502. According to http://ubsicap.github.io/usfm/ it looks as if 3. ... version. Why is that version not used for PT 8.0.100.1?
anon716631 346 提出 已提問 4月 11, 2017
0
1 個回答 233 次瀏覽
各位同事好 我正在協助一個團隊處理他們的 \r 平行參考標記,這些標記會印在章節標題下方 他們希望在某些情況下,這些標記能包含更多資訊,在章節頂部顯示該平行參考具體適用於哪些節(即目標交叉參考的來源) 例如: \s1 Golleeji ... 這是最佳做法嗎?還是有更標準的處理方式?如果這些資訊都放在頁面底部,我很樂意使用標準的交叉參考標記 謝謝 Phil
anon913937 124 提出 已提問 11月 20, 2023
0
2 個回答 45 次瀏覽
我需要將 USX 檔案轉換為 USFM 我看到 Paratext 有一個選項可以進行反向操作,但找不到將 USX 轉換為 USFM 的選項 我需要這樣做,是因為 Proskomma(用於 Scripture App Builder 建置 PWA)拒 ... 卷,理由是它們不符合標準格式 我希望將它們轉換為 USFM 後,Proskomma 處理起來會更容易
yeti 109 提出 已提問 3月 19
0
1 個回答 34 次瀏覽
我們在腳註交叉引用中使用 /xt,並設定為使用書卷的「簡稱」。但在術語表中,我想使用書卷名稱的縮寫,請問有沒有其他 USFM 代碼可以做到這一點?
Lenice 180 提出 已提問 10月 7, 2025
0
0 個回答 51 次瀏覽
目前是否已經存在任何 USFM 的語法高亮工具? 這些工具的形式可能包括 模式匹配器 解析器 語義高亮 我想知道在其他 USFM 編輯器中提供此功能有多困難。
Mercado 203 提出 已提問 10月 24, 2024
Welcome to Support Bible, where you can ask questions and receive answers from other members of the community.
They all joined together constantly in prayer, along with the women and Mary the mother of Jesus, and with his brothers.
Acts 1:14
3,045 個問題
6,005 個回答
5,671 則評論
2,026 位使用者