0 वोट
243 व्यूज़

According to my old notes, I get the cleanest text via copy and paste, with my “View Setting” set to “unformatted”.

I need from PT8 entire chapters in plain text, UTF-8 with clean, original SFM (or USFM if you want) for post-processing with another tool.

I works mostly, but I get typical output like this, see the \p marker:

\c 2\zblo 7974\s1 Marii a ŋʊm Yeesu\p\v 1 Kaŋkǝlǝ́ nɖe na bʊtǝnyarɩ 'baya Ogʊsto a shee ganɔ wàà, ba tʊ̂r bʊtǝna mbʊɖee Room kǝbaja ba jɩ ma, kǝbɛrɛ baŋunii.\v 2 Ŋkǝlǝ́ nɖee ba lee ʊrɩtʊr ʊsǝbaka ɖe ma, Kiriniyɔsɩ a lee ka Nɖiyar ashee Sirii kaatǝna.

Can you see how my paragraph markers never carry any “closing space”. I believe this is against the definition of proper SFM data. I looked it up again in usfmReference2_4.pdf:

SFMs start with a backslash character “” and end with the next space.

For my other tool, working also with Regexes, it would be very very helpful if I could properly define each SFM as “backslash whatever closing-space”.

Please advise for the best way to get clean text with valid markers out of PT8.

पुराना पोस्ट - इसकी मूल भाषा में दिखाया जा रहा है
Paratext में द्वारा (934 अंक)
फिर से दिखाया गया | 243 व्यूज़

1 उत्तर

+2 वोट
सर्वोत्तम उत्तर

I would think that your cleanest text would come directly from the text file that Paratext uses.

पुराना पोस्ट - इसकी मूल भाषा में दिखाया जा रहा है
द्वारा (9.9k अंक)

Thanks anon848905, I could never do this in Fieldworks/Flex, but I just had a look at the project data: Yes, I can copy chunks as needed from the raw data. For example I can easily grab one chapter at a time and “harvest” a book and be even faster then I would be when navigating inside PT8.

And yes, I promise, I will only work from copies… It is just too tempting to do cut and paste rather than copy and paste - because I can see what I have grabbed already. And just too routine to hit, YES, when closing the file and the editor goes “You have unsaved changes…”.

So thread closed, user happy. Thank you.

पुराना पोस्ट - इसकी मूल भाषा में दिखाया जा रहा है

संबंधित प्रश्न

+1 मत
3 उत्तर 446 व्यूज़
Exporting the Wordlist to HTML is a nice feature, except that for many users including me, working with HTML is ... that are currently marked as spelled correctly in PT? anon101508
anon101508 117 पूछा गया मार्च 11, 2019
Paratext में
0 वोट
1 उत्तर 233 व्यूज़
मुझे लगता है कि Paratext विंडो में प्रदर्शित होने वाले टेक्स्ट को बदलने का एक तरीका है, बिना नीचे मौजूद SFM फ़ाइल को ... है? क्या मैं इसकी सटीक याद रख रहा हूँ? धन्यवाद, james_post
[Moderator]
james_post
2.1k
पूछा गया फ़रवरी 4, 2022
Paratext में
0 वोट
4 उत्तर 846 व्यूज़
हमारी टीम ने नए नियम (New Testament) के प्रकाशित अनुवाद के संशोधन का कार्य शुरू किया। शुरुआत में हम आशावादी थे, ... हेडर को किसी भी स्थिति में मैन्युअल संशोधन की आवश्यकता होती है।
bit 495 पूछा गया दिसंबर 2, 2020
0 वोट
2 उत्तर 127 व्यूज़
मेरे पास एक परियोजना है जिसमें .SFC फ़ाइलें हैं। मैं इन्हें .SFM फ़ाइलों में बदलना चाहता/चाहती हूँ। इन दोनों फ़ाइल प्रकारों के ... बस फ़ाइल का नाम बदलकर काम खत्म कर सकता/सकती हूँ?)
Alex W. 191 पूछा गया नवंबर 27, 2024
Paratext में
0 वोट
1 उत्तर 225 व्यूज़
मेरे पास PNG में एक अनुवादक है, जिसका Paratext प्रोजेक्ट SFM फ़ाइलों को PT6 एक्सटेंशन के साथ सहेज रहा है। मैं ... क्या प्रोजेक्ट बनने के बाद एक्सटेंशन बदलना संभव है? anon469049
anon469049 120 पूछा गया जनवरी 14, 2021
Paratext में
Welcome to Support Bible, where you can ask questions and receive answers from other members of the community.
Accept the one whose faith is weak, without quarreling over disputable matters.
Romans 14:1
3,049 प्रश्न
6,007 उत्तर
5,672 टिप्पणियाँ
2,029 उपयोगकर्ता