0 votos
361 visualizações

A project I’m working on wants to have both singular and plural forms of the word for some glossary entries, separated by a slash. So the keyword will look something like this:
\k araba/arabaat\k*
But the markers check comes up with this error:


Yes, it is a non-wordforming character, and I’m OK with that, so how can I make the error go away? A couple of keywords contain commas, and it doesn’t complain about those, so why isn’t it accepting this other punctuation? I’ve verified that it has been validated as a character, validated in the punctuation inventory. I even tried adding it to the “Word break characters” list in the Language Settings.

What small thing am I missing?

Publicação antiga - exibida no idioma original
Paratext por (1,4k pontos) | 361 visualizações

3 Respostas

0 votos
Melhor resposta

So now I have a configuration problem. Before I made this change the only character in the “Other Characters” tab of the Language Settings was a “-” in the Word-medial punctuation field. In spite of this character being there, in the word list, words that had a hyphen were broken, e.g. “hukum”, “al” and “l” were in the word list, but not their combined forms. As soon as I added the “/” to the word-medial punctuation, I guess Paratext recalculated the word list, and it seems to have discovered that there are a lot more words in this language, including “al-hukum” and “l-hukum”, all of which are marked as unknown spelling state, a big problem for this project nearing complete Bible publication! To fix this problem, I had to move the “-” to the Word break characters field. But once I did that, the “-” is now marked as an error in the glossary keywords:


So is the underlying problem that the keyword SFM expects only wordforming characters? But it’s fine with spaces and commas, so that it can have phrases? Sounds kind of contradictory to me…
Publicação antiga - exibida no idioma original
por (1,4k pontos)

Do not move - to the Word break characters When you so that - will disappear when you view it in Print preview mode. Keep it where it is supposed to be, in Word medial punctuation. (Word break characters is intended for Thai and some other South-East Asian scripts that do not have spaces between words.)

Now that Wordlist recognizes - as a word forming character, you really should complete the Wordlist checks again. In particular you should use the Incorrectly joined or split words check to find inconsistencies in the use of the hyphen to join words.

However if you want to quickly approve all hyphenated words you can do that easily by filtering using hyphen and approving all of the words that appear. (If putting just a hyphen in the Filter box does not give you what you need to see, use regex:-)

Publicação antiga - exibida no idioma original
0 votos

Have you tried adding the / as a “Word-medial” punctuation in the Language Settings?

Publicação antiga - exibida no idioma original
por (9,9k pontos)
0 votos

That does seem to solve the problem. What I don’t understand is why it doesn’t complain about the comma as part of the keyword, since it’s not in the Word-medial punctuation list. Is it a different kind of character somehow?

Publicação antiga - exibida no idioma original
por (1,4k pontos)

I didn’t look into the history as to why, but the check currently skips commas in the keyword of the glossary entries.

From the best I can tell, the checks were added for DBL.

John+Wickberg

Publicação antiga - exibida no idioma original

Perguntas relacionadas

+2 votos
0 respostas 172 visualizações
Eu tenho um problema relacionado à questão Caracteres não formadores de palavra inesperados no glossário , mas estou movendo ... opção é negar todas as mensagens para este erro?
BruceBeatham 129 perguntada Nov 9, 2020
0 votos
1 resposta 153 visualizações
Estou enfrentando este erro na minha página de Glossário ao executar a verificação "Nejat-dooya" contains unexpected non- ... resolver este erro. Preciso da sua ajuda. Obrigado.
TSNgaihte 220 perguntada Set 24, 2024
0 votos
0 respostas 189 visualizações
This week I discovered that Alt-X does work for Unicode characters where the code is longer than 4 hex digits. The ... desired character when typing two periods: ..-->\ud804\udf
[Expert]
sewhite
3,3k
perguntada Ago 16, 2019
+1 voto
0 respostas 160 visualizações
Some Unicode characters have a code longer than 4 hex digits. For instance David Rowe was trying to help someone ... on that character gives the UTF16 coding equivalent for it.
[Expert]
sewhite
3,3k
perguntada Mar 18, 2019
0 votos
0 respostas 46 visualizações
Parece-me que a opção de espaços em branco e caracteres ocultos (WHC) impactará significativamente nossos projetos. ... outros e, talvez, receber um retorno dos desenvolvedores.
Kent Spielmann 1,8k perguntada Nov 19, 2024
Welcome to Support Bible, where you can ask questions and receive answers from other members of the community.
And let us consider how we may spur one another on toward love and good deeds, not giving up meeting together, as some are in the habit of doing, but encouraging one another—and all the more as you see the Day approaching.
Hebrews 10:24-25
3,045 perguntas
6,005 respostas
5,671 comentários
2,026 usuários