¡Muchas gracias por esto!
Acabo de probarlo, pero aún no hace exactamente lo que me gustaría que hiciera.
Esta es mi expresión regular, tal como la adapté:
regex:(?<!\w|\\[cv] )(kopo|menas|misika|foa|faev|siks|seven|eit|nain|ten|ileven|tuelv|tetin|fotin|fiftin|seventin|eitin|naintin|tuenti|teti|foti|fifti|siksti|seventi|eiti|nainti|andled|taosen|[0-9])
Tal como está, esto encuentra muchas más instancias de las que debería, por ejemplo, prefijos verbales que comienzan con ‹ten›. Sería genial si pudieramos limitar la búsqueda a palabras que estén:
- al principio de una palabra o precedidas por un guion
- y al final de una palabra o seguidas por un guion
Por otro lado, dado que tenemos tono gramatical a nivel de cláusula, la mayoría de las vocales pueden terminar con un diacrítico (áàāa᷄â), pero esas ocurrencias no se encuentran. Parece que cada vez que uso expresiones regulares en la función Buscar de Paratext, Ignorar diacríticos y puntos vocálicos no funciona, y por lo tanto Paratext no encuentra, por ejemplo, ‘sevén’.
¿Alguna idea de cómo podría resolver estos problemas?
¡Gracias de nuevo!
P. D.: Soy consciente de los Ajustes de números y todo está configurado para el NT, pero no parece haber una forma de incluir números escritos en palabras. (¿O sí la hay?)
Thank you very much for this!
I just tried it out, but it doesn't quite do yet what I would like it to do.
Here is my RegEx, as I adapted it:
regex:(?<!\w|\\[cv] )(kopo|menas|misika|foa|faev|siks|seven|eit|nain|ten|ileven|tuelv|tetin|fotin|fiftin|seventin|eitin|naintin|tuenti|teti|foti|fifti|siksti|seventi|eiti|nainti|andled|taosen|[0-9])
As it is, this finds far more instances than it should, for example, verb prefixes starting with ‹ten›. It would be great if we could limit the search to words being:
- at the beginning of a word or preceded by a hyphen
- and at the end of word or followed by a hyphen
On the other hand, since we have grammatical tone on clause level, most of the vowels can end up with a diacritic (áàāa᷄â), but such occurrences are not found. It seems that whenever I use RegEx in Paratext's Find, Ignore Diacritics And Vowel Points, does not work, and therefore Paratext doesn't find ‘sevén’, for example.
Any idea how I could solve these issues?
Thanks again!
PS: I'm aware of the Numbers Settings and it is all set up for the NT, but there doesn't seem to be a way to include numbers written in words. (Or is there?)
Traducción automática desde English