Files
iol-qwen2.5-14b-sft-awq/rag_resources/book_examples.jsonl

112 lines
720 KiB
Plaintext
Raw Normal View History

{"id":"book_02_01","context":"Here are several words and phrases in the Hmong Daw language written in Shong Lue Yang's script and the missionaries' alphabet, as well as their English translations:\n\n. | \"16B09 1pt\"16B11\"16B32OcFFL\"16B1D 1pt\"16B0B\"16B30OcFFL\"16B2C | kev ntsuas no | ‘degree’\n. | \"16B05\"16B1F | hauv | ‘inside’\n. | 1pt\"16B05\"16B36OcFFL\"16B21 1pt\"16B0F\"16B32OcFFL\"16B21 1pt\"16B0B\"16B30OcFFL\"16B2F | raug raws cai | ‘legal’\n. | \"16B0D\"16B25 1pt\"16B07\"16B32OcFFL\"16B26 | hloov mus | ‘transfer’\n. | 1pt\"16B11\"16B30OcFFL\"16B23 | qhua | ‘guest’\n. | 1pt\"16B13\"16B36OcFFL\"16B24 1pt\"16B13\"16B32OcFFL\"16B1E 1pt\"16B17\"16B36OcFFL\"16B2C | yog los nag | ‘it is raining’\n. | \"16B19 1pt\"16B16\"16B32OcFFL\"16B24 | kwv yees | ‘guess’\n. | 1pt\"16B03\"16B32OcFFL\"16B21 1pt\"16B09\"16B36OcFFL\"16B2F \"16B07\"16B1E | ris ceg luv | ‘Bermuda shorts’\n. | 1pt\"16B0D\"16B36OcFFL\"16B2C | [blank] | ‘bird’\n. | 1pt\"16B19\"16B30OcFFL\"16B2F | [blank] | ‘lobster’\n. | 1pt\"16B0B\"16B32OcFFL\"16B1F 1pt\"16B07\"16B32OcFFL\"16B1E | [blank] | ‘speak’\n. | \"16B13\"16B23 1pt\"16B11\"16B36OcFFL\"16B26 \"16B03 | [blank] | ‘dizzy’\n. | [blank] | hluav | ‘ash’\n. | [blank] | li cas | ‘how?’\n. | [blank] | neeg ntse | ‘smart, wise’\n. | [blank] | yawg | ‘grandfather’\n\nIn the missionaries' alphabet, the letter w represents a specific vowel. The letters g, s, v at the ends of the syllables are not consonants; instead, they denote tones (specific ways of pronouncing the vowels).","query":"- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"First thing we need to notice is that we do not have to provide English translations. This is one of the main characteristics of writing system problems. Moreover, the fact that we are not asked to provide English translations means that the translations are probably not relevant to solving the problem.\n\nNotebulbonThis is not always true; it is just a rule of thumb. It is possible for some problems that, although no translations are required, they can still be relevant – for example, based on semantic considerations.\n\nWe also need to keep in mind that for writing-systems problems the writing direction is relevant (from left to right or from right to left). Moreover, we notice that in Hmong the characters are grouped in clusters of one or two, while in the Latin transcription, they are grouped in syllables. Therefore, we can deduce that each group of characters represents a syllable.\n\nWe can begin by noticing the diacritics placed above some syllables. Since the mark 1pt\"25CC\"16B36OcFFL appears the greatest number of times, we can begin with it and observe that it is transliterated by the letter g at the end of the syllable. Moreover, reading the footnote, we find out that this letter does not represent a consonant or a vowel per se, but rather the syllable tone. Finally, based on example 3, in which the syllable containing the g tone appears first, we deduce that the writing direction for syllables is from left to right.\n\nUsing a similar reasoning, we identify the four possible tone marks: (1pt\"25CC\"16B30OcFFL), g (1pt\"25CC\"16B36OcFFL), s (1pt\"25CC\"16B32OcFFL), v (\"25CC). It is important to notice that the lack of a diacritic mark in Hmong is not equivalent to the lack of tone in the Latin transcription. If there is no diacritic above the syllable in Hmong, the Latin correspondent is the tone v, while if there is no tone marking in the Latin transcription (the syllable ends in a vowel and not in g, s, or v), then the Hmong syllable will have a dot on top of the syllable. Therefore, we can consider the tone marking, to some extent, as an abugida system, in which the default tone is v and the change in tone or the lack thereof is marked by diacritics.\n\nWe are left to find out how the syllable is formed, i.e., which character represents the consonant and which character represents the vowel.
{"id":"book_02_02","context":"The following are some inscriptions in the Luwian language. They correspond to some names of regions: Khamatu, Palaa, names of cities: Kurkuma, Tuvanava and names of kings: Varpalava, Tarkumuva.\n\n1. | \"145EC \"145B1\"14578\"144CA\"145EC\"14411 |\n4. | \"14578\"144CA\"14413\"14506\n\n2. | \"145DC \"145B1\"145DC\"14485\"14502 |\n5. | \"1445B \"145B1\"145DC\"1447F\"145EC\"14411\n\n3. | \"14462\"145EC\"14424\"145EC\"14502 |\n6. | \"144EF\"14485\"14462\"14506","query":"- Determine the correct correspondences.\n\n- Write in Luwian:\n\n- king Parta\n\n- king Artur\n\n- city Tartu\n\n- region Tuva\n\n- city Narva\n\n- [blank]","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"According to the introduction, the six inscriptions correspond to three categories of words: names of kings, cities, and regions. Moreover, we notice that the last character of each inscription does not appear anywhere else inside the inscription. Therefore, we can assume that these characters denote the idea of ‘king’, ‘city’, and ‘region’; thus we can divide the six inscriptions into three categories based on the last character:\n\nI | II | III\n\"145EC \"145B1\"14578\"144CA\"145EC\"14411 | \"145DC \"145B1\"145DC\"14485\"14502 | \"144EF\"14485\"14462\"14506\n\"1445B \"145B1\"145DC\"1447F\"145EC\"14411 | \"14462\"145EC\"14424\"145EC\"14502 | \"14578\"144CA\"14413\"14506\n\nWe can safely assume that this writing system is not alphabetic since each inscription has four or five characters, while their transcriptions have between five and nine characters. Moreover, it is unlikely that this system is an abugida or an abjad since there do not seem to be any diacritics appended (or similar characters). Therefore, it is most likely a syllabic system. To check this, we can try to divide the words in Latin transcription into syllables to check if the number of characters matches the number of syllables. (If you are unsure how to do this, see the discussion in chap-phonetics.)\n\nKha-ma-tu and Pa-la-a each have three syllables. Therefore, we know for sure they correspond to group III (because it is the only group in which both inscriptions have four syllables – three for the actual name and one to show the category).\n\nKur-ku-ma and Tu-va-na-va have three and four syllables and the only category that matches it is II, so we deduce that this corresponds to the cities. Moreover, since the number of syllables is different, we can already make the correct correspondences: 2 – Kurkuma and 3 – Tuvanava.\n\nThe last group is that for kings and both words have indeed four syllables (Var-pa-la-va and Tar-ku-mu-va). We get:\n\nKings | Cities | Regions\n\"145EC \"145B1\"14578\"144CA\"145EC\"14411 | \"145DC \"145B1\"145DC\"14485\"14502 | \"144EF\"14485\"14462\"14506\n| Kurkuma |\n\"1445B \"145B1\"145DC\"1447F\"145EC\"14411 | \"14462\"145EC\"14424\"145EC\"14502 | \"14578\"144CA\"14413\"14506\n| Tuvanava |\n\nLooking at the script representation of Kurkuma, we notice that the first two characters are very similar and they only differ by a little line placed on the bottom-right. Therefore, most likely, those two characters represent the syllables kur and ku and the little line on the bottom-right marks the consonant r at the end of the syllable. This is also confirmed by the fact that the names of the two kings both start with a syllable ending in r (Var-pa-la-va and Tar-ku-mu-va) and in both cases the first character has that line on the bottom right. Therefore, we deduce that the syllables are written from left to right and at the end we write the character showing the category they belong to (regions, cities or kings).\n\nThe rest of the correspondences are easily determined: the first character of Tuvarnava corresponds to the syllable tu and this syllable is also found in one of the regions' names (third character). The only region that contains the syllable tu is Khamatu, thus 6 – Khamatu and 4 – Palaa.\n\nAmong the two kings' names, one
{"id":"book_02_03","context":"Here are some words related to the mythology of the Tagbanwa people, written in the traditional script. They represent deities (Mangindusa, Bugwasin, Tungkuyanin, Tumangkuyun), names of spirits (Kiyabusan), rituals and words related to rituals (Kapupusan, kadiyang), as well as mythical places (Balugu). Their Latin transcriptions are given in random order:Note: Due to the contest taking place online, a slightly different format of the problem was used.\n\n1. | 2. | 3. | 4. | 5. | 6. | 7. | 8.\n\n\"1770\n\n\"1767 \"1773\n\n\"1765 \"1772\n\n\"176B\n\n|\n\n\"1770 \"1772\n\n\"176F\n\n\"1764\n\n\"176A \"1773\n\n|\n\n\"1768 \"1772\n\n\"176C\n\n\"1763 \"1773\n\n\"1766 \"1773\n\n|\n\n\"176C \"1773\n\n\"1763 \"1773\n\n\"176B\n\n\"1766 \"1773\n\n|\n\n\"1770\n\n\"176A \"1773\n\n\"176C \"1773\n\n\"1763 \"1772\n\n|\n\n\"1764 \"1773\n\n\"176E \"1773\n\n\"176A\n\n|\n\n\"176C\n\n\"1767 \"1772\n\n\"1763\n\n|\n\n\"1770\n\n\"1769 \"1773\n\n\"1769 \"1773\n\n\"1763\n\n- balugu\n\n- bugawasin\n\n- kadiyang\n\n- kapupusan\n\n- kiyabusan\n\n- mangindusa\n\n- tumangkuyun\n\n- tungkuyanin\n\nng = ‘ng’ in ‘king’.","query":"- Determine the correct correspondences.\n\n- Write in Tagbanwa:\n\n- mapintatan (‘to charm’)\n\n- panalangin (‘prayer’)\n\n- supisinti (‘lifestyle’)","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"We begin again by attempting to deduce what type of writing system this could be. We know for sure that it is not an alphabet since we do not have any three-letter words (corresponding to examples 6 or 7). Moreover, it is extremely unlikely that this is a picto-, ideo- or logographic system since (1) the characters are rather simplistic and similar to one another (by adding semicircles \"25CC\"1772 or \"25CC\"1773), which can be considered as diacritics), (2) we do not know the specific meaning of the words so we cannot correlate them with some pictographic or ideographic characters, and (3) using four characters to represent a single word would be rather many.\n\nWe are left with the possibilities of a syllabary, abjad, or abugida, in which each character would represent a syllable or a consonant-vowel pair (CV).\n\nWe can start by assuming it is a syllabic system and we syllabify the words. Based on this, the eight words are: ba-lu-gu, bu-ga-wa-sin, ka-di-yang, ka-pu-pu-san, ki-ya-bu-san, ma-ngin-du-sa, tu-mang-ku-yun, tung-ku-ya-nin.\n\nNotebulbonAt first sight, it may seem more likely for an English-speaking person that the word mangindusa be syllabified as mang-in-du-san and not ma-ngin-du-sal since the sound ng is not found at the beginning of the syllable in English. Either way, the number of syllables does not change so we can create a frequency table with the number of characters and syllables.\n\n# syllables | # words\n3 | 2\n4 | 6\n\nIf we first assume that each character represents a CV group, the resulting word-splitting would be ba-lu-gu, bu-ga-wa-si-n, ka-di-ya-ng, ka-pu-pu-sa-n, ki-ya-bu-sa-n, ma-ngi-n-du-sa, tu-ma-ng-ku-yu-n, tu-ng-ku-ya-ni-n, resulting in the following frequency table:\n\n# CV groups | # words\n3 | 1\n4 | 1\n5 | 4\n6 | 2\n\nSince we do not have any word represented by five or six characters, we deduce that this system cannot be based on CV groups.\n\nThe first observation is that in example 8 we have two consecutive identical characters (second and third). If we look at the given transcriptions, only one has two identical syllables, ka-pu-pu-san. Thus, we deduce \"1769 \"1773 = pu.\n\nWe must not forget that we have not yet confirmed the writing direction: it can be either top to bottom or bottom to top. Knowing that 8 = kapupusan, we deduce that the two other characters represent ka and san (not necessarily in this order). In order to find out the writing direction, we look at the last character (from top to bottom). This also appears as the last character in word 7. Thus, we have two possible cases:\n\n- Case 1. Writing from top to bottom ᝣ = san. None of the t
{"id":"book_02_04","context":"Below are some Japanese words written in the tenji system (a Japanese version of the Braille system), together with their Latin transcriptions in random order:\n\na. | <braille:u|b|sh> | d. | <braille:gh|with|s>\nb. | <braille:wh|ed> | e. | <braille:ow|b>\nc. | <braille:ch|o|k> | f. | <braille:a|o|h>\n\natari, haiku, katana, kimono, koi, sake","query":"- Determine the correct correspondences, knowing that:\nkaraoke = <braille:ch|e|i|ed>\n\n- Write in Latin script:\n<braille:ch|e|q> and <braille:a|l|for>.\n\n- Write in tenji: samurai and miso.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"We start, again, by determining the type of writing system. We already know this is not alphabetic (since we have words represented by two tenji characters and we have no two-letter words), and it is obvious we cannot talk about a picto-/ideo- or logographic system. Therefore, it most likely is a syllabic system in which each tenji character represents a syllable (or a CV group – in this case, the two are equivalent since each syllable of the given words has the structure (C)V). We infer that the four characters in karaoke represent ka, ra, o, and ke, but we still do not know in which order.\n\nWe notice that ka appears one more time as a first syllable in the word katana, while ke appears as a last syllable in the word sake. Since the first character of karaoke is the same as the first character of c., we have two possibilities:\n\n- this character represents ke and c. = sake, which is impossible since c. has three characters, not two;\n\n- this character is ka, the writing direction is left-to-right, and c. = katana. Moreover, b. = sake and we can deduce the characters for ta, na, sa.\n\nBased on this information, we can easily make the rest of the correspondences as follows: ta appears in only one other word (atari), so f. = atari and we deduce the characters for a and ri. We are left with three words to match (koi, haiku, kimono). Out of them, only one has two syllables and, therefore e. = koi and we deduce the characters for ko and i. Knowing the character for i, which must also appear in the word haiku, we can make the last two correspondences: a. = haiku, d. = kimono.\n\nWe again make a table to check whether there are any patterns based on the vowel or consonant in the syllable structure.\n\n| a | e | i | o | u\n| <braille:a> | | <braille:b> | <braille:i> |\nh | <braille:u> | | | |\nk | <braille:ch> | <braille:ed> | <braille:gh> | <braille:ow> | <braille:sh>\nm | | | | <braille:with> |\nn | <braille:k> | | | <braille:s> |\nr | <braille:e> | | <braille:h> | |\ns | <braille:wh> | | | |\nt | <braille:o> | | | |\n\nIn this case, we notice that each character represents a combination of a vowel and a consonant, the vowel being marked on the first three dots and the consonant on the last three dots. We can better illustrate this as follows:\n\n| | a | e | i | o | u\n| darkgray | lightgray <braille:vowela> | lightgray<braille:vowele> | lightgray <braille:voweli> | lightgray <braille:vowelo> | lightgray<braille:vowelu>\nh | lightgray<braille:consh> | <braille:u> | | | |\nk | lightgray<braille:consk> | <braille:ch> | <braille:ed> | <braille:gh> | <braille:ow> | <braille:sh>\nm | lightgray<braille:consm> | | | | <braille:with> |\nn | lightgray<braille:consn> | <braille:k> | | | <braille:s> |\nr | lightgray<braille:consr> | <braille:e> | | <braille:h> | |\ns | lightgray<braille:conss> | <braille:wh> | | | |\nt | lightgray<braille:const> | <braille:o> | | | |\n\nwhere the first row and column (highlighted) represent the individual characters and in order to obtain a CV syllable, we simply overlap the two components. Based on these rules, we can solve all the tasks.\n\n-\n\n- haiku\n\n- sake\n\n- katana\n\n- kimono\n\n- koi\n\n- atari\n\n-\n\n- <braille:ch|e|q> = karate\n\n- <braille:a|l|for> = anime\n\n-\n\n- samurai = <braille:wh|y|e|b>\n\n- miso = <braille:of|w>","source":"langsci_420","problem_group_id":"langsci420:2.
{"id":"book_02_05","context":"On her visit to Armenia, Millie has gotten lost in Yerevan, the nation's capital. She is now at the metro station named Shengavit, but her friends are waiting for her at the station named Barekamutyun. Other names of stations that can be found on the map below are: Gortsaranayin, Zoravar Andranik, Charbakh and Garegin Njdehi Hraparak.\n\n[VISUAL OMITTED: images/Armenian_Metro]","query":"- Assuming Millie takes a train in the correct direction, which will be the first stop after Shengavit? Write the name transcribed into English.\n\n- After boarding at Shengavit, how many stops will it take Millie to get to Barekamutyun? Don't include Shengavit itself in the number of stops.\n\n- What is the name (transcribed into English) of the end station on the short, five-station line that is currently under construction?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"2.5. Armenian\n\n- Gortsaranayin\n\n- 7\n\n- Avtogortsaran (the character resembling the letter S can be inferred to mean t from the title of the map, where the last word is metropoliten).","source":"langsci_420","problem_group_id":"langsci420:2.5","chapter":2,"chapter_title":"Writing systems","section":12,"section_title":"Practice problems","topic":"writing systems and script decipherment","language":"Armenian","author":"Dragomir R. Radev","competition":"NACLO","year":2010,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c02_s01_p01","book_method_c02_s02_p01","book_method_c02_s03_p01","book_method_c02_s04_p01","book_method_c02_s05_p01","book_method_c02_s06_p01","book_method_c02_s07_p01","book_method_c02_s08_p01","book_method_c02_s09_p01","book_method_c02_s10_p01","book_method_c02_s11_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Writing systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":false,"visual_dependency_reasons":["included_image"],"source_file":"chapters/02-WritingSystems.tex","source_line_start":652,"source_line_end":666,"solution_line_start":1100,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Writing systems\nPractice problems\nwriting systems and script decipherment\nThis practice problem belongs to the book's Writing systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nOn her visit to Armenia, Millie has gotten lost in Yerevan, the nation's capital. She is now at the metro station named Shengavit, but her friends are waiting for her at the station named Barekamutyun. Other names of stations that can be found on the map below are: Gortsaranayin, Zoravar Andranik, Charbakh and Garegin Njdehi Hraparak.\n\n[VISUAL OMITTED: images/Armenian_Metro]\n- Assuming Millie takes a train in the correct direction, which will be the first stop after Shengavit? Write the name transcribed into English.\n\n- After boarding at Shengavit, how many stops will it take Millie to get to Barekamutyun? Don't include Shengavit itself in the number of stops.\n\n- What is the name (transcribed into English) of the end station on the short, five-station line that is currently under construction?"}
{"id":"book_02_06","context":"Here are some Irish words written in the Ogham alphabet and their transcriptions in the Latin alphabet (together with their English translations) in random order:\n\n. | \"1688\"1693\"1690\"168C\"1686\"1682\"1690\"1689\"1686 | . | grá (‘love’)\n. | \"168D\"168F\"1690 \"168B\"1691 \"1689\"1686\"168F\"1691\"1694 | . | teaghlach (‘family’)\n. | \"1685\"1693\"1690\"168F\"1688 | . | Éire (‘Ireland’)\n. | \"168C\"168F\"1690 | . | neart (‘strength’)\n. | \"1684\"1694\"1691\"1689\"1686\"1690\"1694\"1685 | . | saol (‘life’)\n. | \"1693\"1694\"168F\"1693 | . | síocháin (‘peace’)\n. | \"1684\"1690\"1691\"1682\"168F | . | grá mo chroi (‘love of my heart’)","query":"- Determine the correct correspondences.\n\n- Below is the Ogham spelling of the Irish for ‘I love you’. Write it down in Latin alphabet transliteration. You can ignore accents for this task.\n\n\"1688\"1690 \"168B\"1693 \"1694 \"1685\"168C\"168F\"1690 \"1682\"1693\"1690\"1688","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"2.6. Ogham\n\n-\n\n- B\n\n- G\n\n- D\n\n- A\n\n- F\n\n- C\n\n- E\n\n- ta me i ngra leat (in reality it is Tá mé i ngrá leat).\n\nRules:\n\nWe can classify the characters depending on the number of dots or lines as well as their position or direction (vertical or diagonal, above or below the horizontal line):\n\n| 1 line | 2 lines | 3 lines | 4 lines | 5 lines\nvertical, below | | l | | s | n\nvertical, above | h | | t | c |\ndiagonal | m | g | | | r\ndots | a | o | | e | i\n\nThe accents are not marked (á = a).","source":"langsci_420","problem_group_id":"langsci420:2.6","chapter":2,"chapter_title":"Writing systems","section":12,"section_title":"Practice problems","topic":"writing systems and script decipherment","language":"Ogham","author":"Babette Verhoeven-Newsome","competition":"UKLO","year":2021,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c02_s01_p01","book_method_c02_s02_p01","book_method_c02_s03_p01","book_method_c02_s04_p01","book_method_c02_s05_p01","book_method_c02_s06_p01","book_method_c02_s07_p01","book_method_c02_s08_p01","book_method_c02_s09_p01","book_method_c02_s10_p01","book_method_c02_s11_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Writing systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/02-WritingSystems.tex","source_line_start":668,"source_line_end":690,"solution_line_start":1111,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Writing systems\nPractice problems\nwriting systems and script decipherment\nThis practice problem belongs to the book's Writing systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some Irish words written in the Ogham alphabet and their transcriptions in the Latin alphabet (together with their English translations) in random order:\n\n. | \"1688\"1693\"1690\"168C\"1686\"1682\"1690\"1689\"1686 | . | grá (‘love’)\n. | \"168D\"168F\"1690 \"168B\"1691 \"1689\"1686\"168F\"1691\"1694 | . | teaghlach (‘family’)\n. | \"1685\"1693\"1690\"168F\"1688 | . | Éire (‘Ireland’)\n. | \"168C\"168F\"1690 | . | neart (‘strength’)\n. | \"1684\"1694\"1691\"1689\"1686\"1690\"1694\"1685 | . | saol (‘life’)\n. | \"1693\"1694\"168F\"1693 | . | síocháin (‘peace’)\n. | \"1684\"1690\"1691\"1682\"168F | . | grá mo chroi (‘love of my heart’)\n- Determine the correct correspondences.\n\n- Below is the Ogham spelling of the Irish for ‘I love you’. Write it down in Latin alphabet transliteration. You can ignore accents for this ta
{"id":"book_02_07","context":"Before the Braille tactile writing system was well established in the United States, the New York Point system (NYP) was widely used in American blind education. NYP was developed in the 1860s by William Bell Walt for the New York Institute for the Blind and was intended to fix the shortcomings he perceived in the French and English Braille standards. The next six decades in blind education became known as the ``War of the Dots\", as bitter feuds developed between proponents of this homegrown system and more international Braille-based systems. NYP finally met its end after a series of public hearings convinced educational authorities that there should be a single standard for the entire English-speaking world.\n\nExperts from both sides weighed in on the systems' merits. The proponents of NYP argued that allowing letters to vary in size (from a 2x1 grid to a 2x4 grid, rather than a fixed 3x2 grid) allowed the most frequent letters to use fewer columns, resulting in space (and cost!) savings when publishing texts for the blind. For example, the number of dots needed to write the following names in each system:\n\ncompat=1.11,\n/pgfplots/ybar legend/.style=/pgfplots/legend image code/.code=\n\nplot coordinates (0cm,0.8em);,\n\ncoordinates (Pat,7) (Mary,13) (Eileen,13) (Sally,15) (Kimberly,24) (Catherine,19);\ndots needed for NYP\n\ncoordinates (Pat,10) (Mary,14) (Eileen,16) (Sally,16) (Kimberly,24) (Catherine,25);\ndots needed for Braille\n\nThey also pointed out that NYP had a distinct series of capital letters, whereas Braille only had a “capital” punctuation mark.\n\nOn the Braille side, experts such as Helen Keller wrote that the NYP capitalization system was unintuitive and confusing (“I have often mistaken D for j, I for b and Y for double o in signatures, and I waste time looking at initial letters over and over again”), and that using Braille allowed her to correspond with blind people from all over the world.\n\nThe following 12 words in NYP represent, in random order, the names: Ashley, Barb, Carl, Dave, Elena, Fred, Gerald, Heather, Ivan, Jack, Kathy, Lisa.\n\n- | | |\n| | |\n|\n|\n|\n| | |\n| |\n\n- | | |\n| | | |\n|\n|\n| |\n|\n\n- | | |\n| | |\n| |\n| |\n|\n|\n|\n|\n\n- | | |\n| | |\n|\n|\n|\n| |\n|\n\n- | | |\n| | | |\n|\n| |\n| | | |\n| |\n\n- | | |\n| | |\n\n|\n| |\n|\n|\n| |\n|\n\n- | | |\n| | |\n\n|\n| |\n|\n\n- | | |\n| | | |\n|\n|\n|\n\n- | | | |\n| | |\n\n|\n|\n|\n|\n|\n|\n\n- | | |\n| | | |\n|\n|\n| | |\n| |\n\n- | | |\n| | | | |\n| | |\n|\n| |\n| |\n\n- | | |\n| | |\n|\n|\n| |\n| |","query":"- Determine the correct correspondences.\n\n- Write in NYP: Billy, Ethan, Iggie, Orson, Sasha, Tim.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"2.7. New York Point\n\n-\n\n- Kathy\n\n- Elena\n\n- Ivan\n\n- Carl\n\n- Jack\n\n- Gerald\n\n- Lisa\n\n- Fred\n\n- Heather\n\n- Barb\n\n- Ashley\n\n- Dave\n\n-\nBilly | – | | | |\n| | |\n|\n| |\n| | |\n| |\n\nEthan | – | | | |\n| | | |\n| |\n| |\n|\n\nIggie | – | | | |\n| |\n| |\n| | | |\n| |\n\nOrson | – | | | |\n| | | |\n| |\n| |\n| |\n|\n\nSasha | – | | | |\n| | |\n|\n| | |\n| | |\n|\n\nTim | – | | | |\n| | |\n\n|\n|\n\nRules:\n\nForming the capital letter: all capital letters are four columns long and are formed by appending dots to the lowercase letter until it is four columns long, according to the following pattern:\n\n- If the last column of the lowercase letter has a dot in the upper row, add the extra dots on the lower row.\n\n- If the last column of the lowercase letter has a dot in the lower row or both dots, add the extra dots on the upper row.","source":"langsci_420","problem_group_id":"langsci420:2.7","chapter":2,"chapter_title":"Writing systems","section":12,"section_title":"Practice problems","topic":"writing systems and script decipherment","language":"New York Point","author":"Patrick Littell","competition":"UKLO","year":2011,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solu
{"id":"book_02_08","context":"Sikkim state in India has 11 official languages. Amongst these, ten are given below (the eleventh one is English) written in the Lepcha script, as well as in the Latin script:\n\nnepaalii | 1pt\"1C0D\"1C2COcFFL\"1C0E\"1C28\"1C271pt\"1C1C\"1C36OcFFL | newaar | 1pt\"1C0D\"1C2COcFFL1pt\"1C22\"1C32OcFFL\"1C28\n\nlepchaa | 1pt\"1C1C\"1C2C\"1C31OcFFL\"1C06\"1C28 | raai | \"1C1B\"1C28\"1C27\"1C23\n\nsikkim | 1pt\"1C0C\"1C2C\"1C30OcFFL\"1C25\"1C34\"1C29\"1C191pt\"1C00\"1C2C\"1C36OcFFL | gurung | \"1C03\"1C2A\"1C34\"1C1B\"1C2A\n\ntaamaang | \"1C0A\"1C28\"1C34\"1C15\"1C28 | magar | \"1C151pt\"1C03\"1C32OcFFL\n\nliimbu | \"1C27\n1pt1pt\"1C1C\"1C2EOcFFL0.15em\"1C36OcFFL\"1C13\"1C2A | sunwaar | 1pt\"1C20\"1C30OcFFL\"1C2A1pt\"1C22\"1C32OcFFL\"1C28\n\nVowel doubling denotes length. ch = ‘ch’ in ‘chop’; ng = ‘ng’ in ‘king’; w = ‘v’ in ‘van’.","query":"- One of the languages above has, in reality, two names, and its name written in the Latin script does not match the name written in the Lepcha script. Its name, transliterated from Lepcha, is drenzoongkee. Which language is this?\n\n- The Lepcha speakers, who call themselves roong haagiit (or, in the Lepcha script, \"1C34\"1C29\"1C1B \"1C1D\"1C28\"1C271pt1pt\"1C03\"1C33OcFFL0.15em\"1C36OcFFL) are composed of four main distinct communities: 1pt\"1C1B\"1C30\"1C2COcFFL\"1C34\"1C29\"1C19\"1C15\"1C2A, 1pt\"1C0A\"1C2EOcFFL\"1C28\"1C34\"1C20\"1C28\"1C15\"1C2A, \"1C27\"1C1D1pt\"1C1C\"1C2EOcFFL\"1C28\"1C15\"1C2A, and \"1C29\"1C0E\"1C25\"1C15\"1C2A. Transcribe these four community names into the Latin script.\n\n- Sikkim boasts the Kaangchenzoonggaa, the third highest peak in the world, which, in Tibetan, means ‘the five treasures of the high snow’. Transcribe the name of this peak in Lepcha.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"2.8. Lepcha\n\n- Sikkim language\n\n- renzoongmu taamsaangmu hilaammu proomu\n\n- \"1C34\"1C00\"1C281pt\"1C06\"1C2C\"1C30OcFFL\"1C34\"1C29\"1C19\"1C03\"1C28\n\nRules:\n\n- Abugida script, left-to-right.\n\n- Consonants at the beginning of the syllable:\n\n\"1C23 | \"1C13 | \"1C0C | \"1C03 | \"1C1D | \"1C06 | \"1C00 | \"1C1C\n| b | d | g | h | ch | k | l\n\"1C15 | \"1C0D | \"1C0E | \"1C1B | \"1C20 | \"1C0A | \"1C22 | \"1C19\nm | n | p | r | s | t | w | z\n\n- Vowels are marked by diacritics. The default vowel is a.\n\n\"25CC | \"25CC\"1C28 | 1pt\"25CC\"1C2COcFFL | 1pt\"25CC\"1C2C\"1C36OcFFL | \"1C27\"25CC | \"1C271pt\"25CC\"1C36OcFFL | \"1C29\"25CC | \"25CC\"1C2A\na | aa | e | ee | i | ii | oo | u\n\n- Consonants at the end of the syllable (codas) are also marked by diacritics:\n\n1pt\"25CC\"1C2EOcFFL | 1pt\"25CC\"1C30OcFFL | \"1C34\"25CC | 1pt\"25CC\"1C31OcFFL | 1pt\"25CC\"1C32OcFFL | 1pt\"25CC\"1C33OcFFL\n-m | -n | -ng | -p | -r | -t\n\n- If the syllable onset has the structure Cr, that r is marked as \"25CC\"1C25. Compare:\n\n\"1C00\"1C25 | 1pt\"1C00\"1C32OcFFL | 1pt\"1C00\"1C32OcFFL\"1C25\nkra | kar | krar","source":"langsci_420","problem_group_id":"langsci420:2.8","chapter":2,"chapter_title":"Writing systems","section":12,"section_title":"Practice problems","topic":"writing systems and script decipherment","language":"Lepcha","author":"Monojit Choudhury","competition":"PLO","year":2015,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c02_s01_p01","book_method_c02_s02_p01","book_method_c02_s03_p01","book_method_c02_s04_p01","book_method_c02_s05_p01","book_method_c02_s06_p01","book_method_c02_s07_p01","book_method_c02_s08_p01","book_method_c02_s09_p01","book_method_c02_s10_p01","book_method_c02_s11_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Writing systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"so
{"id":"book_02_09","context":"Arabic and Hebrew are today two different, mutually unintelligible languages. However, they share both grammatical similarities and several lexical correspondences. Besides loanwords (mostly from Arabic to Hebrew) from different historical periods, scholars have identified over a thousand cognates (words with a common etymological origin from Proto-Semitic, which is the partially reconstructed common ancestor language spoken around 6,000 years ago).\n\nThe two lists below show pairs of cognates, but the pairs are mixed: Arabic on the left, and Hebrew on the right. The first match (1–A) is given for you.\n\nشمس | 1. | 9999Indent | A. | שׁמשׁ\n\nغضب | 2. | | B. | כּלב\n\nولد | 3. | | C. | מלך\n\nأرض | 4. | | D. | עבד\n\nبطن | 5. | | E. | ארץ\n\nكلب | 6. | | F. | עצב\n\nحبل | 7. | | G. | קרן\n\nملك | 8. | | H. | ילד\n\nقرن | 9. | | I. | חבל\n\nعبد | 10. | | J. | בּטן\n\nThanks to regular and consistent sound changes, we have some easily identifiable patterns. Take for example the following eight words from the list above (transcribed in the Latin script):\n\nArabic | Hebrew | Translation | Arabic | Hebrew | Translation\n(r)1-3(l)4-6\nkalb | kelev | ‘dog’ | shams | shemesh | ‘sun’\nmalik | melekh | ‘king’ | qarn | qeren | ‘horn’\n'arḍ | 'erets | ‘land, earth’ | ghaḍab | ʕetsev | ‘anger, sadness’\nʕabd | ʕeved | ‘slave’ | walad | yeled | ‘child’\n\nThe apostrophe (') in both languages represents a glottal stop /ʔ/. In Arabic it is written by a hamza, which often `sits' on top of an alif; in Hebrew, it is represented by an alef and often omitted in pronunciation. Here, you should just consider it a consonant and treat it like you would treat any other consonant!\n\nThe symbol ʕ represents a voiced pharyngeal fricative /ʕ/, which is an odd sound made by contracting the muscles in the throat. It gives Arabic its unique flavour we can easily hear. In Modern Hebrew, it is silent and almost only ever appears in writing. Just think of it as an ordinary consonant!","query":"- What is the transliteration (in both Arabic and Hebrew) of the two word pairs from the lists 1-10 and A-J not included in the table showing transliterations?\n\n- The chart below shows a few letters in both scripts with their transliterations. Note that some letters may have different forms depending on the context they appear in.\n\n- Hebrew:\n\nר | ץ / צ | ע | ן / נ | ךּ / כּ | י\n(1) | (2) | (3) | (4) | k | (5)\n\nט | ח | ד | ב | בּ | א\nt | ẖ | (6) | (7) | b | ' (alef)\n\n- Arabic:\n\nـط / ـطـ / طـ | ـر / ر | ـد / د | ـح / ـحـ / حـ | ـب / ـبـ / بـ\nṭ | (8) | (9) | ḥ | (10)\n\nأ | ـو / و | ـن / ـنـ / نـ | ـل / ـلـ / لـ | ـك / ـكـ / كـ | ض\n' (alif) | (11) | (12) | (13) | (14) | (15)\n\n- Fill in the gaps (1–15).\n\n- Pair the matching cognates 2-10 and B-J from the first list (words transcribed in Arabic and Hebrew).\n\n- If ‘thousand’ in Arabic is ألف and the final letter f in Hebrew is ף, what is the transliteration of אלף – also meaning ‘thousand’ in Hebrew?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"2.9. Arabic and Hebrew\n\n- ḥabl ẖevel\nbaṭn beten\n\nNotebulbonIn Hebrew all vowels shown are e. In Arabic it is impossible to deduce the vowels based on the data, so alternative versions are also accepted (such as ḥabal or ḥabil).\n\n-\n\n- r\n\n- ts\n\n- ʕ (ayin)\n\n- n\n\n- y\n\n- d\n\n- v\n\n- r\n\n- d\n\n- b\n\n- w\n\n- n\n\n- l\n\n- k\n\n- ḍ\n\n-\n\n- A\n\n- F\n\n- H\n\n- E\n\n- J\n\n- B\n\n- I\n\n- C\n\n- G\n\n- D\n\n- 'elef","source":"langsci_420","problem_group_id":"langsci420:2.9","chapter":2,"chapter_title":"Writing systems","section":12,"section_title":"Practice problems","topic":"writing systems and script decipherment","language":"Arabic – Hebrew","author":"Gábor Parti","competition":"HKLO","year":2020,"soluti
{"id":"book_02_10","context":"Here are some Javanese words in the Javanese script, Latin script, and their English translations:\n\n. | \"A9A5\"A9BC\"A99A\"A98F\"A9B6\"A9A0\"A9C0 | penyakit | ‘disease’\n. | \"A986\"A981 \"A992\"A9BF\"A9B6\"A9B1\"A9C0 | Inggris | ‘England’\n. |\n[VISUAL OMITTED: images/Java-stacked/3.png]\n| traktor | ‘tractor’\n. |\n[VISUAL OMITTED: images/Java-stacked/4.png]\n| panyumbang | ‘donor’\n. |\n[VISUAL OMITTED: images/Java-stacked/5.png]\n| rembulan | ‘moon’\n. |\n[VISUAL OMITTED: images/Java-stacked/6.png]\n| tansah | ‘always’\n. | \"A984\"A9BA \"A9A9\"A9AB\"A9B6\"A98F | Amérika | ‘America’\n. | \"A994\"A9BD\"A9A7\"A9B8\"A9A0\"A9C0 | ngrebut | ‘to grab’\n. | \"A9B2\"A9B6\"A9A7\"A9B8\"A9BA \"A98F\"A9B4\"A9A0 | ibukota | ‘capital’\n. |\n[VISUAL OMITTED: images/Java-stacked/10.png]\n| Argentina | ‘Argentina’\n. | \"A9B1\"A9BD\"A9BA \"A994\"A9BA \"A994 | srengéngé | ‘sun’\n. |\n[VISUAL OMITTED: images/Java-stacked/12.png]\n| palsu | ‘false’\n. | \"A989\"A989\"A981\"A992\"A9A4\"A9C0 | rerenggan | ‘decoration’\n. | \"A9B2\"A981\"A9B1\"A9AD\"A9C0 | angsal | ‘to acquire’\n. | \"A9B2 \"A9B6 \"A981 \"A992\"A9B6\"A983 | inggih | ‘yes’\n. | \"A98F\"A9BC\"A989\"A9A5\"A9C0 | [blank] | ‘often’\n. |\n[VISUAL OMITTED: images/Java-stacked/17.png]\n| [blank] | ‘letter, script’\n. |\n[VISUAL OMITTED: images/Java-stacked/18.png]\n| [blank] | ‘to unload’\n. |\n[VISUAL OMITTED: images/Java-stacked/19.png]\n| [blank] | ‘to examine’\n. | \"A9A9\"A9B8\"A9AB\"A9B8\"A994\"A9BA \"A98F | [blank] | ‘to cancel’\n. | [blank] | nyolong | ‘to steal’\n. | [blank] | sepalih | ‘half’\n. | [blank] | trengginas | ‘lively’\n. | [blank] | Antartika | ‘Antarctica’\n. | [blank] | Istanbul | ‘Istanbul’\n\nny and ng are consonants; é is a vowel.","query":"- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"2.10. Javanese\n\n-\n\n- kerep\n\n- aksara\n\n- mbongkar\n\n- mrikso\n\n- murungaké\n\n- \"A9BA \"A99A\"A9B4\"A9BA \"A9AD\"A981\"A9B4\n\n- \"A9B1\"A9BC\"A9A5\"A9AD\"A9B6\"A983\n\n- \"A9A0\"A9BD\"A981\"A992\"A9B6\"A9A4\"A9B1\"A9C0\n\n-\n[VISUAL OMITTED: images/Java-stacked/sol9.png]\n\n-\n[VISUAL OMITTED: images/Java-stacked/sol10.png]\n\nRules:\n\n- Abugida script, left-to-right, with the default vowel a.\n\n- Syllables have the structure C_1(C_2)V(C_3), taking into account the following syllabification rules:\n\n- ...VC^aC^bV... ...VC^a_3-C^b_1V... if C^a is ng, h, or r;\n\n- ...VC^aC^bV... ...V-C^a_1C^b_1V... otherwise;\n\n- C_1\n\n\"A9B2 | \"A9A7 | \"A992 | \"A98F | \"A9AD | \"A9A9 | \"A9A4\n| b | g | k | l | m | n\n| | | | | |\n\"A994 | \"A99A | \"A9A5 | \"A9AB | \"A9B1 | \"A9A0 |\nng | ny | p | r | s | t |\n\nNotebulbonThe combination re has its own character: \"A989\n\n- The vowel change is shown using diacritics:\n\n\"25CC | \"25CC\"A9BC | \"A9BA \"25CC | \"25CC\"A9B6 | \"A9BA \"25CC\"A9B4 | \"25CC\"A9B8\nCa | Ce | Cé | Ci | Co | Cu\n\n- C_2:In fact, this special form of the consonant marks the fact that the vowel of the preceding consonant is deleted.\n\n[VISUAL OMITTED: images/Java-stacked/b-stack.png]\n| \"25CC\"A9BF | \"25CC\"A9BD |\n[VISUAL OMITTED: images/Java-stacked/s-stack.png]\n|\n[VISUAL OMITTED: images/Java-stacked/t-stack.png]\n\nb | abcrabc | re | s | t\n\n- C_3:\n\n\"25CC\"A981 | \"25CC\"A983 | \"25CC\"A982 | C\"A9C0\nng | h | r | else (C)\nThis is basically the vowel-removing mark, showing that the previous consonant does not have a vowel.\n\n- Special characters for capital letters:\n\n\"A984 | \"A986\n\nA | I","source":"langsci_420","problem_group_id":"langsci420:2.10","chapter":2,"chapter_title":"Writing systems","section":12,"section_title":"Practice problems","topic":"writing systems and script decipherment","language":"Javanese","author":"Tae Hun Lee","competition":"NACLO","year":2016,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_
{"id":"book_02_11","context":"Here are some Thai words written in the Thai script and their Latin transcriptions (together with their English translations) given in random order:\n\n- หวาย\n\n- ท่าน\n\n- วาย\n\n- ว่าย\n\n- กาย\n\n- ทัน\n\n- กาว\n\n- ถาด\n\n- ถาก\n\n- วัย\n\n- หลาว\n\n- ทาน\n\n- หลัง\n\n- ลาว\n\n-\n\nA. | tʰà:k | ‘to clear (a field)’ | H. | ka:w | ‘glue’\n\nB. | vâ:y | ‘to swim’ | I. | la:w | ‘Laotian’\n\nC. | lǎ:w | ‘javelin’ | J. | tʰâ:n | ‘you (formal)’\n\nD. | va:y | ‘to end’ | K. | tʰa:n | ‘charity’\n\nE. | vǎ:y | ‘rattan’ | L. | vay | ‘age’\n\nF. | tʰan | ‘to have time’ | M. | tʰà:t | ‘tray’\n\nG. | lǎŋ | ‘back’ | N. | ka:y | ‘body’\n\nA colon (:) after a vowel indicates length. The marks above vowels denote tones. This problem features four tones: medium (a), rising (ǎ), falling (â), low (à).\n\ntʰ and ŋ are consonants.","query":"- Determine the correct correspondences.\n\n- Write in Thai:\n\n15. | vǎ:n | ‘sweet’ | 17. | tʰàk | ‘to knit’\n\n16. | ya:ŋ | ‘rubber’ | 18. | vâ:w | ‘kite’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"2.11. Thai\n\n-\n\n- E\n\n- J\n\n- D\n\n- B\n\n- N\n\n- F\n\n- H\n\n- M\n\n- A\n\n- L\n\n- C\n\n- K\n\n- G\n\n- I\n\n-\n\n- หวาน\n\n- ยาง\n\n- ถัก\n\n- ว่าว\n\nRules:\n\n- Writing direction from left to right.\n\n- The character ว represents two consonants: v, if in the beginning of the word; and w, if at the end of the word.\n\n- The sound th corresponds to two Thai characters: ท and ถ. The latter is used to mark the low tone.\n\n- Vowels: a = \"25CC\"0E31 (placed on the first consonant), a: = า\n\n- Tone:\n\n- Medium – default tone (unmarked);\n\n- Rising – ห placed at the beginning of the word;\n\n- Falling – \"25CC\"0E48 placed on the first consonant;\n\n- Low – appears only in words starting with tʰ. In this case, the tone is marked by using the character ถ to mark the consonant tʰ (rather than ท).","source":"langsci_420","problem_group_id":"langsci420:2.11","chapter":2,"chapter_title":"Writing systems","section":12,"section_title":"Practice problems","topic":"writing systems and script decipherment","language":"Thai","author":"Sergey Dmitrenko","competition":"MSK","year":2001,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c02_s01_p01","book_method_c02_s02_p01","book_method_c02_s03_p01","book_method_c02_s04_p01","book_method_c02_s05_p01","book_method_c02_s06_p01","book_method_c02_s07_p01","book_method_c02_s08_p01","book_method_c02_s09_p01","book_method_c02_s10_p01","book_method_c02_s11_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Writing systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/02-WritingSystems.tex","source_line_start":1043,"source_line_end":1095,"solution_line_start":1450,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Writing systems\nPractice problems\nwriting systems and script decipherment\nThis practice problem belongs to the book's Writing systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some Thai words written in the Thai script and their Latin transcriptions (together with their English translations) given in random order:\n\n- หวาย\n\n- ท่าน\n\n- วาย\n\n- ว่าย\n\n- กาย\n\n- ทัน\n\n- กาว\n\n- ถาด\n\n- ถาก\n\n- วัย\n\n- หลาว\n\n- ทา
{"id":"book_03_01","context":"Here are 25 half-lines of Somali poetry written in a metre known as masafo:\n\n- ogaadeen ha ii dirin\n\n- duul haad amxaaraa\n\n- kaa dooni maayee\n\n- amba waa ku daba geli\n\n- dakanka iyo qaankee\n\n- anaa been dabaadee\n\n- galbeed uga dareershaan\n\n- dalkaad adigu joogtiyo\n\n- dar alliyo heshiis iyo\n\n- mase waa dayoobeen\n\n- dacalkaaga kuma shuban\n\n- miyaan duudsiyaayaa\n\n- doodaye maxaad oran\n\n- daliilkii ku siiyaye\n\n- miyaad iigu duurxuli\n\n- dorraad adigu kama dhigin\n\n- ma deldelin raggoodii\n\n- deelqaadkan aad tiri\n\n- diigaanyo ciidana\n\n- wax ma kala dillaallaa\n\n- duunyada ka qaadoo\n\n- diinkiyo dugaaggiyo\n\n- dildillaaca waaberi\n\n- dibnahaaga kama qiran\n\n- hobyo wixii ka soo degey\n\n- [blank]\n\nTo help you understand the structure of masafo, here are ten half-lines which were constructed from genuine masafo half-lines by random rearrangement of words within the half-line. Some of them might conform to the rule of versification, but the majority do not:\n\n- u anigaa lehe diin\n\n- waad nimankaad ma diidi\n\n- qoran daftarkaaga kuma\n\n- fuushaan kaama dusha\n\n- helo dabacayuun kulaan\n\n- kuu miyuu tari wax dafir\n\n- kuu daalasaayee nin\n\n- shareecada dikrigiyo\n\n- dumarkii furayaan ma\n\n- ogaadee diyaar kuu","query":"- Describe the structure of a masafo half-line.\n\n- Here are ten more masafo half-lines. Five of them are genuine, and five of them have been obtained by random rearrangement. Which is which?\n\n- war ismaaciil daarood\n\n- dir miyaad wadaagtaan\n\n- labadaad ka duudiye\n\n- ka jannadaad daahiye\n\n- adiga iyo deriskaa\n\n- digaxaarka mariyoo\n\n- ciid iyo doolo diraac\n\n- nooma keeneen darka\n\n- kala deyaayaa miyaan\n\n- wuxuun kaa danqaabaan","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"Firstly, we notice that there is no footnote about any of the sounds, so we can consider that there are no diphthongs and that two consecutive identical vowels (aa, ii, etc.) most likely denote a long vowel.\n\nNext, we need to syllabify the structures. We can do that using the rules above (VV V.V, VCV V.CV, VCCV VC.CV). We use a dash to mark those places where the word boundary could make a difference for the syllabification (for example, if we want to syllabify the phrase come inside [kʌm ɪnsaɪd], if we are to take into account the word boundary, we would get [kʌm ɪn.saɪd]. However, if the word boundary is ignored, which happens quite often in rapid speech, we would get [kʌ.mɪn.saɪd]).\n\nThe verses become:\n\n- o.gaa.deen.ha.ii.di.rin\n\n- duul.haad-am.xaa.raa\n\n- kaa.doo.ni.maa.yee\n\n- am.ba.waa.ku.da.ba.ge.li\n\n- da.kan.ka.i.yo.qaan.kee\n\n- a.naa.been.da.baa.dee\n\n- gal.beed-u.ga.da.reer.shaan\n\n- dal.kaad-a.di.gu.joog.ti.yo\n\n- dar-al.li.yo.hes.hii.s-i.yo\n\n- ma.se.waa.da.yoo.been\n\n- da.cal.kaa.ga.ku.ma-shu.ban\n\n- mi.yaan.duud.si.yaa.yaa\n\n- doo.da.ye.ma.xaad-o.ran\n\n- da.liil.kii.ku.sii.ya.ye\n\n- mi.yaad-ii.gu.duur.xu.li\n\n- dor.raad-a.di.gu.ka.ma-dhi.gin\n\n- ma.del.de.lin.rag.goo.dii\n\n- deel.qaad.kan-aad.ti.ri\n\n- dii.gaan.yo.cii.da.na\n\n- wax.ma.ka.la.dil.laal.laa\n\n- duun.ya.da.ka.qaa.doo\n\n- diin.ki.yo.du.gaag.gi.yo\n\n- dil.dil.laa.ca.waa.be.ri\n\n- dib.na.haa.ga.ka.ma.qi.ran\n\n- hob.yo.wi.xii.ka.soo.de.gey\n\n- [blank]\n\nIn the second verse, haad-am means that there is a word boundary between haad and am. If we ignore it, the syllables become haa.dam.\n\nIn order to simplify the problem, we can replace each syllable with the following notations (based on the syllable typology described above):\n\n- = syllable with short vowel and no coda = (C)V\n\n- = syllable with long vowel and no coda = (C)VV\n\n- = syllable with short vowel and coda = (C)VC\n\n- = syllable with long vowel and coda = (C)VVC\n\nThe corpus becomes:\n\n- 1.2.4.1.2.1.3\n\n- 4.haad-am.2.2\n\n- 2.2.1.2.2\n\n- 3.1.2.1.1.1.1.1\n\n- 1.3.1.1.1.4.2\n\n- 1.2.4.1.2.2\n\n- 3.beed-u.1.1.4.4\n\n- 3.kaad-a.1.1.4.1.1\n\n- dar-al.1.1.3.hiis-i.1\n\n- 1.1.2.1.2.4\n\n-
{"id":"book_03_02","context":"Given below are some words in Sarangani Manobo. In each word, the location of the stress is marked with ˈ at the beginning of the stressed syllable:\n\nˈbaso, deˈitek, meneˈnoo, beˈgas, leˈkat, ˈotaw, eteˈbay, bineleˈsan, ˈdeget, ˈbenget, miˈneles, ˈikan, ˈdoen","query":"- Mark the stress in the following words:\n\n- mengabat\n\n- tadon\n\n- migbasa\n\n- belegkong\n\n- iselem\n\n- mola\n\n- benal\n\n- medaet\n\n- Given a new word in Sarangani Manobo, how would you identify the syllables in the word? Once you have identified the syllables, how would you determine which syllable is stressed?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"Syllabification has already been discussed above; we can apply the usual rules (VV V.V, VCV V.CV, VCCV VC.CV).\n\nThe first step is to notice where the stress is usually placed. We know it is most likely placed in a window either at the beginning or at the end of the word. Looking at the word bineleˈsan, we infer that, most likely, the stress window is at the end of the word; if it were in the beginning, it would be four syllables long, which is unlikely. Assuming the stress window is at the end of the word, we notice that the stress can only fall on the last or penultimate syllable.\n\nThe next step is splitting the words into two groups, based on the stress position:\n\nPenultimate syllable | Final syllable\nˈba.so | bi.ne.le.ˈsan\nˈde.get | e.te.ˈbay\nˈben.get | le.ˈkat\nmi.ˈne.les | be.ˈgas\nˈi.kan | me.ne.ˈnoo\nˈdo.en |\nˈo.taw |\nde.ˈi.tek |\n\nSince most of the words have the stress placed on the penultimate syllable, we can assume that this is the default position and, in some cases, the stress shifts to the last syllable. The first hypothesis is that the stress moves to the last syllable if it contains a long vowel (me.ne.ˈnoo), but this is the only word that contains a long vowel: such a rule is unlikely to help us with the task. Another idea to consider is that all words in which the stress is on the last syllable end in a closed syllable (with a coda) and furthermore in all these cases the penultimate syllable is open (lacks a coda). Therefore, we could hypothesise that the stress falls, in general, on the penultimate syllable, and moves to the last syllable if the latter is closed whilst the penultimate syllable is open. This hypothesis is also quickly rejected since we have words such as de.ˈi.tek and ˈo.taw in which the penultimate syllable is open and the last syllable is closed, but the stress still falls on the penultimate syllable.\n\nTherefore, vowel length and syllable type seem to not play an important role in stress placement. Looking at the type of vowel, we notice that all words in which the stress falls on the last syllable contain the vowels a or o. Perhaps stress prefers a syllable that contains these two vowels? However, this also proves to be wrong since we have words such as ˈi.kan.\n\nUp until now, we have tried finding a reason for stress to move to the last syllable rather than the penultimate (that is, stress is attracted to the final syllable by virtue of some property of that syllable, such as a long vowel, coda, or the vowels a or o). Perhaps, instead, it is the opposite: could stress actually be repulsed from the penultimate syllable to the last one in some contexts? We can observe that in all words with stress on the last syllable, the penultimate syllable contains the vowel e. Therefore, we need to consider the hypothesis that stress prefers to be on a syllable that does not contain the vowel e.\n\nWe can try and classify the words based on the presence of the vowel e in the last or penultimate syllable:\n\n| Penultimate syll.\n(lr)2-3\nLast syll. | V=e | Ve\nV=e | ˈdeget | deˈitek\n| ˈbenget | ˈdoen\n| miˈneles |\nVe | meneˈnoo | ˈbaso\n| beˈgas | ˈotaw\n| leˈkat | ˈikan\n| eteˈbay |\n| bineleˈsan |\n\nWe notice that if only one of the last two syllables contains the vowel e, stress is placed o
{"id":"book_03_03","context":"The Kakawin poems of Old Javanese were long narrative tales made up of four-line stanzas. In the tradition of early Sanskrit poetry, each line was made up of a precise pattern of heavy and light syllables. The first section of the 11th-century CE Kakawin poem Arjunawiwāha (‘The Marriage of Arjuna’) consisted of lines which followed the śārdūlawikrīdita metre. In this metre, all lines have the same pattern: each consists of 19 syllables in an exact pattern of heavy and light syllables, except for the last syllable which can be either heavy or light. The first three syllables of a śārdūlawikrīdita line are all heavy.\n\nBelow are two stanzas from the opening section of the Arjunawiwāha:\n\n- lakṣmī niŋ suraloka sampun ayaśâŋrĕñcĕm tapa mwaŋ brata\n\n- akweh saŋ pinilih pituŋ siki tikāŋ antuk niŋ okir mulat\n\n- rwêkāŋ ādi Tilottamā pamĕkas iŋ kocap lawan Suprabā\n\n- tapwan marma tuhun lĕhĕŋ lĕhĕŋa saŋkê rūpa saŋ hyaŋ Ratih\n\n-\n\n- tambenyân liniŋir kĕtêkin inamĕr deniŋ watĕk dewata\n\n- sampūrna pwa ya mapradakṣina ta yâmūjâmidĕr pintiga\n\n- hyaŋ Brahmā dumadak caturmuka batārêndrâmahâkweh mata\n\n- eraŋ miŋgĕka kociwâmbĕk ira yan kālanyan uŋgw iŋ wuri\n\nā, â, ê, ĕ, ī, ū are vowels. ŋ, ñ, ṣ, y are consonants. You should assume that long vowels ā, ī, ū (and possibly others) occur only in heavy syllables.","query":"- Here are four more lines from the first section of the Arjunawiwāha (in the śārdūlawikrīdita metre), with their component parts (labelled A–D) in random order:\n=-5pt\n\n-\n\n-\n\n- yêkā rakwa\n\n- kapwa tâmurṣita\n\n- Indra sĕdĕŋ amwit\n\n- kinon hyaŋ\n\n-\n\n-\n\n- daśagunan\n\n- tan sora\n\n- pwa tĕkap nikā\n\n- rūpanya dentânaku\n\n-\n\n-\n\n- widyādarī mūr tĕhĕr\n\n- sinambahakĕn iŋ\n\n- liŋ hyaŋ\n\n- Śakra nahan\n\n-\n\n-\n\n- lokika\n\n- tan sangkêŋ\n\n- lwir saŋgrahêŋ\n\n- wiṣaya prayojñananira\n\n- For each verse, place the parts A-D in the right order to obtain the original verse.\n\n- Below are given four more words which appear in different lines from the first section of the Arjunawiwāha:\n\n- paramārthapandita\n\n- ametmetâśrayā\n\n- santosâhĕlĕtan\n\n- candanâpāyunan\n\n- For each of them, write the number (from 1 to 19) of the syllable in their respective lines of which these words begin (e.g., 1 if the word is at the start of the line, 3 if the word's first syllable is the third in the line).","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"3.3. Old Javanese\n\n-\n\n- yêkā rakwa kinon hyaŋ Indra sĕdĕŋ amwit kapwa tâmurṣita\n\n- tan sora pwa tĕkap nikā daśagunan rūpanya dentânaku\n\n- liŋ hyaŋ Śakra nahan sinambahakĕn iŋ widyādarī mūr tĕhĕr\n\n- tan sangkêŋ wiṣaya prayojñananira lwir saŋgrahêŋ lokika\n\n-\n\n- paramārthapandita 4th syllable\n\n- ametmetâśrayā 11th syllable\n\n- santosâhĕlĕtan 1st syllable\n\n- candanâpāyunan 14th syllable\n\nRules:\n\nThe structure of the line is as follows:\n\nH.H.H.L.L. H.L.H.L.L. L.H.H.H.L. H.H.L.(H/L)\n\nHeavy syllables are those which contain one of the vowels â, ê, e, o, ā, ī, ū, or those that have a coda. Syllabification does not respect word boundaries.","source":"langsci_420","problem_group_id":"langsci420:3.3","chapter":3,"chapter_title":"Phonetics","section":7,"section_title":"Practice problems","topic":"phonetics, stress, tone, and versification","language":"Old Javanese","author":"Michael Salter","competition":"UKLO","year":2021,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c03_s01_p01","book_method_c03_s02_p01","book_method_c03_s02_p02","book_method_c03_s02_p03","book_method_c03_s02_p04","book_method_c03_s02_p05","book_method_c03_s02_p06","book_method_c03_s03_p01","book_method_c03_s04_p01","book_method_c03_s05_p01","book_method_c03_s05_p02","book_met
{"id":"book_03_04","context":"Here are some Chuvash words, transcribed in Latin script. The stress is marked by a prime symbol before the respective syllable.\n\naˈvallăh | ‘antiquity’ | malašneˈhi | ‘future’\nasărhaˈnullă | ‘sensitive’ | mĕskĕnˈlen | ‘to respect’\nănsărˈtran | ‘unexpected’ | nušalanˈtar | ‘to make suffer’\nˈăšăn | ‘to warm up’ | ˈpĕlĕtlĕ | ‘cloudy’\nˈvărlăh | ‘seed’ | ˈpitĕrĕnčĕk | ‘closed’\nˈĕmĕrlĕh | ‘for life’ | suˈnarșă | ‘hunter’\njüˈșek | ‘sour, bitter’ | čuˈralăh | ‘slavery’\nkansĕrˈle | ‘to trip’ | čuhănˈlan | ‘to become poor’\nkĕrkunieˈhi | ‘autumn’ |\n\nă and ĕ are extra-short vowels, which are pronounced shorter than the other vowels in the language. ü and y are vowels; ș, š and č are consonants.","query":"- Mark the stress in the following words:\n\nvĕltrentărri | ‘tit (bird)’ | jyvărlăh | ‘difficulty’\nvișmine | ‘overmorrow’ | măkărălčăk | ‘convex’\nilĕrtüllĕ | ‘tempting’ |","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"3.4. Chuvash\n\n-\nvĕltrentărˈri | ‘tit (bird)’ | ˈjyvărlăh | ‘difficulty’\nvișmiˈne | ‘overmorrow’ | ˈmăkărălčăk | ‘convex’\nilĕrˈtüllĕ | ‘tempting’ |\n\nRules:\n\nThe stress is generally placed on the last syllable. If this contains an extra-short vowel, the stress instead falls on the previous syllable; this is repeated until the syllable does not contain an extra-short vowel. If all vowels are extra-short, the stress is placed on the first syllable.\n\nThis rule can be rephrased as follows: stress falls on the rightmost syllable that does not contain an extra-short vowel. If all vowels are extra-short, it falls on the first syllable.","source":"langsci_420","problem_group_id":"langsci420:3.4","chapter":3,"chapter_title":"Phonetics","section":7,"section_title":"Practice problems","topic":"phonetics, stress, tone, and versification","language":"Chuvash","author":"Artūrs Semeņuks","competition":"LLO","year":2013,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c03_s01_p01","book_method_c03_s02_p01","book_method_c03_s02_p02","book_method_c03_s02_p03","book_method_c03_s02_p04","book_method_c03_s02_p05","book_method_c03_s02_p06","book_method_c03_s03_p01","book_method_c03_s04_p01","book_method_c03_s05_p01","book_method_c03_s05_p02","book_method_c03_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonetics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/03-phonetics.tex","source_line_start":986,"source_line_end":1016,"solution_line_start":1266,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonetics\nPractice problems\nphonetics, stress, tone, and versification\nThis practice problem belongs to the book's Phonetics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some Chuvash words, transcribed in Latin script. The stress is marked by a prime symbol before the respective syllable.\n\naˈvallăh | ‘antiquity’ | malašneˈhi | ‘future’\nasărhaˈnullă | ‘sensitive’ | mĕskĕnˈlen | ‘to respect’\nănsărˈtran | ‘unexpected’ | nušalanˈtar | ‘to make suffer’\nˈăšăn | ‘to warm up’ | ˈpĕlĕtlĕ | ‘cloudy’\nˈvărlăh | ‘seed’ | ˈpitĕrĕnčĕk | ‘closed’\nˈĕmĕrlĕh | ‘for life’ | suˈnarșă | ‘hunter’\njüˈșek | ‘sour, bitter’ | čuˈralăh | ‘slavery’\nkansĕrˈle | ‘to trip’ | čuhănˈlan | ‘to become poor’\nkĕrkunieˈhi |
{"id":"book_03_05","context":"In Ancient Greece, poetry was the most celebrated art. There were several poetic forms, each with its own rules of rhythm, metre and rhyme. In order to describe these forms, the Greeks established the art of prosody, regulating the structure of verses. In this problem, you will decipher the structure of some poems of Ancient Europe.\n\nA metrical foot is a sequence of short (S) and long (L) syllables repeated in a verse. A line consists of several feet, as in the following example:\nL-S-X | L-S-X | L-S-X\n\nThe verse above contains nine syllables and three metrical feet. The metrical foot, in this case, is formed by three syllables L-S-X, where X is a syllable that can be either short or long. All the verses of a poem written in this metre will have the same structure. In Ancient Greek, a syllable counts as long if it contains a long vowel or a diphthong, or if it ends in a consonant. Syllabification does not necessarily take into account word boundaries.\n\nBelow are eight verses, transliterated into Latin script, from Oedipus Tyrannus (or Oidipous Tūrannos, in Ancient Greek), Sophocles' tragedy. These are written in the most common metre at that time. The number after the verse represents the verse's number in the original work.\n\nō tekna, Kadmou tou palai neā tropʰē, | (1)\ntinas potʰ' edrās tāsde moi tʰoazdetē | (2)\nʰiktēriois kladoisin eksestemmenoi; | (3)\npolis d' ʰomou men tʰūmiāmatōn gemei, | (4)\nʰomou de paianōn te kai stenagmatōn. | (5)\nēn ʰēmin, ōnaks, Lāïos potʰ' ʰēgemōn | (103)\npoion logon; leg' autʰis, ōs māllon matʰō | (358)\noukʰi ksunēkas prostʰen; ʰē 'kpeirā legōn; | (359)\n\nIn Ancient Greek, ai, oi, ei, au, eu, ou are diphthongs; pʰ, tʰ and kʰ are consonants. ʰ before a vowel shows that it is preceded by a puff of air similar to ‘h’. Treat sequences of a vowel and i as diphthongs, unless the i appears as ï, in which case it forms its own syllable. The apostrophe ' marks a vowel that has been deleted (or elided).\n\nIn Latin, qu is a single consonant pronounced like ‘c’ followed by ‘w’; ph and f are pronounced the same as each other, y is a vowel, ae is a diphthong. u between two vowels behaves like a consonant.\n\nThe mark \"25CC\"304 above a vowel denotes length.","query":"- Describe the syllabification rules and the structure of the metre in Oedipus Tyrannus. How many metrical feet does a verse have?\n\n- The following four verses were written by other important writers of Ancient Greek. Only one of them is written in the same metre as that of Sophocles. Which one is it?\n\n(Aeschylus, Prometheus bound, 543)\ntʰeit' ema gnōmā kratos antipalon Zeus\n\n(Aristophanes, The clouds, 609)\nprōta men kʰairein Atʰēnaioisi kai tois...\n\n(Euripides, Hippolytus, 1054)\nei pōs dunaimēn, ōs son ekʰtʰairō karā.\n\n(Herodas, Mimiamb, 14)\no pēlos akʰris ignuōn prosestēken:\n\n- Latin civilization perpetuated many aspects of Greek culture. In particular, the metre used by Sophocles and other Greek writers was later taken over and adapted by Latin writers. The following excerpt is from Medea, the tragedy by Seneca. The verse structure is identical to that of Sophocles, but one of the feet presents a slight modification. What is the modification and in which foot is it?\n\nLūcīna, custos, quaeque domitūram fretī | (2)\nTiphyn nouam frēnāre docuistī ratem… | (3)","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"3.5. Ancient Greek\n\n- Syllabification follows the usual rules: ((VV V.V, VCV V.CV, VCCV VC.CV, VCCCV VC.CCV). Moreover, based on the footnotes, we know which vowels can form diphthongs.\n\n- A light syllable (L) contains a short vowel and does not have a coda.\n\n- A heavy syllable (H): contains a long vowel OR a diphthong OR has a coda.\n\n- Each verse contains 12 syllables and has the structure:\nX H L H | X H L H | X H L H\n\nwhere X represents a syllable which can be either light or heavy.\n\nThus, the verse
{"id":"book_03_06","context":"Here are some words in Fijian, with the primary and secondary stresses marked:\n\nláko, paràimarí:, tálo, βináka, kilá:, nrè:nré:, atómi, perèsiténdi, mìnisìterí:, mbàsikètepólo, mbè:léti, Seŋái, taràusése, parò:karámu, mì:sìniŋgáni, ndàirèkitá:\n\nThe marks \"25CC\"300 and \"25CC\"301 above a vowel mark the primary and secondary stress, respectively. Two consecutive vowels form a diphthong. The mark : after a vowel denotes length.","query":"- Mark the primary and secondary stresses in the following words:\nmbelembo:tomu, mbasa:, ndikonesi, ndoketa:, palasita:, terenisisita:","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"3.6. Fijian\n\n-\n\n- mbèlembò:tómu\n\n- mbasá:\n\n- ndìkonési\n\n- ndòketá:\n\n- palàsitá:\n\n- terènisìsitá:\n\n- Primary stress is generally placed on the penultimate vowel. However, if the final syllable contains a long vowel (or a diphthong), i.e., if it's a heavy syllable, then the stress falls on it.\nFor the secondary stresses, the following algorithm is used:\n\n- Assign secondary stress to all other syllables that contain either a long vowel or a diphthong;\n\n- Once this has been done, count syllables from right to left. If this procedure reveals a sequence of two unstressed syllables, assign secondary stress to the leftmost of these.\n\nFor example, in the word mbelembo:tomu, we have the following step-by-step results:\n\n- marking primary stress mbelembo:tómu\n\n- marking secondary stress on all other heavy syllables: mbelembò:tómu\n\n- among the remaining stressed syllables, we look for the first group of two consecutive syllables from right to left: mbelembò:tómu. The leftmost syllable in this group then receives secondary stress mbèlembò:tómu. This last step is repeated until there are no more pairs of consecutive unstressed syllables.\n\nThe rule for secondary stress can be rephrased as follows: all heavy syllables which do not bear primary stress receive secondary stress. For the rest of the syllables, every other syllable from right to left receives secondary stress. The counting is reset when encountering a stress mark.","source":"langsci_420","problem_group_id":"langsci420:3.6","chapter":3,"chapter_title":"Phonetics","section":7,"section_title":"Practice problems","topic":"phonetics, stress, tone, and versification","language":"Fijian","author":"Roxana Dincă","competition":"RoLO","year":2015,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c03_s01_p01","book_method_c03_s02_p01","book_method_c03_s02_p02","book_method_c03_s02_p03","book_method_c03_s02_p04","book_method_c03_s02_p05","book_method_c03_s02_p06","book_method_c03_s03_p01","book_method_c03_s04_p01","book_method_c03_s05_p01","book_method_c03_s05_p02","book_method_c03_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonetics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/03-phonetics.tex","source_line_start":1085,"source_line_end":1102,"solution_line_start":1315,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonetics\nPractice problems\nphonetics, stress, tone, and versification\nThis practice problem belongs to the book's Phonetics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words in Fijian, with the primary and secondary stresses marked:\n\nláko, paràimarí:, tálo, βináka, kilá:, nrè:nré:, atómi, perèsiténdi, mìnisìterí:, mbàsikètepólo, mbè:léti, Seŋái, taràusése, parò:karámu, mì:sìniŋgáni, ndàirèkitá:\n\nThe marks
{"id":"book_03_07","context":"A professor of Old Norse philology, while explaining to his students the principles of Old Icelandic versification, invited them to analyse several lines from an Old Icelandic manual on versification written in the 13th century, containing lists of the names of mythological characters and objects, sometimes connected with each other by the word ok (‘and’). Below are these few lines and the professor's instructions to the students. There are some gaps in the instructions.\n\n- Randverr, Rökkvi, || Reifnir, Leifnir\n\n- Gaurekr ok Húnn, || Gjúki, Buðli\n\n- Þórr ok Hildolfr, || Hermóðr, Sigi\n\n- Byrvill, Kílmundr, || Beimi, Jórekr\n\n- Svalinn ok Randi, || Saurnir, Borði\n\n- Hildr ok Skeggöld, || Hrund, Geirdriful\n\n- Viðarr ok Baldr, || Váli ok Heimdallr\n\n- Frigg ok Freyja, || Fulla ok Snotra\n\n- Skávær, Skáviðr, || Skirfir, Virfir\n\nThe professor's instructions:\nIn Old Icelandic versification, the main poetic technique is alliteration, i.e., the repetition of the initial sounds of words. You need to know the following about Old Icelandic alliteration:\n\n- It affects only content words;\n\n- There are four positions in each line that are significant for alliteration, although the number of content words in a line can, in principle, exceed four; two such positions must be in the first half-line and two in the second (the boundary between the half-lines is indicated by the sign |);\n\n- Position [blank] always alliterates, while position [blank] never alliterates;\n\n- Out of the remaining positions ([blank] and [blank]), one must alliterate (it does not matter which one); sometimes they both alliterate.\n\nj = ‘y’ in ‘year’, h = ‘h’ in ‘hat’, þ = ‘th’ in ‘thin’, ð = ‘th’ in ‘that’; ö, y, æ, œ are vowels. The mark \"25CC\"301 above a vowel denotes length.","query":"- Fill in the blanks in the professor's instructions.\n\nAfter the students figured out the above lines, the professor said:\n\nNow look at another text, which lists the names of various parts of the ship, intended for use in certain kinds of poetic diction. At first glance, it will seem to you that the text contains gross violations of the rules, but in fact, everything is in order. Although all content words in this text [blank], for the purposes of alliteration, [blank] behave as if they were [blank] and never alliterate with each other, with [blank] or with [blank]. By the way, pay attention to line A[blank], which I showed you earlier.\n\nAfter some thought, the professor added:\nAnd keep in mind that in line B [blank] you are dealing with a rare example of “extra” alliteration, whilst lines B [blank] and B [blank], despite having more than four content words, do not show any “extra” alliterations.\n\n- Segl, skör, sigla, || sviðvís, stýri\n\n- sýjur, saumför, || súð ok skautreip\n\n- stag, stafn, stjórnvið, || stuðill ok sikulgjörð\n\n- snotra ok sólborð, || sess, skutr ok strengr\n\n- Söx, stœðingar, || sviptingr ok skaut\n\n- Fill in the blanks in the second part of the professor's instructions.\n\n- Here is one more line from the same text, which the professor absent-mindedly forgot to show the students. Mark the positions of the alliteration in it and explain your choice.\n\n- spíkr, siglutré, || saumr, lokstólpar","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"3.7. Old Norse\n\n-\n\n- third\n\n- last/fourth\n\n- first\n\n- second\n\n-\n\n- start with the same sound (s)\n\n- the combination of s with a voiceless consonant\n\n- a single sound\n\n- s\n\n- the combination of s with a voiced consonant\n\n- 9\n\n- 3\n\n- 1\n\n- 4\n\n- B6. spíkr, siglutré, || saumr, lokstólpar\nsp represents a single sound and does not alliterate.\n\nRules:\n\nThe positions that alliterate in the verses given are marked in a box. The word ok is not a content word (it is written in grey):\n\n- Randverr, Rökkvi, || Reifnir, Leifnir\n\n- Gaurekr grayok Húnn, || Gj
{"id":"book_03_08","context":"Here are some words in Chickasaw and their meanings. The words are given in IPA notation. Both primary and secondary stresses are marked.\ntʃoˈka:ˌno | ‘fly’ | ˌmaʃˈko:ˌkiʔ | ‘(name of a tribe)’\nˌlokˈtʃok | ‘mud’ | ˌokˌfokˈkol | ‘(type of snail)’\nˈa:ˌtʃomˌpaʔ | ‘local store’ | taˈla:ˌnomˌpaʔ | ‘telephone’\niˌbiɬˈkan | ‘snot, mucus’ | ˈna:ɬtoˌkaʔ | ‘policeman’\nˌʃanˈtiʔ | ‘rat’ | aˈbo:koˌʃiʔ | ‘river’\nˈsa:ɬkoˌna | ‘earthworm’ | noˌtakˈfa | ‘jaw’\nˌtʃonˈkaʃ | ‘heart’ | tʃiˌkaʃˈʃaʔ | ‘Chickasaw’\nfaˈla:t | ‘crow’ | ˌokˈtʃa:ˌlinˌtʃiʔ | ‘saviour’\n\nThe mark : after a vowel denotes length. ʃ, ɬ and ʔ are consonants. The marks \"25CC\"030D and \"25CC\"0329 before a syllable mark the primary and secondary stress, respectively.","query":"- If you are given a new Chickasaw word, how would you identify the syllables and determine which syllables get primary stress and which ones get secondary stress?\n\n- Here are some more words in Chickasaw:\n\ntaʔossa:pontaʔ | ‘finance company’\nʃimmano:liʔ | ‘(name of a tribe)’\nkanannak | ‘(type of lizard)’\nintikba:t | ‘sibling’\nokta:k | ‘prairie’\n\n- Mark the primary and secondary stress(es).","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"3.8. Chickasaw\n\n-\ntaˌʔosˈsa:ˌponˌtaʔ | ‘finance company’\nˌʃimmaˈno:ˌliʔ | ‘(name of a tribe)’\nkaˌnanˈnak | ‘(type of lizard)’\nˌinˌtikˈba:t | ‘sibling’\nˌokˈta:k | ‘prairie’\n\nRules:\n\nSyllabification follows the usual rules: (VCV V.CV, VCCV VC.CV). Note that tʃ is a single sound (voiceless post-alveolar affricate).\n\n- Primary stress:\n\n- is always placed on one of the last three syllables;\n\n- if one of the syllables has a long vowel, that syllable receives primary stress;\n\n- otherwise, primary stress is placed on the last syllable.\n\n- Secondary stress:\n\n- is placed on all syllables that have a coda;\n\n- if the last syllable does not bear primary stress, it will receive secondary stress, even if it does not have a coda.","source":"langsci_420","problem_group_id":"langsci420:3.8","chapter":3,"chapter_title":"Phonetics","section":7,"section_title":"Practice problems","topic":"phonetics, stress, tone, and versification","language":"Chickasaw","author":"Saujas Vaduguru","competition":"PLO","year":2019,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c03_s01_p01","book_method_c03_s02_p01","book_method_c03_s02_p02","book_method_c03_s02_p03","book_method_c03_s02_p04","book_method_c03_s02_p05","book_method_c03_s02_p06","book_method_c03_s03_p01","book_method_c03_s04_p01","book_method_c03_s05_p01","book_method_c03_s05_p02","book_method_c03_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonetics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/03-phonetics.tex","source_line_start":1164,"source_line_end":1195,"solution_line_start":1405,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonetics\nPractice problems\nphonetics, stress, tone, and versification\nThis practice problem belongs to the book's Phonetics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words in Chickasaw and their meanings. The words are given in IPA notation. Both primary and secondary stresses are marked.\ntʃoˈka:ˌno | ‘fly’ | ˌmaʃˈko:ˌkiʔ | ‘(name of a tribe)’\nˌlokˈtʃok | ‘mud’ | ˌokˌfokˈkol | ‘(type of snail)’\nˈa:ˌtʃomˌpaʔ | ‘
{"id":"book_03_09","context":"Below are some words written in Ligurian along with their English translations. The stress is indicated by the mark \"25CC\"0301 above the vowel. Two of the words have their stress marked on the wrong syllable.\n\nɔ:ʒél:i | ‘birds’ | sitɛ́: | ‘city’\npásta | ‘pasta’ | skwád:ra | ‘team’\nvjú:vet:a | ‘purple’ | damáskina | ‘pear’\nnɔ́stru | ‘our’ | venín | ‘poison’\ndu:seméŋte | ‘sweetly’ | kutél:u | ‘knife’\nkúm:e | ‘how’ | mejʒíŋ:a | ‘medicine’\ndát:ɔw | ‘date (fruit)’ | pe:tená: | ‘to comb’\nba:ʒú | ‘kiss’ | májstra | ‘teacher’\npú:vje | ‘dust’ | rám:u | ‘copper’\ntaramɔ́t:u | ‘earthquake’ | agýs:u | ‘sharp’\nagysá: | ‘to sharpen’ | béstja | ‘beast’\n\nbulak:u | ‘bucket’ | abityd:ine | ‘habit’\nrystegu | ‘rustic’ | akɔrdju | ‘agreement’\nfyrmine | ‘lightning’ | ɛ:gwa | ‘water’\n\nIn Ligurian, both consonants and vowels can be long (these are marked with the sign : placed after the sound); ʒ, j, ŋ and w are consonants; ɔ, y and ɛ are vowels.","query":"- Identify the two words from the list above in which stress is marked on the wrong syllable and write them with their stress on the correct syllable.\n\n- Mark the stress in the following words:","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"3.9. Ligurian\n\n- vju:vét:a and bá:ʒu\n\n- bulák:u abitýd:ine rýstegu akɔ́rdju fýrmine ɛ́:gwa\n\nRules:\n\n- Stress always falls before a long consonant.\n\n- If there are no long consonants, the stress falls on the last heavy syllable (a syllable counts as heavy if it contains either a coda or a long vowel).\n\nThe rule can be simplified if we treat long consonants as sequences of identical consonants broken up by a syllable boundary.","source":"langsci_420","problem_group_id":"langsci420:3.9","chapter":3,"chapter_title":"Phonetics","section":7,"section_title":"Practice problems","topic":"phonetics, stress, tone, and versification","language":"Ligurian","author":"Kevin Liang","competition":"UKLO","year":2020,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c03_s01_p01","book_method_c03_s02_p01","book_method_c03_s02_p02","book_method_c03_s02_p03","book_method_c03_s02_p04","book_method_c03_s02_p05","book_method_c03_s02_p06","book_method_c03_s03_p01","book_method_c03_s04_p01","book_method_c03_s05_p01","book_method_c03_s05_p02","book_method_c03_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonetics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/03-phonetics.tex","source_line_start":1197,"source_line_end":1231,"solution_line_start":1440,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonetics\nPractice problems\nphonetics, stress, tone, and versification\nThis practice problem belongs to the book's Phonetics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nBelow are some words written in Ligurian along with their English translations. The stress is indicated by the mark \"25CC\"0301 above the vowel. Two of the words have their stress marked on the wrong syllable.\n\nɔ:ʒél:i | ‘birds’ | sitɛ́: | ‘city’\npásta | ‘pasta’ | skwád:ra | ‘team’\nvjú:vet:a | ‘purple’ | damáskina | ‘pear’\nnɔ́stru | ‘our’ | venín | ‘poison’\ndu:seméŋte | ‘sweetly’ | kutél:u | ‘knife’\nkúm:e | ‘how’ | mejʒíŋ:a | ‘medicine’\ndát:ɔw | ‘date (fruit)’ | pe:tená: | ‘to comb’\nba:ʒú | ‘kiss’ | májstra | ‘teacher’\npú
{"id":"book_04_01","context":"Here are some Irish verb forms for the imperative and present indicative, as well as their English translations:\n\nImperative | Present indicative\n(lr)1-2(lr)3-4\nfan | ‘Stay!’ | fanaim | ‘I stay’\ncuir | ‘Put!’ | cuireann sé | ‘he puts’\nceannaigh | ‘Buy!’ | ceannaíonn tú | ‘yousg buy’\ncreid | ‘Believe!’ | creidim | ‘I believe’\ncríochnaigh | ‘End!’ | críochnaíonn sé | ‘he ends’\ndéan | ‘Do!’ | déanann sí | ‘she does’\nsmaoinigh | ‘Think!’ | smaoiníonn sibh | ‘youpl think’\nól | ‘Drink!’ | ólann sé | ‘he drinks’\noibrigh | ‘Work!’ | oibríonn siad | ‘they work’\nfág | ‘Leave!’ | fágann muid | ‘we leave’\néirigh | ‘Raise!’ | éiríonn sí | ‘she raises’\nlig | ‘Let!’ | ligeann tú | ‘yousg let’\ntosaigh | ‘Start!’ | tosaím | ‘I start’\nith | ‘Eat!’ | itheann sé | ‘he eats’","query":"- Translate into Irish:\n\n- ‘yousg believe’\n\n- ‘yousg stay’\n\n- ‘I end’\n\n- ‘I work’\n\n- ‘I put’\n\n- ‘I drink’\n\n- ‘I think’\n\n- ‘yousg start’\n\n- [blank]","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"The first step in this type of problems is to classify the different forms based on their characteristics. In this case, we can classify the forms in the right column based on person. Nevertheless, we notice that all forms, except for those in the 1st person singular (1sg), end in nn and are followed by the subject pronoun, while forms in the 1sg are composed of one word (do not include the subject pronoun).\n\nWe can start with the 1sg forms.\n\nImperative | Present indicative\n(r)1-2(l)3-4\nfan | ‘Stay!’ | fanaim | ‘I stay’\ncreid | ‘Believe!’ | creidim | ‘I believe’\ntosaigh | ‘Start!’ | tosaím | ‘I start’\n\nWe observe that there are three ways in which we can construct the 1sg form: adding the suffix -aim, adding the suffix -im or replacing the ending -igh with -ím.\n\nSimilarly, we will want to observe how the other verb forms are formed. We ignore the subject pronouns added in the end but focus on the verb form. We get:\n\n- 2sg: add the suffix -eann or replace -igh -íonn;\n\n- 3sg, masc.: add suffixes -eann, -ann or replace\n-igh -íonn;\n\n- 3sg, fem.: add suffix -ann or replace -igh -íonn;\n\n- 1pl: add suffix -ann;\n\n- 2pl: replace -igh -íonn;\n\n- 3pl: replace -igh -íonn;\n\nWe notice that all these verb forms are obtained through the same three transformations, independent of person: adding the suffixes -eann, -ann or replacing -igh -íonn. Therefore, most likely, in this problem, there are only two verb forms: the form for 1sg and the form for all the others (in which case, to avoid ambiguity, the subject pronoun is added after the verb). This is also supported by the way in which the task is phrased since all the verb forms that we need to translate are only 1sg or 2sg.\n\nMoreover, just like in the case of 1sg, there are three possible transformations. Therefore, we can assume that Irish verbs can be classified into three different categories, each of them having its own way of expressing these forms. We can easily figure out that the verbs which end in -igh in their imperative form the 1sg present indicative by replacing -igh with -ím and the other forms by replacing -igh with -íonn. We are left to discover the environment for the other two transformations. For this, we classify the verbs into two categories, based on the way they form their non-1sg form:\n\n-eann | -ann\ncuir ‘to put’ | déan ‘to do’\nlig ‘to eat’ | ól ‘to drink’\nith ‘to let’ | fág ‘to leave’\n\nThe classification of these forms seems to not take into account any semantic features (related to meaning) since there seems to be nothing in common between the verbs {‘to put’, ‘to eat’, ‘to let’} compared with {‘to do’, ‘to drink’, ‘to leave’}. Therefore, most likely there are some phonetic char
{"id":"book_04_02","context":"Here are 11 words in the four dialects of the Roro language and their English translations:\n\nDialect |\n(lr)1-4\nHisiu | Delena | Kivori | Paitana | Translation\n[blank] | aitau | [blank] | [blank] | ‘three’\naihi | aisi | aihi | aisi | ‘crab’\ncici | sisi | čiči | cici | ‘meat’\nebeoahi | ebeoasi | ebeoahi | ebeoaci | ‘he ran’\nhiabu | siabu | hiabu | ciabu | ‘smoke’\nnihe | nite | nihe | nite | ‘tooth’\nicu | [blank] | [blank] | icu | ‘nose’\nmaciu | [blank] | [blank] | [blank] | ‘tree’\nmoihana | moitana | moihana | [blank] | ‘look!’\nmahi | [blank] | [blank] | maci | ‘beast’\ncubu | subu | čubu | cubu | ‘grass’","query":"- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"In linguistics problems featuring related languages or dialects, the first step is usually figuring out which sounds are different across the data and which remain the same. We start by writing the words for which all four forms are given:\n\nDialect |\n(lr)1-4\nHisiu | Delena | Kivori | Paitana | Translation\naihi | aisi | aihi | aisi | ‘crab’\ncici | sisi | čiči | cici | ‘meat’\nebeoahi | ebeoasi | ebeoahi | ebeoaci | ‘he ran’\nhiabu | siabu | hiabu | ciabu | ‘smoke’\nnihe | nite | nihe | nite | ‘tooth’\ncubu | subu | čubu | cubu | ‘grass’\n\nWe should notice that there are some words, like ‘crab’ and ‘he ran’, that have h in Hisiu and Kivori, s in Delena, and c in Paitana. In others, like ‘tooth’, Hisiu and Kivori h correspond to t in Delena and Paitana.\n\nSumming up, we can write the following correspondences, numbered for convenience:\n-\n\n| Hisiu | Delena | Kivori | Paitana\n(1) | h | s | h | c\n(2) | c | s | č | c\n(3) | h | t | h | t\n\nIn other problems of this type, we would need to find the environment in which h\nbecomes s or t in Delena (and c or t\nin Paitana), but, before doing this, it is important to check the tasks\nand see whether we need this type of generalisation.\n\nDialect\n(lr)1-4\nHisiu | Delena | Kivori | Paitana | Translation\n[blank] | aitau | [blank] | [blank] | ‘three’\nicu | [blank] | [blank] | icu | ‘nose’\nmaciu | [blank] | [blank] | [blank] | ‘tree’\nmoihana | moitana | moihana | [blank] | ‘look!’\nmahi | [blank] | [blank] | maci | ‘beast’\n\nOn the first row, the only Delena consonant that undergoes any transformation is t and, according to the rules above, there is only one transformation that yields the sound t in Delena (rule 3). Similarly, for tasks 4–8 there is only one possible transformation of the sound c in Hisiu (rule 2). The issue emerges for tasks 9–11 where we are given Hisiu words that contain the sound h. Nevertheless, for task 9 we know that the sound h in Hisiu becomes t in Delena (thus, we know that this is rule 3). As for tasks 10–11, we notice that the Hisiu h becomes c in Paitana (thus following rule 1). Now we have enough information to solve the tasks and there is no need to identify the environments in which each transformation takes place.\n\n-\nDialect\n(lr)1-4\nHisiu | Delena | Kivori | Paitana | Translation\naihau | aitau | aihau | aitau | ‘three’\nicu | isu | iču | icu | ‘nose’\nmaciu | masiu | mačiu | maciu | ‘tree’\nmoihana | moitana | moihana | moitana | ‘look!’\nmahi | masi | mahi | maci | ‘beast’\n\nRules:\n\nHisiu | Delena | Kivori | Paitana\nh | s | h | c\nc | s | č | c\nh | t | h | t","source":"langsci_420","problem_group_id":"langsci420:4.2","chapter":4,"chapter_title":"Phonology","section":3,"section_title":"Complementary distributionComplementary distribution","topic":"phonological rules and sound correspondences","language":"Roro","author":"Vladimir I. Belikov","competition":"MSK","year":1991,"solution_tier":"detailed_worked","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c04_s03_p01","book_method_c04_s03_p02"],"method_link_type":"auth
{"id":"book_04_03","context":"- Below are some Indonesian verbs in their active and passive forms and their English translations. Fill in the blanks.\n\nActive | Passive | Translation\nmeŋuji | diuji | ‘to test’\nmeŋeja | dieja | ‘to spell’\nmeŋgaruk | digaruk | ‘to scratch’\nmendapat | didapat | ‘to obtain’\nmemberi | diberi | ‘to give’\nmenulis | ditulis | ‘to write’\nmemutus | diputus | ‘to cut off’\n[blank] | dibuat | ‘to make’\n[blank] | dipilih | ‘to choose’\n\n-1.5ŋ = ‘ng’ in ‘king’.\n\n- Below are some Mandar words in their active and passive forms, and their English translations. Fill in the blanks.\n\nActive | Passive | Translation\nmambatta | dibatta | ‘to split’\nmandeŋŋeq | dideŋŋeq | ‘to carry on the back’\nmaŋidaŋ | diidaŋ | ‘to crave’\nmappasuŋ | dipasuŋ | ‘to send out’\nmattunu | ditunu | ‘to burn’\nmassiraq | disiraq | ‘to tie’\n[blank] | ditimbe | ‘to throw’\n[blank] | dipande | ‘to feed’\n\nŋ = ‘ng’ in ‘king’.\n\n- Below are some words in Quechua in their nominative, genitive and locative (preposition ‘in’) form, as well as their English translations. Fill in the blanks.\n\nNominative | Genitive | Locative | Translation\nkam | kamba | 808080 | ‘yousg’\natam | 808080 | atambi | ‘frog’\nhatum | [blank] | [blank] | ‘the big one’\nsinik | sinikpa | 808080 | ‘porcupine’\nčilis | čilispa | 808080 | ‘streamless region’\nsača | 808080 | sačapi | ‘jungle’\npunǰa | 808080 | punǰapi | ‘day’\n\n-1.5č = ‘ch’ in ‘chop’.\n\n- Given below are some words in Zoque in their base forms and their 1sg possessive (‘my’), as well as their English translations. Fill in the blanks.\n\nBase | Possessed | Translation\nburru | mburru | ‘donkey’\npama | mbama | ‘clothing’\ntatah | ndatah | ‘father’\nfaha | faha | ‘belt’\nsis | sis | ‘meat’\nflawta | [blank] | ‘harmonica’\nšapun | šapun | ‘soap’\ndisko | [blank] | ‘phonograph record’\nkayu | ŋgayu | ‘horse’\nkopak | [blank] | ‘head’\n\nŋ = ‘ng’ in ‘king’, š = ‘sh’ in ‘shop’.\n\n- Below are given some nouns in Lunyole in their singular and plural forms, as well as their English translations. Fill in the blanks.\n\nSingular | Plural | Translation\noludaalo | endaalo | ‘day’\noluboyooboyo | emboyooboyo | ‘hullabaloo’\nolufudu | efudu | ‘rainbow’\nolukalala | ekalala | ‘list’\nolusosi | [blank] | ‘mountain’\nolubafu | [blank] | ‘rib’\nolupagi | [blank] | ‘spoke (of a bike)’\nolutambi | [blank] | ‘candle’\n\n- All five of the languages in this problem display processes that avoid a specific type of sound combination. Fill in the blanks to describe this generalisation. The blanks should be chosen from: vowel, consonant, nasal, voiced consonant, voiceless consonant.\n\nAvoid having a [blank] directly followed by a [blank].","query":"Complete every requested item in the problem context.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"- We notice that the active form is formed from the passive form by replacing the prefix di with one of the prefixes: meŋ, men, or mem. Moreover, we notice that, in certain situations, the first vowel of the root is dropped. We split the given words based on these transformations:\n\n| meŋ | men | mem\nUnchanged stem | -uji | -dapat | -beri\n| -eja | |\n| -garuk | |\nStem drops first consonant | | -tulis | -putus\n\nLooking at the three types of prefix (which only differ by the type of nasal consonant), we expect a nasal assimilation. Thus we notice, indeed, that the place of articulation of the final nasal of the prefix assimilates to the first consonant of the stem. If the stem begins with a vowel, the nasal used is ŋ. Moreover, we notice that if the stem begins with a voiceless stop, it gets dropped. Hence, the blanks are (1) membuat and (2) memilih.\n\nWe can write the rules in different ways:\n\n- using words: di meN (N is a nasal): N
{"id":"book_04_04","context":"Here are some verbal forms in Valley Yokutsin four different forms and their English translations:\n\nDubitative | Passive voice | Non-future | Imperative | Translation\nDubitative | Passive voice | Non-future | Imperative | Translation\ndo:sol | [blank] | doshin | [blank] | ‘to report’\n[blank] | dubut | dubhun | dubka | ‘to conduct’\nyawa:lal | yawa:lit | yawalhin | yawalka | ‘to follow’\nlogwol | logwit | logiwhin | logiwka | ‘to pulverise’\nwo:nol | wo:nit | wonhin | [blank] | ‘to hide’\nxatal | xatit | xathin | [blank] | ‘to eat’\n[blank] | [blank] | [blank] | toyixka | ‘to treat’\n[blank] | ʔopo:tit | ʔopothin | ʔopotko | ‘to get out of bed’\nʔugnal | [blank] | ʔugunhun | [blank] | ‘to drink’\n[blank] | [blank] | [blank] | ʔilikka | ‘to sing’\n[blank] | lihmit | lihimhin | [blank] | ‘to run’\n[blank] | luklut | [blank] | [blank] | ‘to bury’\n[blank] | koʔit | [blank] | koʔko | ‘to throw’\nme:kal | [blank] | [blank] | [blank] | ‘to swallow’\n\nk, t, x, y, ʔ are consonants. The mark : after a vowel denotes length.","query":"- Fill in the blanks. If you believe that some blanks could allow multiple answers, write them all.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"We begin by segmenting the forms in order to figure out which part is the stem and which are the morphemes corresponding to the four verb forms. We notice that, for two of the verbs, we are given all four forms:\n\nDubitative | Passive voice | Non-future | Imperative | Translation\nyawa:lal | yawa:lit | yawalhin | yawalka | ‘to follow’\nlogwol | logwit | logiwhin | logiwka | ‘to pulverise’\n\nComparing these forms, we deduce that the dubitative is formed by adding the suffixes -al or -ol, the passive voice is formed using the suffix -it, the non-future using -hin and the imperative using -ka. Moreover, we notice that the stems can undergo some changes (for the verb ‘to follow’ there are two possible stems yawa:l and yawal, while for the verb ‘to pulverise’ there is logw and logiw). Moreover, we notice that, in both cases, the dubitative and passive voice use the same stem, while the other two forms use the “modified” stem.\n\nLooking at the other words in the corpus, we notice that each of the four forms has two possible suffixes: -al and -ol for\ndubitative, -it and -ut for passive voice, -hin and -hun for non-future, and -ka and -ko for imperative.\n\nSince the passive voice is the one with the most examples, we can start with it. We make a table in which we split the words into two classes, based on the choice of the suffix used in forming the passive voice.\n\n-it | -ut\nyawa:lit | ‘to follow’ | dubut | ‘to conduct’\nlogwit | ‘to pulverise’ | luklut | ‘to bury’\nwo:nit | ‘to hide’ | |\nxatit | ‘to eat’ | |\nɁopo:tit | ‘to get out of bed’ | |\nlihmit | ‘to run’ | |\nkoɁit | ‘to throw’ | |\n\nWe can easily see that the suffix -ut is used for the verbs whose stem contains u, so this can be viewed as vowel harmony, being a total assimilation of i to u. This phenomenon can also be written as a phonological rule as follows:\n\nitutuC0#\n\nSince the suffixes corresponding to non-future also feature the vowels i and u, we expect them to be chosen based on the same rule (or a very similar rule). Indeed, analysing the given examples, we notice that -hun is used if the last vowel of the stem is u, while -hin is used otherwise.\n\nWe use the same process for the other two verb forms, the dubitative and the imperative, by making a table for each of them. Moreover, considering that for passive and non-future, the rule was purely phonological, not semantic (it does not depend on the meaning of the word, but rather on the form of the word), in the following table we do not include the translation of the verbs.\n\nDubitative | Imperative\n(lr)1-2(lr)3-4\n-al | -ol | -ka | -ko\nyawa:lal | do:sol | dubka | Ɂopotko\nxatal | logwol
{"id":"book_04_05","context":"Given below are some words in Evenki in five different cases: nominative singular (e.g., ‘the dog’), nominative plural (e.g., ‘the dogs’), directional-locative singular (e.g., ‘to the dog’), possessive 1sg (e.g., ‘my dog’), possessive 1pl (e.g., ‘our dog’), as well as their English translations:\n\nNom. sg | Nom. pl | Dir-loc. sg | Pos. 1sg | Pos. 1pl | Transl.\nbitəg | bitəgsəl | bitəglə | bitəgwi | bitəgmʉn | ‘book’\nbɵ:s | bɵ:ssɵl | bɵ:slɵ | bɵ:swi | [blank] | ‘cloth’\nudun | udunsul | [blank] | udunbi | udunmun | ‘rain’\nigga | iggasal | iggala | iggawi | iggamun | ‘flower’\nixʉldʉ:r | ixʉldʉ:rsʉl | ixʉldʉ:rlɵ | ixʉldʉ:rwi | ixʉldʉ:rmʉn | ‘shovel’\nnʉxʉn | nʉxʉnsʉl | nʉxʉnlɵ | nʉxʉnbi | nʉxʉnmʉn | ‘brother’\noron | oronsol | oronlo | oronbi | oronmun | ‘place’\nsatan | satansal | satanla | satanbi | satanmun | ‘candy’\ntəggəŋ | təggəŋsəl | təggəŋlə | təggəŋbi | təggəŋmʉn | ‘car’\nʉ:ŋkʉ | ʉ:ŋkʉsʉl | ʉ:ŋkʉlɵ | ʉ:ŋkʉwi | ʉ:ŋkʉmʉn | ‘towel’\nxocco | xoccosol | xoccolo | xoccowi | xoccomun | ‘shop’\niggə | iggəsəl | [blank] | [blank] | iggəmʉn | ‘tail’\njʉ: | [blank] | jʉ:lɵ | [blank] | jʉ:mʉn | ‘house’\nxə:m | [blank] | [blank] | [blank] | [blank] | ‘meal’\ndo:son | [blank] | [blank] | [blank] | [blank] | ‘salt’\n\nNom. sg | Nom. pl | Dir-loc. sg | Pos. 1sg | Pos. 1pl | Translation\numatta | umattasul | umattalo | umattawi | umattamun | ‘egg’\na:gun | a:gunsal | a:gunla | a:gunbi | a:gunmun | ‘hat’\nʉrəl | ʉrəlsʉl | ʉrəllɵ | ʉrəlwi | ʉrəlmʉn | ‘child’\nmoriŋ | moriŋsol | moriŋlo | moriŋbi | moriŋmun | ‘horse’\nxɵ:ggʉ | [blank] | [blank] | [blank] | [blank] | ‘leg’\noʃitta | [blank] | [blank] | [blank] | [blank] | ‘star’\nxərʉ:ldi: | [blank] | [blank] | [blank] | [blank] | ‘quarrel’\n\nʉ and ɵ are vowels pronounced like u and o respectively, but with the tongue placed centrally (central vowels); a is similar to ‘u’ in ‘cut’, but the tongue is placed more towards the back (back vowel); ə = ‘ea’ in ‘pearl’, ŋ = ‘ng’ in ‘king’, ʃ = ‘sh’ in ‘shop’.\n\nThe mark : after a vowel denotes length.","query":"- Fill in the blanks.\n\n- You are given some more Evenki words in the same five forms:\n\n- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"We start by noticing that the nominative singular can be considered the base form. From this word, all other forms are derived. Moreover, we notice that there are no changes (alternations) in this form, so we can consider it as the stem of all other forms.\n\nThe next step is separating the segments that are added to obtain the other forms. So for the nominative plural we have the suffixes -sal, -səl, -sol, -sul, -sɵl, -sʉl. Such a variety of suffixes which only differ by the vowel is a strong indicator towards some process of vowel harmony. Therefore, we can classify the words based on the suffix they select for the nominative plural form.\n\n-sal | -səl | -sol | -sul | -sɵl | -sʉl\nigga | bitəg | oron | udun | bɵ:s | ixʉldʉ:r\nsatan | təggəŋ | xocco | | | nʉxʉn\n| iggə | | | | ʉ:ŋkʉ\n\nIndeed, the supposition of vowel harmony is easily confirmed. We can observe that the vowel in the suffix is identical to the last vowel of the stem. Therefore, we deduce that the nominative plural suffix is -sVl, where V is the last vowel of the stem. Since the suffix repeats the final stem vowel in all cases, we can talk here about a total assimilation rather than vowel harmony.\n\nWe can do the same thing for the other three forms. For the directional-locative singular form, we find the suffixes -la, -lə, -lo, -lɵ.\n\n-la | -lə | -lo | -lɵ\nigga | bitəg | oron | bɵ:s\nsatan | təggəŋ | xocco | ixʉldʉ:r\n| | | nʉxʉn\n| | | ʉ:ŋkʉ\n| | | jʉ:\n\nIn this case, we can assume that, similar to
{"id":"book_04_06","context":"Here are some words and expressions in Guoyu, the Taiwanese dialect of Mandarin Chinese, as well as their correspondences in the secret language La-Mi (both transcribed into Latin script):\n\nGuoyu | La-Mi | Translation\ne hiau | le i liau hi | ‘capable’\nbe tsai | [blank] | ‘to go shopping’\n[blank] | lat tit | ‘to hit’\ntsin tiam | [blank] | ‘very tired’\n[blank] | laŋ gin | ‘human’\ngi | [blank] | ‘justice’\npiaʔ | liaʔ piʔ | ‘wall’\nkam tsia | lam kin lia tsi | ‘sugarcane’\npɔŋ hɔŋ | lɔŋ pin lɔŋ hin | ‘gust (of wind)’\nho keʔ | [blank] | ‘guest of honour’\npak kak | lak pit lak kit | ‘to clean’\ntsap ap | [blank] | ‘ten boxes’\n\nAll vowel combinations are pronounced as a single syllable (syllables are separated by blanks); p, t, k, ts, ʔ are consonants; ɔ is a vowel; ŋ = ‘ng’ in ‘king’.","query":"- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"4.6. La-Mi\n\n-\nGuoyu | La-Mi | Translation\nbe tsai | le bi lai tsi | ‘to go shopping’\ntat | lat tit | ‘to hit’\ntsin tiam | lin tsin liam tin | ‘very tired’\ngaŋ | laŋ gin | ‘human’\ngi | li gi | ‘justice’\nho keʔ | lo hi leʔ kiʔ | ‘guest of honour’\ntsap ap | lap tsit lap it | ‘ten boxes’\n\nRules:\n\nWe note with C1, V and C2 the onset,\nnucleus and coda of the syllable, respectively. Now each syllable can be\nwritten as (C1)V(C2). The transformation\nis: C1VC2 lVC2\nC1iX, where X depends on C2, as\nfollows:\n\nC2 | X\n|\nʔ | ʔ\nm, n, ŋ | n\np, t, k | t\n\nWe deduce that X is identical with C2, but with an assimilated place of articulation (alveolar). Exception: ʔ (which remains unchanged). Another explanation is:\n\n- ʔʔ\n\n- +nasal n\n\n- +stop\n–voicet","source":"langsci_420","problem_group_id":"langsci420:4.6","chapter":4,"chapter_title":"Phonology","section":7,"section_title":"Practice problems","topic":"phonological rules and sound correspondences","language":"La-Mi","author":"Evgeniya Korovina","competition":"MSK","year":2013,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c04_s01_p01","book_method_c04_s02_p01","book_method_c04_s03_p01","book_method_c04_s03_p02","book_method_c04_s04_p01","book_method_c04_s05_p01","book_method_c04_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/04-phonology.tex","source_line_start":1458,"source_line_end":1488,"solution_line_start":1866,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonology\nPractice problems\nphonological rules and sound correspondences\nThis practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words and expressions in Guoyu, the Taiwanese dialect of Mandarin Chinese, as well as their correspondences in the secret language La-Mi (both transcribed into Latin script):\n\nGuoyu | La-Mi | Translation\ne hiau | le i liau hi | ‘capable’\nbe tsai | [blank] | ‘to go shopping’\n[blank] | lat tit | ‘to hit’\ntsin tiam | [blank] | ‘very tired’\n[blank] | laŋ gin | ‘human’\ngi | [blank] | ‘justice’\npiaʔ | liaʔ piʔ | ‘wall’\nkam tsia | lam kin lia tsi | ‘sugarcane’\npɔŋ hɔŋ | lɔŋ pin lɔŋ hin | ‘gust (of wind)’\nho keʔ | [blank] | ‘guest of honour’\npak kak | lak pit lak kit | ‘to clean’\ntsap ap | [blank] | ‘ten boxes’\n\nAll vowel combinations are pronounced as a single syllable (syllables are separated by
{"id":"book_04_07","context":"Here are some verbal forms in Tolaki in their active and passive voice and their English translations:\n\nActive | Passive | Translation | Active | Passive | Translation\n(lr)1-3(lr)4-6Active | Passive | Translation | Active | Passive | Translation\n(lr)1-3(lr)4-6alo | inalo | ‘take’ | wala | niwala | ‘enclose’\ndaga | nidaga | ‘guard’ | baho | [blank] | ‘bathe’\nehe | inehe | ‘want’ | inu | [blank] | ‘drink’\ngeru | nigeru | ‘scrape’ | kulisi | [blank] | ‘dig’\nhunu | hinunu | ‘burn’ | mala | [blank] | ‘shorten’\nluarako | niluarako | ‘grab’ | paho | [blank] | ‘plant’\noli | inoli | ‘fly’ | ruru | [blank] | ‘collect’\nsaru | sinaru | ‘borrow’ | solongako | [blank] | ‘empty’\ntena | tinena | ‘order’ | usa | [blank] | ‘crush’\n\nnahu | ninahu | ‘cook’\n\nw = ‘v’ in ‘van’","query":"- Fill in the blanks.\n\n- At first, the author wanted to include the following example, but they changed their mind believing it can be confusing. Why might this be?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"4.7. Tolaki\n\n-\nbaho | nibaho | ‘bathe’\ninu | ininu | ‘drink’\nkulisi | kinulisi | ‘dig’\nmala | nimala | ‘shorten’\npaho | pinaho | ‘plant’\nruru | niruru | ‘collect’\nsolongako | sinolongako | ‘empty’\nusa | inusa | ‘crush’\n\n- Because it is not known which n is part of the stem and which one is part of the affix. Thus, it can either be the prefix ni- added to the word nahu, or the infix -in-, added after the first consonant of the word nahu.\n\nRules:\n\n- If the lexeme starts with a vowel, add in- at the beginning;\n\n- If the lexeme starts with a voiced consonant, add ni- at the beginning.\n\n- If the lexeme starts with a voiceless consonant, add -in- after the first consonant.","source":"langsci_420","problem_group_id":"langsci420:4.7","chapter":4,"chapter_title":"Phonology","section":7,"section_title":"Practice problems","topic":"phonological rules and sound correspondences","language":"Tolaki","author":"Peter Arkadiev","competition":"MSK","year":2016,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c04_s01_p01","book_method_c04_s02_p01","book_method_c04_s03_p01","book_method_c04_s03_p02","book_method_c04_s04_p01","book_method_c04_s05_p01","book_method_c04_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/04-phonology.tex","source_line_start":1490,"source_line_end":1521,"solution_line_start":1926,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonology\nPractice problems\nphonological rules and sound correspondences\nThis practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some verbal forms in Tolaki in their active and passive voice and their English translations:\n\nActive | Passive | Translation | Active | Passive | Translation\n(lr)1-3(lr)4-6Active | Passive | Translation | Active | Passive | Translation\n(lr)1-3(lr)4-6alo | inalo | ‘take’ | wala | niwala | ‘enclose’\ndaga | nidaga | ‘guard’ | baho | [blank] | ‘bathe’\nehe | inehe | ‘want’ | inu | [blank] | ‘drink’\ngeru | nigeru | ‘scrape’ | kulisi | [blank] | ‘dig’\nhunu | hinunu | ‘burn’ | mala | [blank] | ‘shorten’\nluarako | niluarako | ‘grab’ | paho | [blank] | ‘plant’\noli | inoli | ‘fly’ | ruru | [blank] | ‘collect’\nsaru | sinaru | ‘borrow’ | solongako | [blank] | ‘emp
{"id":"book_04_08","context":"Minangkabau is a language of Indonesia that features a number of “play languages” that people use for fun, like Pig Latin in English. One of these “play languages” is Sorba. Here are some examples of standard Minangkabau words and their Sorba play language equivalents:Source: John Henderson, University of Western Australia, with the assistance of Sophie Crouch. Based on Crouch (2008, 2009) and data from the MPI EVA Minangkabau corpus.\n\nMinang- | | | Minang- | |\nkabau | Sorba | English | kabau | Sorba | English\n(r)1-3(l)4-6Minang- | | | Minang- | |\nkabau | Sorba | English | kabau | Sorba | English\n(r)1-3(l)4-6raso | sora | ‘taste’ | mangecek | cermange | ‘talk’\nrokok | koro | ‘cigarette’ | bakilek | lerbaki | ‘lightning’\nrayo | yora | ‘celebrate’ | sawah | warsa | ‘rice field’\nsusu | sursu | ‘milk’ | pitih | tirpi | ‘money’\nbaso | sorba | ‘language’ | manangih | ngirmana | ‘cry’\nlamo | morla | ‘long time’ | urang | raru | ‘person’\nmati | tirma | ‘dead’ | cubadak | darcuba | ‘jackfruit’\nbulan | larbu | ‘month’ | iko | kori | ‘this’\nminum | nurmi | ‘drink’ | gata-gata | targa-targa | ‘flirtatious’\nlilin | lirli | ‘wax’ | maha-maha | harma-harma | ‘expensive’\nmintak | tarmin | ‘request’ | campua | purcam | ‘mix’\napa | para | ‘father’ | | |","query":"- Write the Sorba equivalents of the following words:\n\nrancak (‘nice’) | jadi (‘happen’)\nmakan (‘eat’) | marokok (‘smoking’)\nampek (‘hundred’) | limpik-limpik (‘stuck together’)\ndapua (‘kitchen’) |\n\n- If you know a Sorba word, can you work backwards to a single standard Minangkabau word? Demonstrate with the Sorba word lore (‘good’).\n\n- Another “play language” is Solabar. The rules for converting a standard Minangkabau word to Solabar can be worked out from the following examples:\n\nMinangkabau | Solabar | English\n(r)1-3\nbaso | solabar | ‘language’\ncampua | pulacar | ‘mix’\nmakan | kalamar | ‘eat’\n\n- What is the Solabar equivalent of the Sorba word tirpi (‘money’)?\n\n- In writing Minangkabau, does the sequence ng represent one sound or two sounds? Provide evidence that supports your answer.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"4.8. Sorba\n\n- caran, dirja, karma, kormaro, peram, kormaro, peram, pirlim-pirlim, purda\n\n- The word cannot be uniquely determined since the coda of the last syllable disappears. Moreover, it is not known if the r in lore is part of the root or if it is added, thus lore can result from any of the following words: relo, elo, reloC, eloC (where C can be any consonant).\n\n- tilapir\n\n- A single sound, since the word manangih becomes ngirmana in Sorba. If ng was two sounds, the word would have become girmanan.\n\nRules:\n\nIn order to form the Sorba word, we need to take the last syllable of the word, remove its coda (only keep the first consonant – if any – and the first vowel) and add it to the beginning of the word, separated by an r. If the word already begins with an r, no extra r is added. Thus, a word such as ...(C)V(V)(C) becomes (C)Vr.... It is important to notice that two consecutive vowels will form a diphthong, according to the example cam.pua pu-r.cam. If it were a hiatus, we would get cam.pu.a a-r.cam.pu.\n\nIn order to form the Solabar word, the same process as in Sorba is applied, but the connecting r is replaced by la (resulting in (C)Vla...).","source":"langsci_420","problem_group_id":"langsci420:4.8","chapter":4,"chapter_title":"Phonology","section":7,"section_title":"Practice problems","topic":"phonological rules and sound correspondences","language":"Sorba","author":"John Henderson","competition":"UKLO","year":2010,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c04_s01_p01","book_method_c04_s02_p01","book_method_c04_s03_p01","
{"id":"book_04_09","context":"Here are some Arabic nouns in their definite and indefinite form, as well as their English translations:\n\nIndefinite | Definite | Translation | Indefinite | Definite | Translation\n(lr)1-3(lr)4-6\nšams | aššams | ‘sun’ | ɣayma | alɣayma | ‘cloud’\nqamar | alqamar | ‘moon’ | maṭar | almaṭar | ‘rain’\nnaǯm | annaǯm | ‘star’ | ṭaqs | aṭṭaqs | ‘weather’\nfaǯr | alfaǯr | ‘dawn’ | ǯafāf | alǯafāf | ‘draught’\nyawm | alyawm | ‘day’ | bard | albard | ‘coldness’\nð̣alām | að̣ð̣alām | ‘darkness’ | taham | attaham | ‘heat’\nsamāʹ | assamāʹ | ‘sky’ |\n\nArabic linguists classify the consonants into ``lunar'' and ``solar'' consonants. It is known that š and t are solar consonants, while q and m are lunar consonants.\n\nš = ‘sh’ in ‘shop’, ǯ = ‘j’ in ‘judge’, y = ‘y’ in ‘year’, ð = ‘th’ in ‘that’, θ = ‘th’ in ‘thin’, q = ‘c’ in ‘car’, x = ‘ch’ in ‘loch’, ɣ is similar to x, but voiced. ʕ and ' are consonants.\n\nA dot below a consonant denotes its special pronunciation (so-called emphatic). A bar above the vowel denotes length.","query":"- Write the definite form of the following nouns:\n\n- muðannab (‘comet’)\n\n- barq (‘lightning’)\n\n- θalǯ (‘ice’)\n\n- nār (‘fire’)\n\n- ḍaw' (‘light’)\n\n- layla (‘night’)\n\n- ɣurūb (‘sunrise’)\n\n- šitā' (‘winter’)\n\n- rabīʕ (‘spring’)\n\n- ṣayf (‘summer’)\n\n- xarīf (‘autumn’)\n\n- [blank]\n\n- Classify the consonants b, ḍ, f, ɣ, n, r, θ, y and z into the two categories proposed by Arabic linguists (solar and lunar).\n\n- Arab language historians know that the sound represented by one of the Arabic letters was, in time, replaced by another one (in this problem the modern variant is used). Determine which letter it is.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"4.9. Arabic\n\n-\nalmuðannab | albarq | aθθalǯ\nannār | aḍḍawʹ | allayla\nalɣurūb | aššitāʹ | arrabīʕ\naṣṣayf | alxarīf |\n\n- Lunar consonants: b, f, ɣ, y\n\nSolar consonants: ḍ, n, r, θ, z\n\n- ǯ\n\nRules:\n\nSolar consonants include the coronal consonants. In their case, the definite form is constructed by adding the prefix aX- (where X is the first consonant of the word). Lunar consonants are all the rest (labial and dorsal), and if a word begins with a lunar consonant, the definite form is simply obtained by adding the prefix al-. Alternatively, we can write:\n\nl\nC\n[+coronal]\n\nC\n[+coronal]\n\nThe consonant ǯ, although coronal, does not assimilate the prefix al-; therefore, most likely, in the past, it was\npronounced as a dorsal (probably as a voiced palatal plosive).","source":"langsci_420","problem_group_id":"langsci420:4.9","chapter":4,"chapter_title":"Phonology","section":7,"section_title":"Practice problems","topic":"phonological rules and sound correspondences","language":"Arabic","author":"Anton Somin","competition":"Elementy","year":null,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c04_s01_p01","book_method_c04_s02_p01","book_method_c04_s03_p01","book_method_c04_s03_p02","book_method_c04_s04_p01","book_method_c04_s05_p01","book_method_c04_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/04-phonology.tex","source_line_start":1580,"source_line_end":1630,"solution_line_start":1969,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonology\nPractice problems\nphonological rules and
{"id":"book_04_10","context":"Sesotho is primarily spoken in two countries: South Africa and Lesotho. As a result, two different orthographies are used for this language. For example, the word ‘ostrich’ is written mpjhe in South Africa, but mpshe in Lesotho.\n\nBelow are given some words in Sesotho. Some of them are written in one of the orthographies (whether South Africa or Lesotho), while others are written in both orthographies. Two of the words are the same in both orthographies.\n\noache\nKholu\nkalima\nkutloisiso\nkadima\nkgwedi\nnwa\nyohle\n'nete\nKgodu\nlula\nhlompshoa\nnkwe\nya\ntitjhere\nphela\nnkoe\ntjhelete\ncha\nnnete\nMokgatjhane\nnngwaya\nchelete\nntate\n'me\nea\na","query":"- For each of the words, determine in which orthography it is written. For the words written in only one of the two orthographies, provide their equivalent in the other one.\n\n- In which orthography are the words joang and shwa written?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"4.10. Sesotho\n\n-\n\nSouth Africa | Lesotho\ndula | lula\nhlompjhwa | hlompshoa\nkadima | kalima\nKgodu | Kholu\nkgwedi | khoeli\nkutlwisiso | kutloisiso\nmme | 'me\nMokgatjhane | Mokhachane\nnkwe | nkoe\nnnete | 'nete\n\nSouth Africa | Lesotho\nnngwaya | 'ngoaea\nntate\nnwa | noa\nphela\ntitjhere | tichere\ntjha | cha\ntjhelete | chelete\nwatjhe | oache\nya | ea\nyohle | eohle\n\nThe words in bold are those that do not appear in the dataset.\n\n- joang – Lesotho (in South Africa it is written as\njwang)\n\nshwa – South Africa (in Lesotho it is written as\nshoa)\n\nRules:\nWe have the following sound correspondences:\n\nSouth Africa | Lesotho\ndi | li\ndu | lu\nkg | kh\nmm | 'm\nnn | 'n\npjh | psh\ntjh | ch\nw + vowel | o + vowel\ny + vowel | e + vowel","source":"langsci_420","problem_group_id":"langsci420:4.10","chapter":4,"chapter_title":"Phonology","section":7,"section_title":"Practice problems","topic":"phonological rules and sound correspondences","language":"Sesotho","author":"Tamila Krashtan","competition":"UkrLO","year":2021,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c04_s01_p01","book_method_c04_s02_p01","book_method_c04_s03_p01","book_method_c04_s03_p02","book_method_c04_s04_p01","book_method_c04_s05_p01","book_method_c04_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/04-phonology.tex","source_line_start":1632,"source_line_end":1674,"solution_line_start":2017,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonology\nPractice problems\nphonological rules and sound correspondences\nThis practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nSesotho is primarily spoken in two countries: South Africa and Lesotho. As a result, two different orthographies are used for this language. For example, the word ‘ostrich’ is written mpjhe in South Africa, but mpshe in Lesotho.\n\nBelow are given some words in Sesotho. Some of them are written in one of the orthographies (whether South Africa or Lesotho), while others are written in both orthographies. Two of the words are the same in both orthographies.\n\noache\nKholu\nkalima\nkutloisiso\nkadima\nkgwedi\nnwa\nyohle\n'nete\nKgodu\nlula\nhlompshoa\nnkwe\nya\ntitjhere\nphela\nnkoe\ntjhelete\ncha\nnnete\nMokgatjhane\nnngwaya\nchelete\nntate\n'me\nea\na\n- For each of the words, determine in which orthography it is written. For the words written in only one of the two orthographies, provide their equivalent in the other
{"id":"book_04_11","context":"The Dutch language uses suffixes to form the diminutives of nouns. Any Dutch noun can be transposed to its diminutive form. Below are some Dutch words together with their diminutive forms and their English translation:\n\nWord | Diminutive | Transl. | Word | Diminutive | Transl.\n(lr)1-3(lr)4-6\nleeuwerik | leeuwerikje | ‘lark’ | dag | dagje | ‘day’\ntrom | trommetje | ‘drum’ | stro | strootje | ‘straw’\npeer | peertje | ‘pear’ | vrucht | vruchtje | ‘fruit’\nsnor | snorretje | ‘moustache’ | steeg | [blank] | ‘alley’\nhuis | huisje | ‘house’ | steg | [blank] | ‘way’\npotlood | potloodje | ‘pencil’ | bioscoop | [blank] | ‘cinema’\nparaplu | parapluutje | ‘umbrella’ | deur | [blank] | ‘door’\nviool | viooltje | ‘violin’ | auto | [blank] | ‘car’\ntuin | tuintje | ‘garden’ | zoon | [blank] | ‘son’\nster | sterretje | ‘star’ | zon | [blank] | ‘sun’\nkomkommer | komkommertje | ‘cucumber’ | mus | [blank] | ‘sparrow’\nsla | slaatje | ‘salad’ | winkel | [blank] | ‘store’\nkam | kammetje | ‘comb’ | bal | [blank] | ‘ball’\nweb | webje | ‘internet’ | ballet | [blank] | ‘ballet’\npin | pinnetje | ‘pin’ | pyjama | [blank] | ‘pyjamas’\nverhaal | verhaaltje | ‘story’ | schim | [blank] | ‘ghost’\nwodka | wodkaatje | ‘vodka’ | [blank] | petje | ‘beret’","query":"- Fill in the blanks.\n\n- The word vlootje is a homonym, representing the diminutive of two different words. Which are these words?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"4.11. Dutch\n\n-\n\n- steegje\n\n- stegje\n\n- bioscoopje\n\n- deurtje\n\n- autootje\n\n- zoontje\n\n- zonnetje\n\n- musje\n\n- winkeltje\n\n- balletje\n\n- balletje\n\n- pyjamaatje\n\n- schimmetje\n\n- pet\n\n- vlo and vloot\n\nRules:\n\nThe diminutive depends on the last sound of the word. We have the following cases:\n\n- vowel add Vtje (where V is the last vowel of the word);\n\n- obstruent (stop or fricative) add -je;\n\n- sonorant (nasal or liquid – m, n, l, r):\n\n- if the word has only one syllable and the vowel is short, add the suffix -Cetje (where C is the last consonant);\n\n- else, add -tje (if (1) the word has only one syllable and contains a long vowel or a diphthong or (2) the word contains more than a syllable).","source":"langsci_420","problem_group_id":"langsci420:4.11","chapter":4,"chapter_title":"Phonology","section":7,"section_title":"Practice problems","topic":"phonological rules and sound correspondences","language":"Dutch","author":"Ksenia Gilyarova","competition":"MSK","year":2003,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c04_s01_p01","book_method_c04_s02_p01","book_method_c04_s03_p01","book_method_c04_s03_p02","book_method_c04_s04_p01","book_method_c04_s05_p01","book_method_c04_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/04-phonology.tex","source_line_start":1676,"source_line_end":1708,"solution_line_start":2090,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Phonology\nPractice problems\nphonological rules and sound correspondences\nThis practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nThe Dutch language uses suffixes to form the diminutives of nouns. Any Dutch noun can be transposed to its diminutive form. Below are some Dutch words together with their diminutive forms and their English translation:\n\nWord | Diminut
{"id":"book_04_12","context":"Here are some words in Finnish (F) and Estonian (E), declined for the nominative, genitive, and illative cases:Source: Adapted after a problem by Andrei Zaliznyak (published in Задачи лингвистических олимпиад 1965–1975 (Problems for the Linguistics Olympiad 1965–1975), Moscow, 2007.\n\n| Nominative | Genitive | Illative\n(lr)2-3(lr)4-5(lr)6-7\nEnglish | F | E | F | E | F | E\n‘people’ | rahvas | rahvas | [blank] | rahvas | [blank] | [blank]\n‘naked’ | [blank] | [blank] | paljaan | [blank] | paljaaseen | paljasse\n‘row’ | [blank] | toores | tuoreen | toore | tuoreeseen | tooresse\n‘axe’ | kirves | kirves | [blank] | [blank] | kirveeseen | [blank]\n‘ready’ | valmis | [blank] | valmiin | valmi | [blank] | valmisse\n‘part’ | osa | [blank] | osan | osa | osaan | ossa\n‘city’ | linna | [blank] | linnan | linna | [blank] | linna\n‘village’ | külä | külä | külän | küla | külään | [blank]\n‘shelter’ | maja | maja | majan | maja | majaan | majja\n‘ace’ | [blank] | äss | [blank] | [blank] | ässään | ässa\n‘wheel’ | püörä | [blank] | [blank] | [blank] | [blank] | [blank]\n‘snow’ | lumi | lumi | lumen | [blank] | lumeen | lumme\n‘horn’ | sarvi | sarv | [blank] | sarve | sarveen | sarve\n‘cape’ | niemi | neem | niemen | neeme | [blank] | neeme\n‘hackberry’ | [blank] | toom | [blank] | toome | [blank] | [blank]\n‘sea’ | [blank] | [blank] | [blank] | [blank] | [blank] | merre\n\nFor this problem the orthography of Finnish has been slightly changed. In reality, the character denoted here by ü is written as y.","query":"- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"4.12. Finnish – Estonian\n\n-\n\n- rahvaan\n\n- rahvaaseen\n\n- rahvasse\n\n- paljas\n\n- paljas\n\n- palja\n\n- tuores\n\n- kirveen\n\n- kirve\n\n- kirvesse\n\n- valmis\n\n- valmiiseen\n\n- osa\n\n- linn\n\n- linnaan\n\n- külla\n\n- ässä\n\n- ässän\n\n- ässa\n\n- pöör\n\n- püörän\n\n- pööra\n\n- püörään\n\n- pööra\n\n- lume\n\n- sarven\n\n- niemeen\n\n- tuomi\n\n- tuomen\n\n- tuomeen\n\n- toome\n\n- meri\n\n- meri\n\n- meren\n\n- mere\n\n- mereen\n\nRules:\n\nWe divide the nouns (nominative, Finnish) into three classes: ending in s (preceded by a vowel), ending in i, and ending in another vowel. We get:\n\n| Nominative | Genitive | Illative\n(lr)2-3(lr)4-5(lr)6-7\n| F | E | F | E | F | E\nClass I | -Vs | -Vs | -VVn | -V | -VVseen | -Vsse\nClass II | -V | * | -Vn | -V | -VVn | -V\nClass III | -i | * | -en | -e | -een | -e\n\n* To form the Estonian nominative, if the Finnish root contains a diphthong or a consonant cluster, the final vowel is dropped in Estonian (-V). Else, the form is identical to the Finnish one (exception: ä a / _ #).\n\n- Diphthongs in Finnish become long vowels in Estonian: V1V2 V2V2.\n\n- If the Estonian nominative ends in a vowel, the consonant before it is\ndoubled in the illative. (-CV -CCV or Ci -CCe).","source":"langsci_420","problem_group_id":"langsci420:4.12","chapter":4,"chapter_title":"Phonology","section":7,"section_title":"Practice problems","topic":"phonological rules and sound correspondences","language":"Finnish – Estonian","author":"Vlad A. Neacșu","competition":"HKLO","year":2021,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c04_s01_p01","book_method_c04_s02_p01","book_method_c04_s03_p01","book_method_c04_s03_p02","book_method_c04_s04_p01","book_method_c04_s05_p01","book_method_c04_s06_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Phonology chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/04-phonology.tex","source_line_start":1710,
{"id":"book_04_13","context":"Given below are some verbal roots in Bari, together with two different forms, as well as their English translations. The shaded cells represent forms that are not essential for solving the problem (they may exist).\n\nRoot | Form 1 | Form 2 | Tone | Translation\ndé̱ʔ | dí̱lí̱kí̱n | 808080 | | ‘to bend’\nkú̱r | kú̱rà̱kí̱n | kú̱rà̱râ̱ʔ | | ‘to borrow’\n'dó̱k | 'dú̱kú̱kí̱n | 808080 | | ‘to carry’\nmók | mòkákìn | mòkáràʔ | | ‘to catch’\n[blank] | tú̱kú̱kí̱n | tó̱kó̱rô̱ʔ | | ‘to cut with an axe’\n[blank] | [blank] | tòjúpùrùʔ | | ‘to dress’\nyúk | yùkúkìn | yùkúrùʔ | | ‘to shepherd’\n'dép | 'dépákín | [blank] | | ‘to hold’\ngáʔ | [blank] | 808080 | g | ‘to seek’\nlú̱sà̱k | [blank] | 808080 | | ‘to defrost’\nsà̱pû̱k | [blank] | [blank] | | ‘to return’\n'yút | 'yùtúkìn | [blank] | | ‘to seed’\ntò̱kû̱ | tò̱kú̱kì̱n | tò̱kú̱à̱rà̱ʔ | t | ‘to preach’\nbú̱dú̱ | bú̱dú̱kí̱n | 808080 | | ‘to reach the top’\nbáʔ | bàlákìn | 808080 | | ‘to punish’\nsó̱n | sú̱nyú̱kí̱n | só̱nyó̱rô̱ʔ | | ‘to send (something)’\nyà̱kî̱ | yà̱kí̱kì̱n | yà̱kí̱à̱rà̱ʔ | | ‘to send (someone)’\ndòdông' | dòdóng'àkìn | dòdóng'àràʔ | | ‘to shake’\n[blank] | 'bórókín | 808080 | | ‘to smear’\nlì̱lî̱ng' | lì̱lí̱ng'à̱kì̱n | [blank] | | ‘to exterminate’\nré̱m | rí̱mí̱kí̱n | 808080 | | ‘to inject’\nbérén | [blank] | 808080 | | ‘to poison’\n[blank] | lókín | 808080 | | ‘to dry in the sun’\ndwán | dwànyákìn | 808080 | | ‘to open’\nlák | [blank] | lákárâʔ | | ‘to untie’\ndó̱k | [blank] | 808080 | g | ‘to unpack’\n\n'b, 'd, 'y, ng', ny, y, ʔ are consonants. A line below a vowel denotes that the vowel is pronounced with an advanced tongue root (+ATR). The marks \"25CC\"301, \"25CC\"300 and \"25CC\"302 above a vowel denote high, low and falling tones, respectively.","query":"- Bari verbs can be classified into two groups, based on their tone. Fill in the column “Tone”, specifying whether the verb has the tone g (it behaves like gáʔ and dó̱k) or t (it behaves like tò̱kû̱).\n\n- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"4.13. Bari\n\n-\n\nRoot | Form 1 | Form 2 | Tone | Translation\ndé̱ʔ | dí̱lí̱kí̱n | 808080 | g | ‘to bend’\nkú̱r | kú̱rà̱kí̱n | kú̱rà̱râ̱ʔ | g | ‘to borrow’\n'dó̱k | 'dú̱kú̱kí̱n | 808080 | g | ‘to carry’\nmók | mòkákìn | mòkáràʔ | t | ‘to catch’\ntó̱k | tú̱kú̱kí̱n | tó̱kó̱rô̱ʔ | g | ‘to cut with an axe’\ntòjûp | tòjúpùkìn | tòjúpùrùʔ | t | ‘to dress’\nyúk | yùkúkìn | yùkúrùʔ | t | ‘to shepherd’\n'dép | 'dépákín | 'dépárâʔ | g | ‘to hold’\ngáʔ | gálákín | 808080 | g | ‘to seek’\nlú̱sà̱k | lú̱sà̱kà̱kí̱n | 808080 | g | ‘to defrost’\nsà̱pû̱k | sà̱pú̱kà̱kì̱n | sà̱pú̱kà̱rà̱ʔ | t | ‘to return’\n'yút | 'yùtúkìn | 'yùtúrùʔ | t | ‘to seed’\ntò̱kû̱ | tò̱kú̱kì̱n | tò̱kú̱à̱rà̱ʔ | t | ‘to preach’\nbú̱dú̱ | bú̱dú̱kí̱n | 808080 | g | ‘to reach the top’\nbáʔ | bàlákìn | 808080 | t | ‘to punish’\nsó̱n | sú̱nyú̱kí̱n | só̱nyó̱rô̱ʔ | g | ‘to send (something)’\nyà̱kî̱ | yà̱kí̱kì̱n | yà̱kí̱à̱rà̱ʔ | t | ‘to send (someone)’\ndòdông' | dòdóng'àkìn | dòdóng'àràʔ | t | ‘to shake’\n'bóró | 'bórókín | 808080 | g | ‘to smear’\nlì̱lî̱ng' | lì̱lí̱ng'à̱kì̱n | lì̱lí̱ng'à̱rà̱ʔ | t | ‘to exterminate’\nré̱m | rí̱mí̱kí̱n | 808080 | g | ‘to inject’\nbérén | bérényákín | 808080 | g | ‘to poison’\nló | lókín | 808080 | g | ‘to sundry’\ndwán | dwànyákìn | 808080 | t | ‘to open’\nlák | lákákín | lákárâʔ |
{"id":"book_04_14","context":"Here are some words and phrases in Cushillococa Ticuna and their English translations:\n\nˈka̰1 a5 tɨ3 | ‘ka̰1 tree leaves’\nˈku43 te4 e3 ɟa̰1 | ‘your husband's sister’\nˈku43 ʔa̰1 | ‘your mouth’\nˈku43 ʔu4 ne1 | ‘your entire body’\nna4 ˈme43 ʔe5 tʃi1 | ‘it is really good’\nna4 ˈbu3 ʔu1 ra1 | ‘it is sort of immature’\nˈto5 ne1 | ‘owl monkey's tree trunk’\nˈtɨ2 ʔe1 a1 ne1 | ‘cassava garden’\nˈto̰1 ʔtʃi5 ru1 | ‘owl monkey's clothes’\nˈto1 ʔo̰1 | ‘other one's mouth’\nˈto̰1 ʔo5 tʃi1 | ‘really an owl monkey’\nˈtʃau1 ʔtʃi5 ru1 | ‘my clothes’\nˈtʃo1 ʔma̰1 ne1 | ‘my wife's tree trunk’\nˈtʃo1 me4 na2 ʔã2 | ‘my stick’\nˈto1 bɨ2 | ‘other one's high-starch food’\nˈtʃo1 pa3 tɨ4 | ‘my fingernail’\nˈtʃau1 e3 ɟa̰1 te4 | ‘my sister's husband’\nna4 ˈtʃḭ1 bɨ2 | ‘its high-starch food is delicious’\nna4 ˈtʃo5 o1 ne1 ʔɨ1 ra1 | ‘its garden is sort of white’\nˈŋo3 ʔo̰1 a1 ne1 | ‘place where there are lots of ŋo3 ʔo̰1’\n\ntʃ, ɟ, ŋ, and ʔ are consonants; ɨ is a vowel; au is a diphthong: consider it as one vowel. The mark ˈ indicates that the following syllable is stressed.\n\"25CC1, \"25CC2, \"25CC3, \"25CC4, \"25CC5, and \"25CC43 denote tones of the preceding syllable. Pitches of the tones:\n\nlow = \"25CC1 < \"25CC2 < \"25CC3 < \"25CC4 < \"25CC5 = high; \"25CC43 = \"25CC4 \"25CC3\n\nA tilde below a vowel (e.g., a̰) denotes creaky voice (a type of phonation that is often perceived as low-pitched and “rough”). A tilde over a vowel (e.g., ã) denotes a nasal sound.\n\nAn ‘owl monkey’ is a type of monkey. ‘Cassava’ is a woody plant native to South America. A ‘ka̰1 tree’ is a kind of fruit tree. ‘ŋo3 ʔo̰1’ is a kind of fish.","query":"- What is the literal translation of ˈŋo3 ʔo̰1 a1 ne1?\n\n- Translate into English:\n\n- ˈka5 ne1\n\n- na4 ˈtʃo̰1 o5 tɨ3\n\n- ˈŋo3 ʔo̰1 ʔɨ5 tʃi1\n\n- ˈto1 o1 ne1\n\n- ˈto̰1 ʔo4 ne1\n\n- ˈtʃau1 ne1\n\n- Translate into Cushillococa Ticuna:\n\n- ‘it is sort of delicious’\n\n- ‘its clothes are really white’\n\n- ‘my husband's entire body’\n\n- ‘my high-starch food’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"4.14. Cushillococa Ticuna\n\n- ‘ŋo3 ʔo̰1 ('s) garden’\n\n-\n\n- ‘ka̰1 tree trunk’\n\n- ‘its leaves are white’\n\n- ‘really a ŋo3 ʔo̰1’\n\n- ‘other one's garden’\n\n- ‘owl monkey's entire body’\n\n- ‘my tree trunk’\n\n-\n\n- na4 ˈtʃi5 ʔi1 ra1\n\n- na4 ˈtʃo̰1 ʔtʃi5 ru1 ʔɨ5 tʃi1\n\n- ˈtʃau1 te4 ʔɨ4 ne1\n\n- ˈtʃo1 bɨ2\n\nRules:\n\n- The possessive and the adjective are placed before the noun.\n\n- The possessive is marked by: to1 (‘other one's’), ku43 (‘your’), tʃau1 (‘his’).\n\n- tʃau1 becomes tʃo1, if it is before a bilabial consonant (p, b, m). The glottal stop does not block the transformation. Thus:\n\ntʃau1 tʃo1 / _(ʔ)Cbilabial\n\n- The stress falls on the first syllable. If the first syllable is na4 (‘it is’), the stress shifts onto the second syllable.\n\n- Phonological processes:\n\n- a o / ˈo(ʔ) _ (a becomes o if it follows a stressed syllable that contains o, if there is no consonant between them (except for glottal stop));\n\n- ɨ V / ˈ V(ʔ) _ (ɨ fully assimilates to the preceding vowel if this vowel is in a stressed syllable and between them there is no other consonant (except for glottal stop));\n\n- ˈ V̰1 ˈ V5 / _ (C)V1 (tone dissimilation – a pharyngealised vowel with tone 1 in a stressed syllable will become non-pharyngealised and with tone 5 if it is before another syllable with tone 1).\n\n*Supplementary: Dictionary of base forms\n\n*Adjectives\n\n- bu3 = ‘immature’\n\n- me43 = ‘good’\n\n- tʃḭ1 = ‘delicious’\n\n- tʃo̰1 = ‘white’\n\n*Nouns\n\n- a1 ne1 = ‘garden’\n\n- a5 tɨ3 = ‘leaves’\n\n- ʔa̰1 = ‘mouth’\n\n- bɨ2 = ‘high-starch food’\n\n- e3 ɟa̰1
{"id":"book_05_01","context":"Here are some words in Zulu and their English translations:\n\numdwebi | ‘painter’ | abazingeli | ‘hunters’ | umbulali | ‘killer’\nabadwebi | ‘painters’ | zingela | ‘to hunt’ | ababazi | ‘carvers’","query":"- Translate into Zulu: ‘to paint’, ‘hunter’, ‘killers’, ‘to kill’, ‘carver’, ‘to carve’.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"We notice that we have three types of words: singular nouns, plural nouns and verbs which are semantically related to the nouns. Therefore, we can create the following table:\n\n| Noun sg | Noun pl | Verb\n‘painter’ | umdwebi | abadwebi |\n‘hunter’ | | abazingeli | zingela\n‘killer’ | umbulali | |\n‘carver’ | | ababazi |\n\nFrom the table, we can easily deduce that the singular noun is formed using the circumfix um- -i, while the plural is formed with the circumfix aba- -i. The verb is formed by adding the suffix -a.\n\nAnother explanation, in order to avoid the idea of a circumfix, is that the singular and plural are formed using the prefixes um- and aba-, respectively, while the verb is formed by replacing the final vowel (i) with the vowel a.\n\nThus, the answers are:\n\n-\n| Noun sg | Noun pl | Verb\n‘painter’ | umdwebi | abadwebi | dweba\n‘hunter’ | umzingeli | abazingeli | zingela\n‘killer’ | umbulali | ababulali | bulala\n‘carver’ | umbazi | ababazi | baza","source":"langsci_420","problem_group_id":"langsci420:5.1","chapter":5,"chapter_title":"Noun and noun phrase","section":2,"section_title":"Basic principle of morphological analysis","topic":"noun morphology and noun phrases","language":"Zulu","author":"Vlad A. Neacșu","competition":"original","year":null,"solution_tier":"detailed_worked","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s02_p01"],"method_link_type":"author_placement","method_description":"The author places this worked example under “Basic principle of morphological analysis” in the Noun and noun phrase chapter, so it illustrates that method or topic.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":107,"source_line_end":121,"solution_line_start":122,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun phrase\nBasic principle of morphological analysis\nnoun morphology and noun phrases\nThe author places this worked example under “Basic principle of morphological analysis” in the Noun and noun phrase chapter, so it illustrates that method or topic.\nHere are some words in Zulu and their English translations:\n\numdwebi | ‘painter’ | abazingeli | ‘hunters’ | umbulali | ‘killer’\nabadwebi | ‘painters’ | zingela | ‘to hunt’ | ababazi | ‘carvers’\n- Translate into Zulu: ‘to paint’, ‘hunter’, ‘killers’, ‘to kill’, ‘carver’, ‘to carve’."}
{"id":"book_05_02","context":"Here are some words in Swedish and their English translations:\n\nen flaska | ‘a bottle’ | hunden | ‘the dog’ | hyllor | ‘shelves’\nen stol | ‘a chair’ | flaskorna | ‘the bottles’ | kattar | ‘cats’\nen hund | ‘a dog’ | stolarna | ‘the chairs’ | bilen | ‘the car’\nflaskor | ‘bottles’ | hundarna | ‘the dogs’ | hyllan | ‘the shelf’\nstolar | ‘chairs’ | en bil | ‘a car’ | katten | ‘the cat’\nhundar | ‘dogs’ | en hylla | ‘a shelf’ | bilarna | ‘the cars’\nflaskan | ‘the bottle’ | en katt | ‘a cat’ | hyllorna | ‘the shelves’\nstolen | ‘the chair’ | bilar | ‘cars’ | kattarna | ‘the cats’","query":"- Here are some more words in Swedish and their English translations:\nen flicka = ‘a girl’ bussarna = ‘the buses’\n\n- Translate into Swedish: ‘the girl’, ‘girls’, ‘the girls’, ‘a bus’, ‘the bus’, ‘buses’.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"Just as in the previous problem, we first notice the forms of the given nouns. We realise that each noun is given in four different forms: definite singular, indefinite singular, definite plural, and indefinite plural. Therefore, in order to facilitate the analysis of the data and notice the similarities between them, we make the following table:\n\nIndef. sg | Def. sg | Indef. pl | Def. pl | Translation\nen flaska | flaskan | flaskor | flaskorna | ‘bottle’\nen stol | stolen | stolar | stolarna | ‘chair’\nen hund | hunden | hundar | hundarna | ‘dog’\nen bil | bilen | bilar | bilarna | ‘car’\nen hylla | hyllan | hyllor | hyllorna | ‘shelf’\nen katt | katten | kattar | kattarna | ‘cat’\n\nBased on this table, we can consider the indefinite singular form as the stem (excluding the procliticThe term proclitic refers to the fact that the definite article is placed in front of the noun. This contrasts with the term enclitic, meaning it is placed after the noun. article en). Moreover, we notice that the definite plural is derived from the indefinite plural by adding the suffix -na. We are left to discover how to form the definite singular and the indefinite plural.\n\nWe notice that the def. sg is formed using the suffixes -n or -en. Therefore, we need to discover in which context each of them is used. It could be either a semantic context, in which the variation is driven by the meaning of the word, or a phonological context, in which the choice of the allomorph is dictated by the phonological structure of the word. In this case, we can easily notice that the suffix -en is used if the base form ends in a consonant, while -n is used if the base form ends in a vowel, thus the distinction is purely phonological. Another explanation for the definite singular endings can include a phonological process, an elision, by considering the suffix -en as the sole suffix for def. sg, with the additional feature that en n / V _.\n\nApplying the same thought process for the indefinite plural, we notice that the two suffixes are -ar and -or, the first being used if the stem ends in a consonant and the latter if the stem ends in a vowel. Moreover, if the stem ends in a vowel, the vowel is dropped when the suffix -or is added (alternatively: the suffix is -ar and -V + -ar -or – i.e., when the suffix -ar is added after a vowel, it merges with it and results in the suffix -or). Thus, we can write the rules that govern the formation of these noun forms in Swedish (we used the abbreviations S = stem, C = consonant, V = vowel):\n\nIndef. sg | Def. sg | Indef. pl | Def. pl\nen S-C | S-C-en | S-C-ar | S-C-arna\nen S-V | S-V-n | S-or | S-orna\n\nThus, the answers to the task are:\n\n-\n\n- ‘the girl’ = flickan\n\n- ‘girls’ = flickor\n\n- ‘the girls’ = flickorna\n\n- ‘a bus’ = en buss\n\n- ‘the bus’ = bussen\n\n- ‘buses’ = bussar","source":"langsci_420","problem_group_id":"langsci420:5.2","chapter":5,"chapter_title":"Noun and nou
{"id":"book_05_03","context":"Consider the following word forms in Māori:\n\nForm I | Form II | Form III\ninu | inumia | inumaŋa\nhopu | hopukia | hopukaŋa\neke | ekeŋia | ekeŋaŋa\nɸera | ɸerahia | ɸerahaŋa\naɸi | aɸitia | aɸitaŋa\ntupu | tupuria | tupuraŋa","query":"- Explain how the forms are constructed.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"We can easily observe that, in order to obtain Form III from Form II we just replace the suffix -ia with -aŋa. Moreover, we notice that Form II derives from Form I by adding the suffix -Cia, where C is a consonant. The only thing left to do is figure out how the consonant is chosen.\n\nThe first thing we notice is that there are no two examples which use the same consonant; furthermore, the consonant does not seem to be related in any way to the structure of the word (to the phonological characteristics of the other sounds). Finally, we are sure that the choice of the consonant cannot be related to the meaning since the translations are not given in the data. Therefore, since the consonant does not seem to follow any pattern, we can consider it as being part of the stem, thus being a disfix.\n\nIf we consider that consonant as a disfix, we can easily figure out all the rules that generate the three forms:\n\n-\n\n- Form I: elision of the last consonant of the stem;\n\n- Form II: add suffix -ia;\n\n- Form III: add suffix -aŋa.","source":"langsci_420","problem_group_id":"langsci420:5.3","chapter":5,"chapter_title":"Noun and noun phrase","section":2,"section_title":"Basic principle of morphological analysis","topic":"noun morphology and noun phrases","language":"Māori","author":"Vlad A. Neacșu","competition":"original","year":null,"solution_tier":"detailed_worked","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s02_p01"],"method_link_type":"author_placement","method_description":"The author places this worked example under “Basic principle of morphological analysis” in the Noun and noun phrase chapter, so it illustrates that method or topic.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":235,"source_line_end":255,"solution_line_start":256,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun phrase\nBasic principle of morphological analysis\nnoun morphology and noun phrases\nThe author places this worked example under “Basic principle of morphological analysis” in the Noun and noun phrase chapter, so it illustrates that method or topic.\nConsider the following word forms in Māori:\n\nForm I | Form II | Form III\ninu | inumia | inumaŋa\nhopu | hopukia | hopukaŋa\neke | ekeŋia | ekeŋaŋa\nɸera | ɸerahia | ɸerahaŋa\naɸi | aɸitia | aɸitaŋa\ntupu | tupuria | tupuraŋa\n- Explain how the forms are constructed."}
{"id":"book_05_04","context":"Here are some sentences in Bulgarian (written in Latin script) and their English translations in random order:\n\n. | Veshterǎt nahrani maymunata. | . | ‘Your son watched you.’\n. | Kamilata vǎrvya. | . | ‘The girl hugged the cat.’\n. | Momicheto pregǎrna kotkata. | . | ‘You dressed yourself.’\n. | Veshtitsata prokle kotkata. | . | ‘The cat scratched you.’\n. | Kotkata prokle tvoya sin. | . | ‘You fed the son.’\n. | Ti nahrani sina. | . | ‘The witch cursed the cat.’\n. | Kotkata te odraska. | . | ‘The camel walked.’\n. | Ti skochi. | . | ‘The cat cursed your son.’\n. | Tvoyat sin te gleda. | . | ‘The wizard fed the monkey.’\n. | Veshterǎt pregǎrna edna kamila. | . | ‘The son dressed your baby.’\n. | Ti se obleche. | . | ‘You jumped.’\n. | Sinǎt obleche tvoeto bebe. | . | ‘The wizard hugged a camel.’\n\nǎ ‘u’ in ‘but’.","query":"- Determine the correct correspondences.\n\n- Translate into English:\n\n- Maymunata gleda tvoyata veshtitsa.\n\n- Tvoyata kamila obleche edno momiche.\n\n- Veshterǎt se prokle.\n\n- Ti pregǎrna bebeto.\n\n- Ti vǎrvya.\n\n- Ti prokle edin veshter.\n\n- Translate into Bulgarian:\n\n- ‘The witch dressed you.’\n\n- ‘The baby watched the girl.’\n\n- ‘The monkey jumped.’\n\n- ‘You hugged a son.’\n\n- ‘Your son dressed a baby.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"There are multiple possible starting points when approaching this problem, but one of the most common (and generally applicable) is to notice that there are only two sentences in Bulgarian which have two words (2 and 8). Therefore, we can assume that these are the simplest sentences and, most likely, correspond to the shortest sentences in English, i.e., those that have only a subject and a verb. Among all the English sentences, the only ones that follow this pattern are K and G. We have different ways to figure out which is which: we can use the similarity between the Bulgarian and English words (kamilata – ‘camel’) or we can notice that ti appears in two other sentences and kamilata does not occur in any other sentence, while in English we have got two other sentences that begin with ‘you’, but none that begin with ‘camel’. Since neither of the two verbs ever occurs again in the data and the first word does, we discover that this one is the subject and the last word is the verb. Hence, we deduce that 2-G and 8-K and the word order is S-V (Subject-Verb).\n\nFrom here, we can continue with the other two sentences which contain the subject ti = ‘you’, which are 6, 11 and C, E. In order to match them, we can notice that the last word in sentence 6 occurs (in a slightly changed form) as the first word of sentence 12, so we can deduce it is a noun (since it is a subject in sentence 12), thus it cannot mean ‘myself’. Therefore, 6-E and 11-C. Moreover, we understand that sina = ‘son’, nahrani = ‘to feed’, obleche = ‘to dress’. Since obleche and nahrani occur one more time in another sentence, it follows that 1-I and 12-J (we can notice again the similar words bebe – ‘baby’). Based on this information, we can easily match all the other sentences. We get:\n\n. | #1 | #2. | ‘#3’\n\n. | Veshterǎt nahrani maymunata. | . | ‘I’The wizard fed the monkey.\n. | Kamilata vǎrvya. | . | ‘G’The camel walked.\n. | Momicheto pregǎrna kotkata. | . | ‘B’The girl hugged the cat.\n. | Veshtitsata prokle kotkata. | . | ‘F’The witch cursed the cat.\n. | Kotkata prokle tvoya sin. | . | ‘H’The cat cursed your son.\n. | Ti nahrani sina. | . | ‘E’You fed the son.\n. | Kotkata te odraska. | . | ‘D’The cat scratched you.\n. | Ti skochi. | . | ‘K’You jumped.\n. | Tvoyat sin te gleda. | . | ‘A’Your son watched you.\n. | Veshterǎt pregǎrna edna kamila. | . | ‘L’The wizard hugged a camel.\n. | Ti se obleche. | . | ‘C’You dressed yourself.\n. | Sinǎt obleche tvoeto bebe. | . | <20><>
{"id":"book_05_05","context":"Here are some phrases in Japanese and their English translations:\n\nisha kyūnin | ‘9 doctors’\ngakusei sannin | ‘3 students’\nhon yonsatsu | ‘4 books’\ninu kyūhiki | ‘9 dogs’\nkami hachimai | ‘8 sheets of paper’\nmagajin nanasatsu | ‘7 magazines’\nneko nihiki | ‘2 cats’\npurēto yonmai | ‘4 plates’\nratto gohiki | ‘5 rats’\numa rokutō | ‘6 horses’\nzō rokutō | ‘6 elephants’","query":"- Translate into English:purēto rokumai, isha gonin, uma yontō.\n\n- Here are some more Japanese words:\n\nmangabon | ‘comic books’ | piza | ‘pizzas’\nkaeru | ‘frogs’ | ushi | ‘cows’\n\n- Translate into Japanese: ‘2 comic books’, ‘5 pizzas’, ‘7 frogs’, ‘9 cows’.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"Comparing examples ‘9 doctors’ and ‘9 dogs’, we notice that the only part that is repeated is the morpheme kyū- from the beginning of the second word. We can deduce that the number is represented by the first part of the second word. Moreover, comparing the last two examples, (‘6 horses’, ‘6 elephants’), we notice that only the first word is different, so we can assume that this one represents the noun. Therefore, the phrase structure is Noun Number-X.\n\nBased on this structure, we can infer all the numbers and nouns in Japanese, as follows:\n\nJapanese numbers are:\n\n- ni-\n\n- san-\n\n- yon-\n\n- go-\n\n- roku-\n\n- nana-\n\n- hachi-\n\n- kyū-\n\n-\n\nNote that in order to figure out the morpheme for ‘6’, we need to check the examples in task (a) in order to figure out which part corresponds to the number and which to the particle X.\n\nNouns are:\n\n- isha = ‘doctor’\n\n- hon = ‘book’\n\n- kami = ‘sheet of paper’\n\n- neko = ‘cat’\n\n- ratto = ‘rat’\n\n- zō = ‘elephant’\n\n- gakusei = ‘student’\n\n- inu = ‘dog’\n\n- magajin = ‘magazine’\n\n- purēto = ‘plate’\n\n- uma = ‘horse’\n\n-\n\nThe only morpheme left to analyse is X. We notice that it can have five different forms: -nin (in the phrases ‘9 doctors’, ‘5 students’), -satsu (‘4 books’, ‘7 magazines’), -mai (‘8 sheets of paper’, ‘4 plates’), -hiki (‘9 dogs’, ‘2 cats’, ‘5 rats’), and -tō (‘6 horses’, ‘6 elephants’). Therefore, we deduce that they represent classifiers and correspond to: -nin for humans, -satsu for bound materials, -mai for flat objects, -hiki for small animals and -tō for large animals. Now we can write the rules and answer the tasks.\n\nRules:\n\n- Structure: Noun Number–Class\n\n- Class:\n\n- -nin humans\n\n- -satsu prints\n\n- -mai flat objects\n\n- -hiki small animals\n\n- -tō large animals\n\n-\n\n-\n\n- purēto rokumai = ‘6 plates’\n\n- isha gonin = ‘5 doctors’\n\n- uma yontō = ‘4 horses’\n\n-\n\n- ‘2 comic books’ = mangabon nisatsu\n\n- ‘5 pizzas’ = piza gomai\n\n- ‘7 frogs’ = kaeru nanahiki\n\n- ‘9 cows’ = ushi kyūtō","source":"langsci_420","problem_group_id":"langsci420:5.5","chapter":5,"chapter_title":"Noun and noun phrase","section":4,"section_title":"Classifiers","topic":"noun morphology and noun phrases","language":"Japanese","author":"Vlad A. Neacșu","competition":"RoLO","year":2017,"solution_tier":"detailed_worked","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s04_p01"],"method_link_type":"author_placement","method_description":"The author places this worked example under “Classifiers” in the Noun and noun phrase chapter, so it illustrates that method or topic.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":477,"source_line_end":507,"solution_line_start":508,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun phrase\nClassifiers\nnoun morphology and noun phrases\nThe
{"id":"book_05_06","context":"Here are some phrases in Fijian and their English translations:\n\n. | na uluqu | ‘my head’\n. | na nona wau | ‘her weapon (she owns)’\n. | na memunī bia | ‘yourpl beer’\n. | na kemudrau itukutuku | ‘yourdu story (about you two)’\n. | na nona motokaa | ‘her car’\n. | na meda tī | ‘ourincl tea’\n. | na kelemu | ‘yoursg belly’\n. | na nona dio | ‘her oyster (she'll sell)’\n. | na kequ uvi | ‘my yam’\n. | na noqu itukutuku | ‘my story (I tell)’\n. | na watiqu | ‘my spouse’\n. | na kemunī vuaka | ‘yourpl pig (you'll eat)’\n. | na nomu kato | ‘yoursg basket’\n. | na tamana | ‘her father’\n. | na memudrau dio | ‘yourdu oyster (you'll slurp)’\n. | na nodra vuaka | ‘their pig (they raise)’\n. | na keda wau | ‘ourincl weapon (we'll be hit with)’\n. | na kedra raisi | ‘their rice’\n\n‘Weapon’ refers to a club-like tool. A ‘yam’ is an edible starchy root. ‘Kava’ is a ceremonial drink.\n\nThe subscripts sg, du, pl refer to singular (one person), dual (two persons), and plural (more than two persons), respectively. The subscript incl means inclusive (‘ourincl’ = belonging to me, you, and them, in contrast with ‘ourexcl’ = belonging to me and them, but not you).","query":"- Now the Fijian words are given to you. Your task is to translate the phrase in the table into Fijian:\n\n| Fijian | English | English phrase to translate\n\n. | uto | ‘heart’ | ‘my heart’\n. | yaqona | ‘kava’ | ‘her kava (she's drinking)’\n. | yaqona | ‘kava’ | ‘her kava (drunk in her honour)’\n. | draunikau | ‘witchcraft’ | ‘my witchcraft (used on/against me)’\n. | draunikau | ‘witchcraft’ | ‘yourdu witchcraft (you're making)’\n. | dali | ‘rope’ | ‘yoursg rope (you own)’\n. | dali | ‘rope’ | ‘yourpl rope (restraining you)’\n. | ika | ‘fish’ | ‘yourdu fish’\n. | wai | ‘water’ | ‘yourpl water’\n. | luve | ‘child’ | ‘her child’\n. | waqa | ‘canoe’ | ‘ourincl canoe’\n. | yapolo | ‘apple’ | ‘their apple (they'll sell)’\n. | maqo | ‘mango’ | ‘their mango (for drinking)’\n\n- Explain your translation 21. (Why did you translate it this way?)\n\n- The word for ‘coconut’ is niu. List all the ways to say ‘my coconut’ and explain what they could mean.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"The first step we need to take is to determine the structure of the Fijian phrases. We can easily notice that all examples start with the word na followed by either one or two words. Separating the structures that contain a single word (i.e., the Fijian words for ‘my head’, ‘yoursg belly’, ‘my husband’, ‘her father’), we notice that all of them represent inalienable possessions. We therefore expect them to have the possessor marked directly on the noun, as an affix.\n\nIndeed, comparing the structures ‘my head’ = na uluqu and ‘my husband’ = na watiqu, we discover that they both share the suffix -qu, which we then can deduce to mean 1sg possession. Therefore, for the inalienable possession, the phrase structure is na Noun–Poss.\n\nIn order to determine the structure of the rest of the phrases, we compare examples that contain the same noun (for example, 8 and 15, both containing the noun which means ‘oyster’). We notice that both of them have the last word in common (dio), so the last word represents the noun, the possessed. Moreover, comparing examples 9 and 10 (both containing the possessive ‘my’), we notice that they have in common the suffix -qu attached to the second word. Consequently, we can deduce that the structure of the alienable possession in Fijian is na X–Poss Noun. From here, we can easily determine the possession suffixes (applying similar processes as in Problem 5.5 where we had to separate the number from the classifier). We obtain:\n\n- 1sg = -qu\n\n- 2sg = -mu\n\n- 3sg = -na\n\n- 1pl incl. = -da\n\n- 2du = -mu
{"id":"book_05_07","context":"Here are some phrases in Ancient Greek (in a Latin-based transcription) and their English translations in random order:\n\n. | ho tōn hyiōn dulos | . | ‘the donkey of the master’\n. | hoi tōn dulōn cyrioi | . | ‘the brothers of the merchant’\n. | hoi tu emporu adelphoi | . | ‘the merchants of the donkeys’\n. | hoi tōn onōn emporoi | . | ‘the sons of the masters’\n. | ho tu cyriu onos | . | ‘the slave of the sons’\n. | ho tu oicu cyrios | . | ‘the masters of the slaves’\n. | ho tōn adelphōn oicos | . | ‘the house of the brothers’\n. | hoi tōn cyriōn hyioi | . | ‘the master of the house’\n\n-\nō denotes a long o.","query":"- Determine the correct correspondences.\n\n- Translate into Ancient Greek: ‘the houses of the merchants’, ‘the donkeys of the slave’.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"- Step 1. We notice that all Ancient Greek phrases have the structure [ho/hoi] [tu/tōn] X Y, where X and Y represent the two nouns (the possessor and the possessed). Moreover, we notice that these nouns change their form. Furthermore, the first two words (which probably represent articles or possession markers) each have two forms. Checking the English translations, we notice that nouns appear both as singular and plural. Therefore, we can assume that the two markers from the beginning of the phrase change form, agreeing with the number of the noun.\n\n- Step 2. We make a frequency table based on the stem of the nouns (for now, we disregard the endings, which are variable).\n\nGreek | Freq. | English | Freq.\n(lr)1-2 (lr)3-4\nhyi- | 2 | ‘donkey’ | 2\ndul- | 2 | ‘master’ | 4\ncyri- | 4 | ‘brother’ | 2\nempor- | 2 | ‘merchant’ | 2\nadelph- | 2 | ‘son’ | 2\non- | 2 | ‘slave’ | 2\noic- | 2 | ‘house’ | 2\n\nSince both in Greek and in English we have only one noun that appears four times (all the rest appearing only two times each), we infer that cyri- = ‘master’.\n\n- Step 3. We separate all phrases which contain the noun master:\n\n2. | hoi tōn dulōn cyrioi | Indent | A. | ‘the donkey of the master’\n5. | ho tu cyriu onos | | D. | ‘the sons of the masters’\n6. | ho tu oicu cyrios | | F. | ‘the masters of the slaves’\n8. | hoi tōn cyriōn hyioi | | H. | ‘the master of the house’\n\nIn Greek, the nouns that appear together with ‘master’ are dul-, on-, oic-, hyi-, while in English they are ‘donkey’, ‘son’, ‘slave’, ‘house’. The only two nouns that do not appear here are ‘merchant’ and ‘brother’, so these must correspond to the two Greek nouns that do not occur: empor- and adelph-. Moreover, we notice that phrase 3 in Greek and phrase B in English contain both of these nouns. Therefore, we deduce the correspondence 3-B.\n\n- Step 4. We separate the sentences that contain the nouns ‘brother’ or ‘merchant’.\n\n3. | hoi tu emporu adelphoi | Ind | | ‘the brothers of the merchant’\n4. | hoi tōn onōn emporoi | | C. | ‘the merchants of the donkeys’\n7. | ho tōn adelphōn oicos | | G. | ‘the house of the brothers’\n\nNotebulbonThe fact that we didn't mention the letter in front of structure B. means that this phrase is already matched, its Greek correspondence being phrase 3.\n\nSeparating again the nouns which appear in these structures, we deduce that ‘donkey’ and ‘house’ are on- and oic- though we do not know which is which, and similarly with other word pairs.\n\n- Step 5. Recap.\n\nBased on the information above, we know that:\n\n- cyri- = ‘master’\n\n- {on-/oic-} = {‘donkey’/‘house’}\n\n- {empor-/adelph-} = {‘merchant’/‘brother’}\n\n- {dul-/hyi-} = {‘son’/‘slave’}.\n\nMoreover, we notice that there is one phrase in which the nouns ‘slave’ and ‘son’ co-occur, therefore 1-E.\n\n- Step 6. Based on this information, we split the phrases into subgroups:\n\n1. | ho tōn hyiōn dulos | | ‘the slave of the sons’\n3. | hoi tu empo
{"id":"book_05_08","context":"Brent Berlin and Paul Kay have studied colour terms of more than 100 languages by travelling the world and asking the locals to describe photos of different colours (BerlinandKay1969). After examining the results, they concluded that colour perception in all of these languages is governed by the same law.\n\nHere are the colour terms in 12 of the languages the two scientists examined:Source: Adapted from a problem by Ksenia Gilyarova (Elementy).\n\n| Fitzroy | | Upper | | |\nEnglish | River | Nupe | Pyramid | Ibo | Tzeltal | Hanunó'o\n‘white’ | bura | bókùṇ | [blank] | nzu | sak | (ma)biru\n‘blue’ | guru | dòfa | [blank] | [blank] | yaš | [blank]\n‘yellow’ | kalmur | wọṇjiṇ | [blank] | odo | k'an | (ma)raraʔ\n‘brown’ | [blank] | dzúfú | mola | uhie | [blank] | (ma)raraʔ\n‘black’ | [blank] | ẓìkò | [blank] | oji | ʔihk' | (ma)lagtiʔ\n‘red’ | kiran | dzúfú | [blank] | [blank] | cah | [blank]\n‘green’ | [blank] | álígà | muli | oji | yaš | (ma)latuy\n\nEnglish | Bari | Jalé | Hausa | Nasioi | Daza | Ibibio\n‘white’ | -kwe | hóló | fări | kakara | cuo | àfíá\n‘blue’ | -murye | siŋ | shuḍi | mutaŋa | zẹdẹ | [blank]\n‘yellow’ | -forong | [blank] | nawaya | [blank] | mini | ńdàídàt\n‘brown’ | -jere | [blank] | ja | [blank] | maaḍo | [blank]\n‘black’ | -rnö | siŋ | bāḳi | [blank] | yasko | έbubit\n‘red’ | -tor | hóló | [blank] | erereŋ | [blank] | ńdàídàt\n‘green’ | -ngem | [blank] | algashi | [blank] | [blank] | àwàwà","query":"- Fill in the blanks.\n\n- Below are some colour terms in three other languages that follow the hypothesis of Berlin & Kay:\n\n- Urhobo: ‘black’ = ɔbyibi, ‘brown’ = ɔBaBare, ‘green’ = ɔbyibi, ‘yellow’ = 5do;\n\n- N'gombe: ‘white’ = bopu, ‘red’ = bopu;\n\n- Tanna Island: ‘yellow’ = laulau, ‘black’ = rapen, ‘brown’ = laulau, ‘blue’ = ramimera.\n\n- For each language, specify which of the 12 languages above it most resembles. Explain your answer.\n\n- Here are some colour terms in two more languages:\n\n- Acehnese: ‘white’ = iĵu, ‘red’ = pirã, ‘green’ = prãna, ‘yellow’ = iĵu, ‘blue’ = prãna;\n\n- Alabama: ‘green’ = okchakko, ‘red’ = homma, ‘brown’ = laana, ‘blue’ = okchakko, ‘yellow’ = laana.\n\n- Explain why they do not follow the hypothesis of Berlin & Kay.\n\n- Formulate the hypothesis of Berlin & Kay.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"- Step 1. Looking closely at the table, we notice that almost all the languages in the table (save for English and Bari) have some colour terms which share the same name. Therefore, we deduce that, in order to fill in the blanks, we need to reuse some of the words in that language which are already given; in other words, we need to figure out which colour terms are translated identically in that language. This is also signalled by the fact that there does not seem to be any morphological process going on which allows us to derive some colour terms from others and there are few or no common morphemes.\n\nIn order to do so, we classify each language based on the number of distinct colour terms (i.e., colour terms which have different names in that language) and reorder the table in descending order based on the number of distinct colour terms.\n\n| 7 terms | 6 terms | 5 terms\n(lr)2-2(lr)3-4(lr)5-6\nEnglish | Bari | Nupe | Hausa | Tzeltal | Daza\n‘white’ | -kwe | bókùṇ | fări | sak | cuo\n‘blue’ | -murye | dòfa | shuḍi | yaš | zẹdẹ\n‘yellow’ | -forong | wọṇjiṇ | nawaya | k'an | mini\n‘brown’ | -jere | dzúfú | ja | | maaḍo\n‘black’ | -rnö | ẓìkò | bāḳi | ʔihk' | yasko\n‘red’ | -tor | dzúfú | | cah |\n‘green’ | -ngem | álígà | algashi | yaš |\n\n| 4 terms\n(lr)2-5\nEnglish | Fitzroy River | Ibo | Hanunó'o | Ibibio\n‘white’ | bura | nzu | (ma)biru | àfíá\n‘blue’ | gu
{"id":"book_05_09","context":"Here are some words in Ulwa and their English translations in random order:\n\nsuulu, suukilu, suumanalu, mismatu, miskatu, onkinayan, onkayan, onyan\n\n‘bow’, ‘yoursg cat’, ‘my dog’, ‘our bow’, ‘his cat’, ‘dog’, ‘his bow’, ‘yourpl dog’","query":"- Determine the correct correspondences.\n\n- Translate into English: suumalu and miskanatu.\n\n- Translate into Ulwa: ‘cat’, ‘my cat’, ‘yoursg bow’, ‘their bow’.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"5.9. Ulwa\n\n-\n\n- suulu = ‘dog’\n\n- suukilu = ‘my dog’\n\n- suumanalu = ‘yourpl dog’\n\n- onyan = ‘bow’\n\n- miskatu = ‘his cat’\n\n- onkinayan = ‘our bow’\n\n- mismatu = ‘yoursg cat’\n\n- onkayan = ‘his bow’\n\n-\n\n- suumalu = ‘yoursg dog’\n\n- miskanatu = ‘their cat’\n\n-\n\n- ‘cat’ = mistu\n\n- ‘my cat’ = miskitu\n\n- ‘yoursg bow’ = onmayan\n\n- ‘their bow’ = onkanayan\n\nRules:\n\n- The possessive is marked by an infix placed before the last syllable. The structure of the infix is Person–Number.\n\n- Person: -ki- = 1 i -ma- = 2 -ka- = 3\n\n- Number: = sg -na- = pl","source":"langsci_420","problem_group_id":"langsci420:5.9","chapter":5,"chapter_title":"Noun and noun phrase","section":9,"section_title":"Practice problems","topic":"noun morphology and noun phrases","language":"Ulwa","author":"Peter Arkadiev","competition":"TurLom","year":2003,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s01_p01","book_method_c05_s01_p02","book_method_c05_s02_p01","book_method_c05_s03_p01","book_method_c05_s03_p02","book_method_c05_s04_p01","book_method_c05_s05_p01","book_method_c05_s06_p01","book_method_c05_s07_p01","book_method_c05_s08_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":1247,"source_line_end":1263,"solution_line_start":1651,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun phrase\nPractice problems\nnoun morphology and noun phrases\nThis practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words in Ulwa and their English translations in random order:\n\nsuulu, suukilu, suumanalu, mismatu, miskatu, onkinayan, onkayan, onyan\n\n‘bow’, ‘yoursg cat’, ‘my dog’, ‘our bow’, ‘his cat’, ‘dog’, ‘his bow’, ‘yourpl dog’\n- Determine the correct correspondences.\n\n- Translate into English: suumalu and miskanatu.\n\n- Translate into Ulwa: ‘cat’, ‘my cat’, ‘yoursg bow’, ‘their bow’."}
{"id":"book_05_10","context":"Here are some phrases in Palauan and their English translations:\n\neru ęl buil | ‘2 months’ | kltiu ęl hong | ‘9 books’\n\nede ęl sils | ‘3 days’ | kllolem ęl lius | ‘6 coconuts’\n\ntede ęl chad | ‘3 people’ | teai ęl ngalęk | ‘8 children’\n\nkllolem ęl malk | ‘6 chickens’ | ongeru ęl buil | ‘February’\n\nteim ęl sensei | ‘5 teachers’ | ongede ęl ureor | ‘Wednesday’\n\neim ęl rak | ‘5 years’ | etiu ęl klębęse | ‘9 nights’\n\ntęruich me a tede ęl buik | ‘13 boys’\n\ntęruich me a euid ęl sikang | ‘17 hours’","query":"- Translate into English:\n\n- telolem ęl sensei\n\n- tęruich me a etiu ęl buil\n\n- tęruich me a ongeru ęl buil\n\n- ongeim ęl ureor\n\n- Translate into Palauan:\n\n- ‘8 days’\n\n- ‘19 people’\n\n- ‘7 teachers’\n\n- ‘June’\n\n- ‘August’\n\n- For each of the following, write the Palauan word that would be used to translate the word ‘3’:\n\n- ‘3 hours’\n\n- ‘3 girls’\n\n- ‘3 dolphins’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"5.10. Palauan\n\n-\n\n- ‘6 teachers’\n\n- ‘19 months’\n\n- ‘December’\n\n- ‘Friday’\n\n-\n\n- eai e̜l sils\n\n- te̜ruich me a tetiu e̜l chad\n\n- teuid e̜l sensei\n\n- ongelolem e̜l buil\n\n- ongeai e̜l buil\n\n-\n\n-\n\n- ede\n\n- tede\n\n- klde\n\nRules:\n\n- Structure: Numeral + e̜l + Noun\n\n- Numerals have the following prefixes:\n\n- e- = time periods (days, months, years)\n\n- te- = people\n\n- kl- = non-human nouns (which do not refer to persons)\n\n- onge- = ordinal numerals\n\n- For numbers higher than 10, the prefix is attached to the units.\n\n- 10 + X = tęruich me a X","source":"langsci_420","problem_group_id":"langsci420:5.10","chapter":5,"chapter_title":"Noun and noun phrase","section":9,"section_title":"Practice problems","topic":"noun morphology and noun phrases","language":"Palauan","author":"Michael Salter","competition":"NACLO","year":2018,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s01_p01","book_method_c05_s01_p02","book_method_c05_s02_p01","book_method_c05_s03_p01","book_method_c05_s03_p02","book_method_c05_s04_p01","book_method_c05_s05_p01","book_method_c05_s06_p01","book_method_c05_s07_p01","book_method_c05_s08_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":1265,"source_line_end":1310,"solution_line_start":1697,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun phrase\nPractice problems\nnoun morphology and noun phrases\nThis practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some phrases in Palauan and their English translations:\n\neru ęl buil | ‘2 months’ | kltiu ęl hong | ‘9 books’\n\nede ęl sils | ‘3 days’ | kllolem ęl lius | ‘6 coconuts’\n\ntede ęl chad | ‘3 people’ | teai ęl ngalęk | ‘8 children’\n\nkllolem ęl malk | ‘6 chickens’ | ongeru ęl buil | ‘February’\n\nteim ęl sensei | ‘5 teachers’ | ongede ęl ureor | ‘Wednesday’\n\neim ęl rak | ‘5 years’ | etiu ęl klębęse | ‘9 nights’\n\ntęruich me a tede ęl buik | ‘13 boys’\n\ntęruich me a euid ęl sikang | ‘17 hours’\n- Translate into English:\n\n- telolem ęl sensei\n\n- tęruich me a etiu ęl buil\n\n- tęruich me a ongeru ęl buil\n\n- ongeim ęl ureor\n\n- Translate into Palauan:\n\n- ‘8 days’\n\
{"id":"book_05_11","context":"Here are some sentences in Norwegian and their English translations in random order:\n. | #1 | . | ‘#2’\n\n. | Bussen stanser her. | . | ‘A woman has the apple.’\n. | Jeg har en bil. | . | ‘I have an apple.’\n. | Bilen stanser her. | . | ‘The bus stops here.’\n. | Jeg har eplet. | . | ‘The woman has cars.’\n. | Jeg har et eple. | . | ‘The car stops here.’\n. | En kvinne har eplet. | . | ‘I have buses.’\n. | Kvinna har biler. | . | ‘The woman has the cars.’\n. | Kvinna har bilene. | . | ‘I have the apple.’\n. | Jeg har busser. | . | ‘The women stop here.’\n. | Kvinnene stanser her. | . | ‘I have a car.’","query":"- Determine the correct correspondences.\n\n- Nouns in Norwegian can belong to one of three classes: masculine, feminine, or neuter. The class determines how the noun can be used with determiners (words such as ‘the’, ‘a’, ‘an’) and be made plural. The nouns you encountered above are all regular and feature examples of all three classes:\n\nkvinne – feminine, bil – masculine, eple – neuter\n\n- Here are three more regular Norwegian nouns and their translations:\njente (feminine) = ‘girl’, hund (masculine) = ‘dog’, hotell (neuter) = ‘hotel’\n\n- Translate into Norwegian:\n\n- ‘The girl stops here.’\n\n- ‘A girl has a hotel.’\n\n- ‘I have the dogs.’\n\n- ‘The girl has dogs.’\n\n- Here are some more Norwegian words without any information about the classes the slightly irregular nouns belong to: sko = ‘shoe’, mann = ‘man’, ikke = ‘not’\n\n- Translate into English:\n\n- Mennene har epler.\n\n- Kvinna har ikke skoene.\n\n- Jeg har ikke eplene.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"5.11. Norwegian\n\n-\n\n- C.\n\n- J.\n\n- E.\n\n- H.\n\n- B.\n\n- A.\n\n- D.\n\n- G.\n\n- F.\n\n- I.\n\n-\n\n- Jenta stanser her.\n\n- En jente har et hotell.\n\n- Jeg har hundene.\n\n- Jenta har hunder.\n\n-\n\n- ‘The men have apples.’\n\n- ‘The woman does not have the shoes.’\n\n- ‘I do not have the apples.’\n\nRules:\n\nSentence structure: SOV (Subject-Object-Verb)\n\n| Last | sg indef. | sg def. | pl def. | pl indef.\n| letter | ‘a dog’ | ‘the dog’ | ‘the dogs’ | ‘dogs’\nNeuter | e | et X-e | X-et | X-ene | X-er\n| e | et X | X-et | X-ene | X-er\nFeminine | e | en X-e | X-a | X-ene |\n| e | en X | X-a | X-ene |\nMasculine | e | en X-e | X-en | X-ene | X-er\n| e | en X | X-en | X-ene | X-er\n\nNeuter | Feminine | Masculine\n‘apple’ | ‘woman’ | ‘bus’\n‘hotel’ | ‘girl’ | ‘car’\n| | ‘dog’\n\nNotebulbonThe nouns for ‘man’ and ‘shoe’ are not included in the table, since their gender cannot be determined.","source":"langsci_420","problem_group_id":"langsci420:5.11","chapter":5,"chapter_title":"Noun and noun phrase","section":9,"section_title":"Practice problems","topic":"noun morphology and noun phrases","language":"Norwegian","author":"Babette Verhoeven-Newsome","competition":"NACLO","year":2017,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s01_p01","book_method_c05_s01_p02","book_method_c05_s02_p01","book_method_c05_s03_p01","book_method_c05_s03_p02","book_method_c05_s04_p01","book_method_c05_s05_p01","book_method_c05_s06_p01","book_method_c05_s07_p01","book_method_c05_s08_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":1312,"source_line_end":1372,"solution_line_start":1750,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun
{"id":"book_05_12","context":"Here are some words in Afrihili and their English translations:\n\nmmmmmm texttrtoothadu ‘tooth’\nafidi ‘machine’\najamuri ‘republic’\nakalini ‘pen’\namadu ‘dentist’\namkate ‘bread’\namola ‘children’\namukamo ‘kingdom’\naturesine ‘bouquet’\nemelisini ‘fleet’\nemeli ‘ship’\nenti ‘date tree’\neshuli ‘principal’\neture ‘flowers’\nijamura ‘president’\nikalini ‘pens’\nilengi ‘horses’\nimukazi ‘girls’\nisabamatu ‘cobbler’\nishule ‘school’\nolengi ‘horse’\noluganda ‘dialect’\nomola ‘child’\nomukazi ‘girl’\nomuntundu ‘dwarf’\nomuntu ‘man’\nuruzindi ‘stream’\nuruzi ‘river’","query":"- Translate into English: ajamura, amkamate, oluga.\n\n- Translate into Afrihili: ‘machinist’, ‘ships’, ‘flower’, ‘group of girls’, ‘date fruit’, ‘shoe’, ‘king’.\n\n- Below are three more Afrihili words and three options for a likely translation of the word:\n\nimulenzi | a. ‘fruit’ | b. ‘boys’ | c. ‘bridge’\n\naposino | a. ‘baggage’ | b. ‘classroom’ | c. ‘parent’\n\niwelemase | a. ‘book’ | b. ‘library’ | c. ‘librarian’\n\n- Pick the translation most likely to be correct and explain your choice.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"5.12. Afrihili\n\n-\n\n- ajamura = ‘presidents’\n\n- amkamate = ‘baker’\n\n- oluga = ‘language’\n\n-\n\n- ‘machinist’ = afimadi\n\n- ‘ships’ = imeli\n\n- ‘flower’ = ature\n\n- ‘group of girls’ = omukazisini\n\n- ‘date fruit’ = entindi\n\n- ‘shoe’ = isabatu\n\n- ‘king’ = omukama\n\n- [t]llQ\nimulenzi | = b. | ‘boys’ (first and last vowels are identical plural)\naposino | = a. | ‘baggage’ (affix -sin- collective noun)\niwelemase | = c. | ‘librarian’ (affix -ma- profession)\n\nRules:\n\n- All nouns begin and end with a vowel (V_1RV_2)\n\n- Singular: V_1 V_2, plural: V_1 V_2 (V_1 becomes V_2 V_2RV_2)\n\n- Derived nouns (from a basic noun V_1RV_2):\n\n- head of an organisation: V_2RV_1 (first and last vowels switch places);\n\n- profession: infix -ma- is inserted before the last syllable of the stem (V_1C...CV_2 V_1C...maCV_2);\n\n- collective noun: V_1RV_2sinV_2 (add suffix -sinV, where V is the last vowel of the stem);\n\n- diminutive: V_1RV_2ndV_2","source":"langsci_420","problem_group_id":"langsci420:5.12","chapter":5,"chapter_title":"Noun and noun phrase","section":9,"section_title":"Practice problems","topic":"noun morphology and noun phrases","language":"Afrihili","author":"Michael Salter & Aleka Blackwell","competition":"UKLO","year":2019,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s01_p01","book_method_c05_s01_p02","book_method_c05_s02_p01","book_method_c05_s03_p01","book_method_c05_s03_p02","book_method_c05_s04_p01","book_method_c05_s05_p01","book_method_c05_s06_p01","book_method_c05_s07_p01","book_method_c05_s08_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":1376,"source_line_end":1425,"solution_line_start":1821,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun phrase\nPractice problems\nnoun morphology and noun phrases\nThis practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words in Afrihili and their English translations:\n\nmmmmmm texttrtoothadu ‘tooth’\nafidi ‘machine’\najam
{"id":"book_05_13","context":"Here are some phrases in Maltese and their English translations:\nserp aħdar | ‘green snake’ | mogħża sewda | ‘[blank] [blank]’\n\nmogħżiet bojod | ‘white goats’ | ktieb aħmar | ‘[blank] [blank]’\n\nkowt blu | ‘blue coat’ | mejda bajda | ‘[blank] [blank]’\n\nmejda kannella | ‘brown chair’ | baqra [blank] | ‘blue cow’\n\nżiemel iswed | ‘black horse’ | fjuri [blank] | ‘red flowers’\n\nfjura vjola | ‘purple flower’ | kelb [blank] | ‘brown dog’\n\nkarozza ħamra | ‘red car’ | kotba [blank] | ‘yellow books’\n\nrħula ħodor | ‘green settlments’ | siġra [blank] | ‘green tree’\n\nkowtijiet roża | ‘pink coats’ | mwejjed [blank] | ‘purple chairs’\n\nqomos blu | ‘blue shirts’ | tuffieħa [blank] | ‘yellow apple’\n\nfenek isfar | ‘yellow rabbit’ | [blank] [blank] | ‘red snake’\n\nqomos sowod | ‘black shirts’ |","query":"- Fill in the blanks. Each blank corresponds to a single word.\n\n- Based on the data given, one cannot translate ‘white book’. Why not?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"5.13. Maltese\n\n-\n\n- ‘goat’\n\n- ‘black’\n\n- ‘book’\n\n- ‘red’\n\n- ‘chair’\n\n- ‘white’\n\n- blu\n\n- ħomor\n\n- kannella\n\n- sofor\n\n- ħadra\n\n- vjola\n\n- safra\n\n- serp\n\n- aħmar\n\n- We do not know the first vowel of the stem for ‘white’ (V_1).\n\nRules:\nThere are two types of colours: invariable (‘pink’, ‘purple’, ‘brown’, ‘blue’) – which have the same form in all contexts – and variable (‘green’, ‘white’, ‘black’, ‘red’, ‘yellow’).\n\nNouns are divided into three categories:\n\n- Singular, end in a consonant;\n\n- Singular, end in a;\n\n- Plural.\n\nBased on these categories, the adjective is declined as follows:\n\n- V_1C_1C_2V_2C_3\n\n- C_1V_2C_2C_3a\n\n- C_1oC_2oC_3\n\nwhere C and V refer to a consonant and a vowel, respectively.\n\nWe notice here that V_1 appears only for Cat. I. Therefore, knowing the forms for Cat. II and III is not sufficient to deduce the form for Cat. I; similarly, only knowing the form for Cat. III is not enough to deduce the forms for Cat. I and II.","source":"langsci_420","problem_group_id":"langsci420:5.13","chapter":5,"chapter_title":"Noun and noun phrase","section":9,"section_title":"Practice problems","topic":"noun morphology and noun phrases","language":"Maltese","author":"Simona Strizhevskaya","competition":"LLO","year":2020,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s01_p01","book_method_c05_s01_p02","book_method_c05_s02_p01","book_method_c05_s03_p01","book_method_c05_s03_p02","book_method_c05_s04_p01","book_method_c05_s05_p01","book_method_c05_s06_p01","book_method_c05_s07_p01","book_method_c05_s08_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":1427,"source_line_end":1449,"solution_line_start":1870,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun phrase\nPractice problems\nnoun morphology and noun phrases\nThis practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some phrases in Maltese and their English translations:\nserp aħdar | ‘green snake’ | mogħża sewda | ‘[blank] [blank]’\n\nmogħżiet bojod | ‘white goats’ | ktieb aħmar | ‘[blank] [blank]’\n\nkowt blu | ‘blue coat’ | mejda bajda | ‘[
{"id":"book_05_14","context":"Here are some phrases in Latvian and their English translations:\n\n. | augsts ozols | ‘tall oak’\n. | vecs grāmatu veikals | ‘old bookstore’\n. | veca meža | ‘of the old wood’\n. | stikla galds | ‘glass table’\n. | bruņinieka cimds | ‘knight's glove’\n. | sudraba ābols | ‘silver apple’\n. | autora teksts | ‘author's text’\n. | operāciju galds | ‘surgery table’\n. | pretīgu piena ēdienu | ‘of disgusting dairy products’\n. | labs institūts | ‘good institute’\n. | laba bērnu ārsta | ‘of the good paediatrician’\n. | grāmatu veikala | ‘[blank]’\n. | [blank] kolektīvs | ‘authors' group’\n. | [blank] turnīrs | ‘knight tournament’\n. | pretīgs [blank] | ‘[blank] child’\n. | balts [blank] | ‘white silver’\n. | [blank] zara | ‘of the oak branch’\n. | [blank] | ‘of the oak wood’\n. | [blank] ārstu | ‘of the institute's doctors’","query":"- Fill in the blanks. Some blanks may correspond to multiple words. If you think some blanks can be filled in different ways, write all possibilities and explain the difference between them.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"5.14. Latvian\n\n-\n\n- ‘of the bookshop’\n\n- autoru\n\n- bruņinieku\n\n- bērns\n\n- ‘disgusting’\n\n- sudrabs\n\n- ozola\n\n- ozolu\n\n- institūta or institūtu (the former is used if we refer to the doctors from a single institute, while the latter is used to refer to the doctors from different institutes)\n\n- Suffixes of the noun head:\n\n-s = nominative -a = genitive sg -u = genitive pl\n\n- Determiners can be split into three groups:\n\n- Those which agree with the noun head (receive the same suffix). In this category, we include all the qualifying adjectives (‘tall’, ‘old’, ‘good’, ‘disgusting’).\n\n- Those which receive the ending -a – they are represented by Latvian nouns in genitive singular (‘dairy products’ = products of milk).\n\n- Those receiving the ending -u – they are Latvian nouns in genitive plural (‘paediatrician’ = doctor of children, ‘bookshop’ = shop of books).\n\nThe difference between the last two groups is purely semantic, depending on whether they refer to a singular or a plural noun (a ‘surgery table’ is a table for surgeries since multiple surgeries are performed on the same table; a ‘paediatrician’ is a doctor of children because they treat more children, not a single one, etc.)","source":"langsci_420","problem_group_id":"langsci420:5.14","chapter":5,"chapter_title":"Noun and noun phrase","section":9,"section_title":"Practice problems","topic":"noun morphology and noun phrases","language":"Latvian","author":"Maria Rubinstein","competition":"MSK","year":1999,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s01_p01","book_method_c05_s01_p02","book_method_c05_s02_p01","book_method_c05_s03_p01","book_method_c05_s03_p02","book_method_c05_s04_p01","book_method_c05_s05_p01","book_method_c05_s06_p01","book_method_c05_s07_p01","book_method_c05_s08_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":1451,"source_line_end":1481,"solution_line_start":1924,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun phrase\nPractice problems\nnoun morphology and noun phrases\nThis practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some
{"id":"book_05_15","context":"Below are 12 Ilocano words written in the Baybayin script, as well as their English translations given in random order:\n\n1. | \"1703\"1712\"1706 |\n7. | \"170D\"1713\"170D\"1713\"1704\"1714\n\n2. | \"1703\"1712\"1706\"1714\"1703\"1712\"1706 |\n8. | \"170D\"1713\"170D\"1714\"170D\"1713\"170D\"1713\"1704\"1714\n\n3. | \"1703\"1713\"170B\"1712\"1706 |\n9. | \"170D\"1713\"170B\"1713\"170D\"1714\"170D\"1713\"170D\"1713\"1704\"1714\n\n4. | \"1703\"1713\"170B\"1712\"1706\"1714\"1703\"1712\"1706 |\n10. | \"1704\"1713\"170B\"1706\"1705\"1714\n\n5. | \"170D\"1704\"1714\"1710\"1703\"1714 |\n11. | \"1710\"1709\"1706\"1714\n\n6. | \"170D\"1713\"170B\"1704\"1714\"170D\"1704\"1714\"1710\"1703\"1714 |\n12. | \"1710\"1713\"170B\"1709\"1706\"1714\n\n‘to look’, ‘is skipping with joy’, ‘is becoming a skeleton’, ‘a skeleton’, ‘to buy’, ‘various skeletons’, ‘various appearances’, ‘to reach the top’, ‘is looking’, ‘appearance’, ‘summit’, ‘happiness’, ‘skeleton’","query":"- Determine the correct correspondences.\n\n- Fill in the blanks.\n\n\"170D\"1713\"170B\"1713\"170D\"1713\"1704\"1714 | ‘[blank]’\n\n\"1710\"1709\"1714\"1710\"1709\"1706\"1714 | ‘[blank]’\n\n\"1710\"1713\"170B\"1709\"1714\"1710\"1709\"1706\"1714 | ‘[blank]’\n\n[blank] | ‘(a) purchase’\n\n[blank] | ‘is buying’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"5.15. Ilocano\n\n-\n\n- ‘appearance’\n\n- ‘various appearances’\n\n- ‘to look’\n\n- ‘is looking’\n\n- ‘happiness’\n\n- ‘is skipping for joy’\n\n- ‘skeleton’\n\n- ‘various skeletons’\n\n- ‘is becoming a skeleton’\n\n- ‘to buy’\n\n- ‘summit’\n\n- ‘to reach the top’\n\n-\n\n- ‘to become a skeleton’\n\n- ‘various summits’\n\n- ‘is reaching the top’\n\n- \"1704\"1706\"1705\"1714\n\n- \"1704\"1713\"170B\"1706\"1714\"1704\"1706\"1705\"1714\n\nRules:\n\n- Stems:\n\n- ‘appearance’ = \"1703\"1712\"1706\n\n- ‘skeleton’ = \"170D\"1713\"170D\"1713\"1704\"1714\n\n- ‘purchase’ = \"1704\"1706\"1705\"1714\n\n- ‘happiness’ = \"170D\"1704\"1714\"1710\"1703\"1714\n\n- ‘summit / top’ = \"1710\"1709\"1706\"1714\n\n-\n\n- Derivation processes:\n\n- Partial reduplication: copy the first two symbols and add them to the beginning of the word. The first symbol retains its diacritic, while the diacritic of the second symbol is replaced by a plus (\"25CC\"1714).\n\n- Epenthesis: insert \"170B after the first symbol. The diacritic of the first symbol moves to this one and the first symbol receives an underdot (\"25CC\"1713).\n\nWith these two processes, we can obtain three different transformations starting from the singular noun:\n\n- Plural noun (‘various...’) (process a)\n\n- Infinitive verb (process b)\n\n- 3sg present cont. verb (process a, followed by b)","source":"langsci_420","problem_group_id":"langsci420:5.15","chapter":5,"chapter_title":"Noun and noun phrase","section":9,"section_title":"Practice problems","topic":"noun morphology and noun phrases","language":"Ilocano","author":"Patrick Littell","competition":"NACLO","year":2008,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s01_p01","book_method_c05_s01_p02","book_method_c05_s02_p01","book_method_c05_s03_p01","book_method_c05_s03_p02","book_method_c05_s04_p01","book_method_c05_s05_p01","book_method_c05_s06_p01","book_method_c05_s07_p01","book_method_c05_s08_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":1483,"source_line_end":1526,"solution_line_start":1961,"license":"CC-BY-4.0","attribu
{"id":"book_05_16","context":"Below are some number phrases in Irish and their English equivalents:\n\n. | garra amháin | ‘1 garden’\n. | gasúr déag | ‘11 boys’\n. | ocht mballa is dhá fichid | ‘48 walls’\n. | dhá gharra déag is ceithre fichid | ‘92 gardens’\n. | trí bhád | ‘3 boats’\n. | seacht ndoras déag | ‘17 doors’\n. | seacht mbád déag is dhá fichid | ‘57 boats’\n. | naoi nduine déag is fiche | ‘39 people’\n. | ceithre fichid doras | ‘80 doors’\n. | cúig bhalla | ‘5 walls’\n. | sé ghasúr is trí fichid | ‘66 boys’\n. | deich mbád | ‘10 boats’\n. | sé dhuine | ‘6 people’\n. | trí dhoras is dhá fichid | ‘43 doors’\n. | garra is ceithre fichid | ‘81 gardens’","query":"- Translate into English:\n\n- naoi mbád déag is ceithre fichid\n\n- sé dhuine déag\n\n- naoi nduine\n\n- fiche gasúr\n\n- garra déag is fiche\n\n- Translate into Irish:\n\n- ‘2 boys’\n\n- ‘38 walls’\n\n- ‘14 walls’\n\n- ‘71 doors’\n\n- ‘21 boats’\n\n- ‘90 people’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"5.16. Irish\n\n-\n\n- ‘99 boats’\n\n- ‘16 people’\n\n- ‘9 people’\n\n- ‘20 boys’\n\n- ‘31 gardens’\n\n-\n\n-\n\n- dhá ghasúr\n\n- ocht mballa déag is fiche\n\n- ceithre bhalla déag\n\n- doras déag is trí fichid\n\n- bád is fiche\n\n- deich nduine is ceithre fichid\n\nRules:\n\nIrish uses base 20; numbers are written as:\n\n(U) + (10) + (20X) (U) (déag) (is X fichid)\n\n- U is between 2 and 9\n\n- number 10 has two forms: déag – used if there is a unit number (from 2 to 9) – and deich – used only for 10 and multiples of 10 (30, 50, etc.)\n\n- if X = 1, is X fichid becomes is fiche\n\nPhrase structure: we can consider each structure to have four parts: I + II + III + IV.\n\n- Part I is for the units (or deich).\n\n- Part II is always the noun.\n\n- Part III can only be filled by two words: amháin (meaning 1, only for a singular noun) or déag (but not deich).\n\n- Part IV is always a multiple of 20 (structures is X fichidis fiche).\n\nThe only exception takes place when the number of objects is a multiple of 20 (20, 40, etc.). In this case, Part I is occupied by X fichidfiche and the noun is placed at the end (in this case the particle is is not used). Therefore, the examples in the problem (and task (a)) can be analysed as follows:\n\n| I | II | III | IV | Translation\n| I | II | III | IV | Translation\n1. | | garra | amháin | | ‘1 garden’\n2. | | gasúr | déag | | ‘11 boys’\n3. | ocht | mballa | | is dhá fichid | ‘48 walls’\n4. | dhá | gharra | déag | is ceithre fichid | ‘92 gardens’\n5. | trí | bhád | | | ‘3 boats’\n6. | seacht | ndoras | déag | | ‘17 doors’\n7. | seacht | mbád | déag | is dhá fichid | ‘57 boats’\n8. | naoi | nduine | déag | is fiche | ‘39 people’\n10. | cúig | bhalla | | | ‘5 walls’\n11. | sé | ghasúr | | is trí fichid | ‘66 boys’\n12. | deich | mbád | | | ‘10 boats’\n13. | sé | dhuine | | | ‘6 people’\n14. | trí | dhoras | | is dhá fichid | ‘43 doors’\n15. | | garra | | is ceithre fichid | ‘81 gardens’\n16. | naoi | mbád | déag | is ceithre fichid | ‘99 boats’\n17. | sé | dhuine | déag | | ‘16 people’\n18. | naoi | nduine | | | ‘9 people’\n20. | | garra | déag | is fiche | ‘31 gardens’\n9. | ceithre fichid | doras | | | ‘80 doors’\n19. | fiche | gasúr | | | ‘20 boys’\n\nWe separated examples 9 and 19 to highlight the case in which the number of objects is a multiple of 20.\n\nMoreover, we can notice from the table that the difference between, for example, 20 and 21 (or 40/41, 60/61 etc.) is only based on the position of the noun.\n\nLastly, we notice that the noun has a variable form. In this case, it undergoes an initial consonant mutation. We notice three types of initial consonants:\n\n- simple consonants (b, d, g) – used if the number of units is 1 or if the number of objects is a multiple
{"id":"book_05_17","context":"Here are some phrases spoken by two speakers of Iaai and their English translations, in random order:\n\n*Speaker 1 (male, 50 years old):\n\n. | hoom hu | . | ‘your fire’\n. | belem mââng | . | ‘my water’\n. | waau haalee Aiawa | . | ‘my bus’\n. | uutap taben than | . | ‘your mango’\n. | tabik kar | . | ‘your boat’\n. | anyik sawakiny | . | ‘the chief's chair’\n. | anyim meic | . | ‘my necklace’\n. | belik köiö | . | ‘Aiawa's cat’\n\n*Speaker 2 (male, 12 years old):\n\n. | belen koka | . | ‘your goat’\n. | anyik tang | . | ‘his yam’\n. | anyim karopëë | . | ‘Kua's mother’\n. | haaleem nani | . | ‘his watermelon’\n. | an koko | . | ‘his car’\n. | hinyö anyi Kua | . | ‘my basket’\n. | belik nu | . | ‘his Coke’\n. | anyin loto | . | ‘my coconut’\n. | an waajem | . | ‘your dugout’\n\nâ, ë, ö are vowels. A ‘dugout’ is a long, narrow canoe made of a tree trunk. A ‘yam’ is a starchy vegetable, similar to a sweet potato. ‘Coke’ is a carbonated beverage sold by The Coca-Cola Company, an American multinational company. ‘Aiawa’ and ‘Kua’ are names of people.","query":"- For each speaker, determine the correct correspondences.\n\n- Translate into English:\n\n- belem waajem\n\n- karopëë hoon hinyö\n\n- Mention any meaning that is not reflected in the literal translation.\n\n- Given below are some English words and their Iaai translations:\n\n- ‘dog’ = kuli\n\n- ‘tea’ = trii\n\n- ‘canoe’ = ok\n\n- Translate into Iaai in all possible ways:\n\n- ‘the cat's tea’\n\n- ‘Kua's coconut’\n\n- ‘his canoe’\n\n- ‘my dog’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"5.17. Iaai\n\n-\n\n- E.\n\n- D.\n\n- H.\n\n- F.\n\n- C.\n\n- G.\n\n- A.\n\n- B.\n\n- O.\n\n- N.\n\n- Q.\n\n- I.\n\n- J.\n\n- K.\n\n- P.\n\n- M.\n\n- L.\n\n-\n\n- ‘your watermelon (for drinking)’\n\n- ‘the mother's dugout’\n\n-\n\n- trii belen waau\n\n- nu a Kua or nu bele Kua\n\n- hoon ok or anyin ok\n\n- haaleik kuli\n\nRules:\n\nStructure:\n\n- X's Y =\n\nC–Poss Y | (X = pronoun)\nY C–Poss X | (X = common noun)\nY C X | (X = proper noun)\n\n- Poss:\n\n-k | (X = 1sg) e i / _ k\n-m | (X = 2sg)\n-n | (X = 3sg)\n\n- C:\n\na | Y = food*\nbele | Y = drinks*\nhaalee | Y = animals\nhoo | Y = boats\ntabe | Y = things you can sit on/in\nanyi | otherwise\n\n* Fruits can be either ‘food’ or ‘drink’ depending on how the speaker intends them to be consumed.\n\nIn the case of younger generations (Speaker 2), these types of noun also fall into the anyi category (new generations tend to simplify the classifier system and give up on very specific classifiers, preferring to use the general classifier).","source":"langsci_420","problem_group_id":"langsci420:5.17","chapter":5,"chapter_title":"Noun and noun phrase","section":9,"section_title":"Practice problems","topic":"noun morphology and noun phrases","language":"Iaai","author":"Rujul Gandhi","competition":"APLO","year":2020,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c05_s01_p01","book_method_c05_s01_p02","book_method_c05_s02_p01","book_method_c05_s03_p01","book_method_c05_s03_p02","book_method_c05_s04_p01","book_method_c05_s05_p01","book_method_c05_s06_p01","book_method_c05_s07_p01","book_method_c05_s08_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Noun and noun phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/05-Noun.tex","source_line_start":1577,"source_line_end":1646,"solution_line_start":2114,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Noun and noun ph
{"id":"book_06_01","context":"Here are some verbal forms in Swahili and their English translations:\n\n. | Ninasema. | ‘I speak.’\n. | Wunasema. | ‘You speak.’\n. | Anasema. | ‘She speaks.’\n. | Wanasema. | ‘They speak.’\n. | Ninaona. | ‘I see.’\n. | Niliona. | ‘I saw.’\n. | Ninawaona. | ‘I see them.’\n. | Niliwuona. | ‘I saw you.’\n. | Ananiona. | ‘She sees me.’\n. | Wutakaniona. | ‘You will see me.’\n. | [blank] | ‘She saw them.’\n. | [blank] | ‘I will see you.’\n. | [blank] | ‘She saw me.’","query":"- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"We notice that all English examples containing the verb ‘to see’ end in ona in Swahili. Similarly, all English examples which contain the subject ‘I’ start with ni in Swahili. Separating these two morphemes, we obtain:\n\n. | Ni-nasema. | ‘I speak.’\n. | Wunasema. | ‘You speak.’\n. | Anasema. | ‘She speaks.’\n. | Wanasema. | ‘They speak.’\n. | Ni-na-ona. | ‘I see.’\n. | Ni-li-ona. | ‘I saw.’\n. | Ni-nawa-ona. | ‘I see them.’\n. | Ni-liwu-ona. | ‘I saw you.’\n. | Anani-ona. | ‘She sees me.’\n. | Wutakani-ona. | ‘You will see me.’\n\nIn examples 5 and 6 we are left with only one unidentified morpheme (na and li, respectively). The two sentences differ only in terms of their tense (past vs. present); therefore, we infer that na is the present-tense marker, while li is the past-tense marker. Separating these two morphemes, we obtain:\n\n. | Ni-na-sema. | ‘I speak.’\n. | Wu-na-sema. | ‘You speak.’\n. | A-na-sema. | ‘She speaks.’\n. | Wa-na-sema. | ‘They speak.’\n. | Ni-na-ona. | ‘I see.’\n. | Ni-li-ona. | ‘I saw.’\n. | Ni-na-wa-ona. | ‘I see them.’\n. | Ni-li-wu-ona. | ‘I saw you.’\n. | A-na-ni-ona. | ‘She sees me.’\n. | Wutakani-ona. | ‘You will see me.’\n\nNow it is easy to identify that the morpheme order is Subject – Tense – Object – Stem (we can abbreviate it as S-Tense-O-V). Moreover, we notice that the subject and object markers are identical. Therefore, we can segment example 10 as well (knowing that ‘you’ is -wu-, and ‘I/me’ is -ni-) and we get Wu-taka-ni-ona. Therefore, the future marker is taka.\n\nOnce all the rules are discovered, it is time to structure them neatly. Usually, in this type of problem, we start by writing down the morpheme order, followed by explaining each of the morphemes.\n\n*Rules: Option 1\n\n- Structure: S–Tense–O–V\n\n- S: 1sg = ni, 2sg = wu, 3sg = a, 3pl = wa\n\n- Tense: present = na, past = li, future = taka\n\n- O: 1sg = ni, 2sg = wu, 3sg = a, 3pl = wa\n\n- Verb: ‘speak’ = sema, ‘see’ = ona\n\nAnother option is combining the order of the morphemes with their structure, in a single table.\n\n*Rules: Option 2\n\nS | Tense | O | V\n\nni | 1sg\nwu | 2sg\na | 3sg\nwa | 3pl\n|\n\nli | past\nna | present\ntaka | future\n|\n\nni | 1sg\nwu | 2sg\na | 3sg\nwa | 3pl\n|\n\nsema | speak\nona | see\n\nThe main disadvantage of these two options (in the case of this problem) is that we have to write the pronoun markers twice (once for S and once for O), although they are identical.\n\nUsually, when we work with pronoun markers, it is favourable to structure them in a table in which we write the person in columns and the number in rows (or vice versa). Therefore, the subject markers become:\n\n| 1 | 2 | 3\nsg | ni | wu | a\npl | | | wa\n\nIn this way, we can show the fact that the subject is identical to the object.\n\n*Rules: Option 3\n\n- Structure: S–Tense–O–V\n\n- S = O\n| 1 | 2 | 3\nsg | ni | wu | a\npl | | | wa\n\n- Tense: present = na, past = li, future = taka\n\n- Verb: ‘speak’ = sema, ‘see’ = ona\n\nAll these options for writing out the rules are correct and complete, and they would all receive the maximum score. But, depending on the problem, one of them fits better in the sense that it is more succinct and helps save some time.\n\nOnce the rules are written, we ca
{"id":"book_06_02","context":"Here are some verbal forms in Dabida and their English translations in random order:\n\ndichakaδana, βichanirasha, kuchanikunda, dichakurasha, dicharashana,\nβichamukunda, muchadikaδa, βichakaδana\n‘we will argue’, ‘youpl will beat us’, ‘we will curse yousg’, ‘they will fight’, ‘they will curse me’,\n‘they will fall in love with youpl’, ‘we will fight’, ‘yousg will fall in love with me’","query":"- Determine the correct correspondences.\n\n- Translate into English:\n\n- nichakukaδa\n\n- βichakundana\n\n- Translate into Dabida:\n\n- ‘youpl will argue’\n\n- ‘yousg will curse them’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"- Step 1. Segmenting: when segmenting the verbs, we are not yet too concerned with the English translations. For this problem, we can start with the special characters since they are the easiest to follow. If we start with the letter β, we notice that it appears in three examples and in each of these examples we can separate the morpheme βicha. Nevertheless, we need to notice that the morpheme -cha- appears in every single given example. Therefore, it most likely represents a separate morpheme. Separating the morphemes -cha- and -βi- we get:\n\ndi-cha-kaδana, βi-cha-nirasha, ku-cha-nikunda, di-cha-kurasha, di-cha-rashana,\nβi-cha-mukunda, mu-cha-dikaδa, βi-cha-kaδana\n\nNext, we notice the repeated strings -rasha-, -kunda-, and -kaδa-, obtaining:\n\ndi-cha-kaδa-na, βi-cha-ni-rasha, ku-cha-ni-kunda, di-cha-ku-rasha, di-cha-rasha-na,\nβi-cha-mu-kunda, mu-cha-di-kaδa, βi-cha-kaδa-na\n\n- Step 2. Now we can deduce the Dabida verb structure:\n\ndi\nβi\nku\nmu\n|\ncha\n|\ndi\nni\nku\nmu\n\n|\nkaδa\nrasha\nkunda\n|\nna\n\nRememberbulbIn case of morphemes 3 and 5, it is important to also mark (the null morpheme), meaning that the morpheme is optional and it does not appear in all examples.\n\n- Step 3. Checking the English examples, we notice only three variables: subject (‘yousg’, ‘we’, ‘youpl’, ‘they’), object (, ‘me’, ‘yousg’, ‘us’, ‘youpl’) and verb (‘to fight’, ‘to argue’, ‘to beat’, ‘to curse’, ‘to fall in love’). The tense is not a variable since all examples are in the future tense. We can probably assume that the morpheme -cha-, which appears in every single example, is the mark of the future.\n\nRememberbulbFor a correct and complete set of rules, we do not need to write the meaning of that morpheme, since it appears in all examples. Actually, we have no proof that it indicates future tense; we have no evidence as to what its function is as there are no contrasting examples without it.\n\nMoreover, we can make some preliminary observations in order to help deduce what is the purpose of each morpheme. We notice that morphemes 1 and 3 are extremely similar (both of them can be -di-, -ku-, -mu-). Therefore, we can guess that these mark the subject and the object (both have the same markers, similar to the previous problem). Moreover, the long morpheme is, usually, the stem (morpheme 4). Nevertheless, we notice that in Dabida we have only three stems, while in English we have five.\n\nWe make a frequency table for morphemes 1 and 3 (which we presumed correspond to the pronouns):\n\n| Dabida\n2-3\nMorpheme | 1st | 3rd\ndi | 3 | 1\nβi | 3 |\nku | 1 | 1\nmu | 1 | 1\nni | | 2\n| | 3\n\n| English\n2-3\nPronoun | S | O\n1sg | | 2\n2sg | 1 | 1\n1pl | 3 | 1\n2pl | 1 | 1\n3pl | 3 |\n| | 3\n\nIn other words, this table shows, for example, that the morpheme di in Dabida appears in three examples in the first position and in a single example in the third position. In English, we have only one pronoun which matches this 3-and-1 pattern, namely the second person singular (2sg, ‘yousg’).\n\nWe can easily notice that the first morpheme in Dabida corresponds to the subject in English while the third morpheme corresponds to the object. Moreover, we can infer that -di- = 1pl, -ni- = 1
{"id":"book_06_03","context":"Here are some verbal forms in Ge'ez and their English translations:\n\ntawalada | ‘He was born.’\ntawaladu | ‘They were born.’\ntawaladna | ‘We were born.’\ntawaladkəmu | ‘Youpl were born.’\nqatalkəwo | ‘I killed him.’\nqatalkomu | ‘Yousg killed themm.’\nqatalomu | ‘He killed themm.’\nqatalon | ‘He killed themf.’\nqatalnon | ‘We killed themf.’\nqatalkəməwon | ‘Youpl killed themf.’\nqataləwo | ‘Theym killed him.’\nqataləwomu | ‘Theym killed themm.’\n\nThe subscripts m and f refer to masculine and feminine, respectively.","query":"- Translate into English:\n\n- tawaladku\n\n- qatalkəwon\n\n- qatalo\n\n- [blank]\n\n- Translate into Ge'ez:\n\n- ‘Yousg were born.’\n\n- ‘Youpl killed him.’\n\n- ‘We killed themm.’\n\n- ‘Theym killed themf.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"We easily notice the verb stem: tawalad- = ‘to be born’ and qatal- = ‘to kill’. Moreover, checking the English translations, we notice that the remaining morpheme needs to mark the subject and the object (all the other parameters are constant – tense, aspect, mood, etc.).\n\nSince, at first glance, we cannot identify any patterns, we include all these morphemes in a table in which we mark the combinations:\n\nO/S | 1sg | 2sg | 3sg.m | 1pl | 2pl | 3pl.m\n| | | -a | -na | -kǝmu | -u\n3sg.m | -kǝwo | | | | | -ǝwo\n3pl.m | | -komu | -omu | | | -ǝwomu\n3pl.f | | | -on | -non | -kǝmǝwon |\n\nLooking at the row in which the object is 3pl.m (‘themm’), we notice that all the entries end in -omu. Therefore, we can assume that this morpheme marks the 3pl.m object. Similarly, we can separate the object marker for 3pl.f = -on. If we separate these markers, the table becomes:\n\nO/S | 1sg | 2sg | 3sg.m | 1pl | 2pl | 3pl.m\n| | | -a | -na | -kǝmu | -u\n3sg.m | -kǝwo | | | | | -ǝwo\n3pl.m | | -k-omu | -omu | | | -ǝw-omu\n3pl.f | | | -on | -n-on | -kǝmǝw-on |\n\nIf we look at the column 3sg.m, we notice that the subject marker is null if there is an object, and it is -a if there is no object (if the verb is intransitive). Therefore, we can assume that, when an object marker is added, a.\n\nSimilarly, looking at the columns 2pl and 3pl.m, we deduce that when an object marker is added, u ǝw. Based on these two phonological rules, we can fill in the table with all the other combinations of subject and object:\n\nO/S | 1sg | 2sg | 3sg.m | 1pl | 2pl | 3pl.m\n| -ku | -ka | -a | -na | -kǝmu | -u\n3sg.m | -kǝwo | -ko | -o | -no | -kǝmǝwo | -ǝwo\n3pl.m | -kǝwomu | -komu | -omu | -nomu | -kǝmǝwomu | -ǝwomu\n3pl.f | -kǝwon | -kon | -on | -non | -kǝmǝwon | -ǝwon\n\nThus, we can write the rules and solve the tasks:\n\nRules:\n\n- Structure: V-S-O\n\n- Verb: tawalad- = ‘to be born’, qatal- = ‘to kill’\n\n- S and O:\n\n| 1 | 2 | 3M | 3F\nsg | S: -ku | S: -ka | S: -a, O: -o |\npl | S: -na | S: -kǝmu | S: -u, O: -omu | O: -on\n\nWhen an object morpheme is added, the final vowel of the subject changes: a and u ǝw.\n\n-\n\n- ‘I was born.’\n\n- ‘I killed themf.’\n\n- ‘He killed him.’\n\n-\n\n- tawaladka\n\n- qatalkǝmǝwo\n\n- qatalnomu\n\n- qatalǝwon","source":"langsci_420","problem_group_id":"langsci420:6.3","chapter":6,"chapter_title":"Verb and verb phrase","section":6,"section_title":"Patterns for arguments","topic":"verb morphology and argument structure","language":"Ge'ez","author":"Peter Arkadiev","competition":"TurLom","year":2007,"solution_tier":"detailed_worked","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s06_p01"],"method_link_type":"author_placement","method_description":"The author places this worked example under “Patterns for arguments” in the Verb and verb phrase chapter, so it illustrates that method or topic.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/06-Verb.tex","source_line_start":574,"source_lin
{"id":"book_06_04","context":"Here are some verbal forms in Itelmen and their English translations:\n\naniaķzоvŏmnen | ‘He was asking me.’\n\nnk'aniaķzozvŏmnen | ‘They would ask me.’\n\nnaniaķzoneʔn | ‘They were asking them.’\n\nk'añchpnen | ‘He would have taught him.’\n\nnañchpvŏmnеʔn | ‘They have taught us.’\n\nañchpķzoznеʔn | ‘He teaches them.’","query":"- Translate into English:\n\n- аñchpnеʔn\n\n- nk'аniаķzovŏmnеn\n\n- nanianen\n\n- Translate into Itelmen:\n\n- ‘He has asked them.’\n\n- ‘They ask us.’\n\n- ‘They would have taught me.’\n\n- ‘He would ask him.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"The first step is to segment the verbs. In this process, two of the morphemes can be a bit problematic:\n\n- The final morpheme (-nen / -neʔn): perhaps, at first sight, we would be tempted to say that -ne- is a separate morpheme which appears in all examples, followed by the morpheme -ʔ-, which is optional, and finally followed by -n which appears in all examples. Although this explanation is correct and is applicable to all examples, it unnecessarily overcomplicates the verb structure, for which reason it is more convenient to treat the whole structure as if it was a single morpheme.\n\n- A similar problem appears in the case of the morpheme -z- which is sometimes placed after the morpheme -ķzo-. Although we would perhaps be tempted to analyse it as a separate morpheme, we notice that it always appears after -ķzo- and nowhere else. Therefore, we will analyse -ķzoz- as a whole, not treating -z- as a separate morpheme.\n\nOf course, these observations are preliminary. If it turns out that this hypothesis does not work, we can come back and re-segment the verbs.\n\nBased on this, the verb segmentation is:\n\nania-ķzо-vŏm-nen | ‘He was asking me.’\n\nn-k'-ania-ķzoz-vŏm-nen | ‘They would ask me.’\n\nn-ania-ķzo-neʔn | ‘They were asking them.’\n\nk'-añchp-nen | ‘He would have taught him.’\n\nn-añchp-vŏm-nеʔn | ‘They have taught us.’\n\nañchp-ķzoz-nеʔn | ‘He teaches them.’\n\nAnd the Itelmen verb structure can be written as:\n\nI | II | III | IV | V | VI\nn- | -k'- | -ania- | -ķzо- | -vŏm- | -nen\n| | -añchp- | -ķzоz- | | neʔn\n| | | | |\n\nSince only morphemes III and VI are pervasive, it is very likely that one of these is the stem. Morpheme VI has only two forms, which are highly similar to one another, so, most likely, morpheme III is the stem. Indeed, based on the given examples, we can suggest that -ania- = ‘to ask’ and -añchp- = ‘to teach’.\n\nMoreover, looking at the examples that contain morpheme I (n-), we notice that it marks the 3pl subject.\n\nNotebulbonWe cannot know for sure whether this morpheme marks the subject 3pl or only that the subject is plural (not necessarily 3rd person), since, in all examples, the subject is 3rd person.\n\nMorpheme II occurs only in the structures ‘They would ask me’ and ‘He would have taught him’, and these two have in common the (conditional) mood. Moreover, we notice that there are no other examples in this mood.\n\nTherefore, we can summarise our findings:\n\nI | II | III | IV | V | VI\nn- = S3pl | -k'- = Cond. | -ania- = ‘to ask’ | -ķzо- | -vŏm- | -nen\n= S3sg | = Ind. | -añchp- = ‘to teach’ | -ķzоz- | | neʔn\n| | | | |\n\nFor the morpheme V, we can make a table in which we compare the examples that contain that morpheme, and those which do not:\n\n-vŏm- |\n‘He was asking me.’ | ‘They were asking them.’\n‘They would ask me.’ | ‘He would have taught him.’\n‘They have taught us.’ | ‘He teaches them.’\n\nThe only difference we can notice is that the morpheme occurs every time the object is in the 1st person (singular or plural). Therefore, we deduce that morpheme V marks the person of the object (-vŏm- = O1 and = O3). Moreover, since this morpheme only marks the person (not the number), we expect that one of the remain
{"id":"book_06_05","context":"Here are some verbal forms in Proto-Algonquian (in a simplified transcription) and their English translations:\n\n. | kewa:pameθehm | ‘I see yousg.’\n. | kewa:pameθehmwa: | ‘I see youpl.’\n. | newa:pama:ehma | ‘I see him.’\n. | newa:pama:ehmaki | ‘I see them.’\n. | kewa:pameθehmwa:ena:n | ‘We see youpl.’\n. | newa:pama:ehmena:na | ‘We see him.’\n. | kewa:pamiehm | ‘Yousg see me.’\n. | kewa:pama:ehma | ‘Yousg see him.’\n. | kewa:pamiehmwa: | ‘Youpl see me.’\n. | kewa:pamiehmwa:ena:n | ‘Youpl see us.’\n. | newa:pamekwehmena:naki | ‘They see us.’\n\nThe mark : after a vowel denotes length. θ = ‘th’ in ‘thin’. All ‘we’ pronouns in this problem refer to ‘weexcl’ (‘we exclusive’, meaning ‘me’ and ‘them’, not including the listener).","query":"- Translate into English: kewa:pamiehmena:n.\n\n- Translate into Proto-Algonquian: ‘We see them’ and ‘They see me’.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"After segmentation, we obtain the following verb structure in Proto-Algonquian:\n\nI | II | III | IV | V\nk- | ewa:pam | -eθ- | -ehm- | -a\nn- | | -a:- | | -aki\n| | -i- | | -wa:ena:n\n| | -ekw- | | -ena:na\n| | | | -ena:naki\n| | | |\n\nMorphemes II and IV are constant, so we do not need to offer a translation for them. Nevertheless, we can easily assume that one of them represents the stem (‘to see’) and the other one the tense (present). As mentioned at the beginning of this chapter, the stem tends to be the longest morpheme, so we can assume that morpheme II is the stem (of the verb), while morpheme IV represents the tense (present indicative).\n\nMorpheme I: we make a table in order to highlight the contrast between the two possible forms:\n\nk- | n-\n‘I see yousg.’ | ‘I see him.’\n‘I see youpl.’ | ‘I see them.’\n‘We see youpl.’ | ‘We see him.’\n‘Yousg see him.’ | ‘They see us.’\n‘Yousg see me.’ |\n‘Youpl see me.’ |\n‘Youpl see us.’ |\n\nWe infer that k- marks the presence of the 2nd person (either as a subject or as an object), while n- marks the absence thereof.\n\nAnalysing morpheme V, we infer that -ena:n- = 1pl, -wa:- = 2pl, -a- = 3sg, -aki- = 3pl. Moreover, we notice that persons 1sg and 2sg are unmarked. One of the interesting elements is that these morphemes have a specific order, independent of their role as a subject or object, namely -wa:- is always the first, followed by -ena:n- and, finally, by -a- or -aki-. In other words, it seems as if these morphemes are placed following a pronominal hierarchy: 2 > 1 > 3. The fact that this hierarchy is focused on the listener (2nd person) is also supported by the fact that the first morpheme refers strictly to the 2nd person.\n\nThe only morpheme left to analyse is morpheme III. In order to figure out its function (although we expect it to be the additional morpheme pointing towards whether the hierarchy is respected or not), we make the following table:\n\n-eθ- | -a:- | -i- | -ekw-\n‘I see yousg.’ | ‘I see him.’ | ‘Yousg see me.’ | ‘They see us.’\n‘I see youpl.’ | ‘I see them.’ | ‘Youpl see me.’ |\n‘We see youpl.’ | ‘We see him.’ | ‘Youpl see us.’ |\n| ‘Yousg see him.’ | |\n\nSince in the case of pronominal hierarchies the number is usually not relevant, but only the person, we can reduce the table above to only the essential data, i.e., the person of the subject and the object:\n\n-eθ- | -a:- | -i- | -ekw-\nS1 O2 | S1 O3 | S2 O1 | S3 O1\n| S2 O3 | |\n\nEven if the concept of pronoun hierarchy is unknown, the table above can help us solve most of the problem. Nevertheless, if, for example, we were asked to translate an example like S3 O2, we would not know which form of morpheme III to use.\n\nAnother possible explanation for these data is that morpheme -a:- is used if the object is in the 3rd person (independent of the subject) and -ekw- is used for S3 (independent of the object)
{"id":"book_06_06","context":"Here are some verbal forms in Tabasaran and their English translations. The verb endings have been separated from the verb root with a hyphen.\n\nbisun-čačwu | ‘we caught youpl’ | ilbicun-za | ‘I turned’\nʁapun-čwa | ‘youpl said’ | kčwuχun-vu | ‘yousg slipped’\nʁarʁun-zu | ‘I froze’ | kčwuχun-za | ‘I sleighed’\nʁiliχun-čwa | ‘youpl worked’ | uldugun-zu | ‘I got lost’\nʁip'un-za | ‘I ate’ | ursun-čwa | ‘youpl jumped’\nʁurč:wun-vazu | ‘yousg beat me’ | qergun-zu | ‘I woke up’\ndaqun-za | ‘I stretched’ | šadʁaxun-čwu | ‘youpl got happy’\nduʁmišʁaxun-zu | ‘I was born’ | ergun-vu | ‘yousg got tired’","query":"- You are given some additional verb roots:\naqun = ‘to fall’, ʁilirq'un = ‘to get scared’, uč'wun = ‘to enter’\n\n- Translate into English:\n\n- aqun-za\n\n- aqun-zu\n\n- bisun-čwazu\n\n- ʁilirq'un-ču\n\n- uč'wun-va\n\n- [blank]\n\n- You are given some additional verb roots:\nʁalrʁun = ‘to inflate’, dusun = ‘to stay’, ʁergun = ‘to escape’\n\n- Translate into Tabasaran:\n\n- ‘yousg escaped’\n\n- ‘I got scared’\n\n- ‘I beat yousg’\n\n- ‘we inflated’\n\n- ‘we stayed’\n\n- [blank]","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"It is easy to notice that the arguments are marked at the end by the morphemes č (1pl), čw (2pl), z (1sg), v (2sg). The only challenge in this problem is figuring out the choice between the vowels a and u which follow them.\n\nWe notice that in transitive phrases (which have both subject and object), the subject always has the vowel a, while the object has the vowel u. Therefore, for transitive verbs, the structure of the phrase is V-SaOu.\n\nFor intransitive verbs, we notice the examples kčwuχun-vu and kčwuχun-za, which have the same verbal stem in Tabasaran, but are translated as ‘to slip’ and ‘to sleigh’. Moreover, they differ only in the final vowel a vs. u. We notice that the difference between the two verbs is the degree of volition: ‘to slip’ is an involuntary or accidental action, while ‘to sleigh’ is used to describe a voluntary action, done on purpose, over which we have control. We therefore notice that the vowel a is used if the action is done on purpose (it represents an agentive marker), while u is used if the action is done by mistake or unintentionally (patientive marker). Moreover, this reasoning also explains the choice of vowels in the case of transitive constructions, since, in these, the subject is the same as the agent and the object is the same as the patient.\n\nTherefore, the rules are:\n\n- Structure: V-Sx(Ox)\n\n- x: a = agent marker, u = patient marker\n\n- S = O: z = 1sg, v = 2sg, č = 1pl, čw = 2pl, or as a table:\n\n| 1 | 2\nsg | z | v\npl | č | čw\n\n*Answers:\n\n1. and 2.: in both cases, the stem is ‘to fall’, but one is marked as agentive, while the other is marked as patientive. The action of falling is accidental, so example 2 is translated as ‘I fell’. In order to find the right verb for the agentive marking, we need to think of the signification of falling: you are standing and then you end up sitting or lying. Therefore, a suitable verb for it is ‘to sit’, ‘to lie’. Thus:\n\n-\n\n- ‘I sat’\n\n- ‘I fell’\n\n- ‘youpl caught me’\n\n- ‘we got scared’\n\n- ‘yousg entered’\n\n-\n\n- ʁergun-va\n\n- ʁilirq'un-zu\n\n- ʁurč:wun-zavu\n\n- ʁalrʁun-ča\n\n- dusun-ča","source":"langsci_420","problem_group_id":"langsci420:6.6","chapter":6,"chapter_title":"Verb and verb phrase","section":8,"section_title":"Verb semantics","topic":"verb morphology and argument structure","language":"Tabasaran","author":"Yakov Testelets","competition":"MSK","year":1998,"solution_tier":"detailed_worked","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"a
{"id":"book_06_07","context":"Here are two sets of Swahili verbs and their English translations, in random order:\n. | #1 | . | ‘#2’\n\n*Set I.\n\n. | Alikula. | . | ‘He ate.’\n. | Atacheza. | . | ‘He will play.’\n. | Mlifahamu. | . | ‘I eat.’\n. | Mnapika. | . | ‘I played.’\n. | Nilicheza. | . | ‘I cook.’\n. | Ninakula. | . | ‘I will cook.’\n. | Ninapika. | . | ‘They understand.’\n. | Nitapika. | . | ‘They will cook.’\n. | Tulifahamu. | . | ‘They played.’\n. | Unacheza. | . | ‘We understood.’\n. | Utapika. | . | ‘Youpl understood.’\n. | Wanafahamu.Htuku | . | ‘Youpl cook.’\n. | Watapika. | . | ‘Yousg play.’\n. | Walicheza. | . | ‘Yousg will cook.’\n\n*Set II.\n\n. | Hakucheza. | . | ‘He did not play.’\n. | Hamkupika. | . | ‘He will not cook.’\n. | Hamli. | . | ‘He will not play.’\n. | Hatacheza. | . | ‘I did not play.’\n. | Hatapika. | . | ‘I do not eat.’\n. | Hatukufahamu.Wna | . | ‘I will not fear.’\n. | Hatupiki. | . | ‘They do not fear.’\n. | Hawachi. | . | ‘They do not understand.’\n. | Hawafahamu. | . | ‘We did not understand.’\n. | Huchezi. | . | ‘We do not cook.’\n. | Sikucheza. | . | ‘Youpl do not eat.’\n. | Sili. | . | ‘Youpl did not cook.’\n. | Sitakucha. | . | ‘Yousg do not play.’","query":"- Determine the correct correspondences for each of the two sets.\n\n- Knowing that Ninatembelea means ‘I visit’ and Ninakufa means ‘I die’, translate into Swahili:\n\n- ‘Yousg visit.’\n\n- ‘Yousg do not visit.’\n\n- ‘Yousg did not visit.’\n\n- ‘Yousg will visit.’\n\n- ‘He dies.’\n\n- ‘He does not die.’\n\n- ‘He died.’\n\n- ‘He will not die.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.7. Swahili\n\n- Set I:\n\n- A.\n\n- B.\n\n- K.\n\n- L.\n\n- D.\n\n- C.\n\n- E.\n\n- F.\n\n- J.\n\n- M.\n\n- N.\n\n- G.\n\n- H.\n\n- I.\n\n-\n\n- Set II:\n\n- O.\n\n- Z.\n\n- Y.\n\n- Q.\n\n- P.\n\n- W.\n\n- X.\n\n- U.\n\n- V.\n\n- AA.\n\n- R.\n\n- S.\n\n- T.\n\n-\n\n- unatembelea\n\n- hutembelei\n\n- hukutembelea\n\n- utatembelea\n\n- anakufa\n\n- hafi\n\n- alikufa\n\n- hatakufa\n\n- x\n\nRules:\n\n- Negation: ha-\n\n- Subject:\n\n| 1 | 2 | 3\nsg | ni | u | a\npl | tu | m | wa\n\n- Tense/Negation:\n\n| Past | Present | Future\nAffirmative | li | na | ta\nNegative | ku | | ta\n\n- Stem\n\n- Other changes:\n\n- Negative marker can merge with the subject marker:\n\n- ha- + -ni- si-\n\n- ha- + -V- hV- (V = a or u)\n\n- For present negative:\n\n- last vowel of the stem (a) becomes i.\n\n- if the stem starts with the morpheme ku-, it gets dropped.","source":"langsci_420","problem_group_id":"langsci420:6.7","chapter":6,"chapter_title":"Verb and verb phrase","section":9,"section_title":"Practice problems","topic":"verb morphology and argument structure","language":"Swahili","author":"Catherine Sheard","competition":"UKLO","year":2014,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s01_p01","book_method_c06_s02_p01","book_method_c06_s03_p01","book_method_c06_s03_p02","book_method_c06_s03_p03","book_method_c06_s03_p04","book_method_c06_s04_p01","book_method_c06_s06_p01","book_method_c06_s07_p01","book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/06-Verb.tex","source_line_start":1187,"source_line_end":1247,"solution_line_start":1637,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Verb and verb phrase\nPractice problems\nverb morphology and argument structure\nThis practice problem be
{"id":"book_06_08","context":"On the following page we can see some drawings depicting Jovino's experiences on a summer morning.\n\nLater that day, Jovino met his brother Gracilian and told him what happened (the English sentences correspond to the numbers in the figures above; the Tariana sentences are given in random order):\n\n. | ‘My dog bathed.’ | . | nahã pisanane kuphe nañhasika\n. | ‘The fishpl died.’ | . | tʃãɾi kuphene diyanamahka\n. | ‘The man cooked the fishpl.’ | . | pisana dipitaka\n. | ‘The dog stole his fishsg.’ | . | mawaɾi nuhã tʃinu diwhãmahka\n. | ‘Their cats ate the fishsg.’ | . | duhã tʃãɾi mawaɾine diinusika\n. | ‘The cat bathed.’ | . | mawaɾine nahã pisana nawhãpidaka\n. | ‘The snakes bit their cat.’ | . | nuhã tʃinu dipitasika\n. | ‘The snake bit my dog.’ | . | mawaɾi tʃãɾi diwhãmahka\n. | ‘My dog died.’ | . | kuphene nayãmika\n. | ‘Her husband ate.’ | . | nuhã tʃinu diyãmipidaka\n. | ‘The snake bit the man.’ | . | duhã tʃãɾi diñhaka\n. | ‘Her husband killed the snakes.’ | . | tʃinu dihã kuphe diituka\n\n[VISUAL OMITTED: images/Tariana_main.png]","query":"- Determine the correct correspondences.\n\n- How would Jovino describe the following situations in Tariana?\n\n[VISUAL OMITTED: images/Tariana_13.png] | [VISUAL OMITTED: images/Tariana_15.png]\n13. | ‘Her cat died.’ | 15. | ‘Her husband stole their dog.’\n\n[VISUAL OMITTED: images/Tariana_14.png] | [VISUAL OMITTED: images/Tariana_16.png]\n14. | ‘The fishpl bit my cat.’ | 16. | ‘The men cooked my fishsg.’\n\n- Translate into English the following sentences and explain in which situation might Jovino utter them:\n\n- dihã tʃinune nañhamahka\n\n- pisana nahã kuphene diinuka\n\n- mawaɾi tʃãɾi diwhãsika\n\n- duhã kuphene nañhapidaka","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.8. Tariana\n\n-\n\n- G.\n\n- I.\n\n- B.\n\n- L.\n\n- A.\n\n- C.\n\n- F.\n\n- D.\n\n- J.\n\n- K.\n\n- H.\n\n- E.\n\n-\n\n- duhã pisana diyãmisika\n\n- kuphene nuhã pisana nawhãka\n\n- duhã tʃãɾi nahã tʃinu diitupidaka\n\n- tʃãɾine nuhã kuphe nayanaka\n\n-\n\n- ‘His dogs ate.’ (Jovino heard the dogs eating, tearing the meat.)\n\n- ‘The cat killed their fishpl.’ (Jovino witnessed the action.)\n\n- ‘The snake bit the man.’ (Jovino saw the bleeding wound.)\n\n- ‘Her fishpl ate.’ (Someone told that to Jovino.)\n\nRules:\n\n- Word order: SOV, Possessor-Possessed\n\n- Possessives: nuhã = 1sg, dihã = 3sg masc, duhã = 3sg fem, nahã = 3pl\n\n- -ne = plural (for nouns)\n\n- Verb:\n\n- Subject number (di- = sg, na- = pl)\n\n- Stem\n\n- Evidential markers:\n\n- = visual (Jovino was a witness.)\n\n- -mah- = sensory non-visual (hearing/smell)\n\n- -si- = inferential (Jovino sees the result of the action.)\n\n- -pida- = reportative (Jovino finds out about it from someone else.)\n\n- -ka – Alternatively, this mark can be combined with the evidential markers, resulting in: -ka, -mahka, -sika, -pidaka).","source":"langsci_420","problem_group_id":"langsci420:6.8","chapter":6,"chapter_title":"Verb and verb phrase","section":9,"section_title":"Practice problems","topic":"verb morphology and argument structure","language":"Tariana","author":"Michaela Svatošová","competition":"ČLO","year":2019,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s01_p01","book_method_c06_s02_p01","book_method_c06_s03_p01","book_method_c06_s03_p02","book_method_c06_s03_p03","book_method_c06_s03_p04","book_method_c06_s04_p01","book_method_c06_s06_p01","book_method_c06_s07_p01","book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked me
{"id":"book_06_09","context":"Here are some verbal forms in Gee and their English translations in random order:\n\n. | baipaʔmedo | . | ‘I came first.’\n. | baitenedoleʔ | . | ‘Will youpl not come first?’\n. | biʔʃunerisa | . | ‘Yousg do not go.’\n. | biʔteme | . | ‘Youpl will only talk.’\n. | biʔtemirisadoleʔ | . | ‘I only ran.’\n. | dospaʔmi | . | ‘Will I not go?’\n. | dospaʔmiduʔaleʔ | . | ‘Do youpl only run?’\n. | dosʃuneduʔa | . | ‘Yousg only talked.’\n. | dosʃuneduʔadoleʔ | . | ‘Yousg will not come.’\n. | meʔpaʔmerisaleʔ | . | ‘Do yousg talk first?’\n. | meʔʃumeduʔa | . | ‘Did I not just run?’\n. | meʔtemiduʔa | . | ‘Youpl run.’","query":"- Determine the correct correspondences.\n\n- Translate into English:\n\n- meʔpaʔmi\n\n- baiʃune\n\n- biʔʃunidoleʔ\n\n- meʔtemeleʔ\n\n- Translate into Gee:\n\n- ‘Do I talk?’\n\n- ‘Yousg will only run.’\n\n- ‘Youpl did not go first.’\n\n- ‘Do we just not come?’\n\n- ‘Yousg talk first.’\n\n- ‘Will I not run first?’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.9. Gee\n\n-\n\n- C.\n\n- F.\n\n- A.\n\n- I.\n\n- B.\n\n- L.\n\n- G.\n\n- E.\n\n- K.\n\n- J.\n\n- H.\n\n- D.\n\n-\n\n- ‘Youpl talk.’\n\n- ‘Did we not come?’\n\n- ‘I went.’\n\n- ‘Will yousg talk?’\n\n-\n\n- meʔpaʔneleʔ\n\n- dostemeduʔa\n\n- baiʃumirisado\n\n- biʔpaʔniduʔadoleʔ\n\n- meʔpaʔmerisa\n\n- dostenerisadoleʔ\n\nRules:\n\n| | Subject | | |\n(lr)3-4\nStem | Tense | Pers. | Number | Adverb | Neg | Q\n\nbiʔ (‘come’)\nbai (‘go’)\ndos (run)\nmeʔ (‘talk’)\n|\n\nʃu (past)\npaʔ (present)\nte (future)\n|\n\nn (1)\nm (2)\n|\n\ne (sg)\ni (pl)\n|\n\nduʔa (‘only’)\nrisa (‘first’)\n|\n\ndo\n|\n\nleʔ","source":"langsci_420","problem_group_id":"langsci420:6.9","chapter":6,"chapter_title":"Verb and verb phrase","section":9,"section_title":"Practice problems","topic":"verb morphology and argument structure","language":"Gee","author":"Paul Helmer","competition":"RoLO","year":2018,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s01_p01","book_method_c06_s02_p01","book_method_c06_s03_p01","book_method_c06_s03_p02","book_method_c06_s03_p03","book_method_c06_s03_p04","book_method_c06_s04_p01","book_method_c06_s06_p01","book_method_c06_s07_p01","book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/06-Verb.tex","source_line_start":1298,"source_line_end":1339,"solution_line_start":1797,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Verb and verb phrase\nPractice problems\nverb morphology and argument structure\nThis practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some verbal forms in Gee and their English translations in random order:\n\n. | baipaʔmedo | . | ‘I came first.’\n. | baitenedoleʔ | . | ‘Will youpl not come first?’\n. | biʔʃunerisa | . | ‘Yousg do not go.’\n. | biʔteme | . | ‘Youpl will only talk.’\n. | biʔtemirisadoleʔ | . | ‘I only ran.’\n. | dospaʔmi | . | ‘Will I not go?’\n. | dospaʔmiduʔaleʔ | . | ‘Do youpl only run?’\n. | dosʃuneduʔa | . | ‘Yousg only talked.’\n. | dosʃuneduʔadoleʔ | . | ‘Yousg will not come.’\n. | meʔpaʔmerisaleʔ | . | ‘Do yousg talk first?’\n. | meʔʃumeduʔa | . | ‘Did I not just run?’\n. | meʔtemiduʔa | . | ‘Youpl run.’\n- De
{"id":"book_06_10","context":"Here are some sentences in Gyarung and their English translations:\n\n. | ŋəñe no tast'on | ‘We will take care of yousg.’\n. | no ŋəñe kəust'oi | ‘Yousg will take care of us.’\n. | wəjonǯəskə ŋəñe wəst'oi | ‘They two will take care of us.’\n. | ŋənǯe ño tast'oñ | ‘We two will take care of youpl.’\n. | wəjoñek nǯo təust'ončh | ‘They will take care of you two.’\n. | wəjok ŋəñe wəst'oi | ‘He will take care of us.’\n. | wəjok no təust'on | ‘He will take care of yousg.’\n. | wəjonǯəskə ño təust'oñ | ‘They two will take care of youpl.’\n. | ŋəñe ño tast'oñ | ‘We will take care of youpl.’\n. | wəjoñek ŋənǯe wəst'očh | ‘They will take care of us two.’\n. | ño ŋa kəust'oŋ | ‘Youpl will take care of me.’\n\n. | #1 | ‘#2’\n\nFor tasks (b) and (c), you are given the following additional sentences:\n\n. | wəjonǯəskə ŋənǯe nɐrə ño t'has wəst'oi\n| ‘They two will take care of us two and youpl together.’\n. | wəjok no nɐrə wəjoñe t'has təust'oñ\n| ‘He will take care of yousg and them together.’\n. | no ŋənǯe nɐrə wəjo t'has kəust'oi\n| ‘Yousg will take care of us two and him together.’\n\nčh, ñ, ŋ, t', t'h and ǯ are consonants; ɐ and ə are vowels.","query":"- Translate into English:\n\n- no ŋa kəust'oŋ\n\n- wəjonǯəskə no təust'on\n\n- ño ŋənǯe kəust'očh\n\n[resume]\n\n- Here is a Gyarung sentence, in which a single word is missing:\n18. ŋəñe _______ nɐrə wəjo t'has tast'ončh\n\n- Fill in the missing word and translate the sentence into English.\n\n- Translate into Gyarung:\n\n- ‘I will take care of you two.’\n\n- ‘They two will take care of me.’\n\n- ‘They will take care of me and you two together.’\n\n- ‘Yousg will take care of me and him together.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.10. Gyarung\n\n-\n\n- ‘Yousg will take care of me.’\n\n- ‘They two will take care of yousg.’\n\n- ‘Youpl will take care of us two.’\n\n- 18. Missing word: no\n18. Translation: ‘We will take care of yousg and him together.’\n\n-\n\n- ŋa nǯo tast'ončh\n\n- wəjonǯəskə ŋa wəst'oŋ\n\n- wəjoñek ŋa nɐrə nǯo t'has wəst'oi\n\n- no ŋa nɐrə wəjo t'has kəust'očh\n\nRules:\n\n- Structure: S O (nɐrə O' t'has) V\nFor complex sentences, V refers to the sum of the objects (yousg + me = us two; yousg + us two = us; him + yousg = you two, etc.)\n\n- Pronouns (S = O):\n\n| sg | du | pl\n1 | ŋa | ŋənǯe | ŋañe\n2 | no | nǯo | ño\n3 | wəjok | wəjonǯəskə | wəjoñek\n\nfinal k is dropped if it's an object\n\n- Verb:\n\n- Information about 2nd person: t = O2, k = S2, w = 2\n\n- Subject and object person: a = S1–O2, ə = S3–O1, ə = S2–O1 or S3–O2\n\nNotebulbonAn alternative (and simpler) explanation can combine the first two morphemes into one. We can write: ta = S1–O2, təu = S3–O2, wə = S3–O1, kəu = S2–O1.\n\n- abc\n\n- st'o – it most likely represents the stem, possibly including the TAM marker\n\n- Information about the subject:\n\n| sg | du | pl\n1 | ŋ | čh | i\n2 | n | nčh | ñ","source":"langsci_420","problem_group_id":"langsci420:6.10","chapter":6,"chapter_title":"Verb and verb phrase","section":9,"section_title":"Practice problems","topic":"verb morphology and argument structure","language":"Gyarung","author":"Svetlana Britova","competition":"MSK","year":1998,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":2,"method_ids":["book_method_c06_s01_p01","book_method_c06_s02_p01","book_method_c06_s03_p01","book_method_c06_s03_p02","book_method_c06_s03_p03","book_method_c06_s03_p04","book_method_c06_s04_p01","book_method_c06_s06_p01","book_method_c06_s07_p01","book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book'
{"id":"book_06_11","context":"Here are some sentences in Hakhun and their English translations:\n\n. | ŋa ka kɤ ne | ‘Do I go?’\n. | nɤ ʒip tuʔ ne | ‘Did yousg sleep?’\n. | ŋabə ati lapkʰi tɤʔ ne | ‘Did I see him?’\n. | nirum kəmə nuʔrum cʰam ki ne | ‘Do we know youpl?’\n. | nɤbə ŋa lapkʰi rɤ ne | ‘Do yousg see me?’\n. | tarum kəmə nɤ lan tʰu ne | ‘Did they beat yousg?’\n. | nuʔrum kəmə ati lapkʰi kan ne | ‘Do youpl see him?’\n. | nɤbə ati cʰam tuʔ ne | ‘Did yousg know him?’\n. | tarum kəmə nirum lapkʰi ri ne | ‘Do they see us?’\n. | ati kəmə ŋa lapkʰi tʰɤ ne | ‘Did he see me?’\n\ncʰ, kʰ, ŋ, tʰ, ʒ and ʔ are consonants; ə and ɤ are vowels.","query":"- Translate into English:\n\n- nɤ ʒip ku ne\n\n- ati kəmə nirum lapkʰi tʰi ne\n\n- tarum kəmə nuʔrum cʰam ran ne\n\n- nirum kəmə tarum lan ki ne\n\n- nirum kəmə nɤ cʰam tiʔ ne\n\n- nirum ka tiʔ ne\n\n- Translate into Hakhun:\n\n- ‘Did I beat yousg?’\n\n- ‘Did they seem me?’\n\n- ‘Does he know yousg?’\n\n- ‘Do youpl sleep?’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.11. Hakhun\n\n-\n\n- ‘Do yousg sleep?’\n\n- ‘Did he see us?’\n\n- ‘Do they know youpl?’\n\n- ‘Do we beat them?’\n\n- ‘Did we know yousg?’\n\n- ‘Did we go?’\n\n-\n\n- ŋabə nɤ lan tɤʔ ne\n\n- tarum kəmə ŋa lapkʰi tʰɤ ne\n\n- ati kəmə nɤ cʰam ru ne\n\n- nuʔrum ʒip kan ne\n\nRules:\n\n- Word order: S O V X ne\n\n- S: in transitive sentences, it receives the suffix -bə (if S is 1sg or 2sg) or the word kəmə placed after the subject (otherwise).\n\n- X: hierarchy 1 > 2 > 3:\n\nt- -ʔ | past, S>O | ɤ | 1sg\ntʰ- | past, S<O | i | 1pl\nk- | present, S>O | u | 1 & 2sg\nr- | present, S>O | an | 1 & 2pl","source":"langsci_420","problem_group_id":"langsci420:6.11","chapter":6,"chapter_title":"Verb and verb phrase","section":9,"section_title":"Practice problems","topic":"verb morphology and argument structure","language":"Hakhun","author":"Peter Arkadiev","competition":"IOL","year":2018,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s01_p01","book_method_c06_s02_p01","book_method_c06_s03_p01","book_method_c06_s03_p02","book_method_c06_s03_p03","book_method_c06_s03_p04","book_method_c06_s04_p01","book_method_c06_s06_p01","book_method_c06_s07_p01","book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/06-Verb.tex","source_line_start":1403,"source_line_end":1445,"solution_line_start":1929,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Verb and verb phrase\nPractice problems\nverb morphology and argument structure\nThis practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some sentences in Hakhun and their English translations:\n\n. | ŋa ka kɤ ne | ‘Do I go?’\n. | nɤ ʒip tuʔ ne | ‘Did yousg sleep?’\n. | ŋabə ati lapkʰi tɤʔ ne | ‘Did I see him?’\n. | nirum kəmə nuʔrum cʰam ki ne | ‘Do we know youpl?’\n. | nɤbə ŋa lapkʰi rɤ ne | ‘Do yousg see me?’\n. | tarum kəmə nɤ lan tʰu ne | ‘Did they beat yousg?’\n. | nuʔrum kəmə ati lapkʰi kan ne | ‘Do youpl see him?’\n. | nɤbə ati cʰam tuʔ ne | ‘Did yousg know him?’\n. | tarum kəmə nirum lapkʰi ri ne | ‘Do they see us?’\n. | ati kəmə ŋa lapkʰi tʰɤ ne | ‘Did he see me?’\n\ncʰ, kʰ, ŋ, tʰ, ʒ
{"id":"book_06_12","context":"Here are some verbal forms in Cree (in the Plains Cree dialect) and their English translations:\n\n. | kiwīminahitin | ‘I want to make you drink.’\n. | ninanāskomik | ‘He thanks me.’\n. | kiwīminahāw | ‘You want to make him drink.’\n. | kikīwāpamin | ‘You saw me.’\n. | nikīminahikwak | ‘They made me drink.’\n. | nikananāskomāw | ‘I will thank him.’\n. | kikīnanāskomik | ‘He thanked you.’\n. | kikaminahāwak | ‘You will make them drink.’\n\nA bar above a vowel denotes length.","query":"- Translate into English:\n\n- niwīwāpamāwak\n\n- kiminahin\n\n- ninanāskomikwak\n\n- kikawāpamik\n\n- Translate into Cree:\n\n- ‘I saw you.’\n\n- ‘I want to thank him.’\n\n- ‘You will thank me.’\n\n- ‘I make them drink.’\n\n- ‘He wants to see me.’\n\n- ‘They see you.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.12. Cree\n\n-\n\n- ‘I want to see them.’\n\n- ‘You make me drink.’\n\n- ‘They thank me.’\n\n- ‘He will see you.’\n\n-\n\n- kikīwāpamitin\n\n- niwīnanāskomāw\n\n- kikananāskomin\n\n- niminahāwak\n\n- niwīwāpamik\n\n- kiwāpamikwak\n\nRules:\n\n*Option 1 – Detailed\n\n- Existence of 2sg:\n\n- ki = exists (S or O)\n\n- ni = does not exist\n\n- TAM markers:\n\n- wī = volitive (‘to want to’)\n\n- kī = past\n\n- = present\n\n- ka = future\n\n- Stem\n\n- minah = ‘to make drink’\n\n- nanāskom = ‘to thank’\n\n- wāpam = ‘to see’\n\n- Information about 3rd person\n\n- ik = S3sg\n\n- ikwak = S3pl\n\n- āw = O3sg\n\n- āwak = O3pl\n\n- in = 3 & S2sg\n\n- itin = 3 & O2sg\n\n*Option 2 – Condensed\n\n2nd pers. | TAM | Stem | S/O\n\nki\nni\n|\n\n= present\nwī = volitive\nkī = past\nka = future\n|\n\nminah = ‘to make drink’\nnanāskom = ‘to thank’\n\nwāpam = ‘to see’\n|\n\n-ik(wak) = S3sg(pl)\n-āw(ak) = O3sg(pl)\n-in = 3 & S2sg\n-itin = 3 & O2sg","source":"langsci_420","problem_group_id":"langsci420:6.12","chapter":6,"chapter_title":"Verb and verb phrase","section":9,"section_title":"Practice problems","topic":"verb morphology and argument structure","language":"Cree","author":"Ivan Derzhanski","competition":"MSK","year":2008,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s01_p01","book_method_c06_s02_p01","book_method_c06_s03_p01","book_method_c06_s03_p02","book_method_c06_s03_p03","book_method_c06_s03_p04","book_method_c06_s04_p01","book_method_c06_s06_p01","book_method_c06_s07_p01","book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/06-Verb.tex","source_line_start":1447,"source_line_end":1489,"solution_line_start":1967,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Verb and verb phrase\nPractice problems\nverb morphology and argument structure\nThis practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some verbal forms in Cree (in the Plains Cree dialect) and their English translations:\n\n. | kiwīminahitin | ‘I want to make you drink.’\n. | ninanāskomik | ‘He thanks me.’\n. | kiwīminahāw | ‘You want to make him drink.’\n. | kikīwāpamin | ‘You saw me.’\n. | nikīminahikwak | ‘They made me drink.’\n. | nikananāskomāw | ‘I will thank him.’\n. | kikīnanāskomik | ‘He thanked you.’\n. | kikaminahāwak | ‘You will make them drink.’\n\nA bar above a vowel denotes length.\n- Translate into
{"id":"book_06_13","context":"Here are some sentences in Ainu (in the Shizunai dialect) and their English translations:\n\n. | ikupa as wa isam | ‘We drank.’\n. | inkartek an wa an | ‘I was glancing.’\n. | e inkar wa an | ‘Yousg were seeing.’\n. | inu wa isam | ‘He listened.’\n. | iperepa wa oka | ‘They were feeding.’\n. | e ipe wa an | ‘Yousg were eating.’\n. | eci inuruypa wa oka | ‘Youpl were listening a lot.’\n. | cie koretek wa isam | ‘We lent yousg.’\n. | cieci nukarruypa wa isam | ‘We stared at youpl.’\n. | eun nurepa wa oka | ‘Yousg were telling us.’\n. | un etekpa wa oka | ‘He was tasting us.’\n. | ecien nutek wa an | ‘Youpl were listening to me a little.’\n. | an yaynu wa isam | ‘I thought.’\n. | an eruypa wa oka | ‘I was devouring them.’\n. | inuruypa as wa isam | ‘We listened a lot.’\n. | en e wa an | ‘They were eating me.’\n. | e yaykore wa isam | ‘Yousg gave yourself.’\n. | cieci nurepa wa oka | ‘We were telling youpl.’","query":"- Translate into English in all possible ways:\n\n- e nukarepa wa isam\n\n- e koreruy wa an\n\n- ci yaynukarpa wa oka\n\n- nuruypa wa isam\n\n- iperuy an wa isam\n\n- [blank]\n\n- Translate into Ainu:\n\n- ‘He was listening to youpl.’\n\n- ‘We ate.’\n\n- ‘Yousg were thinking a lot.’\n\n- ‘They were staring.’\n\n- ‘We were borrowing him.’\n\n- ‘I glanced at them.’\n\n- ‘Youpl fed yourselves.’\n\n- ‘Yousg were chattering to us.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.13. Ainu\n\n-\n\n- ‘Yousg showed them.’\n\n- ‘He was giving a lot to yousg.’ / ‘They were giving a lot to yousg.’ / ‘Yousg were giving a lot to him.’\nYou can also use other ways to express ‘give a lot’, such as ‘be generous’ etc.\n\n- ‘We were seeing ourselves.’\n\n- ‘He listened a lot to them.’ / ‘They listened a lot to them.’\n\n- ‘I ate a lot.’\n\n-\n\n- eci nupa wa oka\n\n- ipepa as wa isam\n\n- e yaynuruy wa an\n\n- inkarruypa wa oka\n\n- ci koretekre wa an\n\n- an nukartekpa wa isam\n\n- eci yayerepa wa isam\n\n- eun nureruypa wa oka\n\nRules:\n\n- Tense/Aspect: placed at the end of the structure.\n\n- past simple (perfective): wa isam\n\n- past continuous (imperfective):\n\n- wa an – if the verb is considered singularThe plurality of the verb is determined by the subject (for intransitive verbs) or by the object (for transitive verbs). In other words, the verb plurality follows an ergative marking – see Section morphoalign.\n\n- wa oka – if the verb is considered pluralSee previous footnote.\n\n- Pronoun markers\n\nPerson | S (Intransitive) | S (Transitive) | Object\n1 sg | -an | an- | en-\n2 sg | e- | e- | e-\n3 sg | | |\n1 pl | -as | ci- | un-\n2 pl | eci- | eci- | eci-\n3 pl | | |\n\n- The hyphen that precedes or follows the morpheme shows its position with respect to the verb (before or after). In the case of transitive verbs, if both arguments are placed before the verb, they are placed in the order subject – object and they fuse together into a single word.\n\n- Marker for verb plurality (as defined above): suffix -pa is placed after the verbal affixes (if any).\n\n- Verbal affixes:\n\n- yay- = reflexive (‘to think’ = ‘to hear yourself’)\n\n- -ruy = intensifier (can be translated by ‘a lot’, but it can also be lexicalised in the choice of verb: ‘to eat’ – ‘to devour’, ‘to see’ – ‘to stare’)\n\n- -tek = mitigator (the opposite of -ruy: ‘to eat’ – ‘to taste’, ‘to see’ – ‘to glance’)\n\n- –(r)e: causative (‘to eat’ to make someone eat = ‘to feed’, ‘to see’ to make someone see = ‘to show’). Additionally, -re e / r _\n\n- Verbal stems: Each verb has two different stems, for transitive and intransitive:\n\nEnglish | Intransitive | Transitive\n‘to eat’ | ipe | e\n‘to see’ | inkar | nukar\n‘to listen’ | inu | nu\n‘to drink’ | iku |\n‘to give’ | | koreIn reality, the stem kore (‘t
{"id":"book_06_14","context":"Here are some verbal forms in Rotokas and their English translations in random order:\n\n. | aloravirovo | . | ‘I went’\n. | iparaepa | . | ‘I threw it’\n. | ourovo | . | ‘I just talked’\n. | oraoupaveiepa | . | ‘I just devoured it’\n. | orareoveiepo | . | ‘I just confessed’\n. | reoraepo | . | ‘he moved it’\n. | reoraviroepo | . | ‘he just took it’\n. | rupupaveiepo | . | ‘we two just discussed’\n. | rururova | . | ‘we two were just swimming’\n. | vikirava | . | ‘we two were getting married’","query":"- Determine the correct correspondences, knowing that:\n\noraruruveiepa | = | ‘we two moved ourselves’\n\naloparovo | = | ‘he was just eating it’\n\n- Translate into English:\n\n- rupuraepo\n\n- ouparava\n\n- reoparoepa\n\n- oraruruveviroepa\n\n- Translate into Rotokas:\n\n- ‘I was devouring him’\n\n- ‘he was just getting married’\n\n- ‘we two just jumped’\n\n- ‘we two arrived’\n\n- ‘we two ate it’\n\n- [blank]","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.14. Rotokas\n\n-\n\n- D.\n\n- A.\n\n- G.\n\n- J.\n\n- H.\n\n- C.\n\n- E.\n\n- I.\n\n- F.\n\n- B.\n\n-\n\n- ‘I just swam’\n\n- ‘I was taking him’\n\n- ‘he was talking’\n\n- ‘we two emigrated’\n\n-\n\n- aloparavirova\n\n- oraouparoepo\n\n- oravikiveiepo\n\n- ipaveviroepa\n\n- aloveva\n\nRules:\n\n- ora- = reciprocal / reflexive (‘to move’ ‘to move oneself’; ‘to take’ to take oneself = ‘to marry’; ‘to speak’ ‘to discuss’; ‘to throw’ ‘to jump’);\n\n- Stem: -alo- = ‘to eat’, -ipa- = ‘to go’, -ou- = ‘to take’, -reo- = ‘to speak’, -rupu- = ‘to swim’, -ruru- = ‘to move’, -viki- = ‘to throw’;\n\n- -pa- = imperfect (progressive aspect);\n\n- Subject: -ra- = 1sg, -ro- = 3sg, -ve- = 1d;\n\n- -viro- = intensifier / to do till the end (‘to eat’ ‘to devour’; ‘to talk’ ‘to confess’; ‘to walk’ ‘to arrive’ = to walk till the end; ‘to move oneself’ ‘to emigrate’);\n\n- Transitivity: -v- = transitive, -(i)ep- = intransitive (ep iep / e _);\n\n- -a = far past, -o = recent past (‘just’).\n\nNotebulbonIn this language, the reflexive is considered intransitive (it takes the marker -ep-), while in Ainu (previous problem), the reflexive is considered transitive (it uses the transitive form of the stem).","source":"langsci_420","problem_group_id":"langsci420:6.14","chapter":6,"chapter_title":"Verb and verb phrase","section":9,"section_title":"Practice problems","topic":"verb morphology and argument structure","language":"Rotokas","author":"Theodor Cucu","competition":"RoLO","year":2019,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s01_p01","book_method_c06_s02_p01","book_method_c06_s03_p01","book_method_c06_s03_p02","book_method_c06_s03_p03","book_method_c06_s03_p04","book_method_c06_s04_p01","book_method_c06_s06_p01","book_method_c06_s07_p01","book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/06-Verb.tex","source_line_start":1542,"source_line_end":1585,"solution_line_start":2144,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Verb and verb phrase\nPractice problems\nverb morphology and argument structure\nThis practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some verbal forms in Rotokas and their English translations in random
{"id":"book_06_15","context":"Here are some sentences in Dinka (in the Agar dialect) and their English translations:\n\n. | dàam báng | ‘Do I catch the chief?’\n. | báng àdɔ́ɔm jò | ‘It's the chief that the dog catches.’\n. | rów àpíik wèng | ‘It's the hippo that the cow pushes.’\n. | rów àdɔ̀m jó | ‘The hippo catches the dog.’\n. | tɛ̀ɛɛt wéng | ‘Do I curse the cow?’\n. | ghɛ̂ɛn àgèl rów | ‘I protect the hippo.’\n. | tèeet kwàc ghɛ̂ɛn | ‘Does the leopard curse me?’\n. | gèeet bàng jó | ‘Does the chief cook the dog?’\n. | jó àgéeet bàng | ‘It's the dog that the chief cooks.’\n. | gèel jò kwác | ‘Does the dog protect the leopard?’\n. | jó àlɔ̀ɔk ê | ‘The dog washes him.’\n. | gɔ̀ɔɔr kwác | ‘Do I seek the leopard?’\n. | báng àgòor kwác | ‘The chief seeks the leopard.’\n. | rów àwèc wéng | ‘The hippo hits the cow.’\n. | wéng àwèc ê | ‘The cow hits him.’\n\n- ‘I cook the leopard.’\n\n- ‘Do I cook the dog?’\n\n- ‘The leopard washes the hippo.’\n\n- ‘It's the leopard that the chief washes.’\n\n- ‘Do I wash the chief?’\n\n- ‘I hit him.’\n\n- ‘Does the hippo push the dog?’\n\n- ‘Does the cow hit the dog?’\n\n- ‘The chief curses him.’\n\n- ‘It's me that the hippo protects.’\n\nVowel doubling and tripling denotes length (short a, medium aa, long aaa). The marks \"25CC\"300, \"25CC\"301, and \"25CC\"302 above the vowel denote low, high, and falling tones respectively.\n\nɛ and ɔ are vowels similar to e and o respectively, but pronounced with a more open mouth (but less open than a).\n\nThe language features two types of vowel phonation, but they were not included in the problem for simplicity.","query":"- Translate into Dinka:","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"6.15. Dinka\n\n-\n\n- ghɛ̂ɛn àgèet kwác\n\n- gɛ̀ɛɛt jó\n\n- kwác àlɔ̀ɔk rów\n\n- kwác àlɔ́ɔɔk bàng\n\n- làaak báng\n\n- ghɛ̂ɛn àwèc ê\n\n- pìik ròw jó\n\n- wèec wèng jó\n\n- báng àtèet ê\n\n- ghɛ̂ɛn àgéel ròw\n\nRules:\n\n- Sentence structure:\n\n- Active (e.g., ‘The dog catches the hippo.’) – word order S V O\n\n- Active with focused object (e.g., ‘It's the hippo that the dog catches.’) – word order O V S;\nthe tone of the subject becomes low.\n\n- Interrogative (e.g., ‘Does the dog catch the hippo?’) – word order V S O;\nthe tone of the subject becomes low.\n\n- Verb\n\n- Stems:\n\n- dɔ̀m = ‘to catch’\n\n- gèl = ‘to protect’\n\n- gèet = ‘to cook’\n\n- gòor = ‘to seek’\n\n- lɔ̀ɔk = ‘to wash’\n\n- pìk = ‘to push’\n\n- tèet = ‘to curse’\n\n- wèc = ‘to hit’\n\n- Active:\nAdd prefix à-.\n\n- Interrogative:\nVowels lengthen by one degree (short medium long).\nIf the subject is 1sg, vowel opens by one degree (i e ɛ a and u o ɔ a).\n\n- Active (focused object):\nAdd prefix à-.\nThe vowel lengthens by one degree.\nThe tone becomes high.","source":"langsci_420","problem_group_id":"langsci420:6.15","chapter":6,"chapter_title":"Verb and verb phrase","section":9,"section_title":"Practice problems","topic":"verb morphology and argument structure","language":"Dinka","author":"Michal Láznička","competition":"ČLO","year":2019,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c06_s01_p01","book_method_c06_s02_p01","book_method_c06_s03_p01","book_method_c06_s03_p02","book_method_c06_s03_p03","book_method_c06_s03_p04","book_method_c06_s04_p01","book_method_c06_s06_p01","book_method_c06_s07_p01","book_method_c06_s08_p01","book_method_c06_s08_p02"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Verb and verb phrase chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only
{"id":"book_07_01","context":"Here are some sentences in Nung and their English translations:\n\n. | Cáu ca vửhn nhahng kíhn.\n| ‘I was about to continue to eat it.’\n. | Cáu cháhn slờng páy mi?\n| ‘Do I truly want to go?’\n. | Cáu mi slày kíhn.\n| ‘I don't have to eat it.’\n. | Cáu ngám hẻht pehn tế.\n| ‘I did it like that just now.’\n. | Cáu tan đohc hảhn mưhng.\n| ‘I only see you.’\n. | Cáu vửhn nhahng bô sạhm tảhng hẻht hơn.\n| ‘I also continue to build the house alone.’\n. | Da kíhn!\n| ‘Don't eat it!’\n. | Da khải hơn!\n| ‘Don't sell the house!’\n. | Mưhn chớng ca cháhn fải khải.\n| ‘Then she truly was about to have to sell it.’\n. | Mưhn mi cháhn đày non.\n| ‘She truly can't sleep.’\n. | Mưhn náhc-thày chớng bô sạhm kíhn.\n| ‘Then she also just previously ate it.’\n. | Mưhng náhc-thày slờng tảhng páy.\n| ‘You wanted to go alone just previously.’","query":"- Translate into English:\n\n- Cáu cháhn đày non.\n\n- Da páy non!\n\n- Mưhn bô sạhm mi slờng hẻht hơn mi?\n\n- Mưhn ngám bô sạhm páy hơn.\n\n- Translate into Nung:\n\n- ‘I wasn't about to eat it just previously.’\n\n- ‘She didn't have to eat it alone like that just now.’\n\n- ‘The house truly can't eat you.’\n\n- ‘Then were you also about to go just previously?’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"- Step 1. We notice the repetition of the first word, which we can easily correlate with the subject pronoun (cáu = ‘I’, mưhng = ‘you’, mưhn = ‘she’). This is also confirmed by sentence 5 where we notice that the object has the same form (therefore, S = O).\n\nMoreover, we notice that sentences 7 and 8 both start with da and both are in the (negative) imperative mood. We therefore deduce that da is the negative imperative marker.\n\n- Step 2. In sentence 7, we have only one word remaining to be translated (kíhn). This must be the verb (with semantic content), so kíhn = ‘to eat’. This is also confirmed by sentences 1, 3, and 11.\n\nIn sentence 8, we have only two remaining words, so one must be the verb, while the other must represent the noun ‘house’. Comparing with sentence 9, we deduce that khải = ‘to sell’ and hơn = ‘house’. Moreover, we infer that for negative imperative sentences the word order is Da V O. Furthermore, since in sentence 7 ‘it’ is not translated, we infer that the 3sg pronominal object is unmarked.\n\nSince we have already identified two verbs, we continue to focus on the verbs. From sentences 2 and 12, we deduce that pȧy = ‘to go’.\n\n- Step 3. Based on sentence comparison, we can also identify most of the vocabulary, especially the adverbs: chớng = ‘then’, náhc-thày = ‘just previously’, bô sạhm = ‘also’, cháhn = ‘truly’, tảhng = ‘alone’.\n\nBased on the same principle, we can identify some modal verbs: ca = ‘was about to’, vửhn nhahng = ‘continue to’.\n\nLastly, since we already know that cháhn = ‘truly’, we are left with slờng = ‘want to’.\n\n- Step 4. Based on the above words, sentence 6 is left with only one untranslated word, mainly hẻht = ‘to build’. On the other hand, we notice the same word in sentence 4, this time meaning ‘to do’ (we can assume it also represents a verb). Therefore, we deduce that in Nung, similar to many languages derived from Latin, ‘to build a house’ = ‘to do (make) a house’.\n\nComparing sentences 3 and 10, we infer that mi is a negative marker. Nevertheless, the same marker occurs in sentence 2. Since the only thing left undiscovered in sentence 2 is the interrogation, we deduce that mi is also an interrogative marker. Moreover, we notice that if mi marks a question, it will be placed at the very end of the sentence (it is rather common that the interrogative marker is placed at the end of the sentence). Therefore, mi represents two distinct morphemes
{"id":"book_07_02","context":"The following 16 sentences represent the translations of four different English sentences into four different languages, in random order:\n\n- Bayi jugumbil baŋgul gúdaŋgu buŗan.\n\n- 'ua hi'o te tamari'i.\n\n- Bayi yaŗa buŗan.\n\n- Pol'eʔi hina nawta.\n\n- Cíq'ămqalnim peexne 'áyatne.\n\n- 'ua hi'o te 'ūrī 'i te vahine.\n\n- Háma peexne.\n\n- Bayi gagara baŋgul yaŗaŋgu buŗan.\n\n- Met'aii pol'eʔ nawta.\n\n- 'áyatnim peexne hámane.\n\n- 'ua hi'o te tamari'i 'i te 'āva'e.\n\n- Pol'eʔi nawta.\n\n- 'ua hi'o te vahine 'i te tamari'i.\n\n- Hámanim peexne hísemtuksne.\n\n- Bayi yaŗa baŋgul jugumbilŋgu buŗan.\n\n- Tsu'itsui met'ai nawta.","query":"- Group the 16 sentences into four groups, based on the language they are in.\n\n- Group the 16 sentences into four groups, based on their meaning.\n\n- Here are eight more sentences:\n\n- 'ua hi'o 'i te vahine.\n\n- Met'ai pol'eʔi nawta.\n\n- Bayi gúda buŗan.\n\n- Cíq'ămqalnim peexne.\n\n- 'ua te hi'o 'i te tamari vahine.\n\n- Bayi jugumbilŋgu baŋgul yaŗa buŗan.\n\n- Pol'eʔi pol'eʔ nawta.\n\n- Peexne 'áyatnim háma.\n\n- Out of these sentences, six are wrong. Which are these and why are they wrong?\n\n- Translate the two correct sentences from task (c) into the other three languages.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"- Step 1. The sentence grouping based on language can be easily done considering that all sentences in Language 1 (Dyirbal) contain the word buŗan, all sentences in Language 2 (Tahitian) start with 'ua hi'o, all sentences in Language 3 (Nez-Perce) contain the word peexne, and all sentences in Language 4 (Wappo) end in nawta.\n\n- Step 2. Based on the previous grouping, we have the categories:\n\n- Lang. 1 (Dyirbal)\n\n-\n\n- 1. Bayi jugumbil baŋgul gúdaŋgu buŗan.\n\n- 3. Bayi yaŗa buŗan.\n\n- 8. Bayi gagara baŋgul yaŗaŋgu buŗan.\n\n- 15. Bayi yaŗa baŋgul jugumbilŋgu buŗan.\n\n- Lang. 2 (Tahitian)\n\n-\n\n- 2. 'ua hi'o te tamari'i.\n\n- 6. 'ua hi'o te 'ūrī 'i te vahine.\n\n- 11. 'ua hi'o te tamari'i 'i te 'āva'e.\n\n- 13. 'ua hi'o te vahine 'i te tamari'i.\n\n- Lang. 3 (Nez-Perce)\n\n-\n\n- 5. Cíq'ămqalnim peexne 'áyatne.\n\n- 7. Háma peexne.\n\n- 10. 'áyatnim peexne hámane.\n\n- 14. Hámanim peexne hísemtuksne.\n\n- Lang. 4 (Wappo)\n\n-\n\n- 4. Pol'eʔi hina nawta.\n\n- 9. Met'aii pol'eʔ nawta.\n\n- 12. Pol'eʔi nawta.\n\n- 16. Tsu'itsui met'ai nawta.\n\nWe notice that in each language there is one sentence which is shorter than the others. We assume that these sentences are translations of each other, so 3 = 2 = 7 = 12.\n\nWe notice that one part of these sentences occurs in all other sentences of that language (buŗan, 'ua hi'o, peexne, nawta). This represents the verb.\n\n- Step 3. Looking at the sentences in each language, we notice that in Dyirbal bayi and buŗan appear in every sentence. Expanding on this, the structure of Dyirbal sentences is:\n\nBayi X buŗan. for short sentences\nBayi X baŋgul Y-ŋgu buŗan. for long sentences\n\nDoing the same for the other languages, we get:\n\nLanguage | Short sentence | Long sentence\nDyirbal | Bayi X buŗan. | Bayi Y baŋgul Z-ŋgu buŗan.\nTahitian | 'ua hi'o te A. | 'ua hi'o te B 'i te C.\nNez-Perce | M peexne. | N-nim peexne P-ne.\nWappo | R-i nawta. | S-i T nawta.\n\n- Step 4. By analysing the nouns in sentences 2, 3, 7, 12, we know that yaŗa = tamari'i = pol'eʔi = háma.\n\nWe notice that, in each language, these nouns appear in two more sentences. Therefore, each language has a sentence which does not contain these words, so 1 = 6 = 5 = 16.\n\nEach of these sentences contains two nouns: one which occurs in one more sentence, and one which does not occur in any other place. Checking the sentences in which that noun also occurs, we deduce 15 = 13 = 10 = 9 and jugumbil = vahine = met'ai = 'áyat.\n\nWe are left with the other noun from sentences 1=6=5=16, so gúda = 'ūrī = cíq'ămqal = tsu'itsu.\n\nNow we are left with on
{"id":"book_07_03","context":"Here are some sentences in Luiseño written in the International Phonetic Alphabet and their English translations:\n\n. | nawitmalqajwukalaqpoki:k | ‘The girl does not walk home.’\n. | jaʔaʃpolo:v | ‘The man is good.’\n. | hu:ʔunikatqajtʃipomkat | ‘The teacher is not a liar.’\n. | haxʂuxetʃiqʂuŋa:li | ‘Who hits the woman?’\n. | jaʔaʃwukalaq | ‘The man walks.’\n. | to:wqʂuʂuŋa:lihu:ʔunikat | ‘Does the teacher see the woman?’\n. | ʔiviʂuŋa:lnona:jixetʃiq | ‘This woman hits my father.’\n. | nona:jiʂuxetʃiqʔiviʂuŋa:l | ‘Does this woman hit my father?’\n. | ʔiviʂuŋa:lxetʃiqnona:ji | ‘This woman hits my father.’\n. | hu:ʔunikattʃipomkat | ‘The teacher is a liar.’\n. | ʔivihu:ʔunikatnona:jito:wq | ‘This teacher sees my father.’\n. | hu:ʔunikatʂuto:wqʂuŋa:li | ‘Does the teacher see the woman?’","query":"- Translate into English:\n\n- jaʔaʃwukalaqpoki:k\n\n- xetʃiqʂuʂuŋa:linona:j\n\n- haxʂuqajtʃipomkat\n\n- ʂuŋa:liʂuto:wqhu:ʔunikat\n\n- Translate into Luiseño. Use vertical lines to represent word spaces (a|b):\n\n- ‘Is the teacher a liar?’\n\n- ‘The teacher sees the woman.’\n\n- ‘This girl does not see my father.’\n\n- ‘Who is good?’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.3. Luiseño\n\n-\n\n- ‘The man walks home.’\n\n- ‘Does my father hit the woman?’\n\n- ‘Who is not a liar?’\n\n- ‘Does the teacher see the woman?’\n\n-\n\n- hu:ʔunikat | ʂu | tʃipomkat\n\n- hu:ʔunikat | to:wq | ʂuŋa:li\n\n- ʔivi | nawitmal | qaj | to:wq | nona:ji\n\n- hax | ʂu | polo:v\n\nNotebulbonAny other word order is accepted, as long as it follows the rules below.\n\nRules:\n\n- Flexible word order; Det. – Noun; Neg. – Verb;\n\n- The interrogative particle is placed before the verb. If the verb is first in the sentence, the particle is placed after it; i.e., the interrogative particle is always second;\n\n- -i = object marker.","source":"langsci_420","problem_group_id":"langsci420:7.3","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Luiseño","author":"Richard Hudson","competition":"UKLO","year":2012,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c07_s01_p01","book_method_c07_s02_p01","book_method_c07_s02_p02","book_method_c07_s03_p01","book_method_c07_s04_p01","book_method_c07_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/07-Syntax.tex","source_line_start":636,"source_line_end":673,"solution_line_start":1018,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Syntax\nPractice problems\nsyntax, word order, focus, and alignment\nThis practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some sentences in Luiseño written in the International Phonetic Alphabet and their English translations:\n\n. | nawitmalqajwukalaqpoki:k | ‘The girl does not walk home.’\n. | jaʔaʃpolo:v | ‘The man is good.’\n. | hu:ʔunikatqajtʃipomkat | ‘The teacher is not a liar.’\n. | haxʂuxetʃiqʂuŋa:li | ‘Who hits the woman?’\n. | jaʔaʃwukalaq | ‘The man walks.’\n. | to:wqʂuʂuŋa:lihu:ʔunikat | ‘Does the teacher see the woman?’\n. | ʔiviʂuŋa:lnona:jixetʃiq | ‘This woman hits my father.’\n. | nona:jiʂuxetʃiqʔiviʂuŋa:l | ‘Does this woman hit my father?’\n. | ʔiviʂuŋa:lxetʃiqn
{"id":"book_07_04","context":"Here are some Beja sentences and their English translations in random order. Two Beja sentences have the same English translation.\n\n. | Tak rihan. | . | ‘I saw a man that is strong.’\n. | Yaas rihan. | . | ‘I know a man that I saw.’\n. | Akra tak rihan. | . | ‘I saw a man that is small.’\n. | Dabalo yaas rihan. | . | ‘I saw a small dog.’\n. | Tak akraab rihan. | . | ‘I saw a strong man.’\n. | Tak dabaloob rihan. | . | ‘I saw a dog.’\n. | Tak akteen. | . | ‘I saw a man.’\n. | Rihane tak akteen. | . | ‘I know a man.’\n9. | Tak rihaneeb akteen. | |","query":"- Determine the correct correspondences.\n\n- Here are some more words from the Beja language with their translations:\naraw = ‘friend’, mek = ‘donkey’, kwati = ‘happy’\n\n- Translate the following sentences into Beja. If there are different ways to translate the sentence, show all the alternatives.\n\n- ‘I saw a donkey.’\n\n- ‘I saw a happy man.’\n\n- ‘I know a strong donkey.’\n\n- ‘I saw a friend that is happy.’\n\n- ‘I know a dog that is small.’\n\n- ‘I saw a donkey that I know.’\n\n- Translate the following sentences into English. One of them has a mistake. Write the correct version of this sentence.\n\n- Kwati mek rihan.\n\n- Akraab araw akteen.\n\n- Akteene yaas rihan.\n\n- Mek dabaloob akteen.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.4. Beja\n\n-\n\n- G\n\n- F\n\n- E\n\n- D\n\n- A\n\n- C\n\n- H\n\n- B\n\n- B\n\n-\n\n- Mek rihan.\n\n- Kwati tak rihan.\n\n- Akra mek akteen.\n\n- Araw kwatiib rihan.\n\n- Yaas dabaloob akteen.\n\n- Akteene mek rihan. or Mek akteeneeb rihan.\n\n-\n\n- ‘I saw a happy donkey.’\n\n- Two options:\n\n- ‘I know a friend that is strong.’ (Correct: Araw akraab akteen.)\n\n- ‘I know a strong friend.’ (Correct: Akra araw akteen.)\n\n- ‘I saw a dog that I know.’\n\n- ‘I know a donkey that is small.’\n\nRules:\n\n- Three sentence patterns (V = Verb, A = Adjective, N = Noun):\n\n- ‘I’ V ‘a’ (A) N. (A) N V.\n\n- ‘I’ V ‘a’ N ‘that is’ A. N A-V*b V, where V* refers to the last vowel of the word.\n\n- ‘I’ V ‘a’ N ‘that I’ V′. V′-e N V or N V′-eeb V.\n\n- The parts separated by hyphen (-) are suffixes.","source":"langsci_420","problem_group_id":"langsci420:7.4","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Beja","author":"Harold Somers & Richard Hudson","competition":"NACLO","year":2018,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c07_s01_p01","book_method_c07_s02_p01","book_method_c07_s02_p02","book_method_c07_s03_p01","book_method_c07_s04_p01","book_method_c07_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/07-Syntax.tex","source_line_start":675,"source_line_end":719,"solution_line_start":1048,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Syntax\nPractice problems\nsyntax, word order, focus, and alignment\nThis practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some Beja sentences and their English translations in random order. Two Beja sentences have the same English translation.\n\n. | Tak rihan. | . | ‘I saw a man that is strong.’\n. | Yaas rihan. | . | ‘I know a man that I saw.’\n. | Akra tak rihan. | . | ‘I saw a man that is small.’\n. | Dabalo yaas rihan. | . | ‘I saw a small
{"id":"book_07_05","context":"Here are some sentences in Mundari and their English translations:\n. | senkena-ñ\n| ‘I left.’\n. | koɽa-eʔ senkena\n| ‘The man left.’\n. | otere-m dubkena\n| ‘Yousg sat on the ground.’\n. | coke-ñ lelkiʔia\n| ‘I saw the frog.’\n. | pulis honko-eʔ lelkedkoa\n| ‘The policeman saw the children.’\n. | biŋ coke-ʔ huakiʔia\n| ‘The snake bit the frog.’\n. | seta pulisko-eʔ huakedkoa\n| ‘The dog bit the policemen.’\n. | biŋ setaʔre-m sabkiʔia\n| ‘Yousg caught the snake in the morning.’\n. | pulisko kumbuɽu hola-ko sabkiʔia\n| ‘The policemen caught the thief yesterday.’\n. | kuɽiko honko hature-ko ʈokoeʔkedkoa\n| ‘The women scolded the children in the village.’","query":"- Translate into English:\n\n- kumbuɽuko-ko dubkena\n\n- hola-ñ senkena\n\n- biŋko-m lelkedkoa\n\n- hon seta setaʔre-ʔ ʈokoeʔkiʔia\n\n- koɽa coke-ʔ sabkiʔia\n\n- Translate into Mundari:\n\n- ‘They left.’\n\n- ‘The woman sat on the ground.’\n\n- ‘The thieves saw the men.’\n\n- ‘The dogs bit the thief.’\n\n- ‘He caught the frogs yesterday.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.5. Mundari\n\n-\n\n- ‘The thieves sat.’\n\n- ‘I left yesterday.’\n\n- ‘Yousg saw the snakes.’\n\n- ‘The child scolded the dog in the morning.’\n\n- ‘The man caught the frog.’\n\n-\n\n- senkena-ko\n\n- kuɽi otere-ʔ dubkena\n\n- kumbuɽuko koɽako-ko lelkedkoa\n\n- setako kumbuɽu-ko huakiʔia\n\n- cokeko hola-eʔ sabkedkoa\n\nRules:\n\n- Word order: S O (Location/Time) V\n\n- Plural: -ko added to the end of the noun (before the hyphen).\n\n- Verbal suffixes:\n\n- -kena = intransitive verb;\n\n- -kedkoa = transitive verb, plural object;\n\n- -kiʔia = transitive verb, singular object.\n\n- The agreement between verb and subject is marked through a suffix separated by a hyphen. It is attached to the word before the verb. If the sentence only contains a verb (one word), it is attached to the verb instead. The forms of this suffix are:\n\n- -ñ = 1sg;\n\n- -m = 2sg;\n\n- -eʔ = 3sg (eʔ ʔ / e _);\n\n- -ko = 3pl.","source":"langsci_420","problem_group_id":"langsci420:7.5","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Mundari","author":"Peter Arkadiev","competition":"MSK","year":2014,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c07_s01_p01","book_method_c07_s02_p01","book_method_c07_s02_p02","book_method_c07_s03_p01","book_method_c07_s04_p01","book_method_c07_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/07-Syntax.tex","source_line_start":721,"source_line_end":757,"solution_line_start":1118,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Syntax\nPractice problems\nsyntax, word order, focus, and alignment\nThis practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some sentences in Mundari and their English translations:\n. | senkena-ñ\n| ‘I left.’\n. | koɽa-eʔ senkena\n| ‘The man left.’\n. | otere-m dubkena\n| ‘Yousg sat on the ground.’\n. | coke-ñ lelkiʔia\n| ‘I saw the frog.’\n. | pulis honko-eʔ lelkedkoa\n| ‘The policeman saw the children.’\n. | biŋ coke-ʔ huakiʔia\n| ‘The snake bit the frog.’\n. | seta pulisko-eʔ huakedkoa\n| ‘The dog bit the policemen.’\n. | biŋ setaʔre-m sabkiʔia\n| ‘Yousg caught the snake in t
{"id":"book_07_06","context":"Here are some sentences in Swahili and their English translations:\n\n. | Mtu ana watoto wazuri.\n| ‘The man has good children.’\n. | Mto mrefu una visiwa vikubwa.\n| ‘The long river has large islands.’\n. | Wafalme wana vijiko vidogo.\n| ‘The kings have small spoons.’\n. | Watoto wabaya wana miwavuli midogo.\n| ‘The bad children have small umbrellas.’\n. | Kijiko kikubwa kinatosha.\n| ‘The large spoon is enough.’\n. | Mwavuli una mfuko mdogo.\n| ‘The umbrella has a small bag.’\n. | Kisiwa kikubwa kina mfalme mbaya.\n| ‘The large island has a bad king.’\n. | Watu wana mifuko mikubwa.\n| ‘The men have large bags.’\n. | Viazi vibaya vinatosha.\n| ‘The bad potatoes are enough.’\n. | Mtoto ana mwavuli mkubwa.\n| ‘The child has a large umbrella.’\n. | Mito mizuri mirefu inatosha.\n| ‘The good long rivers are enough.’\n. | Mtoto mdogo ana kiazi kizuri.\n| ‘The small child has a good potato.’","query":"- Translate into Swahili:\n\n- ‘The small children have good spoons.’\n\n- ‘The long umbrella is enough.’\n\n- ‘The bad potato has a good bag.’\n\n- ‘The good kings are enough.’\n\n- ‘The long island has bad rivers.’\n\n- ‘The spoons have long bags.’\n\n- If the Swahili word for ‘the prince’ is mkuu, what do you think the word for ‘the princes’ is? Explain.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.6. Swahili\n\n-\n\n- Watoto wadogo wana vijiko vizuri.\n\n- Mwavuli mrefu unatosha.\n\n- Kiazi kibaya kina mfuko mzuri.\n\n- Wafalme wazuri wanatosha.\n\n- Kisiwa kirefu kina mito mibaya.\n\n- Vijiko vina mifuko mirefu.\n\n- wakuu. ‘Prince’ belongs to Class 1 (human), so the plural is formed by replacing the singular prefix m- with the plural wa-.\n\nRules:\n\n- Word order: SVO, Noun – Adj.\n\n- Nouns are grouped into three classes:\n\n- Human nouns: ‘child’, ‘king’, ‘man’;\n\n- Non-human nouns: ‘umbrella’, ‘river’, ‘bag’;\n\n- Non-human nouns: ‘spoon’, ‘potato’, ‘island’;\n\n- The class and number are marked by a prefix on the noun. The adjective agrees with the noun in class and number, while the verb is conjugated according to the class and number of the subject, as follows:\n\n| Class 1 | Class 2 | Class 3\n(lr)2-3(lr)4-5(lr)6-7\n| sg | pl | sg | pl | sg | pl\nNounAdj. | m- | wa- | m- | mi- | ki- | vi-\nVerb | a- | wa- | u- | i- | ki- | vi-\n\nNotebulbonThe identification of noun classes which are marked by a prefix is specific to the Bantoid languages. Depending on the language, the nouns can be classified based on certain semantic considerations (similar to the way in which classifiers work – see chap-noun), but not necessarily. For example, in this problem, there is no semantic reason to discriminate between Classes 2 and 3. Moreover, we need not find a discriminator, since there are no new words whose class we need to determine. The only distinction we need to make is that Class 1 only includes human nouns, in order to be able to differentiate between Class 1 and Class 2, which use the same singular marker.","source":"langsci_420","problem_group_id":"langsci420:7.6","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Swahili","author":"Harold Somers","competition":"NACLO","year":2011,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c07_s01_p01","book_method_c07_s02_p01","book_method_c07_s02_p02","book_method_c07_s03_p01","book_method_c07_s04_p01","book_method_c07_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"c
{"id":"book_07_07","context":"Here are some sentences in Arabic and their English translations:\n\n. | 'aḥraqa lmudarrisu lḥayawāna\n| ‘The teacher burnt the monster.’\n. | 'abda'a lqāmūsa llaðī 'aḥraqtuhu\n| ‘He created the dictionary that I burnt.’\n. | 'aðlaltu lmudarrisa llaðī 'aṣammaka\n| ‘I beat [=did beat] the teacher who surprised yousg.’\n. | 'axraǧtu lxādima llaðī 'aṣmamtahu\n| ‘I brought the servant whom yousg surprised.’\n. | 'aðalla lḥayawānu lkalba llaðī 'afazzahu\n| ‘The monster beat [=did beat] the dog which scared him.’\n\n', ð, ǧ, h, ḥ, q, ṣ, x are consonants. A bar above a vowel denotes length.","query":"- One of the sentences above is ambiguous and can be translated into English in a different way. Which sentence is it and what is the alternative translation?\n\n- Translate into English:\n\n- 'abda'tuhu\n\n- 'axraǧta lmudarrisa llaðī 'afazzaka\n\n- 'aṣamma lxādimu lkalba llaðī 'aðallahu lmudarrisu\n\n- 'aḥraqtu lḥayawāna llaðī 'aðalla lxādima\n\n- Translate into Arabic:\n\n- ‘Yousg scared the servant who surprised the monster.’\n\n- ‘The dog brought the teacher who beat [=did beat] yousg.’\n\n- ‘I burnt the dictionary that yousg created.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.7. Arabic\n\n- Sentence 5. ‘The monster beat [=did beat] the dog which he scared.’\n\n-\n\n- ‘I created him.’\n\n- ‘Yousg brought the teacher who scared yousg.’\n\n- ‘The servant surprised the dog which the teacher beat [=did beat].’\n\n- ‘I burnt the monster which beat [=did beat] the servant.’\n\n-\n\n- 'afzazta lxādima llaðī 'aṣamma lḥayawāna\n\n- 'axraǧa lkalbu lmudarrisa llaðī 'aðallaka\n\n- 'aḥraqtu lqāmūsa llaðī 'abda'tahu\n\nRules:\n\n- Word order: VSO; the relative clause is introduced by llaðī (‘which’‘who’‘whom’) and the word order inside it is identical (VSO).\n\n- Noun: receives the suffixes -u (subject) or -a (object).\n\n- Verb:\n\n- Verb stem is represented by three consonants (C1-C2-C3), while the conjugation is done through transfixes (specific to Semitic languages). For example: ‘to create’ = b—d—', ‘to scare’ = f—z—z, etc.\n\n- Subject is marked by the following transfixes:\n\n- 1sg: 'a—C1C2—a—C3—tu\n\n- 2sg: 'a—C1C2—a—C3—ta\n\n- 3sg: 'a—C1C2—a—C3—a (if C2 C3)\n\n- 3sg:'a—C1—a—C2C3—a (if C2 = C3)\n\n- Object is marked as a suffix to the verb: 2sg = -ka, 3sg = -hu only if it is not already expressed by noun.","source":"langsci_420","problem_group_id":"langsci420:7.7","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Arabic","author":"Grigory Durnovo","competition":"MSK","year":1997,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c07_s01_p01","book_method_c07_s02_p01","book_method_c07_s02_p02","book_method_c07_s03_p01","book_method_c07_s04_p01","book_method_c07_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/07-Syntax.tex","source_line_start":795,"source_line_end":828,"solution_line_start":1205,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Syntax\nPractice problems\nsyntax, word order, focus, and alignment\nThis practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some sentences in Arabic and their English translations:\n\n. | 'aḥraqa l
{"id":"book_07_08","context":"Here are some sentences in Welsh and their English translations:\n\n. | Mae tad canllaith gan ei fanon e.\n| ‘His queen has a good father.’\n. | Mae banon ganllaith gan ei blentyn e.\n| ‘His child has a good queen.’\n. | Mae brawd teg gan ei gyfaill e.\n| ‘His friend has a beautiful brother.’\n. | Mae tywysoges deg gan 'y nhad i.\n| ‘My father has a beautiful princess.’\n. | Mae cyfaill penffol gan 'y newynes i.\n| ‘My witch has a stupid friend.’\n. | Mae plentyn talentog gan 'y manon i.\n| ‘My queen has a talented child.’\n. | Mae dewynes gall gan 'y nghyfaill i.\n| ‘My friend has a wise witch.’\n\nc = ‘c’ in ‘car’.","query":"- Translate into English:\n\n- Mae banon deg gan ei frawd e.\n\n- Mae tywysoges gall gan ei ddewynes e.\n\n- Mae cyfaill canllaith gan 'y nhywysoges i.\n\n- Translate into Welsh:\n\n- ‘His father has a stupid princess.’\n\n- ‘His princess has a wise father.’\n\n- ‘My child has a talented witch.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.8. Welsh\n\n-\n\n- ‘His brother has a beautiful queen.’\n\n- ‘His witch has a wise princess.’\n\n- ‘My princess has a good friend.’\n\n-\n\n- Mae tywysoges benffol gan ei dad e.\n\n- Mae tad call gan ei dywysoges e.\n\n- Mae dewynes dalentog gan 'y mhlentyn i.\n\nRules:\n\n- Word order: Mae [O Adj.] gan S.\n\n- The possessive surrounds the noun: ‘his X’ = ei X e, ‘my X’ = 'y X i.\n\n- Noun undergoes initial consonant mutation based on context:\n\nNo possessive | 1sg poss. | 3sg poss.\nb | m | f\np | mh | b\nd | n | dd\nt | nh | d\ng | ng |\nc | ngh | g\n\nThus, we observe the following rules: for 1sg poss., voiced stops become nasals, preserving the place of articulation (b m, d n, g ng), while voiceless stop become aspirated nasals with the same place of articulation (p mh, t nh, c ngh).\n\nFor 3sg poss., we cannot deduce the transformation rule for voiced stops, but in the case of the voiceless ones, they become voiced (p b, t d, c g).\n\nThe adjective undergoes an initial consonant mutation as well. In the masculine it will have a voiceless stop as the initial consonant, while if it is feminine, it will be voiced (e.g., ‘beautiful’: teg + ‘brother’, deg + ‘princess’).","source":"langsci_420","problem_group_id":"langsci420:7.8","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Welsh","author":"Timur Maisak","competition":"MSK","year":1998,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c07_s01_p01","book_method_c07_s02_p01","book_method_c07_s02_p02","book_method_c07_s03_p01","book_method_c07_s04_p01","book_method_c07_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/07-Syntax.tex","source_line_start":830,"source_line_end":862,"solution_line_start":1245,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Syntax\nPractice problems\nsyntax, word order, focus, and alignment\nThis practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some sentences in Welsh and their English translations:\n\n. | Mae tad canllaith gan ei fanon e.\n| ‘His queen has a good father.’\n. | Mae banon ganllaith gan ei blentyn e.\n| ‘His child has a good queen.’\n. | Mae brawd teg gan ei gyfaill e.\n| ‘His friend has a beautiful brother.’\n. | Mae tywysoges deg gan 'y nhad i.\n| ‘My father ha
{"id":"book_07_09","context":"Here are some sentences in Tadaksahak and their English translations:\n\n. | aɣagon cidi\n| ‘I swallowed the salt.’\n. | atezelmez hamu\n| ‘He will have the meat swallowed (by someone).’\n. | atedini a\n| ‘He will take it.’\n. | hamu anetubuz\n| ‘The meat was not taken.’\n. | jifa atetukuš\n| ‘The corpse will be taken out.’\n. | amanokal anešukuš cidi\n| ‘The chief didn't have the salt taken out.’\n. | aɣakaw hamu\n| ‘I took out the meat.’\n. | itegzem\n| ‘They were slaughtered.’\n. | aɣasezegzem a\n| ‘I'm not having him slaughtered.’\n. | anešišu aryen\n| ‘He didn't have the water drunk (by anybody).’\n. | feji abnin aryen\n| ‘The sheep is drinking the water.’\n. | idumbu feji\n| ‘They slaughtered the sheep.’\n. | cidi atetegmi\n| ‘The salt will be looked for.’\n. | amanokal abtuswud\n| ‘The chief is being watched.’\n. | cidi asetefred\n| ‘The salt is not being gathered.’\n. | amanokal asegmi i\n| ‘The chief had them looked for.’\n\nʒ = ‘s’ in ‘vision’, š = ‘sh’ in ‘shop’, ɣ is a consonant.","query":"- Translate into English:\n\n- aryen anetišu\n\n- aɣasuswud feji\n\n- cidi atetelmez\n\n- asedini jifa\n\n- If the stem of the verb ‘to walk’ is iʒuwenket, translate into Tadaksahak:\n\n- ‘He is having the water taken.’\n\n- ‘I'm having them walked.’\n\n- ‘The chief did not drink the water.’\n\n- ‘The salt was not looked for.’\n\n- ‘He will have the salt gathered.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.9. Tadaksahak\n\n-\n\n- ‘The water was not drunk.’\n\n- ‘I had the sheep watched.’\n\n- ‘The salt will be swallowed.’\n\n- ‘He is not taking the corpse.’\n\n-\n\n- abzubuz aryen\n\n- aɣabʒiʒuwenket i\n\n- amanokal anenin aryen\n\n- cidi anetegmi\n\n- atesefred cidi\n\nRules:\n\n- Word order: SVO. If the subject is a pronoun, it is omitted;\n\n- Verb: S–T–V–R;\n\n- S = Subject: a- = 3sg, i- = 3pl, aɣa- = 1sg;\n\n- T = Tense (combined with negation):\n\n| Past | Present | Future\nAffirmative | | -b- | -te-\nNegative | -ne- | -se- |\n\n- V = Voice:\n\n- = active;\n\n- -t- = passive;\n\n- -š- / -z- / -s- / -ʒ- = causative (‘to have someone do...’). If the stem contains any of these four sounds, the same sound is used here. Otherwise, -s- is used.\n\n- R = stem; the stem has two suppletive forms: one for active and another one for passive and causative.","source":"langsci_420","problem_group_id":"langsci420:7.9","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Tadaksahak","author":"Bozhidar Bozhanov","competition":"UKLO","year":2011,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c07_s01_p01","book_method_c07_s02_p01","book_method_c07_s02_p02","book_method_c07_s03_p01","book_method_c07_s04_p01","book_method_c07_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/07-Syntax.tex","source_line_start":864,"source_line_end":912,"solution_line_start":1296,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Syntax\nPractice problems\nsyntax, word order, focus, and alignment\nThis practice problem belongs to the book's Syntax chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some sentences in Tadaksahak and their English translations:\n\n. | aɣagon cidi\n| ‘I swallowed the salt.’\n. | atezelmez hamu\n| ‘He will ha
{"id":"book_07_10","context":"Here are some sentences in Sandawe and their English translations:\n\n. | !'ìnéỳsù kòŋkórìsà xéʔé̥wáá\n| ‘A hunterf brought roosters.’\n. | thíméỳsù kókósà ǁ'èésú\n| ‘A cookf skinned a hen.’\n. | múk'ùmè kókó xéʔé̥wáátshú\n| ‘A cow didn't bring hens.’\n. | kòŋkórì múk'ùmèʔà khàású\n| ‘Roosters hit [=did hit] a cow.’\n. | !'ìnéỳsò k'ámbà khàáyétshógé\n| ‘Apparently, hunters didn't hit a bull.’\n. | thíméỳ !'ìnéỳ xééyétshèégé\n| ‘Apparently, a cookm didn't bring a hunterm.’\n. | !'ìnéỳsò kókógéʔà ǁ'èésú\n| ‘Apparently, hunters skinned a hen.’\n. | !'ìnéỳsò kókóʔà khǎʔḁ́wáá\n| ‘Hunters hit [=did hit] hens.’\n. | kòŋkórì !'ìnéỳà xééyé\n| ‘A rooster brought a hunterm.’\n. | thíméỳsù kókó khǎʔḁ́wáátshúgé\n| ‘Apparently, a cookf didn't hit hens.’\n. | !'ìnéỳ thíméỳsògéà khàáʔíŋ\n| ‘Apparently, a hunterm hit [=did hit] cooks.’\n. | !'ìnéỳsò thíméỳsò ǁ'èéʔíntshó\n| ‘Hunters didn't skin cooks.’\n. | kòŋkórì !'ìnéỳsò xééʔíntshó\n| ‘Roosters didn't bring hunters.’\n. | thíméỳ kòŋkórì khǎʔḁ́wáátshèé\n| ‘A cookm didn't hit roosters.’\n\nGiven below are some more words in Sandawe and their English translations:\n\nŋ!àméỳ = ‘blacksmithm’\nbálóó = ‘to herd’\nthéká = ‘leopard (any gender)’\n\nx, th, tsh, kh, k', ŋ, ŋ!, ʔ, !', and ǁ' are consonants. The marks \"25CC\"301, \"25CC\"300, and \"25CC\"30C above a vowel denote high, low and rising (low high) tones, respectively.\nA circle under a vowel (e.g., ḁ) indicates a devoiced vowel.\nThe subscripts m and f refer to masculine and feminine, respectively.","query":"- Translate into English:\n\n- thíméỳ kòŋkórìgéà ǁ'èéyé\n\n- ŋ!àméỳsù thíméỳsùsà xéésú\n\n- k'ámbà théká khàásútshógé\n\n- múk'ùmè !'ìnéỳsòsà bálóóʔíŋ\n\n- Translate into Sandawe:\n\n- ‘Cooks herded hens.’\n\n- ‘Apparently, a blacksmithf didn't skin leopards.’\n\n- ‘A leopardf didn't herd a rooster.’\n\n- ‘Apparently, a bull didn't bring cooks.’\n\n- ‘Apparently, a hunterm brought blacksmiths.’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.10. Sandawe\n\n-\n\n- ‘Apparently, a cookm skinned a rooster.’\n\n- ‘A blacksmithf brought a cookf.’\n\n- ‘Apparently, bulls didn't hit a leopardf.’\n\n- ‘A cow herded hunters.’\n\n-\n\n- thíméỳsò kókóʔà bálóʔó̥wáá\n\n- ŋ!àméỳsù théká ǁ'ěʔé̥wáátshúgé\n\n- théká kòŋkórì bálóóyétshú\n\n- k'ámbà thíméỳsò xééʔíntshèégé\n\n- !'ìnéỳ ŋ!àméỳsògéà xééʔíŋ\n\nRules:\n\n- Human nouns receive the suffixes: (masc. sg), -sù (fem. sg), -sò (pl);\n\n- Sentence structure:\n\n- Affirmative: S + O—(gé)—XS + V—XO;\n\n- Negative: S + O + V—XO—YS—(gé).\n\n- -gé- marks ‘Apparently’ (non-witnessed evidential);\n\n- XS and YS agree with the subject, while XO agrees with the object, as follows:\n\n| XS | YS | XO\nsg masc. | -à | -tshèé | -yé\nsg fem. | -sà | -tshú | -sú\npl human | -ʔà | -tshó | -ʔín-ʔín -ʔíŋ / _ # (alternatively, in affirmative sentences).\npl non-human | -ʔà | -tshó | -ʔwááV́V́ + -ʔwáá V́ʔV̥́wáá and V̀V́ + -ʔwáá V̌ʔV̥́wáá.\n\nThe morpheme -ʔwáá attracts tone change if V2 has a high tone. In this case, V2 will get devoiced and, if V1 has low tone, it will become rising.","source":"langsci_420","problem_group_id":"langsci420:7.10","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Sandawe","author":"Shen-Chang Huang","competition":"APLO","year":2021,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_id
{"id":"book_07_11","context":"Here are some sentences in Burushaski and their English translations:\n\n. | khue gušiŋanc uwaran.\n| ‘These women will get tired.’\n. | ise ṣiqar iγurci.\n| ‘That wasp will drown.’\n. | biṭayue amin dasin musarkan?\n| ‘Which girl will the shamans let in?’\n. | γeniṣ muwalo.\n| ‘The queen will fall.’\n. | ue dasiwance šugulimuc usarkan.\n| ‘Those girls will let the friendsf in.’\n. | guse γurqune ṣiqarišo uγarki.\n| ‘This frog will catch the wasps.’\n. | qhudaae ice j̣akuyo uyeeci.\n| ‘The god will see those donkeys.’\n. | khine hilese belišo uγarki.\n| ‘This boy will catch the rams.’\n. | hoolalase amic talabuudomuc uyeeci?\n| ‘Which spiders will the butterfly see?’\n. | ue thamišue γeniṣanc uyaranan.\n| ‘Those kings will deceive the queens.’\n. | hilešue šugulo isarkan.\n| ‘The boys will let the friendm in.’\n. | γaṣepe khine biṭan iyarani.\n| ‘The magpie will deceive this shaman.’\n\nγ, j̣, ŋ, ṣ, š, and ṭ are consonants. The subscripts m and f refer to masculine and feminine, respectively.","query":"- Translate into English:\n\n- ice belišo uwalan.\n\n- qhudaamuce tham iyaranan.\n\n- talabuudue khine gus muyeeci.\n\n- amin guse γurquyo uγarko?\n\n- Translate into Burushaski:\n\n- ‘Those shamans will drown.’\n\n- ‘Which magpies will the women catch?’\n\n- ‘The kings will see these butterflies.’\n\n- ‘Which friendm will let the boys in?’\n\n- ‘That boy will deceive the friendf.’\n\n- ‘The queen will let that girl in.’\n\n- ‘This girl will see the friendsm.’\n\n- ‘The wasp will deceive that frog.’\n\n- ‘Which donkey will get tired?’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"7.11. Burushaski\n\n-\n\n- ‘Those rams will fall.’\n\n- ‘The gods will deceive the king.’\n\n- ‘The spider will see this woman.’\n\n- ‘Which woman will catch the frogs?’\n\n-\n\n- ue biṭayo uγurcan.\n\n- gušiŋance amic γaṣepišo uγarkan?\n\n- thamišue guce hoolalašo uyeecan.\n\n- amin šugulue hilešo usarki?\n\n- ine hilese šuguli muyarani.\n\n- γeniṣe ine dasin musarko.\n\n- khine dasine šugulomuc uyeeco.\n\n- ṣiqare ise γurqun iyarani.\n\n- amis j̣akun iwari?\n\nRules:\n\n- Word order: SOV, Modifier – Noun\n\n- Modifiers:\n\n| Human | Non-human\n(lr)2-3(lr)4-5\n| sg | pl | sg | pl\n‘this’ | khine | khue | guse | guce\n‘that’ | ine | ue | ise | ice\n‘which’ | amin | | amis | amic\n\n- Noun plural:\n\n- Masculine (human) and non-human nouns:\n\n- if singular ends in n: -n -yo;\n\n- if singular ends in s: -s -šo;\n\n- if singular ends in another consonant: C -Cišo;\n\n- if singular ends in a vowel: -V -Vmuc;\n\n- Feminine (human):\n\n- if singular ends in a vowel: -V -Vmuc;\n\n- if singular ends in a consonant – nouns behave irregularly (they receive the suffix -anc, but some consonant alterations may occur). Nevertheless, the problem does not require us to infer any plural form from this category;\n\n- Ergative marker: -e added after the plural marker; o u / _ e;\n\n- Verb: receives a prefix and a suffix. The prefix agrees with the absolutive argument of the verb, while the suffix agrees with the nominative argument of the verb, as follows:\n\n| non-human / | |\n| masc. human | fem. human | plural\nprefix | i- | mu- | u-\nsuffix | -i | -o | -an","source":"langsci_420","problem_group_id":"langsci420:7.11","chapter":7,"chapter_title":"Syntax","section":6,"section_title":"Practice problems","topic":"syntax, word order, focus, and alignment","language":"Burushaski","author":"Danylo Mysak","competition":"UkrLO","year":2019,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c07_s01_p01","book_method_c07_s02_p01","book_method_c07_s02_p02","book_method_c07_s03_p01","book_method_c07_s04_p01","book_method_c07_s05_p01"],"method_link_type":"chapter_candidate_methods
{"id":"book_08_01","context":"Here are some words and phrases in Lango and their English translations in random order:\n\ndyè ɔ̀t, dyè tyɛ̀n, gìn, gìn wìc, ɲíg, ɲíg wàŋ, ɔ̀t cɛ̀m, wìc ɔ̀t\n‘eyeball’, ‘grain’, ‘roof’, ‘garment’, ‘floor’, ‘restaurant’, ‘sole of foot’, ‘hat’","query":"- Determine the correct correspondences.\n\n- Translate into English: cɛ̀m and dyè.\n\n- Translate into Lango: ‘window’.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"- Step 1. Create the graph (see fig:Lango-step1). We can notice that ɔ̀t appears three times, so we can start with it.\n\nCaption: Graph corresponding to Step 1.\n\n(1) at (0,0) dyè;\n(2) at (4,0) ɔ̀t;\n(3) at (8,0) wìc;\n(4) at (4,-2) cɛ̀m;\n(1) – (2) node[anchor=south,inner sep=3pt,midway] ;\n(3) – (2) node[anchor=south,inner sep=3pt,midway] ;\n(2) – (4) node[inner sep=3pt,midway,right] ;\n\nWe continue to connect each word and obtain the final graphs in fig:Lango-step2. Note that for this problem there are two independent subgraphs.\n\nCaption: Complete graph of the words and phrases in Lango.\n\n2(1) at (0,0) dyè;\n(2) at (2.75,0) ɔ̀t;\n(3) at (5.5,0) wìc;\n(4) at (2.75,-2) cɛ̀m;\n(5) at (-2.75,0) tyɛ̀n;\n(6) at (8.25,0) gìn;\n(7) at (5.5,-2) ɲíg;\n(8) at (8.25,-2) wàŋ;\n(1) – (2) node[anchor=south,inner sep=3pt,midway] ;\n(3) – (2) node[anchor=south,inner sep=3pt,midway] ;\n(2) – (4) node[anchor=south,inner sep=3pt,midway,right] ;\n(1) – (5) node[anchor=south,inner sep=3pt,midway] ;\n(6) – (3) node[anchor=south,inner sep=3pt,midway] ;\n(7) – (8) node[anchor=south,inner sep=3pt,midway] ;\n\nNotice that we have two separate subgraphs: the first one, which has a T-shape, and the second one which only has two nodes. Moreover, notice that the words gìn and ɲíg are underlined, meaning they also appear as single words in the corpus.\n\n- Step 2. It is time to try and form a partial graph using the English words. At first sight, we can certainly correlate the word pairs: ‘roof’ – ‘floor’ (top part and bottom part of the room/house), as well as ‘hat’ – ‘roof’ (the covering of the head and of the room/house). Since the word ‘garment’ is already given, we can consider ‘hat’ to be the ‘head garmentthe garment of the top part’. Thus, we can create the partial graph shown in fig:Lango-EN-partial.\n\n(1) at (0,0) bottom;\n(2) at (3,0) house;\n(3) at (6,0) top;\n(4) at (9,0) garment;\n(1) – (2) node[anchor=south,inner sep=3pt,midway] floor;\n(2) – (3) node[anchor=south,inner sep=3pt,midway] roof;\n(4) – (3) node[anchor=south,inner sep=3pt,midway] hat;\n\nCaption: Partial graph of the English words.\n\nWe need to notice, in this case, that the connection between the nodes was not done using arrows, but lines, since the order of the constituents is not relevant.\n\nComparing the two graphs (Figures fig:Lango-step2 and fig:Lango-EN-partial), notice that the partial graph contains a chain of four words (‘garment’ – ‘top’ – ‘house’ – ‘bottom’), and one of the end-nodes is underlined (meaning it is given in the corpus). Looking at the complete graph of Lango words (fig:Lango-step2), we know for sure that the partial graph of English words (fig:Lango-EN-partial) cannot be part of the small subgraph, since it only has two nodes. In the big subgraph (the T-shaped one), only one node is underlined, namely gìn. Thus, we deduce that gìn = ‘garment’, and the four nodes (‘garment’ – ‘top’ – ‘house’ – ‘bottom’) can either be gìn – wìc – ɔ̀t – dyè, or gìn – wìc – ɔ̀t – cɛ̀m. Either way, the first three words are identical, so we deduce that gìn = ‘garment’, gìn wìc = ‘hat’, wìc = ‘top’, wìc ɔ̀t = ‘roof’, ɔ̀t = ‘house’. Following this, the proposed graph becomes the one shown in fig:Lango-step3.\n\n2(1) at (0,0) dyè;\n(2) at (2.75,0) ‘house’;\n(3) at (5.5,0) ‘top’;\n(4) at
{"id":"book_08_02","context":"Here are some words and phrases in Guaraní and their English translations in random order:\n\n. | jaxy | . | ‘water’\n. | jaxy-tata | . | ‘brave’\n. | jaxy endy | . | ‘thumb’\n. | kuã guaxu | . | ‘liver, heart’\n. | kuã regua | . | ‘fire’\n. | py'a | . | ‘smoke’\n. | py'a guaxu | . | ‘pregnant’\n. | tata | . | ‘ring (jewellery)’\n. | tata endy | . | ‘moonlight’\n. | tata rataxĩ | . | ‘firelight’\n. | ye guaxu | . | ‘moon’\n. | yvy rataxĩ | . | ‘good soil’\n. | yvy porã | . | ‘dust’\n. | yy | . | ‘star’","query":"- Determine the correct correspondences.\n\n- Translate into English:\n\n- guaxu\n\n- porã\n\n- rataxĩ\n\n- regua\n\n- ye\n\n- Translate into Guaraní:\n\n- ‘calm, relaxed’\n\n- ‘fog’\n\n- [blank]","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"First notice that in English we have words referring to organs (‘liver, heart’) and emotions (‘brave’ and, in task (c), ‘calm’). So we can expect that these two are connected. Nevertheless, the first step is constructing the graphs for the Guaraní words (see fig:Guarani-step1).\n\nCaption: Complete graph of the words and phrases in Guaraní.\n\n(1) at (0,0) endy;\n(2) at (2,0) tata;\n(3) at (4,0) rataxĩ;\n(4) at (6,0) yvy;\n(5) at (8,0) porã;\n(6) at (1,2) jaxy;\n(2) – (1) node[anchor=south,inner sep=3pt,midway] ;\n(2) – (3) node[anchor=south,inner sep=3pt,midway] ;\n(4) – (3) node[anchor=south,inner sep=3pt,midway] ;\n(5) – (4) node[anchor=south,inner sep=3pt,midway] ;\n(6) – (1) node[anchor=south,inner sep=3pt,midway,left] ;\n(6) – (2) node[anchor=south,inner sep=3pt,midway,right] ;\n\n(1) at (0,0) ye;\n(2) at (2,0) guaxu;\n(3) at (4,0) kuã;\n(4) at (6,0) regua;\n(5) at (3,-1.5) yy;\n(6) at (2,2) py'a;\n(1) – (2) node[anchor=south,inner sep=3pt,midway] ;\n(3) – (2) node[anchor=south,inner sep=3pt,midway] ;\n(3) – (4) node[anchor=south,inner sep=3pt,midway] ;\n(6) – (2) node[anchor=south,inner sep=3pt,midway,right] ;\n\nNotebulbonIn Appendix appendix:3 I present a hand-drawn graph in order to show what such a graph might look like in reality, when solving a problem.\n\nWe notice that, in Guaraní, we have three independent subgraphs. For the partial graph of English words, we have the words: ‘moon’, ‘fire’, ‘moonlight’, and ‘firelight’. These can be arranged in a graph as shown in fig:Guarani-step2.\n\n(1) at (2,0) ‘fire’;\n(2) at (4,0) ‘light’;\n(3) at (6,0) ‘moon’;\n(1) – (2) node[anchor=south,inner sep=3pt,midway] ;\n(2) – (3) node[anchor=south,inner sep=3pt,midway] ;\n\nCaption: Partial graph of the English words ‘moon’, ‘fire’, ‘moonlight’, and ‘firelight’.\n\nThis is an ideal partial graph since we have two base words (nodes), ‘fire’, and ‘moon’, which are found in the corpus and are both connected to the same word (‘light’). According to the Guaraní graph in fig:Guarani-step1, the only two nodes close to one another that are found in the corpus (are underlined) are jaxy and tata, and both of them are connected to the word endy. Thus, we deduce that endy = ‘light’, and {jaxy, tata} = {‘fire’, ‘moon’}By this notation, we mean that jaxy and tata correspond to ‘fire’ and ‘moon’, but we do not know which is which.. In order to determine which is which, we notice that jaxy is not connected to anything else, while tata is further connected to rataxĩ. In English, we have the word ‘smoke’ which is clearly connected to ‘fire’, so tata = ‘fire’ and jaxy = ‘moon’. Moreover, from the graph, we notice that jaxy and tata combine with one another, thus, in English, we need to find a word formed by combining the words ‘moon’ and ‘fire’. The only one which is semantically close to that is ‘star’ ( = ‘fire moon’). Adding this information, our graph will look like that in fig:Guarani-step3.\n\nCaption: Partially solved graph.\n\n(1) at (0,0) endy\n(‘
{"id":"book_08_03","context":"Here are some words in Basque and their English translations in random order:\n\nigogailu, artzain, lantegi, lantalde, bizitegi, taldekide, erizain, garbigailu, ikastalde,\nbizikide, garbitegi, ikaskide, lankide, eritegi, artalde\n‘classmate’, ‘flatmate’, ‘flock of sheep’, ‘crew’, ‘elevator’, ‘clinic’, ‘factory’, ‘nurse’, ‘home’, ‘shepherd’,\n‘wash-house’, ‘colleague’, ‘washing machine’, ‘team member’, ‘class (of students)’","query":"- Determine the correct correspondences.\n\n- How is the word artalde different from the other Basque words?\n\n- Translate the word ‘sheep-pen’ into Basque, knowing that it has the same feature as the word artalde.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"8.3. Basque\n\n-\n\n- igogailu = ‘elevator’\n\n- artzain = ‘shepherd’\n\n- lantegi = ‘factory’\n\n- lantalde = ‘crew’\n\n- bizitegi = ‘home’\n\n- taldekide = ‘team member’\n\n- erizain = ‘nurse’\n\n- garbigailu = ‘washing machine’\n\n- ikastalde = ‘class of (students)’\n\n- bizikide = ‘flatmate’\n\n- garbitegi = ‘wash-house’\n\n- ikaskide = ‘classmate’\n\n- lankide = ‘colleague’\n\n- eritegi = ‘clinic’\n\n- artalde = ‘flock of sheep’\n\n- -t + t- -t- (if a morpheme which ends in t is joined to another that starts with t, one of the two t's is dropped).\n\n- ‘sheep-pen’ = artegi\n\nRules:\n\nEach Basque word is composed of two morphemes (the first one shows the semantic field, while the other the category):\n\n1st morpheme | 2nd morpheme\nigo- = ‘to lift’ |\nlan- = ‘to work’ | -zain = ‘worker’In reality, a more accurate translation would be ‘keeper’.\nbizi- = ‘to live’ | -tegi = ‘place’\nart- = ‘sheep’ | -talde = ‘collective’\neri- = ‘sick’ | -kide = ‘member’\ngarbi- = ‘to wash’ | -gailu = ‘machine’\nikas- = ‘to learn’ |","source":"langsci_420","problem_group_id":"langsci420:8.3","chapter":8,"chapter_title":"Semantics","section":3,"section_title":"Practice problems","topic":"semantics and graph-based matching","language":"Basque","author":"Natalia Zaika","competition":"MSK","year":2012,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c08_s01_p01","book_method_c08_s02_p01","book_method_c08_s02_p02","book_method_c08_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/08-Semantics.tex","source_line_start":547,"source_line_end":562,"solution_line_start":754,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Semantics\nPractice problems\nsemantics and graph-based matching\nThis practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words in Basque and their English translations in random order:\n\nigogailu, artzain, lantegi, lantalde, bizitegi, taldekide, erizain, garbigailu, ikastalde,\nbizikide, garbitegi, ikaskide, lankide, eritegi, artalde\n‘classmate’, ‘flatmate’, ‘flock of sheep’, ‘crew’, ‘elevator’, ‘clinic’, ‘factory’, ‘nurse’, ‘home’, ‘shepherd’,\n‘wash-house’, ‘colleague’, ‘washing machine’, ‘team member’, ‘class (of students)’\n- Determine the correct correspondences.\n\n- How is the word artalde different from the other Basque words?\n\n- Translate the word ‘sheep-pen’ into Basque, knowing that it has the same feature as the word artalde."}
{"id":"book_08_04","context":"Here are some words in Turkish and their English translations in random order:\n\ngözlemci, döndürmek, gündöndü, gözlükcü, şarkıcı, çocukluk, gözlemek, pazar, pazartesi, cumartesi, güneşli\n‘Saturday’, ‘Sunday’, ‘Monday’, ‘observer’, ‘singer’, ‘to observe’, ‘to rotate’, ‘sunny’, ‘sunflower’, ‘optician’, ‘childhood’","query":"- Determine the correct correspondences.\n\n- Translate into Turkish:\n\n- ‘observation’\n\n- ‘child’\n\n- ‘the state of being a singer’\n\n- ‘spectacles’\n\n- ‘Friday’\n\n- [blank]","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"8.4. Turkish\n\n- The morphemes have been separated by a hyphen (-) and their meaning is given, in order, between brackets:\n\ngöz-lem-ci | ‘observer’ | (eye + abstract-noun + agent-marker)\ngöz-lük-cü | ‘optician’ | (eye + state-of-being + agent-marker)\ngöz-le(m)-mek | ‘to observe’ | (eye + abstract-noun + verb-marker)\ndöndür-mek | ‘to rotate’ | (rotation + verb-marker)\ngün(eş)-li | ‘sunny’ | (sun + adjective-marker)\ngün-döndü | ‘sunflower’ | (sun + rotation)\nşarkı-cı | ‘singer’ | (song + agent-marker)\nçocuk-luk | ‘childhood’ | (child + state-of-being)\npazar | ‘Sunday’ |\npazar-tesi | ‘Monday’ | (Sunday + tomorrow)\ncumar-tesi | ‘Saturday’ | (Friday + tomorrow)\n\nNotebulbonSince Turkish is a language displaying vowel harmony, the form of the suffixes will differ depending on the word (ci/cı/cü or lük/luk).\n\n-\n\n- gözlem\n\n- çocuk\n\n- şarikcilik\n\n- gözlük\n\n- cumarIn reality, it is cuma, but cumar is the answer as can be deduced from the given data.","source":"langsci_420","problem_group_id":"langsci420:8.4","chapter":8,"chapter_title":"Semantics","section":3,"section_title":"Practice problems","topic":"semantics and graph-based matching","language":"Turkish","author":"Monojit Choudhury","competition":"PLO","year":2014,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c08_s01_p01","book_method_c08_s02_p01","book_method_c08_s02_p02","book_method_c08_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/08-Semantics.tex","source_line_start":566,"source_line_end":586,"solution_line_start":800,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Semantics\nPractice problems\nsemantics and graph-based matching\nThis practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words in Turkish and their English translations in random order:\n\ngözlemci, döndürmek, gündöndü, gözlükcü, şarkıcı, çocukluk, gözlemek, pazar, pazartesi, cumartesi, güneşli\n‘Saturday’, ‘Sunday’, ‘Monday’, ‘observer’, ‘singer’, ‘to observe’, ‘to rotate’, ‘sunny’, ‘sunflower’, ‘optician’, ‘childhood’\n- Determine the correct correspondences.\n\n- Translate into Turkish:\n\n- ‘observation’\n\n- ‘child’\n\n- ‘the state of being a singer’\n\n- ‘spectacles’\n\n- ‘Friday’\n\n- [blank]"}
{"id":"book_08_05","context":"Here are some words and phrases in Chinese and their English translations in random order:\n\nhóng, míngbái, báishì, huáng, hēishì, shìqing huángle, yăn, yănhóng, hóngshì, hóngyán, hēihuà, báiyăn, báihuà, hēibái fēnmíng, yán, hēibái, huángle\n‘encoding’, ‘black-and-white’, ‘face’, ‘yellow’, ‘funeral’, ‘bankruptcy’, ‘to clarify’, ‘to dislike’, ‘young woman’, ‘failure’, ‘decoding’, ‘wedding’, ‘eye’, ‘black market’, ‘it's written in black and white’, ‘jealousy’, ‘red’","query":"- Determine the correct correspondences.\n\n- The word báishì can have two meanings in Chinese, although only one is reflected in the correspondences above. What is the other meaning?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"8.5. Chinese\n\n-\n\n- hóng = ‘red’\n\n- míngbái = ‘to clarify’\n\n- báishì = ‘funeral’\n\n- huáng = ‘yellow’\n\n- hēishì = ‘black market’\n\n- shìqing huángle = ‘bankruptcy’\n\n- yǎn = ‘eye’\n\n- yǎnhóng = ‘jealousy’\n\n- hóngshì = ‘wedding’\n\n- hóngyán = ‘young woman’\n\n- hēihuà = ‘encoding’\n\n- báiyǎn = ‘to dislike’\n\n- báihuà = ‘decoding’\n\n- hēibái fēnmíng = ‘it is written in black and white’\n\n- yán = ‘face’\n\n- hēibái = ‘black-and-white’\n\n- huángle = ‘failure’\n\n- ‘white market’","source":"langsci_420","problem_group_id":"langsci420:8.5","chapter":8,"chapter_title":"Semantics","section":3,"section_title":"Practice problems","topic":"semantics and graph-based matching","language":"Chinese","author":"Roxana Dincă","competition":"RoLO","year":2015,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c08_s01_p01","book_method_c08_s02_p01","book_method_c08_s02_p02","book_method_c08_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/08-Semantics.tex","source_line_start":588,"source_line_end":600,"solution_line_start":833,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Semantics\nPractice problems\nsemantics and graph-based matching\nThis practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words and phrases in Chinese and their English translations in random order:\n\nhóng, míngbái, báishì, huáng, hēishì, shìqing huángle, yăn, yănhóng, hóngshì, hóngyán, hēihuà, báiyăn, báihuà, hēibái fēnmíng, yán, hēibái, huángle\n‘encoding’, ‘black-and-white’, ‘face’, ‘yellow’, ‘funeral’, ‘bankruptcy’, ‘to clarify’, ‘to dislike’, ‘young woman’, ‘failure’, ‘decoding’, ‘wedding’, ‘eye’, ‘black market’, ‘it's written in black and white’, ‘jealousy’, ‘red’\n- Determine the correct correspondences.\n\n- The word báishì can have two meanings in Chinese, although only one is reflected in the correspondences above. What is the other meaning?"}
{"id":"book_08_06","context":"Here are some phrases in Hausa and their English translations in random order:\n\n. | bàakín rúwáa | . | ‘talkative boy’\n. | bàbbán bàakín tsúntsúu | . | ‘white camel’\n. | bàbbán yátsàa | . | ‘big beak’\n. | bàbbán ràkúmín rúwáa | . | ‘thumb’\n. | bàbbán yár sháanúu | . | ‘fruit’\n. | bákín cíkìi | . | ‘fork’\n. | bákín ràagóo | . | ‘pregnant woman’\n. | cóokàlíi mài yátsàa | . | ‘sorrow’\n. | dán ràagóo | . | ‘estuary’\n. | dán mài bàbbán bàakíi | . | ‘lamb’\n. | fárin ràkúmíi | . | ‘black sheep’\n. | rúwán bíshíyàa | . | ‘(tree) sap’\n. | yár mài bàbbán cíkìi | . | ‘tsunami’\n. | yár bíshíyàa | . | ‘big heifer’\n\nAn ‘estuary’ is a wide part of the river, similar to a funnel. A ‘heifer’ is a young female cow.","query":"- Determine the correct correspondences.\n\n- Translate into English:\n\n- dán sháanúu\n\n- fárín cíkìi\n\n- yár mài bákin bàakíi\n\n- Translate into Hausa:\n\n- ‘girl who has a spoon’\n\n- ‘crow’\n\n- ‘river’\n\n- ‘(tree) branch’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"8.6. Hausa\n\n-\n\n- I.\n\n- C.\n\n- D.\n\n- M.\n\n- N.\n\n- H.\n\n- K.\n\n- F.\n\n- J.\n\n- A.\n\n- B.\n\n- L.\n\n- G.\n\n- E.\n\n- 15. ‘calf’ (boy + cow)\n\n- 16. ‘happiness’ (white + stomach) – based on ‘sorrow’ = black + stomach\n\n- 17. ‘impolite/naughty girl’ (girl + with + black + mouth)\n\n- 18. yár mài cóokàlíi (girl + with + spoon)\n\n- 19. bákín tsúntsúu (bird + black)\n\n- 20. bàbbán rúwáa (big + water)\n\n- 21. yátsàn bíshíyàa (finger + tree)\n\nRules:\n\n- The determiners come before the head noun;\n\n- The words yár = ‘female, woman’ and mài = ‘with’ are invariable;\n\n- All other words end in -n if they are not the head noun or they double the final vowel if they are the head noun. An alternative explanation is that they end in -n, unless they are phrase-final, in which case the -n is removed and the last vowel is doubled.","source":"langsci_420","problem_group_id":"langsci420:8.6","chapter":8,"chapter_title":"Semantics","section":3,"section_title":"Practice problems","topic":"semantics and graph-based matching","language":"Hausa","author":"Paul Helmer","competition":"RoLO","year":2019,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c08_s01_p01","book_method_c08_s02_p01","book_method_c08_s02_p02","book_method_c08_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/08-Semantics.tex","source_line_start":604,"source_line_end":652,"solution_line_start":863,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Semantics\nPractice problems\nsemantics and graph-based matching\nThis practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some phrases in Hausa and their English translations in random order:\n\n. | bàakín rúwáa | . | ‘talkative boy’\n. | bàbbán bàakín tsúntsúu | . | ‘white camel’\n. | bàbbán yátsàa | . | ‘big beak’\n. | bàbbán ràkúmín rúwáa | . | ‘thumb’\n. | bàbbán yár sháanúu | . | ‘fruit’\n. | bákín cíkìi | . | ‘fork’\n. | bákín ràagóo | . | ‘pregnant woman’\n. | cóokàlíi mài yátsàa | . | ‘sorrow’\n. | dán ràagóo | . | ‘estuary’\n. | dán mài bàbbán bàakíi | . | ‘lamb’\n. | fárin ràkúmíi | . | ‘black
{"id":"book_08_07","context":"Here are some words and phrases in Tetum and their English translations in random order:\n\n. | ai boot | . | ‘eyelid’\n. | ai fuan boot | . | ‘scapula’\n. | ai fuan musan | . | ‘leaf’\n. | ai tahan | . | ‘big fruit’\n. | ibun kulit boot | . | ‘big tree’\n. | kbas | . | ‘auricle’\n. | kbas tahan | . | ‘skin’\n. | kulit | . | ‘seed’\n. | matan kulit | . | ‘shoulder’\n. | matan musan | . | ‘eyeball’\n. | tilun tahan | . | ‘big lip’\n\nThe ‘scapula’ (or shoulder blade) is the large flat bone that is part of the shoulder joint. The ‘auricle’ is the visible part of the ear.","query":"- Determine the correct correspondences.\n\n- Translate into English:\n\n- matan fuan\n\n- ai fuan kulit\n\n- tilun boot\n\n- One of the phrases has the same translation as one of the phrases 1–11.\n\n- Translate into Tetum:\n\n- ‘mouth’\n\n- ‘big eye’\n\n- ‘(tree) bark’\n\n- ‘grain’\n\n- Two of the words above can be combined to construct a phrase meaning ‘impolite person’. Which ones are these?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"8.7. Tetum\n\n-\n\n- E.\n\n- D.\n\n- H.\n\n- C.\n\n- K.\n\n- I.\n\n- B.\n\n- G.\n\n- A.\n\n- J.\n\n- F.\n\n-\n\n-\n\n- ‘eyeball’\n\n- ‘(fruit) peel’\n\n- ‘big ear’\n\n-\n\n- ibun\n\n- matan boot\n\n- ai kulit\n\n- musan\n\n- ibun boot (‘big mouth’)","source":"langsci_420","problem_group_id":"langsci420:8.7","chapter":8,"chapter_title":"Semantics","section":3,"section_title":"Practice problems","topic":"semantics and graph-based matching","language":"Tetum","author":"Aleksejs Peguševs","competition":"RoLO","year":2020,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c08_s01_p01","book_method_c08_s02_p01","book_method_c08_s02_p02","book_method_c08_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/08-Semantics.tex","source_line_start":656,"source_line_end":702,"solution_line_start":904,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Semantics\nPractice problems\nsemantics and graph-based matching\nThis practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words and phrases in Tetum and their English translations in random order:\n\n. | ai boot | . | ‘eyelid’\n. | ai fuan boot | . | ‘scapula’\n. | ai fuan musan | . | ‘leaf’\n. | ai tahan | . | ‘big fruit’\n. | ibun kulit boot | . | ‘big tree’\n. | kbas | . | ‘auricle’\n. | kbas tahan | . | ‘skin’\n. | kulit | . | ‘seed’\n. | matan kulit | . | ‘shoulder’\n. | matan musan | . | ‘eyeball’\n. | tilun tahan | . | ‘big lip’\n\nThe ‘scapula’ (or shoulder blade) is the large flat bone that is part of the shoulder joint. The ‘auricle’ is the visible part of the ear.\n- Determine the correct correspondences.\n\n- Translate into English:\n\n- matan fuan\n\n- ai fuan kulit\n\n- tilun boot\n\n- One of the phrases has the same translation as one of the phrases 1–11.\n\n- Translate into Tetum:\n\n- ‘mouth’\n\n- ‘big eye’\n\n- ‘(tree) bark’\n\n- ‘grain’\n\n- Two of the words above can be combined to construct a phrase meaning ‘impolite person’. Which ones are these?"}
{"id":"book_08_08","context":"Here are some words and phrases in Malagasy and their English translations in random order:\n\n. | mahandohalika | . | ‘grandson’\n. | lohalika | . | ‘ankle’\n. | zafin-kitrokely | . | ‘shoot of rice (departing from the stem)’\n. | hafaladia | . | ‘up to the sole’\n. | zafim-bary | . | ‘rice field’\n. | kitrokely | . | ‘great-great-great-grandson’\n. | zafim-paladia | . | ‘one who can get on his knees’\n. | zafy | . | ‘great-great-great-great–grandson’\n. | tanim-bary | . | ‘knee’\n. | mahambozona | . | ‘one who can carry something on his neck’\n\ny = ‘i’ in ‘pit’.","query":"- Determine the correct correspondences.\n\n- Translate into English:\n\n- tany\n\n- vozona\n\n- halohalika\n\n- Translate into Malagasy:\n\n- ‘great-great-grandson’\n\n- ‘sole’\n\n- ‘up to the ankle’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"8.8. Malagasy\n\n-\n\n- G.\n\n- I.\n\n- F.\n\n- D.\n\n- C.\n\n- B.\n\n- H.\n\n- A.\n\n- E.\n\n- J.\n\n-\n\n- ‘field’\n\n- ‘neck’\n\n- ‘up to the knee’\n\n-\n\n- zafin-dohalika\n\n- faladia\n\n- hakitrokely","source":"langsci_420","problem_group_id":"langsci420:8.8","chapter":8,"chapter_title":"Semantics","section":3,"section_title":"Practice problems","topic":"semantics and graph-based matching","language":"Malagasy","author":"Alexey Kretov","competition":"MSK","year":2011,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c08_s01_p01","book_method_c08_s02_p01","book_method_c08_s02_p02","book_method_c08_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/08-Semantics.tex","source_line_start":706,"source_line_end":747,"solution_line_start":946,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Semantics\nPractice problems\nsemantics and graph-based matching\nThis practice problem belongs to the book's Semantics chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some words and phrases in Malagasy and their English translations in random order:\n\n. | mahandohalika | . | ‘grandson’\n. | lohalika | . | ‘ankle’\n. | zafin-kitrokely | . | ‘shoot of rice (departing from the stem)’\n. | hafaladia | . | ‘up to the sole’\n. | zafim-bary | . | ‘rice field’\n. | kitrokely | . | ‘great-great-great-grandson’\n. | zafim-paladia | . | ‘one who can get on his knees’\n. | zafy | . | ‘great-great-great-great–grandson’\n. | tanim-bary | . | ‘knee’\n. | mahambozona | . | ‘one who can carry something on his neck’\n\ny = ‘i’ in ‘pit’.\n- Determine the correct correspondences.\n\n- Translate into English:\n\n- tany\n\n- vozona\n\n- halohalika\n\n- Translate into Malagasy:\n\n- ‘great-great-grandson’\n\n- ‘sole’\n\n- ‘up to the ankle’"}
{"id":"book_09_01","context":"Here are some Quenya numbers:\n\nneldë | 3 | enquë yucainen | 26\ncanta | 4 | minë nelcainen | 31\nlempë | 5 | cancainen | 40\notso | 7 | atta tolcainen | 82\ntolto | 8 | atta tolcainen tuxa | 182\nnelcëa | 13 | nertë nelcainen lemtuxa | 539\nencëa | 16 | |","query":"- Write in numerals:\n\n- tolcëa\n\n- enquë cancainen\n\n- cancainen neltuxa\n\n- lempë tolcainen\n\n- tuxa\n\n- [blank]\n\n- Write in Quenya: 1, 70, 192, 385.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"text_to_num","eval_type":"single","reasoning_trace":"We notice that numbers 13 and 16 both end in the suffix -cëa, while numbers bigger than 26 have a different structure. Thus, comparing the words for 3 and 13, we can assume it is a base-10 language and that numbers 10 + X are written as X-cëa. We further notice that in order to form the number 13, only a part of the word for 3 is used (nel).\n\nComparing examples 82 and 182, they differ only by the word tuxa placed at the end; therefore, we can deduce that tuxa means 100 (which further confirms that the number system for this language is base 10). Moreover, looking at the number 539, we notice that the last word is lemtuxa, where tuxa = 100, and lem is the first part of the number 5. Thus, in this case as well, only a part of the stem of the unit is used. Furthermore, we notice that in the right column, the second word always ends in cainen. We can assume that this is the suffix which marks the tens (10X), whence we deduce that the word order in Quenya is units-tens-hundreds.\n\nOnly based on these observations, we can write the following rules:\n\n10+X = X′ -cëa 100X+10Y+Z = Z Y′-cainen X′-tuxa\n\nBy X′ and Y′ we mean that a different, truncated, form of the word is used.\n\nBased on these observations, we can create a table with the form of each digit in different contexts:\n\nDigit | X | 10+X | 10X | 100X\n1 | minë | | |\n2 | atta | | yu- |\n3 | neldë | nel- | nel- |\n4 | canta | | can- |\n5 | lempë | | | lem-\n6 | enquë | en- | |\n7 | otso | | |\n8 | tolto | | tol- |\n9 | nertë | | |\n\nThe four columns represent the form in which the respective digit is used if it represents the units, if it appears with the suffix -cëa (meaning 10+X), if it appears with the suffix -cainen (10X), or if it appears with the suffix -tuxa (100X).\n\nWe notice that the same form of 3 appears both in the case of 10+X and 10X. Therefore, we deduce that both contexts use the same form. Moreover, we have no reason not to assume that the same form will also be used in the case of hundreds. In order to deduce how that form is constructed, we compare it with the full form in the first column (the units). We notice that the short form (which we designated by X′ and Y′) represents the first syllable of the unit. The only exception is the digit 2, where the form used for 20 is yu-. It appears that in Quenya 1 and 2 are irregular and have different forms in different contexts. This is not uncommon cross-linguistically.\n\nThus, we can solve the tasks and write the rules.\n\nRules:\nDigits are single words. In compound words, only the stem of the digit is used, which is represented by the first syllable (in the notation below, X′ is the stem/first syllable of X). The digit 2 has a special form, yu-. Thus:\n\n10+X = X′-cëa 100X+10Y+Z = Z Y′-cainen X′-tuxa\n\n-\n\n- 18\n\n- 46\n\n- 340\n\n- 85\n\n- 100\n\n-\n\n- 1 = minë\n\n- 70 = otcainen\n\n- 192 = atta nercainen tuxa\n\n- 385 = lempë tolcainen neltuxa\n\nIn situations where the base is unknown, a simple method to get some additional information is to count how many morphemes there are. If in a particular language we count 11 digits, we expect the base to be, most likely, 10 or 12. Usually, this method is just an estimation, and the result should probably be taken with an error margin of ±2 because: (1) it is possible that we misidentified some of the digits, and (2) it is possible that the problem does not feature all the digits or even some digits might have different for
{"id":"book_09_02","context":"Here are two equalities in Embera Chami:\n\n() | umbea + huasoma kwimane | = | omme huasoma omme\n() | omme huasoma kwimane + huasoma abba | = | kwimane huasoma","query":"- Write the equalities above with numerals.\n\n- Write in Embera Chami: 1, 5, 17, 23.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"text_to_num","eval_type":"single","reasoning_trace":"At first sight, it might seem a very difficult problem without an obvious starting point and with very little information given. Nevertheless, if we check the structure of the numbers, we notice that there are two types: single words or numbers like X huasoma Y. We can assume that the second type will represent bigger numbers, and that huasoma is the base, these numbers representing Xhuasoma + Y (or Y × huasoma + X). Moreover, we notice that the same word can represent both X and Y, therefore X and Y are the slots where the digits are placed. Based on these rules, we can try to count the number of digits that occur in the problem, and we notice that there are only four (umbea, kwimane, omme, abba). Moreover, they have a constant form (there are no changes or added or deleted morphemes). We can then assume that the base is 5 and that huasoma = 5.\n\nThe next thing we need to do is figure out the order of the constituents, i.e., figure out whether X huasoma Y means 5X + Y or 5Y + X. Looking at equality (1), we have two cases:\n\n- huasoma X = 5X\n\n- huasoma X = 5+X\n\nAssuming case a. is true, eq. (1) becomes:\n\numbea+5kwimane = omme+5omme\n\nThis seems unlikely, since we know that umbea is, most likely, a digit. Thus, if huasoma kwimane meant 5X, then umbea + huasoma kwimane should be equal to umbea huasoma kwimane (i.e., umbea + 5kwimane). Generally, the carryoverWe use the term carryover for the following situation: when you add 17 and 19, you add the units first (7 and 9 to get 16) and “carry over” the 1 (from 16) to the tens. is a strong strategy to discover certain digits.\n\nThus, case b. must be correct, and now we know that huasoma X = 5 + X, so we can deduce that X huasoma = 5X and, extrapolating, X huasoma Y = 5X + Y.\n\nMoreover, from eq. (1), we see that the number resulting from the addition has the multiplier omme. Since this is the result of a sum between a unit and a number like base + X, the result can either also be base + X (if there is no carryover) or 2base + X (if there is carryover). Since we know that there is carryover (the number on the right has a multiplier, since it has the structure 5X+Y), we deduce that omme = 2.\n\nIf we denote the remaining three digits by U, K and A (corresponding to their first letter), we can rewrite the equalities as follows:\n\n- U+(5+K)=12\n\n- (10+K)+(5+A)=5K\n\nRearranging equality (1), we get: U + K = 7.\n\nKnowing that U and K are digits smaller than 5 (the base), U and K can only correspond to 3 and 4, not necessarily in this order. Thus, A can only be 1 since it is the only remaining digit. Therefore, abba = 1.\n\nReplacing this in eq. (2) gives us: 10 + K + 6 = 5K 4K = 16. So K = 4 and, subsequently, U = 3.\n\nThus, the rules are:\n\n- 1 = abba, 2 = omme, 3 = umbea, 4 = kwimane, 5 = huasoma;\n\n- 5X + Y = X huasoma Y.\n\n-\n\n- 3+9=12\n\n- 14+6=20\n\n-\n\n- 1 = abba\n\n- 5 = huasoma\n\n- 17 = 35+2 = umbea huasoma omme\n\n- 23 = 45+3 = kwimane huasoma umbea","source":"langsci_420","problem_group_id":"langsci420:9.2","chapter":9,"chapter_title":"Number systems","section":1,"section_title":"Introduction","topic":"number systems","language":"Embera Chami","author":"Vlad A. Neacșu","competition":"PLO","year":2022,"solution_tier":"detailed_worked","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02"],"method_link_type":"author_placement","method_description":"The author places this worked example under “Introduction” in the Number systems chapter, so it illustrates that method or topic.","string_only_usable":true,"visual_dependency_reasons
{"id":"book_09_03","context":"Yup'ik people have an interesting concept when it comes to counting – the words for the numbers can be broken down into meaningful parts which may be related to body parts. For example, the word for 5, talliman, means ‘arm’ and the word for 6, arvinlegen, means ‘cross over’, as you need to change hands to go on counting.\n\nThe Yup'ik people often include geometry in the border patterns of their traditional garments, called \"parkas\". One such pattern comes from a 33 square, as represented below. This is a magic square, constructed by placing the digits 1 to 9 within the cells such that the sum of all the digits in every row, column, and diagonal is the same.\n[VISUAL OMITTED: images/Yupik_square.jpg]\n\nTo help you fill in the magic square, the following clues are given in Yup'ik.\nHint: 294 in Yup'ik is yuinaat qula cetaman qula cetaman.\n\n- Rows:\n\n- Yuinaat yuinaq cetaman qula malruk\n\n- Yuinaat akimiaq malruk akimiaq malruk\n\n- Yuinaat yuinaak malruk akimiaq atauciq\n\n- Columns:\n\n- Yuinaat yuinaq atauciq akimiaq pingayun\n\n- Yuinaat yuinaak malruk yuinaat malrunglegen qula atauciq\n\n- Yuinaat qula pingayun akimiaq atauciq","query":"- Fill in the numbers missing from the magic square above. One digit is already given (a2 = 9).\n\n- Write in Yup'ik the number given in the diagonal from top left (the number formed by the digits a1-b2-c3).","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"*1. Solving the magic square\n\nThis is in reality the easiest part. It is known that in a magic square the middle number must be 5 (you can attempt a mathematical proof, it is rather easy), hence b2 = 5. Moreover, it is known that the sum on every row, column and diagonal must be 15. Since on the middle column we already have 9 and 5, we deduce that c2 = 1.\n\nOn the first row, we already have the digit 9, so the sum of the other two digits must be 6. We have three possibilities: 5 and 1 (impossible, since 5 is already used), 3 and 3 (impossible, since we cannot repeat digits), or 4 and 2 (this is therefore the only possible option). Therefore, the first row can be either 492 or 294. Since in the introduction we are given the Yup'ik name for 294, which does not appear in the crossword clues (hence, it doesn't appear in the square), we deduce that the first row must be 492, so a1 = 4, a3 = 2.\n\nSince we are told that the sum on the diagonals is also constant (so, 15), we can easily deduce that c1 = 8 and c3 = 6, which makes filling in the rest of the square trivial. In the end we get:\n\n[VISUAL OMITTED: images/Yupik_Square-solved.png]\n\n*2. Solving the number problem\n\nOnce the square is filled in, we can extract all the information in a table, transforming the problem into a classic one, in which we are given some numbers spelled out in Yup'ik:\n\n276 | yuinaat qula pingayun akimiaq atauciq\n294 | yuinaat qula cetaman qula cetaman\n357 | yuinaat akimiaq malruk akimiaq malruk\n438 | yuinaat yuinaq atauciq akimiaq pingayun\n492 | yuinaat yuinaq cetaman qula malruk\n816 | yuinaat yuinaak malruk akimiaq atauciq\n951 | yuinaat yuinaak malruk yuinaat malrunglegen qula atauciq\n\nThe first important observation is based on the last word in every number. We have four types of numbers: ending in atauciq (276, 816, 951), ending in malruk (357, 492), ending in cetaman (294), and ending in pingayun (438). Looking closely, one notices that these numbers can be grouped based on their remainder when divided by 5 (i.e., modulo 5). Thus, we can deduce that:\n\natauciq = 1, malruk = 2, pingayun = 3, cetaman = 4\n\nReplacing these numbers, we get:\n\n276 | yuinaat qula 3 akimiaq 1\n294 | yuinaat qula 4 qula 4\n357 | yuinaat akimiaq 2 akimiaq 2\n438 | yuinaat yuinaq 1 akimiaq 3\n492 | yuinaat yuinaq 4 qula 2\n816 | yuinaat yuinaak 2 akimiaq 1\n951 | yuinaat yuinaak 2 yuinaat malrunglegen qula 1\n\nSince we assumed that the last number is added, we can simply subtract it (from the number representation) and delete it (
{"id":"book_09_04","context":"| Umbu-Ungu\n10 | rureponga talu\n15 | malapunga yepoko\n20 | supu\n21 | tokapunga telu\n27 | alapunga yepoko\n30 | polangipunga talu\n\n| Umbu-Ungu\n35 | tokapu rureponga yepoko\n40 | tokapu malapu\n48 | tokapu talu\n50 | tokapu alapunga talu\n69 | tokapu talu tokapunga telu\n79 | tokapu talu polangipunga yepoko\n97 | tokapu yepoko alapunga telu\n\ntelu < yepoko","query":"- Write in numerals:\n\n- tokapu polanigpu\n\n- tokapu talu rureponga telu\n\n- tokapu yepoko malapunga talu\n\n- tokapu yepoko polangipunga telu\n\n- Write in Umbu-Ungu: 13, 66, 72, 76, 95.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"text_to_num","eval_type":"single","reasoning_trace":"Initial observations:\n\n- supu is a single word, so most likely is a multiple of the base; therefore the base can be 4, 5, 10, or 20;\n\n- 10 is not a single word, so it is unlikely that the base is 5, 10 or 20. Therefore, it seems to be a base-4 language;\n\n- tokapu occurs in the number 35 (but it does not occur in 30), so it most likely means 32. 31, 33 and 34 do not seem to make sense as single words, since they are extremely unlikely as bases, and it can't be 35 either since there are other words following it. Moreover, 32 confirms the hypothesis that the language is base-4, since it is a multiple of 4;\n\n- General structure: (tokapu) (X) (Y-pu(nga)) (Z).\n\nBased on this structure, we can count the digits. These appear as X or Z (but we notice that they do not occur as Y, which means that Y-pu(nga) is a single word). There are only three words appearing in X and Z positions: talu, telu, yepoko.\n\nMoreover, we notice that 48 is tokapu talu. Most likely, talu is a multiplier for tokapu, and, since the numbers smaller than 48 do not contain this structure, we deduce that talu = 2. Indeed, for numbers 35 and 40, tokapu occurs without a multiplier (1 is implicit), so the first multiplier that ought to appear is 2. Based on the same logic, yepoko must mean 3 and, knowing that telu < yepoko, we get that telu is 1.\n\nReplacing these numbers in the given data, we get:\n\n| Umbu-Ungu\n10 | rureponga 2\n15 | malapunga 3\n20 | supu\n21 | tokapunga 1\n27 | alapunga 3\n30 | polangipunga 2\n\n| Umbu-Ungu\n35 | tokapu rureponga 3\n40 | tokapu malapu\n48 | tokapu 2\n50 | tokapu alapunga 2\n69 | tokapu 2tokapunga 1\n79 | tokapu 2 polangipunga 3\n97 | tokapu 3 alapunga 1\n\nBased on the number 48, since we assumed that 2 is a multiplier, we deduce that tokapu = 24. Moreover, based on the first column, by subtraction, we obtain the numbers: rureponga = 8, malapunga = 12, supu = 20, tokapunga = 20, alapunga = 24, polangipunga = 28.\n\nHowever, we notice that we have two different words for 20. Nevertheless, all words than end in -punga also have a units digit (1, 2 or 3), so it is possible that each word has two forms, one for when it appears alone and the other that appears when a units digit is added.\nThus, based on the words we know already, 35 is written as ‘24 8 3’ (and, indeed, 35 = 24 + 8 + 3).\n\nAn interesting thing happens with the number 40. Knowing that tokapu = 24, it must be that malapu is 16 (but we also know that malapunga = 12). Thus, we deduce the following rule: each multiple of 4 is a single or base word (ending in -pu). When a units digit (1, 2, or 3) is added, the following multiple of 4 is used to which the suffix -nga is added. Therefore, 24 = tokapu and 28 = alapu, but 25 = alapunga telu (basically, (28-4) + 1), 26 = alapunga talu (28-4 + 2) and 27 = alapunga yepoko (28-4 + 3).\n\nThis is a rather common phenomenon called overcounting, in which the numbers are regarded as going towards.... Thus, 27 can be translated literally as ‘three (units) towards 28’ (meaning that it is three units past 24). Overcounting occasionally occurs Indo-European languages as well (e.g., in German, the clock time 7.30 is read as halb acht (meaning ‘half eight’) – which is to say, half an hour has passed towards 8 o'clock).\n\nA last observation concerns the numbers 48, 50 and 69. We notice that
{"id":"book_09_05","context":"The perfect squares from 1 to 100 are written in Huli below, in random order:\n\n- ngui ki, ngui tebone-gonaga waragaria\n\n- mbira\n\n- ngui dau, ngui waragane-gonaga waragaria\n\n- nguira-ni pira\n\n- nguira-ni mbira\n\n- dira\n\n- maria\n\n- ngui tebo, ngui mane-gonaga maria\n\n- ngui ma, ngui dauni-gonaga maria\n\n- ngui waraga, ngui kane-gonaga pira","query":"- For each of them, write its corresponding value.\n\n- Here are four consecutive numbers written in Huli, in ascending order:\n\n- ngui ka, ngui haline-gonaga bearia\n\n- ngui ka, ngui haline-gonaga hombearia\n\n- ngui ka, ngui haline-gonaga haleria\n\n- ngui ka, ngui haline-gonaga deria\n\n- Write their corresponding values.\n\n- Write in Huli: 2, 4, 6, 7, 22, 44, 66, 77, 88, 173.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"The first step is figuring out the structure of Huli numbers. We have three types of structures: single words (dira, maria, mbira), structures like nguira-ni X and structures like ngui X, ngui Y-gonaga Z. At first sight, we would expect the numbers represented by single words to be the smallest (digits) – taking this with a grain of salt, since some of them could also represent orders, e.g., 100; numbers like nguira-ni X are the second smallest ones (we can probably assimilate them with the type base + X), while the last category represents the biggest numbers (base A + B) – although we still do not know why there are three digits in these structures and not only two.\n\nOnce the structures are identified, we know exactly where the digits are placed in these structures, so we can try counting them in order to get an estimate of the base. We get the morphemes: ki, tebone, waragaria, mbira, dau, waragane, pira, dira, maria, tebo, mane, ma, dauni, waraga, kane. Nevertheless, we notice that digits can have different forms (since we find the triplets ma – mane – maria and waraga – waragane – waragaria, each of them following the same pattern: X – X-ne – X-ria), and furthermore, those without any suffix appear only after ngui, those with the suffix -ne appear only with the ending -gonaga, while those with the suffix -ra/-ria appear only at the very end, after nguira-ni or if they are single structures. Therefore, we can assume that they denote the same digit, and each digit has three different forms, depending on the context. Thus, we are left with 10 morphemes: ki, mbi(ra), dau, pi(ra), di(ra), dau(ni), ka(ne), tebo, waraga, ma, so we could expect this language to be base-10. Nevertheless, if we look at task (b), we notice four more morphemes occur: bea(ria), hombea(ria), hale(ria), de(ria). Therefore, the total number of digits is 14. Since base 13 is extremely unlikely, as well as base 14, it is most likely one of the bases 12, 15 or 16 (among which, base 15 is the most likely one since we have discovered 14 morphemes).\n\nFurthermore, since the base is bigger than 12, we certainly know that 1, 4, and 9 are digits (so they will be represented by a single word). Therefore, based on the previous observation according to which X < nguira-na X < ngui X, ngui Y-gonaga Z, we can already split the numbers into categories. Thus:\n\n{mbira, dira, maria} must correspond to {1, 4, 9} – the use of curly brackets shows that we are not sure about the exact order,\nand\n{nguira-ni mbira, nguira-ni pira} = {16, 25}.\n\nSince mbira appears in both structures, we can try obtaining (by subtraction) the value of nguira-ni, which, most likely, represents the base.\n\nWe can do this by considering all the possible cases, as follows, and calculating the difference between them:\n\n| | mbira\n3-5\n| | 1 | 4 | 9\n3-5\nnguira-ni | 16 | 15 | 12 | 7\nmbira | 25 | 24 | 21 | 16\n\nThus, the values for nguira-ni, and, implicitly, for the base are 7, 12, 15, 16, 21, 24. Since we know we have roughly 14 digits, we can exclude the bases 7, 21 and 24. Moreover, since 16 appears in the corpus (and it is not a single word), it is unlikely t
{"id":"book_09_06","context":"Here are some numbers in Yoruba:\n\nèji | 2 | ẹẹ́rìndilogóji | 36\nẹ̀rin | 4 | ẹ̀rìndogóji | 44\nàrun | 5 | àádorin | 70\nẹ̀rinlá | 14 | ẹẹ́tàdilogórin | 77\neéjìdilogun | 18 | ẹ̀tàdogórin | 83\n\nThe marks above the vowels denote tones; e and ẹ are distinct vowels.","query":"- Write in numerals:\n\n- àádota\n\n- àrùndogórin\n\n- aárùndilogórin\n\n- ẹ̀tàdogórun\n\n- òkándilogóji\n\n- [blank]\n\n- Write in Yoruba: 12, 45, 57, 90, 99.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"text_to_num","eval_type":"single","reasoning_trace":"Firstly, we can notice the pair 4 and 14, in which the only difference is the suffix -lá. We can therefore deduce that 10 + X = X-lá, so this system is most likely base-10.\n\nNext, we have the pair 36 and 44, in which the only difference is -il-. Moreover, the first part of the word is highly similar to the word for 4 (in the case of 44 it is actually identical, while in the case of 36 it is slightly mutated), and the same thing goes for the pair 77 and 83. We notice that a common aspect of these numbers is that they are symmetrical with respect to the closest tens (thus, 36 and 44 are symmetrical with respect to 40, i.e., 36 = 40 - 4 and 44 = 40 + 4, while 77 and 83 are symmetrical with respect to 80, i.e. ±3). This can make us think of a subtractive system. Moreover, the last morpheme of 36 and 44 is óji (which resembles èji = 2), while the last morpheme of 77 and 83 is órin (ẹ̀rin). Since this number cannot refer to the units (we know already that the units are 4 and 3, respectively), it most likely represents the tens. Thus, 40 = óji, and 80 = órin. Therefore, we realise it is actually a base-20 system, and the structure of the numbers is:\n\n- X-lá = 10 + X\n\n- U-dog-Z = 20Z + U\n\n- U-dilog-Z = 20Z - U\n\nMoreover, we notice that 70 has a special form, also derived from órin = 80, so, most probably, it represents 80 - 10.\n\nThus, we notice that the digits have different forms depending on the context in which they appear (single units, 10 + X, added units, subtracted units, 20X, 20X-10). Moreover, we already noticed that 10+X uses the same form of X as the single unit, so we can combine these two forms.\n\nCombining everything into a table, we get:\n\n| Unit / 10+X | + units | - units | 20X | 20X-10\n1 | | | | un |\n2 | èji | | eéjì | óji |\n3 | | ẹ̀tà | ẹẹ́tà | |\n4 | ẹ̀rin | ẹ̀rìn | ẹẹ́rìn | órin | àádorin\n5 | àrun | | | |\n\nSince we have all the different forms of the digit 4, we can see the transformations that take place. In order to form the added units, the second vowel receives a grave accent; to form the subtracted units, the first vowel is doubled (of which the first is toneless, while the second has an acute accent), and the final vowel gets a grave accent. In order to form 20X, the first vowel becomes ó, while for the formation of 20X-10, the first vowel is replaced by àádo-.\n\nThe only exception is the number 20 (201), but, as explained before, we can expect that digit 1 might have some irregular forms.\n\nAnother simpler way to write the rules is to notice that each digit has the structure V̀1CV2(n). We can easily derive the rest of the forms: + units = V̀1CV̀2(n), - units = V1V́1CV̀2(n), 20X = ó CV2(n), 20X-10 = àádo CV2(n), once again, mentioning that for 1 the form 20X is irregular (un).\n\nThus, we can solve all the tasks:\n\n-\n\n- 50\n\n- 85\n\n- 75\n\n- 103\n\n- 39\n\nNotebulbonIn order to figure out the meaning of òkán in task (e), we need to look at the table above (which we previously filled in with the additional information that we discovered in task (a)) and we notice that 1 is the only digit for which we do not know the subtracted units form (column 3 from the table above). So we can conclude that òkán is the subtracted form of 1. Moreover, we see again that its form is irregular.\n\n-\n\n- 12 = èjilá\n\n- 90 = àádorun\n\n- 57 = ẹẹ́tàdilogóta\n\n- 45 = àrùndogóji\n\
{"id":"book_09_07","context":"Here are some examples of how to tell the time in Czech:\n\nza pět minut osm | ‘five minutes to eight’\nza deset minut osm | ‘ten minutes to eight’\nčtvrt na osm | ‘quarter past seven’\nza sedm minut osm | ‘seven minutes to eight’\nza osm minut čtvrt na osm | ‘seven minutes past seven’\nza deset minut čtvrt na sedm | ‘five minutes past six’\npůl osmé | ‘half past seven’\npůl deváté | ‘half past eight’\nza deset minut půl šesté | ‘twenty minutes to five’\nčtvrt na deset | ‘quarter past nine’","query":"- Translate into Czech:\n\n- ‘twenty-three minutes past five’\n\n- ‘ten minutes to nine’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"We notice three types of structures:\n\n- čtvrt na ... = ‘quarter past ...’\n\n- půl ...-é = ‘half past ...’\n\n- za ... minut (čtvrt na / půl) ... – otherwise\n\nIt is obvious that minut = ‘minute(s)’, so if we compare the first two examples, we obtain: osm = 8, deset = 10 and pět = 5. Moreover, we deduce that za X minut Y = ‘X minutes to Y’.\n\nWe notice that osm occurs in the phrase ‘half past seven’, but not in ‘half past eight’. This indicates that Czech also features overcounting when it comes to telling the time (similar to German). Therefore, půl X-é = ‘half past (X-1)’, from which we deduce that devát = 9.\n\nMoreover, looking at the example ‘quarter past seven’, in which the word osm = 8 occurs again, we deduce that overcounting is also applied for quarters, so čtvrt na X = ‘quarter past (X-1)’ (or ‘quarter towards X’).\n\nThe only structures we are left to analyse are those following the pattern za X minut ... na Y. We can separate them from the rest and replace the words we already know:\n\nza osm minut čtvrt na osm | ‘seven minutes past seven’\nza deset minut čtvrt na sedm | ‘five minutes past six’\nza deset minut půl šesté | ‘twenty minutes past five’\n\nWe already know that the hour is placed in the end, so we can directly replace the structures čtvrt na X and půl šesté. Moreover, we know all the words, so we notice that ‘seven minutes past seven’ is written as za ‘8 minutes quarter past seven’, and ‘five minutes past six’ is written as za ‘10 minutes quarter past six’. We therefore notice that the Czech system is based on overcounting. Therefore, ‘seven minutes past seven’ is translated as ‘8 minutes to [quarter towards 8]’ or ‘8 minutes to [quarter past 7]’.\n\nThe same thing is noticed in the structure za deset minut půl šesté which is literally translated as ‘10 minutes to [half towards 6]’ or ‘10 minutes to [half past 5]’.\n\nSo we now know all the possible structures so we can write the rules and solve the tasks.\n\nRules:\n\n- půl X-é = ‘half past (X-1)’\n\n- čtvrt na X = ‘quarter past (X-1)’\n\n- za X minut Y = ‘X minutes to Y’, where Y can be a full hour (o'clock) or any of the two structures above.\n\n- 1. ‘23 past 5’ = ‘7 minutes to [half past five]’ = ‘7 minutes to [half towards 6]’ = za sedm minut půl šesté\n\n- 2. ‘10 past 9’ = ‘5 minutes to [quarter past nine]’ = ‘5 minutes to [quarter towards 10]’ = za pět minut čtvrt na deset\n\nIn this case, overcounting is used for all the structures, but, as mentioned above, there are languages in which only certain structures use overcounting. For example, in German, overcounting is only used in the half-past constructions – halb X = ‘half past (X-1)’.","source":"langsci_420","problem_group_id":"langsci420:9.7","chapter":9,"chapter_title":"Number systems","section":5,"section_title":"Time","topic":"number systems","language":"Czech","author":"Mirjam Fried","competition":"Princeton","year":null,"solution_tier":"detailed_worked","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s05_p01"],"method_link_type":"author_placem
{"id":"book_09_08","context":"Here are some Swahili phrases and their English translations in random order:\n\n. | jumamosi, saa moja usiku | . | ‘Sunday, 1.00 AM’\n. | jumapili, saa tatu na robu asubuhi | . | ‘Sunday, 7.30 AM’\n. | jumamosi, saa saba usiku | . | ‘Sunday, 9.15 AM’\n. | jumamosi, saa mbili na robu usiku | . | ‘Tuesday, 12.15 PM’\n. | jumanne, saa tano na nusu usiku | . | ‘Tuesday, 11.30 PM’\n. | jumanne, saa sita na robu asubuhi | . | ‘Saturday, 10.30 AM’\n. | jumamosi, saa nne na nusu asubuhi | . | ‘Saturday, 7.00 PM’\n. | jumapili, saa moja na nusu asubuhi | . | ‘Saturday, 8.15 PM’","query":"- Determine the correct correspondences.\n\n- Translate into English:\n\n- jumatano, saa moja na robu asubuhi\n\n- jumapili, saa nne na nusu asubuhi\n\n- Translate into Swahili:\n\n- ‘Monday, 12.15 AM’\n\n- ‘Monday, 10.00 PM’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"translation","eval_type":"single","reasoning_trace":"The first step is noticing the structure of the phrases. All of them start with a word separated by a comma, which has the prefix juma-. This most likely represents the day (it is unlikely that all that variety of times would use the same one-word structure). Therefore, the structure after the comma must be the time. For this, we notice two types of structures:\n\nsaa X usiku and saa X na {robu/nusu} {usiku/asubuhi}\n\nWe can assume that the simplest structure refers to the o'clock times, which is also reinforced by the fact that we only have two such structures in both Swahili and English. Thus:\n\njumamosi, saa moja usiku | ‘Sunday, 1.00 AM’\njumamosi, saa saba usiku | ‘Saturday, 7.00 PM’\n\nCuriously, in Swahili, the same name seems to be used to express different days in English. We will have to explain why that is in due course.\n\nAmong the remaining examples, only one more contains the hour 7 (although it is AM, instead of PM), and only one other example starts with either saa saba or saa moja (since we do not know the correspondence between the two). Since phrase 8 contains saa moja, we deduce that it corresponds to translation B. Moreover, we deduce that saa moja means 7 o'clock, so we can make three correspondences:\n\n1. | jumamosi, saa moja usiku | G. | ‘Saturday, 7.00 PM’\n3. | jumamosi, saa saba usiku | A. | ‘Sunday, 1.00 AM’\n8. | jumapili, saa moja na nusu asubuhi | B. | ‘Sunday, 7.30 AM’\n\nBased on example 8, we can infer that na nusu means ‘half past’. Moreover, most likely, usiku and asubuhi correspond, in some way, to the notion of AM and PM, but it is not a one-to-one correspondence, since usiku is used for both 7 PM and 1 AM.\n\nWe only have two more examples which contain na nusu (5 and 7), and the English times correspond to 10.30 AM and 11.30 PM. Since usiku is used for both 7 PM and 1 AM, we can assume that it will also be used for 11.30 PM, so that usiku suggests the idea of ‘evening/night’.\n\nWe get two more correspondences:\n\n1. | jumamosi, saa moja usiku | G. | ‘Saturday, 7.00 PM’\n3. | jumamosi, saa saba usiku | A. | ‘Sunday, 1.00 AM’\n8. | jumapili, saa moja na nusu asubuhi | B. | ‘Sunday, 7.30 AM’\n5. | jumanne, saa tano na nusu usiku | E. | ‘Tuesday, 11.30 PM’\n7. | jumamosi, saa nne na nusu asubuhi | F. | ‘Saturday, 10.30 AM’\n\nThe last three correspondences can be easily made. Only one of the phrases contains the word usiku, and among the times 8.15 PM, 9.15 AM, 12.15 PM, the only one that refers to evening time is 8.15 PM. Therefore, 4=H. Among the remaining two phrases, one starts with jumanne, which seems to mean ‘Tuesday’. Therefore, we have the correspondences:\n\n1. | jumamosi, saa moja usiku | G. | ‘Saturday, 7.00 PM’\n3. | jumamosi, saa saba usiku | A. | ‘Sunday, 1.00 AM’\n8. | jumapili, saa moja na nusu asubuhi | B. | ‘Sunday, 7.30 AM’\n5. | jumanne, saa tano na nusu usiku | E. | ‘Tuesday, 11.30 PM’\n7. | jumamosi, saa nne na nusu asubuhi | F. | ‘Saturday, 10.30 AM’\n2. | jumapili, saa tatu na robu asubu
{"id":"book_09_09","context":"Here are some numbers written in Danish:\n\ntre | 3 | tredive | 30\n\nfire | 4 | fyrre | 40\n\nfem | 5 | syvoghalvtreds | 57\n\nseks | 6 | tres | 60\n\nsyv | 7 | otteoghalvfjerds | 78\n\ntyve | 20 | firs | 80","query":"- Write in numerals:\n\n- treogtyve\n\n- seksoghalvtreds\n\n- fireogtres\n\n- femoghalvfjerds\n\n- syvoghalvfems\n\n- [blank]\n\n- Write in Danish: 8, 27, 36, 65, 98.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"text_to_num","eval_type":"single","reasoning_trace":"9.9. Danish\n\n-\n\n- 23\n\n- 56\n\n- 64\n\n- 75\n\n- 97\n\n-\n\n- 8 = otte\n\n- 27 = syvogtyve\n\n- 36 = seksogtredive\n\n- 65 = femogtres\n\n- 98 = otteoghalvfems\n\nRules:\n\nThe Danish system is based on scores (20s) rather than tens, thus: tres (3) treds (20 3 = 60); fire (4) firs (20 4 = 80). Note that there are special words for 20, 30, and 40.\n\nThe particle halv- attached before means -10 (‘halfway towards’), while noting that the word might undergo some phonological changes (tres – treds, firs – fjerds): treds (60) halvtreds (60 - 10 = 50).\n\nThe numbers are formed following the structure UogS (U represents units and S scores; og means ‘and’).","source":"langsci_420","problem_group_id":"langsci420:9.9","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Danish","author":"Michael Swan","competition":"UKLO","year":2012,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/09-Numbers.tex","source_line_start":1148,"source_line_end":1176,"solution_line_start":1433,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some numbers written in Danish:\n\ntre | 3 | tredive | 30\n\nfire | 4 | fyrre | 40\n\nfem | 5 | syvoghalvtreds | 57\n\nseks | 6 | tres | 60\n\nsyv | 7 | otteoghalvfjerds | 78\n\ntyve | 20 | firs | 80\n- Write in numerals:\n\n- treogtyve\n\n- seksoghalvtreds\n\n- fireogtres\n\n- femoghalvfjerds\n\n- syvoghalvfems\n\n- [blank]\n\n- Write in Danish: 8, 27, 36, 65, 98."}
{"id":"book_09_10","context":"The following expressions show how to tell the time in Estonian:\n\n[VISUAL OMITTED: figures/Estonian_1.pdf] | [VISUAL OMITTED: figures/Estonian_2.pdf] | [VISUAL OMITTED: figures/Estonian_3.pdf]\nKell on üks | Kell on kaks | Veerand kaks\n\n[VISUAL OMITTED: figures/Estonian_4.pdf] | [VISUAL OMITTED: figures/Estonian_5.pdf] | [VISUAL OMITTED: figures/Estonian_6.pdf]\nPool neli | Kolmveerand üksteist | Viis minutit üks läbi\n\nHere are some numbers in Estonian:\n\n- 6 = kuus\n\n- 7 = seitse\n\n- 8 = kaheksa\n\n- 9 = üheksa\n\n- 10 = kümme","query":"- What do the following Estonian time expressions mean? Write with numbers:\n\n- Kakskümmend viis minutit üheksa läbi\n\n- Veerand neli\n\n- Pool kolm\n\n- Kolmveerand kaksteist\n\n- Kolmkümmend viis minutit kuus läbi\n\n- Write the following times in Estonian:\n\n- 8:45\n\n- 4:15\n\n- 11:30\n\n- 7:05\n\n- 12:30","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"numbers","eval_type":"single","reasoning_trace":"9.10. Estonian\n\n-\n\n- 9:25\n\n- 3:15\n\n- 2:30\n\n- 11:45\n\n- 6:35\n\n-\n\n- kolmveerand üheksa\n\n- veerand viis\n\n- pool kaksteist\n\n- viis minutit seitse läbi\n\n- pool üks\n\n-\n\nRules:\n\n- X:00 = kell on X\n\n- X:15 = veerand (X+1)\n\n- X:30 = pool (X+1)\n\n- X:45 = kolmveerand (X+1)\n\n- Else, X:Y = Y minutit X läbi\n\n- 10+X = X-teist\n\n- 10X+Y = X-kümmend Y\n\n-","source":"langsci_420","problem_group_id":"langsci420:9.10","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Estonian","author":"Babette Verhoeven-Newsome","competition":"UKLO","year":2014,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":false,"visual_dependency_reasons":["included_image"],"source_file":"chapters/09-Numbers.tex","source_line_start":1178,"source_line_end":1219,"solution_line_start":1469,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nThe following expressions show how to tell the time in Estonian:\n\n[VISUAL OMITTED: figures/Estonian_1.pdf] | [VISUAL OMITTED: figures/Estonian_2.pdf] | [VISUAL OMITTED: figures/Estonian_3.pdf]\nKell on üks | Kell on kaks | Veerand kaks\n\n[VISUAL OMITTED: figures/Estonian_4.pdf] | [VISUAL OMITTED: figures/Estonian_5.pdf] | [VISUAL OMITTED: figures/Estonian_6.pdf]\nPool neli | Kolmveerand üksteist | Viis minutit üks läbi\n\nHere are some numbers in Estonian:\n\n- 6 = kuus\n\n- 7 = seitse\n\n- 8 = kaheksa\n\n- 9 = üheksa\n\n- 10 = kümme\n- What do the following Estonian time expressions mean? Write with numbers:\n\n- Kakskümmend viis minutit üheksa läbi\n\n- Veerand neli\n\n- Pool kolm\n\n- Kolmveerand kaksteist\n\n- Kolmkümmend viis minutit kuus läbi\n\n- Write the following times in Estonian:\n\n- 8:45\n\n- 4:15\n\n- 11:30\n\n- 7:05\n\n- 12:30"}
{"id":"book_09_11","context":"Here are some equalities written in Waorani. Each sequence printed in bold represents one number from 1 to 10.\n\n- mẽña mẽña mẽña mẽña + mẽña go mẽña = ãẽmãẽmpoke go aroke 2\n\n- aroke2 + mẽña2 = ãẽmãẽmpoke\n\n- ãẽmãẽmpoke go aroke2 = mẽña go mẽña ãẽmãẽmpoke mẽña go mẽña\n\n- mẽña ãẽmãẽmpoke = tipãẽmpoke","query":"- Write in Waorani the numbers from 4 to 10.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"numbers","eval_type":"single","reasoning_trace":"9.11. Waorani\n\n-\n\n- mẽña go mẽña\n\n- ãẽmãẽmpoke\n\n- ãẽmãẽmpoke go aroke\n\n- ãẽmãẽmpoke go mẽña\n\n- mẽña mẽña mẽña mẽña\n\n- ãẽmãẽmpoke mẽña go mẽña\n\n- tipãẽmpoke\n\n-","source":"langsci_420","problem_group_id":"langsci420:9.11","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Waorani","author":"Dragomir R. Radev","competition":"UKLO","year":2012,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/09-Numbers.tex","source_line_start":1221,"source_line_end":1234,"solution_line_start":1511,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some equalities written in Waorani. Each sequence printed in bold represents one number from 1 to 10.\n\n- mẽña mẽña mẽña mẽña + mẽña go mẽña = ãẽmãẽmpoke go aroke 2\n\n- aroke2 + mẽña2 = ãẽmãẽmpoke\n\n- ãẽmãẽmpoke go aroke2 = mẽña go mẽña ãẽmãẽmpoke mẽña go mẽña\n\n- mẽña ãẽmãẽmpoke = tipãẽmpoke\n- Write in Waorani the numbers from 4 to 10."}
{"id":"book_09_12","context":"Here are some numbers in Selkup and their values in random order:\n\nsomplylasar εj šitty, muktyssar εj ukkyr, sompylasar εj sompyla, šittysar, ukkyr ca muktyssar, šitty ca tɛ̄sar, sompylasar εj sel’cy, ukkyr ca tōn\n20, 38, 52, 55, 57, 59, 61, 99\n\nA bar above a vowel denotes length.","query":"- Determine the correct correspondences.\n\n- Write in Selkup: 41, 48, 77, 98.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"9.12. Selkup\n\n-\nsomplylasar εj šitty | = 52 | muktyssar εj ukkyr | = 61\nsompylasar εj sompyla | = 55 | šittysar | = 20\nukkyr ca muktyssar | = 59 | šitty ca tɛ̄sar | = 38\nsompylasar εj sel’cy | = 57 | ukkyr ca tōn | = 99\n\n-\n41 = tɛ̄sar εj ukkyr | 48 = šitty ca sompylasar\n77 = sel’cysar εj sel’cy | 98 = šitty ca tōn\n\nRules:\n\n- Base-10 subtractive system.\n\n- 10X = X-sar 100 = tōn\n\n- 10X + Y = X-sar εj Y Y = (1,7)\n\n- 10X + Y = (10-Y) ca (X+1)-sar Y=(8,9)","source":"langsci_420","problem_group_id":"langsci420:9.12","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Selkup","author":"Svetlana Burlak","competition":"MSK","year":1991,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/09-Numbers.tex","source_line_start":1237,"source_line_end":1253,"solution_line_start":1530,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some numbers in Selkup and their values in random order:\n\nsomplylasar εj šitty, muktyssar εj ukkyr, sompylasar εj sompyla, šittysar, ukkyr ca muktyssar, šitty ca tɛ̄sar, sompylasar εj sel’cy, ukkyr ca tōn\n20, 38, 52, 55, 57, 59, 61, 99\n\nA bar above a vowel denotes length.\n- Determine the correct correspondences.\n\n- Write in Selkup: 41, 48, 77, 98."}
{"id":"book_09_13","context":"Below are given the following equalities written in Vambon in random order:\n31 = 3 32 = 6 ... 39 = 27\n\n0in\n\n- takhem ambalop = emkelop\n\n- takhem hitulop = silutop\n\n- takhem javet = emsanop\n\n- takhem kumuk = emalin\n\n- takhem mben = emben\n\n- takhem muyop = emhitulop\n\n- takhem sanop = takhem\n\n- takhem sanopkunip = kumuk\n\n- takhem takhem = javet\n\n- [blank]","query":"- Write in numerals:\n\n- (emnggokmit + nggokmit) sanopkunip = kalit\n\n- emambalop - emjavet = hitulop\n\n- Write in Vambon:\n\n- 20 + 22 - 26 = 16\n\n- 12 + 13 = 25\n\n- You are given the following Vambon words:\n\nkalit | ‘nose’ | kelop | ‘eye’\n\nkumuk | ‘wrist’ | muyop | ‘elbow’\n\nnggokmit | ‘neck’ | sanopkunip | ‘ring finger’\n\nsilutop | ‘ear’ |\n\n- Which body parts do the following words refer to?\n\n- ambalop\n\n- javet\n\n- malin\n\n- mben\n\n- sanop\n\n- [blank]\n\n- What does the prefix em- mean in Vambon?","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"text_to_num","eval_type":"single","reasoning_trace":"9.13. Vambon\n\nBody-part-based system, centred on 14.\n\n- a. (17+11) 2 = 14 b. 23-19=4\n\n- c. emuyop + emkumuk – emsanopkunip = emsilutop\n0in\n\n- d. silutop + kelop = emtakhem\n\n-\n\n- ‘thumb’\n\n- ‘arm’\n\n- ‘shoulder’\n\n- ‘forearm’\n\n- ‘little finger’\n\n-\n\n- The prefix em- means ‘other / opposite’. Note: em e / _ m.","source":"langsci_420","problem_group_id":"langsci420:9.13","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Vambon","author":"Alexander Piperski","competition":"Elementy","year":null,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/09-Numbers.tex","source_line_start":1255,"source_line_end":1313,"solution_line_start":1556,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nBelow are given the following equalities written in Vambon in random order:\n31 = 3 32 = 6 ... 39 = 27\n\n0in\n\n- takhem ambalop = emkelop\n\n- takhem hitulop = silutop\n\n- takhem javet = emsanop\n\n- takhem kumuk = emalin\n\n- takhem mben = emben\n\n- takhem muyop = emhitulop\n\n- takhem sanop = takhem\n\n- takhem sanopkunip = kumuk\n\n- takhem takhem = javet\n\n- [blank]\n- Write in numerals:\n\n- (emnggokmit + nggokmit) sanopkunip = kalit\n\n- emambalop - emjavet = hitulop\n\n- Write in Vambon:\n\n- 20 + 22 - 26 = 16\n\n- 12 + 13 = 25\n\n- You are given the following Vambon words:\n\nkalit | ‘nose’ | kelop | ‘eye’\n\nkumuk | ‘wrist’ | muyop | ‘elbow’\n\nnggokmit | ‘neck’ | sanopkunip | ‘ring finger’\n\nsilutop | ‘ear’ |\n\n- Which body parts do the following words refer to?\n\n- ambalop\n\n- javet\n\n- malin\n\n- mben\n\n- sanop\n\n- [blank]\n\n- What does the prefix em- mean in Vambon?"}
{"id":"book_09_14","context":"Here are some Alamblak numbers and their numerical values in random order:\n\n0em\n\n- yima hosfirpati tir hosfirpat\n\n- yima yohtti tir hosfi rpat\n\n- yima hosfi hosf\n\n- yima hosfi tir hosf\n\n- yima yohtti tir yohtti rpat\n\n- yima hosfirpati tir hosfi hosfirpat\n\n- yima hosfirpati tir hosfirpati hosfirpat\n\n- yima yohtti tir hosfirpati rpat\n\n- yima hosfihosfi tir yohtti hosfihosf\n\n- yima hosfi tir hosfi hosf\n\n26, 31, 36, 42, 50, 52, 73, 75, 78, 89\nMoreover, it is known that:\n\n1 = rpat | 2 = hosf | 3 = hosfirpat | 4 = hosfihosf\n\n5 = tir yohtt | 6 = tir yohtti rpat | 11 = tir hosfi rpat","query":"- Determine the correct correspondences.\n\n- Write in numerals:\n\n- yima hosfirpati hosfihosf + yima yohtti tir hosf =\n* = yima hosfihosfi tir hosfi hosfihosf\n\n- tir yohtti hosf + tir hosfi hosf = tir hosfirpati hosfihosf\n\n- Write in Alamblak: 21, 48, 83.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"9.14. Alamblak\n\n-\n\n- 71\n\n- 31\n\n- 42\n\n- 50\n\n- 26\n\n- 73\n\n- 78\n\n- 36\n\n- 89\n\n- 52\n\n- k. 64+30=97 l. 7+12=19\n\n- 21 = yima yohtti rpat\n\n- 48 = yima hosfi tir yohtti hosfirpat\n\n- 83 = yima hosfihosfi hosfirpat\n\nRules:\n\n- General structure: yima X-i tir Y-i Z = 20X + 5Y + Z\n\n- The digit 1 has two different forms: rpat (if it is Z), yohtti (for X or Y)\n\n- Thus, multiplication is implied, while addition is marked by the suffix -i","source":"langsci_420","problem_group_id":"langsci420:9.14","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Alamblak","author":"Roxana Dincă","competition":"RoLO","year":2013,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/09-Numbers.tex","source_line_start":1315,"source_line_end":1354,"solution_line_start":1581,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some Alamblak numbers and their numerical values in random order:\n\n0em\n\n- yima hosfirpati tir hosfirpat\n\n- yima yohtti tir hosfi rpat\n\n- yima hosfi hosf\n\n- yima hosfi tir hosf\n\n- yima yohtti tir yohtti rpat\n\n- yima hosfirpati tir hosfi hosfirpat\n\n- yima hosfirpati tir hosfirpati hosfirpat\n\n- yima yohtti tir hosfirpati rpat\n\n- yima hosfihosfi tir yohtti hosfihosf\n\n- yima hosfi tir hosfi hosf\n\n26, 31, 36, 42, 50, 52, 73, 75, 78, 89\nMoreover, it is known that:\n\n1 = rpat | 2 = hosf | 3 = hosfirpat | 4 = hosfihosf\n\n5 = tir yohtt | 6 = tir yohtti rpat | 11 = tir hosfi rpat\n- Determine the correct correspondences.\n\n- Write in numerals:\n\n- yima hosfirpati hosfihosf + yima yohtti tir hosf =\n* = yima hosfihosfi tir hosfi hosfihosf\n\n- tir yohtti hosf + tir hosfi hosf = tir hosfirpati hosfihosf\n\n- Write in Alamblak: 21, 48, 83."}
{"id":"book_09_15","context":"Below are given the first four multiples of the number efi tʃumtʃum eku bab eku iŋki, written in Chabu in ascending order (if the number is X, the four numbers below represent 2X, 3X, 4X, and 5X, respectively):\n\n- bab ef eku efi tʃumtʃum eku iŋki\n\n- ink ufe kor eku bab eku bab\n\n- ink ufe kor eku bab ef eku bab\n\n- bab ufe kor","query":"- Write in Chabu all the divisors of the number bab eku iŋki ufe kor (including 1). If you consider some of them can be written in different ways, write all the possibilities.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"numbers","eval_type":"single","reasoning_trace":"9.15. Chabu\n\n-\n\n- iŋki / ink\n\n- bab\n\n- bab eku iŋki\n\n- bab eku bab\n\n- efi tʃumtʃum\n\n- efi tʃumtʃum eku iŋki\n\n- 10 bab ef\n\n- 12 bab ef eku bab\n\n- 15 bab ef eku efi tʃumtʃum\n\n- 20 ink ufe kor\n\n- 30 ink ufe kor eku bab ef\n\n-\n\nRules:\n\n- General structure:\n\n- 2 [+ X] = bab [eku X] (X < 3)\n\n- 5 [+ X] = efi tʃumtʃum [eku X] (X < 5)\n\n- 10 [+ X] = bab ef [eku X] (X < 10)\n\n- 20Y [+ X] = Y ufe kor [eku X] (X < 20)\n\n- Y is written even if it is 1 (20 = ink ufe kor).\n\n- The digit 1 has two forms: ink (if it is a multiplier) or iŋki (if it is added). In task (a), 1 has two alternative spellings, since we do not know which one to choose if it appears as a single word.","source":"langsci_420","problem_group_id":"langsci420:9.15","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Chabu","author":"Danylo Mysak","competition":"UkrLO","year":2018,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/09-Numbers.tex","source_line_start":1356,"source_line_end":1367,"solution_line_start":1613,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nBelow are given the first four multiples of the number efi tʃumtʃum eku bab eku iŋki, written in Chabu in ascending order (if the number is X, the four numbers below represent 2X, 3X, 4X, and 5X, respectively):\n\n- bab ef eku efi tʃumtʃum eku iŋki\n\n- ink ufe kor eku bab eku bab\n\n- ink ufe kor eku bab ef eku bab\n\n- bab ufe kor\n- Write in Chabu all the divisors of the number bab eku iŋki ufe kor (including 1). If you consider some of them can be written in different ways, write all the possibilities."}
{"id":"book_09_16","context":"Here are some equalities written in Tifal. It is known that none of the numbers in the problem are greater than 30.\n\n0in\n\n- asumano aleeb = bokob\n\n- asumano ataling = tadang\n\n- bokob ataling = ataling madi\n\n- bokob asumano = nakal madi\n\n- asumano feet = feet madi\n\n- ataling ataling = tadang madi\n\n- asumano + ataling = feet\n\n- feet + miit = feet madi\n\n- tadang + ataling = tadang madi\n\n- [blank]","query":"- Write in numerals:\n\n- beeti + nakal = beeti madi\n\n- bokob + maakob = feet\n\n- awok awok = asumano madi\n\n- Write in Tifal the results of the following equalities:\n\n- tadang + miit =\n\n- ataling madi - aleeb =","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"text_to_num","eval_type":"single","reasoning_trace":"9.16. Tifal\n\n-\n\n- 9+10=19\n\n- 6+1=7\n\n- 55=25\n\n-\n\n- tadang + miit = aleeb madi\n\n- ataling madi – aleeb = bokob madi\n\nRules:\n\nBody-part-based system, centred on 14. X madi = 28 – X.","source":"langsci_420","problem_group_id":"langsci420:9.16","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Tifal","author":"Svetlana Burlak & Peter Arkadiev","competition":"MSK","year":2017,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/09-Numbers.tex","source_line_start":1369,"source_line_end":1399,"solution_line_start":1650,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some equalities written in Tifal. It is known that none of the numbers in the problem are greater than 30.\n\n0in\n\n- asumano aleeb = bokob\n\n- asumano ataling = tadang\n\n- bokob ataling = ataling madi\n\n- bokob asumano = nakal madi\n\n- asumano feet = feet madi\n\n- ataling ataling = tadang madi\n\n- asumano + ataling = feet\n\n- feet + miit = feet madi\n\n- tadang + ataling = tadang madi\n\n- [blank]\n- Write in numerals:\n\n- beeti + nakal = beeti madi\n\n- bokob + maakob = feet\n\n- awok awok = asumano madi\n\n- Write in Tifal the results of the following equalities:\n\n- tadang + miit =\n\n- ataling madi - aleeb ="}
{"id":"book_09_17","context":"Here are some numbers in Mansi (written in Latin script):\n\nńollow | 8\n\natxujplow | 15\n\natlow nopъl ontъllow | 49\n\natlow | 50\n\nontъlsāt ontъllow | 99\n\nxōtsātn xōtlow nopъl at | 555\n\nontъllowsāt | 900\n\nontъllowsāt ńollowxujplow | 918","query":"- Write in numerals:\n\n6.0pt plus 2.0pt minus 1.5pt\n\n- atsātn at\n\n- ńolsāt nopъl xōt\n\n- ontъllowsātn ontъllowxujplow\n\n- Write in Mansi: 58, 80, 716.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"text_to_num","eval_type":"single","reasoning_trace":"9.17. Mansi\n\n-\n\n- 405\n\n- 76\n\n- 819\n\n- 58 = xōtlow nopъl ńollow\n\n- 80 = ńolsāt\n\n- 716 = ńollowsātn xōtxujplow\n\nRules:\nSystem based largely on overcounting. Tens are formed from units: adding the suffix -low (for 50, 60) or replacing the suffix -low with -sāt (for 80, 90).\n\n- 10 + X = X-xujplow\n\n- 10X + Y = 10(X+1) nopъl Y\n\n- 90 + X = ontъlsāt X\n\n- 100X = X-sāt\n\n- 100X + Y = (X+1)-sātn Y\n\n- 900 + X = ontъllowsāt X","source":"langsci_420","problem_group_id":"langsci420:9.17","chapter":9,"chapter_title":"Number systems","section":6,"section_title":"Practice problems","topic":"number systems","language":"Mansi","author":"Ivan Derzhanski","competition":"IOL","year":2005,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c09_s01_p01","book_method_c09_s01_p02","book_method_c09_s02_p01","book_method_c09_s03_p01","book_method_c09_s04_p01","book_method_c09_s05_p01"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":true,"visual_dependency_reasons":[],"source_file":"chapters/09-Numbers.tex","source_line_start":1401,"source_line_end":1428,"solution_line_start":1673,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Number systems\nPractice problems\nnumber systems\nThis practice problem belongs to the book's Number systems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nHere are some numbers in Mansi (written in Latin script):\n\nńollow | 8\n\natxujplow | 15\n\natlow nopъl ontъllow | 49\n\natlow | 50\n\nontъlsāt ontъllow | 99\n\nxōtsātn xōtlow nopъl at | 555\n\nontъllowsāt | 900\n\nontъllowsāt ńollowxujplow | 918\n- Write in numerals:\n\n6.0pt plus 2.0pt minus 1.5pt\n\n- atsātn at\n\n- ńolsāt nopъl xōt\n\n- ontъllowsātn ontъllowxujplow\n\n- Write in Mansi: 58, 80, 716."}
{"id":"book_10_01","context":"Manam Pile is a Malayo-Polynesian language spoken on Manam Island off the coast of Papua New Guinea. Manam is one of the most active volcanoes in the world. Below, a Manam islander describes the relative locations of the houses shown on the map.\n\n- Onkau pera kana auta ieno, Kulu pera kana ilau ieno.\n\n- Mombwa pera kana ata ieno, Kulu pera kana awa ieno.\n\n- Tola pera kana auta ieno, Sala pera kana ilau ieno.\n\n- Sulung pera kana awa ieno, Tola pera kana ata ieno.\n\n- Sala pera kana awa ieno, Mombwa pera kana ata ieno.\n\n- Pita pera kana ilau ieno, Sulung pera kana auta ieno.\n\n- Sala pera kana awa ilau ieno, Onkau pera kana ata auta ieno.\n\n- Butokang pera kana awa auta ieno, Pita pera kana ata ilau ieno.\n\n[VISUAL OMITTED: images/Manam.png]","query":"- Onkau's, Mombwa's and Kulu's houses have already been located on the map above. Who lives in the other five houses (A-E)?\n\n- Arongo is building his new house in the location marked with an X on the map. In three Manam Pile sentences like the ones on the previous page, describe the location of Arongo's house in relation to the three closest houses (A-C).","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"The first step is identifying the structure of the sentences. This is:\n\nH1 pera kana X1 ieno, H2 pera kana X2 ieno.\nwhere H1 and H2 refer to the names of the houses. Therefore, X1 and X2 must represent the words for directions. We notice that these can be: ata, auta, awa, ilau. Moreover, based on the examples 7 and 8, we notice that they can also combine with one another: ata ilau, awa auta, awa ilau, ata auta. Therefore, we deduce that there are two main directions or axes: ata–awa and ilau–auta, which can combine together (similarly to how the directions north–south and east–west can combine to form directions like north-east, north-west, etc.). Furthermore, we notice that X1 is always the opposite of X2 (similar to the English sentences ‘A is due east, and B due west’ or ‘A is north-east, and B south-west’).\n\nBased on this information, we can rewrite the eight sentences above in a simpler way:\n\n- ‘Onkau is to the’ auta ‘of Kulu.’\n\n- ‘Mombwa is to the’ ata ‘of Kulu.’\n\n- ‘Tola is to the’ auta ‘of Sala.’\n\n- ‘Sulung is to the’ awa ‘of Tola.’\n\n- ‘Sala is to the’ awa ‘of Mombwa.’\n\n- ‘Pita is to the’ ilau ‘of Sulung.’\n\n- ‘Sala is to the’ awa ilau ‘of Onkau.’\n\n- ‘Butokang is to the’ awa auta ‘of Pita.’\n\nThe first hypothesis that comes to mind is that indeed, ata, awa, ilau, and auta represent the ‘north’, ‘south’, ‘east’, and ‘west’. Thus, based on sentence 1, we deduce that auta = ‘north’, and based on 2 we deduce that ata = ‘west’. Moreover, knowing already that the two main directions are ata – awa and ilau – auta, we can also deduce the other two cardinal points, mainly awa = ‘east’ and ilau = ‘south’.\n\nThus, from sentence 7, we find out that Sala's house is to the south-east of Onkau's, so Sala's house can only be E, which is also confirmed by sentence 5, which mentions that Sala's house is to the east of Mombwa's. Next, from sentence 3, we deduce that Tola's house is to the north of Sala's, so Tola's house can only be A or C, but, since sentence 4 says that Sulang's house is to the east of Tola's, Tola's house needs to be C and Sulung's A (if Tola's house were A, there is no other house to the east of it). From sentence 6, Pita's house is to the south of Sulung's, so it is most likely D. Finally, according to sentence 8, Butokang's house is to the north-east of Pita's, but, at the same time, we know that Butokang's house must be B, since it is the only house left. Thus, we reach a contradiction which points to the fact that, most likely, the four words are not cardinal points.\n\nIf we carefully read the introduction, we learn that the island is in fact a volcano. Taking this into account, we can interpret the fir
{"id":"book_10_02","context":"Below you can find part of an Arawak family tree. Three family members – two men and a woman (not necessarily in this order) – describe their family:\n\n- De to Fatan. Onikhan to dajo. Mithakotoan ken Kolhen to dakhithonon. Tholhady to dato. Ematonoa to dathi.\n\n- De to Sobole. Bokoa to dakhithi. Kolhen ken Moty to dajonon. Balhose ken Konoko to dathinon. Onikhan ken Ylhydaba to dakythynon. Fatan ken Mithakotoan to dajaboathonon. Sareke to dajorodatho. Ematonoa to dadokothi.\n\n- De to Balhose. Ylhydaba to dajo. Kolhen to daretho. Sobole ken Bokoa to daithinon. Konoko to dakhithi. Moty to dajorodatho. Mithakotoan ken Fatan to darebiathonon. Sareke to dato.\n\n[VISUAL OMITTED: figures/Arawak.pdf]\n\nIn the diagram above, triangles represent men and circles represent women. Horizontal lines represent siblings and vertical lines children. The equals sign denotes marriage.","query":"- Supply the tree with names. If multiple options are possible, provide them all.\n\n- Three more people (from the same family) describe their family:\n\n- De to Ajonym. Fatan ken Kolhen to [blank]. Onikhan to dakythy. Mithakotoan to dajo. [blank] to dadokothi.\n\n- De to [blank]. Balhose ken Konoko to daithinon. Holholho ken Sobole ken [blank] to dalykynthinon. Moty to dato. [blank] to dalykyntho.\n\n- De to Kolhen. [blank] to dajo. [blank] to darethi. Ematonoa to\n[blank]. Sareke to dato. Sobole ken Bokoa to [blank]. Konoko to [blank].\n\n- Fill each gap with exactly one word.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"Firstly, we look at the structure of the sentences. Each sentence, except for the first, follows the pattern Name to X. So we know that X represents the kinship term. We deduce that ko = ‘is/are’, and the first sentence most likely means ‘I am X’, therefore de = ‘I’. Moreover, we notice that if in the rest of the sentences there are more names co-occurring, they are separated by ken, so this word most likely means ‘and’. Last but not least, we notice that every time that a sentence contains more than one name, the kinship term ends in non, so we deduce that the suffix -non marks the plural.\n\nBased on this we can make a table in which we show the relations between the persons mentioned in the data:\n\n| Fatan | Sobole | Balhose\n| Fatan | Sobole | Balhose\nFatan | 808080 | dajaboatho | darebiatho\nOnikhan | dajo | dakythy |\nMithakotoan | dakhitho | dajaboatho | darebiatho\nKolhen | dakhitho | dajo | daretho\nTholhady | dato | |\nEmatonoa | dathi | dadokothi |\nSobole | | 808080 | daithi\nBokoa | | dakhithi | daithi\nMoty | | dajo | dajorodatho\nBalhose | | dathi | 808080\nKonoko | | dathi | dakhithi\nYlhydaba | | dakythy | dajo\nSareke | | dajorodatho | dato\n\nMoreover, we notice that two names are missing: Ajonym and Holholho (both of them found in task (b)).\n\nGenerally, the first step in solving this type of problem, once the table above is made, is assuming that if two persons have the same relation with a third, then those two persons belong to the same generation (e.g., if A and B are m to person X – or X is m to persons A and B – then, most likely, A and B are part of the same generation, i.e., they are on the same level of the tree). In this case, we notice from the diagram that we have three generations: the top one, which has three members, the middle one (siz members) and the bottom one (six members).\n\nStarting with Fatan, we notice that they are the same relation (dakhitho) to both Mithakotoan and Kolhen, so we can assume that these two belong to the same generation. Similarly, we deduce that Fatan and Mithakotoan belong to the same generation (since they both are dajaboatho to Sobole). Knowing that Mithakotoan and Kolhen as well as Mithakotoan and Fatan belong to the same generation, we can deduce that all three of them are part of the same generation.\n\nUsing the same thought process, we get the following pairs of persons belonging to the same generation: Kolhen <20><>
{"id":"book_10_03","context":"In the picture below, note that both you and the speaker are facing the paper. The bird is to the left of everything else and the kangaroo is to the right of everything else. The cat is behind everything else and the kangaroo is in front of everything else.\n\n[VISUAL OMITTED: images/Bardi.png]\n\nHere are some Bardi sentences describing the scene:\n\n- Aamba bornkony yaawardon.\n\n- Baawa joorroonggony garrabalgoon.\n\n- Boorroo alaboor yaawardon.\n\n- Iila alaboor ooranygoon.\n\n- Iila baybirrony aambon.\n\n- Minyaw baybirrony baawon.\n\n- Oorany joorroonggony baawon.\n\n- Yaawarda bornkony aambon.","query":"- Based on these, determine the following correspondences:\n\n. | Aarlgoodony | . | ‘bird’\n. | Aamba | . | ‘child’\n. | Alaboor | . | ‘cat’\n. | Baawa | . | ‘dog’\n. | Baybirrony | . | ‘horse’\n. | Boorroo | . | ‘kangaroo’\n. | Bornkony | . | ‘man’\n. | Garrabal | . | ‘woman’\n. | Iila | . | ‘next to’\n. | Joorroonggony | . | ‘behind’\n. | Minyaw | . | ‘in front of’\n. | Oorany | . | ‘to the left of’\n. | Yaawarda | . | ‘to the right of’","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"match_letters","eval_type":"single","reasoning_trace":"10.3. Bardi\n\n- 1-l, 2-g, 3-k, 4-b, 5-j, 6-f, 7-i, 8-a, 9-d, 10-m, 11-c, 12-h, 13-e.","source":"langsci_420","problem_group_id":"langsci420:10.3","chapter":10,"chapter_title":"Other types of problems","section":3,"section_title":"Practice problems","topic":"orientation, kinship, and other structural problems","language":"Bardi","author":"Catherine Sheard","competition":"UKLO","year":2012,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c10_s01_p01","book_method_c10_s02_p01","book_method_c10_s02_p02","book_method_c10_s02_p03"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":false,"visual_dependency_reasons":["included_image"],"source_file":"chapters/10-OtherProblems.tex","source_line_start":351,"source_line_end":389,"solution_line_start":651,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Other types of problems\nPractice problems\norientation, kinship, and other structural problems\nThis practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nIn the picture below, note that both you and the speaker are facing the paper. The bird is to the left of everything else and the kangaroo is to the right of everything else. The cat is behind everything else and the kangaroo is in front of everything else.\n\n[VISUAL OMITTED: images/Bardi.png]\n\nHere are some Bardi sentences describing the scene:\n\n- Aamba bornkony yaawardon.\n\n- Baawa joorroonggony garrabalgoon.\n\n- Boorroo alaboor yaawardon.\n\n- Iila alaboor ooranygoon.\n\n- Iila baybirrony aambon.\n\n- Minyaw baybirrony baawon.\n\n- Oorany joorroonggony baawon.\n\n- Yaawarda bornkony aambon.\n- Based on these, determine the following correspondences:\n\n. | Aarlgoodony | . | ‘bird’\n. | Aamba | . | ‘child’\n. | Alaboor | . | ‘cat’\n. | Baawa | . | ‘dog’\n. | Baybirrony | . | ‘horse’\n. | Boorroo | . | ‘kangaroo’\n. | Bornkony | . | ‘man’\n. | Garrabal | . | ‘woman’\n. | Iila | . | ‘next to’\n. | Joorroonggony | . | ‘behind’\n. | Minyaw | . | ‘in front of’\n. | Oorany | . | ‘to the left of’\n. | Yaawarda | . | ‘to the right of’"}
{"id":"book_10_04","context":"Below is shown the kinship tree of a Kharia family in which circles represent women and squares represent men. The age of each person is written under their name.\n\n[VISUAL OMITTED: figures/Kharia.pdf]\n\nEach member of this family says something about their family in Kharia:\n\n- Bhaiiɲaʔ ɲimi Thuyu.\n\n- Didiiɲaʔ ɲimi Muni. Ɖonkuiiɲaʔ ɲimi Mariya.\n\n- Sowiɲaʔ ɲimi Nuh.\n\n- Didikiiɲaʔ ɲimiki Olem oɖoyoʔ no Ewa.\n\n- Konon bahiniɲaʔ ɲimi Olem. Didikiiɲaʔ ɲimiki Ewa oɖoyoɁ no Kolo.\n\n- Ɖonkuiiɲaʔ ɲimi Muni. Dad=iɲaʔ ɲimi Dele.\n\n- Sowiɲaʔ ɲimi Thuyu. Beʔʈiɲaʔ ɲimi Anil.\n\n- Dadakiiɲaʔ ɲimiki Beni oɖoyoʔ no Sim.\n\n- Beʔʈiɲaʔ ɲimi Beni. Ɖonkuiiɲaʔ ɲimi Mariya.\n\n- Bhaikiiɲaʔ ɲimiki Beni oɖoyoʔ no Anil. Apaiɲaʔ ɲimi Dele.\n\n- Konon bahinkiiɲaʔ ɲimiki Kolo, Ewa oɖoyoʔ no Olem.\n\n- Konon bahinkiiɲaʔ ɲimiki Muni oɖoyoʔ no Kepka.\n\n- Apaiɲaʔ ɲimi Nuh. Dad=iɲaʔ ɲimi Sim.","query":"- Assign each of the sentences above to the person who uttered it.\n\n- Fill in the blanks:\n\n- Muni: “[blank] ɲimi Dele.”\n\n- Kepka: “[blank] ɲimi Nuh.”\n\n- Ewa: “Konon bahinkiiɲaʔ ɲimiki [blank].”\n\n- Anil: “Dad=iɲaʔ ɲimi [blank].”\n\n- Beni: “[blank] ɲimi Dele.”\n\n- A few years later, Kepka has a son named Caitu. Fill in the blanks:\n\n- Sim: “[blank] ɲimi Caitu.”\n\n- Kepka: “[blank] ɲimi Caitu.”\n\n- Caitu: “[blank] ɲimiki Ewa, Olem oɖoyoʔ no Kolo.”","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"10.4. Kharia\n\n-\n\n- Dele\n\n- Kepka\n\n- Mariya\n\n- Anil\n\n- Beni\n\n- Thuyu\n\n- Rut\n\n- Olem\n\n- Muni\n\n- Ewa\n\n- Sim\n\n- Nuh\n\n- Kolo\n\n-\n\n- Sowiɲaʔ\n\n- Dad=iɲaʔ\n\n- Kolo oɖoyoʔ no Olem\n\n- Beni\n\n- Apaiɲaʔ\n\n-\n\n-\n\n- Bhaiiɲaʔ\n\n- Beʔʈiɲaʔ\n\n- Didikiiɲaʔ\n\nRules:\n\nSentence structure: [Kinship term]–(pl)–iɲaʔ ɲimi–(pl) [Persons]\n\n- pl = -ki- = plural marker (if there is more than one person)\n\n- Persons: Their names; if it's more than one person, oɖoyoʔ no = ‘and’ is added before the last one.\n\n- Kinship terms:\n\n- bhai = ‘younger brother’\n\n- dad= = ‘older brother’ (plural: dadaki)\n\n- konon bahin = ‘younger sister’\n\n- didi = ‘older sister’\n\n- beʔʈ = ‘son’\n\n- apa = ‘father’\n\n- sow = ‘husband’\n\n- đonkui = ‘brother's wife’","source":"langsci_420","problem_group_id":"langsci420:10.4","chapter":10,"chapter_title":"Other types of problems","section":3,"section_title":"Practice problems","topic":"orientation, kinship, and other structural problems","language":"Kharia","author":"Barbora Dohnalová","competition":"ČLO","year":2021,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c10_s01_p01","book_method_c10_s02_p01","book_method_c10_s02_p02","book_method_c10_s02_p03"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":false,"visual_dependency_reasons":["included_image"],"source_file":"chapters/10-OtherProblems.tex","source_line_start":391,"source_line_end":432,"solution_line_start":659,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Other types of problems\nPractice problems\norientation, kinship, and other structural problems\nThis practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nBelow is shown the kinship tree of a Kharia family in which circles represent women and squares represent men. The age of each person is written under their name.\n\
{"id":"book_10_05","context":"A tourist travels to a village on the course of the river Benoit Martinus Ambala (Kalimantan Island, Indonesia), in order to learn the Embaloh language. He will live at the Chief's House (see map). On the first day, the Chief takes his guest outside, points towards the north, south, east and west and says “urait, kalaut, anait, suali”. The tourist wrote down in his own dictionary: urait = ‘north’, kalaut = ‘south’, anait = ‘east’, suali = ‘west’.\n\n[VISUAL OMITTED: images/Embolah_EN.png]\n\nThe next day, the tourist wanted to visit the Sanctuary – the place where all the important tribal ceremonies take place. He took his dictionary and compass, but he left his map at home. Exiting the Chief's House, he started walking north and reached the Shaman's House. He asked the Shaman: “How can I get to the Sanctuary?” “Keep going urait,” replied the Shaman. “So, I should head north” thought the tourist, checking his dictionary. He crossed the river, but then he got lost in a Rice Field, so he decided to return to the Shaman. A plantation worker guided him: “Towards the Shaman's House, go suali.” “That means west?! Weird!” thought the tourist, but nevertheless he headed west. However, the river didn't come up and the rice fields were slowly replaced by coconut plantations and the tourist realised he was completely lost. “Whatever it is,” a worker on the coconut plantation started comforting the tourist, “keep going kalaut. You will get to the School and the teacher will explain everything.”\n\nChecking his dictionary, he headed south and he indeed reached the school. “Chief's House anait?” the tourist asked the teacher. “No, anait Diamond Mine. Chief's House suali” he replied. The tourist, humbled, headed west and found himself at the Diamond Mine. Extremely angry, he asked: “How can I finally reach the Chief's House or at least the School?” “Chief's House suali, but School andoor.” Unfortunately, this last word does not appear in the tourist's dictionary.","query":"- Explain why the tourist got lost and explain how the orientation system of this tribe works, as well as what each direction means.\n\n- Describe in Embolah the directions:\n\n- from the Sanctuary to the Shaman's House;\n\n- from the Bamboo Forest to the Shaman's House;\n\n- from the Shaman's House to the Rice Field.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"10.5. Embaloh\n\n- On the first day, when he was told the four directions, he automatically assumed that they represented cardinal directions. In reality, they are based on the topography of the area and represent the relative positions with respect to the river. As such:\n\n- andoor = ‘towards (closer to) the river’\n\n- anait = ‘away from the river’\n\n- urait = ‘upstream’\n\n- kalaut = ‘downstream’\n\n- suali = ‘across the river’\n\n-\n\n- kalaut\n\n- andoor\n\n- suali","source":"langsci_420","problem_group_id":"langsci420:10.5","chapter":10,"chapter_title":"Other types of problems","section":3,"section_title":"Practice problems","topic":"orientation, kinship, and other structural problems","language":"Embaloh","author":"Ksenia Gilyarova","competition":"MSK","year":2006,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c10_s01_p01","book_method_c10_s02_p01","book_method_c10_s02_p02","book_method_c10_s02_p03"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":false,"visual_dependency_reasons":["included_image"],"source_file":"chapters/10-OtherProblems.tex","source_line_start":434,"source_line_end":454,"solution_line_start":726,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu,
{"id":"book_10_06","context":"The picture below represents a field divided into 49 squares (7x7), aligned with north at the top and east on the right. In some of the squares there are rocks, indicated by a black circle ●.\n\n[VISUAL OMITTED: figures/Hungarian_Squares.pdf]\n\nThere are four Hungarian friends - A, B, C and D – standing in the field, each in a different square not containing a rock, and each facing in one of the four cardinal directions (north, south, east, west). Each person makes some statements describing the position of the rocks. For instance, A's first statement means ‘To the east (behind me) there is one stone’. References to directions are to be understood as describing a single line in the field: ‘due east’, ‘directly behind me’, and so on. No directions describe a more complex spatial relationship.\n\n- Délre két kő van.\nKeletere (mögöttem) egy kő van.\nJobbra nincs kő.\n\n- Délre (balra) nincs kő.\nÉszakra egy kő van.\nMögöttem két kő van.\n\n- Északra (előttem) nincs kő.\nNyugatra egy kő van.\nJobbra két kő van.\n\n- Nyugatra (jobbra) két kő van.\nÉszakra egy kő van.\nBalra nincs kő.","query":"- Which square is occupied by each of B, C and D? Draw an arrow (like the one under A) to show the direction they are facing.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"10.6. Hungarian\n\n[VISUAL OMITTED: figures/Hungarian_Squares_solution.pdf]\n\n- előttem = ‘front’\n\n- mögöttem = ‘back’\n\n- balra = ‘left’\n\n- jobbra = ‘right’\n\n- északra = ‘north’\n\n- délre = ‘south’\n\n- nyugatra = ‘west’\n\n- keletere = ‘east’\n\n- nincs = ‘0’\n\n- egy = ‘1’\n\n- két = ‘2’\n\n-","source":"langsci_420","problem_group_id":"langsci420:10.6","chapter":10,"chapter_title":"Other types of problems","section":3,"section_title":"Practice problems","topic":"orientation, kinship, and other structural problems","language":"Hungarian","author":"Adam Hesterberg","competition":"UKLO","year":2014,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c10_s01_p01","book_method_c10_s02_p01","book_method_c10_s02_p02","book_method_c10_s02_p03"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":false,"visual_dependency_reasons":["included_image"],"source_file":"chapters/10-OtherProblems.tex","source_line_start":456,"source_line_end":476,"solution_line_start":750,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Other types of problems\nPractice problems\norientation, kinship, and other structural problems\nThis practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nThe picture below represents a field divided into 49 squares (7x7), aligned with north at the top and east on the right. In some of the squares there are rocks, indicated by a black circle ●.\n\n[VISUAL OMITTED: figures/Hungarian_Squares.pdf]\n\nThere are four Hungarian friends - A, B, C and D – standing in the field, each in a different square not containing a rock, and each facing in one of the four cardinal directions (north, south, east, west). Each person makes some statements describing the position of the rocks. For instance, A's first statement means ‘To the east (behind me) there is one stone’. References to directions are to be understood as describing a single line in the field: ‘due east’, ‘directly behind me’, and so on. No directions describe a more complex spatial relationship.\n\n- Délre két kő van.\nKeletere (mögöttem)
{"id":"book_10_07","context":"Two linguists, Dr. David Lovelang and Dr. Matt Hateword were studying the language spoken by the Tabaq people in South Sudan. While Dr. Lovelang was focused on the phonology of the language, Dr. Hateword was concerned with the way their kinship system works. To further his studies, he chose ten members of a big family and asked them to say a name first and then use the kinship term they'd use to describe that person. He carefully wrote everything down in a notebook and drew the following diagram:\n\n[VISUAL OMITTED: figures/Tabaq.pdf]\n\nShortly after, Dr. Hateword had to give up on his research, leaving all his scribbles as well as the blank diagram to his colleague. At first, Dr. Lovelang had no clue as to how he could fill in the diagram, but, after a closer look at the information in the notebook, he managed to fill it in.\n\nHere are the scribbles from the notebook:\n\n- Rowa: ít̪ɛ̀ t̪ɛ́ɛ̀r\nMinni, t̪ɔ̀ɔ̀d̪ʊ̀\nNadwah, ít̪ɛ̀-n-t̪ɔ̀ɔ̀d̪ʊ̀-t̪ɛ́ɛ̀r\n\n- Kuwa: Salva, ʊ́t̪ɛ́-kɔ̀t̪ʊ̀\nAbdalla, t̪ɔ̀ɔ̀d̪ʊ̀\nRowa, t̪ɔ̀ɔ̀d̪ʊ̀-t̪ɛ́ɛ̀r\n\n- Salva: Abir, wɔ́ɔ́\nMalak, áɲá-n-t̪ɔ̀ɔ̀d̪ʊ̀\nNadwah, ít̪ɛ̀ t̪ɛ́ɛ̀r\n\n- Sihan: Sadiq, ít̪ɛ̀\nMinni, ít̪ɛ̀-n-t̪ɔ̀ɔ̀d̪ʊ̀-kɔ̀t̪ʊ̀\nKuwa, áfá\n\n- Malak: Minni, ít̪ɛ̀\nSihan, màà\nAbdalla, t̪íì\n\n- Sadiq: Salva, t̪ɔ̀ɔ̀d̪ʊ̀\nAbir, màà\nSihan, ít̪ɛ̀ t̪ɛ́ɛ̀r\n\n- Minni: Rowa, màà\nSadiq, t̪íì\nAbir, wɔ́ɔ́\n\n- Nadwah: Kuwa, wɔ́ɔ́\nAbdalla, fáàfá\nRowa, áɲá\n\n- Abir: Malak, ʊ́t̪ɛ́\nNadwah, ʊ́t̪ɛ́\nSadiq, t̪ɔ̀ɔ̀dʊ̀-kɔ̀t̪ʊ̀\n\n- Abdalla: Kuwa, áfá\nMalak, ít̪ɛ̀-n-t̪ɔ̀ɔ̀d̪ʊ̀\nSalva, ít̪ɛ̀-n-t̪ɔ̀ɔ̀dʊ̀-kɔ̀t̪ʊ̀\n\nWhile filling in the diagram, Dr. Lovelang noticed that, in Tabaq, certain kinship terms can be expressed using two different terms, and one of them is derived from the other. Moreover, he noticed somewhere in the notebook the following information: Sadiq is a man. and Rowa has children.","query":"=3pt\n\n- Fill in the diagram above with the names of the ten family members (in the diagram above circles represents women and squares men).\n\n- Write in Tabaq all the possible kinship terms that denote the relation between the following persons:\n\n- Sadiq to Salva\n\n- Abir to Nadwah\n\n- Salva to Sihan\n\n- Nadwah to Minni\n- Malak to Kuwa\n\n- Abdalla to Rowa\n\n- Abdalla to Salva","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"other","eval_type":"single","reasoning_trace":"10.7. Tabaq\n\n- The names in the diagram are, from left to right:\n\n- Top generation: Abir, Kuwa\n\n- Middle generation: Sihan, Rowa, Sadiq, Abdalla\n\n- Bottom generation: Malak, Minni, Nadwah, Salva\n\n-\n\n- áfá\n\n- wɔ́ɔ́\n\n- ít̪ɛ̀-n-t̪ɔ̀ɔ̀d̪ʊ̀ or ít̪ɛ̀-n-t̪ɔ̀ɔ̀d̪ʊ̀-kɔ̀t̪ʊ̀\n\n- tíì-n-t̪ɔ̀ɔ̀d̪ʊ̀ or tíì-n-t̪ɔ̀ɔ̀d̪ʊ̀-t̪ɛ́ɛ̀r\n\n- ʊ́t̪ɛ́ or ʊ́t̪ɛ́-t̪ɛ́ɛ̀r\n\n- ít̪ɛ̀ or ít̪ɛ̀-kɔ̀t̪ʊ̀\n\n- fáàfá\n\nRules:\n\nThe kinship terms are:\n\n- ít̪ɛ̀* = ‘sibling’\n\n- t̪ɔ̀ɔ̀d̪ʊ* = ‘child’\n\n- ʊ́t̪ɛ́* = ‘grandchild’\n\n- ít̪ɛ̀-n-t̪ɔ̀ɔ̀d̪ʊ̀* = ‘nephew, niece’\n\n- wɔ́ɔ́ = ‘grandparent’\n\n- áɲá = ‘father's sister’\n\n- màà = ‘mother, mother's sister’\n\n- fáàfá = ‘father's brother’\n\n- tíì = ‘mother's brother’\n\n- áfá = ‘father’\n\nThe words marked with * can receive the suffixes -t̪ɛ́ɛ̀r and -kɔ̀t̪ʊ̀, in order to mark the gender (feminine and masculine, respectively).\n\nThe structure X-n-t̪ɔ̀ɔ̀d̪ʊ is translated as ‘X's child’ (thus, a nephew/niece is actually translated as the ‘child of the sibling’). The words màà, áɲá, tíì, and fáàfá can be combined with -n-t̪ɔ̀ɔ̀d̪ʊ̀ to express ‘cousin’.","source":"langsci_420","problem_group_id":"langsci420:10.7","chapter":10,"chapter_title":"Other types of problems","section":3,"section_title":"Practice
{"id":"book_10_08","context":"A linguist came to Salu Leang (Sulawesi) to study the Aralle-Tabulahan language. He visited various hamlets of Salu Leang (see the map below) and asked local residents: Umba laungngola? ‘Where are you going?’\n\n[VISUAL OMITTED: images/Aralle_EN.jpg]\n\nBelow are the answers he got. There are gaps in some of them.\n\n- In Kahangang hamlet:\n\n- Lamaoä' bete' di Bulung.\n\n- Lamaoä' sau di Kota.\n\n- Lamaoä' [blank] di Palempang.\n\n- In Kombeng hamlet:\n\n- Lamaoä' pano di Pahihuang.\n\n- Lamaoä' tama di Sohongang.\n\n- Lamaoä' naung di Tamonseng.\n\n- Lamaoä' [blank] di Palempang.\n\n- In Kota hamlet:\n\n- Lamaoä' dai' di Kombeng.\n\n- Lamaoä' dai' di Palempang.\n\n- Lamaoä' naung di Pikung.\n\n- Lamaoä' [blank] di Bulung.\n\n- Lamaoä' [blank] di Sohongang.\n\n- In Palempang hamlet:\n\n- Lamaoä' bete' di Kahangang.\n\n- Lamaoä' dai' di Kombeng.\n\n- Lamaoä' pano di Panampo.\n\n- Lamaoä' sau di Sohongang.\n\n- Lamaoä' [blank] di Bulung.\n\n- Lamaoä' [blank] di Kota.\n\n- Lamaoä' [blank] di Pahihuang.\n\n- In Pahihuang hamlet:\n\n- Lamaoä' naung di Bulung.\n\n- Lamaoä' naung di Pikung.\n\n- In Bulung hamlet:\n\n- Lamaoä' pano di Pahihuang.\n\n- Lamaoä' pano di Panampo.\n\n- Lamaoä' [blank] di Kota.\n\n- Lamaoä' [blank] di Pikung.\n\n- In Panampo hamlet:\n\n- Lamaoä' tama di Kahangang.\n\n- Lamaoä' pano di Tamonseng.\n\n- Lamaoä' [blank] di Kota.\n\n- In Pikung hamlet:\n\n- Lamaoä' pano di Kota.\n\n- Lamaoä' dai' di Pahihuang.\n\n- Lamaoä' sau di Sohongang.\n\n- Lamaoä' [blank] di Bulung.\n\n- Lamaoä' [blank] di Kahangang.\n\n- Lamaoä' [blank] di Panampo.\n\n- In Sohongang hamlet:\n\n- Lamaoä' bete' di Bulung.\n\n- Lamaoä' tama di Kahangang.\n\n- Lamaoä' tama di Kota.\n\n- Lamaoä' dai' di Pahihuang.\n\n- In Tamonseng hamlet:\n\n- Lamaoä' pano di Pahihuang.\n\n- Lamaoä' pano di Panampo\n\n- Lamaoä' [blank] di Kahangang.\n\n- Lamaoä' [blank] di Palempang.","query":"- Fill in the blanks.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"10.8. Aralle-Tabulahan\n\n-\n\n- dai'\n\n- dai'\n\n- bete'\n\n- sau\n\n- naung\n\n- sau\n\n- naung\n\n- tama\n\n- naung\n\n- pano\n\n- bete'\n\n- tama\n\n- dai'\n\n- bete'\n\n- dai'\n\n-\n\nRules:\n\nBasic directions are:\n\n- bete' = ‘across the river’\n\n- tama = ‘upstream’\n\n- sau = ‘downstream’\n\n- pano = ‘on a flat road’\n\n- dai' = ‘upwards’\n\n- naung = ‘downwards’","source":"langsci_420","problem_group_id":"langsci420:10.8","chapter":10,"chapter_title":"Other types of problems","section":3,"section_title":"Practice problems","topic":"orientation, kinship, and other structural problems","language":"Aralle-Tabulahan","author":"Ksenia Gilyarova","competition":"IOL","year":2016,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c10_s01_p01","book_method_c10_s02_p01","book_method_c10_s02_p02","book_method_c10_s02_p03"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":false,"visual_dependency_reasons":["included_image"],"source_file":"chapters/10-OtherProblems.tex","source_line_start":534,"source_line_end":622,"solution_line_start":823,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Other types of problems\nPractice problems\norientation, kinship, and other structural problems\nThis practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nA linguist came to Salu Leang (Sulawesi) to study the Aralle-Tabulahan language. He visited various hamlets of Salu Leang (se
{"id":"book_10_09","context":"Below three Akan men who belong to one family introduce themselves and some members of their family (see the family tree):\n\n- Yɛfrɛ me Enu. Yɛfrɛ me banom Thema ne Yaw ne Ama. Yɛfrɛ me yere Kunto. Yɛfrɛ me nuanom Awotwi ne Nsia. Yɛfrɛ me wɔfaase Berko.\n\n- Yɛfrɛ me Kofi. Yɛfrɛ me nua Esi. Yɛfrɛ me agya Ofori. Yɛfrɛ me sewaanom Dubaku ne Kunto. Yɛfrɛ me sewaabanom Yaw ne Ama ne Kobina.\n\n- Yɛfrɛ me Kobina. Yɛfrɛ me ɛnanom Dubaku ne Kunto. Yɛfrɛ me nuanom Yaw ne Ama. Yɛfrɛ me wɔfa Ofori. Yɛfrɛ me yere Efua.\n\n[VISUAL OMITTED: figures/akan.pdf]\n[VISUAL OMITTED: figures/akan_legend.pdf]","query":"- Supply the family tree with names.\n\n- Here are some more statements by two other men from this family:\n\n- Yɛfrɛ me Yaw. Yɛfrɛ me ɛnanom [blank]. Yɛfrɛ me [blank] Nsia ne [blank]. Yɛfrɛ me nuanom Thema ne [blank]. Yɛfrɛ me [blank] Awotwi. Yɛfrɛ me [blank] Ofori. Yɛfrɛ me [blank] Esi ne [blank]. Yɛfrɛ me [blank] Berko.\n\n- Yɛfrɛ me [blank]. Yɛfrɛ me banom Kofi ne [blank]. Yɛfrɛ me\n[blank] Yaw ne [blank]. Yɛfrɛ me [blank] Kunto ne [blank].\n\n- Fill in the gaps. Some gaps contain more than one word.","answer":[],"explanation":"","work_lang":"eng_Latn","task_lang":[],"task_type":"fill_blanks","eval_type":"single","reasoning_trace":"10.9. Akan\n\n- [VISUAL OMITTED: figures/akan_solutions.pdf]\n\n-\n\n- Dubaku ne Kunto\n\n- agyanom\n\n- Enu\n\n- Ama ne Kobina\n\n- sewaa\n\n- wɔfa\n\n- wɔfabanom\n\n- Kofi\n\n- sewaaba\n\n- Ofori\n\n- Esi\n\n- wɔfaasenom\n\n- Ama ne Kobina\n\n- nuanom\n\n- Dubako\n\nRules:\n\n- Yɛfrɛ me N. = ‘My name is N.’\n\n- Yɛfrɛ me R N. = ‘My R's name is N.’ (R = kinship term)\n\n- X ne Y = ‘X and Y’\n\n- -nom = pl\n\n- Kinship terms:\n\n- ɛna = ‘mother, mother's sister’\n\n- wɔfaase = ‘sister's child’\n\n- sewaa = ‘father's sister’\n\n- sewaaba = ‘father's sister's child’\n\n- nua = ‘sibling, parallel cousin’\n\n- agya = ‘father, father's brother’\n\n- ba = ‘child, brother's child’\n\n- wɔfa = ‘mother's brother’\n\n- wɔfaba = ‘mother's brother's child’\n\n- yere = ‘wife’","source":"langsci_420","problem_group_id":"langsci420:10.9","chapter":10,"chapter_title":"Other types of problems","section":3,"section_title":"Practice problems","topic":"orientation, kinship, and other structural problems","language":"Akan","author":"Ksenia Gilyarova","competition":"IOL","year":2018,"solution_tier":"practice_solution","answer_extraction_status":"not_separated_from_solution","assignment_environment_count":1,"method_ids":["book_method_c10_s01_p01","book_method_c10_s02_p01","book_method_c10_s02_p02","book_method_c10_s02_p03"],"method_link_type":"chapter_candidate_methods","method_description":"This practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.","string_only_usable":false,"visual_dependency_reasons":["included_image"],"source_file":"chapters/10-OtherProblems.tex","source_line_start":624,"source_line_end":646,"solution_line_start":868,"license":"CC-BY-4.0","attribution":"Vlad A. Neacșu, Linguistics Olympiad: Training guide, Language Science Press","retrieval_text":"Other types of problems\nPractice problems\norientation, kinship, and other structural problems\nThis practice problem belongs to the book's Other types of problems chapter and is intended to exercise methods from that chapter. Retrieval should select the most relevant linked method chunk.\nBelow three Akan men who belong to one family introduce themselves and some members of their family (see the family tree):\n\n- Yɛfrɛ me Enu. Yɛfrɛ me banom Thema ne Yaw ne Ama. Yɛfrɛ me yere Kunto. Yɛfrɛ me nuanom Awotwi ne Nsia. Yɛfrɛ me wɔfaase Berko.\n\n- Yɛfrɛ me Kofi. Yɛfrɛ me nua Esi. Yɛfrɛ me agya Ofori. Yɛfrɛ me sewaanom Dubaku ne Kunto. Yɛfrɛ me sewaabanom Yaw ne Ama ne Kobina.\n\n- Yɛfrɛ me Kobina. Yɛfrɛ me ɛnanom Dubaku ne Kunto. Yɛfrɛ me nuanom