Turkish (Türkçe [ˈtyɾctʃɛ], Türk dili [ˈtyɾc dilɪ], also known as Türkiye Türkçesi 'Turkish of Turkey') is the most widely spoken of the Turkic languages, with around 90 million speakers. It is the national language of Turkey and one of two official languages of Cyprus. Significant smaller groups of Turkish speakers also exist in Germany, Austria, Bulgaria, North Macedonia, Greece, other parts of Europe, the South Caucasus, and some parts of Central Asia, Iraq, and Syria. Turkish is the 18th-most spoken language in the world.
To the west, the influence of Ottoman Turkish—the variety of the Turkish language that was used as the administrative and literary language of the Ottoman Empire—spread as the Ottoman Empire expanded. In 1928, as one of Atatürk's reforms in the early years of the Republic of Turkey, the Perso-Arabic script-based Ottoman Turkish alphabet was replaced with the Latin script-based Turkish alphabet.
Turkish grammar is characterized by extensive agglutination and is generally very regular. The basic word order is subject–object–verb. Turkish has no noun classes or grammatical gender. Other notable grammatical features include evidentiality, converbs, and a variety of tenses, aspects, and moods.
Turkish phonology is marked by vowel harmony; native words typically contain only front vowels (⟨e, i, ö, ü⟩) or only back vowels (⟨a, ı, o, u⟩), a phenomenon known as palatal harmony. In addition, certain suffixes exhibit a fourfold harmony that also includes rounding harmony among high vowels. While most suffixes agree with vowel harmony, there are some exceptions. Loanwords are not required to follow this rule.
Turkish has a word accent system that has been described as either stress- or pitch-based. Most native words are accented on the last syllable, though this can vary due to suffixes. Meanwhile, loanwords vary considerably in this regard.
Over 80% of Turkish vocabulary stems from native Turkic, with significant contributions from Ottoman-era Arabic, French, and Persian loanwords, as well as smaller contributions from Italian, Greek, and English. Many loanwords, especially those from Arabic, were replaced by Turkish coinages during the Turkish language reform due to prevailing ideas of linguistic purism.
The language makes usage of honorifics and has a strong T–V distinction which distinguishes varying levels of politeness, social distance, age, courtesy or familiarity toward the addressee. The plural second-person pronoun and verb forms can be used for referring to a single person out of respect.
Contents
Classification
Turkish is a member of the Oghuz group of the Turkic family. Other members include Azerbaijani, spoken in Azerbaijan and north-west Iran, Gagauz of Gagauzia, Qashqai of south Iran, and Turkmen of Turkmenistan.
Historically, the Turkic language family was considered part of the larger Altaic language family, along with the Japanese, Korean, Mongolic, and Tungusic language families. Some linguists have also proposed including other language families.
Altaic theory has fallen out of favour since the 1960s, and a majority of linguists now consider Turkic languages to be unrelated to any other language family, though the Altaic hypothesis still has a small degree of support from individual linguists. The nineteenth-century Ural-Altaic theory, which grouped Turkish with Finnish, Hungarian and Altaic languages, is considered even less plausible in light of Altaic's rejection. The theory was based mostly on the fact these languages share three features: agglutination, vowel harmony and lack of grammatical gender.
History
The earliest known Old Turkic inscriptions are the three monumental Orkhon inscriptions found in modern Mongolia. Erected in honour of the prince Kul Tigin and his brother Emperor Bilge Khagan, these date back to the Second Turkic Khaganate (dated 682–744 CE). After the discovery and excavation of these monuments and associated stone slabs by Russian archaeologists in the wider area surrounding the Orkhon Valley between 1889 and 1893, it became established that the language on the inscriptions was the Old Turkic language written using the Old Turkic alphabet, which has also been referred to as "Turkic runes" or "runiform" due to a superficial similarity to the Germanic runic alphabets.
With the Turkic expansion during Early Middle Ages (c. 6th–11th centuries), peoples speaking Turkic languages spread across Central Asia, covering a vast geographical region stretching from Siberia all the way to Europe and the Mediterranean. The Seljuqs of the Oghuz Turks, in particular, brought their language, Oghuz—the direct ancestor of today's Turkish language—into Anatolia during the 11th century. Also during the 11th century, an early linguist of the Turkic languages, Mahmud al-Kashgari from the Kara-Khanid Khanate, published the first comprehensive Turkic language dictionary and map of the geographical distribution of Turkic speakers in the Dīwān Lughāt al-Turk (ديوان لغات الترك).
Ottoman Turkish
Following the adoption of Islam around the year 950 by the Kara-Khanid Khanate and the Seljuq Turks, who are both regarded as the ethnic and cultural ancestors of the Ottomans, the administrative language of these states acquired a large collection of loanwords from Arabic and Persian. Turkish literature during the Ottoman period, particularly Divan poetry, was heavily influenced by Persian, including the adoption of Persian poetic meters and a great quantity of imported Persian words. The literary and official language during the Ottoman Empire period (c. 1299–1922) is termed Ottoman Turkish, which borrowed heavily from Persian and Arabic that differed considerably from today's modern Turkish, was largely unintelligible to the period's everyday Turkish. The everyday Turkish, known as kaba Türkçe or 'vulgar Turkish', spoken by the less-educated, lower and also rural members of Ottoman society, contained a higher percentage of native vocabulary and served as the basis for the modern Turkish language.
While visiting the region between Adıyaman and Adana, Evliya Çelebi recorded the "Turkman language" and compared it with his own Turkish:
Language reform and modern Turkish
After the foundation of the modern state of Turkey and the script reform, the Turkish Language Association (TDK) was established in 1932 under the patronage of Mustafa Kemal Atatürk, with the aim of conducting research on Turkish. One of the tasks of the newly established association was to initiate a language reform to replace loanwords of Arabic and Persian origins with Turkish equivalents. By banning the usage of imported words in the press, the association succeeded in removing several hundred foreign words from the language. While most of the words introduced to the language by the TDK were newly derived from Turkic roots, it also opted for reviving Old Turkish words which had not been used for centuries. In 1935, the TDK published a bilingual Ottoman-Turkish/Pure Turkish dictionary that documents the results of the language reform.
Owing to this sudden change in the language, older and younger people in Turkey started to differ in their vocabularies. While the generations born before the 1940s tend to use the older terms of Arabic and Persian origins, the younger generations favor new expressions. It is considered particularly ironic that Atatürk himself, in his lengthy speech to the new Parliament in 1927, used the formal style of Ottoman Turkish that had been common at the time amongst statesmen and the educated strata of society in the setting of formal speeches and documents. After the language reform, the Turkish education system discontinued the teaching of literary Ottoman Turkish, and over time the speaking and writing ability of society atrophied to the point that later generations of Turkish speakers would perceive the speech as sounding so alien that it had to be "translated" three times into modern Turkish: first in 1963, again in 1986, and most recently in 1995.
The past few decades have seen the continuing work of the TDK to coin new Turkish words to express new concepts and technologies as they enter the language, mostly from English. Many of these new words, particularly information technology terms, have received widespread acceptance. However, the TDK is occasionally criticized for coining words which sound contrived and artificial. Some earlier changes—such as Turkic bölem to replace the Arabic-derived fırka, ("political party")—also failed to meet with popular approval (the Arabic loanword fırka has been replaced by the French loanword parti).
Some examples of modern Turkish words and the old loanwords are:
Geographic distribution
Turkish is natively spoken by the Turkish people in Turkey and by the Turkish diaspora in some 30 other countries. The Turkish language is mutually intelligible with Azerbaijani. In particular, Turkish-speaking minorities exist in countries that formerly (in whole or part) belonged to the Ottoman Empire, such as Iraq, Bulgaria, Cyprus, Greece (primarily in Western Thrace), the Republic of North Macedonia, Romania, and Serbia. More than two million Turkish speakers live in Germany; and there are significant Turkish-speaking communities in the United States, France, the Netherlands, Austria, Belgium, Switzerland, and the United Kingdom. Due to the cultural assimilation of Turkish immigrants in host countries, not all ethnic members of the diaspora speak the language with native fluency.
In 2005, 93% of the population of Turkey were native speakers of Turkish, about 67 million at the time, with Kurdish languages making up most of the remainder.
Azerbaijani is the official language of Azerbaijan and is mutually intelligible with Turkish. Speakers of the two languages can usually understand each other, particularly in everyday conversations. Turkey and Azerbaijan have very good relations, and many Turkish companies and government agencies invest in Azerbaijan. Consequently, Turkey exerts significant influence over Azerbaijan. However, the growing presence of Turkish in Azerbaijan, coupled with the tendency of many children to use Turkish words instead of Azerbaijani ones due to satellite TV, has raised concerns that the unique characteristics of Azerbaijani may be eroded. Many bookstores sell Turkish books alongside Azerbaijani ones. Agalar Mahmadov, a leading intellectual, has expressed concern that Turkish has "already started to take over the national and natural dialects of Azerbaijan." Nevertheless, Turkish is not as prevalent as Russian as a foreign language.
Official status
Turkish is the official language of Turkey and is one of the official languages of Cyprus. Turkish has official status in 38 municipalities in Kosovo, including Mamusha,, two in the Republic of North Macedonia and two in Iraq. Cyprus has requested the European Union to add Turkish as an official language, as it is one of the two official languages of the country.
In Turkey, the regulatory body for Turkish is the Turkish Language Association (Türk Dil Kurumu or TDK), which was founded in 1932 under the name Türk Dili Tetkik Cemiyeti ("Society for Research on the Turkish Language"). The Turkish Language Association was influenced by the ideology of linguistic purism: indeed one of its primary tasks was the replacement of loanwords and of foreign grammatical constructions with equivalents of Turkish origin. These changes, together with the adoption of the new Turkish alphabet in 1928, shaped the modern Turkish language spoken today. The TDK became an independent body in 1951, with the lifting of the requirement that it should be presided over by the Minister of Education. This status continued until August 1983, when it was again made into a governmental body in the constitution of 1982, following the military coup d'état of 1980.
Dialects
Standard Turkish is based on the dialect of Istanbul. This Istanbul Turkish (İstanbul Türkçesi) constitutes the model of written and spoken Turkish, as recommended by Ziya Gökalp, Ömer Seyfettin and others.
Dialectal variation persists, in spite of the levelling influence of the standard used in mass media and in the Turkish education system since the 1930s. Academic researchers from Turkey often refer to Turkish dialects as ağız or şive, leading to an ambiguity with the linguistic concept of accent, which is also covered with these words. Several universities, as well as a dedicated work-group of the Turkish Language Association, carry out projects investigating Turkish dialects. As of 2002 work continued on the compilation and publication of their research as a comprehensive dialect-atlas of the Turkish language. Although the Ottoman alphabet, being more phonetically ambiguous than the Latin script, encoded for many of the dialectal variations between Turkish dialects, the modern Latin script fails to do this. Examples of this are the presence of the nasal velar sound [ŋ] in certain eastern dialects of Turkish which was represented by the Ottoman letter ⟨ڭ⟩ but that was merged into ⟨n⟩ in the Latin script. Additionally are letters such as ⟨خ ,ق ,غ⟩ which make the sounds [ɣ], [q], and [x], respectively in certain eastern dialects but that are merged into [g], [k], and [h] in western dialects and are therefore defectively represented in the Latin alphabet for speakers of eastern dialects.
Some immigrants to Turkey from Rumelia speak Rumelian Turkish, which includes the distinct dialects of Ludogorie, Dinler, and Adakale, which show the influence of the theorized Balkan sprachbund. Kıbrıs Türkçesi is the name for Cypriot Turkish and is spoken by the Turkish Cypriots. Edirne is the dialect of Edirne. Ege is spoken in the Aegean region, with its usage extending to Antalya. The nomadic Yörüks of the Mediterranean Region of Turkey also have their own dialect of Turkish. This group is not to be confused with the Yuruk nomads of Macedonia, Greece, and European Turkey, who speak Balkan Gagauz Turkish.
The Meskhetian Turks who live in Kazakhstan, Azerbaijan and Russia as well as in several Central Asian countries, also speak an Eastern Anatolian dialect of Turkish, originating in the areas of Kars, Ardahan, Artvin, Diyarbakır and Erzurum and sharing similarities with Azerbaijani, the language of Azerbaijan.
Phonology
Consonants
The phoneme that is usually referred to as yumuşak g ("soft g"), written ⟨ğ⟩ in Turkish orthography, represents a vowel sequence or a rather weak bilabial approximant between rounded vowels, a weak palatal approximant between unrounded front vowels, and a vowel sequence elsewhere. It never occurs at the beginning of a word or a syllable, but always follows a vowel. When word-final or preceding another consonant, it lengthens the preceding vowel.
In native Turkic words, the sounds [c], [ɟ], and [l] are mainly in complementary distribution with [k], [ɡ], and [ɫ]; the former set occurs adjacent to front vowels and the latter adjacent to back vowels. The distribution of these phonemes is often unpredictable, however, in foreign borrowings and proper nouns. In such words, [c], [ɟ], and [l] often occur with back vowels: some examples are given below. However, there are minimal pairs that distinguish between these sounds, such as kar /kaɾ/ "snow" vs kâr /caɾ/ "profit".
Turkish orthography reflects final-obstruent devoicing, a form of consonant mutation whereby the voiced obstruents /b d d͡ʒ ɡ/ are devoiced to /p t t͡ʃ k/ at the end of a word. At least one source claims Turkish consonants are laryngeally-specified three-way fortis-lenis (aspirated/neutral/voiced) like Armenian, though they only appear this way at the end of a syllable or at the start of certain suffixes, where they can interact with other morphemes. Some words end in an underlying voiced consonant that remains voiced at the end of the word. Similarly, suffixes beginning with a voiced consonant become voiceless only after a word ending in an underlying "aspirated" consonant. Other words, such as kanat ("wing") end in an underlying "neutral" consonant that becomes devoiced only at the end of a word. Certain words and suffixes such as sanat ("art") and the relative suffix -ki contain an underlying "aspirated" stop that does not voice between vowels.
Native nouns of two or more syllables that end in /k/ in dictionary form are nearly all /g/ in underlying form. However, most verbs and monosyllabic nouns are underlyingly /k/.
Vowels
The vowels of the Turkish language are, in their alphabetical order, ⟨a⟩, ⟨e⟩, ⟨ı⟩, ⟨i⟩, ⟨o⟩, ⟨ö⟩, ⟨u⟩, ⟨ü⟩. The Turkish vowel system can be considered as being three-dimensional, where vowels are characterised by how and where they are articulated focusing on three key features: front and back, rounded and unrounded and vowel height. Vowels are classified [±back], [±round] and [±high].
The only vowels in hiatus in the language are found in loanwords and may be categorised as falling diphthongs usually analyzed as a sequence of /j/ and a vowel.
The principle of vowel harmony, which permeates Turkish word-formation and suffixation, is due to the natural human tendency towards economy of muscular effort. This principle is expressed in Turkish through three rules:
If the first vowel of a word is a back vowel, any subsequent vowel is also a back vowel; if the first is a front vowel, any subsequent vowel is also a front vowel.
If the first vowel is unrounded, so too are subsequent vowels.
If the first vowel is rounded, subsequent vowels are either rounded and close or unrounded and open.
The second and third rules minimize muscular effort during speech. More specifically, they are related to the phenomenon of labial assimilation: if the lips are rounded (a process that requires muscular effort) for the first vowel they may stay rounded for subsequent vowels. If they are unrounded for the first vowel, the speaker does not make the additional muscular effort to round them subsequently.
Grammatical affixes have "a chameleon-like quality", and obey one of the following patterns of vowel harmony:
twofold (-e/-a): In his more recent works Lewis prefers to omit the superscripts, on the grounds that "there is no need for this once the principle has been grasped" (Lewis [2001]). The locative case suffix, for example, is -de after front vowels and -da after back vowels. The notation -de² is a convenient shorthand for this pattern.
Word-accent
With the exceptions stated below, Turkish words are oxytone (accented on the last syllable).
Exceptions to word-accent rules
Place-names are not oxytone: Anadolú (Anatolia), İstánbul. Most place names are accented on their first syllable as in Páris. This holds true when place names are spelled the same way as common nouns, which are oxytone: mısír (maize), Mísır (Egypt), sirkecı̇́ (vinegar-seller), Sı̇́rkeci (district in Istanbul), bebék (doll, baby), Bébek (district in Istanbul), ordú (army), Órdu (a Turkish city on the Black Sea).
Foreign nouns usually retain their original accentuation, e.g., lokánta (< Italian locanda "restaurant"), gazéte (< Italian gazzetta "newspaper")
Some words about family members and living creatures have irregular accentuation: ánne (mother), görúmce (husband's sister), çekı̇́rge (grasshopper), karínca (ant), kokárca (skunk)
Adverbs are usually accented on the first syllable, e.g., şı̇́mdi (now), sónra (after), ánsızın (suddenly), gérçekten (really), (but gerçektén (from reality)), kíşın (during winter)
Compound words are accented on the end of the first element, e.g., çırílçıplak (stark naked), bakán (minister), báşbakan (prime minister)
Diminutives constructed by suffix –cik are accented on the first syllable, e.g., úfacık (very tiny)
Words with enclitic suffixes, –le (meaning "with"), –ken (meaning "while"), –ce (creating an adverb), –leyin (meaning "in" or "during"), –me (negating the verbal stem), –yor (denoting the present tense)
Enclitic words, which shift the accentuation to the previous syllable, e.g., ol- (meaning to be), mi (denoting a question), gibi (meaning similar to), için (for), ki (that), de (too)
Syntax
Sentence groups
Turkish has two groups of sentences: verbal and nominal sentences. In the case of a verbal sentence, the predicate is a finite verb, while the predicate in nominal sentence will have either no overt verb or a verb in the form of the copula ol or y (variants of "be"). Examples of both are given below:
The two groups of sentences have different ways of forming negation. A nominal sentence can be negated with the addition of the word değil. For example, the sentence above would become Necla öğretmen değil ('Necla is not a teacher'). However, the verbal sentence requires the addition of a negative suffix -me to the verb (the suffix comes after the stem but before the tense): Necla okula gitmedi ('Necla did not go to school').
In the case of a verbal sentence, an interrogative clitic mi is added after the verb and stands alone, for example Necla okula gitti mi? ('Did Necla go to school?'). In the case of a nominal sentence, then mi comes after the predicate but before the personal ending, so for example Necla, siz öğretmen misiniz? ('Necla, are you [formal, plural] a teacher?').
Word order
Word order in simple Turkish sentences is generally subject–object–verb, as in Korean and Latin, but unlike English, for verbal sentences and subject-predicate for nominal sentences. However, as Turkish possesses a case-marking system, and most grammatical relations are shown using morphological markers, often the SOV structure has diminished relevance and may vary. The SOV structure may thus be considered a "pragmatic word order" of language, one that does not rely on word order for grammatical purposes.
Consider the following simple sentence which demonstrates that the focus in Turkish is on the element that immediately precedes the verb:
The postpredicate position signifies what is referred to as background information in Turkish—information that is assumed to be known to both the speaker and the listener, or information that is included in the context. Consider the following examples:
There has been some debate among linguists whether Turkish is a subject-prominent (like English) or topic-prominent (like Japanese and Korean) language, with recent scholarship implying that it is indeed both subject and topic-prominent. This has direct implications for word order as it is possible for the subject to be included in the verb-phrase in Turkish. There can be S/O inversion in sentences where the topic is of greater importance than the subject.
Grammar
Turkish is an agglutinative language and frequently uses affixes, and specifically suffixes, or endings. One word can have many affixes and these can also be used to create new words, such as creating a verb from a noun, or a noun from a verbal root (see the section on Word formation). Most affixes indicate the grammatical function of the word.
The only native prefixes are alliterative intensifying syllables used with adjectives or adverbs: for example sımsıcak ("boiling hot" < sıcak) and masmavi ("bright blue" < mavi).
The extensive use of affixes can give rise to long words, e.g. Çekoslovakyalılaştıramadıklarımızdanmışsınızcasına, meaning "In the manner of you being one of those that we apparently couldn't manage to convert to Czechoslovak". While this case is contrived, long words frequently occur in normal Turkish, as in this heading of a newspaper obituary column: Bayramlaşamadıklarımız (Bayram [festival]-Recipr-Impot-Partic-Plur-PossPl1; "Those of our number with whom we cannot exchange the season's greetings"). Another example can be seen in the final word of this heading of the online Turkish Spelling Guide (İmlâ Kılavuzu): Dilde birlik, ulusal birliğin vazgeçilemezlerindendir ("Unity in language is among the indispensables [dispense-Pass-Impot-Plur-PossS3-Abl-Copula] of national unity ~ Linguistic unity is a sine qua non of national unity").
Nouns
Turkish does not have grammatical gender and the sex of a person does not affect the forms of words. The third-person pronoun o may refer to "he", "she" or "it." Despite this lack, Turkish still has ways of indicating gender in nouns:
Most domestic animals have male and female forms, e.g., aygır ("stallion"), kısrak ("mare"), boğa ("bull"), inek ("cow").
For other animals, the sex may be indicated by adding the word erkek ("male") or dişi ("female") before the corresponding noun, e.g., dişi kedi ("female cat").
For people, the female sex may be indicated by adding the word kız ("girl") or kadın ("woman"), e.g., kadın kahraman ("heroine") instead of kahraman ("hero").
Some foreign words of French or Arabic origin already have separate female forms, e.g., aktris ("actress"), kâtibe ("female clerk").
The Serbo-Croat feminine suffix –ica is used in three borrowings: kraliçe (queen), imparatoriçe ("empress") and çariçe ("tsarina"). This suffix was also used in the neologism tanrıça ("goddess") (< tanrı ("god")).
There is no definite article in Turkish, but definiteness of the object is implied when the accusative ending is used (see below). Turkish nouns decline by taking case endings. There are six noun cases in Turkish, with all the endings following vowel harmony (shown in the table using the shorthand superscript notation). Since the postposition ile often gets suffixed onto the noun, some analyze it as an instrumental case, although in formal speech it takes the genitive with personal pronouns, singular demonstratives, and interrogative kim. The plural marker -ler ² immediately follows the noun before any case or other affixes (e.g. köylerin "of the villages").
The accusative case marker is used only for definite objects; compare (bir) ağaç gördük "we saw a tree" with ağacı gördük "we saw the tree". The plural marker -ler ² is generally not used when a class or category is meant: ağaç gördük can equally well mean "we saw trees [as we walked through the forest]"—as opposed to ağaçları gördük "we saw the trees [in question]".
Personal pronouns
The Turkish personal pronouns in the nominative case are ben (1s), sen (2s), o (3s), biz (1pl), siz (2pl, or 2h), and onlar (3pl). They are declined regularly with some exceptions: benim (1s gen.); bizim (1pl gen.); bana (1s dat.); sana (2s dat.); and the oblique forms of o use the root on. As mentioned before, all demonstrative singular and personal pronouns take the genitive when ile is affixed onto it: benimle (1s ins.), bizimle (1pl ins.); but onunla (3s ins.), onlarla (3pl ins.). All other pronouns (reflexive kendi and so on) are declined regularly.
Two nouns, or groups of nouns, may be joined in either of two ways:
definite (possessive) compound (belirtili tamlama). E.g. Türkiye'nin sesi "the voice of Turkey (radio station)": the voice belonging to Turkey. Here the relationship is shown by the genitive ending -in4 added to the first noun; the second noun has the third-person suffix of possession -(s)i4.
indefinite (qualifying) compound (belirtisiz tamlama). E.g. Türkiye Cumhuriyeti "Turkey-Republic = the Republic of Turkey": not the republic belonging to Turkey, but the Republic that is Turkey. Here the first noun has no ending; but the second noun has the ending (s)i4—the same as in definite compounds.
The following table illustrates these principles. In some cases, the constituents of the compounds are themselves compounds; for clarity these subsidiary compounds are marked with [square brackets]. The suffixes involved in the linking are underlined. If the second noun group already had a possessive suffix (because it is a compound by itself), no further suffix is added.
As the last example shows, the qualifying expression may be a substantival sentence rather than a noun or noun group.
There is a third way of linking the nouns where both nouns take no suffixes (takısız tamlama). However, in this case the first noun acts as an adjective, e.g. Demir kapı (iron gate), elma yanak ("apple cheek", i.e. red cheek), kömür göz ("coal eye", i.e. black eye) :
Adjectives
Adjectives are hard to distinguish from nouns, with a great majority of them being able to take the same morphology that nouns do. For example, büyük ("big, old") can become büyüklerim ("my elders"). The only large class of exceptions to this are those formed with the suffixes -si, -(i)msi, -(i)mtrak, the Arabic-derived nisba suffix -î, the Persian-derived Persian -ane and -varî, and certain more recent borrowings such as demokratik ("democratic") and kültürel ("cultural").
Adjectives always precede their nouns, with two exceptions:
The words kare ("square") and küp ("cube") follow unit names such as in bir metre küp or bir metreküp ("one cubic meter")
merhum ("the late") may come after the name of the deceased, mirroring Arabic usage.
Comparison of adjectives is achieved through putting the noun being compared against in the ablative case: ağır ("heavy"), kurşundan ağır ("heavier than lead"). Less … than is translated by putting
az ("little") between the noun and adjective: kurşundan az ağır. daha ("more") may be inserted for emphasis: kurşundan daha ağır, kurşundan daha az ağır. However, daha is necessary when there is no noun being compared against: bu çekiç daha ucuz, öteki daha sağlam ("this hammer is cheaper, that one is stronger").
The superlative is expressed through the adverb en ("most"): en az verimli toprak ("the least fertile soil").
Certain adjectives can form reduplicated intensive forms, where the first syllable is prefixed to the word, with the letters ⟨m, p, n, s⟩ replacing its coda if it exists. The prefix also takes the word's accent.
açık ("open") → apaçık ("explicit, obvious, absolutely clear")
başka ("different") → bambaşka ("totally different")
Verbs
The copula *imek, equivalent to English to be, is typically omitted in informal speech in sentences of the form "A = B" (where A is in the third person), though it is used for this purpose in formal speech or writing. In informal speech, the copula is instead used to mark supposition, emphasis, surety, or confidence.
Other uses for the copula in informal speech include when
the predicate is a relative clause
the subject is a pronoun understood from the context
the subject is a noun which follows the predicate,
the subject is a phrase containing a postposition, and the predicate is introduced by ki ("that").
Bundan dolayıdır ki gitmedim. ("It is because of this that I did not go.")
Only in this last situation is the copula strictly necessary.
The copula is negated by suffixing it to the end of the adverb değil ("not").
Turkish verbs indicate person. They can be made negative, potential ("can"), or non-potential ("cannot"). Furthermore, Turkish verbs show tense (present, past, future, and aorist), mood (conditional, imperative, inferential, necessitative, and optative), and aspect. The inferential suffix -miş4 is also glossed as a direct evidential or a mirative. Negation is expressed by the suffix -me²- immediately following the stem.
Verb tenses
(For the sake of simplicity the term "tense" is used here throughout, although for some forms "aspect" or "mood" might be more appropriate.) There are nine simple and 20 compound tenses in Turkish. The nine simple tenses are: simple past (di'li geçmiş), inferential past (miş'li geçmiş), present continuous, simple present (aorist), future, optative, subjunctive, necessitative ("must") and imperative. There are three groups of compound forms. "Story" (hikaye) is the witnessed past of the above forms (except command), referral (rivayet) is the unwitnessed past of the above forms (except simple past and command), conditional (koşul) is the conditional form of the first five basic tenses. In the example below, the second person singular of the verb gitmek ("go"), stem gid-/git-, is shown.
There are also so-called combined verbs, which are created by suffixing certain verb stems (like bil or ver) to the original stem of a verb. Bil is the suffix for the sufficiency mood. It is the equivalent of the English auxiliary verbs "able to", "can" or "may". Ver is the suffix for the swiftness mood, kal for the perpetuity mood and yaz for the approach ("almost") mood. Thus, while gittin means "you went", gidebildin means "you could go" and gidiverdin means "you went swiftly". The tenses of the combined verbs are formed the same way as for simple verbs.
Turkish verbs have attributive forms, including present, similar to the English present participle (with the ending -en2); future (-ecek2); indirect/inferential past (-miş4); and aorist (-er2 or -ir4).
The most important function of some of these attributive verbs is to form modifying phrases equivalent to the relative clauses found in most European languages. The subject of the verb in an -en2 form is (possibly implicitly) in the third person (he/she/it/they); this form, when used in a modifying phrase, does not change according to number. The other attributive forms used in these constructions are the future (-ecek2) and an older form (-dik4), which covers both present and past meanings. These two forms take "personal endings," which have the same form as the possessive suffixes but indicate the person and possibly number of the subject of the attributive verb; for example, yediğim means "what I eat," yediğin means "what you eat," and so on. The use of these "personal or relative participles" is illustrated in the following table, in which the examples are presented according to the grammatical case which would be seen in the equivalent English relative clause.
Vocabulary
The latest 2011 edition of Güncel Türkçe Sözlük (Current Turkish Dictionary), the official dictionary of the Turkish language published by Turkish Language Association, contains 117,000 words organized into 93,000 entries.
Word origins
Around 86% of the Turkish vocabulary is of Turkic origin. The majority of the core vocabulary and the most commonly used words in Turkish, including those first acquired by children as they learn to speak, derive from Turkic. Nevertheless, Turkish vocabulary contains a significant number of loanwords from other languages, in which around 14% of Turkish words are of foreign origin. According to the Turkish Language Association, 6,463 of these foreign words come from Arabic, 4,974 from French, 1,374 from Persian, 632 from Italian, 538 from English, 399 from Greek, and 147 from Latin.
In Turkish, there are many pairs of synonyms where one word is of foreign origin and the other of Turkic origin. These pairs are the result of the enrichment of the Turkish vocabulary with loanwords from Arabic, Persian and French, and of the Turkish language reform initiated in the early 20th century that aimed to restore foreign-origin words with Turkic equivalents.
Word formation
Turkish extensively uses agglutination to form new words from nouns and verbal stems. The majority of Turkish words originate from the application of derivative suffixes to a relatively small set of core vocabulary.
Turkish obeys certain principles when it comes to suffixation. Most suffixes in Turkish will have more than one form, depending on the vowels and consonants in the root- vowel harmony rules will apply; consonant-initial suffixes will follow the voiced/ voiceless character of the consonant in the final unit of the root; and in the case of vowel-initial suffixes an additional consonant may be inserted if the root ends in a vowel, or the suffix may lose its initial vowel. There is also a prescribed order of affixation of suffixes- as a rule of thumb, derivative suffixes precede inflectional suffixes which are followed by clitics, as can be seen in the example set of words derived from a substantive root below:
Another example, starting from a verbal root:
New words are also frequently formed by compounding two existing words into a new one, as in German. Compounds can be of two types- bare and (s)I. The bare compounds, both nouns and adjectives are effectively two words juxtaposed without the addition of suffixes for example the word for girlfriend kızarkadaş (kız+arkadaş) or black pepper karabiber (kara+biber). A few examples of compound words are given below:
However, the majority of compound words in Turkish are (s)I compounds, which means that the second word will be marked by the 3rd person possessive suffix. A few such examples are given in the table below (note vowel harmony):
Idiomatic language
Turkish has a wide variety of idioms derived from body parts. Unlike Western metaphors that connect the heart with love, Turkish speakers more often conceptualize the heart (yürek (the more commonly used one of the two in idioms) or kalp) as a container, bearer, or experiencer of negative emotions such as sadness, pity, distress, or fear. The eye is similarly used in a great number of idioms. Common metaphors for the eye include using it to represent a compartment or division of a physical object, a hole or gap, an object of love, a person or experiencer, perception, hunger, or a mental state.
Writing system
Turkish is written using a version of Latin script introduced in 1928 by Atatürk to replace the Ottoman Turkish alphabet, a version of Perso-Arabic script. The Ottoman alphabet marked only three different vowels—long ā, ū and ī—and included several redundant consonants, such as variants of z (which were distinguished in Arabic but not in Turkish). The omission of short vowels in the Arabic script was claimed to make it particularly unsuitable for Turkish, which has eight vowels.
The reform of the script was an important step in the cultural reforms of the period. The task of preparing the new alphabet and selecting the necessary modifications for sounds specific to Turkish was entrusted to a Language Commission composed of prominent linguists, academics, and writers. The introduction of the new Turkish alphabet was supported by public education centers opened throughout the country, cooperation with publishing companies, and encouragement by Atatürk himself, who toured the country teaching the new letters to the public. As a result, there was a dramatic increase in literacy from its original, pre-modern levels.
The Latin alphabet was applied to the Turkish language for educational purposes even before the 20th-century reform. Instances include a 1635 Latin-Albanian dictionary by Frang Bardhi, who also incorporated several sayings in the Turkish language, as an appendix to his work (e.g. alma agatsdan irak duschamas—"An apple does not fall far from its tree").
Turkish now has an alphabet suited to the sounds of the language: the spelling is largely phonemic, with one letter corresponding to each phoneme. Most of the letters are used approximately as in English, the main exceptions being ⟨c⟩, which denotes [dʒ] (⟨j⟩ being used for the [ʒ] found in Persian and European loans); and the undotted ⟨ı⟩, representing [ɯ]. As in German, ⟨ö⟩ and ⟨ü⟩ represent [ø] and [y]. The letter ⟨ğ⟩, in principle, denotes [ɣ] but has the property of lengthening the preceding vowel and assimilating any subsequent vowel. The letters ⟨ş⟩ and ⟨ç⟩ represent [ʃ] and [tʃ], respectively. A circumflex is written over back vowels following ⟨k⟩ and ⟨g⟩ when these consonants represent [c] and [ɟ]—almost exclusively in Arabic and Persian loans.
The Turkish alphabet consists of 29 letters (q, w, x omitted and ç, ş, ğ, ı, ö, ü added); the complete list is:
Sample texts
Dostlar Beni Hatırlasın
Dostlar Beni Hatırlasın is a Turkish folk poem by the world-renowned poet and ashik Âşık Veysel Şatıroğlu (1894–1973).
İnsan Hakları Evrensel Bildirisi
Article 1 of the Universal Declaration of Human Rights in Turkish:
Bütün insanlar hür, haysiyet ve haklar bakımından eşit doğarlar. Akıl ve vicdana sahiptirler ve birbirlerine karşı kardeşlik zihniyeti ile hareket etmelidirler.
Article 1 of the Universal Declaration of Human Rights in English:
All human beings are born free and equal in dignity and rights. They are endowed with reason and conscience and should act towards one another in a spirit of brotherhood.
International Phonetic Alphabet transcription:
[byˈt̪ʰyn̪ in̪s̪än̪ˈɫ̪äɾ̞̊ hyɾ̞̊ häjs̪iˈje̞t̪ ve̞ häk̠ˈɫ̪äɾ‿bäk̠ʰɯmɯn̪ˈd̪än̪ e̞ˈʃit̪ d̪o̞.äˈɫ̪äɾ̞̊ ‖ äˈk̠ʰɯɫ̪ ve̞ vid͡ʒd̪äˈn̪ä sä(h)ipt̪ʰiɾˈl̠ɛɾ̞̊ ve̞ biɾbiɾl̠e̞ɾiˈn̪e̞ k̠ʰäɾˈʃɯ k̠ʰäɾd̪e̞ʃˈl̠ik̟ z̪ihn̪ije̞ˈt̪ʰi‿iˈl̠e̞ häɾe̞ˈk̟ʰe̞t̪ e̞t̪me̞l̠id̪iɾˈl̠ɛɾ̞̊ ‖]
Turkish computer keyboard
Turkish language uses two standardised keyboard layouts, known as Turkish Q (QWERTY) and Turkish F, with Turkish Q being the most common.




