Models of natural languages and language characteristics are widely used in many computer science applications such as data security, language identification, spell checking, data compression, authorship attribution and speech recognition. In the scope of this study, a large scale corpus is created and used to discover language characteristics of Turkish. Word and letter based analyses are made on this corpus to build a base for several NLP studies. In the author identification part, we used two different methods based on word n- grams to identify author of an anonymous text. For 16 authors, training and test set articles are collected, and mentioned two methods are applied on these article sets. Finally, obtained results from two methods are compared with each other and most successful method is determined. This study can help professionals working on author identification, corpus linguistics, n-gram analysis, cryptanalysis, and speech recognition.
Les informations fournies dans la section « Synopsis » peuvent faire référence à une autre édition de ce titre.
Models of natural languages and language characteristics are widely used in many computer science applications such as data security, language identification, spell checking, data compression, authorship attribution and speech recognition. In the scope of this study, a large scale corpus is created and used to discover language characteristics of Turkish. Word and letter based analyses are made on this corpus to build a base for several NLP studies. In the author identification part, we used two different methods based on word n- grams to identify author of an anonymous text. For 16 authors, training and test set articles are collected, and mentioned two methods are applied on these article sets. Finally, obtained results from two methods are compared with each other and most successful method is determined. This study can help professionals working on author identification, corpus linguistics, n-gram analysis, cryptanalysis, and speech recognition.
Feri?tah Örücü: She had received the B.S. and M.S. degrees in Comp Eng from DEU, Turkey. She has been a Ph.D. student and a Res Asst of Dept of Comp Eng of DEU. Gökhan Dalk?l?ç: He had received M.S. degrees in Comp Sci from USC, and from Ege Univ CI, Ph.D. degree in Comp Eng from DEU. He has been an Asst Prof of the Dept of Comp Eng of DEU.
Les informations fournies dans la section « A propos du livre » peuvent faire référence à une autre édition de ce titre.
Vendeur : BuchWeltWeit Ludwig Meier e.K., Bergisch Gladbach, Allemagne
Taschenbuch. Etat : Neu. This item is printed on demand - it takes 3-4 days longer - Neuware -Models of natural languages and language characteristics are widely used in many computer science applications such as data security, language identification, spell checking, data compression, authorship attribution and speech recognition. In the scope of this study, a large scale corpus is created and used to discover language characteristics of Turkish. Word and letter based analyses are made on this corpus to build a base for several NLP studies. In the author identification part, we used two different methods based on word n- grams to identify author of an anonymous text. For 16 authors, training and test set articles are collected, and mentioned two methods are applied on these article sets. Finally, obtained results from two methods are compared with each other and most successful method is determined. This study can help professionals working on author identification, corpus linguistics, n-gram analysis, cryptanalysis, and speech recognition. 80 pp. Englisch. N° de réf. du vendeur 9783838385075
Quantité disponible : 2 disponible(s)
Vendeur : moluna, Greven, Allemagne
Etat : New. N° de réf. du vendeur 5418759
Quantité disponible : Plus de 20 disponibles
Vendeur : buchversandmimpf2000, Emtmannsberg, BAYE, Allemagne
Taschenbuch. Etat : Neu. This item is printed on demand - Print on Demand Titel. Neuware -Models of natural languages and language characteristics are widely used in many computer science applications such as data security, language identification, spell checking, data compression, authorship attribution and speech recognition. In the scope of this study, a large scale corpus is created and used to discover language characteristics of Turkish. Word and letter based analyses are made on this corpus to build a base for several NLP studies. In the author identification part, we used two different methods based on word n- grams to identify author of an anonymous text. For 16 authors, training and test set articles are collected, and mentioned two methods are applied on these article sets. Finally, obtained results from two methods are compared with each other and most successful method is determined. This study can help professionals working on author identification, corpus linguistics, n-gram analysis, cryptanalysis, and speech recognition.VDM Verlag, Dudweiler Landstraße 99, 66123 Saarbrücken 80 pp. Englisch. N° de réf. du vendeur 9783838385075
Quantité disponible : 1 disponible(s)
Vendeur : AHA-BUCH GmbH, Einbeck, Allemagne
Taschenbuch. Etat : Neu. nach der Bestellung gedruckt Neuware - Printed after ordering - Models of natural languages and language characteristics are widely used in many computer science applications such as data security, language identification, spell checking, data compression, authorship attribution and speech recognition. In the scope of this study, a large scale corpus is created and used to discover language characteristics of Turkish. Word and letter based analyses are made on this corpus to build a base for several NLP studies. In the author identification part, we used two different methods based on word n- grams to identify author of an anonymous text. For 16 authors, training and test set articles are collected, and mentioned two methods are applied on these article sets. Finally, obtained results from two methods are compared with each other and most successful method is determined. This study can help professionals working on author identification, corpus linguistics, n-gram analysis, cryptanalysis, and speech recognition. N° de réf. du vendeur 9783838385075
Quantité disponible : 1 disponible(s)
Vendeur : preigu, Osnabrück, Allemagne
Taschenbuch. Etat : Neu. Characteristics of Contemporary Printed Turkish | Letter and Word Characteristics of Contemporary Printed Turkish and Author Identification | Fer¿¿Tah Örücü (u. a.) | Taschenbuch | 80 S. | Englisch | 2010 | LAP LAMBERT Academic Publishing | EAN 9783838385075 | Verantwortliche Person für die EU: preigu GmbH & Co. KG, Lengericher Landstr. 19, 49078 Osnabrück, mail[at]preigu[dot]de | Anbieter: preigu. N° de réf. du vendeur 107485005
Quantité disponible : 5 disponible(s)
Vendeur : Mispah books, Redhill, SURRE, Royaume-Uni
Paperback. Etat : Like New. LIKE NEW. SHIPS FROM MULTIPLE LOCATIONS. book. N° de réf. du vendeur ERICA79038383850716
Quantité disponible : 1 disponible(s)