To see the other types of publications on this topic, follow the link: Documents linguistics.

Journal articles on the topic 'Documents linguistics'

Create a spot-on reference in APA, MLA, Chicago, Harvard, and other styles

Select a source type:

Consult the top 50 journal articles for your research on the topic 'Documents linguistics.'

Next to every source in the list of references, there is an 'Add to bibliography' button. Press on it, and we will generate automatically the bibliographic reference to the chosen work in the citation style you need: APA, MLA, Harvard, Chicago, Vancouver, etc.

You can also download the full text of the academic publication as pdf and read online its abstract whenever available in the metadata.

Browse journal articles on a wide variety of disciplines and organise your bibliography correctly.

1

Khachatryan, Robert. "On Questioned Document Examination in Forensic Linguistics." Armenian Folia Anglistika 11, no. 1 (13) (2015): 76–83. http://dx.doi.org/10.46991/afa/2015.11.1.076.

Full text
Abstract:
The overarching objective of this article is to contemplate on the significance of questioned document examination in forensic linguistics. Questioned document examination (QDE) is a forensic linguistics discipline pertaining to disputed documents and applying a variety of linguistic methods and tools to answer questions about a disputed document. More specifically, this article elaborates on the categorization of legal documents that are instrumental in the process of establishing the authenticity of documents in dispute.
APA, Harvard, Vancouver, ISO, and other styles
2

Wang, Jian Xiong, Yuan Yong Feng, and Ke Qi. "A Survey on Computational Linguistics in Design Documents." Applied Mechanics and Materials 58-60 (June 2011): 1630–35. http://dx.doi.org/10.4028/www.scientific.net/amm.58-60.1630.

Full text
Abstract:
Design is one of the main activities in industrial manufacture. Researchers in the design field have an increasing number of opportunities to analyse design documents. Some researchers have sought to explore the natural language in these documents, or the design documents. This paper briefly reviews previous research in design document. By describing and analyzing the existing methods, it identifies the gap for the computational linguistics in design documents.
APA, Harvard, Vancouver, ISO, and other styles
3

Eick, Stephen G., Justin Mauger, and Alan Ratner. "A Visualization Testbed for Analyzing the Performance of Computational Linguistics Algorithms." Information Visualization 6, no. 1 (2007): 64–74. http://dx.doi.org/10.1057/palgrave.ivs.9500141.

Full text
Abstract:
We have built an AJAX-enabled browser-based testbed for evaluating the performance of computational linguistics algorithms. Our testbed consists of a visualization system and analysis portal. Our focus is on algorithms that classify and cluster documents by assigning weights to words and scoring each document against high-dimensional reference concept vectors. The testbed visualization and algorithm analysis techniques include Confusion Matrices, ROC Curves, Document Visualizations showing word importance, and Interactive Reports. A unique aspect of our testbed is document visualizations built
APA, Harvard, Vancouver, ISO, and other styles
4

Belous, Elena. "Interactive Documents: Language Features and Document Status." Vestnik Volgogradskogo gosudarstvennogo universiteta. Serija 2. Jazykoznanije, no. 1 (April 2021): 168–80. http://dx.doi.org/10.15688/jvolsu2.2021.1.14.

Full text
Abstract:
The research is carried out in line with the current problems of modern document linguistics, related to the study of formation peculiarities, design and functioning of new types of documents. It is shown that currently there is a change in the structure of the document in two directions: the unification of the document form and the creation of documents without a clear structure. The concept of "interactive document" is introduced. It refers to a form of hypertext representation, a special material structure (code, program, existing in an electronic environment) created by a person to store a
APA, Harvard, Vancouver, ISO, and other styles
5

Akhmedshaeva, Mavlyuda. ""LEGAL LINGUISTICS AND LINGUISTIC EXPERTISE OF DRAFT NORMATIVE LEGAL DOCUMENTS: SOME THEORETICAL AND LEGAL ISSUES "." Jurisprudence 5, no. 1 (2025): 8–16. https://doi.org/10.51788/tsul.jurisprudence.5.1./lzfk9429.

Full text
Abstract:
"The article discusses the role and significance of language issues in ensuring the quality and effectiveness of legislative documents, as well as the importance of legal linguistics in ensuring that the language of normative legal documents complies with the rules and requirements of the state language. This field reflects the interconnection between law and language, and the article explores the development prospects of this discipline in our country’s scientific field. At the same time, the article analyzes the varying perspectives of scholars on the necessity of ensuring that the language
APA, Harvard, Vancouver, ISO, and other styles
6

Alivernini, Sergio. "“kišib-gu10 zi-ra-ab”: Annul my Sealed Tablet!" Altorientalische Forschungen 50, no. 2 (2023): 141–49. http://dx.doi.org/10.1515/aofo-2023-0011.

Full text
Abstract:
Abstract The practice of annulling or destroying written documents is well documented during the III Dynasty of Ur (2112–2004 BC), where several documents record the expression “kišib PN zi-re-dam,” the sealed document is to be annulled/destroyed. This practice is recorded in three different types of administrative documents: loan texts, “orders” requesting the annulment of another document, and documents whose annulment takes place only after another document has arrived. The aim of this article is to study the documents that record this practice and to provide a description of the administra
APA, Harvard, Vancouver, ISO, and other styles
7

Fitria, Tira Nur. "FORENSIC LINGUISTICS: CONTRIBUTION OF LINGUISTICS IN LEGAL CONTEXT." PRASASTI: Journal of Linguistics 9, no. 1 (2024): 117. http://dx.doi.org/10.20961/prasasti.v9i1.71527.

Full text
Abstract:
This research describes the contribution of linguistics in forensic linguistics, especially in a legal context. This research is library research. The analysis shows that forensic linguistics applies language analysis and linguistic theories in linguistic events involved in the legal process, including products, interactions in the judicial process, and interactions between individuals that result in certain legal impacts. Forensic linguistic analysis involves linguistic fields, including phonetics, semantics, discourse and pragmatics, stylistics, morphological, syntactical, and sociological.
APA, Harvard, Vancouver, ISO, and other styles
8

Petyko, Marton, Lucia Busso, Tim Grant, and Sarah Atkins. "The Aston Forensic Linguistic Databank (FoLD)." Language and Law=Linguagem e Direito 9, no. 1 (2022): 9–24. http://dx.doi.org/10.21747/21833745/lanlaw/9_1a1.

Full text
Abstract:
The Aston Forensic Linguistic Databank (FoLD) is a permanent,controlled access online repository for forensic linguistic data. We broadlyunderstand forensic linguistics as any academic research with a potential toimprove the delivery of justice through the analysis of language. FoLD thuscomprises a wide range of datasets with relevance to forensic linguistics andlanguage and law, including commercial extortion letters, investigative interviewsin police and other contexts, legal documents, forum posts from far-right onlinegroups, and comment threads from political blogs. This paper outlines how
APA, Harvard, Vancouver, ISO, and other styles
9

Al-Arab, Zeinab E., Ahmed M. Gadallah, and Hesham M. Hefny. "An Enhanced Fuzzy Information Retrieval Model Based on Linguistics." Applied Mechanics and Materials 519-520 (February 2014): 853–56. http://dx.doi.org/10.4028/www.scientific.net/amm.519-520.853.

Full text
Abstract:
The paper proposes a linguistic based fuzzy ontology information retrieval model. The model deals with linguistic based queries in multi domains. Such linguistics are user defined, reflecting his subjective view. The model also proposes a ranking algorithm that ranks the set of relevant documents according to some criteria such as their relevance degree, confidence degree, and updating degree.
APA, Harvard, Vancouver, ISO, and other styles
10

Garabík, Radovan. "Corpus of Slovak Legislative Documents." Journal of Linguistics/Jazykovedný casopis 73, no. 2 (2022): 175–89. http://dx.doi.org/10.2478/jazcas-2023-0004.

Full text
Abstract:
Abstract The article describes the construction of the corpus of Slovak legislative documents. By analyzing several statistical values of the source metadata and documents, we efficiently improve corpus quality. We describe the methods used to clean up small variations in metadata, length based discrimination of document and examine the effectiveness of several strategies of deduplication. The corpus is a part of a comparable corpus of legislative documents of seven languages, created in the Multilingual Resources for CEF.AT in the Legal Domain (MARCELL) project.
APA, Harvard, Vancouver, ISO, and other styles
11

Nasharuddin, Nurul Amelina, Muhamad Taufik Abdullah, Azreen Azman, Rabiah Abdul Kadir, and Enrique Herrera-Viedma. "Feature-based Similarity Method for Aligning the Malay and English News Document." INTERNATIONAL JOURNAL OF COMPUTERS & TECHNOLOGY 11, no. 4 (2013): 2410–21. http://dx.doi.org/10.24297/ijct.v11i4.3125.

Full text
Abstract:
Corpus-based translation approach can be used to obtain reliable translation knowledge in addition to the use of dictionaries or machine translation. But the availability of such corpus is very limited especially for the low-resources languages. Many works have been reported for the alignments of multilingual documents especially among the European languages, but less focusing on the languages with less linguistics resources. One of the challenges is to align the available multilingual documents for the creation of comparable corpus for these kinds of languages. This article describes an align
APA, Harvard, Vancouver, ISO, and other styles
12

Gorban, Oksana A., and Elena M. Sheptukhina. "Genre determination of grammatical expression of textual chronotope in regional documents of the 18th century." International Journal “Speech Genres” 18, no. 4 (40) (2023): 330–36. http://dx.doi.org/10.18500/2311-0740-2023-18-4-40-330-336.

Full text
Abstract:
The article reveals several current issues of documental linguistics and genre theory – grammatical means of expressing/actualizing textual chronotope in the 18th cent. official documents from the Archive Fund of the Mikhailovsky Stanichny ataman of the State Archive of the Volgograd Region, which are described with the procedure of genre related analysis. By comparing organization and grammatical representation of text chronotope in administrative, informative and registration documents, there was revealed a tendency for typological genre implications defined by the document properties and th
APA, Harvard, Vancouver, ISO, and other styles
13

Kosova, Marina. "Development of Theoretical Approaches to Modern Document Studies in Linguistic School of Professor Sofiya P. Lopushanskaya." Vestnik Volgogradskogo gosudarstvennogo universiteta. Serija 2. Jazykoznanije, no. 6 (February 2024): 66–79. http://dx.doi.org/10.15688/jvolsu2.2023.6.5.

Full text
Abstract:
The article presents fundamental positions of the linguistic school by Sofiya P. Lopushanskaya, Professor of Volgograd State University, and describes their theoretical advancement as being applied in the discourse studies on "Document text: history and modern state". It explains that the study of business and administrative communication should be based on a clarified definition of the term document as its main constituent. The major goal of documental linguistics is to process a diverse linguistic analysis, discover objective knowledge and reason linguistic comprehension of the documental na
APA, Harvard, Vancouver, ISO, and other styles
14

Rodríguez Cortez, Josafat Jonathan. "Unión de palabras en documentos novohispanos del siglo XVI: ¿influencias orales o tradiciones escriturales? = Union of words in novo-hispanic documents of 16th century: Oral influences or writing traditions?" Estudios Humanísticos. Filología, no. 42 (December 18, 2020): 107–30. http://dx.doi.org/10.18002/ehf.v0i42.6024.

Full text
Abstract:
Mediante los resultados del análisis de 20 documentos del siglo XVI, este artículo muestra que la unión y separación de palabras en la escritura no sólo se debe a los tipos de letra utilizados en su época, sino a marcadas influencias lingüísticas, principalmente fonológicas. Proponemos aquí que en la escritura del español novohispano convergieron por igual motivaciones lingüísticas como extralingüísticas. Through the results of the analysis of 20 documents of 16th Century, this article shows that union and separation of words is not only due to the kind of writing used in that time, but due to
APA, Harvard, Vancouver, ISO, and other styles
15

Kowalski, Paweł. "Multilingual Dictionary of Keywords as a Tool for the Digital Bibliographic Database of World Slavic Linguistics." Rasprave Instituta za hrvatski jezik i jezikoslovlje 46, no. 2 (2020): 783–95. http://dx.doi.org/10.31724/rihjj.46.2.18.

Full text
Abstract:
The paper presents the structure of a multilingual dictionary of keywords, which is an integral part of the bibliographic database of Slavic linguistics iSybislaw representing the digital information retrieval system (www.isybislaw.ispan.waw.pl). The lexical units (keywords) of the language of keywords used in the system are represented primarily by linguistic terms. In spite of a different denotation – the keywords directly denote sets of documents, and indirectly the non-documentary reality, while the terms denote elements of linguistic reality – they are formally equal with linguistic terms
APA, Harvard, Vancouver, ISO, and other styles
16

Newmeyer, Frederick J. "American Linguists Look at Swiss Linguistics, 1925–1940." Historiographia Linguistica 42, no. 1 (2015): 107–18. http://dx.doi.org/10.1075/hl.42.1.06new.

Full text
Abstract:
Summary Swiss linguistic research did not have a major impact on American linguistics in the inter-war period. Nevertheless, there was a perhaps surprising awareness of the results of Swiss scholars among American linguists active in that period. This paper documents both their numerous references to the work of the linguists of the Geneva School as well as the recognition given to Swiss scholars in general by the Linguistic Society of America.
APA, Harvard, Vancouver, ISO, and other styles
17

Am, St Asriati, Nurlaili Nurlaili, Ridwan Ridwan, Ega Safira, and St Asmayanti AM. "An Overview of Linguistics and Education: A Bibliometric Analysis." IDEAS: Journal on English Language Teaching and Learning, Linguistics and Literature 12, no. 2 (2024): 1763–79. https://doi.org/10.24256/ideas.v12i2.5115.

Full text
Abstract:
This study attempts to identify the main linguistic education topic that is an overview on linguistics and education: A bibliometric analysis from 2015 to 2023. It employed a quantitative method with a bibliometric literature review. The 309 documents (48 papers, 218 articles, and 43 scientific papers) maintained in the Mendeley applications formed the source of the research data. This study took the data from Google Scholar to obtain some researches related to the topic by using Public or Perish application. VOS viewer software was applied for the bibliometric analysis of the data. Numerous p
APA, Harvard, Vancouver, ISO, and other styles
18

Mustafa-Elhadi, Widad. "La contribution de la terminologie à la conception théorique des langages documentaires et à l’indexation de documents." Meta 37, no. 3 (2002): 465–73. http://dx.doi.org/10.7202/002699ar.

Full text
Abstract:
Abstract This paper attempts to show the contribution of certain linguistic phenomena, and reciprocally, the contribution of linguistics to the analysis of semantic relationships used by the documentalists in conceiving thesauri. Abstract representation of extra-linguistic reality involves understanding of certain designation processes which are useful for linguistic research. Taking into account these elements would help achieving a better collaboration between linguists, documentalists and information systems designers.
APA, Harvard, Vancouver, ISO, and other styles
19

Adil Elshiekh Abdalla. "Forensic Linguistics and its Role in Crime Investigation: Descriptive Study." JALL | Journal of Arabic Linguistics and Literature 2, no. 2 (2022): 55–75. http://dx.doi.org/10.59202/jall.v2i2.343.

Full text
Abstract:
This desrcrptive study aims at discussing the discipline of forensic linguistics in terms of its definition, history , scope, and its role in detection of the crime. The study revealed that this decipine is considr ed one of the most contemporary scientific fields that deal with the analysis of linguistic evidence with aim to clarify any ambiguity exist in any judicial process, especially in the investigation of crimes and legal cases. The forensic linguistics also involves in probing the crucial legal documents and other linguistic evidence, such as handwritten texts prior to suicide attempt,
APA, Harvard, Vancouver, ISO, and other styles
20

VLASIUK, Liudmyla, and Olga DEMYDENKO. "LINGUISTIC INDEXATION AS WAY OF MEDIA TEXT CLUSTERIZATION." Folia Philologica, no. 3 (2022): 37–41. http://dx.doi.org/10.17721/folia.philologica/2022/3/5.

Full text
Abstract:
Linguistic indexation is a highly complex phenomenon due to collecting, sorting and storing data aimed at providing high-speed, high-quality and accurate search of information. That is why it has quickly turned into one of the core problems for researchers, especially when it comes to examining it in the media text, which complicates this process for scholars engaged in linguistic studies. Profound research of the linguistic text’s indexation is predetrmined by the necessity to structure of the information system. The stated phenomenon has gained high relevance in linguistics over the recent y
APA, Harvard, Vancouver, ISO, and other styles
21

Xigui, QIU. "NOTES ET DOCUMENTS." Cahiers de Linguistique Asie Orientale 22, no. 1 (1993): 107–17. http://dx.doi.org/10.1163/19606028-90000365.

Full text
APA, Harvard, Vancouver, ISO, and other styles
22

HUYNH, Sabine, and Sabine HUYNH. "Notes et documents." Cahiers de Linguistique Asie Orientale 37, no. 2 (2008): 223–40. http://dx.doi.org/10.1163/1960602808x00082.

Full text
Abstract:
Le contact entre les communautés vietnamienne et française durant l'occupation française au Viêt Nam a entraîné l'adaptation phonologique d'un nombre important d'emprunts au français. En vietnamien, ces emprunts se sont vu attribuer des tons. La littérature scientifique sur cette question (ainsi que sur celle du contact linguistique français-vietnamien) est jusqu'ici limitée et se fonde sur un nombre restreint d'emprunts. L'analyse phonologique et statistique de 600 mots vietnamiens d'origine française nous permet d'étudier les mécanismes d'attribution de tons à des syllabes qui en étaient dép
APA, Harvard, Vancouver, ISO, and other styles
23

GALAMBOS, Imre, and Imre GALAMBOS. "Notes et Documents." Cahiers de Linguistique Asie Orientale 40, no. 1 (2011): 73–108. http://dx.doi.org/10.1163/1960602811x00088.

Full text
Abstract:
From the 11th century onwards, the Xixia state grew into a major power along the northwestern frontier of Song China. This article examines the Tangut perspective of their geo-political environment as it is reflected in a translation of a Chinese military treatise called Jiangyuan, a work attributed to Zhuge Liang. The concluding part of the original text presents the traditional Sino-centric worldview with the four barbarian tribes (Yi, Man, Rong and Di) around the empire. The Tangut translation, however, omits three of the four tribes and discusses only the Northern Di, thus adapting the Chi
APA, Harvard, Vancouver, ISO, and other styles
24

ZHU, Xiaonong, and Xiaonong ZHU. "Notes et Documents." Cahiers de Linguistique Asie Orientale 41, no. 1 (2012): 81–106. http://dx.doi.org/10.1163/1960602812x00032.

Full text
Abstract:
This article reports on two previously unknown tonal types discovered in recent fieldwork in China: the "double circumflex tone" and the "back dipping tone". The double circumflex tone has a "fall-rise-fall" pitch curve with two turning points in both perceptual and acoustic terms. The back dipping tone has a higher tonal head and a later turning point than an ordinary dipping tone. This paper lists the 32 localities where DCTs are found. It further discusses major characteristics of the two tones and their phonological contrast with the more commonly known dipping tone.
APA, Harvard, Vancouver, ISO, and other styles
25

Tsypina, Alla V. "Methods of Legal Linguistics and their Use for Solving Communicative Problems in the Conduct of Certain investigative Actions." Gaps in Russian Legislation 18, no. 2 (2025): 90–96. https://doi.org/10.33693/2541-8025-2025-18-2-90-96.

Full text
Abstract:
The purpose of the study is to substantiate the need to use in the production of investigative actions based on the communicative aspect, the synergy of language and law and the resolution of criminal procedural issues related to the use of certain linguistic methods. The article is devoted to improving the activities of law enforcement agencies through the integration of knowledge from various scientific fields, particularly linguistics and law. The article explores the techniques of forensic linguistics used in the preliminary investigation phase, focusing on the examination of language patt
APA, Harvard, Vancouver, ISO, and other styles
26

Dedinkin, A. L. "Legal Discourse as a Multi-Dimensional Integrated Phenomenon and Legal Linguistics as a Syncretic Science." Bulletin of Kemerovo State University 23, no. 1 (2021): 220–28. http://dx.doi.org/10.21603/2078-8975-2021-23-1-220-228.

Full text
Abstract:
The article introduces legal discourse as part of a complex communicative activity. It is an integrative interdisciplinary phenomenon on the border of jurisprudence and linguistics. The research objective was to establish the constituent parts of legal discourse, which includes legal texts, related scientific literature, and other documents. Legal linguistics is a generalizing discipline that studies the interaction of language and law. The line between legal discourse and other discourses is hard to define. Legal discourse is characterized by unified subjects, procedures, circumstances, and i
APA, Harvard, Vancouver, ISO, and other styles
27

Keller-Cohen, Deborah. "Literate practices in a modern credit union." Language in Society 16, no. 1 (1987): 7–23. http://dx.doi.org/10.1017/s0047404500012100.

Full text
Abstract:
ABSTRACTModern bureaucratic institutions are notorious for producing documents that are difficult to understand. Much attention has been paid to the language of these materials; little is known about the contexts in which these documents are used and their potential effects on functional literacy. Drawing on research in a midwestern credit union, this paper discusses several factors that seem to characterize how and why credit unions and their members use credit union documents: the characteristics of document availability, the structure of interactions in which documents are used, attitudes a
APA, Harvard, Vancouver, ISO, and other styles
28

Rajabi, Taghi, Elham Alayiaboozar, and Beheshti Moluksadat Hosseini. "Conceptual Map of Linguistic Terminology." Journal of Language and Linguistics 1, no. 31 (2021): 117–35. https://doi.org/10.5281/zenodo.14035731.

Full text
Abstract:
A bilingual Persian-English terminology for linguistics has not yet been compiled. A linguistic terminology plays a crucial role in organizing and indexing information in scientific documents, standardizing vocabulary, and facilitating the search and retrieval of information in linguistic databases and related linguistic research. The most significant challenge in producing and compiling a bilingual Persian-English linguistic terminology is designing a conceptual map of the field of linguistics. Finding suitable equivalents in Persian for English terms, as well as the existence of various pers
APA, Harvard, Vancouver, ISO, and other styles
29

Vukčević, Miodrag M. "Turns of the centuries. The Transkribus automated tool for recognition, transcription and translation of handwritten historical documents." Babel. Revue internationale de la traduction / International Journal of Translation 66, no. 2 (2020): 294–310. http://dx.doi.org/10.1075/babel.00159.vuk.

Full text
Abstract:
Abstract The translation of handwritten historical documents faces many challenges due to variation in the writing style, local language, and an inevitable language change. Even the transliteration from Cyrillic to Latin characters is standardized by the bijective transliteration standard ISO 9. This presentation introduces a number of tools offered by Transkribus for the automated processing of documents, such as Handwritten Text Recognition (HTR) and Document Understanding, which are needed for the translation of historical documents. Next to the problem of decoding handwritten documents, wr
APA, Harvard, Vancouver, ISO, and other styles
30

Bowman, A. "Imaging incised documents." Literary and Linguistic Computing 12, no. 3 (1997): 169–76. http://dx.doi.org/10.1093/llc/12.3.169.

Full text
APA, Harvard, Vancouver, ISO, and other styles
31

Murugova, Elen, Galina Matveeva, and George Myasischev. "International communication of the Ancient Russian state (pragmalinguistic aspect)." SHS Web of Conferences 69 (2019): 00074. http://dx.doi.org/10.1051/shsconf/20196900074.

Full text
Abstract:
The article deals with studying business communication between the City of Novgorod and the Hanseatic League. The “Methods and study” section provides a detailed theoretical description of the linguistic personality of a diplomat of the 13-15th centuries. The main methods and directions of study and reconstruction of the personality under study are substantiated. The “Results and Discussion” section describes an experiment conducted on the basis of pragmatic linguistics and its results. Using some documents of the 13-15th centuries as the background, we have reconstructed the speech image of a
APA, Harvard, Vancouver, ISO, and other styles
32

Filimonov, Daniil, Andrey Svetlov, Oksana Gorban, and Marina Kosova. "Automation of Archival Documents Meta Tagging." Mathematical Physics and Computer Simulation, no. 4 (February 2021): 56–68. http://dx.doi.org/10.15688/mpcm.jvolsu.2020.4.6.

Full text
Abstract:
The main goal of this project is to create a corpus of documents from the «Mikhailovsky stanichny ataman» archival fund. The methods of corpus linguistics seem to be the most optimal in this case, since they involve the processing of a large number of texts in order to solve a wide variety of linguistic problems. Our group joined the team of philologists to provide the technical and software part of the project. The main task for us is to create a document corpus engine, that is, software that solves the tasks of storing a database of marked-up texts, executing queries to this database, and al
APA, Harvard, Vancouver, ISO, and other styles
33

Wu, Chuan, Evangelos Kanoulas, and Maarten de Rijke. "It all starts with entities: A Salient entity topic model." Natural Language Engineering 26, no. 5 (2019): 531–49. http://dx.doi.org/10.1017/s1351324919000585.

Full text
Abstract:
AbstractEntities play an essential role in understanding textual documents, regardless of whether the documents are short, such as tweets, or long, such as news articles. In short textual documents, all entities mentioned are usually considered equally important because of the limited amount of information. In long textual documents, however, not all entities are equally important: some are salient and others are not. Traditional entity topic models (ETMs) focus on ways to incorporate entity information into topic models to better explain the generative process of documents. However, entities
APA, Harvard, Vancouver, ISO, and other styles
34

Ebeling, Signe O., and Alois Heuboeck. "Encoding document information in a corpus of student writing: the British Academic Written English corpus." Corpora 2, no. 2 (2007): 241–56. http://dx.doi.org/10.3366/cor.2007.2.2.241.

Full text
Abstract:
The information contained in a document is only partly represented by the wording of the text; in addition, features of formatting and layout can be combined to lend specific functionality to chunks of text (e.g., section headings, highlighting, enumeration through list formatting, etc.). Such functional features, although based on the ‘objective’ typographical surface of the document, are often inconsistently realised and encoded only implicitly, i.e., they depend on deciphering by a competent reader. They are characteristic of documents produced with standard text-processing tools. We discus
APA, Harvard, Vancouver, ISO, and other styles
35

MOTA, PEDRO, MAXINE ESKENAZI, and LUÍSA COHEUR. "MUSED: A multimedia multi-document dataset for topic segmentation." Natural Language Engineering 24, no. 6 (2018): 921–46. http://dx.doi.org/10.1017/s1351324918000359.

Full text
Abstract:
AbstractResearch on topic segmentation has recently focused on segmenting documents by taking advantage of documents covering the same topics. In order to properly evaluate such approaches, a dataset of related documents is needed. However, existing datasets are limited in the number of related documents per domain. In addition, most of the available datasets do not consider documents from different media sources (PowerPoints, videos, etc.), which pose specific challenges to segmentation. We fill this gap with the MUltimedia SEgmentation Dataset (MUSED), a collection of documents manually segm
APA, Harvard, Vancouver, ISO, and other styles
36

RAHIMI, RAZIEH, AZADEH SHAKERY, JAVID DADASHKARIMI, MOZHDEH ARIANNEZHAD, MOSTAFA DEHGHANI, and HOSSEIN NASR ESFAHANI. "Building a multi-domain comparable corpus using a learning to rank method." Natural Language Engineering 22, no. 4 (2016): 627–53. http://dx.doi.org/10.1017/s1351324916000164.

Full text
Abstract:
AbstractComparable corpora are key translation resources for both languages and domains with limited linguistic resources. The existing approaches for building comparable corpora are mostly based on ranking candidate documents in the target language for each source document using a cross-lingual retrieval model. These approaches also exploit other evidence of document similarity, such as proper names and publication dates, to build more reliable alignments. However, the importance of each evidence in the scores of candidate target documents is determined heuristically. In this paper, we employ
APA, Harvard, Vancouver, ISO, and other styles
37

ASTRI, Zul, Nurdin NONI, Abd HALIM, and Fhadli NOER. "A Comprehensive Guide to Corpus Linguistics: A Book Review of Corpora in Applied Linguistics." Research and Innovation in Applied Linguistics-Electronic Journal 2, no. 2 (2024): 174. http://dx.doi.org/10.31963/rial.v2i2.4861.

Full text
Abstract:
Corpora have emerged as a transformative tool in the field of linguistics, providing researchers with a robust, data-driven approach to understanding language use. By compiling large collections of texts-ranging from written documents to transcriptions of spoken language-corpus linguistics enables the analysis of authentic language in its natural context. Corpora in applied linguistics have revolutionized the study of language by providing a data-driven approach to understanding linguistic phenomena. By analyzing large collections of naturally occurring texts, researchers can uncover patterns,
APA, Harvard, Vancouver, ISO, and other styles
38

Bhayro, Siam. "On Performatives in Aramaic Documents." Aramaic Studies 11, no. 1 (2013): 47–52. http://dx.doi.org/10.1163/17455227-13110101.

Full text
Abstract:
This article argues for the presence of performative sentences in early Aramaic documents, giving two examples—one from the Wadi Murabbaʿat divorce document and the other from an Elephantine letter. It is suggested that the use of a compound, consisting of a preposition and a demonstrative pronoun, to reinforce the performative function of the participle, parallels the use of “hereby” in English.
APA, Harvard, Vancouver, ISO, and other styles
39

Brun, Caroline, and Frédérique Segond. "Semantic Encoding of Electronic Documents." International Journal of Corpus Linguistics 6, no. 1 (2001): 79–96. http://dx.doi.org/10.1075/ijcl.6.1.04bru.

Full text
Abstract:
This paper presents an unsupervised, all-words, word sense disambiguation system for English. The system associates a word with its meaning in a given context using an electronic dictionary as a tagged corpora in order to extract semantic disambiguation rules. The methodology attempts to avoid the data acquisition bottleneck observed in word sense disambiguation techniques. Semantic rules are used as input of a semantic application program encoding a linguistic strategy in order to select the best rule to apply. The semantic rule extraction process as well as the application program is describ
APA, Harvard, Vancouver, ISO, and other styles
40

Fuertes Olivera, Pedro A., Silvia Montero Martínez, and Mercedes Garcia de Quesada. "Modelos culturales y discursivos en la traducción de textos de comercio internacional." Babel. Revue internationale de la traduction / International Journal of Translation 51, no. 4 (2005): 357–79. http://dx.doi.org/10.1075/babel.51.4.06fue.

Full text
Abstract:
Resumen La traducción de las Convenciones y las Leyes Modelo, documentos prototípicos emitidos por las Naciones Unidas en relación con la legislación reguladora del Comercio Internacional, constituye un caso de documentos iguales. Éstos muestran una gran dependencia entre el texto en lengua origen y el texto en lengua meta. Este artículo analiza algunos ejemplos de estos documentos en inglés y en español con el objetivo de proponer soluciones a algunos problemas de traducción relacionados con la semiótica de la cultura. En concreto, nuestro objetivo es, por un lado, analizar los modelos discur
APA, Harvard, Vancouver, ISO, and other styles
41

ZHANG, WEN, TAKETOSHI YOSHIDA, and XIJIN TANG. "DISTRIBUTION OF MULTI-WORDS IN CHINESE AND ENGLISH DOCUMENTS." International Journal of Information Technology & Decision Making 08, no. 02 (2009): 249–65. http://dx.doi.org/10.1142/s0219622009003399.

Full text
Abstract:
As a hybrid of N-gram in natural language processing and collocation in statistical linguistics, multi-word is becoming a hot topic in area of text mining and information retrieval. In this paper, a study concerning distribution of multi-words is carried out to explore a theoretical basis for probabilistic term-weighting scheme. Specifically, the Poisson distribution, zero-inflated binomial distribution, and G-distribution are comparatively studied on a task of predicting probabilities of multi-words' occurrences using these distributions, for both technical multi-words and nontechnical multi-
APA, Harvard, Vancouver, ISO, and other styles
42

Sietsema, Brian M., and G. H. R. Horsley. "New Documents Illustrating Early Christianity, Vol. 5: Linguistic Essays." Language 67, no. 4 (1991): 866. http://dx.doi.org/10.2307/415095.

Full text
APA, Harvard, Vancouver, ISO, and other styles
43

NASERASADI, ALI, HAMID KHOSRAVI, and FARAMARZ SADEGHI. "Extractive multi-document summarization based on textual entailment and sentence compression via knapsack problem." Natural Language Engineering 25, no. 1 (2018): 121–46. http://dx.doi.org/10.1017/s1351324918000414.

Full text
Abstract:
AbstractBy increasing the amount of data in computer networks, searching and finding suitable information will be harder for users. One of the most widespread forms of information on such networks are textual documents. So exploring these documents to get information about their content is difficult and sometimes impossible. Multi-document text summarization systems are an aid to producing a summary with a fixed and predefined length, while covering the maximum content of the input documents. This paper presents a novel method for multi-document extractive summarization based on textual entail
APA, Harvard, Vancouver, ISO, and other styles
44

Guenée, Bernard. "Documents insérés et documents abrégés dans la Chronique du religieux de Saint- Denis." Bibliothèque de l'école des chartes 152, no. 2 (1994): 375–428. http://dx.doi.org/10.3406/bec.1994.450736.

Full text
APA, Harvard, Vancouver, ISO, and other styles
45

Bakhteev, Dmitry V., and Alexey V. Antropov. "Forensic examination of anonymous handwritten documents in order to obtain data on the identity of their author and performer." Nowa Kodyfikacja Prawa Karnego 49 (April 18, 2019): 9–22. http://dx.doi.org/10.19195/2084-5065.49.2.

Full text
Abstract:
The anonymous document as an object of forensic research has an information field that is subject to study by three groups of scientific disciplines: handwriting, technical re­search of documents and forensic linguistics. Stability and reproducibility characteristics of handwriting and written speech allows us to identify the author and performer of an anonymous message, as demonstrated by the example of the authorsʼ practice.
APA, Harvard, Vancouver, ISO, and other styles
46

Ringlstetter, Christoph, Klaus U. Schulz, and Stoyan Mihov. "Orthographic Errors in Web Pages: Toward Cleaner Web Corpora." Computational Linguistics 32, no. 3 (2006): 295–340. http://dx.doi.org/10.1162/coli.2006.32.3.295.

Full text
Abstract:
Since the Web by far represents the largest public repository of natural language texts, recent experiments, methods, and tools in the area of corpus linguistics often use the Web as a corpus. For applications where high accuracy is crucial, the problem has to be faced that a non-negligible number of orthographic and grammatical errors occur in Web documents. In this article we investigate the distribution of orthographic errors of various types in Web pages. As a by-product, methods are developed for efficiently detecting erroneous pages and for marking orthographic errors in acceptable Web d
APA, Harvard, Vancouver, ISO, and other styles
47

C Ravago, Joan, Daisy O Casipit, Allan Jay A Esteban, et al. "Situating Philippine English in University Internationalization Efforts." International Research Journal of Multidisciplinary Scope 06, no. 03 (2025): 569–79. https://doi.org/10.47857/irjms.2025.v06i03.05085.

Full text
Abstract:
Abstract English is central to internationalization in higher education institutions, yet the role of its nativized varieties, such as the Philippine English (PhE), remains under examined. Employing theoretical perspectives such as variationist linguistics and corpus linguistics, this study examines the features and extent of PhE use in a Philippine university’s internationalization efforts by analysing institutional documents, including 50 invitations, 7 programs, 6 memoranda, and 26 certificates. Findings reveal lexical innovations such as Philippine as an adjective, indigenous borrowings, a
APA, Harvard, Vancouver, ISO, and other styles
48

Mandzhikova, Larisa B. "Документы Комиссии калмыцких дел как источник по изучению организации делопроизводства и документооборота в Калмыкии в XIX в. (на примере рассмотрения прошения калмыцкого владельца Э. Ц. Кичикова о дозволении носить бронзовую медаль)". Oriental Studies 15, № 5 (2022): 1126–35. http://dx.doi.org/10.22162/2619-0990-2022-63-5-1126-1135.

Full text
Abstract:
Introduction. The issues pertaining to records keeping and management in 19th-century Russia’s government agencies have been well studied. However, there are no fundamental scientific works covering the phenomena in Kalmykia. The paper analyzes activities by the Astrakhan Kalmyk Affairs Commission for the actual document flow procedures and types of documents to have resulted from the Commission’s work. The review procedure of a petition filed by landlord E. Ts. Kichikov and dealing with the bronze medal awarded to the Russian nobility and merchants in memory of the Patriotic War of 1812 — and
APA, Harvard, Vancouver, ISO, and other styles
49

MONZ, CHRISTOF. "Machine learning for query formulation in question answering." Natural Language Engineering 17, no. 4 (2011): 425–54. http://dx.doi.org/10.1017/s1351324910000276.

Full text
Abstract:
AbstractResearch on question answering dates back to the 1960s but has more recently been revisited as part of TREC's evaluation campaigns, where question answering is addressed as a subarea of information retrieval that focuses on specific answers to a user's information need. Whereas document retrieval systems aim to return the documents that are most relevant to a user's query, question answering systems aim to return actual answers to a users question. Despite this difference, question answering systems rely on information retrieval components to identify documents that contain an answer t
APA, Harvard, Vancouver, ISO, and other styles
50

Burg, J. F. M., and R. P. van de Riet. "Enhancing CASE Environments by Using Linguistics." International Journal of Software Engineering and Knowledge Engineering 08, no. 04 (1998): 435–48. http://dx.doi.org/10.1142/s0218194098000248.

Full text
Abstract:
In this paper it is argued that CASE environments could and should be enhanced considerably by using theories and knowledge from linguistics. The environments should 'know' about the language of their users and the domains they are used for. By basing the modeling techniques supported by the CASE tool on linguistic theories and by incorporating Natural Language parsing and generating tools, the CASE environment is able to handle the users' language in an accurate way. More specifically, the CASE environment deals with the meaning of words, instead of the meaningless strings themselves. These m
APA, Harvard, Vancouver, ISO, and other styles
We offer discounts on all premium plans for authors whose works are included in thematic literature selections. Contact us to get a unique promo code!