چکیده:
Preparing a dictionary, especially comprehensive dictionaries that are more detailed and richer in terms of references, is a very difficult and time-consuming manual task. Furthermore, due to the presence of multiple lexicographers at different times with various abilities and literary tastes, the possibility of inconsistency and errors in understanding and adhering to lexicographical formats and styles is very high. Providing software that takes over part of the work and, while assisting definers, limits the occurrence of the aforementioned errors, can significantly contribute to easier, more accurate, and faster dictionary preparation. In this article, we introduce the FarhangYar software, which has been developed as a web-based tool for producing a comprehensive Persian language dictionary at the Academy of Persian Language and Literature. This software helps lexicographers focus on preparing content, regardless of peripheral issues related to formatting and visual appearance. By following lexicographical styles, FarhangYar prevents many lexicographer errors and provides complex features such as automatic reference establishment; by connecting to a linguistic corpus, it creates a suitable environment for searching this corpus and transferring lemmas and examples into the dictionary. Maintaining and transferring the creation and editing history of an entry among definers, editors, and managers is another important feature of this software. Additionally, by defining different access levels for users, the security of the produced content is ensured. Another essential function of using FarhangYar is the creation of a computational corpus of the Persian language, which can be used as an independent resource in other Persian language processing research. In this article, while examining the problems of traditional manual lexicography, we will describe the features and capabilities of FarhangYar and address issues such as the interaction between lexicography and corpus linguistics, collaborative lexicography, simultaneous access to data in a web environment, solving reference problems, managing entries and examples, user management, management of comments and histories, and other prominent features of this software.
خلاصه ماشینی:
The workflow in preparing the corpus is briefly as follows: the experts of the lexicography group of the Academy of Persian Language and Literature, by carefully studying each source, proceed to collect and register phrases as evidence that possess suitable entries for lexicography.
Lexicography Status Lexicography in the group, prior to the use of Farhangyar, was such that each lexicographer was given a set of related entries that had been selected during the corpus preparation process.
Then, with the help of the group's software at the time, the lexicographer would proceed to extract evidences (with the problems and issues mentioned in the previous section) in the form of an exported Excel file.
The main issues and problems that existed at this stage were: Desktop Semantic distinctions refer precisely to the various meanings that a written form of an entry can have, and the separation and definition of these meanings can be considered the primary task in defining an entry.
Word Copy and Paste With the addition of a new evidence or distinction 1- Lexicographers getting involved in irrelevant matters: An important issue in the group's lexicography process was that lexicographers had to spend a significant portion of their time on matters that were not professionally related to them.
Finally, with the effort of the developer and the cooperation of the lexicography group, a web-based software version was provided to the group in Mehr 1390, which significantly reduced the problems of lexicographers, especially regarding searching the corpus.