Conversion of Lexicon-Grammar tables to LMF. Application to French
Summary
Describes the conversion of French verb Lexicon-Grammar tables into the LMF format, enhancing interoperability and standardization for NLP dictionaries.
View Cached Full Text
Cached at: 05/15/26, 06:24 AM
# Conversion of Lexicon-Grammar tables to LMF. Application to French Source: [https://arxiv.org/abs/2605.14816](https://arxiv.org/abs/2605.14816) [View PDF](https://arxiv.org/pdf/2605.14816) > Abstract:We describe the first experiment of conversion of Lexicon\-Grammar tables for French verbs into the Lexical Markup Framework \(LMF\) format\. The Lexicon\-Grammar of the French language is currently one of the major sources of lexical and syntactic information for French\. Its conversion into an interoperable representation format according to the LMF standard makes it usable in different contexts, thus contributing to the standardization and interoperability of natural language processing dictionaries\. We briefly introduce the Lexicon\-Grammar and the derived dictionaries; we analyse the main difficulties faced during the conversion; and we describe the resulting resource\. ## Submission history From: Eric Laporte \[[view email](https://arxiv.org/show-email/d4eef2af/2605.14816)\] **\[v1\]**Thu, 14 May 2026 13:28:24 UTC \(789 KB\)
Similar Articles
French parsing enhanced with a word clustering method based on a syntactic lexicon
This article evaluates the integration of data from the French syntactic lexicon Lexicon-Grammar into a probabilistic parser, using word clustering methods on verbs to improve parsing accuracy for French.
On the Use of LLMs for Specialised Terminology: A Good Alternative to Corpora?
This study evaluates four proprietary LLMs (GPT-4o, GPT-5.2, Claude Sonnet 4.5, DeepSeek) for specialized terminology translation from English to French across two domains, comparing prompting strategies. Results show Claude Sonnet 4.5 performs best, but LLMs cannot yet replace specialized corpora.
Luth-2: New State-of-the-Art French Small Language Models
Luth-2 releases two French small language models (0.8B and 2B) that achieve state-of-the-art results on French benchmarks for their size, with open-weights and data on Hugging Face.
From Lexicon to AI: A Structured-Data Pipeline for Specialized Conversational Systems in Low-Resource Languages
Presents a systematic methodology for converting Hindi WordNet into 1.25 million instruction-response pairs to fine-tune a 12B-parameter language model using LoRA, demonstrating improved pedagogical effectiveness for specialized conversational systems in low-resource languages.
Analyzing and Encoding the Al-Mawrid Arabic-English Dictionary with the ISO Language Markup Framework and TEI Lex-0
This paper presents a methodology for digitizing the Al-Mawrid Arabic-English dictionary using ISO LMF and TEI Lex-0 standards, achieving high parsing accuracy and precision, and addressing gaps in Arabic lexical infrastructure.