Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Reusable Component Retrieval: A Semantic Search Approach for Low-Resource Languages

Domain:

natural language processing
Creator:
NazTauAyeTam
Publisher:
Ass
Host:
A common practice among programmers is to reuse existing code, accomplished by performing natural language queries through search engines. The main aim of code retrieval is to search for the most relevant snippet from a corpus of code snippets. However, code retrieval frameworks for low-resource languages are insufficient. Retrieving the most relevant code snippet efficiently can be accomplished only by eliminating the semantic gap between the code snippets residing in the repository and the user’s query (natural language description). The primary objective of the research is to contribute to this field by providing a code search framework that can be extended for low-resource languages. The secondary objective is to provide a code retrieval mechanism that is semantically relevant to the user query and provide programmers with the ability to locate source code that they want to use when developing new applications. The proposed approach is implemented using a web platform to search for source code. As code retrieval is a sophisticated task, the proposed approach incorporates a semantic search mechanism. This research uses a semantic model for code retrieval, which generates meanings or synonyms of words. The proposed model integrates ontologies and Natural Language Processing. System performance measures and classification accuracy are computed using precision, recall, and F1-score. We also compare the proposed approach with state-of-the-art baseline models. The retrieved results are ranked, showing that our approach significantly outperforms robust code matching. Our evaluation shows that semantic matching leads to improved source code retrieval. This study marks a substantial advancement in integrating programming expertise with code retrieval techniques. Moreover, our system lets users know when and how it is used for successful semantic searching.

Visit

doi.org

Tasks

information retrieval

Licenses

https://www.acm.org/publications/policies/copyright_policy#Background

Similar

saisrikar-dev/low-resource-semantic-searchSemantic Alignment Impact on Zero-Shot Retrieval Performance in Low-Resource LanguagesA Proposed Approach for Extracting Semantic and Lexical Relations for Low-Resource Languages: A Case Study of DarijaGUIDE: Creating Semantic Domain Dictionaries for Low-Resource LanguagesLungisanikhan/Enhancing-Semantic-Relatedness-for-Low-Resource-African-Languages-Cross-Lingual Retrieval Augmented Prompt for Low-Resource Languages

saisrikar-dev/low-resource-semantic-search

Multilingual semantic search and query understanding for low-resource languages. # Multilingual Sem

Semantic Alignment Impact on Zero-Shot Retrieval Performance in Low-Resource Languages

Information retrieval across different languages is an increasingly important challenge in natural l

A Proposed Approach for Extracting Semantic and Lexical Relations for Low-Resource Languages: A Case Study of Darija

GUIDE: Creating Semantic Domain Dictionaries for Low-Resource Languages

Over 7,000 of the world's 7,168 living languages are still low-resourced. This paper aims to narrow

Lungisanikhan/Enhancing-Semantic-Relatedness-for-Low-Resource-African-Languages-

Enhancing Semantic Relatedness for Low-Resource African Languages via Transfer Learning and M2M-100

Cross-Lingual Retrieval Augmented Prompt for Low-Resource Languages

Multilingual Pretrained Language Models (MPLMs) have shown their strong multilinguality in recent em