Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Multilingual Evaluation of Semantic Textual Relatedness (2404.09047v1)

Published 13 Apr 2024 in cs.CL

Abstract: The explosive growth of online content demands robust NLP techniques that can capture nuanced meanings and cultural context across diverse languages. Semantic Textual Relatedness (STR) goes beyond superficial word overlap, considering linguistic elements and non-linguistic factors like topic, sentiment, and perspective. Despite its pivotal role, prior NLP research has predominantly focused on English, limiting its applicability across languages. Addressing this gap, our paper dives into capturing deeper connections between sentences beyond simple word overlap. Going beyond English-centric NLP research, we explore STR in Marathi, Hindi, Spanish, and English, unlocking the potential for information retrieval, machine translation, and more. Leveraging the SemEval-2024 shared task, we explore various LLMs across three learning paradigms: supervised, unsupervised, and cross-lingual. Our comprehensive methodology gains promising results, demonstrating the effectiveness of our approach. This work aims to not only showcase our achievements but also inspire further research in multilingual STR, particularly for low-resourced languages.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (6)
  1. Sharvi Endait (5 papers)
  2. Srushti Sonavane (3 papers)
  3. Ridhima Sinare (3 papers)
  4. Pritika Rohera (3 papers)
  5. Advait Naik (1 paper)
  6. Dipali Kadam (6 papers)
Citations (2)

Summary

We haven't generated a summary for this paper yet.