Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
38 tokens/sec
GPT-4o
59 tokens/sec
Gemini 2.5 Pro Pro
41 tokens/sec
o3 Pro
7 tokens/sec
GPT-4.1 Pro
50 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

RFBES at SemEval-2024 Task 8: Investigating Syntactic and Semantic Features for Distinguishing AI-Generated and Human-Written Texts (2402.14838v1)

Published 19 Feb 2024 in cs.CL, cs.AI, and cs.LG

Abstract: Nowadays, the usage of LLMs has increased, and LLMs have been used to generate texts in different languages and for different tasks. Additionally, due to the participation of remarkable companies such as Google and OpenAI, LLMs are now more accessible, and people can easily use them. However, an important issue is how we can detect AI-generated texts from human-written ones. In this article, we have investigated the problem of AI-generated text detection from two different aspects: semantics and syntax. Finally, we presented an AI model that can distinguish AI-generated texts from human-written ones with high accuracy on both multilingual and monolingual tasks using the M4 dataset. According to our results, using a semantic approach would be more helpful for detection. However, there is a lot of room for improvement in the syntactic approach, and it would be a good approach for future work.

Definition Search Book Streamline Icon: https://streamlinehq.com
References (7)
User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (5)
  1. Mohammad Heydari Rad (1 paper)
  2. Farhan Farsi (2 papers)
  3. Shayan Bali (1 paper)
  4. Romina Etezadi (5 papers)
  5. Mehrnoush Shamsfard (20 papers)
Citations (2)