Papers
Topics
Authors
Recent
Search
2000 character limit reached

AIpom at SemEval-2024 Task 8: Detecting AI-produced Outputs in M4

Published 28 Mar 2024 in cs.CL | (2403.19354v1)

Abstract: This paper describes AIpom, a system designed to detect a boundary between human-written and machine-generated text (SemEval-2024 Task 8, Subtask C: Human-Machine Mixed Text Detection). We propose a two-stage pipeline combining predictions from an instruction-tuned decoder-only model and encoder-only sequence taggers. AIpom is ranked second on the leaderboard while achieving a Mean Absolute Error of 15.94. Ablation studies confirm the benefits of pipelining encoder and decoder models, particularly in terms of improved performance.

Definition Search Book Streamline Icon: https://streamlinehq.com
References (17)
  1. Scaling Instruction-Finetuned Language Models. arXiv preprint arXiv:2210.11416.
  2. Automatic Detection of Hybrid Human-Machine Text Boundaries.
  3. Real or Fake Text?: Investigating Human Ability to Detect Boundaries between Human-written and Machine-Generated Text. In Proceedings of the AAAI Conference on Artificial Intelligence, volume 37, pages 12763–12771.
  4. GLTR: Statistical detection and visualization of generated text. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics: System Demonstrations, pages 111–116, Florence, Italy. Association for Computational Linguistics.
  5. DeBERTav3: Improving deBERTa using ELECTRA-style pre-training with gradient-disentangled embedding sharing. In The Eleventh International Conference on Learning Representations.
  6. Lora: Low-rank adaptation of large language models. In International Conference on Learning Representations.
  7. Automatic detection of machine generated text: A critical survey. In Proceedings of the 28th International Conference on Computational Linguistics, pages 2296–2309, Barcelona, Spain (Online). International Committee on Computational Linguistics.
  8. Mistral 7b. arXiv preprint arXiv:2310.06825.
  9. Diederik Kingma and Jimmy Ba. 2015. Adam: A method for stochastic optimization. In International Conference on Learning Representations (ICLR), San Diega, CA, USA.
  10. Efficient memory management for large language model serving with pagedattention. In Proceedings of the ACM SIGOPS 29th Symposium on Operating Systems Principles.
  11. Crosslingual Generalization through Multitask Finetuning. arXiv preprint arXiv:2211.01786.
  12. LLaMA: Open and Efficient Foundation Language Models.
  13. Adaku Uchendu. 2023. Reverse Turing Test in the Age of Deepfake Texts. Ph.D. thesis, The Pennsylvania State University.
  14. Semeval-2024 task 8: Multigenerator, multidomain, and multilingual black-box machine-generated text detection. In Proceedings of the 18th International Workshop on Semantic Evaluation, SemEval.
  15. M4: Multi-generator, multi-domain, and multi-lingual black-box machine-generated text detection. In Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers), pages 1369–1407, St. Julian’s, Malta. Association for Computational Linguistics.
  16. Taxonomy of Risks Posed by Language Models. In Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency, pages 214–229.
  17. Transformers: State-of-the-art natural language processing. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations, pages 38–45, Online. Association for Computational Linguistics.
Citations (1)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.