Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

PatentGPT: A Large Language Model for Intellectual Property (2404.18255v5)

Published 28 Apr 2024 in cs.CL and cs.AI

Abstract: In recent years, LLMs(LLMs) have attracted significant attention due to their exceptional performance across a multitude of natural language process tasks, and have been widely applied in various fields. However, the application of LLMs in the Intellectual Property (IP) domain is challenging due to the strong need for specialized knowledge, privacy protection, processing of extremely long text in this field. In this technical report, we present for the first time a low-cost, standardized procedure for training IP-oriented LLMs, meeting the unique requirements of the IP domain. Using this standard process, we have trained the PatentGPT series models based on open-source pretrained models. By evaluating them on the open-source IP-oriented benchmark MOZIP, our domain-specific LLMs outperforms GPT-4, indicating the effectiveness of the proposed training procedure and the expertise of the PatentGPT models in the IP domain. Remarkably, our model surpassed GPT-4 on the 2019 China Patent Agent Qualification Examination, scoring 65 and matching human expert levels. Additionally, the PatentGPT model, which utilizes the SMoE architecture, achieves performance comparable to that of GPT-4 in the IP domain and demonstrates a better cost-performance ratio on long-text tasks, potentially serving as an alternative to GPT-4 within the IP domain.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (27)
  1. Zilong Bai (5 papers)
  2. Linqing Chen (5 papers)
  3. Qijun Cai (2 papers)
  4. Yuan Zhong (70 papers)
  5. Jie Fang (20 papers)
  6. Jing Sun (115 papers)
  7. Weikuan Wang (1 paper)
  8. Lizhi Zhou (2 papers)
  9. Chaochao Wang (3 papers)
  10. Cheng Sun (40 papers)
  11. Jianping Lu (4 papers)
  12. Yixin Wang (103 papers)
  13. Haowen Liu (12 papers)
  14. Peng Xu (357 papers)
  15. Licong Xu (3 papers)
  16. Fu Bian (3 papers)
  17. Xiaolong Gu (1 paper)
  18. Changyang Tu (3 papers)
  19. Cong Wang (310 papers)
  20. Yan Fang (20 papers)
Citations (2)

Summary

We haven't generated a summary for this paper yet.

X Twitter Logo Streamline Icon: https://streamlinehq.com
Youtube Logo Streamline Icon: https://streamlinehq.com