Evaluating Large Language Models for Causal Modeling (2411.15888v1)

Published 24 Nov 2024 in cs.CL

Abstract: In this paper, we consider the process of transforming causal domain knowledge into a representation that aligns more closely with guidelines from causal data science. To this end, we introduce two novel tasks related to distilling causal domain knowledge into causal variables and detecting interaction entities using LLMs. We have determined that contemporary LLMs are helpful tools for conducting causal modeling tasks in collaboration with human experts, as they can provide a wider perspective. Specifically, LLMs, such as GPT-4-turbo and Llama3-70b, perform better in distilling causal domain knowledge into causal variables compared to sparse expert models, such as Mixtral-8x22b. On the contrary, sparse expert models such as Mixtral-8x22b stand out as the most effective in identifying interaction entities. Finally, we highlight the dependency between the domain where the entities are generated and the performance of the chosen LLM for causal modeling.

Collections

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Follow-up Questions

We haven't generated follow-up questions for this paper yet.

Generate Now

Evaluating Large Language Models for Causal Modeling (2411.15888v1)

Collections

Summary

Follow-up Questions

Authors (4)

Don't miss out on important new AI/ML research

Evaluating Large Language Models for Causal Modeling (2411.15888v1)

Collections

Summary

Follow-up Questions

Related Papers

Authors (4)

Don't miss out on important new AI/ML research