NatLogAttack: A Framework for Attacking Natural Language Inference Models with Natural Logic

Zheng, Zi'ou; Zhu, Xiaodan

Computer Science > Computation and Language

arXiv:2307.02849 (cs)

[Submitted on 6 Jul 2023 (v1), last revised 11 Oct 2024 (this version, v2)]

Title:NatLogAttack: A Framework for Attacking Natural Language Inference Models with Natural Logic

Authors:Zi'ou Zheng, Xiaodan Zhu

View PDF HTML (experimental)

Abstract:Reasoning has been a central topic in artificial intelligence from the beginning. The recent progress made on distributed representation and neural networks continues to improve the state-of-the-art performance of natural language inference. However, it remains an open question whether the models perform real reasoning to reach their conclusions or rely on spurious correlations. Adversarial attacks have proven to be an important tool to help evaluate the Achilles' heel of the victim models. In this study, we explore the fundamental problem of developing attack models based on logic formalism. We propose NatLogAttack to perform systematic attacks centring around natural logic, a classical logic formalism that is traceable back to Aristotle's syllogism and has been closely developed for natural language inference. The proposed framework renders both label-preserving and label-flipping attacks. We show that compared to the existing attack models, NatLogAttack generates better adversarial examples with fewer visits to the victim models. The victim models are found to be more vulnerable under the label-flipping setting. NatLogAttack provides a tool to probe the existing and future NLI models' capacity from a key viewpoint and we hope more logic-based attacks will be further explored for understanding the desired property of reasoning.

Comments:	Published as a conference paper at ACL 2023
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2307.02849 [cs.CL]
	(or arXiv:2307.02849v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2307.02849

Submission history

From: Zi'ou Zheng [view email]
[v1] Thu, 6 Jul 2023 08:32:14 UTC (7,516 KB)
[v2] Fri, 11 Oct 2024 00:45:25 UTC (7,516 KB)

Computer Science > Computation and Language

Title:NatLogAttack: A Framework for Attacking Natural Language Inference Models with Natural Logic

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:NatLogAttack: A Framework for Attacking Natural Language Inference Models with Natural Logic

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators