---
title: FaithEval-counterfactual-v1.0
canonical_url: "https://www.modelscope.cn/datasets/Salesforce/FaithEval-counterfactual-v1.0"
md_url: "https://www.modelscope.cn/datasets/Salesforce/FaithEval-counterfactual-v1.0.md"
repository: Salesforce/FaithEval-counterfactual-v1.0
last_updated: 2025-08-15
license: "Apache License 2.0"
storage_size: "1.6 MB"
downloads: 226
stars: 0
---

# FaithEval-counterfactual-v1.0

> FaithEval-counterfactual-v1.0 - Salesforce 在 ModelScope 开源的数据集。FaithEval is a new and comprehensive benchmark dedicated to evaluating contextual faithfulness in LLMs across three diverse tasks: unanswerable, inconsistent, and counterfactual contexts.

Salesforce/FaithEval-counterfactual-v1.0 是 ModelScope 魔搭社区上的数据集，存储大小 1.6 MB，采用 Apache License 2.0 许可。

- **Repository**: Salesforce/FaithEval-counterfactual-v1.0
- **License**: Apache License 2.0
- **Storage size**: 1.6 MB
- **Downloads**: 226
- **Stars**: 0
- **Last updated**: 2025-08-15

Source: https://www.modelscope.cn/datasets/Salesforce/FaithEval-counterfactual-v1.0

---

# FaithEval

FaithEval is a new and comprehensive benchmark dedicated to evaluating contextual faithfulness in LLMs across three diverse tasks: unanswerable, inconsistent, and counterfactual contexts.

[Paper] FaithEval: Can Your Language Model Stay Faithful to Context, Even If "The Moon is Made of Marshmallows", ICLR 2025, https://arxiv.org/abs/2410.03727

[Code and Detailed Instructions] https://github.com/SalesforceAIResearch/FaithEval


## Disclaimer and Ethical Considerations
This release is for research purposes only in support of an academic paper. Our datasets and code are not specifically designed or evaluated for all downstream purposes. We encourage users to consider the common limitations of AI, comply with applicable laws, and leverage best practices when selecting use cases, particularly for high-risk scenarios where errors or misuse could significantly impact people’s lives, rights, or safety. For further guidance on use cases, refer to our AUP and AI AUP.
