LongFact is a comprehensive factuality benchmark created through a dynamic process involving GPT-4, focusing on diverse topics across different fields. The topics were carefully selected to ensure comprehensive factuality coverage. The set includes both LongFact-Concepts (niche concepts) and LongFact-Objects (specific entities). Examples of prompts from these categories are provided, and the full dataset is available on GitHub.




