Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

CultureBank: An Online Community-Driven Knowledge Base Towards Culturally Aware Language Technologies

About

To enhance language models' cultural awareness, we design a generalizable pipeline to construct cultural knowledge bases from different online communities on a massive scale. With the pipeline, we construct CultureBank, a knowledge base built upon users' self-narratives with 12K cultural descriptors sourced from TikTok and 11K from Reddit. Unlike previous cultural knowledge resources, CultureBank contains diverse views on cultural descriptors to allow flexible interpretation of cultural knowledge, and contextualized cultural scenarios to help grounded evaluation. With CultureBank, we evaluate different LLMs' cultural awareness, and identify areas for improvement. We also fine-tune a language model on CultureBank: experiments show that it achieves better performances on two downstream cultural tasks in a zero-shot setting. Finally, we offer recommendations based on our findings for future culturally aware language technologies. The project page is https://culturebank.github.io . The code and model is at https://github.com/SALT-NLP/CultureBank . The released CultureBank dataset is at https://huggingface.co/datasets/SALT-NLP/CultureBank .

Weiyan Shi, Ryan Li, Yutong Zhang, Caleb Ziems, Chunhua yu, Raya Horesh, Rog\'erio Abreu de Paula, Diyi Yang• 2024

Related benchmarks

TaskDatasetResultRank
Cultural AlignmentCulturalBench Easy
Accuracy90.8
24
Cultural AlignmentCulturalBench Hard
Accuracy60
24
Cultural AlignmentWVS
Accuracy68.11
24
Cultural AlignmentPRISM
Rating4.323
24
Cultural AlignmentGlobalOpinionQA
Accuracy56.76
24
Content ModerationContent Moderation Korean (test)
Abusive Rate63.5
4
Cultural UnderstandingKorean Cultural Understanding Benchmark (test)
Abusive Score0.635
4
Cultural UnderstandingArabic Cultural Understanding Benchmark (test)
Hate Score54
3
Content ModerationContent Moderation Arabic (test)
Hate Accuracy54
2
Showing 9 of 9 rows

Other info

Follow for update