ConcurrentQA is a textual multi-hop QA benchmark to require concurrent retrieval over multiple data-distributions (i.e. Wikipedia and email data). This dataset was constructed by researchers at Stanf
# OmniThought-0528 *A High-Quality Chain-of-Thought (CoT) Dataset for Enhanced Model Distillation* ## **Overview** OmniThought-0528 is an advanced version of the **OmniThought** dataset, design
# Dataset Card for Indic STS This dataset is STS benchmark between English and 12 high-resource Indic languages. This was released as a part of [Samanantar](https://arxiv.org/abs/2104.05596) paper.