مجموعات بيانات Hugging Face: دليل المطوِّرين
A Hugging Face dataset is a structured data collection hosted on the Hugging Face Hub, searchable at huggingface.co/datasets — over 300,000 public datasets as of 2026. Load any dataset in one line: from datasets import load_dataset; ds = load_dataset(‘stanfordnlp/imdb’)Each dataset ships with typed splits (train/validation/test), a features schema (text, image, audio, labels), and optional streaming for terabyte-scale files. Push your own data with ds.push_to_hub(‘your-username/your-dataset’) after running huggingface-cli login. A Hugging Face dataset is a versioned, structured data collection stored on the Hugging Face Hub and consumed through the datasets Python library.












