Buckets:
218 GB
2,270 files
Updated about 16 hours ago
Ctrl+K
| Name | Size | Uploaded | Xet hash |
|---|---|---|---|
| batch_0 | 11 items | ||
| batch_1 | 11 items | ||
| batch_10 | 11 items | ||
| batch_11 | 11 items | ||
| batch_12 | 11 items | ||
| batch_13 | 11 items | ||
| batch_14 | 11 items | ||
| batch_15 | 11 items | ||
| batch_16 | 11 items | ||
| batch_17 | 11 items | ||
| batch_18 | 11 items | ||
| batch_19 | 11 items | ||
| batch_2 | 11 items | ||
| batch_20 | 11 items | ||
| batch_21 | 11 items | ||
| batch_22 | 11 items | ||
| batch_23 | 11 items | ||
| batch_24 | 11 items | ||
| batch_25 | 11 items | ||
| batch_26 | 11 items | ||
| batch_27 | 11 items | ||
| batch_28 | 11 items | ||
| batch_29 | 11 items | ||
| batch_3 | 11 items | ||
| batch_30 | 11 items | ||
| batch_31 | 11 items | ||
| batch_32 | 11 items | ||
| batch_33 | 11 items | ||
| batch_34 | 11 items | ||
| batch_35 | 11 items | ||
| batch_36 | 11 items | ||
| batch_37 | 11 items | ||
| batch_38 | 11 items | ||
| batch_39 | 11 items | ||
| batch_4 | 11 items | ||
| batch_40 | 11 items | ||
| batch_41 | 11 items | ||
| batch_42 | 11 items | ||
| batch_43 | 11 items | ||
| batch_44 | 11 items | ||
| batch_45 | 11 items | ||
| batch_46 | 11 items | ||
| batch_47 | 11 items | ||
| batch_48 | 11 items | ||
| batch_49 | 11 items | ||
| batch_5 | 11 items | ||
| batch_6 | 11 items | ||
| batch_7 | 11 items | ||
| batch_8 | 11 items | ||
| batch_9 | 11 items | ||
| .gitattributes | 2.46 kB xet | 19463de8 | |
| README.md | 2.86 kB xet | 8e8032a3 | |
| progress.json | 869 Bytes xet | 1d8b3cd7 |
🎵 Suno Audio Dataset
A comprehensive dataset of 49,698 AI-generated music tracks from Suno, organized in 50 batches of 1000 samples each.
🎧 All audio files are playable directly in the dataset viewer!
Dataset Structure
The dataset is organized into batches (batch_0, batch_1, etc.), each containing up to 1000 audio samples with metadata.
Fields
audio: 🎵 Playable MP3 audio file (click to play in viewer!)id: Unique track identifiertitle: Song titledisplay_name: Creator/artist namehandle: Creator handletags: Music tags, genres, and stylesprompt: Text prompt used for generationduration: Track duration in secondsplay_count: Number of plays on Sunoupvote_count: Community upvotesmodel_name: Suno model version usedcreated_at: Creation timestampstatus: Track statusis_public: Public visibility flag
Usage
Load Entire Dataset
from datasets import load_dataset
# Load all batches
dataset = load_dataset("Humair332/suno-audio")
print(f"Total tracks: {len(dataset['train'])}")
Load Specific Batch
# Load only batch 0
dataset = load_dataset("Humair332/suno-audio", data_dir="batch_0")
Play Audio
# Get audio data
audio_data = dataset['train'][0]['audio']
audio_array = audio_data['array']
sampling_rate = audio_data['sampling_rate']
# Play in Jupyter/Colab
from IPython.display import Audio
Audio(audio_array, rate=sampling_rate)
Filter by Tags
# Filter by genre
rock_songs = dataset['train'].filter(lambda x: 'rock' in x['tags'].lower())
print(f"Found {len(rock_songs)} rock songs")
Most Popular Tracks
# Sort by play count
from datasets import Dataset
df = dataset['train'].to_pandas()
top_tracks = df.nlargest(10, 'play_count')[['title', 'display_name', 'play_count']]
print(top_tracks)
Dataset Statistics
- Total Tracks: 49,698
- Batches: 50
- Batch Size: 1000
- Format: Apache Arrow with embedded MP3 audio
- Audio Format: MP3
- Metadata: Tags, prompts, engagement metrics
Explore the Music! 🎶
Click on the dataset viewer above and browse through the tracks. Click any row to play the audio directly in your browser!
Source
Original dataset: nyuuzyou/suno
License
MIT License
Citation
If you use this dataset, please cite:
@dataset{suno_audio_dataset,
title={Suno Audio Dataset},
author={Humair332},
year={2026},
publisher={Hugging Face},
url={https://huggingface.co/datasets/Humair332/suno-audio}
}
- Total size
- 218 GB
- Files
- 2,270
- Last updated
- Oct 3
- Pre-warmed CDN
- US EU US EU