BUTom21: Still image tomato dataset for detection and segmentation
Open this dataset in the live repository
- Persistent identifier
- doi:10.60507/FK2/DHTEH1
- Published version
- 1.0
- Publication date
- 2026-07-17
- License
- CC BY 4.0
Description
BUTom21 dataset is a fully hand-annotated still image dataset of tomatoes. It contains 123, 72, 98 images respectively in the training, validation, and evaluation subsets. The dataset contains RGB and noisy depth images captured across 2 capture days and two different RealSense D435i cameras. This creates a fully pixel-wise annotated dataset for tomato detection and segmentation in a real-world robotics scenario.
Creators
- Halstead, Michael
- Guclu, Esra
- Farag, Mohamed
- Pallotta, Enrico
- Hund, Christian
- Roscher, Ribana
- Bennewitz, Maren
- Gall, Juergen
- McCool, Chris
Keywords
Tomato Detection, Tomato Segmentation, sweeterno, phenotyping
Files
| File | Type | Bytes | Checksum |
|---|---|---|---|
| Halstead_BUTom21_ReadMe.txt | text/plain | 5739 | MD5 99401ee209c44d10c1e714b5b499f094 |
| BUTom21_structure.md | text/markdown | 5419 | MD5 49f756154edca4a025e8fab8e6cc4c22 |
| BUTom21.tar.gz | application/gzip | 500703195 | MD5 09ba15fe38e8b1997f31bce6e9224753 |
Citation
Halstead, Michael; Guclu, Esra; Farag, Mohamed; Pallotta, Enrico; Hund, Christian; Roscher, Ribana; Bennewitz, Maren; Gall, Juergen; McCool, Chris, 2026-07-17, BUTom21: Still image tomato dataset for detection and segmentation, doi:10.60507/FK2/DHTEH1, V1.0
Additional Dataverse fields
| Id | 784 |
|---|---|
| Dataset Type | dataset |
| Internal Version Number | 34 |
| Latest Version Publishing State | RELEASED |
| Release Time | 2026-07-17T11:55:25Z |
| Create Time | 2026-07-14T20:04:38Z |
| Citation Date | 2026-07-17 |
| File Access Request | True |
Export metadata
Static metadata exports available for this published dataset version:
Complete Dataverse metadata
Expected crawler behaviour
Use a stable, truthful User-Agent with product/version and a working contact URL. Across all IP addresses and HTTP connections used by one crawler identity, allow no more than 5 requests in flight and wait at least 20 seconds between request starts. Crawl URLs listed in the catalog sitemap, including file pages and download URLs when they are published, use conditional requests, honor Retry-After, and apply exponential backoff after errors.
The welcome page may link to the interactive repository for human navigation. Automated clients must not treat that human link as a catalog crawl target.
Read the live machine-readable crawler policy before and during a crawl. Stop crawling when it reports CPU or memory utilization at or above 80% and 80% respectively.
