Webtunix Developers

CharacterDetection

Object Detection

CharacterDetection Image Dataset

v1

2024-03-05 12:03am

Generated on Mar 4, 2024

Popular Download Formats

Pascal VOC XML
Common XML annotation format for local data munging (pioneered by ImageNet).
PaliGemma
PaliGemma JSONL format used for fine-tuning PaliGemma, Google's open multimodal vision model.
CreateML JSON
CreateML JSON format is used with Apple's CreateML and Turi Create tools.
Other Formats
Choose another format.

Preprocessing

Auto-Orient: Applied
Isolate Objects: Applied
Static Crop: 25-75% Horizontal Region, 25-75% Vertical Region
Dynamic Crop: Class: 0
Resize: Stretch to 1024x1024
Auto-Adjust Contrast: Using Adaptive Equalization
Grayscale: Applied
Tile: 2 rows x 2 columns
Modify Classes: 1 remapped, 0 dropped
Filter Null: Require all images to contain annotations.

Augmentations

Outputs per training example: 3
Flip: Horizontal
90° Rotate: Clockwise, Counter-Clockwise
Crop: 0% Minimum Zoom, 20% Maximum Zoom
Rotation: Between -15° and +15°
Shear: ±10° Horizontal, ±10° Vertical
Grayscale: Apply to 15% of images
Hue: Between -15° and +15°
Saturation: Between -25% and +25%
Brightness: Between -15% and +15%
Exposure: Between -10% and +10%
Blur: Up to 2.5px
Noise: Up to 0.1% of pixels
Cutout: 3 boxes with 10% size each
Mosaic: Applied
Bounding Box: Flip: Horizontal, Vertical
Bounding Box: Crop: 0% Minimum Zoom, 35% Maximum Zoom
Bounding Box: Shear: ±15° Horizontal, ±17° Vertical
Bounding Box: Blur: Up to 2.5px
Bounding Box: Noise: Up to 0.1% of pixels