speech recognition for Amharic language
# Amharic-ASR-Dataset
speech recognition for Amharic language
--------------------------------------------------------------------
Text data is adobted from Contemporary Amharic Corpus (CACO)
## Length
--------------------------------------
- 2.3 hours
## unsplitted
------------------------------------------------------------
unsplitted row data can be found in `unsplitted` branch of this repository.
## Contents
--------------------------------------------------------------------
```
Amharic-ASR-Dataset
|
├── data
│ └── records
│ ├── train
│ │ └── *.wav
│ ├── val
│ │ └── *.wav
│ ├── test
│ └── *.wav
│
└── README.md
└── chars.txt
└── charset.json
└── linker.txt
└── raw_text_file.txt
```
## Data collection App
-------------------------------
Dataset was recorded using LIG-Aikuma app