2/5/2026
Qlean Dataset Launches a Japanese Single-Speaker Scripted Read Speech Dataset with Transcripts

Visual Bank Inc. (Minato-ku, Tokyo; CEO: Saneyuki Nagai), through its subsidiary amanaimages Inc., has launched a new dataset as part of its AI training data solution, Qlean Dataset: the Japanese Single-Speaker Scripted Read Speech Corpus with Transcripts. The dataset is intended for the development and evaluation of speech- and language-based AI systems, including Automatic Speech Recognition (ASR), Natural Language Processing (NLP), and Large Language Models (LLMs).
The dataset is included in Qlean Dataset’s machine learning dataset lineup, AI Data Recipe. It features Japanese audio recordings in which a native Japanese male speaker reads prepared scripts aloud, with each recording paired with accurate transcripts that faithfully represent the spoken content. The use of scripted speech ensures clear alignment between audio and text, making the dataset suitable for tasks that require explicit correspondence between spoken language and written text.
All recordings follow a single-speaker, read-aloud format, minimizing disfluencies commonly found in spontaneous speech, such as self-corrections or topic drift. This structure supports foundational speech and language processing tasks, including the training and evaluation of speech recognition models and the validation of speech-to-text–based processing pipelines.
Qlean Dataset provides AI training data for both research and commercial development, with usage conditions and rights clearance clearly defined. This dataset is offered as part of that framework to support AI development and evaluation environments that require reliable alignment between Japanese speech and textual data.
Dataset Overview: Japanese Single-Speaker Scripted Read Speech Corpus with Transcripts
Data Type | Audio, Text |
|---|---|
Speaker Attributes | Japanese, Male |
Data Format | Audio: MP3 |
Sampling Rate | 44.1 kHz / 48 kHz |
Sample Details |
Use Case Examples
Research Use Cases
Baseline Evaluation of Japanese ASR Models
The dataset can be used to evaluate recognition accuracy and error patterns in Japanese ASR systems, leveraging single-speaker scripted speech where audio–text correspondence is explicitly defined.
Industrial Use Cases
Validation of LLM and Speech-Language Processing Pipelines with Voice Input
The paired Japanese speech and accurate transcripts can be used to verify preprocessing stages that convert speech to text, as well as downstream pipelines that connect ASR outputs to language models.
Additional Practical Applications
Training and Evaluation Data for Speech and Language Processing Systems
The dataset is suitable for educational purposes, such as learning the fundamentals of speech recognition and speech-to-text conversion, as well as for evaluating and comparing the behavior of existing models.
About Qlean Dataset
Qlean Dataset is a commercial-use-ready AI training data solution provided by Amana Images Inc., a subsidiary of Visual Bank Inc.
It supports a wide range of data types, including images, videos, audio, 3D assets, and text, enabling both research and commercial AI development in a legally safe environment.
Through collaborations with data partners such as Chiba Lotte Marines Co., Ltd. and Toyo Keizai Inc., Qlean Dataset continues to expand its specialized, industry-focused lineup known as the “AI Data Recipe.”
By reducing the operational burden of data collection and preparation, Qlean Dataset helps organizations establish AI development environments that are both legally compliant and risk-free.
▶ Qlean Dataset: https://qleandataset.visual-bank.co.jp/en
▶ AI Data Recipe: https://qleandataset.visual-bank.co.jp/en/lineup




Key Features of Qlean Dataset
Existing datasets deliverable within one business day
Custom data collection and recording services available
▶ Contact: https://qleandataset.visual-bank.co.jp/en/contact
About Visual Bank Inc.
Visual Bank Inc. is a Tokyo-based startup building Next-Generation Data infrastructure to enhance AI development capabilities under the mission “Unlocking Data Accessibility.”
The company operates THE PEN, an AI-assisted creative tool for manga artists and the Qlean Dataset service.
Its subsidiaries include Amana Images Inc., one of Japan’s largest photostock providers; Qlean Dataset, which leads research and development in AI data; and THE PEN Inc., an AI-assisted creative tool for manga artists.
CEO: Saneyuki Nagai
Address: 6F, C-Cube Minami Aoyama Building, 7-1-7 Minami-Aoyama, Minato-ku, Tokyo 107-0062
Corporate Site: https://visual-bank.co.jp/en
Amana Images: https://qleandataset.visual-bank.co.jp/en/company-overview





