2/5/2026

Qlean Dataset Launches a Japanese Single-Speaker Scripted Read Speech Dataset with Transcripts

Visual Bank Inc. (Minato-ku, Tokyo; CEO: Saneyuki Nagai), through its subsidiary amanaimages Inc., has launched a new dataset as part of its AI training data solution, Qlean Dataset: the Japanese Single-Speaker Scripted Read Speech Corpus with Transcripts. The dataset is intended for the development and evaluation of speech- and language-based AI systems, including Automatic Speech Recognition (ASR), Natural Language Processing (NLP), and Large Language Models (LLMs).

The dataset is included in Qlean Dataset’s machine learning dataset lineup, AI Data Recipe. It features Japanese audio recordings in which a native Japanese male speaker reads prepared scripts aloud, with each recording paired with accurate transcripts that faithfully represent the spoken content. The use of scripted speech ensures clear alignment between audio and text, making the dataset suitable for tasks that require explicit correspondence between spoken language and written text.

All recordings follow a single-speaker, read-aloud format, minimizing disfluencies commonly found in spontaneous speech, such as self-corrections or topic drift. This structure supports foundational speech and language processing tasks, including the training and evaluation of speech recognition models and the validation of speech-to-text–based processing pipelines.

Qlean Dataset provides AI training data for both research and commercial development, with usage conditions and rights clearance clearly defined. This dataset is offered as part of that framework to support AI development and evaluation environments that require reliable alignment between Japanese speech and textual data.

Dataset Overview: Japanese Single-Speaker Scripted Read Speech Corpus with Transcripts

Data Type

Audio, Text

Speaker Attributes

Japanese, Male

Data Format

Audio: MP3
Text: txt,json,csv

Sampling Rate

44.1 kHz / 48 kHz

Sample Details

https://qleandataset.visual-bank.co.jp/en/lineup/pn-010

Use Case Examples

Research Use Cases

  • Baseline Evaluation of Japanese ASR Models
    The dataset can be used to evaluate recognition accuracy and error patterns in Japanese ASR systems, leveraging single-speaker scripted speech where audio–text correspondence is explicitly defined.

Industrial Use Cases

  • Validation of LLM and Speech-Language Processing Pipelines with Voice Input
    The paired Japanese speech and accurate transcripts can be used to verify preprocessing stages that convert speech to text, as well as downstream pipelines that connect ASR outputs to language models.

Additional Practical Applications

  • Training and Evaluation Data for Speech and Language Processing Systems
    The dataset is suitable for educational purposes, such as learning the fundamentals of speech recognition and speech-to-text conversion, as well as for evaluating and comparing the behavior of existing models.

About Qlean Dataset

Qlean Dataset is a commercial-use-ready AI training data solution provided by Amana Images Inc., a subsidiary of Visual Bank Inc.
It supports a wide range of data types, including images, videos, audio, 3D assets, and text, enabling both research and commercial AI development in a legally safe environment.
Through collaborations with data partners such as Chiba Lotte Marines Co., Ltd. and Toyo Keizai Inc., Qlean Dataset continues to expand its specialized, industry-focused lineup known as the “AI Data Recipe.”
By reducing the operational burden of data collection and preparation, Qlean Dataset helps organizations establish AI development environments that are both legally compliant and risk-free.

▶ Qlean Dataset: https://qleandataset.visual-bank.co.jp/en
▶ AI Data Recipe: https://qleandataset.visual-bank.co.jp/en/lineup

Key Features of Qlean Dataset

  • Existing datasets deliverable within one business day

  • Custom data collection and recording services available

▶ Contact: https://qleandataset.visual-bank.co.jp/en/contact

About Visual Bank Inc.

Visual Bank Inc. is a Tokyo-based startup building Next-Generation Data infrastructure to enhance AI development capabilities under the mission “Unlocking Data Accessibility.”
The company operates THE PEN, an AI-assisted creative tool for manga artists and the Qlean Dataset service.
Its subsidiaries include Amana Images Inc., one of Japan’s largest photostock providers; Qlean Dataset, which leads research and development in AI data; and THE PEN Inc., an AI-assisted creative tool for manga artists.

CEO: Saneyuki Nagai
Address: 6F, C-Cube Minami Aoyama Building, 7-1-7 Minami-Aoyama, Minato-ku, Tokyo 107-0062
Corporate Site: https://visual-bank.co.jp/en
Amana Images: https://qleandataset.visual-bank.co.jp/en/company-overview

    amana images inc.

    Visual Bank Inc.


    © amanaimages inc.