3/25/2026
Qlean Dataset Launches Single-Speaker Japanese Classical Audio Corpus

Visual Bank Inc. (Minato-ku, Tokyo; Saneyuki Nagai, Representative Director and CEO) has announced the release of the "Japanese Single-Speaker Classical Literature Audio Dataset" through its AI training data solution, "Qlean Dataset," managed by its subsidiary, amanaimages inc. This dataset is specifically optimized for training speech and text models aimed at improving Text-to-Speech (TTS) precision and sophisticated linguistic understanding of classical Japanese expressions.
The dataset consists of high-quality audio recordings of Japanese classical literary works paired with accurate text transcripts. Featuring consistent speech from a single Japanese narrator, the corpus covers a wide range of complex grammatical structures, archaic phrasing, and rhythmic patterns unique to classical literature. This makes it an ideal resource for developing models that replicate specific voice characteristics or for research into prosody control in speech generation (AI Voice) involving non-contemporary Japanese. By enabling deep learning of the correlation between acoustic features and text—such as breath intake, intonation shifts in long-form narration, and phonetic estimation for classical script—the dataset facilitates the generation of natural, context-aware speech.
This release is part of the "AI Data Recipe" lineup, Qlean Dataset’s original series of high-quality data products for AI development. It is designed for practical AI implementation phases, ranging from the automated generation of audiobooks to the deployment of Automatic Speech Recognition (ASR) systems that support the digital archiving of historical documents. Visual Bank and amanaimages remain committed to supporting the research and development of AI capable of understanding and generating diverse expressions by providing premium audio and linguistic data that capture Japan’s cultural heritage.
Dataset Overview: Japanese Single-Speaker Classical Literature Audio Dataset
Data Type: | Audio, Text |
|---|---|
Subject Attributes: | Japanese Speaker |
File Format: | Audio (mp3), Text (txt, json) |
Recording Duration: | 30 seconds to 90 minutes per file |
Audio Sampling Rate: | 44.1kHz / 48kHz |
Sample Details: |
Potential Use Cases
【Research & Academia】
Acoustic Feature Extraction and Prosodic Analysis of Classical Language:
Used to analyze how a consistent speaker manages pitch and pauses relative to archaic grammatical structures and vocabulary, supporting the construction of prosody models specific to classical literature.
【Industrial Application】
Development of High-Precision TTS Models for Entertainment:
Training on stable, long-form data from a single speaker enables the creation of highly consistent and emotive TTS engines suitable for audiobooks and digital content.
Improving Accuracy in ASR for Classical Expressions:
Provides linguistic adaptation for speech recognition models, allowing for higher accuracy in tools and search systems dedicated to transcribing or indexing historical documents.
【Educational & Social Implementation】
Development of Self-Study Apps for Classical Literature:
Allows for the implementation of AI-driven educational tools that compare a student's reading against professional reference audio and transcripts to provide real-time evaluation and feedback.
Accessibility Improvements for the Visually Impaired or Learners with Difficulties:
Facilitates the creation of accessible educational environments and digital textbooks by providing high-quality, consistent audio for complex kanji and historical kana usage.
About Qlean Dataset
Qlean Dataset is a commercially cleared AI training data solution provided by Amana Images, a subsidiary of Visual Bank Group. The platform offers diverse data formats including image, video, audio, 3D, and text, as well as a specialized AI Data Recipe lineup developed through collaborations with major media organizations and data rights holders.
URL:https://qleandataset.visual-bank.co.jp/en




About Visual Bank Inc.
Visual Bank Group is a technology company developing data infrastructure and AI solutions that support advanced AI development. The company operates THE PEN, an AI tool for manga creators, and its subsidiary, amanaimages Inc., provides commercial digital content and AI training data solutions, including Qlean Dataset. Visual Bank is also a selected participant in GENIAC, a Japanese government initiative supporting the advancement of next generation AI technologies.
CEO: Saneyuki Nagai
Website:https://visual-bank.co.jp/en





