publications
International Conference (peer-reviewed)
2026
-
Multi-modal, Multi-task, Multi-criteria Automatic Evaluation Using a Vision Language ModelIn LREC, 2026 -
DISCODE: Distribution-Aware Score Decoder for Robust Automatic Evaluation of Image CaptioningIn AAAI 2026, 2026
2025
-
Building Instruction-Tuning Datasets from Human-Written Instructions with Open-Weight Large Language ModelsIn COLM 2025, 2025 -
HALL-E: Hierarchical Neural Codec Language Model for Minute-Long Zero-Shot Text-to-Speech SynthesisIn ICLR 2025, 2025Proposed an LLM‑based text‑to‑speech model for minute‑long zero-shot synthesis.
2024
-
Continual Pre-Training for Cross-Lingual LLM Adaptation: Enhancing Japanese Language CapabilitiesIn COLM 2024, 2024 -
Journal (peer-reviewed)
2025
- 大規模言語モデルにおける評価バイアスの尤度に基づく緩和 (English title: Likelihood-based Mitigation of Evaluation Bias in Large Language Models)Journal of Natural Language Processing, Jun 2025Best Paper Award.
2024
-
ELP-Adapters: Parameter Efficient Adapter Tuning for Various Speech Processing TasksIEEE/ACM Transactions on Audio, Speech, and Language Processing, 2024
Workshop (peer-reviewed; non-archival)
2025
-
Why We Build Local Large Language Models: An Observational Analysis from 35 Japanese and Multilingual LLMsIn MELT 2025, 2025
Preprint (non-peer-reviewed)
No preprints currently listed.
Domestic Conference and Symposium (non-peer-reviewed)
2026
- 自己回帰性を組み込んだ直接選好最適化 (English title: Autoregressive Direct Preference Optimization)In The 32nd Annual Meeting of the Association for Natural Language Processing, 2026In Japanese. Best Paper Award.
- 対照的デコーディングを用いた指示学習データの合成 (English title: Synthesizing Instruction-Tuning Datasets with Contrastive Decoding)In The 32nd Annual Meeting of the Association for Natural Language Processing, 2026In Japanese.
- 蒸留による日英推論型大規模言語モデル構築戦略の探索 (English title: Exploring Strategies for Building Japanese-English Reasoning Large Language Models through Distillation)In The 32nd Annual Meeting of the Association for Natural Language Processing, 2026In Japanese.
2025
- JUBAKU: 日本文化における偏見評価のための敵対的ベンチマーク (English title: JUBAKU: An Adversarial Benchmark for Evaluating Bias in Japanese Culture)In The 39th Annual Conference of the Japanese Society for Artificial Intelligence, 2025In Japanese. Annual Conference Award.
- 複数タスク・複数項目に跨ったマルチモーダル自動評価手法 (English title: Multi-modal, Multi-task, Multi-criteria Automatic Evaluation)In The 31st Annual Meeting of the Association for Natural Language Processing, 2025In Japanese. Committee Special Award.
- 模倣学習による大規模言語モデルの指示チューニング (English title: Instruction Tuning for Large Language Models through Imitation Learning)In The 31st Annual Meeting of the Association for Natural Language Processing, 2025In Japanese.
- 新聞記事からつくる時事と社会に強い日本語LLM (English title: Building a Japanese LLM Strong in Current Affairs and Society from Newspaper Articles)In The 31st Annual Meeting of the Association for Natural Language Processing, 2025In Japanese.
- Swallowコーパスv2: 教育的な日本語ウェブコーパスの構築 (English title: Swallow Corpus v2: Building an Educational Japanese Web Corpus)In The 31st Annual Meeting of the Association for Natural Language Processing, 2025In Japanese.
2024
- LLMに日本語テキストを学習させる意義 (English title: The Value of Training LLMs on Japanese Text)In The 261st NL Research Presentation, 2024In Japanese. Best Research Award.
- 大規模言語モデルにおける評価バイアスの尤度に基づく緩和 (English title: Likelihood-based Mitigation of Evaluation Bias in Large Language Models)In The 30th Annual Meeting of the Association for Natural Language Processing, 2024In Japanese. Young Researcher Award.
- Swallowコーパス: 日本語大規模ウェブコーパスの構築 (English title: Swallow Corpus: A Large-Scale Japanese Web Corpus)In The 30th Annual Meeting of the Association for Natural Language Processing, 2024In Japanese. Outstanding Paper Award.
- 大規模言語モデルの日本語能力の効率的な強化: 継続事前学習における語彙拡張と対訳コーパスの活用 (English title: Efficiently Enhancing the Japanese Capabilities of Large Language Models: Vocabulary Expansion and Parallel Corpora in Continual Pre-training)In The 30th Annual Meeting of the Association for Natural Language Processing, 2024In Japanese.
- 継続事前学習による日本語に強い大規模言語モデルの構築 (English title: Building a Japanese-Strong Large Language Model through Continual Pre-training)In The 30th Annual Meeting of the Association for Natural Language Processing, 2024In Japanese. Outstanding Paper Award.