Multimodal Evaluation & Metricsマルチモーダルモデルの評価
Building reliable evaluation protocols for vision-language models, with a focus on grounding and cultural nuance in Japanese contexts.画像と言語の対応や、日本語の文化的なニュアンスを捉える、信頼性の高い視覚言語モデルの評価手法を開発しています。
RESEARCHER / VISION & LANGUAGE研究者 / 視覚と言語
Doctoral Student · Vision and Language Researcher博士後期課程 · 視覚と言語の研究
Exploring the intersection of computer vision and natural language processing, with a focus on multimodal evaluation metrics and context-aware image captioning.コンピュータビジョンと自然言語処理の交差領域を研究しています。特に、マルチモーダルモデルの評価指標と、文脈を考慮した画像説明生成に取り組んでいます。
Explore my research研究業績を見る
Building reliable evaluation protocols for vision-language models, with a focus on grounding and cultural nuance in Japanese contexts.画像と言語の対応や、日本語の文化的なニュアンスを捉える、信頼性の高い視覚言語モデルの評価手法を開発しています。
Curating instruction datasets and training recipes that keep open-weight models aligned and deployable.オープンウェイトモデルが指示に沿って動作し、実環境でも活用できるよう、指示データセットの整備と学習手法の開発に取り組んでいます。
Teaching models to produce and consume structured outputs such as captions, diagrams, and graphs for real-world tasks.画像説明文・図・グラフなどの構造を持つ情報を生成・理解するモデルを通じて、実世界の課題への応用を目指しています。
Jan 2026 – Present2026年1月 – 現在
Applied Research Engineer (Internship)応用研究エンジニア(インターン)
Jun 2024 – Present2024年6月 – 現在
Research Assistantリサーチアシスタント
Jul 2024 – Mar 20252024年7月 – 2025年3月
Engineering Internshipエンジニアリング・インターン
Jul 2022 – Aug 20242022年7月 – 2024年8月
Research Part-timer研究パートタイマー
Vision and Language research at the Language Information Access Technology Team.言語情報アクセス技術チームで視覚と言語の研究に従事。
Jun 2023 – Jan 20242023年6月 – 2024年1月
Research Internship研究インターン
Dec 2021 – Dec 20222021年12月 – 2022年12月
Research Assistantリサーチアシスタント
Aug 2022 – Sep 20222022年8月 – 2022年9月
Summer Internshipサマーインターン
Parameter-efficient pre-training methods for CLIP in Vision and Language tasks.視覚言語タスクに向けた、CLIP のパラメータ効率のよい事前学習手法を研究。
Feb 2021 – Mar 20222021年2月 – 2022年3月
Strategic AI GroupStrategic AI Group
SNS data analysis and annotation guideline creation.SNS データの分析とアノテーションガイドラインの作成。
2024 – Present2024年 – 現在
Vision and Language, Evaluation視覚と言語・評価手法
Focusing on novel evaluation metrics for multimodal systems and cross-modal representation learning.マルチモーダルシステムの新たな評価指標と、モダリティを横断する表現学習を研究。
Advisors:指導教員: Naoaki Okazaki岡崎 直観
2022 – 20242022年 – 2024年
Vision and Language: Image Captioning視覚と言語:画像説明生成
Developed context-aware image captioning models that generate descriptions based on user preferences.ユーザーの好みを考慮し、文脈に応じた説明文を生成する画像説明モデルを開発。
Advisors:指導教員: Naoaki Okazaki岡崎 直観
2018 – 20222018年 – 2022年
NLP: Grammatical Error Correction自然言語処理:文法誤り訂正
Created improved evaluation metrics for grammatical error correction systems.文法誤り訂正システムの評価指標の改善に取り組む。
Advisors:指導教員: Naoaki Okazaki, Masahiro Kaneko (Mentor)岡崎 直観、金子 正弘(メンター)
2025-11-26
Japanese Symposium on Open Large Language Models · Tokyo, Japan · Poster Presentation, Oral SessionJapanese Symposium on Open Large Language Models · 東京 · ポスター発表・口頭発表
Python · PyTorch · JavaPython・PyTorch・Java
Computer Vision · Natural Language Processing · Multimodal Learning · Image Captioning · Evaluation Metricsコンピュータビジョン・自然言語処理・マルチモーダル学習・画像説明生成・評価指標
Japanese (Native) · English (Professional)日本語(母語)・英語(業務・研究)