Aug 2026🗣️We presented 5 papers at YANS 2026!Code2Figure: 学術論文における実装コードに基づくモデル図生成の初期検討Where Do Vision-Language Models Preserve and Use Depth Information?Activation Steering における文章崩壊の抑制に向けた初期検討視覚言語モデルを用いた AI エージェントによる悪質サイトの検知と警告能力の評価視覚言語モデルと病徴属性を用いた植物病害診断手法の検討
Aug 2026🏆Our paper was selected as a Best Paper Finalist at MIPR 2026!Document-Grounded Coaching Agent for Video Skill Assessment
Aug 2026🎉Our paper has been accepted to MIPR 2026!Document-Grounded Coaching Agent for Video Skill Assessment
Aug 2026🗣️We presented 3 papers at MIRU 2026!Generation and Evaluation of Editable Graphical Abstracts for Academic PapersDA-LoRA: 植物病害診断におけるドメイン特徴の視覚的分析のための敵対的画像生成密集圃場群における超解像衛星画像特徴と空間補正を用いた病害率推定の初期検討
Aug 2026🌱I joined LENS as a committee member and helped launch it!Learning and Exchange Network for Vision Studies (LENS)
Jul 2026🎉Our paper has been accepted to ACL SRW 2026!A11y-Compressor: A Framework for Enhancing the Efficiency of GUI Agent Observations through Visual Context Reconstruction and Redundancy Reduction
Jun 2026🎉Our paper has been accepted to CVPRF 2026!SciGA: A Comprehensive Dataset for Designing Graphical Abstracts in Academic Papers
Mar 2026🗣️We presented 2 papers at NLP 2026!SciGA-Vec: 学術論文におけるベクタ画像形式の Graphical Abstract データセットCompressed-a11y: 視覚的文脈の再構成と冗長性削減による GUI エージェント観測の効率化
Sep 2025🏆Our paper received the ABEJA Sponsor Award at YANS 2025!GenGA: 学術論文における編集可能な Graphical Abstract の自動生成に関する初期検討
Sep 2025🗣️We presented 3 papers at YANS 2025!GenGA: 学術論文における編集可能な Graphical Abstract の自動生成に関する初期検討ローカル LLM を用いた AI エージェントの現状と課題WRIME-TC: 時間的文脈による書き手と読み手の感情分析の強化
Nov 2024🎤I gave a talk at the Computer Vision Paper Reading!Long-CLIP: Unlocking the Long-Text Capability of CLIP