课题基金 / 基金详情

Development of learning subspace-based methods for pattern recognition

Development of learning subspace-based methods for pattern recognition
基于学习子空间的模式识别方法的开发
批准号:
22K17960
负责人:
SALESDESOUZA LINCON
金额:
$3.0万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Early-Career Scientists
财政年份:
2022
资助国家:
日本
项目状态:
未结题
起止时间:
2022-04-01 至 2026-03-31

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
在2022财年,我们致力于解决传统深度神经网络框架独立处理图像集的问题,而不考虑集中图像的底层特征分布和方差。为了克服这一限制,我们设计了一种新的子空间学习方法,称为格拉斯曼学习互子空间方法(G-LMSM),这是一种可以集成到深度神经网络中的神经网络层。G-LMSM将图像集映射到低维输入子空间表示中,然后使用正则角度的相似性度量(可解释且计算效率高的度量)将其与字典子空间匹配。G-LMSM的关键思想是将字典子空间作为Grassmann流形上的点来学习,Grassmann流形是一种光滑的非线性流形,可以捕获子空间的几何结构。这种学习是用黎曼随机梯度下降优化的,稳定,有效,理论上有很好的基础。该方法在三个不同的任务上进行了评估:手部形状识别、面部识别和面部情绪识别。我们的实验结果表明,G-LMSM在所有三个任务上都优于最先进的方法,这表明它有潜力提高图像集对象识别的深度框架的性能。
英文摘要
In fiscal year 2022, we worked to address the problem that traditional deep neural network frameworks process image sets independently, without considering the underlying feature distribution and the variance of the images in the set. To overcome this limitation, we devised a new subspace learning method called Grassmannian learning mutual subspace method (G-LMSM), which is an NN layer that can be integrated into deep neural networks.G-LMSM maps the image set into a low-dimensional input subspace representation, which is then matched with dictionary subspaces using a similarity metric of their canonical angles, an interpretable and computationally efficient metric. The key idea of G-LMSM is to learn dictionary subspaces as points on the Grassmann manifold, which is a smooth, non-linear manifold that captures the geometric structure of subspaces. This learning is optimized with Riemannian stochastic gradient descent, which is stable, efficient, and theoretically well-grounded.The proposed method was evaluated on three different tasks: hand shape recognition, face identification, and facial emotion recognition. Our experimental results showed that G-LMSM outperformed state-of-the-art methods on all three tasks, demonstrating its potential to improve the performance of deep frameworks for object recognition from image sets.
期刊论文(4)
专著(0)
科研奖励(0)
会议论文
DOI: 10.1109/cvprw56347.2022.00534
发表时间: 2022-06
期刊: 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)
影响因子: --
作者: [Bojan Batalo;L. S. Souza;B. Gatto;Naoya Sogi;K. Fukui]
通讯作者: Bojan Batalo;L. S. Souza;B. Gatto;Naoya Sogi;K. Fukui
DOI: 10.1016/j.neucom.2022.10.040
发表时间: 2021-11
期刊: Neurocomputing
影响因子: 6
作者: [L. S. Souza;Naoya Sogi;B. Gatto;Takumi Kobayashi;K. Fukui]
通讯作者: L. S. Souza;Naoya Sogi;B. Gatto;Takumi Kobayashi;K. Fukui
DOI: 10.1016/j.mlwa.2022.100407
发表时间: 2022-09
期刊: Machine Learning with Applications
影响因子: --
作者: [Bojan Batalo;L. S. Souza;B. Gatto;Naoya Sogi;K. Fukui]
通讯作者: Bojan Batalo;L. S. Souza;B. Gatto;Naoya Sogi;K. Fukui
Environmental sound classification based on CNN latent subspaces
基于CNN潜在子空间的环境声音分类
DOI: --
发表时间: 2022
期刊:
影响因子: --
作者: [Maha Mahyub, Lincon S. Souza, Bojan Batalo, Kazuhiro Fukui]
通讯作者: Kazuhiro Fukui
海外基金