检索规则说明:AND代表“并且”;OR代表“或者”;NOT代表“不包含”;(注意必须大写,运算符两边需空一格)
检 索 范 例 :范例一: (K=图书馆学 OR K=情报学) AND A=范并思 范例二:J=计算机应用与软件 AND (U=C++ OR U=Basic) NOT M=Visual
作 者:杨帆 韩巧玲[1,2,3,4] 赵文迪[1,2,3,4] 赵玥[1,2,3,4] Fan Yang;Qiaoling Han;Wendi Zhao;Yue Zhao(School of technology,Beijing Forestry University,Beijing 100083,China;Key Lab of State Forestry Administration for Forestry Equipment and Automation,Beijing 100083,China;Beijing Laboratory of Urban and Rural Ecological Environment,Beijing 100083,China;Research Center for Intelligent Forestry,Beijing Forestry University,Beijing 100083,China)
机构地区:[1]北京林业大学工学院,北京100083 [2]林业装备与自动化国家林业局重点实验室,北京100083 [3]城乡生态环境北京实验室,北京100083 [4]北京林业大学智慧林业研究中心,北京100083
出 处:《遗传》2024年第8期661-670,共10页Hereditas(Beijing)
基 金:国家自然科学基金面上项目(编号:32071838);国家自然科学基金青年科学基金项目(编号:32101590)资助。
摘 要:酶功能的识别对理解生命活动的机制、推进生命科学的发展有重要作用。然而现有的酶EC编号预测方法,并未充分利用蛋白质序列信息,在识别精度上仍有所不足。针对上述问题,本研究提出一种基于层级特征和全局特征的EC编号预测网络(EC number prediction network using hierarchical features and global features,ECPN-HFGF)。该方法首先通过残差网络提取蛋白质序列通用特征,并通过层级特征提取模块和全局特征提取模块进一步提取蛋白质序列的层级特征和全局特征,之后结合两种特征信息的预测结果,采用多任务学习框架,实现酶EC编号的精确预测。计算实验结果表明,ECPN-HFGF方法在蛋白质序列EC编号预测任务上性能最佳,宏观F1值和微观F1值分别达到95.5%和99.0%。ECPN-HFGF方法能有效结合蛋白质序列的层级特征和全局特征,快速准确预测蛋白质序列EC编号,比当前常用方法预测精确度更高,能够为酶学研究和酶工程应用的发展提供一种高效的思路和方法。The identification of enzyme functions plays a crucial role in understanding the mechanisms of biological activities and advancing the development of life sciences.However,existing enzyme EC number prediction methods did not fully utilize protein sequence information and still had shortcomings in identification accuracy.To address this issue,we proposed an EC number prediction network using hierarchical features and global features(ECPN-HFGF).This method first utilized residual networks to extract generic features from protein sequences,and then employed hierarchical feature extraction modules and global feature extraction modules to further extract hierarchical and global features of protein sequences.Subsequently,the prediction results of both feature types were combined,and a multitask learning framework was utilized to achieve accurate prediction of enzyme EC numbers.Experimental results indicated that the ECPN-HFGF method performed best in the task of predicting EC numbers for protein sequences,achieving macro F1 and micro F1 scores of 95.5%and 99.0%,respectively.The ECPN-HFGF method effectively combined hierarchical and global features of protein sequences,allowing for rapid and accurate EC number prediction.Compared to current commonly used methods,this method offers significantly higher prediction accuracy,providing an efficient approach for the advancement of enzymology research and enzyme engineering applications.
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在链接到云南高校图书馆文献保障联盟下载...
云南高校图书馆联盟文献共享服务平台 版权所有©
您的IP:216.73.216.49