ThePlanteomeProjectLaurelCooper,AustinMeier,JustinL.
Elser,JustinPreece,XuXu,RyanS.
Kitchen,BotongQu,EugeneZhang,SinisaTodorovic,PankajJaiswalOregonStateUniversity,Corvallis,OR,USAMarie-AngéliqueLaporte,ElizabethArnaudBioversityInternational,Montpellier,FranceSethCarbon,ChrisMungallLawrenceBerkeleyNationalLaboratory,Berkeley,CA,USABarrySmithUniversityatBuffalo,Buffalo,NY,USAGeorgiosGkoutosUniversityofBirmingham,UKandUniversityofAberystwyth,UKJohnDoonanUniversityofAberystwyth,UKAbstract—ThePlanteomeprojectisacentralizedonlineplantinformaticsportalwhichprovidessemanticintegrationofwidelydiversedatasetswiththegoalofplantimprovement.
Traditionalplantbreedingmethodsforcropimprovementmaybecombinedwithnext-generationanalysismethodsandautomatedscoringoftraitsandphenotypestodevelopimprovedvarieties.
ThePlanteomeproject(www.
planteome.
org)developsandhostsasuiteofreferenceontologiesforplantsassociatedwithagrowingcorpusofgenomicsdata.
Dataannotationslinkingphenotypesandgermplasmtogenomicsresourcesareachievedbydatatransformationandmappingspecies-specificcontrolledvocabulariestothereferenceontologies.
Analysisandannotationtoolsarebeingdevelopedtofacilitatestudiesofplanttraits,phenotypes,diseases,genefunctionandexpressionandgeneticdiversitydataacrossawiderangeofplantspecies.
TheprojectdatabaseandtheonlineresourcesprovideresearcherstoolstosearchandbrowseandaccessremotelyviaAPIsforsemanticintegrationinannotationtoolsanddatarepositoriesprovidingresourcesforplantbiology,breeding,genomicsandgenetics.
Keywords—ontology;traitsphenotype;semantic;dataintegration,plantsI.
INTRODUCTIONA.
RationaleItisestimatedthattheworldpopulationisprojectedtoreach9.
6billionpeopleinnextfewdecades(http://www.
wri.
org/blog/2013/12/global-food-challenge-explained-18-graphics).
Therefore,thechallengeishowtofeedthisgrowingpopulation,whileprotectingtheearth'senvironment.
Traditionalplantbreedingmethodsforplantimprovementmaybecombinedwithnext-generationanalysismethods,includingthehigh-throughputandautomatedscoringoftraitsandphenotypestodevelopimprovedvarieties.
Datafromhigh-throughputsequencing,transcriptomic,proteomic,phenomicandgenomeannotationprojectscanbelinkedtogermplasmresourcesthroughtheuseofinteroperable,referencevocabularies(ontologies).
Inthisway,theknowledgegainedfromthenext-generationdatacanbeutilizedforcropimprovement.
B.
WhatisthePlanteomeThePlanteomeProject(www.
planteome.
org)isacentralizedonlineinformaticsportalanddatabase,consistingofasuiteofreferenceontologiesforplants,anassociatedcorpusofplantgenomicsandphenomicsdata,andtoolsfordataanalysisandannotation.
Analysesofthesedatasetsfromgeneticandgenomicstudieshavethepotentialtoimproveourunderstandingofthemolecularbasisofeconomicallyrelevanttraits.
Inordertoutilizethisdata,researchersmustbeabletoconnecttherelevantplanttraitsofinteresttothespatialandtemporalexpressionpatternsofgenes,andelucidatetheirrolesinbiologicalprocessesinplants.
C.
GoalsofthePlanteomeProject:1.
Asuiteofinterrelatedreferenceontologiestodescribemajorknowledgedomainsofplantbiology,comprisingplantphenotypeandtraits,environments,andbioticandabioticstresses.
2.
Standards,workflowsandtoolsforannotationofplantgenomicsdata,andmetadataforcurationandimprovedannotationofgenes,genomes,phenotypeandgermplasm.
3.
ThePlanteomebrowseranddatabase,acentralized,onlineinformaticsportalandrepositorywherereferenceontologiesforplantsareusedtoaccessdataresourcesforplanttraits,phenotypes,diseases,geneexpressionandgeneticdiversitydataacrossawiderangeofplantspecies.
4.
OutreachinvolvingtheplantresearchcommunityandK-12andundergraduatestudents.
II.
THESCOPEOFTHEPLANTEOMEThescopeoftheontologiesinthePlanteomeprojectrangesfromabroadoverviewofplantenvironmentsandtaxonomy,tothecellularandmolecularlevelofexpressedgenesandtheirbiologicalfunctions.
ThePlanteomeontologies,describedinmoredetailbelow,consistofthePlantOntology(PO)[1-6],PlantTraitOntology(TO)[7,8],thePlantEnvironmentOntology(EO)[7]andthePlantStressOntology(PSO).
ThePlanteomeprojectimportsandintegrateswithrelevantreferenceontologiesdevelopedbycollaboratinggroups;theGeneOntology(GO)[9,10],thePhenotypicQualitiesOntology(PATO)[11],theEnvironmentOntology(ENVO)[12],andtheChemicalEntitiesofBiologicalInterest(ChEBI)[13].
Inaddition,thePlanteomeintegratesandmapsspecies-orclade-specificapplicationontologiesdevelopedbytheCropOntology(CO)project[14].
Togetherthissuiteofreferenceontologiescanbeusedtofullyannotateandlinktogetherthevitalplantknowledgedomain.
Thecentralreferenceontologyforplantanatomyandplantdevelopmentalstages,thePlantOntology(PO)[1-6]grewoutoftheneedtocreateassociationsbetweenstandardizedterminologyforplantsandgenomicsdata,andwasbasedtheworkdonetodeveloptheGeneOntologyinthelate1990s[9,10].
ThePOisrecognizedworldwideasthereferenceontologyforplantstructuresanddevelopmentalstages,andislinkedtodatafromawidevarietyofplants,fromtraditionalmodelspeciestothecropplantsthatfeedtheworld'sgrowingpopulation.
Plantimprovementreliesonanalysesofplanttraitsandphenotypes.
Forthesepurposes,thePlantTraitOntology(TO)[9,10]describesawiderangeofprecomposedplanttraitsconsistentwithEntity(E)-Quality(Q)statementsandleadstoanunderstandingofthemolecularprocessesthatunderliethem.
Eachtraitisameasurableorobservablecharacteristicofaplantstructure(PO:000901),aplantcellularcomponent(GO:0005575),oraplantstructuredevelopmentstage(PO:0009012),aswellasplantbiologicalprocesses(GO:0008150)andmolecularfunctions(GO:0003674).
TheTOencompassesninebroad,upper-levelcategoriesofplanttraits:biochemicaltrait(TO:0000277),biologicalprocesstrait(TO:0000283),plantgrowthanddevelopmenttrait(TO:0000357),plantmorphologytrait(TO:0000017),qualitytrait(TO:0000597),statureorvigortrait(TO:0000133),sterilityorfertilitytrait(TO:0000392),stresstrait(TO:0000164)andyieldtrait(TO:0000371).
ThePlantEnvironmentOntology(EO)isusedtodescribetheplantgrowthconditionsandstudytypesandcanbecombinedwiththetermsfromtheotherreferenceontologiestofullyannotateaplantphenotypedescription.
Inadditiontothereferenceontologies,thePlanteomeworkscloselywithdevelopersofthespecies-specificvocabulariessuchastheCropOntology[14]tointegratetheirterms,createmappingstothereferenceontologiesandlinkphenotypesandgermplasmtogenomicsresources.
III.
DEVELOPMENTOFTHEPLANTEOMEONTOLOGYNETWORKThedevelopmentofthePlanteomeProjectontologynetworkisafundamentalchangeinthewayofthinkingaboutontologiesforplants.
Inthepreviousproject,thePlantOntology(http://www.
plantontology.
org/),asinglereferenceontologywasdevelopedandusedtoannotateplantgenomicdatatoontologytermsdescribingplantstructuresandplantdevelopmentalstages.
Theadditionoftheotherreferenceandspecies-specifcontologiesforplantsenrichestheannotationenvironmentsoamorecompletepictureofthemetadataofplantpheotypescanbeexpressed.
Inordertocreatethenetwork,ontologytermsintheTOandthespecies-specifccroptraitontologieshavebeen'decomposed'intothecorrespondingEntity(E)-Quality(Q)statementswhichutilizetermsfromtheotherreferenceontologies,suchasPOandGOfortheentitiesandPATOforthequalities.
Inthisway,anetworkisformedwhichlinksallthevariousontologiestogether.
Oneofthelessonslearnedindevelopingthisnetworkisthatsomeofthereferenceontologiesandvocabulariesdevelopedbyourcollaborators(suchasChEBI,andtheNCBITaxonomy)aresolargethattheyarecumbersometodisplayonourbrowser.
Forthese,wehavedevelopedscripttoextractarelevant"slim"versionwhichcontainstheneededterms.
IV.
PLANTEOMEANNOTATIONDATABASEThePlanteomedatabaseprovidesontologytermsanddefinitionsalongwiththeassociated'annotations'[15],betweentheontologytermsanddatasourcedfromnumerousplantgenomicsdatasets.
ThePlanteome1.
0BetaRelease(Nov.
2015)containsabout47millionannotationslinkingreferenceontologytermstodataobjectsrepresentinggenes,genemodels,proteins,RNAs,germplasmandquantitativetraitloci(QTLs)from87differentplantspecies.
Thesedataarecurrentlycontributedby29differentdatasources.
Planteomecuratorsandresearchersatvariouscollaboratingdatabasegroupsworkcloselytodeveloptheannotationfilesinthestandardizeddataformatdatabase.
Thedatabaseisaccessibleonline(http://planteome.
org/)andalsoavailableforbulkdownload(http://palea.
cgrb.
oregonstate.
edu/viewsvn/associations/).
TheannotationdatabaseincludesfunctionalGeneOntologyannotationsfor60species.
Thesepredictionsweredoneusingtwomethods.
ThefirstmethodutilizedanInterProScan[16]toidentifyproteindomains.
TheresultinganalysisfileswerethenparsedtoassociatetheproteindomainstoGOterms.
ThesecondmethodwastoprojectontologyannotationsbasedonFig.
1.
AnnotationofRicebrd1mutantwithreferenceontologytermstocapturethephenotype.
Thericeplantimageisadaptedwithpermissionfrom[19]JohnWileyandSons.
orthologytoArabidopsisthalianagenes.
OrthologywaspredictedwithInParanoid[17],aprogramthattakesreciprocalBLASToutputandusespairwisesimilarityscorestodetermineorthologousclustersofgenes.
Thisisfollowedbycreatinggenesuperclustersbypoolingspecies-pairclusterswithcommongenes.
Theorthologoussuperclustersofthe60specieswerecomparedwiththeknownannotationfilesforArabidopsisthalianaforGO,andnewannotationfilesweregenerated.
PlanteomeistheonlyonlinesourceprovidingGOfunctionalannotationofgenesidentifiedformanyofthesespecies.
V.
CASESTUDYEXAMPLE:PHENOTYPEANNOTATIONOFRICEBRASSINOSTEROID(BR)-DEFICIENTDWARFMUTANTBrassinosteroid(BR)-deficient(brd1)dwarfmutantsofricewerecharacterizedtodeterminetherolesthatBRsplayinnormalplantgrowthanddevelopmentinamonocotplant[19].
Fig.
1showsanexampleofhowthereferenceontologiescanbeusedtoannotatethephenotypeofa(BR)-Deficientdwarfmutantrice,brd1-1.
ThisimageisacompliationofontologytermsfromvariousPlanteomereferenceontologiesthathavebeenusedtoannotatetheexpressionofbrd1(Os03g0602300)inthePlanteomedatabase.
Theseannotationswerecontributedfromavarietyofsources,suchasGramene(http://www.
gramene.
org/),EnsemblPlants(http://plants.
ensembl.
org/index.
html),andTheRiceAnnotationProject(RAP)(http://rapdb.
dna.
affrc.
go.
jp/)andcanbeusedtodescribeallaspectsofthebrd1mutantphenotype.
GatheringtheannotationstogetherinaunifiedplatformsuchasthePlanteomeallowsthedatatobemadeaccessibleandfacilitatesgenediscoverythroughinter-andintra-speciescomparisons.
VI.
PLANTEOMETOOLSFORCOLLABORATIONANDONTOLOGYINTEGRATIONThePlanteomeprojectisdevelopinganumberoftoolstoincreaseaccesstotheontologytermsandtoincreasetheinteroperabilityoftheannotateddata.
AllthePlanteomeontologiesarepublicallyavailableandaremaintainedatthePlanteomeGitHubsite(https://github.
com/Planteome)forsharingandtrackingrevsions.
Thissitefacilitatescommunityfeedback;userscanmakecomments,requesttermsandsuggestchangestothePlanteomeontologies.
Inaddition,thePlanteomeGitHubsitealsofeaturesspecies-specificvocabulariessuchasthosefromCropOntology(http://www.
cropontology.
org/).
AnothernewtoolwhichisunderdevelopmentisaTraitOntology-specific(http://to.
termgenie.
org/)instanceoftheTermGenietool[20].
TermGenieusesapattern-basedapproachtorapidlygeneratenewtermsandplacethemappropriatelywithintheontologystructure.
AlltermsarereviewedbyaPlanteomecuratorbeforethefinalcommittotheontology.
TermGeniecanbeusedtoquicklyobtainaTOtermforannotation,ifanappriopriateonedoesnotalreadyexist.
Planteomeisdevelopinganapplicationprogramminginterface(API)thatwillallowcollaboratorstoaccessandusethehosteddataintheirwebsitesandapplications.
ThefirsttwoAPImethods–currentlyaccessiblefromthePlanteomedevelopmentenvironment–queryPlanteome-hostedontologiesforterms,termdefinitions,andotherattributes,returningtheminJSONformat.
The"search"methodisfastenoughtobeusedinanautocompletesearchbox.
AllthePlanteomereferenceandspecies-specificontologiesareavailablethroughtheAPIservice.
Currently,theAPIonlyservestheterminformation,butthePlanteomeprojectplanstoaddAPImethodstoaccessannotationdata,aswell.
ThePlanteomeprojectiscollaboratingwiththeBisqueImageAnalysisEnvironment(CenterforBio-ImageInformatics,UCSB;http://www.
cyverse.
org/bisque)onintegratedimagesegmentationandontologyannotationfeatures.
ThePlanteomeprojectalreadyhostssuchatoolasadesktopapplication;AnnotationofImageSegmentswithOntologies(AISO;http://planteome.
org/node/3),butwewishtomoveitsfunctionalityonlineasamodulewithinBisque,takingadvantageofitssharedCyVerseauthentication,datastore,andcomputationinfrastructure.
Theontologydataitselfwillbeservedfromexternalservices,suchasthePlanteomeAPI.
VII.
CONCLUSIONSThePlanteomeprojectisacentralizedonlineplantinformaticsportalandwhichintegratesreferenceontologiesforplants,andspecies-specificcontrolledvocabularieswithalargeandgrowingcorpusofplantgenomicsdata.
Thisplatformprovidessemanticintegrationofwidelydiversedatasetswiththegoalofplantimprovement.
ACKNOWLEDGMENTFundingforthePlanteomeprojectisprovidedbytheNationalScienceFoundationawardIOS#1340112REFERENCES[1]Jaiswal,P,SAvraham,KIlic,EAKellogg,SMcCouch,APujar,etal.
,2005.
PlantOntology(PO):AControlledVocabularyofPlantStructuresandGrowthStages.
CompFunctGenomics,.
6(7--‐8):p.
388-97(references)[2]Pujar,A,PJaiswal,EAKellogg,KIlic,LVincent,SAvraham,etal.
2006.
Whole-‐plantgrowthstageontologyforangiospermsanditsapplicationinplantbiology.
PlantPhysiol,142(2):p.
414--‐28.
[3]Ilic,K,EAKellogg,PJaiswal,FZapata,PFStevens,LPVincent,etal.
,2007.
Theplantstructureontology,aunifiedvocabularyofanatomyandmorphologyofafloweringplant.
PlantPhysiol.
143(2):p.
587--‐599.
[4]Avraham,S,CWTung,KIlic,PJaiswal,EAKellogg,SMcCouch,etal.
,2008.
ThePlantOntologyDatabase:acommunityresourceforplantstructureanddevelopmentalstagescontrolledvocabularyandannotations.
NucleicAcidsRes.
,36(Databaseissue):p.
D449--‐54.
.
[5]CooperL,WallsRL,ElserJ,GandolfoMA,StevensonDW,SmithB,etal.
(2013)ThePlantOntologyasatoolforcomparativeplantanatomyandgenomicanalyses.
PlantandCellPhysiology54:e1–e1[6]CooperLandJaiswalP(2016)ThePlantOntology:AToolforPlantGenomics.
InDEdwards,ed,PlantBioinformatics.
SpringerNewYork,pp89–114[7]JaiswalP,WareD,NiJ,ChangK,ZhaoW,SchmidtS,etal.
(2002)Gramene:developmentandintegrationoftraitandgeneontologiesforrice.
ComparativeandFunctionalGenomics3:132–136.
[8]ArnaudE,CooperL,ShresthaR,MendaN,NelsonRT,MatteisL,etal.
(2012)TowardsareferencePlantTraitOntologyformodelingknowledgeofplanttraitsandphenotypes.
ProceedingsoftheInternationalConferenceonKnowledgeEngineeringandOntologyDevelopment.
Barcelona,Spain,pp220–225.
[9]AshburnerM,BallCA,BlakeJA,BotsteinD,ButlerH,CherryJM,etal.
(2000)GeneOntology:toolfortheunificationofbiology.
NatGenet25:25–29.
[10]TheGeneOntologyConsortium(2014)GeneOntologyConsortium:goingforward.
NucleicAcidsResearch.
doi:10.
1093/nar/gku1179.
[11]GkoutosG,GreenE,MallonA-M,HancockJ,DavidsonD(2004)Usingontologiestodescribemousephenotypes.
GenomeBiol6:R8[12]ButtigiegP,MorrisonN,SmithB,MungallC,LewisS(2013)Theenvironmentontology:contextualisingbiologicalandbiomedicalentities.
JournalofBiomedicalSemantics4:43[13]HastingsJ,OwenG,DekkerA,EnnisM,KaleN,MuthukrishnanV,etal.
(2016)ChEBIin2016:Improvedservicesandanexpandingcollectionofmetabolites.
NucleicAcidsResearch44:D1214–D1219[14]Shrestha,R,Davenport,GFBruskiewich,R,Arnaud,E.
(2011)Developmentofcropontologyforsharingcropphenotypicinformation.
Droughtphenotypingincrops:fromtheorytopractice.
pp171–179[15]HillDP,SmithB,McAndrews-HillMS,BlakeJ(2008)GeneOntologyannotations:whattheymeanandwheretheycomefrom.
BMCBioinformatics9:S2[16]QuevillonE,SilventoinenV,PillaiS,etal.
2005.
InterProScan:proteindomainsidentifier.
NucleicAcidsResearch.
33(WebServerissue):W116-W120.
doi:10.
1093/nar/gki442.
[17]RemmM,StormCEVandSonnhammerELL(2001).
AutomaticClusteringofOrthologsandIn-paralogsfromPairwiseSpeciesComparisons.
JMB,314:1041-1052.
[18]Altschul,SF,Madden,TL,Schffer,AA,Zhang,J,Zhang,Z,Miller,W,etal.
(1997).
GappedBLASTandPSI-BLAST:anewgenerationofproteindatabasesearchprograms.
NucleicAcidsRes.
25:3389-3402.
[19]Hong,Z,Ueguchi-Tanaka,M,Shimizu-Sato,S,Inukai,Y,Fujioka,S,Shimada,Y,etal(2002)Loss-of-functionofaricebrassinosteroidbiosyntheticenzyme,C-6oxidase,preventstheorganizedarrangementandpolarelongationofcellsintheleavesandstem.
ThePlantJournal32:495–508[20]Dietze,H,Berardini,T,Foulger,R,Hill,D,Lomax,J,OsumiSutherland,D,RoncagliaP,MungallC(2014)TermGenie-Awebapplicationforpattern-basedontologyclassgeneration.
JournalofBiomedicalSemantics5:48[21]LingutlaN,PreeceJ,TodorovicS,CooperL,MooreL,JaiswalP(2014)AISO:AnnotationofImageSegmentswithOntologies.
JournalofBiomedicalSemantics5:50
无忧云怎么样?无忧云值不值得购买?无忧云,无忧云是一家成立于2017年的老牌商家旗下的服务器销售品牌,现由深圳市云上无忧网络科技有限公司运营,是正规持证IDC/ISP/IRCS商家,主要销售国内、中国香港、国外服务器产品,线路有腾讯云国外线路、自营香港CN2线路等,都是中国大陆直连线路,非常适合免备案建站业务需求和各种负载较高的项目,同时国内服务器也有多个BGP以及高防节点。目前,四川雅安机房,4...
青果网络QG.NET定位为高效多云管理服务商,已拥有工信部颁发的全网云计算/CDN/IDC/ISP/IP-VPN等多项资质,是CNNIC/APNIC联盟的成员之一,2019年荣获国家高薪技术企业、福建省省级高新技术企业双项荣誉。那么青果网络作为国内主流的IDC厂商之一,那么其旗下美国洛杉矶CN2 GIA线路云服务器到底怎么样?官方网站:https://www.qg.net/CPU内存系统盘流量宽带...
欧路云 主要运行弹性云服务器,可自由定制配置,可选加拿大的480G超高防系列,也可以选择美国(200G高防)系列,也有速度直逼内地的香港CN2系列。所有配置都可以在下单的时候自行根据项目 需求来定制自由升级降级 (降级按天数配置费用 退款回预存款)。由专业人员提供一系列的技术支持!官方网站:https://www.oulucloud.com/云服务器(主机测评专属优惠)全场8折 优惠码:zhuji...
www.meansys为你推荐
vc组合金钟大奖VC组合的两个人分别叫什么?www.hao360.cn每次打开电脑桌面都出现以下图标,打开后链接指向www.hao.360.cn。怎么彻底删除?比肩工场比肩夺财,行官杀制比是什么意思?www.jjwxc.net晋江文学网 的网址是什么?陈嘉垣马德钟狼吻案事件是怎么回事曲妙玲张婉悠香艳版《白蛇传》是电影还是写真集?rawtoolsU盘显示是RAW格式怎么办月神谭适合12岁男孩的网名,要非主流的,帮吗找找,谢啦月神谭有没有什么好看的小说?拒绝言情小说!www.bbb336.comwww.zzfyx.com大家感觉这个网站咋样,给俺看看呀。多提意见哦。哈哈。
移动服务器租用 免费申请域名 中国万网域名 老左 flashfxp怎么用 主机评测 美国主机评论 BWH cloudstack tk域名 申请空间 免费个人空间申请 太原联通测速平台 e蜗 什么是刀片服务器 adroit cdn加速原理 太原网通测速平台 申请免费空间和域名 免费mysql数据库 更多