Search results for "Biological data"

showing 10 items of 53 documents

The Protein Structure Context of PolyQ Regions.

2016

Proteins containing glutamine repeats (polyQ) are known to be structurally unstable. Abnormal expansion of polyQ in some proteins exceeding a certain threshold leads to neurodegenerative disease, a symptom of which are protein aggregates. This has led to extensive research of the structure of polyQ stretches. However, the accumulation of contradictory results suggests that protein context might be of importance. Here we aimed to evaluate the structural context of polyQ regions in proteins by analysing the secondary structure of polyQ proteins and their homologs. The results revealed that the secondary structure in polyQ vicinity is predominantly random coil or helix. Importantly, the region…

Models MolecularProtein Conformation alpha-HelicalProtein Structure ComparisonProtein StructureSaccharomyces cerevisiae ProteinsGlutaminelcsh:MedicineNerve Tissue ProteinsSaccharomyces cerevisiaePlant ScienceResearch and Analysis MethodsBiochemistryPlant Roots570 Life sciencesDatabase and Informatics MethodsProtein Structure DatabasesMacromolecular Structure AnalysisHumansProtein Interaction Domains and MotifsAmino AcidsDatabases ProteinProtein Interactionslcsh:ScienceMolecular BiologyMediator ComplexOrganic CompoundsPlant AnatomyAcidic Amino AcidsOrganic Chemistrylcsh:RChemical CompoundsBiology and Life SciencesProteinsRoot StructureChemistryBiological DatabasesProtein-Protein InteractionsPhysical Scienceslcsh:QStructural ProteinsProtein Structure DeterminationPeptidesResearch Article570 BiowissenschaftenPLoS ONE
researchProduct

A novel approach to investigate the evolution of structured tandem repeat protein families by exon duplication.

2020

Tandem Repeat Proteins (TRPs) are ubiquitous in cells and are enriched in eukaryotes. They contributed to the evolution of organism complexity, specializing for functions that require quick adaptability such as immunity-related functions. To investigate the hypothesis of repeat protein evolution through exon duplication and rearrangement, we designed a tool to analyze the relationships between exon/intron patterns and structural symmetries. The tool allows comparison of the structure fragments as defined by exon/intron boundaries from Ensembl against the structural element repetitions from RepeatsDB. The all-against-all pairwise structural alignment between fragments and comparison of the t…

Protein familyStructural alignmentBiological data visualizationExonComputational biologyBiologyEvolution Molecular03 medical and health sciencesExonProtein structureTandem repeatStructural BiologyGene duplicationAnimalsHumans030304 developmental biology0303 health sciences030302 biochemistry & molecular biologyIntronProteinsExonsProtein superfamilyClassificationIntronsBiological data visualization; Classification; Exon; Protein evolution; Protein structure; Repeat proteinTandem Repeat SequencesRepeat proteinProtein structureProtein evolutionJournal of structural biology
researchProduct

Missing value imputation in proximity extension assay-based targeted proteomics data

2020

Targeted proteomics utilizing antibody-based proximity extension assays provides sensitive and highly specific quantifications of plasma protein levels. Multivariate analysis of this data is hampered by frequent missing values (random or left censored), calling for imputation approaches. While appropriate missing-value imputation methods exist, benchmarks of their performance in targeted proteomics data are lacking. Here, we assessed the performance of two methods for imputation of values missing completely at random, the previously top-benchmarked ‘missForest’ and the recently published ‘GSimp’ method. Evaluation was accomplished by comparing imputed with remeasured relative concentrations…

ProteomicsMaleMultivariate analysisProtein ExpressionBiochemistryProtein expressionDatabase and Informatics MethodsLimit of DetectionStatisticsMedicine and Health SciencesBiochemical SimulationsImputation (statistics)Immune ResponseMathematicsMultidisciplinaryProteomic DatabasesQREukaryotaBlood ProteinsVenous ThromboembolismPlantsMiddle AgedLegumesTargeted proteomicssymbolsEngineering and TechnologyMedicineFemaleAlgorithmsResearch ArticleQuality ControlAdultScienceImmunologyResearch and Analysis Methodssymbols.namesakeSigns and SymptomsBiasIndustrial EngineeringProtein Concentration AssaysGene Expression and Vector TechniquesMissing value imputationHumansMolecular Biology TechniquesMolecular BiologyAgedInflammationMolecular Biology Assays and Analysis TechniquesInterleukin-6OrganismsPeasBiology and Life SciencesComputational BiologyMissing dataPearson product-moment correlation coefficientBiological DatabasesMultivariate AnalysisClinical MedicineVenous thromboembolismPLOS ONE
researchProduct

CiliaCarta: An integrated and validated compendium of ciliary genes

2019

The cilium is an essential organelle at the surface of mammalian cells whose dysfunction causes a wide range of genetic diseases collectively called ciliopathies. The current rate at which new ciliopathy genes are identified suggests that many ciliary components remain undiscovered. We generated and rigorously analyzed genomic, proteomic, transcriptomic and evolutionary data and systematically integrated these using Bayesian statistics into a predictive score for ciliary function. This resulted in 285 candidate ciliary genes. We generated independent experimental evidence of ciliary associations for 24 out of 36 analyzed candidate proteins using multiple cell and animal model systems (mouse…

ProteomicsSensory ReceptorsNematodaSocial SciencesCiliopathiesBiochemistrySensory disorders Donders Center for Medical Neuroscience [Radboudumc 12]Transcriptome0302 clinical medicineAnimal CellsPsychologyRETINAL PHOTORECEPTOR CELLSExomeNeurons0303 health sciences030302 biochemistry & molecular biologyEukaryotaGenomicsPRIMARY CILIUMthecilium3. Good healthNucleic acidsGenetic interferenceOsteichthyesMedicineEpigeneticsCellular Structures and OrganellesCellular Typesproteomic databasesSensory Receptor CellsScienceeducationCiliary genesLEBER CONGENITAL AMAUROSISGenomics03 medical and health sciencesGeneticsCiliaCaenorhabditis elegansIDENTIFICATIONMUTATIONSEmbryosciliaOrganismsBiology and Life SciencesBayes TheoremMolecular Sequence Annotationmedicine.diseaseInvertebratesFishciliary proteomeAnimal StudiesCaenorhabditisGene expressionembryos030217 neurology & neurosurgeryDevelopmental BiologyNeurosciencePhotoreceptorsCandidate geneEmbryologyOligonucleotidesMorpholinoDatabase and Informatics MethodsRNA interferenceBayesian classifierTRANSITION ZONEZebrafishAntisense OligonucleotidesZebrafishGeneticsMultidisciplinarySpectrometric Identification of ProteinsProteomic DatabasesNucleotidesCiliumQStable Isotope Labeling by Amino Acids in Cell CultureRphotoreceptorsMetabolic Disorders Radboud Institute for Molecular Life Sciences [Radboudumc 6]Animal ModelsPhenotypeINTRAFLAGELLAR TRANSPORTDIFFERENTIATIONPhenotypeExperimental Organism SystemsCaenorhabditis ElegansVertebratesSensory PerceptionResearch ArticleSignal TransductionEXPRESSIONStable isotope labeling by amino acids in cell cultureComputational biologyBiologyResearch and Analysis MethodsSOLUTE-CARRIER-PROTEINModel OrganismsmedicineAnimalsdata integration030304 developmental biologyAfferent NeuronsReproducibility of ResultsCell Biologyzebrafishbiology.organism_classificationCiliopathyRenal disorders Radboud Institute for Molecular Life Sciences [Radboudumc 11]Biological DatabasesCellular NeuroscienceRNAOSCP1CiliaCartaPLoS ONE
researchProduct

A New Linear Initialization in SOM for Biomolecular Data

2009

In the past decade, the amount of data in biological field has become larger and larger; Bio-techniques for analysis of biological data have been developed and new tools have been introduced. Several computational methods are based on unsupervised neural network algorithms that are widely used for multiple purposes including clustering and visualization, i.e. the Self Organizing Maps (SOM). Unfortunately, even though this method is unsupervised, the performances in terms of quality of result and learning speed are strongly dependent from the neuron weights initialization. In this paper we present a new initialization technique based on a totally connected undirected graph, that report relat…

Self-organizing mapBiological dataArtificial neural networkbusiness.industryComputer scienceUnsupervised learningInitializationPattern recognitionArtificial intelligencebusinessCluster analysisField (computer science)Visualization
researchProduct

The BioDICE Taverna plugin for clustering and visualization of biological data: a workflow for molecular compounds exploration

2014

Background: In many experimental pipelines, clustering of multidimensional biological datasets is used to detect hidden structures in unlabelled input data. Taverna is a popular workflow management system that is used to design and execute scientific workflows and aid in silico experimentation. The availability of fast unsupervised methods for clustering and visualization in the Taverna platform is important to support a data-driven scientific discovery in complex and explorative bioinformatics applications. Results: This work presents a Taverna plugin, the Biological Data Interactive Clustering Explorer (BioDICE), that performs clustering of high-dimensional biological data and provides a …

Self-organizing mapBiological dataMolecular compoundComputer scienceLibrary and Information Sciencescomputer.software_genreComputer Graphics and Computer-Aided DesignClusteringVisualizationComputer Science ApplicationsTavernaWorkflowMolecular compoundsSelf organizing mapKnowledge extractionPlug-inData miningPhysical and Theoretical ChemistryCluster analysiscomputerSoftwareWorkflow management systemVisualizationJournal of Cheminformatics
researchProduct

Classification of Sequences with Deep Artificial Neural Networks: Representation and Architectural Issues

2021

DNA sequences are the basic data type that is processed to perform a generic study of biological data analysis. One key component of the biological analysis is represented by sequence classification, a methodology that is widely used to analyze sequential data of different nature. However, its application to DNA sequences requires a proper representation of such sequences, which is still an open research problem. Machine Learning (ML) methodologies have given a fundamental contribution to the solution of the problem. Among them, recently, also Deep Neural Network (DNN) models have shown strongly encouraging results. In this chapter, we deal with specific classification problems related to t…

SequenceBiological dataSequence classificationSettore INF/01 - InformaticaArtificial neural networkProcess (engineering)Computer sciencebusiness.industryDeep learningBacteria classificationSequence classificationBacteria classificationNucleosome identificationDeep neural networkMachine learningcomputer.software_genreData typeNucleosome identificationComponent (UML)Artificial intelligenceMetagenomicsRepresentation (mathematics)businesscomputer
researchProduct

Novel Combinatorial and Information-Theoretic Alignment-Free Distances for Biological Data Mining

2010

Among the plethora of alignment-free methods for comparing biological sequences, there are some that we have perceived as representative of the novel techniques that have been devised in the past few years and as being of a fundamental nature and of broad interest and applicability, ranging from combinatorics to information theory. In this chapter, we review these alignment free methods, by presenting both their mathematical definitions and the experiments in which they are involved in.

Settore INF/01 - InformaticaComputer scienceAlignment-free distances for biological sequenceBiological data miningData miningcomputer.software_genrecomputer
researchProduct

Pathway analysis of high-throughput biological data within a Bayesian network framework

2011

Abstract Motivation: Most current approaches to high-throughput biological data (HTBD) analysis either perform individual gene/protein analysis or, gene/protein set enrichment analysis for a list of biologically relevant molecules. Bayesian Networks (BNs) capture linear and non-linear interactions, handle stochastic events accounting for noise, and focus on local interactions, which can be related to causal inference. Here, we describe for the first time an algorithm that models biological pathways as BNs and identifies pathways that best explain given HTBD by scoring fitness of each network. Results: Proposed method takes into account the connectivity and relatedness between nodes of the p…

Statistics and ProbabilityComputer scienceHigh-throughput screeningGene regulatory networkcomputer.software_genreModels BiologicalBiochemistrySynthetic dataBiological pathwayBayes' theoremHumansGene Regulatory NetworksCarcinoma Renal CellMolecular BiologyGeneBiological dataMicroarray analysis techniquesGene Expression ProfilingBayesian networkRobustness (evolution)Bayes TheoremPathway analysisKidney NeoplasmsHigh-Throughput Screening AssaysComputer Science ApplicationsGene expression profilingComputational MathematicsComputational Theory and MathematicsCausal inferenceData miningcomputerAlgorithmsSoftwareBioinformatics
researchProduct

Collective behaviours: from biochemical kinetics to electronic circuits

2013

In this work we aim to highlight a close analogy between cooperative behaviors in chemical kinetics and cybernetics; this is realized by using a common language for their description, that is mean-field statistical mechanics. First, we perform a one-to-one mapping between paradigmatic behaviors in chemical kinetics (i.e., non-cooperative, cooperative, ultra-sensitive, anti-cooperative) and in mean-field statistical mechanics (i.e., paramagnetic, high and low temperature ferromagnetic, anti-ferromagnetic). Interestingly, the statistical mechanics approach allows a unified, broad theory for all scenarios and, in particular, Michaelis-Menten, Hill and Adair equations are consistently recovered…

Work (thermodynamics)Biological dataMultidisciplinaryStatistical Mechanics (cond-mat.stat-mech)business.industryComputer scienceKineticsFOS: Physical sciencesAnalogyStatistical mechanicsModels TheoreticalArticleChemical kineticsHumans; Algorithms; Models Theoretical; MultidisciplinaryHumansCyberneticsArtificial intelligenceStatistical physicsElectronicsbusinessAlgorithmsCondensed Matter - Statistical MechanicsElectronic circuitScientific Reports
researchProduct