ontology alignment. ontologies in biomedical research many biomedical ontologies e.g. go, obo,...
TRANSCRIPT
![Page 1: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/1.jpg)
Ontology Alignment
![Page 2: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/2.jpg)
Ontologies in biomedical research many biomedical ontologies
e.g. GO, OBO, SNOMED-CT
practical use of biomedical ontologiese.g. databases annotated with GO
GENE ONTOLOGY (GO)
immune response i- acute-phase response i- anaphylaxis i- antigen presentation i- antigen processing i- cellular defense response i- cytokine metabolism i- cytokine biosynthesis synonym cytokine production … p- regulation of cytokine biosynthesis … … i- B-cell activation i- B-cell differentiation i- B-cell proliferation i- cellular defense response … i- T-cell activation i- activation of natural killer cell activity …
![Page 3: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/3.jpg)
Ontologies with overlapping information
SIGNAL-ONTOLOGY (SigO)
Immune Response i- Allergic Response i- Antigen Processing and Presentation i- B Cell Activation i- B Cell Development i- Complement Signaling synonym complement activation i- Cytokine Response i- Immune Suppression i- Inflammation i- Intestinal Immunity i- Leukotriene Response i- Leukotriene Metabolism i- Natural Killer Cell Response i- T Cell Activation i- T Cell Development i- T Cell Selection in Thymus
GENE ONTOLOGY (GO)
immune response i- acute-phase response i- anaphylaxis i- antigen presentation i- antigen processing i- cellular defense response i- cytokine metabolism i- cytokine biosynthesis synonym cytokine production … p- regulation of cytokine biosynthesis … … i- B-cell activation i- B-cell differentiation i- B-cell proliferation i- cellular defense response … i- T-cell activation i- activation of natural killer cell activity …
![Page 4: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/4.jpg)
Ontologies with overlapping information
Use of multiple ontologies e.g. custom-specific ontology + standard ontology
Bottom-up creation of ontologiesexperts can focus on their domain of expertise
important to know the inter-ontology important to know the inter-ontology relationshipsrelationships
![Page 5: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/5.jpg)
SIGNAL-ONTOLOGY (SigO)
Immune Response i- Allergic Response i- Antigen Processing and Presentation i- B Cell Activation i- B Cell Development i- Complement Signaling synonym complement activation i- Cytokine Response i- Immune Suppression i- Inflammation i- Intestinal Immunity i- Leukotriene Response i- Leukotriene Metabolism i- Natural Killer Cell Response i- T Cell Activation i- T Cell Development i- T Cell Selection in Thymus
GENE ONTOLOGY (GO)
immune response i- acute-phase response i- anaphylaxis i- antigen presentation i- antigen processing i- cellular defense response i- cytokine metabolism i- cytokine biosynthesis synonym cytokine production … p- regulation of cytokine biosynthesis … … i- B-cell activation i- B-cell differentiation i- B-cell proliferation i- cellular defense response … i- T-cell activation i- activation of natural killer cell activity …
![Page 6: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/6.jpg)
Ontology Alignment
equivalent concepts
equivalent relations
is-a relation
SIGNAL-ONTOLOGY (SigO)
Immune Response i- Allergic Response i- Antigen Processing and Presentation i- B Cell Activation i- B Cell Development i- Complement Signaling synonym complement activation i- Cytokine Response i- Immune Suppression i- Inflammation i- Intestinal Immunity i- Leukotriene Response i- Leukotriene Metabolism i- Natural Killer Cell Response i- T Cell Activation i- T Cell Development i- T Cell Selection in Thymus
GENE ONTOLOGY (GO)
immune response i- acute-phase response i- anaphylaxis i- antigen presentation i- antigen processing i- cellular defense response i- cytokine metabolism i- cytokine biosynthesis synonym cytokine production … p- regulation of cytokine biosynthesis … … i- B-cell activation i- B-cell differentiation i- B-cell proliferation i- cellular defense response … i- T-cell activation i- activation of natural killer cell activity …
Defining the relations between the terms in different ontologies
![Page 7: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/7.jpg)
Many experimental systems: Prompt (Stanford SMI) Anchor-Prompt (Stanford SMI) Chimerae (Stanford KSL) Rondo (Stanford U./ULeipzig) MoA (ETRI) Cupid (Microsoft research) Glue (Uof Washington) FCA-merge (UKarlsruhe) IF-Map Artemis (UMilano) T-tree (INRIA Rhone-Alpes) S-MATCH (UTrento)
Coma (ULeipzig) Buster (UBremen) MULTIKAT (INRIA S.A.) ASCO (INRIA S.A.) OLA (INRIA R.A.) Dogma's Methodology ArtGen (Stanford U.) Alimo (ITI-CERTH) Bibster (UKarlruhe) QOM (UKarlsruhe) KILT (INRIA LORRAINE)
![Page 8: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/8.jpg)
According to input KR: OWL, UML, EER, XML, RDF, … components: concepts, relations, instance, axioms
According to process What information is used and how?
According to output 1-1, m-n Similarity vs explicit relations (equivalence, is-a) confidence
Classification
![Page 9: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/9.jpg)
Matchers
![Page 10: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/10.jpg)
Strategies based on linguistic matching Structure-based strategies Constraint-based approaches Instance-based strategies Use of auxiliary information
Matcher Strategies Strategies based on linguistic matchingStrategies based on linguistic matching
SigO: complement signaling synonym complement activation
GO: Complement Activation
![Page 11: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/11.jpg)
Example matchers
Edit distance Number of deletions, insertions, substitutions required to
transform one string into another aaaa baab: edit distance 2
N-gram N-gram : N consecutive characters in a string Similarity based on set comparison of n-grams aaaa : {aa, aa, aa}; baab : {ba, aa, ab}
![Page 12: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/12.jpg)
Matcher Strategies Strategies based on linguistic matching Structure-based strategiesStructure-based strategies Constraint-based approaches Instance-based strategies Use of auxiliary information
![Page 13: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/13.jpg)
Example matchers
Propagation of similarity values Anchored matching
![Page 14: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/14.jpg)
Example matchers
Propagation of similarity values Anchored matching
![Page 15: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/15.jpg)
Example matchers
Propagation of similarity values Anchored matching
![Page 16: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/16.jpg)
Matcher Strategies Strategies based on linguistic matching Structure-based strategies Constraint-based approachesConstraint-based approaches Instance-based strategies Use of auxiliary information
O1O2
Bird
Mammal Mammal
FlyingAnimal
![Page 17: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/17.jpg)
Matcher Strategies Strategies based on linguistic matching Structure-based strategies Constraint-based approachesConstraint-based approaches Instance-based strategies Use of auxiliary information
O1O2
Bird
Mammal Mammal
Stone
![Page 18: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/18.jpg)
Example matchers
Similarities between data types Similarities based on cardinalities
![Page 19: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/19.jpg)
Matcher Strategies Strategies based on linguistic matching Structure-based strategies Constraint-based approaches Instance-based strategiesInstance-based strategies Use of auxiliary information
Ontology
instancecorpus
![Page 20: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/20.jpg)
Example matchers
Instance-based Use life science literature as instances
Structure-based extensions
![Page 21: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/21.jpg)
Learning matchers – instance-based strategies Basic intuition
A similarity measure between concepts can be computed based on the probability that documents about one concept are also about the other concept and vice versa.
Intuition for structure-based extensionsDocuments about a concept are also about their
super-concepts.
(No requirement for previous alignment results.)
![Page 22: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/22.jpg)
Learning matchers - steps Generate corpora
Use concept as query term in PubMed Retrieve most recent PubMed abstracts
Generate text classifiers One classifier per ontology / One classifier per concept
Classification Abstracts related to one ontology are classified by the other
ontology’s classifier(s) and vice versa Calculate similarities
![Page 23: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/23.jpg)
Basic Naïve Bayes matcher Generate corpora Generate classifiers
Naive Bayes classifiers, one per ontology Classification
Abstracts related to one ontology are classified to the concept in the other ontology with highest posterior probability P(C|d)
Calculate similarities
![Page 24: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/24.jpg)
Structural extension ‘Cl’ Generate classifiers
Take (is-a) structure of the ontologies into account when building the classifiers
Extend the set of abstracts associated to a concept by adding the abstracts related to the sub-concepts
C1
C3
C4
C2
![Page 25: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/25.jpg)
Structural extension ‘Sim’
Calculate similarities Take structure of the ontologies into account when
calculating similarities Similarity is computed based on the classifiers applied
to the concepts and their sub-concepts
![Page 26: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/26.jpg)
Matcher Strategies Strategies based linguistic matching Structure-based strategies Constraint-based approaches Instance-based strategies Use of auxiliary informationUse of auxiliary information
thesauri
alignment strategies
dictionary
intermediateontology
![Page 27: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/27.jpg)
Example matchers
Use of WordNet Use WordNet to find synonyms Use WordNet to find ancestors and descendants in the is-
a hierarchy Use of Unified Medical Language System (UMLS)
Includes many ontologies Includes many alignments (not complete) Use UMLS alignments in the computation of the
similarity values
![Page 28: Ontology Alignment. Ontologies in biomedical research many biomedical ontologies e.g. GO, OBO, SNOMED-CT practical use of biomedical ontologies e.g. databases](https://reader030.vdocuments.site/reader030/viewer/2022032415/56649f005503460f94c16b15/html5/thumbnails/28.jpg)
Ontology A
lignment and M
ergning S
ystems