首页 > AI前沿 > Uncertainty in Representation Learning on Knowledge Graphs

Uncertainty in Representation Learning on Knowledge Graphs

arXiv机器学习 2026-10-04 07:02 4 阅读 查看原文

Knowledge graph embedding (KGE) methods represent entities and predicates in continuous vector spaces to infer missing knowledge.

Despite strong benchmark performance, their predictions often lack principled reliability guarantees, limiting their use in high-stakes applications.

Moreover, uncertainty arises throughout the KGE pipeline, from incomplete or probabilistic input knowledge to stochastic training and prediction.

This thesis systematically investigates three sources of uncertainty in KGE:

  • Knowledge uncertainty, arising from incomplete, noisy, or probabilistic input knowledge;
  • Algorithmic uncertainty, induced by randomness in model training;
  • Predictive uncertainty, concerning the reliability of model outputs.

To address algorithmic uncertainty, the thesis demonstrates that models trained under identical settings can produce substantially different predictions and introduces a voting-based aggregation framework to mitigate this instability.

To quantify predictive uncertainty, it adapts conformal prediction to KGE, constructing answer sets with distribution-free coverage guarantees and extending them to provide predicate-conditional reliability guarantees.

To support reasoning under knowledge uncertainty, it develops statistically valid prediction intervals for confidence-scored triples and an embedding-based approach to approximate probabilistic reasoning over statistical ontologies with formal soundness guarantees.

Together, these complementary, model-agnostic methods provide a practical and theoretically grounded approach to uncertainty in KGE, advancing beyond predictive accuracy toward reliable and uncertainty-aware knowledge graph reasoning.