Item request has been placed! ×
Item request cannot be made. ×
loading  Processing Request

Monitoring geometrical properties of word embeddings for detecting the emergence of new topics

Item request has been placed! ×
Item request cannot be made. ×
loading   Processing Request
  • المؤلفون: Christophe, Clément; Velcin, Julien; Cugliari, Jairo; Boumghar, Manel; Suignard, Philippe
  • المصدر:
    Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing ; 2021 Conference on Empirical Methods in Natural Language Processing ; https://hal.science/hal-03699173 ; 2021 Conference on Empirical Methods in Natural Language Processing, Nov 2021, Punta Cana, Dominican Republic. ⟨10.18653/v1/2021.emnlp-main.76⟩
  • الموضوع:
  • نوع التسجيلة:
    conference object
  • اللغة:
    English
  • معلومة اضافية
    • Contributors:
      EDF R&D (EDF R&D); EDF (EDF); Entrepôts, Représentation et Ingénierie des Connaissances (ERIC); Université Lumière - Lyon 2 (UL2)-Université Claude Bernard Lyon 1 (UCBL); Université de Lyon-Université de Lyon
    • بيانات النشر:
      HAL CCSD
    • الموضوع:
      2021
    • Collection:
      Université de Lyon: HAL
    • الموضوع:
    • نبذة مختصرة :
      International audience ; Slow emerging topic detection is a task between event detection, where we aggregate behaviors of different words on short period of time, and language evolution, where we monitor their long term evolution. In this work, we tackle the problem of early detection of slowly emerging new topics. To this end, we gather evidence of weak signals at the word level. We propose to monitor the behavior of words representation in an embedding space and use one of its geometrical properties to characterize the emergence of topics. As evaluation is typically hard for this kind of task, we present a framework for quantitative evaluation. We show positive results that outperform state-ofthe-art methods on two public datasets of press and scientific articles.
    • Relation:
      hal-03699173; https://hal.science/hal-03699173; https://hal.science/hal-03699173/document; https://hal.science/hal-03699173/file/EMNLP_2021_CameraReady__Monitoring_geometrical_properties_of_word_embeddings_for_the_detection_of_new_topic_emergence.pdf
    • الرقم المعرف:
      10.18653/v1/2021.emnlp-main.76
    • Rights:
      info:eu-repo/semantics/OpenAccess
    • الرقم المعرف:
      edsbas.CFB04EAD