kth.sePublikationer KTH
Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Recommendation System for Product Test Failures Using BERT
KTH. Ericsson AB, Stockholm, Sweden.
Ericsson AB, Stockholm, Sweden.
Ericsson AB, Stockholm, Sweden.
Ericsson AB, Stockholm, Sweden.
Visa övriga samt affilieringar
2023 (Engelska)Ingår i: 15th International Conference on Knowledge Discovery and Information Retrieval, KDIR 2023 as part of IC3K 2023 - Proceedings of the 15th International Joint Conference on Knowledge Discovery, Knowledge Engineering and Knowledge Management, INSTICC , 2023, s. 206-213Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

Historical failure records can provide insights to investigate if a similar situation occurred during the troubleshooting process in software. However, in the era of information explosion, massive amounts of data make it unrealistic to rely solely on manual inspection of root causes, not to mention mapping similar records. With the ongoing development and breakthroughs of Natural Language Processing (NLP), we propose an end-to-end recommendation system that can instantly generate a list of similar records given a new raw failure record. The system consists of three stages: 1) general and tailored pre-processing of raw failure records; 2) information retrieval; 3) information re-ranking. In the process of model selection, we undertake a thorough exploration of both frequency-based models and language models. To mitigate issues stemming from imbalances in the available labeled data, we propose an updated Recall@K metric that utilizes an adaptive K. We also develop a multi-stage training pipeline to deal with limited labeled data and investigate how different strategies affect performance. Our comprehensive experiments demonstrate that our two-stage BERT model, fine-tuned on extra domain data, achieves the best score over the baseline models.

Ort, förlag, år, upplaga, sidor
INSTICC , 2023. s. 206-213
Nyckelord [en]
Information Retrieval, Language Models, Recommendation System
Nationell ämneskategori
Datavetenskap (datalogi)
Identifikatorer
URN: urn:nbn:se:kth:diva-341693DOI: 10.5220/0012160800003598Scopus ID: 2-s2.0-85179760636OAI: oai:DiVA.org:kth-341693DiVA, id: diva2:1823050
Konferens
15th International Conference on Knowledge Discovery and Information Retrieval, KDIR 2023 as part of the 15th International Joint Conference on Knowledge Discovery, Knowledge Engineering and Knowledge Management, IC3K 2023, Hybrid, Rome, Italy, Nov 13 2023 - Nov 15 2023
Anmärkning

Part of ISBN 9789897586712

QC 20231229

Tillgänglig från: 2023-12-29 Skapad: 2023-12-29 Senast uppdaterad: 2023-12-29Bibliografiskt granskad

Open Access i DiVA

Fulltext saknas i DiVA

Övriga länkar

Förlagets fulltextScopus

Sök vidare i DiVA

Av författaren/redaktören
Sun, Xiaolong
Av organisationen
KTH
Datavetenskap (datalogi)

Sök vidare utanför DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetricpoäng

doi
urn-nbn
Totalt: 225 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf