Endre søk
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Scalable Unsupervised Feature Selection with Reconstruction Error Guarantees via QMR Decomposition
KTH, Skolan för elektroteknik och datavetenskap (EECS), Intelligenta system, Robotik, perception och lärande, RPL.ORCID-id: 0000-0002-8044-4773
SEB Group, Stockholm, Sweden.
KTH, Skolan för elektroteknik och datavetenskap (EECS), Intelligenta system, Robotik, perception och lärande, RPL.ORCID-id: 0000-0003-2965-2953
2024 (engelsk)Inngår i: CIKM 2024 - Proceedings of the 33rd ACM International Conference on Information and Knowledge Management, Association for Computing Machinery (ACM) , 2024, s. 3658-3662Konferansepaper, Publicerat paper (Fagfellevurdert)
Abstract [en]

Unsupervised feature selection (UFS) methods have garnered significant attention for their capability to eliminate redundant features without relying on class label information. However, their scalability to large datasets remains a challenge, rendering common UFS methods impractical for such applications. To address this issue, we introduce QMR-FS, a greedy forward filtering approach that selects linearly independent features up to a specified relative tolerance, ensuring that any excluded features can be reconstructed from the retained set within this tolerance. This is achieved through the QMR matrix decomposition, which builds upon the well-known QR decomposition. QMR-FS benefits from linear complexity relative to the number of instances and boasts exceptional performance due to its ability to leverage parallelized computation on both CPU and GPU. Despite its greedy nature, QMR-FS achieves comparable classification and clustering accuracies across multiple datasets when compared to other UFS methods, while achieving runtimes approximately 10 times faster than recently proposed scalable UFS methods for datasets ranging from 100 million to 1 billion elements.

sted, utgiver, år, opplag, sider
Association for Computing Machinery (ACM) , 2024. s. 3658-3662
Emneord [en]
feature selection, linear independence, scalability, unsupervised learning
HSV kategori
Identifikatorer
URN: urn:nbn:se:kth:diva-357143DOI: 10.1145/3627673.3679994ISI: 001349579603081Scopus ID: 2-s2.0-85210013171OAI: oai:DiVA.org:kth-357143DiVA, id: diva2:1918220
Konferanse
33rd ACM International Conference on Information and Knowledge Management, CIKM 2024, Boise, United States of America, October 21-25, 2024
Merknad

Part of ISBN 9798400704369

QC 20241205

Tilgjengelig fra: 2024-12-04 Laget: 2024-12-04 Sist oppdatert: 2025-12-08bibliografisk kontrollert

Open Access i DiVA

Fulltekst mangler i DiVA

Andre lenker

Forlagets fulltekstScopus

Person

Ceylan, CiwanKragic, Danica

Søk i DiVA

Av forfatter/redaktør
Ceylan, CiwanKragic, Danica
Av organisasjonen

Søk utenfor DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetric

doi
urn-nbn
Totalt: 118 treff
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf