Endre søk
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Vision Beyond Boundaries: An Initial Design Space of Domain-specific Large Vision Models in Human-robot Interaction
KTH, Skolan för elektroteknik och datavetenskap (EECS), Intelligenta system, Robotik, perception och lärande, RPL.ORCID-id: 0000-0003-1804-6296
University of Bergen Bergen, Norway.
KTH, Skolan för elektroteknik och datavetenskap (EECS), Intelligenta system, Robotik, perception och lärande, RPL.ORCID-id: 0000-0003-2965-2953
2024 (engelsk)Inngår i: MobileHCI 2024 Adjunct Proceedings - Publication of the 26th International Conference on Mobile Human-Computer Interaction, Association for Computing Machinery (ACM) , 2024, artikkel-id 17Konferansepaper, Publicerat paper (Fagfellevurdert)
Abstract [en]

The emergence of large vision models (LVMs) is following in the footsteps of the recent prosperity of Large Language Models (LLMs) in following years. However, there's a noticeable gap in structured research applying LVMs to human-robot interaction (HRI), despite extensive evidence supporting the efficacy of vision models in enhancing interactions between humans and robots. Recognizing the vast and anticipated potential, we introduce an initial design space that incorporates domain-specific LVMs, chosen for their superior performance over normal models. We delve into three primary dimensions: HRI contexts, vision-based tasks, and specific domains. The empirical evaluation was implemented among 15 experts across five evaluated metrics, showcasing the primary efficacy in relevant decision-making scenarios. We explore the process of ideation and potential application scenarios, envisioning this design space as a foundational guideline for future HRI system design, emphasizing accurate domain alignment and model selection.

sted, utgiver, år, opplag, sider
Association for Computing Machinery (ACM) , 2024. artikkel-id 17
Emneord [en]
domain-specific, empirical study, human-robot interaction, large vision models
HSV kategori
Identifikatorer
URN: urn:nbn:se:kth:diva-366772DOI: 10.1145/3640471.3680244ISI: 001327589500017Scopus ID: 2-s2.0-85206148319OAI: oai:DiVA.org:kth-366772DiVA, id: diva2:1983162
Konferanse
26th International Conference on Mobile Human-Computer Interaction, MobileHCI 2024, Melbourne, Australia, September 30 - October 3, 2024
Merknad

Part of ISBN 9798400705069

QC 20250709

Tilgjengelig fra: 2025-07-09 Laget: 2025-07-09 Sist oppdatert: 2025-07-09bibliografisk kontrollert

Open Access i DiVA

Fulltekst mangler i DiVA

Andre lenker

Forlagets fulltekstScopus

Person

Zhang, YuchongKragic, Danica

Søk i DiVA

Av forfatter/redaktør
Zhang, YuchongKragic, Danica
Av organisasjonen

Søk utenfor DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetric

doi
urn-nbn
Totalt: 53 treff
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf