Endre søk
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Question Answering for Visual Navigation in Human-Centered Environments
Moscow Institute of Physics and Technology, Moscow, Russia.
HSE University, Moscow, Russia; Artificial Intelligence Research Institute FRC CSC RAS, Moscow, Russia.
Luleå tekniska universitet, Institutionen för system- och rymdteknik, Datavetenskap.ORCID-id: 0000-0003-0069-640x
Moscow Institute of Physics and Technology, Moscow, Russia; Artificial Intelligence Research Institute FRC CSC RAS, Moscow, Russia.
2021 (engelsk)Inngår i: Advances in Soft Computing: 20th Mexican International Conference on Artificial Intelligence, MICAI 2021, Mexico City, Mexico, October 25–30, 2021, Proceedings, Part II / [ed] Ildar Batyrshin, Alexander Gelbukh, Grigori Sidorov, Springer Nature, 2021, s. 31-45Konferansepaper, Publicerat paper (Fagfellevurdert)
Abstract [en]

In this paper, we propose an HISNav VQA dataset - a challenging dataset for a Visual Question Answering task that is aimed at the needs of Visual Navigation in human-centered environments. The dataset consists of images of various room scenes that were captured using the Habitat virtual environment and of questions important for navigation tasks using only visual information. We also propose a baseline for a HISNav VQA dataset, a Vector Semiotic Architecture, and demonstrate its performance. The Vector Semiotic Architecture is a combination of a Sign-Based World Model and Vector Symbolic Architectures. The Sign-Based World Model allows representing various aspects of an agent’s knowledge, and Vector Symbolic Architectures serve on a low computational level. The Vector Semiotic Architecture addresses the symbol grounding problem that plays an important role in the Visual Question Answering Task.

sted, utgiver, år, opplag, sider
Springer Nature, 2021. s. 31-45
Serie
Lecture Notes in Artificial Intelligence, ISSN 0302-9743, E-ISSN 1611-3349
Emneord [en]
Visual question answering, Semiotic approach, Vector symbolic architecture, Habitat, Visual Navigation
HSV kategori
Forskningsprogram
Kommunikations- och beräkningssystem
Identifikatorer
URN: urn:nbn:se:ltu:diva-90559DOI: 10.1007/978-3-030-89820-5_3ISI: 000769230400003Scopus ID: 2-s2.0-85118354614OAI: oai:DiVA.org:ltu-90559DiVA, id: diva2:1657347
Konferanse
20th Mexican International Conference on Artificial Intelligence (MICAI 2021), Mexico City, Mexico, [ONLINE], October 25–30, 2021
Merknad

Funder: Russian Foundation for Basic Research, RFBR (19-37-90164)

Tilgjengelig fra: 2022-05-10 Laget: 2022-05-10 Sist oppdatert: 2025-10-21bibliografisk kontrollert

Open Access i DiVA

Fulltekst mangler i DiVA

Andre lenker

Forlagets fulltekstScopus

Person

Osipov, Evgeny

Søk i DiVA

Av forfatter/redaktør
Osipov, Evgeny
Av organisasjonen

Søk utenfor DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetric

doi
urn-nbn
Totalt: 52 treff
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf