Qualitative scene descriptions from images for integrated speech and image understanding
Explore this paper's citation graph
Summary
The design and implementation of a high-level computer vision component for the integrated speech and image understanding system QUASI-ACE, a prototype of a ‘situated artificial communicator’ which aims to interact with humans in a natural way given a specific scenario or situation.
- Type
- dissertation
- Published
- 1997-01-01
- Cited by
- 12
- References
- 180
- OpenAlex
- https://openalex.org/W81741686
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:11513446
Keywords
Computer science, Artificial intelligence, Component (thermodynamics), Object (grammar), Computer vision
References
- Automatisches Verstehen gesprochener Sprache
- Representing Procedural Knowledge for Semantic Networks using Neural Nets
- Interpretation von Bild- und Sprachsignalen: ein hybrider Ansatz
- A Hybrid Approach to Identifying Objects from Verbal Descriptions
- Ein System zur automatischen Analyse von Sequenzszintigrammen des Herzens
- Basic Meanings of Spatial Relations: Computation and Evaluation in 3D Space
- The dialog module of the speech recognition and dialog system EVAR
- 3D Dynamic Scene Analysis
- Control Structures for Incorporating Picture-Specific Context in Image Interpretation
- A REPLAI of SOCCER: Recognizing Intentions in the Domain of Soccer Games
- Neat versus scruffy: a review of computational models for spatial expressions
- Ermittlung von begrifflichen Beschreibungen von Geschehen in Straßenverkehrsszenen mit Hilfe unscharfer Mengen
- Psychologie der Objektbenennung
- Sprechen und Situation
- Das menschliche Gedächtnis : das Erinnern von Sprache, Bildern und Handlungen
- Allgemeine Sprachpsychologie : Grundlagen und Probleme
- Objekterkennung mit Neuronalen Netzen
- [서평]Computer Graphics : Principles and Practice
- Conceptual Structures: Information Processing in Mind and Machine
- Qualitative Representation of Spatial Knowledge
Cited by
- Connecting concepts from vision and speech processing
- Multi-modal scene understanding using probabilistic models
- A computational model for the influence of cross modal context upon syntactic parsing
- Bayesian reasoning on qualitative descriptions from images and speech
- Describing Images Using Qualitative Models and Description Logics
- A Computational Model of Concept Generalization in Cross-Modal Reference
- Integration of Vision and Speech Understanding Using Bayesian Networks
- Multilevel Integration of Vision and Speech Understanding Using Bayesian Networks
- Qualitative distances and qualitative description of images for indoor scene description and recognition in robotics
- Ein Raummodell für die Bennung von Objekten in 3D-Szenen
- Sehen und Verstehen: Der Beitrag bildlicher Information zur robusten Sprachverarbeitung
- Qualitative Models of Shape , Size , Orientation and Distance Applied to the Description of Images Containing 2 D Objects