Inferring the Why in Images

Explore this paper's citation graph

Summary

The results suggest that transferring knowledge from language into vision can help machines understand why a person might be performing an action in an image, and recently developed natural language models to mine knowledge stored in massive amounts of text.

Type
preprint
Published
2014-06-20
Cited by
43
References
41

Keywords

Computer science, Artificial intelligence, Computer vision, Computer graphics (images)

References

Cited by

Related papers