Skip to content

TheAwesomeAndy/Awesome-Multimodal-Research

 
 

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 

Repository files navigation

Awesome Multimodal Research Awesome

build license prs

This repo is reorganized from Paul Liang's repo: Reading List for Topics in Multimodal Machine Learning, any suggestions are welcome!

Research Papers

News

[01/2021] OpenAI: We’ve developed two neural networks which have learned by associating text and images. CLIP maps images into categories described in text, and DALL-E creates new images, like this, from text. A step toward systems with deeper understanding of the world. https://openai.com/multimodal/

Recent Workshop

Advances in Language and Vision Research (ALVR), NAACL 2021

Visually Grounded Interaction and Language (ViGIL), NAACL 2021

Wordplay: When Language Meets Games, NeurIPS 2020

NLP Beyond Text, EMNLP 2020

International Challenge on Compositional and Multimodal Perception, ECCV 2020

Multimodal Video Analysis Workshop and Moments in Time Challenge, ECCV 2020

Video Turing Test: Toward Human-Level Video Story Understanding, ECCV 2020

Grand Challenge and Workshop on Human Multimodal Language, ACL 2020

Workshop on Multimodal Learning, CVPR 2020

Language & Vision with applications to Video Understanding, CVPR 2020

International Challenge on Activity Recognition (ActivityNet), CVPR 2020

The End-of-End-to-End A Video Understanding Pentathlon, CVPR 2020

Towards Human-Centric Image/Video Synthesis, and the 4th Look Into Person (LIP) Challenge, CVPR 2020

Visual Question Answering and Dialog, CVPR 2020

Recent Tutorial

Multi-modal Information Extraction from Text, Semi-structured, and Tabular Data on the Web (Cutting-edge), ACL 2020

Achieving Common Ground in Multi-modal Dialogue (Cutting-edge), ACL 2020

Recent Advances in Vision-and-Language Research, CVPR 2020

Neuro-Symbolic Visual Reasoning and Program Synthesis, CVPR 2020

Large Scale Holistic Video Understanding, CVPR 2020

A Comprehensive Tutorial on Video Modeling, CVPR 2020

Related: Code & Team & Dataset

Papers With Code

Research Team

Related Datasets

About

A curated list of Multimodal Related Research.

Resources

License

Stars

Watchers

Forks

Releases

No releases published

Packages

No packages published

Languages

  • Python 100.0%