Meta tags:
description= Workshop on Advances in Language and Vision Research ;
author= ;
Headings (most frequently used words):
in, research, 2021, university, microsoft, from, vqa, to, vln, recent, advances, vision, and, language, cvpr, 00, pdt, conjunction, with, june, 20th, am, pm, location, virtual, tutorial, on, program, utc, organizers, contacts, peter, anderson, yoav, artzi, zhe, gan, xiaodong, he, linjie, li, jingjing, liu, xin, eric, wang, qi, wu, luowei, zhou, google, cornell, jd, com, tsinghua, uc, santa, cruz, of, adelaide,
Text of the page (most frequently used words):
and (18), language (11), video (10), the (8), vln (8), #vision (8), slides (7), #tutorial (6), research (5), 2021 (5), natural (4), embodied (4), vqa2vln (3), microsoft (3), university (3), anderson (3), training (3), program (3), navigation (3), will (3), agents (3), this (3), recent (3), advances (3), from (3), contact (2), com (2), luowei (2), zhou (2), xin (2), eric (2), wang (2), jingjing (2), liu (2), linjie (2), xiaodong (2), zhe (2), gan (2), yoav (2), artzi (2), peter (2), organizers (2), live (2), panel (2), discussion (2), pre (2), 40min (2), for (2), vlp (2), sessions (2), pdt (2), that (2), visual (2), environment (2), computer (2), processing (2), about (2), with (2), cvpr (2), vqa (2), organizing, committee, gmail, contacts, adelaide, santa, cruz, tsinghua, cornell, google, all, speakers, zoom, session, summary, 15min, forward, realistic, 58min, generalizable, methods, 55min, introduction, 42min, robustness, efficiency, extensions, representations, strategies, 50min, opening, remarks, 4min, prerecorded, our, divided, into, two, sub, recording, available, after, utc, long, term, goal, build, intelligent, can, see, rich, around, communicate, understanding, humans, other, act, physical, end, nexus, have, made, tremendous, progress, generating, descriptions, images, videos, answering, questions, them, holding, free, form, conversations, content, most, recently, where, are, trained, perform, various, tasks, egocentric, perception, has, attracted, surge, interest, within, robotics, communities, one, fundamental, topic, was, proposed, not, only, cover, latest, approaches, principles, frontier, but, also, present, comprehensive, overview, field, full, day, event, 00pm, several, middle, breaks, home, toggle, photo, unsplash, nasa, conjunction, june, location, virtual,
Text of the page (random words):
vqa2vln tutorial 2021 from vqa to vln recent advances in vision and language research in conjunction with cvpr 2021 june 20 th 2021 9 00 am 5 00 pm pdt location virtual photo by nasa on unsplash toggle navigation vqa2vln 2021 home program organizers contact cvpr 2021 tutorial on from vqa to vln recent advances in vision and language research a long term goal of ai research is to build intelligent agents that can see the rich visual environment around us communicate this understanding in natural language to humans and other agents and act in a physical or embodied environment to this end recent advances at the nexus of computer vision and natural language processing have made tremendous progress from generating natural language descriptions of images videos to answering questions about them and to holding free form conversations about visual content most recently embodied ai where embodied agents are trained to perform various tasks in egocentric perception has attracted a surge of interest within computer vision natural language processing and robotics communities vision language navigation vln is one fundamental topic in embodied ai that was proposed by anderson and wu et al in this tutorial we will not only cover the latest approaches and principles at the frontier of vision and language research but also present a comprehensive overview of the field of vln the tutorial will be a full day event 9 00 am to 5 00pm with several middle breaks program pdt utc 7 our program is divided into two sub sessions 1 vision and language pre training and 2 vision and language navigation recording of panel discussion will be available after the tutorial prerecorded sessions 4min opening remarks video jingjing liu and xiaodong he 50min representations and training strategies for vlp video slides zhe gan 40min robustness efficiency and extensions for vlp video slides linjie li 40min video and language pre training video slides luowei zhou 42min introduction to vln video slides qi wu 55min generalizable vln methods video slides xin eric wang 58min forward to realistic vln video slides yoav artzi and peter anderson 15min vln summary video slides qi wu live session 16 00 17 00 panel discussion live on zoom video all speakers organizers peter anderson google research yoav artzi cornell university zhe gan microsoft xiaodong he jd com linjie li microsoft jingjing liu tsinghua university xin eric wang uc santa cruz qi wu university of adelaide luowei zhou microsoft contacts contact the organizing committee vqa2vln tutorial gmail com
|