Meta tags:
Headings (most frequently used words):
dibya, ghosh, selected, publications,
Text of the page (most frequently used words):
dibya (20), ghosh (20), sergey (14), levine (14), learning (7), neurips (6), with (5), iclr (4), kumar (4), reinforcement (4), for (4), data (4), singh (3), and (3), abhishek (3), gupta (3), preprint (3), icml (3), policy (3), 2021 (3), aviral (3), 2023 (3), zhang (3), offline (3), from (3), via (3), oral (3), 2018 (2), avi (2), justin (2), representations (2), goal (2), conditioned (2), policies (2), marc (2), bellemare (2), 2020 (2), implicit (2), 2022 (2), anurag (2), ajay (2), pulkit (2), agrawal (2), adaptive (2), amy (2), ben (2), eysenbach (2), latent (2), chethan (2), bhateja (2), pre (2), training (2), see (2), reach (2), the (2), berkeley (2), google (2), aravind, rajeswaran, vikash, divide, conquer, larry, yang, variational, inverse, control, events, general, framework, driven, reward, definition, 2019, actionable, william, fedus, john, martin, yoshua, bengio, hugo, larochelle, catastrophic, interference, atari, 2600, games, stable, off, marlos, machado, nicolas, roux, operator, view, gradient, methods, rishabh, agarwal, under, parameterization, inhibits, efficient, deep, distributionally, meta, qiyang, jason, accelerating, exploration, unlabeled, prior, seohong, park, hiql, states, actions, derek, guo, anikait, manan, tomar, quan, vuong, yevgen, chebotar, robotic, internet, videos, value, function, all, ashwin, reddy, coline, devin, goals, iterated, supervised, jad, rahme, ryan, adams, why, generalization, difficult, epistemic, pomdps, partial, observability, should, trained, passive, intentions, rss, 2024, homer, walke, karl, pertsch, kevin, black, oier, mees, octo, open, source, generalist, robot, annotation, bootstrapping, self, reinforcing, approach, visual, selected, publications, member, technical, staff, pretraining, team, anthropic, received, phd, working, even, earlier, was, brain, montréal, published, work, more, fun, first, name, pronounced, dibbo, dot, edu, twitter, blog, scholar,
Text of the page (random words):
dibya ghosh dibya ghosh i am a member of the technical staff on the pretraining team at anthropic i received my phd from uc berkeley working with sergey levine even earlier i was at google brain montréal see my google scholar for published work and blog for more fun my first name is pronounced dibbo reach me on twitter or at dibya ghosh at berkeley dot edu selected publications annotation bootstrapping a self reinforcing approach to visual pre training preprint dibya ghosh sergey levine octo an open source generalist robot policy rss 2024 dibya ghosh homer walke karl pertsch kevin black oier mees et al reinforcement learning from passive data via latent intentions icml 2023 oral dibya ghosh chethan bhateja sergey levine offline rl policies should be trained to be adaptive icml 2022 oral dibya ghosh anurag ajay pulkit agrawal sergey levine why generalization in rl is difficult epistemic pomdps and implicit partial observability neurips 2021 dibya ghosh jad rahme aviral kumar amy zhang ryan p adams sergey levine learning to reach goals via iterated supervised learning iclr 2021 oral dibya ghosh abhishek gupta ashwin reddy justin fu coline devin ben eysenbach sergey levine see all robotic offline rl from internet videos via value function pre training preprint chethan bhateja derek guo dibya ghosh anikait singh manan tomar quan vuong yevgen chebotar sergey levine aviral kumar hiql offline goal conditioned rl with latent states as actions neurips 2023 seohong park dibya ghosh ben eysenbach sergey levine accelerating exploration with unlabeled prior data neurips 2023 qiyang li jason zhang dibya ghosh amy zhang sergey levine distributionally adaptive meta reinforcement learning neurips 2022 anurag ajay abhishek gupta dibya ghosh sergey levine pulkit agrawal implicit under parameterization inhibits data efficient deep rl iclr 2021 aviral kumar rishabh agarwal dibya ghosh sergey levine an operator view of policy gradient methods neurips 2020 dibya ghosh marlos c machado nicolas le roux representations for stable off policy reinforcement learning icml 2020 dibya ghosh marc g bellemare on catastrophic interference in atari 2600 games preprint william fedus dibya ghosh john d martin marc g bellemare yoshua bengio hugo larochelle learning actionable representations with goal conditioned policies iclr 2019 dibya ghosh abhishek gupta sergey levine variational inverse control with events a general framework for data driven reward definition neurips 2018 justin fu avi singh dibya ghosh larry yang sergey levine divide and conquer reinforcement learning iclr 2018 dibya ghosh avi singh aravind rajeswaran vikash kumar sergey levine
|