Gedas Bertasius
@gedasb7.bsky.social
Assistant Professor at UNC Chapel Hill, previously a postdoc at Meta AI, PhD from UPenn, a basketball enthusiast 🏀. 🔗 https://www.gedasbertasius.com/ 🎓 https://scholar.google.com/citations?hl=en&user=8FWkjw8AAAAJ
Usefully sober fact sheet on Ukraine. Please share with those who need it. understandingwar.org/backgrounder...
Institute for the Study of War
Russian forces currently occupy around 20 percent of Ukraine, leaving the remaining 80 percent of the country under Ukraine's sovereign control. At the current rate of advance, it would take Russian f...
understandingwar.org
Deeply honored & humbled to have received the Presidential #PECASE Award by the @WhiteHouse and @POTUS office! 🙏 Most importantly, very grateful to my amazing mentors, students, postdocs, collaborators, and friends+family for making this possible, and for making the journey worthwhile + beautiful 💙
🎉 Congratulations to Prof. @mohitbansal.bsky.social for receiving the Presidential #PECASE Award by @WhiteHouse, which is the highest honor bestowed by US govt. on outstanding scientists/engineers who show exceptional potential for leadership early in their careers! whitehouse.gov/ostp/news-up...
🚨 We have postdoc openings at UNC 🙂 Exciting+diverse NLP/CV/ML topics**, freedom to create research agenda, competitive funding, very strong students, mentorship for grant writing, collabs w/ many faculty+universities+companies, superb quality of life/weather. Please apply + help spread the word 🙏
Workshop days are always the most engaging and rewarding. Here are my two picks for this weekend: Saturday Video-language video-and-language-workshop-2024.webflow.io Sunday Multimodal algorithmic reasoning marworkshop.github.io/neurips24/ Do you have other recommendations? #NeurIPS #NeurIPS2024
Excited to attend #NeurIPS and give a talk on video-language models for complex video understanding in the First Workshop on Video-Language Models on Saturday at 10:10am PST. Stop by + DM/email if you want to chat about anything related to video-language modeling.
🚨 I’m on the academic job market! j-min.io I work on ✨Multimodal AI✨, advancing reasoning in understanding & generation by: 1⃣ Making it scalable 2⃣ Making it faithful 3⃣ Evaluating + refining it Completing my PhD at UNC (w/ @mohitbansal.bsky.social). Happy to connect (will be at #NeurIPS2024)! 👇🧵
🚨 I am on the faculty job market this year 🚨 I will be presenting at #NeurIPS2024 and am happy to chat in-person or digitally! I work on developing AI agents that can collaborate and communicate robustly with us and each other. More at: esteng.github.io and in thread below 🧵👇
Introducing 🧞Genie 2 🧞 - our most capable large-scale foundation world model, which can generate a diverse array of consistent worlds, playable for up to a minute. We believe Genie 2 could unlock the next wave of capabilities for embodied agents 🧠.
We generate a soundtrack for a silent video, given a text prompt! For example, we can make a cat's meow sound like a lion's roar or a typewriter sound like a piano. Paper: arxiv.org/abs/2411.17698 Webpage: ificl.github.io/MultiFoley/ Led by @czyang.bsky.social! bsky.app/profile/czya...
MultiFoley
Video-Guided Foley Sound Generation with Multimodal Controls
ificl.github.io
🎥 Introducing MultiFoley, a video-aware audio generation method with multimodal controls! 🔊 We can ⌨️Make a typewriter sound like a piano 🎹 🐱Make a cat meow like a lion roars! 🦁 ⏱️Perfectly time existing SFX 💥 to a video. arXiv: arxiv.org/abs/2411.17698 website: ificl.github.io/MultiFoley/
SOTA virtual face re-aging techniques often struggle to preserve identity for large age changes. We present MyTimeMachine, a personalized virtual aging gen. model, trained with ~50 images across 20-40 years. Check out more cool results here: mytimemachine.github.io 1/2 ⬇️👂
My growing list of #computervision researchers on Bsky. Missed you? Let me know. go.bsky.app/M7HGC3Y