We’re happy to announce that we just published a tentative schedule for #WAC2022 on the conference website: wac2022.i3s.univ-cotedazur.f… @WebAudioConf @Laboratoire_I3S @inria_sophia @EURCreates #webaudio #musictech
2
15
Creating, capturing and displaying immersive 3D scenes. Talk by Adrien Boisseau and George Drettakis from the INRIA Graphdeco team. At the 3rd Journée d'études #XR2C2 polyteh@Univ_CotedAzur @inria_sophia @Laboratoire_I3S
2
2
179
Musical metaverse: opportunities and challenges. Talk by Luca Turchet from @UniTrento at the 3rd Journée d'études #XR2C2 @Univ_CotedAzur @Laboratoire_I3S
2
2
142
Informed virtual environments and adaptive sensorial feedback. Talk by Indira Thouvenin of @utcompiegne at the 3rd Journée d'études #XR2C2. @Univ_CotedAzur @Laboratoire_I3S
1
1
110
ViAjeRo: XR and passengers of the future. Talk by Stephen Brewster from @UofGlasgow at the 3rd Journée d'études #XR2C2. @Univ_CotedAzur @Laboratoire_I3S
2
2
138
Computational models for natural visual perception and interaction in XR. Talk by Fabio Solari from @UniGenova at the 3rd Journée d'études #XR2C2. @Univ_CotedAzur @Laboratoire_I3S
1
2
2
123
WAC2022 retweeted
Robot factory in China 🇨🇳
68
195
527
100,301
WAC2022 retweeted
If you want to use YOLO-World in your own project, you can do so through the @roboflow Inference Python package. github: github.com/roboflow/inferenc…
1
2
13
1,034
WAC2022 retweeted
The YOLO-World YouTube tutorial is out! please, let us know what you think! - model architecture - processing images and video in Colab - prompt engineering and detection refinement - pros and cons of the model watch here: piped.video/watch?v=X7gKBGVz… ↓ more resources
11
124
777
91,629
WAC2022 retweeted
The Diffusion Transformer paper, by my former-FAIR-and-current-NYU colleague @sainingxie and former-Berkeley-student-and-current-OpenAI engineer William Peebles, was rejected from CVR2023 for "lack of novelty", accepted at ICCV2023, and apparently forms the basis for Sora. openaccess.thecvf.com/conten…
Here's my take on the Sora technical report, with a good dose of speculation that could be totally off. First of all, really appreciate the team for sharing helpful insights and design decisions – Sora is incredible and is set to transform the video generation community. What we have learned so far: - Architecture: Sora is built on our diffusion transformer (DiT) model (published in ICCV 2023) — it's a diffusion model with a transformer backbone, in short: DiT = [VAE encoder + ViT + DDPM + VAE decoder]. According to the report, it seems there are not much additional bells and whistles. - "Video compressor network": Looks like it's just a VAE but trained on raw video data. Tokenization probably plays a significant role in getting good temporal consistency. By the way, VAE is a ConvNet, so DiT technically is a hybrid model ;) (1/n)
41
389
2,432
784,222
WAC2022 retweeted
30 years of web browser market disruption #CES2024
Visual Capitalist
48
416
1,398
239,929
[LIVE] Forum de l'Alternance | Métiers du numérique Les entretiens vont bon train ! Beaucoup d'opportunités pour nos étudiants ! #alternance #numérique @MstratDigitale @JMCNice @Univ_CotedAzur @SopraSteria_fr @MonacoTelecom
1
2
267
WAC2022 retweeted
One of the most impressive CV works I've seen recently. Also huge kudos to Meta AI for sticking to open sourcing despite the trend increasingly going towards the opposite direction.
Today we're releasing the Segment Anything Model (SAM) — a step toward the first foundation model for image segmentation. SAM is capable of one-click segmentation of any object from any photo or video + zero-shot transfer to other segmentation tasks ➡️ bit.ly/433YuBI
3
57
8,478
WAC2022 retweeted
We've released the Koala into the wild🐨 The Koala is a chatbot finetuned from LLaMA that is specifically optimized for high-quality chat capabilities, using some tricky data sourcing. Our blog post: bair.berkeley.edu/blog/2023/… Web demo: koala.lmsys.org
13
86
467
139,349
Might be the most impressive version of this yet
LERF: Language Embedded Radiance Fields TL;DR: Grounding CLIP vectors volumetrically inside a NeRF allows flexible natural language queries in 3D abs: arxiv.org/abs/2303.09553 project page: lerf.io/
1
4
43
6,362
WAC2022 retweeted
If a person in some parts of China crosses the street in the improper location, the facial recognition technology will immediately send them to a public shame board, connect them to their bank card, and deduct the amount of the fine
97
252
589
253,558
WAC2022 retweeted
Bing Chat is underrated. Most people don't know what it’s capable of, how it differs from ChatGPT and what kind of amazing content it can generate. 🚀 8 compelling ways to talk to The New Bing: 🧵👇
10
61
407
117,437
Deadline extension for for the Call For Paper for the next Web Audio Special Issue of the Journal of the Audio Engineering Society (JAES). Deadline is now March 15, 2023. tinyurl.com/4ufdmt5n
2
3
785
Excited to share SingSong, a system which can generate instrumental accompaniments to pair with input vocals! 📄arxiv.org/abs/2301.12662 🔊g.co/magenta/singsong Work co-led by myself, @antoine_caillon, and @ada_rob as part of @GoogleMagenta and the broader MusicLM project 🧵
133
699
3,119
1,458,994
those recent, powerful music generation models make me wonder why / to what extent we need music source separation.
9
3
27
14,403
👾Vous hésitez à rejoindre le concours de programmation #GamesOnWeb ? Ces gagnants 2022 ont un message pour vous ! @Univ_CotedAzur|@univamu|@UT3PaulSabatier avec @CGI_FR 👉Inscriptions bit.ly/GOW2023 #GOW2023 @MIAGENiceSophia|@PolytechNSophia|@iut_nca|@MiageAixMrs
1
4