9 Apr 2024, 13:51 UTC≈15,800 views29 reactionsread 7 August 2026 🎸 Suno v3 prompt engineering guide
Go to the app homepage, create tab. There's the simple mode (which will generate a song and lyrics, but without the tricks below), and the custom mode wiht more contorl. We tap the second, of course. Now we see a prompt and lyrics window.
1. Workflow.
First generation is a max of 2 minutes. Usually, it can include an intro, verse, and chorus (maybe more if you have a high tempo). …
👍14❤13🔥2
9 Apr 2024, 13:51 UTC≈9,840 views5 reactionsread 7 August 2026 Photo
Suno v3 - The best txt2music model
Recently released Suno v3 is the absolute best txt2music and txt2audio model ever.
Suno v3 is capable of generating actually interesting 2-minute songs in one go (or even potentially indefinitely long ones with the continue function). And yes, precisely songs! Because it also generates vocals, which have been greatly upgraded in the last version. So to put it in perspective, Suno v…
👍4❤1
30 Mar 2024, 17:13 UTC≈7,740 views1 reactionsread 7 August 2026 The paper also shows that SD3-Turbo is better than Midjourney 6 in both image quality and prompt alignment, which is surprising 🫥. Looking forward to the weights release to do a reality check!
@gradientdude
👍1
30 Mar 2024, 17:11 UTC≈7,170 views15 reactionsread 7 August 2026 Photo
⚡️SD3-Turbo: Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation
Following Stable Diffusion 3, my ex-colleagues have published a preprint on SD3 distillation using 4-step, while maintaining quality.
The new method – Latent Adversarial Diffusion Distillation (LADD), which is similar to ADD (see post about it in @ai_newz), but with a number of differences:
️↪️ Both teacher and student…
❤7👍5😍2🔥1
30 Mar 2024, 17:10 UTC≈4,700 views21 reactionsread 7 August 2026 Photo
I'm getting back to juicy posts in english!
@gradientdude
🔥11👍7❤1🏆1😍1
29 Feb 2024, 15:32 UTC≈5,330 views129 reactionsread 7 August 2026 Photo
Staff Research Scientist: Personal Update
I have some exciting news that I'd like to share with you! On Monday, I was promoted to E6, which means I am now a Staff Research Scientist at Meta GenAI.
This was made possible thanks to the significant impact and scope of a Generative AI project that I proposed, led, and completed last year. The project is not yet public, so I can't share details about it right now.
Befo…
🔥93❤18🎉11👍7
20 Jun 2023, 16:03 UTC≈7,580 views35 reactionsread 7 August 2026 Video
Today I will be presenting our CVPR2023 poster "Avatars Grow Legs: Generating Smooth Human Motion from Sparse Tracking Inputs with Diffusion Model".
Learn how to synthesize full body motion based on 3 known points only (head and wrists)!
❱❱ Detailed post about the paper.
Come to chat with me today at 10:30-12:30 PDT, poster #46 if you are at CVPR.
@gradientdude
🔥31❤3🤩1
1 May 2023, 20:47 UTC≈7,690 views24 reactionsread 7 August 2026 Video
Demo of our Avatars Grow Legs model that synthesizes the full 3D body motion based on sparse tracking inputs from the head and the wrists.
More details are in the paper.
@gradientdude
🔥21❤2🥰1
1 May 2023, 20:44 UTC≈6,370 views17 reactionsread 7 August 2026 Video
🦿Avatars Grow Legs
I'm thrilled to share with you my latest research paper (CVPR 2023)! This was a joint effort with my intern at Meta Reality Labs before our team transitioned to GenAI.
Our innovative method, dubbed Avatars Grow Legs (AGRoL), aims to control a 3D avatar's entire body in VR without the need for extra sensors. Typically in VR, your interaction is limited to a headset and two handheld joysticks, leav…
🔥9👍7❤1
1 Dec 2022, 00:48 UTC≈7,080 views30 reactionsread 7 August 2026 Photo
I will be presenting our VisCo Grids paper at NeurIPS tomorrow at 11:00-13:00 in Hall J.
Feel free to come to poster #527 to learn more details if you are attending the conference!
@gradientdude
👍17🔥12❤1
1 Dec 2022, 00:42 UTC≈6,290 views9 reactionsread 7 August 2026 VisCo Grids: Surface Reconstruction with Viscosity and Coarea Grids
The goal of this work it show that network-free grid-based implicit representations can achieve INR (Implicit Neural Representation)-level 3D reconstructions when incorporating suitable priors.
VisCo Grids is a grid-based surface reconstruction algorithm that incorporates well-defined geometric priors: Viscosity and Coarea.
The viscosity loss, i…
👍4🔥3❤2
11 Jul 2022, 13:00 UTC≈6,580 views15 reactionsread 7 August 2026 Image Inpainting: Partial Convolution vs Gated Convolution
Let’s talk about some essential components of the image inpainting networks - convolutions. #fundamentals
It is common in image inpainting model to feed a corrupted image (with some parts masked out) to the generator network. But we don’t want the network layers to rely on empty regions when features are computed. There is a straightforward solutions to t…
👍6🔥5👏4
Showing the 12 most recent of 15 posts we hold for @gradientdude. View and reaction counts are the latest single reading for each post, not a live figure, and a recent post is still accumulating both. A view count marked ≈ was rounded by Telegram before we ever saw it — t.me prints views in full below 1,000 and to three significant figures above, so ≈1,200,000 means somewhere between 1,150,000 and 1,249,999. Unmarked counts are exact. Text is reproduced from the public post preview and truncated for length.