Sep 4, 2026LLaDA-Image Isn't the Diffusion-Language-Model Trick It Sounds LikeAnt Group's LLaDA-Image borrows its name from a diffusion language model, but only half the pipeline actually works that way. Here's the part that's real and the part that's marketing.IInference Desk
Sep 4, 2026SolarWM Trains on Five-Second Clips, Then Runs for an Hour Without BreakingSolarWM trains video world models on five-second clips, then runs stable rollouts for hours by fixing a training mismatch instead of scaling up the model.IInference Desk
Aug 25, 2026The Training Trick That's Cutting AI Image Generator Costs by Up to 63%ByteDance Seed's DiffusionOPSD swaps one vague final-image score for precise per-step training targets, and the GPU savings are the real headline.IInference Desk