10x Faster AI: Revolutionizing LLM Performance
2026-08-20 · 32 min · episode 47 · 15 entities
Asserted relationships
-
0.40
evidence rules-v5
Feed category: News
-
0.40
evidence rules-v5
Feed category: Technology
-
0.40
evidence rules-v5
Feed category: Tech News
-
0.40
evidence rules-v5
Feed author/publisher: Dell Technologies AI Factory with NVIDIA
-
0.38
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/precisionai
-
0.38
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/dellpromax
-
0.38
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/nvidia-ai
Entities found in this episode
concepts 5
-
0.50
evidence rules-v5
Feed category: Tech News
-
0.42
evidence rules-v5
10x Faster AI: Revolutionizing LLM Performance
-
0.40
evidence rules-v5
Feed category: News
-
0.40
evidence rules-v5
Feed category: Tech News
-
0.35
evidence rules-v5
LLM
companys 4
-
0.70
evidence rules-v5
Feed author/publisher: Dell Technologies AI Factory with NVIDIA
-
0.50
evidence rules-v5
Feed category: Technology
-
0.40
evidence rules-v5
Feed category: Technology
-
0.40
evidence rules-v5
Feed author/publisher: Dell Technologies AI Factory with NVIDIA
websites 4
-
0.45
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/precisionai
-
0.45
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/dellpromax
-
0.38
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/precisionai
-
0.38
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/dellpromax
persons 2
-
0.45
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/nvidia-ai
-
0.38
evidence rules-v5
Link in episode "10x Faster AI: Revolutionizing LLM Performance": https://dell.com/nvidia-ai
Episode description as stored
Is slow model performance stopping your AI projects from scaling?
In this episode of Reshaping Workflows with Dell Pro Precision and NVIDIA RTX GPUs , Logan Lawler sits down with Stefano Ermon, academic and Inception CEO, to uncover how diffusion-based LLMs are breaking speed records, generating tokens up to 10x faster than traditional models.
Stefano Ermon walks us through the evolution from academic research at Stanford to launching the Mercury 2 model, now powering high-speed, latency-sensitive enterprise applications. From the science of parallel generative processes to enterprise-grade deployment on NVIDIA hardware, learn how Inception’s models are changing the rules for AI inference, cost, and scalability.
You’ll also hear real-world applications across voice agents, code generation, and search, plus candid discussion about the future of AI model specialization as hardware rapidly advances.
Watch now and see how AI is getting infinitely faster!
You can also watch this and all previous episodes here .
Follow Us
LinkedIn @DellTechnologies
Twitter @DellTech
YouTube @DellTechnologies
LinkedIn @NVIDIA
Twitter @NVIDIA
Instagram @NVIDIA
Facebook @NVIDIA
Presented by Dell and NVIDIA
https://www.dell.com/precisionai
https://www.dell.com/dellpromax
https://www.dell.com/nvidia-ai
Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.