Reshaping Workflows with Dell Pro Precision and NVIDIA RTX PRO GPUs
Reshaping Workflows with Dell Pro Precision and NVIDIA RTX PRO GPUs

10x Faster AI: Revolutionizing LLM Performance

20 August 2026 31:48 Dell Technologies AI Factory with NVIDIA

Listen to episode

About this episode

Is slow model performance stopping your AI projects from scaling?

In this episode of Reshaping Workflows with Dell Pro Precision and NVIDIA RTX GPUs, Logan Lawler sits down with Stefano Ermon, academic and Inception CEO, to uncover how diffusion-based LLMs are breaking speed records, generating tokens up to 10x faster than traditional models.

Stefano Ermon walks us through the evolution from academic research at Stanford to launching the Mercury 2 model, now powering high-speed, latency-sensitive enterprise applications. From the science of parallel generative processes to enterprise-grade deployment on NVIDIA hardware, learn how Inception’s models are changing the rules for AI inference, cost, and scalability.

You’ll also hear real-world applications across voice agents, code generation, and search, plus candid discussion about the future of AI model specialization as hardware rapidly advances.

Watch now and see how AI is getting infinitely faster!

You can also watch this and all previous episodes here.

Follow Us

  • LinkedIn @DellTechnologies
  • Twitter @DellTech
  • YouTube @DellTechnologies
  • LinkedIn @NVIDIA
  • Twitter @NVIDIA
  • Instagram @NVIDIA
  • Facebook @NVIDIA

Presented by Dell and NVIDIA
https://www.dell.com/precisionai

https://www.dell.com/dellpromax

https://www.dell.com/nvidia-ai


Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.

Want to find AI jobs?

Join thousands of AI professionals finding their next opportunity

We respect your inbox. Unsubscribe at any time.

© 2026 Reshaping Workflows with Dell Pro Precision and NVIDIA RTX PRO GPUs. All rights reserved.

Common Questions

Frequently asked questions

Quick answers about how DevFound's AI matching, resumes, and referrals work.

DevFound's AI Copilot ingests your profile, goals, and live job data to deliver curated matches in seconds. Every match includes a resume variant, suggested referrals, and interview prep so you can act immediately. The more feedback you provide, the sharper the Copilot becomes.

AI-led job searches shrink the hours spent sifting through boards and formatting resumes. DevFound pairs automation with your personal outreach, so you reserve energy for interviews and negotiation. Traditional networking still matters, but AI gives you a lift before you even send a message.

Modern AI roles expect comfort with production-grade code, data fluency, and practical ML tooling. The strongest candidates pair deep technical chops with storytelling—translating model impact to product, GTM, and exec partners. Continuous learning keeps you ahead as stacks evolve.

DevFound rewards active seekers. Keep your profile fresh, respond to match quality prompts, and enable alerts so you never miss a role. The AI prioritizes companies and teams that align with your feedback, accelerating both introductions and interview invites.

High-density tech hubs continue to host the deepest AI talent pools, yet distributed teams are catching up fast. Use DevFound filters to hone in on onsite, hybrid, or fully remote roles and watch openings expand across time zones.

DevFound aggregates thousands of remote AI openings and flags the nuances—core hours, async culture, and visa needs—up front. The Copilot also recommends how to position your distributed work experience so hiring managers know you can thrive on a remote team.