Which of these sound most exciting to you?. A new voice model that not only can faithfully transcribe what you actually say, but even clean it up and get at your intent.
Why listen
It goes beyond the title with direct discussion of nvidia, model, writes, including: A new ability for Claude to use a browser window to be able to do tasks for you.
Key takeaways
01A new voice model that not only can faithfully transcribe what you actually say, but even clean it up and get at your intent
02A new video model from Google with massively more controllability for actually generating videos that you can use for real things
03Or a totally different video model that takes you less time to generate the video than it does to watch it
Best for
research-minded practitioners comparing model behavior