{"href":"https://api.simplecast.com/oembed?url=https%3A%2F%2Fthenewstack.simplecast.com%2Fepisodes%2Fnvidia-FLZY21pM","width":444,"version":"1.0","type":"rich","title":"Nvidia","thumbnail_width":300,"thumbnail_url":"https://image.simplecastcdn.com/images/1425ebfd-95bd-4a66-b963-a0b885c75680/bb688835-10e4-4197-b01f-34221ccb5d38/tns-makers-logo-simplecast.jpg","thumbnail_height":300,"provider_url":"https://simplecast.com","provider_name":"Simplecast","html":"<iframe src=\"https://player.simplecast.com/9cadd633-2683-4d9c-8c78-9fc3b17a457e\" height=\"200\" width=\"100%\" title=\"Nvidia\" frameborder=\"0\" scrolling=\"no\"></iframe>","height":200,"description":"In this episode with The New Stack's Frederic Lardinois, NVIDIA’s Joey Conway says advances in AI over the past year have dramatically improved the capabilities of local models, making them practical for enterprise and personal use alongside frontier cloud models. Rather than replacing large models, Conway envisions a “system of models” where specialized local models handle routine, cost-sensitive, or privacy-focused tasks, while larger frontier models tackle more complex reasoning. He explains that organizations can fine-tune smaller open models using domain-specific data, creating expert AI agents that reflect the specialized roles found within businesses. "}