Artwork

Contenuto fornito da Raza Habib. Tutti i contenuti dei podcast, inclusi episodi, grafica e descrizioni dei podcast, vengono caricati e forniti direttamente da Raza Habib o dal partner della piattaforma podcast. Se ritieni che qualcuno stia utilizzando la tua opera protetta da copyright senza la tua autorizzazione, puoi seguire la procedura descritta qui https://it.player.fm/legal.
Player FM - App Podcast
Vai offline con l'app Player FM !

Why Your AI Product Needs Evals with Hamel Husain and Swyx

1:09:02
 
Condividi
 

Manage episode 441766382 series 3586305
Contenuto fornito da Raza Habib. Tutti i contenuti dei podcast, inclusi episodi, grafica e descrizioni dei podcast, vengono caricati e forniti direttamente da Raza Habib o dal partner della piattaforma podcast. Se ritieni che qualcuno stia utilizzando la tua opera protetta da copyright senza la tua autorizzazione, puoi seguire la procedura descritta qui https://it.player.fm/legal.

Hamel Husain is a seasoned AI consultant and engineer with experience at companies like GitHub, DataRobot, and Airbnb. He is a trailblazer in AI development, known for his innovative work in literate programming and AI-assisted development tools. Shawn Wang (aka Swyx) is the host of the Latent Space podcast, the author of the essay 'Rise of the AI Engineer,' and the founder of the AI Engineer World Fair. In this episode, Hamel and Swyx share their unique insights on building effective AI products, the critical importance of evaluations, and their vision for the future of AI engineering.

Chapters
00:00 - Introduction and recent AI advancements

06:14 - The critical role of evals in AI product development

15:33 - Common pitfalls in AI product development

26:33 - Literate programming: A new paradigm for AI development

39:58 - Answer AI and innovative approaches to software development

51:56 - Integrating AI with literate programming environments

58:47 - The importance of understanding AI prompts

01:00:37 - Assessing the current state of AI adoption

01:07:10 - Challenges in evaluating AI models

--------------------------------------------------------------------------------------------------------------------------------------------------
Humanloop is an Integrated Development Environment for Large Language Models. It enables product teams to develop LLM-based applications that are reliable and scalable. To find out more go to humanloop.com

  continue reading

22 episodi

Artwork
iconCondividi
 
Manage episode 441766382 series 3586305
Contenuto fornito da Raza Habib. Tutti i contenuti dei podcast, inclusi episodi, grafica e descrizioni dei podcast, vengono caricati e forniti direttamente da Raza Habib o dal partner della piattaforma podcast. Se ritieni che qualcuno stia utilizzando la tua opera protetta da copyright senza la tua autorizzazione, puoi seguire la procedura descritta qui https://it.player.fm/legal.

Hamel Husain is a seasoned AI consultant and engineer with experience at companies like GitHub, DataRobot, and Airbnb. He is a trailblazer in AI development, known for his innovative work in literate programming and AI-assisted development tools. Shawn Wang (aka Swyx) is the host of the Latent Space podcast, the author of the essay 'Rise of the AI Engineer,' and the founder of the AI Engineer World Fair. In this episode, Hamel and Swyx share their unique insights on building effective AI products, the critical importance of evaluations, and their vision for the future of AI engineering.

Chapters
00:00 - Introduction and recent AI advancements

06:14 - The critical role of evals in AI product development

15:33 - Common pitfalls in AI product development

26:33 - Literate programming: A new paradigm for AI development

39:58 - Answer AI and innovative approaches to software development

51:56 - Integrating AI with literate programming environments

58:47 - The importance of understanding AI prompts

01:00:37 - Assessing the current state of AI adoption

01:07:10 - Challenges in evaluating AI models

--------------------------------------------------------------------------------------------------------------------------------------------------
Humanloop is an Integrated Development Environment for Large Language Models. It enables product teams to develop LLM-based applications that are reliable and scalable. To find out more go to humanloop.com

  continue reading

22 episodi

Tutti gli episodi

×
 
Loading …

Benvenuto su Player FM!

Player FM ricerca sul web podcast di alta qualità che tu possa goderti adesso. È la migliore app di podcast e funziona su Android, iPhone e web. Registrati per sincronizzare le iscrizioni su tutti i tuoi dispositivi.

 

Guida rapida