PodcastIntel
Sign in Get Started Free
Latent Space: The AI Engineer Podcast

Latent Space: The AI Engineer Podcast

Alessio Fanelli · Technology · EN

This podcast explores the technical infrastructure, models, and agents used by leading AI labs, with a focus on AI for science. It features in-depth discussions with prominent AI engineers and researchers. The show is geared towards individuals interested in the practical engineering aspects of...

230
Episodes
120
Guests

Episodes (Page 8)

Oct 18, 2024 · 1h 11m
Drew Houston has spent 400+ hours coding with LLMs and is refocusing Dropbox's 2,500+ employees around AI-native development 17 years after founding the company
1h 11m
Oct 11, 2024 · 1h 56m
Ankur Goyal of Braintrust argues that production AI engineering should start with evals, not operational tooling, following the pattern of successful LLMOps founders with AI/research backgrounds
1h 56m
Oct 3, 2024 · 2h 9m
OpenAI DevDay 2024 focused on developer-facing API announcements including Realtime API, Vision Finetuning, Prompt Caching, and Model Distillation rather than ChatGPT product announcements
2h 9m
Sep 27, 2024 · 1h 29m
OpenAI's o1 release and recent hiring of Noam Brown and Shunyu Yao signals focus on tool-using chain-of-thought and tree-of-thought architectures for Level 3 Agents
1h 29m
Sep 20, 2024 · 1h 9m
Sander Schulhoff's 'The Prompt Report' synthesizes 1,600+ arXiv papers on prompting techniques including few-shot learning, chain-of-thought, tree search, and self-criticism strategies
1h 9m
Sep 13, 2024 · 2h 4m
Michelle Pokrass and OpenAI's DevRel team cover the entire OpenAI product suite including ChatGPT-latest, GPT-4o, o1 models, and how they're delivered via API with Structured Outputs
Michelle Pokrass
2h 4m
Sep 3, 2024 · 1h 5m
AI inference costs decreased 10-100x in 2024, with open models like Llama 3.1 405B costing $3/mtok versus $30/mtok for Claude 3 Opus, and frontier models dropped 400x from 2022-2024
1h 5m
Aug 29, 2024 · 1h 10m
Nicholas Carlini's 'How I Use AI' blog post demonstrates a practical approach focused on individual AI applications rather than broad AGI potential, covering 12 use cases with specific prompts
Nicholas Carlini
1h 10m
Aug 22, 2024 · 1h 5m
Cosine Genie achieved #1 ranking on SWE-Bench Full, Lite, and Verified using GPT-4o fine-tuning at scale on billions of tokens of synthetic data, beating all other agents including Cognition's Devin
Alistair Pullen
1h 5m
Aug 16, 2024 · 58m
Jeremy Howard's Answer.AI ships 1000s of successful AI products with no managers and a team of 12, focusing on practical AI R&D aligned with GPU-poor needs
58m
Aug 7, 2024 · 1h 3m
Meta's Segment Anything 2 (SAM 2) improves image segmentation accuracy while being 6x faster than SAM 1, and elegantly solved video segmentation with 3x fewer interactions than prior approaches
1h 3m
Aug 2, 2024 · 1h 55m
Q2 2024 AI progress analyzed through Four Wars framework: GPU-rich frontier labs (Claude 3.5, Mistral Large), GPU-rich helping GPU-poors (Llama 3.1 synthetic data, Phi 3, Gemma 2), and on-device LL...
1h 55m
Jul 23, 2024 · 1h 5m
Meta released Llama 3.1-405B, the largest open source model trained on 15T tokens beating GPT-4 on benchmarks, with 8B and 70B models also receiving significant spec bumps
1h 5m
Jul 12, 2024 · 58m
Clémentine Fourrier leads HuggingFace's OpenLLM Leaderboard, which standardizes model evaluation using high-quality benchmarks with reproducible, centralized scoring to replace lab-specific reports
58m
Jul 5, 2024 · 1h 44m
Reka AI achieved #7 on LMsys leaderboard with only 20 employees and $60M funding, demonstrating that top-tier model performance no longer requires massive teams like OpenAI (600) or Google (950+ co...
1h 44m
Jun 25, 2024 · 1h 21m
Databricks' DBRX and Imbue's 70B model outperform GPT-4o zero-shot on reasoning/coding benchmarks while using 7x less data than Llama 3 70B
1h 21m
Jun 25, 2024 · 49m
Raza Habib of HumanLoop hosts High Agency podcast, flipping the interview dynamic with Shawn Wang to discuss the AI Engineer World's Fair and the relevance of the 'Rise of the AI Engineer' essay on...
49m
Jun 21, 2024 · 1h 3m
James Brady (Head of Engineering) and Adam Wiggins (Cofounder Ink & Switch, Heroku) from Elicit share hiring strategies for AI engineers, defining the role as conventional engineers with LLM and pr...
James Brady
1h 3m
Jun 11, 2024 · 54m
Mike Conover, who led OSS models at Databricks and created Dolly, founded Brightwave as an AI research assistant for investment professionals and announced $6M seed round led by Alessio and Decibel
54m
Jun 10, 2024 · 4h 29m
Discusses code editing benchmarks (WebArena, Sotopia), OpenDevin agent framework, and tensions between academic research and industry implementation of AI systems
Aman Sanger Graham Neubig Moritz Hardt
4h 29m

Track New Episodes & Guest Appearances

Subscribe to get AI-powered episode summaries, guest detection, and weekly email digests.

Sign Up Free →