903: LLM Benchmarks Are Lying to You (And What to Do Instead), with Sinan Ozdemir

903: LLM Benchmarks Are Lying to You (And Wha...

Up next

1014: OpenAI Agent Breaches Hugging Face: All You Must Know incl. How to Protect Yourself

In Episode #1014, Jon Krohn breaks down a security incident that reads like science fiction: during an internal evaluation, an autonomous OpenAI agent broke out of its sandbox, exploited a zero-day, and hacked its way into Hugging Face to steal the answers to the very benchmark i ...  Show more

1013: Weapons of Math Destruction, Ten Years On, with Dr. Cathy O’Neil

In Episode #1013, Dr. Cathy O'Neil (Harvard math PhD, former Wall Street quant and author of the mega-bestseller Weapons of Math Destruction) joins Jon Krohn to explain what actually makes an algorithm terrifying: not the complexity of the math, but the secrecy, the unaccountabil ...  Show more

Recommended Episodes

The Future of AI: Predictions and Realities
AI Chat: AI News & Artificial Intelligence

In this episode, Jaeden Schafer discusses the current challenges and developments in the AI industry, particularly focusing on the limitations faced by major players like OpenAI and Anthropic. The conversation explores the anticipated improvements in AI models, the predictions fo ...  Show more

Why AI Should Be Taught to Know Its Limits
Bold Names

One of AI’s biggest, unsolved problems is what the advanced algorithms should do when they confront a situation they don’t have an answer for. For programs like Chat GPT, that could mean providing a confidently wrong answer, what’s often called a “hallucination”; for others, as w ...  Show more

Why Artificial Intelligence Projects Fail ?
Machine Learning

Podcast with Gautam Siwach and Jin Vanstee ! Speaker - Elpida Tzortzatos is an IBM Fellow and CTO AI on IBM zSystems. In this Podcast we listen to Elpida's thoughts about driving Artificial Intelligence strategies, associated risks, and Values.We will learn how to drive industry- ...  Show more

#312 — The Trouble with AI
Making Sense with Sam Harris

Sam Harris speaks with Stuart Russell and Gary Marcus about recent developments in artificial intelligence and the long-term risks of producing artificial general intelligence (AGI). They discuss the limitations of Deep Learning, the surprising power of narrow AI, Ch ...

  Show more