Is ChatGPT Getting Worse? with James Zou - #645
2023-09-04 · 42 min · episode 645 · 16 entities
Asserted relationships
-
→ appeared on The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence podcast0.68
evidence rules-v4
Is ChatGPT Getting Worse? with James Zou - #645
-
0.55
evidence rules-v4
Feed author/publisher: Sam Charrington
-
0.50
evidence rules-v4
professor at Stanford University. In
-
0.50
evidence rules-v4
professor at Stanford University. In
-
0.40
evidence rules-v4
Feed category: Science
-
0.40
evidence rules-v4
Feed category: Technology
-
0.40
evidence rules-v4
Feed category: News
-
0.40
evidence rules-v4
Feed category: Tech News
-
0.40
evidence rules-v4
Feed author/publisher: TWIML
Entities found in this episode
companys 7
-
0.70
evidence rules-v4
Feed author/publisher: TWIML
-
0.62
evidence rules-v4
professor at Stanford University. In
-
0.50
evidence rules-v4
Feed category: Technology
-
0.50
evidence rules-v4
professor at Stanford University. In
-
0.50
evidence rules-v4
professor at Stanford University. In
-
0.40
evidence rules-v4
Feed category: Technology
-
0.40
evidence rules-v4
Feed author/publisher: TWIML
concepts 5
-
0.50
evidence rules-v4
Feed category: Science
-
0.50
evidence rules-v4
Feed category: Tech News
-
0.40
evidence rules-v4
Feed category: Science
-
0.40
evidence rules-v4
Feed category: News
-
0.40
evidence rules-v4
Feed category: Tech News
persons 3
-
0.72
evidence rules-v4
Is ChatGPT Getting Worse? with James Zou - #645
-
0.70
evidence rules-v4
Feed author/publisher: Sam Charrington
-
0.55
evidence rules-v4
Feed author/publisher: Sam Charrington
podcasts 1
-
appeared on The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence podcast0.68
evidence rules-v4
Is ChatGPT Getting Worse? with James Zou - #645
Episode description as stored
Today we’re joined by James Zou, an assistant professor at Stanford University. In our conversation with James, we explore the differences in ChatGPT’s behavior over the last few months. We discuss the issues that can arise from inconsistencies in generative AI models, how he tested ChatGPT’s performance in various tasks, drawing comparisons between March 2023 and June 2023 for both GPT-3.5 and GPT-4 versions, and the possible reasons behind the declining performance of these models. James also shared his thoughts on how surgical AI editing akin to CRISPR could potentially revolutionize LLM and AI systems, and how adding monitoring tools can help in tracking behavioral changes in these models. Finally, we discuss James' recent paper on pathology image analysis using Twitter data, in which he explores the challenges of obtaining large medical datasets and data collection, as well as detailing the model’s architecture, training, and the evaluation process.
The complete show notes for this episode can be found at twimlai.com/go/645.