

Recent experiments with ChatGPT's Advanced Voice Mode and Llama 3.1-based chatbots have yielded intriguing results. Cris Giardina tested the ability of ChatGPT Advanced Voice Mode to distinguish between a human and a machine by asking five questions. A similar test was conducted with a Llama 3.1-based chatbot, which was instructed to mimic a human. The outcomes of these tests were surprising. Additionally, Llama 3.1 405B outperformed ChatGPT and Claude 3.5 Sonnet in self-reflection and logical reasoning tasks. In particular, it successfully answered tricky questions such as whether '9.9 is larger than 9.11' and 'how many r’s are in strawberry'. Furthermore, a call featuring Cris Giardina and Diego Cabezas highlighted the abilities and limitations of ChatGPT's Advanced Voice Mode, including attempts to make it sing.