Apple just savaged LLM reasoning claims (again)
"Complex problems exhibit consistently near-zero accuracy, indicating complete reasoning failure"
"Complex problems exhibit consistently near-zero accuracy, indicating complete reasoning failure"