Is ChatGPT Good at NYT Connections?
Dev Detour · 9:18
Large language models are still far worse than a typical human at *New York Times* Connections: after more than 350 puzzles solved one category at a time with game-like feedback, even first-place GPT-4o only won 91 ga...