Do GFlowNets Transfer? Case Study on the Game of 24/42

Published 3 Mar 2025 in cs.AI and cs.CL | (2503.01819v1)

Abstract: Generating diverse solutions is key to human-like reasoning, yet autoregressive LLMs focus on single accurate responses, limiting creativity. GFlowNets optimize solution generation as a flow network, promising greater diversity. Our case study shows their limited zero-shot transferability by fine-tuning small and medium-sized LLMs on the Game of 24 and testing them on the Game of 42 datasets. Results revealed that GFlowNets struggle to maintain solution diversity and accuracy, highlighting key limitations in their cross-task generalization and the need for future research in improved transfer learning capabilities.