Summary
As generative AI becomes increasingly integrated into research workflows, what does "human in the loop" actually mean? This latest installment in the ARF and MSI's
Psychology of Gen AI series explores how researchers and large language models can work together to develop rigorous research designs—and why human judgment remains indispensable throughout the process. By comparing multiple, AI-generated experimental designs and refining them through expert evaluation, the study demonstrates that AI is most valuable as a collaborator that expands possibilities, not as a replacement for methodological expertise.