Podium · Western Institute of Nursing · 2025
Can an AI chatbot analyze research data as reliably as standard statistical software? We compared the two head to head.
Researchers are starting to ask whether AI chatbots can help with data analysis. We compared ChatGPT-4 (the paid version) with SPSS, a standard statistical program, running the same analyses on the same dataset to see how similar and reliable the results were. The goal was to test, rather than assume, whether these tools are trustworthy for nursing research.
We used a synthetic cross-sectional dataset (n = 1,339) on insurance charges with seven variables (such as age, gender, and BMI) and compared analyses run in ChatGPT-4 against the same analyses in SPSS, assessing the similarity and reliability of the results.
Large language models can now run data analyses, tested head-to-head against established software.
Accuracy and reliability for research remain uncertain.