"Tomatoes extend your life by three years," "Limit eggs to one per day," "Red wine contains polyphenols"—these famous pieces ...
Data-peeking, the practice of stopping data collection once significant results are obtained, poses a significant threat to the credibility of research by increasing the likelihood of observing and ...
AI与P-Hacking:自动化造假?当大语言模型(LLMs)作为数据分析助手时,它们是提高了研究效率,还是帮我们自动化了p-hacking?本期视频基于最新前沿工作,详细拆解针对 Claude Code 和 Codex 的统计 ...
Luskova, M, N Buliskeria, A Elminejad, T Havranek, Z Irsova, Š Jurajda and M Kapicka (2026), ‘DP21630 Publication Bias and P-Hacking in the Effect of COVID-19 on Learning‘, CEPR Discussion Paper No.
跑了一整夜,P 值还是 0.5?先别急着怀疑自己是“学术垃圾”,可能只是你还没掌握统计学里的“黑魔法”。 为什么同一份数据,有人跑不出结果,有人却能连发 4 篇顶刊? 为什么喝咖啡和 ...
According to Ethan Mollick on X, Huaxiu Yao cautioned that while AutoResearchClaw—an automated system that turns a single prompt into a full research paper with experiments, citations, and code—shows ...
According to God of Prompt (@godofprompt), the AI research industry faces a systematic problem of benchmark overfitting, with 94% of studies testing on the same six benchmarks. Analysis of code ...
科技日报柏林7月28日电 (记者李山)德国人工智能研究中心(DFKI)研究团队在日前召开的国际机器学习大会上报告称,在可解释人工智能(AI)领域,“X-hacking”是一个此前被普遍忽视的风险 ...
Add a description, image, and links to the p-hacking topic page so that developers can more easily learn about it.
Large language models (LLMs) like ChatGPT are becoming deeply embedded in academic research workflows, and with their rise comes a new threat to scientific integrity - prompt-hacking. In a new opinion ...