"Tomatoes extend your life by three years," "Limit eggs to one per day," "Red wine contains polyphenols"—these famous pieces ...
Data-peeking, the practice of stopping data collection once significant results are obtained, poses a significant threat to the credibility of research by increasing the likelihood of observing and ...
AI与P-Hacking:自动化造假?当大语言模型(LLMs)作为数据分析助手时,它们是提高了研究效率,还是帮我们自动化了p-hacking?本期视频基于最新前沿工作,详细拆解针对 Claude Code 和 Codex 的统计 ...
Luskova, M, N Buliskeria, A Elminejad, T Havranek, Z Irsova, Š Jurajda and M Kapicka (2026), ‘DP21630 Publication Bias and P-Hacking in the Effect of COVID-19 on Learning‘, CEPR Discussion Paper No.
跑了一整夜,P 值还是 0.5?先别急着怀疑自己是“学术垃圾”,可能只是你还没掌握统计学里的“黑魔法”。 为什么同一份数据,有人跑不出结果,有人却能连发 4 篇顶刊? 为什么喝咖啡和 ...
According to Ethan Mollick on X, Huaxiu Yao cautioned that while AutoResearchClaw—an automated system that turns a single prompt into a full research paper with experiments, citations, and code—shows ...
According to God of Prompt (@godofprompt), the AI research industry faces a systematic problem of benchmark overfitting, with 94% of studies testing on the same six benchmarks. Analysis of code ...
科技日报柏林7月28日电 (记者李山)德国人工智能研究中心(DFKI)研究团队在日前召开的国际机器学习大会上报告称,在可解释人工智能(AI)领域,“X-hacking”是一个此前被普遍忽视的风险 ...
Large language models (LLMs) like ChatGPT are becoming deeply embedded in academic research workflows, and with their rise comes a new threat to scientific integrity - prompt-hacking. In a new opinion ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results