230 points by stefanpie 5 days ago | 119 comments | View on ycombinator
fn-mote 2 days ago |
greenflag 2 days ago |
syntamono 1 day ago |
c7b 2 days ago |
Actually, that sounds like an interesting idea for peer review in general, to include an interview between referees and authors. If it saves one round of rebuttals/reactions, it needn't even consume a lot more of everyone's time if you're doing those things properly. What it would undermine would be blindness, but something's gotta give, and it was already on its way out.
mlmonkey 2 days ago |
Personally, I would love to see a conference where people are explicitly encouraged to use LLMs for doing the work and writing the papers, and LLMs are used to review them too.
danieltanfh95 1 day ago |
Actors who already went through the process (or otherwise) gained sufficient reputation or credentials to self-market their own paper can skip journals entirely. That was what OpenAI did. With a sufficiently powerful AI model and correctional pipelines, generating a paper is trivial given some insight.
I think we should be reminded that papers are a channel to distribute papers. Editors are unpaid now, but the economics of a journals are such that the editors are incentivized to curate or distribute papers to schools that pay for the paper. Some perceive quality as a core metric for this. However, in my experience of dealing with computational biology, a paper in so and so journal hardly means a stamp of quality as compared to a paper in some github repo with code to reproduce the paper. This, simply, is broken, because journals and peer reviewers cannot guarantee that data and results in the paper is correct (assuming that it is not maths or theoretical) without reproducing the results in the paper.
None of these are helpful towards students who are already struggling to keep up with the cadence of producing papers.
JasonCEC 2 days ago |
WCSTombs 2 days ago |
> LLMs may be used as general-purpose assistive tools. Whichever tools are used, authors are fully responsible for content on which they are listed as (co-) authors. This includes, but is not limited to, content generated by LLMs that could be construed as plagiarism or scientific misconduct (e.g., fabrication of facts). Low-quality contributions (be they submissions or reviews) that appear to be largely LLM-generated will be closely examined for evidence of the issues mentioned previously, such as scientific misconduct. LLMs are not eligible for authorship. We will periodically revise this policy as new information about the use of LLMs in the scientific process becomes available.
While it doesn't outright encourage using LLMs, it's right at the door, and IMO a policy this weak is actively contributing to the problem the article's author is complaining about. In my opinion any policy weaker than "using LLMs to generate any part of your submission is not allowed and considered a serious breach of ethics" is insane. People like to say that such policies are unenforceable, but that's really not the point (at first), since there are other things like (somewhat ironically) p-hacking that are pretty hard to detect but still widely recognized as unethical. We haven't exactly solved p-hacking either, but at least most of us can agree that p-hacking should be eliminated.
It's hard for me not to read between the lines here. Maybe it's the tinfoil talking, but it being a machine-learning journal, it probably embodies a generally pro-AI philosophy, and thus may not want to discourage too much of it...
It may also be worth noting that this journal apparently uses AI itself on the reviewing side [2]. I'm not claiming this is super unethical or anything as long as the main review is human (although I have concerns), it probably should be part of the conversation.
[1]: https://jmlr.org/tmlr/editorial-policies.html
[2]: https://medium.com/@TmlrOrg/ai-reviews-at-tmlr-for-assessing...
wackget 2 days ago |
Is it standard practice for authors to have to defend their submissions via interview like this? If not, why not?
Does the vetting process vary with the quality of the publisher?
As an outsider, it's extremely worrying that anyone would even attempt to submit an AI-generated paper for publication in an academic journal. At that level I would have assumed literally everybody should know better than to even try.
figassis 1 day ago |
If you cannot answer, you did not author the paper, meaning you are misrepresenting your contribution, and there is already a process for this. And this is actually a really good test for any field. Use AI as much as you want, but you need to be able to explain your work. Applies to SWE as well, you need to understadn what you built, at the code level and system level.
sampo 2 days ago |
He is not an unpaid volunteer.
He's an associate professor at the prestigious Carnegie Mellon University. He is not paid by the journal, but he is paid a salary by the university, and the university expects that a small part of his academic work is to serve as an editor in academic journals.
softwaredoug 2 days ago |
jszymborski 2 days ago |
EDIT: I found a live link on arxiv https://arxiv.org/html/2609.20481v1
Havoc 1 day ago |
That's going to crater the signal to noise ratio of papers so best that be fixed asap. How though...
WD-42 1 day ago |
Calazon 2 days ago |
N_Lens 2 days ago |
The article highlights how only one out of ten paper’s authors were able to answer questions thoroughly and at a high level. This indicates an overwhelming percentage of authors are slopping up their work with AI and submitting it without even reading it.
pc86 1 day ago |
Short of just the general "vibes" type of reputation that follows someone around this type of behavior seems pretty low risk, which is a large part of why people engage in it. Perhaps the risk should increase a couple orders of magnitude to stop it from happening.
ChrisMarshallNY 1 day ago |
I’m wondering if anyone just flat-out admits they used LLMs in their work. I have no problem, doing that, myself, but I also have the luxury of not having my livelihood on the line, and not being overly-concerned about what people think of me.
Eventually, I assume that AI will affect every aspect of the industry, and it will actually be a signal of effectiveness, to claim its use. I could see a day, when claims of not using AI would be like artisans, declaring their work to be “genuine hand-crafted,” and relegated to fairly small, specialized corners of the industry.
lh712 1 day ago |
A year or so ago I completed a PhD in theoretical high energy physics. I did that after working for about 5 years outside of academia [the reasons for that were financial issues in my family]. Because of this hiatus (I suppose) and the fact that I had troubles getting reference letters (my master thesis advisor passed away at quite a young age right at the time when I tried to start applying for positions; we actually agreed to meet in person to discuss next steps but the meeting never happened) I had issues getting accepted into PhD programs and I ended up with an offer from a rather weak institution in my home country (which is in EU). I had no other options to choose from (and I could not wait another year) so I accepted. However, I also looked at the publishing record of the team that I was about to join. Their works did not seem stellar (and I did not expect that) but they seemed to publish regularly, on several topics, and with several collaborating institutions.
When I started I quickly realized that the knowledge/expertise level was much lower than I expected (and I did not expect too much). The group consisted of the group leader and three senior researchers, and I am pretty sure that my knowledge of QFT when I *started* working with them was the best in the group. (Now, I had studied in my own time during some part of the years I was outside of academia, and I think my knowledge at that time exceeded that of an average 1st year hep-theory PhD student in Europe, but it was also nothing spectacular; definitely not what I would call "competent"; which is admittedly a high bar in quantum field theory / hep theory, but something I would expect senior researchers to more or less satisfy.)
I realized that the one topic on which the group was publishing on their own was just a continual rehash of the same thing (applied to various different problems, so perhaps not completely useless) and the other topics, those which seemed more advanced when I originally looked at the publication record, were all done essentially outside the group by other people. The group members contributed with some non-essential help (often resembling a work done by a student, such as finalizing a manuscript or double checking calculations) or with nothing; and were added as coauthors due to some other reasons (I suppose: past loyalties, friend groups, advantage of having foreign or external institution co-authors). [Note: I was not included into any of those collaborations so I am not speaking with a full knowledge of the inner workings there.]
During this time I saw that people being added as coauthors had a very noisy relation to what they did or didn't do on any given paper. I was added as a co-author on a paper that I made essentially no work on [not fair], I was added as a co-author on papers which drew upon some of my earlier results [fair, but I did not subscribe to the overall spirit of the paper or even to the paper being published at all], I was a co-author of papers where number of authors and my positions in the list roughly corresponded to my contribution [as it should be], I was a co-author of papers that were done basically by myself but there were several other co-authors included who made a little to no contribution [and the ordering of the author list was always alphabetical not reflecting the relative contribution].
There is also another aspect of this: If you are an early career researcher (PhD student, PostDoc, and even non-tenured professor) you might have a very limited control over who gets included on the publications you work on, and on what publications your name appears. In my case, I am a very disagreeable person when I think things are being done wrongly and yet I have not managed to refuse from being included on papers I did not want to be a co-author of, or to prevent people who contributed nothing to be included on my papers. [I mean the only way to achieve that meant escalating conflict into levels which (a) would likely disable any further cooperation with the group, and (b) might be even morally questionable concerning the level of distress it would make to other people who just seemed to be fine about "the normal way" things were being done.]
In any case, while there are still some really good people in academia (actually the best people I met in my life were nearly all in academia), in overall (by no means fully indicated in the above paragraphs) I feel very bitter about it, and I think it is so dysfunctional and morally corrupted, that it is nearly inevitable that it collapses in future. The current and future shockwaves of the change in the public sentiment and funding, the evolution of demographics, and the consequences of AI, just hasten the process that would have probably unfolded, sooner or later, anyway.
cgio 1 day ago |
dovholuknf 2 days ago |
logicallee 1 day ago |
These days agents are able to really perform genuine experiments and write up the results. A prompt like this: "You'll work autonomously end to end to select a research task that meaningfully advances the state of the art in AI, is clearly defined and worth performing, that people would be interested in reading, and that you can perform on this hardware" (insert details) " in a week. Carefully log your steps so that your results can be replicated. Then, do a research review and write your paper about it up with correct, cited references. You must check all of your citations. Look up current lists of "Claudisms", (such as use of the word "genuinely", or "load-bearing"), and remove them from your writeup. After writing your writeup, edit it and pare it down, remove anything unnecessary, keep it fast paced and interesting. Also, try to tell a story, be engaging in your writeup. Don't use violent metaphors, remove references to killing, strangulation, etc. Your writeup should be ready to publish and accurately reflect a real experiment with a meaningful result that advances the state of the art and contributes to understanding. Be concise and focus on why it matters."
Okay, so there's the prompt. You can give it to any AI and have a journal-ready publication in a week. I guess you can ask it to add charts and stuff, if you want to be fancy.
If I gave my agent the above prompt, would I be one of the authors? Maybe it's fair to say I guided, facilitated, elicited, or advised it. But it's clear that the AI would be the one that is actually selecting and running the experiment and writing up the results.
Someone could probably get a publication without even reading the paper they wrote their name on. Their only contribution might be editing their name into the PDF.
muh_gradle 1 day ago |
trombuance 1 day ago |
I would keep a private blacklist (shadow ban) the authors who wasted several hours of a reviewer's time to prove they were not legitimate. The existence of such a list would be problematic, though.
Could the same system we use here be applied? Accepted authors could "vouch" for "dead" papers in case they were "auto-killed"?
This system is broken and providing more evidence that it is broken isn't much of a step towards fixing it.