r/MachineLearning
500 articles archived · Visit source ↗ · RSS
-
r/MachineLearning community 11d ago
GoBench: Evaluating LLMs on the game of Go [R]
GoBench evaluates LLMs on 9x9 Go games against a ladder of KataGo opponents, from random to superhuman. It measures general reasoning ability, strongly correlates with ARC-AGI 2 (r=0.83 correlation), and remains highly unsaturated. GPT-6 Astra max achieves 2500 Elo, much lower…
7 -
r/MachineLearning community 11d ago
NeurIPS Reference Check Response[D]
I have received the email for the NeurIPS E&D track reference review checker mentioning the 2 hallucinated references. Where do we respond to this email about hallucinated references? In the openreview, as a comment, or in the mail we have received, mentioning the hallucinated…
31 -
r/MachineLearning community 11d ago
PhD in ML in the AI era, autoresearch [D]
Hello everyone, I am currently a first year PhD student in AI/ML. I am just wondering if I am not wasting my time doing a PhD in academia vs going to a research startup/company department directly. Given the rate of AI progress, and the needs in terms of infrastructure, I feel…
7 -
r/MachineLearning community 11d ago
LARA: small, composable behaviours for frozen LLMs [P]
GitHub: https://github.com/pfekin/LARA I've been working on LARA (Lightweight Additive Residual Adaptation), a research project on making post-training modular for frozen language models. I've also developed a small PyTorch library that implements it. The main idea is to train a…
17 -
r/MachineLearning community 12d ago
TabPFN-3.5 is released as the next SOTA tabular foundation model [N]
Prior Labs released their latest tabular foundation model, TabPFN-3.5 today. The model is top of both TabArena and BeyondArena and SOTA for 1M rows and up to 20k features It comes with: - TabPFN-3.5-Fast (in alpha): This one goes 6x faster than the base model -…
11 -
r/MachineLearning community 12d ago
NeurIPS 2026: handling of multiple venue locations seems bad [D]
There has been a recent post acknowledging that NeurIPS passes for Sydney sold out in minutes. For paper authors (guaranteed 1 pass at their designated location) — we recently received forms to select a preferred venue, but apparently we’re not guaranteed to present there. I…
10 -
r/MachineLearning community 13d ago
How much work in progress can a workshop submission be [R]
Hi, let's suppose I am working on an algorithm that uses principles x to solve problems A and B. I already implemented a very basic algorithm that used principle "x mini" to just solve problem A, ran experiments, but have not yet implemented the full one to solve A and B. I must…
26 -
r/MachineLearning community 13d ago
[D] How do you get preprocessed dataset of a paper [D]
Hi all, I'm trying to reproduce a paper where the reported dataset statistics in Table 1 don't match what I get from the public raw data, even after implementing the preprocessing exactly as described. I've tried all reasonable interpretations of the filtering described in the…
18 -
r/MachineLearning community 13d ago
RSI is not happening [R]
A new paper (I'm not a coauthor BTW -- I just found it interesting) argues, basically, that RSI is not on the horizon, because current (at the time the study was done) agents cannot do open-ended ML research. Specifically, they took some accepted, but unpublished papers from…
17 -
r/MachineLearning community 13d ago
[P] Wine synthesis using VAE [P]
I have created a VAE model using PyTorch on White Wine dataset. Basically, the main goal is to discover a brand-new white wine recipe. It puts all the wines into a latent space, finds the best part where higher bands are located, and then it makes 100 steps with a step size of…
36 -
r/MachineLearning community 13d ago
MS MARCO click-translation expansion tables ("poor man's" DSSM) [P]
TLDR: I made "poor man’s" DSSM (Deep Structured Semantic Model) — the count-based translation table that can enrich the inverted index for full-text search. This trick can improve baseline BM25. So the idea is the following: - You have supervised pairs (query, relevant…
15 -
r/MachineLearning community 13d ago
Duplicating baseline benchmarks [D]
Suppose I create two machine learning models suppose tree and neural network for a task let's suppose regression problem, now suppose I am sending both of this paper to two different journals, now the thing is the baseline models I need to only run once because I have reported…
35 -
-
r/MachineLearning community 14d ago
PhD branding question [R]
I'm starting a PhD where I will be doing Graph ML (somewhere along the lines of graph signal processing/ graph deep learning.) My eventual goal is research scientist at big tech, or whichever company has a strong research division, where I can continue similar AI/ML work. I have…
36 -
-
r/MachineLearning community 14d ago
[Upcoming AMA] Waymo AI Team AMA – Drop Your Questions Early! [D]
Hi r/MachineLearning , Join our AI leads as they answer your questions on foundation models, simulation, and scaling the Waymo Driver. Our AMA thread is officially open, and you can start dropping your questions now. From multimodality and end-to-end architectures to the…
20 -
-
-
r/MachineLearning community 15d ago
Getting Mac Air m2/m3/m4 is it good [D]
So want to switch to macbook after running epochs on my run down vivobook which still works pretty well thanks to its rtx 3050. Now my budget us nearly 80,000-1lac rupees(800-1000$) and buying it second hand is also an option. I could have asked this in any apple Reddit pages…
12 -
r/MachineLearning community 15d ago
How much do tech reports matter for a PhD application? [D]
The title, by tech reports I don't mean arXiv submissions, but reports of a large model, like say Kimi K3, DeepSeek, Gemini, Mistral Leanstral, etc. Is it much above, above, much below, below or equal to a first author A* paper?   submitted by   /u/simple-Flat0263 [link]…
27 -
-
r/MachineLearning community 16d ago
Confusion regarding EMNLP registration [D]
Hi, I posted some months ago and got some very helpful responses (for another conference), but I have some confusions regarding EMNLP (and the way registration works here). To give some context, I used to work as an intern during my undergrads at an Indian uni (final year), and…
36 -
r/MachineLearning community 16d ago
How to handle cofound variables? [D]
edit: confound Hello all, I am working on a object classification with a automotive radar point clouds. I compared many models and feature vectors. Once i used range as feature, all models scored higher f1 in all K validation sets and on the final test set. One particular…
25 -
-
r/MachineLearning community 17d ago
Neurips 2026: site selection email [D]
We just received the email for site selection for our neurips paper. Although it is obviously not an acceptance decision, I wonder whether every single non-withdrawn submission received this email, or this might hint towards a higher acceptance chance for our paper?  …
15 -
r/MachineLearning community 17d ago
ACL Sustainable Reviewing Policy [D]
ACL just announced on X how they are planning to handle the increased submission numbers. Interestingly enough, they call it "the proposal". My understanding is that, in a nutshell, each submission should come with someone who can review, otherwise it may only get a slot through…
21 -
r/MachineLearning community 17d ago
Any tools to turn a codebase into a fine tuning dataset? [D]
I have a few web projects with pretty good UI/UX and I’m wondering if there’s any tool or workflow that can turn an existing codebase into a dataset for fine tuning. For example, given a React/Next.js project with components, pages, styling, etc. or a static html site, I’d like…
7 -
r/MachineLearning community 17d ago
Why is TMLR so slow in recent times [D]
A final-year PhD student here. A few months back, I submitted a solo-authored paper to TMLR. The reviewers were on time and extremely positive, with some minor revisions. After submitting the revised version, there was absolute silence from the reviewers, with just one…
7 -
r/MachineLearning community 17d ago
Anybody working on Test Time Training over here? Lemme work with u pls [D]
I'm an undergrad student who got a taste of research. I love it. I currently have a draft, which me and my mentor have planned for TMLR, and plan to submit it by next month for the first round of review. It was some work on self explanation methods of LLM models. We are…
29 -
r/MachineLearning community 18d ago
What Sante's 83.83 on DiagnosisArena-MCQ actually measures [D]
Ant Ling reports 83.83 on DiagnosisArena-MCQ for Ling-3.0-flash-Sante, its new medical reasoning model. The suffix matters: the task provides case information, examinations and tests, then asks the model to choose from four diagnoses. That result tells us about selecting an…
33 -
r/MachineLearning community 19d ago
Teach ML! Community service project from Stanford [N]
Hi r/machinelearning . Nice to meet you! My name is Chris Piech and I'm a professor at Stanford University in the AI lab. I built a class called Probability for AI: pai.stanford.edu. It starts Oct 9th and applications are due end of Sept. Its (hopefully) cool for a few reasons:…
25 -
r/MachineLearning community 19d ago
ECCV 2026 Social Groups [D]
Hi. I'm visiting ECCV in Malmo and was wondering if there's any medium, like discord, whatsapp etc where people are discussing social activities. I couldn't find a way to connect to people on the official app to discuss common interests, research or otherwise. It'd also be nice…
17