r/MachineLearning
500 articles archived · Visit source ↗ · RSS
-
r/MachineLearning community 5d ago
LinearSolveBench: new benchmark for linear solvers [P]
LinearSolverBench measures the ability of a model or harness to write fast, accurate, and general numerical solvers for large sparse linear systems in C. The goal is to encourage algorithmic advances in numerical methods for solving linear systems of equations.…
27 -
r/MachineLearning community 5d ago
QontoFAQ: A better Information Retrieval Benchmark [R]
Retrieval benchmarks sometimes feel benchmaxxed by models, so we wanted to find a way to tie it as close as possible to my objective: finding the article that answers a product question right . We worked on a new metric which seems more proportional to document relevance, and…
8 -
r/MachineLearning community 6d ago
Understanding and Enhancing Kimi Delta Attention [R]
TLDR: We demonstrate and explain the difference in expressivity of Gated Deltanet (GDN) and Kimi Delta Attention (KDA). We show how the full diagonal gate in KDA can act as a reflection allowing 2D rotations to be carried out in a single step, but only if the range of the gates…
13 -
-
r/MachineLearning community 6d ago
Jev's calibration was measured. The LLMs won [D]
Source: Jev Benchmarks Its training method is literally called "Reinforcement Learning for Calibrated Decisions." Calibration gap vs human labels (lower = better): Yes/no: Jev 5.0, Gemini 3.8 Flash 2.0 Pick-one: Jev 9.8, DeepSeek V4.1 Flash 2.8 Rubric: Jev 19.7, GLM-5.3 12.9 It…
29 -
r/MachineLearning community 6d ago
Systems for Machine Learning[D]
I’m a computer engineering graduate and come from a traditional embedded systems background, with knowledge of microcontrollers, computer architecture and operating systems. Is knowledge of C and C++ programming, Linux networking, memory management , multithreading,…
34 -
r/MachineLearning community 7d ago
Concerns about the ICLR review policy [D]
The ICLR review policy says if your name appears on 3 or more papers, you will need to serve as a reviewer, and it did not say anything about qualifications. I only have 2 so this doesn't apply to me, but I am not sure if I understand it correctly and thus have some concerns. So…
27 -
r/MachineLearning community 7d ago
How is your experience with ICLR LLM Feedback? [D]
Besides the ridiculous number of submissions, interesting to hear your experience with it. For me it had 1-2 valid points, and 3 pages of nitpicking. I have the time, so I can address both types of issues, but wonder if that’s your experience as well. Overall, I think it is an…
19 -
r/MachineLearning community 7d ago
Zero-shot Neural Style Transfer (NST) App [P]
Hi everyone! I spent the last 6 days (today included) developing and deploying a zero-shot nueral style transfer (NST) application using AdaIN (Huang and Belongie, 2017). Feeling kinda proud of it :) if y'all would like to give it a spin you can find it at stylyze.app . Just a…
33 -
r/MachineLearning community 7d ago
Typesafe's JEV model work as an LLM [P]
I built a conversational AI that doesn't generate a single token — it selects from 400 pre-written responses using TypeSafe's Jev, a non-generative model that returns probabilistic judgments instead of text. The technical approach: Traditional LLMs generate responses token by…
25 -
r/MachineLearning community 7d ago
Autograd project [P]
Hello, I'm a 3rd year Highschooler interested in machine learning and for the last few weeks have been working on a small project meant to learn the basics of machine learning. I have implemented a simple tensor library and autograd in c++. It's very simple but i want some…
31 -
r/MachineLearning community 8d ago
Inside sanoTTS — a 294,279-parameter TTS system [P]
How sanoTTS works? I have vibe coded this site to show what's inside sanoTTS? Every tensor shown on the page is a real intermediate value captured from the shipped int8 model while it synthesized an actual sentence; no mock-ups, no stand-in data. Just check this out:…
35 -
-
-
r/MachineLearning community 8d ago
I wanted to watch a neural network learn [P]
I wanted to really see how a neural network learns different functions, so built an interactive demo. You can change the architecture of the network and the function it will try to approximate. A fully-connected network with ReLU activations will create a piecewise linear…
15 -
r/MachineLearning community 8d ago
World Models From Scratch 2: Model Training and Dreaming [P]
I am adding self-contained and accessible videos on how World Models work as well as how you can create one! This is part 2 which gets you to the exciting place where you can play a gameboy goy entirely in a world model!   submitted by   /u/Available_Pressure47 [link]…
17 -
-
r/MachineLearning community 9d ago
DiffusionGemma: How It Generates Text in Parallel (From Scratch in PyTorch) [P]
  submitted by   /u/Winter_Mistake_3185 [link]   [comments]
26 -
r/MachineLearning community 9d ago
JMLR submission experience [D]
Hi just wondering has anyone submitted anything to JMLR before? Especially in the last 2 years? What is the experience like? Background: I am a Comp Sci PhD student, but my secondary supervisor (the one who is actually looking after me) is from the Stats faculty. He has 0 Comp…
8 -
r/MachineLearning community 9d ago
How is RLCD (jev) RL? [D]
Just saw the YouTube presentation and I was left wondering this question. If jev only outputs Choice, Score, or Noul … well those are all perfectly differentiable. (Cross entropy or mse) I don’t know if I’m missing something or if adding RL is just for marketing. Like what would…
34 -
-
r/MachineLearning community 9d ago
How competitive are journals compared to top ai conferences? [D]
Hi everyone, The NeurIPS results aren’t out yet, but I’m expecting my paper to be rejected, so I’m considering submitting it to a journal instead. My scores were 2/3/3 (3/4/4). I feel like I would probably get a similar outcome if I submitted the paper to ICLR or CVPR, so at…
30 -
r/MachineLearning community 10d ago
augmenting large datasets to have more edge case data for training [D]
I have this idea I'd like feedback on. most camera footage for training, is sunny daytime, because that's what cameras record most of the time. The edge cases models actually struggles with, like night, fog, rain or glare, are rare in the data - the idea is to augment it to have…
25 -
r/MachineLearning community 10d ago
AAAI-27 Phase 1 Results [D]
AAAI-27 Phase 1 results are expected on September 24. Anyone else waiting for the decision? Would be useful to keep this thread for updates when people start receiving notifications.   submitted by   /u/BeneficialFish04 [link]   [comments]
22 -
r/MachineLearning community 10d ago
Seeking RA position in the US (Causal/Robust ML) [R]
I never thought it’d come to this, but here we are. I’m a Master’s graduate in the US with solid research experience, and I’m currently looking for an RA/research position. I need to maintain my visa status, so I’m trying to find something ASAP. I haven’t been getting many…
31 -
r/MachineLearning community 10d ago
XGBoost vs Human Markets [P]
What is the generally thought of as the upper limit of the predictive power of XGBoost vs an aggregate of humans? Right now I feed the model the same information that the human market has access to, and the model gets crushed on Top 1 accuracy, it closes the gap a bit but is…
15 -
r/MachineLearning community 11d ago
ICLR 2027 table font sizes [D]
I am preparing an ICLR 2027 submission using the official LaTeX style. Some main-text tables use \footnotesize, and two wide tables also use \resizebox{\linewidth}{!}{...}. I have not modified the style file, margins, page dimensions, or global font settings. The official…
6 -
r/MachineLearning community 11d ago
Question about TMLR [D]
I have a submission under review in TMLR. Less than a month after submission, I have already received 2 reviews. However, a month has passed since those two reviews, and I still haven't received the third. Is this normal? I have already made the suggested changes, but the…
23