How to handle cofound variables? [D]
Mirrored from r/MachineLearning for archival readability. Support the source by reading on the original site.
edit: confound
Hello all,
I am working on a object classification with a automotive radar point clouds. I compared many models and feature vectors.
Once i used range as feature, all models scored higher f1 in all K validation sets and on the final test set.
One particular artifact of a radar, is that as the farther the object is the less number of points it returns to the radar. Although the performance improved and there is no overfit in the classical sense, i am afraid my model is learning the environment not the class distribuiton and even worse, its learning that big range means big object.
How can i stress test this claim? Should i try to split the data sets so range distribution differs? Or not even using the feature at all and accept lower performance?
Would appreciate your insights.
Thank you.
[link] [comments]
More from r/MachineLearning
-
Understanding and Enhancing Kimi Delta Attention [R]
Sep 22
-
Xiaomi releases MiMo-V2.6: "Frontier intelligence, all the modalities, built in public." [N]
Sep 22
-
Paper on ArXiv for a year now, should I disclose about it in ICLR submission? [Discussion]
Sep 22
-
Jev's calibration was measured. The LLMs won [D]
Sep 21
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.