r/MachineLearning
r/MachineLearning is a Reddit community for machine learning researchers and enthusiasts. It features discussions on topics like networking at conferences and the long-term value of AI research.
Profile generated by AI for Anything
![GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]](https://preview.redd.it/tlvz4c3i32eh1.png?width=640&crop=smart&auto=webp&s=aad6aeec9197e26debda00093dd47611e70c5a08)
GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]
2d

The original title is about deep learning for scRNA-seq analysis. Let me rewrite it to be punchy and specific while preserving key facts.
2d
![Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]](https://preview.redd.it/b0u6q9a46ndh1.jpg?width=140&height=98&auto=webp&s=dbb02d2e0fc85305a04e37864167c2578891d46c)
Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]
5d
![PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]](https://external-preview.redd.it/d6rTpW7131dBgTTGfjDXPAIkblduF91pERLr20qfQH4.jpeg?width=140&height=78&auto=webp&s=efc5027d8347fa1bc645f041300b0979e5c14469)
PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]
5d
![The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"](https://preview.redd.it/wvhcgu1q1fdh1.png?width=140&height=88&auto=webp&s=d914b037ae47cf6b529bc66f8af00430d0d590ae)
The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"
6d

Building an XGBoost Pipeline for Cross-Domain Conflict Resolution with Explainable AI
6d
![I trained a vision-language model to play Snake, and so can you. [P]](https://external-preview.redd.it/YWcAyMNI6jxa5S-SYFMgIq4qY5VYLAesOmSGvUtU3as.png?width=140&height=70&auto=webp&s=247aa5820e62edcecbb91eb5a961e10ecb0ae66f)
I trained a vision-language model to play Snake, and so can you. [P]
7d
![[P] RL-training Qwen3.6 to RL-train tool using AI models [P]](https://preview.redd.it/hg7ww6ute8dh1.png?width=140&height=75&auto=webp&s=d9c4aa6843cd8b9f2a480a97e0f8469ec80b4d41)
[P] RL-training Qwen3.6 to RL-train tool using AI models [P]
7d
![LLM hallucination paper(using math) accepted to ICML workshop[R]](https://preview.redd.it/3uyvbtoa76dh1.png?width=140&height=61&auto=webp&s=523d3943b9adbcbbdaca03be35c5e073be075de9)
LLM hallucination paper(using math) accepted to ICML workshop[R]
7d
![Please help me understand figure on subspace similarity in LoRA paper. [D]](https://preview.redd.it/3l5qhbiroech1.png?width=640&crop=smart&auto=webp&s=e5534631f23bcc8d8e89b7fd411120c2b7a84442)
Please help me understand figure on subspace similarity in LoRA paper. [D]
11d

The original title is about IMGNet, a face verification model. Let me rewrite it to be punchy and informative while preserving key facts.
12d
![EMNLP: All of the papers in my review pool being detected as AI [D]](https://preview.redd.it/elmu1a6d0kbh1.png?width=140&height=140&crop=1:1,smart&auto=webp&s=c44181dd02b9668433e47a57a648d98af652fbde)
EMNLP: All of the papers in my review pool being detected as AI [D]
15d
![The original title is "How to get more from your chatbot for less [P]"](https://external-preview.redd.it/qZ4HpTKGOz0DwdVSdExJAW-UpDhB6hu23O6kgI1aNvE.jpeg?width=640&crop=smart&auto=webp&s=0d0ad48f322f4db16ae97f0b70276398c4b3e711)
The original title is "How to get more from your chatbot for less [P]"
17d
![Training transformers where every layer W = V·Uᵀ from initialization reveals a corpus-determined optimal rank - looking for arXiv endorser (cs.LG) [D]](https://external-preview.redd.it/Qfw5SuGCt2d45VbzHurInHB_fbCrPRWPZr4XzFenJcc.png?width=140&height=70&auto=webp&s=6e9379fe0f90d43518578b30abf4563219025786)
Training transformers where every layer W = V·Uᵀ from initialization reveals a corpus-determined optimal rank - looking for arXiv endorser (cs.LG) [D]
17d
![The original title is "Hamiltonian Neural Networks from a Differential Geometry Perspective [D]"](https://external-preview.redd.it/7q8iktqnOmHdHgGNxMCQbvHkXz6extXfcSIuznTr8CA.png?width=640&crop=smart&auto=webp&s=158ee06f289fc1e95a2efb1e71a67adbc515092f)
The original title is "Hamiltonian Neural Networks from a Differential Geometry Perspective [D]"
19d
![P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P]](https://preview.redd.it/v50euf4pymah1.png?width=140&height=81&auto=webp&s=b9a9d3b99087e03cd79f28ebf6ac8622dd9bcc0f)
P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P]
20d
![Loss functions in Instance Representation Learning [R]](https://preview.redd.it/3l7mtxoc3bah1.png?width=140&height=27&auto=webp&s=8426b12f6ec1f44b193529124dee890e0642ad25)
Loss functions in Instance Representation Learning [R]
21d
![I'm trying to implement CALM paper, and I have some questions. [P]](https://preview.redd.it/kr4u22yfx8ah1.png?width=140&height=83&auto=webp&s=784c46c82400e669571b4d8a7dcdc997ad0fba57)
I'm trying to implement CALM paper, and I have some questions. [P]
22d
![NagaTranslate: Building a translation and voice pipeline for low-resource Nagaland creoles (Whisper, VITS, LLMs) [P]](https://preview.redd.it/bu6xsk4hvx9h1.jpg?width=140&height=63&auto=webp&s=0a4c589616d351d10d735940874706494e48d408)
NagaTranslate: Building a translation and voice pipeline for low-resource Nagaland creoles (Whisper, VITS, LLMs) [P]
23d

Attention pathologies stem from norm
26d
![[R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost](https://external-preview.redd.it/q3evP6JeDpAC2MdSQHWYxnCYTqbJkElIQsLFqVSdkss.png?width=640&crop=smart&auto=webp&s=de730fbf7ecace6df0036b21470c16a2d4feacfb)
[R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost
26d
![[R] All Routes Lead to Collapse: attention sinks, representation collapse, and norm stratification are what content-based routing does under a norm-blind metric](https://external-preview.redd.it/q3evP6JeDpAC2MdSQHWYxnCYTqbJkElIQsLFqVSdkss.png?width=640&crop=smart&auto=webp&s=de730fbf7ecace6df0036b21470c16a2d4feacfb)
[R] All Routes Lead to Collapse: attention sinks, representation collapse, and norm stratification are what content-based routing does under a norm-blind metric
26d
![CALHippo - Mapping neurons and glial cells in the human brain hippocampus in 3D using SOTA segmentation and density estimation models [R]](https://preview.redd.it/m8eyacfmbf9h1.gif?width=640&crop=smart&s=1a9d654de34977e02d4c3b3a30f0f9e2d36a5c35)
CALHippo - Mapping neurons and glial cells in the human brain hippocampus in 3D using SOTA segmentation and density estimation models [R]
26d
![High Dimensional, Dynamic Rotary Positional Embedding [P]](https://external-preview.redd.it/Go7zlxhewkLxNN5-ZvZe623w5Zrdi3SXYEIr0JeEGQk.png?width=140&height=75&auto=webp&s=2d3a7ad647024e077a4b7f7b5746c806eba71b8a)
High Dimensional, Dynamic Rotary Positional Embedding [P]
27d