r/MachineLearning
r/MachineLearning

r/MachineLearning

r/MachineLearning is a Reddit community for machine learning researchers and enthusiasts. It features discussions on topics like networking at conferences and the long-term value of AI research.

Profile generated by AI for Anything

GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]

GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]

2d

The original title is about deep learning for scRNA-seq analysis. Let me rewrite it to be punchy and specific while preserving key facts.

The original title is about deep learning for scRNA-seq analysis. Let me rewrite it to be punchy and specific while preserving key facts.

2d

Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]

Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]

5d

PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]

PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]

5d

The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"

The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"

6d

Building an XGBoost Pipeline for Cross-Domain Conflict Resolution with Explainable AI

Building an XGBoost Pipeline for Cross-Domain Conflict Resolution with Explainable AI

6d

I trained a vision-language model to play Snake, and so can you. [P]

I trained a vision-language model to play Snake, and so can you. [P]

7d

[P] RL-training Qwen3.6 to RL-train tool using AI models [P]

[P] RL-training Qwen3.6 to RL-train tool using AI models [P]

7d

LLM hallucination paper(using math) accepted to ICML workshop[R]

LLM hallucination paper(using math) accepted to ICML workshop[R]

7d

Please help me understand figure on subspace similarity in LoRA paper. [D]

Please help me understand figure on subspace similarity in LoRA paper. [D]

11d

The original title is about IMGNet, a face verification model. Let me rewrite it to be punchy and informative while preserving key facts.

The original title is about IMGNet, a face verification model. Let me rewrite it to be punchy and informative while preserving key facts.

12d

EMNLP: All of the papers in my review pool being detected as AI [D]

EMNLP: All of the papers in my review pool being detected as AI [D]

15d

The original title is "How to get more from your chatbot for less [P]"

The original title is "How to get more from your chatbot for less [P]"

17d

Training transformers where every layer W = V·Uᵀ from initialization reveals a corpus-determined optimal rank - looking for arXiv endorser (cs.LG) [D]

Training transformers where every layer W = V·Uᵀ from initialization reveals a corpus-determined optimal rank - looking for arXiv endorser (cs.LG) [D]

17d

The original title is "Hamiltonian Neural Networks from a Differential Geometry Perspective [D]"

The original title is "Hamiltonian Neural Networks from a Differential Geometry Perspective [D]"

19d

P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P]

P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P]

20d

Loss functions in Instance Representation Learning [R]

Loss functions in Instance Representation Learning [R]

21d

I'm trying to implement CALM paper, and I have some questions. [P]

I'm trying to implement CALM paper, and I have some questions. [P]

22d

NagaTranslate: Building a translation and voice pipeline for low-resource Nagaland creoles (Whisper, VITS, LLMs) [P]

NagaTranslate: Building a translation and voice pipeline for low-resource Nagaland creoles (Whisper, VITS, LLMs) [P]

23d

Attention pathologies stem from norm

Attention pathologies stem from norm

26d

[R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost

[R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost

26d

[R] All Routes Lead to Collapse: attention sinks, representation collapse, and norm stratification are what content-based routing does under a norm-blind metric

[R] All Routes Lead to Collapse: attention sinks, representation collapse, and norm stratification are what content-based routing does under a norm-blind metric

26d

CALHippo - Mapping neurons and glial cells in the human brain hippocampus in 3D using SOTA segmentation and density estimation models [R]

CALHippo - Mapping neurons and glial cells in the human brain hippocampus in 3D using SOTA segmentation and density estimation models [R]

26d

High Dimensional, Dynamic Rotary Positional Embedding [P]

High Dimensional, Dynamic Rotary Positional Embedding [P]

27d