r/MachineLearning
r/MachineLearning

r/MachineLearning

r/MachineLearning is a Reddit community for machine learning researchers and enthusiasts. It features discussions on topics like networking at conferences and the long-term value of AI research.

Profile generated by AI for Anything

The original title is "I Compressed Bad Apple into a 3MB Neural Network [P]"

The original title is "I Compressed Bad Apple into a 3MB Neural Network [P]"

9d

The original title is "I created an autonomous boxing benchmark [D]"

The original title is "I created an autonomous boxing benchmark [D]"

10d

Twin: A Possible Solution to AI Context Rebuilding [P]

Twin: A Possible Solution to AI Context Rebuilding [P]

11d

The original title is "I have trained a model to predict my blood sugar [P]"

The original title is "I have trained a model to predict my blood sugar [P]"

13d

Day 9 of self-studying ML — entropy, cross-entropy, and logistic regression notes [D]

Day 9 of self-studying ML — entropy, cross-entropy, and logistic regression notes [D]

14d

MLVC: Multi-platform Learned Video Codec for Real-World Deployment [P]

MLVC: Multi-platform Learned Video Codec for Real-World Deployment [P]

15d

AI Security Leaderboard: benchmarking model robustness [P]

AI Security Leaderboard: benchmarking model robustness [P]

15d

I implemented the YOLO26n model inference from scratch using ARM64 Assembly Language (No framework) [P]

I implemented the YOLO26n model inference from scratch using ARM64 Assembly Language (No framework) [P]

19d

Real task cost across GPT, Claude, Gemini and Kimi, 10.6x spread on models with only 2x price difference [R]

Real task cost across GPT, Claude, Gemini and Kimi, 10.6x spread on models with only 2x price difference [R]

22d

SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]

SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]

23d

Looking for feedback on my GPU-accelerated Snake AI project [P]

Looking for feedback on my GPU-accelerated Snake AI project [P]

23d

The original headline is: "AI system Fable 5 reportedly solves major open problem in Algebraic Geometry"

The original headline is: "AI system Fable 5 reportedly solves major open problem in Algebraic Geometry"

24d

Tri-Net v2: Open-source implementation of our Scientific Reports paper on unified skin lesion and symptom-based monkeypox detection [R]

Tri-Net v2: Open-source implementation of our Scientific Reports paper on unified skin lesion and symptom-based monkeypox detection [R]

24d

Introducing ASCIITermDraw Bench | Testing the ability of VLMs to Generate and Edit ASCII [P]

Introducing ASCIITermDraw Bench | Testing the ability of VLMs to Generate and Edit ASCII [P]

25d

Follow up: GPT-2's vocabulary as a hyperbolic tree — 32,070 tokens in a Poincaré ball you can fly through [P]

Follow up: GPT-2's vocabulary as a hyperbolic tree — 32,070 tokens in a Poincaré ball you can fly through [P]

26d

ASCIITermDraw-Bench | Evaluating VLMs on ASCII Generation and Editing Tasks [P]

ASCIITermDraw-Bench | Evaluating VLMs on ASCII Generation and Editing Tasks [P]

26d

GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]

GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]

26d

The original title is about deep learning for scRNA-seq analysis. Let me rewrite it to be punchy and specific while preserving key facts.

The original title is about deep learning for scRNA-seq analysis. Let me rewrite it to be punchy and specific while preserving key facts.

26d

Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]

Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]

29d

PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]

PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]

29d

The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"

The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"

30d

Building an XGBoost Pipeline for Cross-Domain Conflict Resolution with Explainable AI

Building an XGBoost Pipeline for Cross-Domain Conflict Resolution with Explainable AI

30d

I trained a vision-language model to play Snake, and so can you. [P]

I trained a vision-language model to play Snake, and so can you. [P]

31d

[P] RL-training Qwen3.6 to RL-train tool using AI models [P]

[P] RL-training Qwen3.6 to RL-train tool using AI models [P]

31d