r/MachineLearning
r/MachineLearning

r/MachineLearning

r/MachineLearning is a Reddit community for machine learning researchers and enthusiasts. It features discussions on topics like networking at conferences and the long-term value of AI research.

Profile generated by AI for Anything

Real task cost across GPT, Claude, Gemini and Kimi, 10.6x spread on models with only 2x price difference [R]

Real task cost across GPT, Claude, Gemini and Kimi, 10.6x spread on models with only 2x price difference [R]

1d

SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]

SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]

2d

Looking for feedback on my GPU-accelerated Snake AI project [P]

Looking for feedback on my GPU-accelerated Snake AI project [P]

3d

GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]

GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]

6d

The original title is about deep learning for scRNA-seq analysis. Let me rewrite it to be punchy and specific while preserving key facts.

The original title is about deep learning for scRNA-seq analysis. Let me rewrite it to be punchy and specific while preserving key facts.

6d

Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]

Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]

8d

PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]

PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]

8d

The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"

The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"

9d

Building an XGBoost Pipeline for Cross-Domain Conflict Resolution with Explainable AI

Building an XGBoost Pipeline for Cross-Domain Conflict Resolution with Explainable AI

9d

I trained a vision-language model to play Snake, and so can you. [P]

I trained a vision-language model to play Snake, and so can you. [P]

10d

[P] RL-training Qwen3.6 to RL-train tool using AI models [P]

[P] RL-training Qwen3.6 to RL-train tool using AI models [P]

10d

LLM hallucination paper(using math) accepted to ICML workshop[R]

LLM hallucination paper(using math) accepted to ICML workshop[R]

10d

Please help me understand figure on subspace similarity in LoRA paper. [D]

Please help me understand figure on subspace similarity in LoRA paper. [D]

14d

The original title is about IMGNet, a face verification model. Let me rewrite it to be punchy and informative while preserving key facts.

The original title is about IMGNet, a face verification model. Let me rewrite it to be punchy and informative while preserving key facts.

15d

EMNLP: All of the papers in my review pool being detected as AI [D]

EMNLP: All of the papers in my review pool being detected as AI [D]

18d

The original title is "How to get more from your chatbot for less [P]"

The original title is "How to get more from your chatbot for less [P]"

20d

Training transformers where every layer W = V·Uᵀ from initialization reveals a corpus-determined optimal rank - looking for arXiv endorser (cs.LG) [D]

Training transformers where every layer W = V·Uᵀ from initialization reveals a corpus-determined optimal rank - looking for arXiv endorser (cs.LG) [D]

21d

The original title is "Hamiltonian Neural Networks from a Differential Geometry Perspective [D]"

The original title is "Hamiltonian Neural Networks from a Differential Geometry Perspective [D]"

23d

P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P]

P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P]

23d

Loss functions in Instance Representation Learning [R]

Loss functions in Instance Representation Learning [R]

25d

I'm trying to implement CALM paper, and I have some questions. [P]

I'm trying to implement CALM paper, and I have some questions. [P]

25d

NagaTranslate: Building a translation and voice pipeline for low-resource Nagaland creoles (Whisper, VITS, LLMs) [P]

NagaTranslate: Building a translation and voice pipeline for low-resource Nagaland creoles (Whisper, VITS, LLMs) [P]

26d

Attention pathologies stem from norm

Attention pathologies stem from norm

29d

[R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost

[R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost

29d