r/MachineLearning
r/MachineLearning is a Reddit community for machine learning researchers and enthusiasts. It features discussions on topics like networking at conferences and the long-term value of AI research.
Profile generated by AI for Anything
![Real task cost across GPT, Claude, Gemini and Kimi, 10.6x spread on models with only 2x price difference [R]](https://preview.redd.it/7ejtvp684xeh1.png?width=140&height=65&auto=webp&s=10790ba444afd733ece8c54a8b9da99969a86066)
Real task cost across GPT, Claude, Gemini and Kimi, 10.6x spread on models with only 2x price difference [R]
1d
![SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]](https://preview.redd.it/1457xi9fcqeh1.jpg?width=140&height=90&auto=webp&s=879aad6df9e51a2735d91112d01518ff76ba3cbe)
SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]
2d
![Looking for feedback on my GPU-accelerated Snake AI project [P]](https://preview.redd.it/4k0bf6wgtneh1.gif?width=640&crop=smart&s=7309dc4cdba7df36b615ed9025f212c2b34fd4b0)
Looking for feedback on my GPU-accelerated Snake AI project [P]
3d
![GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]](https://preview.redd.it/tlvz4c3i32eh1.png?width=640&crop=smart&auto=webp&s=aad6aeec9197e26debda00093dd47611e70c5a08)
GPT-2 Small’s embedding geometry around “Trump”: discretized vs. continuous nearest neighbours [P]
6d

The original title is about deep learning for scRNA-seq analysis. Let me rewrite it to be punchy and specific while preserving key facts.
6d
![Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]](https://preview.redd.it/b0u6q9a46ndh1.jpg?width=140&height=98&auto=webp&s=dbb02d2e0fc85305a04e37864167c2578891d46c)
Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]
8d
![PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]](https://external-preview.redd.it/d6rTpW7131dBgTTGfjDXPAIkblduF91pERLr20qfQH4.jpeg?width=140&height=78&auto=webp&s=efc5027d8347fa1bc645f041300b0979e5c14469)
PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]
8d
![The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"](https://preview.redd.it/wvhcgu1q1fdh1.png?width=140&height=88&auto=webp&s=d914b037ae47cf6b529bc66f8af00430d0d590ae)
The original title is: "All major robotics and VLA papers, ranked and benchmarked in a single place [P]"
9d

Building an XGBoost Pipeline for Cross-Domain Conflict Resolution with Explainable AI
9d
![I trained a vision-language model to play Snake, and so can you. [P]](https://external-preview.redd.it/YWcAyMNI6jxa5S-SYFMgIq4qY5VYLAesOmSGvUtU3as.png?width=140&height=70&auto=webp&s=247aa5820e62edcecbb91eb5a961e10ecb0ae66f)
I trained a vision-language model to play Snake, and so can you. [P]
10d
![[P] RL-training Qwen3.6 to RL-train tool using AI models [P]](https://preview.redd.it/hg7ww6ute8dh1.png?width=140&height=75&auto=webp&s=d9c4aa6843cd8b9f2a480a97e0f8469ec80b4d41)
[P] RL-training Qwen3.6 to RL-train tool using AI models [P]
10d
![LLM hallucination paper(using math) accepted to ICML workshop[R]](https://preview.redd.it/3uyvbtoa76dh1.png?width=140&height=61&auto=webp&s=523d3943b9adbcbbdaca03be35c5e073be075de9)
LLM hallucination paper(using math) accepted to ICML workshop[R]
10d
![Please help me understand figure on subspace similarity in LoRA paper. [D]](https://preview.redd.it/3l5qhbiroech1.png?width=640&crop=smart&auto=webp&s=e5534631f23bcc8d8e89b7fd411120c2b7a84442)
Please help me understand figure on subspace similarity in LoRA paper. [D]
14d

The original title is about IMGNet, a face verification model. Let me rewrite it to be punchy and informative while preserving key facts.
15d
![EMNLP: All of the papers in my review pool being detected as AI [D]](https://preview.redd.it/elmu1a6d0kbh1.png?width=140&height=140&crop=1:1,smart&auto=webp&s=c44181dd02b9668433e47a57a648d98af652fbde)
EMNLP: All of the papers in my review pool being detected as AI [D]
18d
![The original title is "How to get more from your chatbot for less [P]"](https://external-preview.redd.it/qZ4HpTKGOz0DwdVSdExJAW-UpDhB6hu23O6kgI1aNvE.jpeg?width=640&crop=smart&auto=webp&s=0d0ad48f322f4db16ae97f0b70276398c4b3e711)
The original title is "How to get more from your chatbot for less [P]"
20d
![Training transformers where every layer W = V·Uᵀ from initialization reveals a corpus-determined optimal rank - looking for arXiv endorser (cs.LG) [D]](https://external-preview.redd.it/Qfw5SuGCt2d45VbzHurInHB_fbCrPRWPZr4XzFenJcc.png?width=140&height=70&auto=webp&s=6e9379fe0f90d43518578b30abf4563219025786)
Training transformers where every layer W = V·Uᵀ from initialization reveals a corpus-determined optimal rank - looking for arXiv endorser (cs.LG) [D]
21d
![The original title is "Hamiltonian Neural Networks from a Differential Geometry Perspective [D]"](https://external-preview.redd.it/7q8iktqnOmHdHgGNxMCQbvHkXz6extXfcSIuznTr8CA.png?width=640&crop=smart&auto=webp&s=158ee06f289fc1e95a2efb1e71a67adbc515092f)
The original title is "Hamiltonian Neural Networks from a Differential Geometry Perspective [D]"
23d
![P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P]](https://preview.redd.it/v50euf4pymah1.png?width=140&height=81&auto=webp&s=b9a9d3b99087e03cd79f28ebf6ac8622dd9bcc0f)
P Moth-Retrieval: Graph-Free Multi-Hop Retrieval via Query-Time Orchestration (Beating Graph-Based Systems on HotpotQA) [P]
23d
![Loss functions in Instance Representation Learning [R]](https://preview.redd.it/3l7mtxoc3bah1.png?width=140&height=27&auto=webp&s=8426b12f6ec1f44b193529124dee890e0642ad25)
Loss functions in Instance Representation Learning [R]
25d
![I'm trying to implement CALM paper, and I have some questions. [P]](https://preview.redd.it/kr4u22yfx8ah1.png?width=140&height=83&auto=webp&s=784c46c82400e669571b4d8a7dcdc997ad0fba57)
I'm trying to implement CALM paper, and I have some questions. [P]
25d
![NagaTranslate: Building a translation and voice pipeline for low-resource Nagaland creoles (Whisper, VITS, LLMs) [P]](https://preview.redd.it/bu6xsk4hvx9h1.jpg?width=140&height=63&auto=webp&s=0a4c589616d351d10d735940874706494e48d408)
NagaTranslate: Building a translation and voice pipeline for low-resource Nagaland creoles (Whisper, VITS, LLMs) [P]
26d

Attention pathologies stem from norm
29d
![[R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost](https://external-preview.redd.it/q3evP6JeDpAC2MdSQHWYxnCYTqbJkElIQsLFqVSdkss.png?width=640&crop=smart&auto=webp&s=de730fbf7ecace6df0036b21470c16a2d4feacfb)
[R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost
29d