01Developer Creates "Handwritten" Language Model to Study the Illusion of IntelligenceREDDIT · /u/Helpful-Series132 · reddit.com · Sep 03, 02:04 AM3d
023 Ways to Enhance Your AI Model’s InterpretabilityRSS · Machine Learning Mastery · machinelearningmastery.com · Sep 01, 12:00 PM4d
04Z.AI Confirms Ox Alpha is a GLM Model and Plans Weight ReleaseREDDIT · /u/pscoutou · reddit.com · Aug 26, 11:43 AM10d
05Meta Releases MobileMoE: Efficient On-Device Mixture-of-Experts ModelsREDDIT · /u/jacek2023 · reddit.com · Aug 24, 06:49 PM12d
06Agnes-AI Releases Agnes-2.5-Pro-Alpha Reasoning Model on Hugging FaceREDDIT · /u/External_Mood4719 · reddit.com · Aug 24, 11:05 AM12d
07Training a 1.57B-Parameter Dreamer 4 World Model for Under $150REDDIT · /u/OtherRaisin3426 · reddit.com · Aug 24, 03:19 AM13d
08SenseTime Releases SenseNova U1.5-Lite with Unified Inference via OPD DistillationREDDIT · /u/SandyL925 · reddit.com · Aug 21, 03:52 AM16d
09On-Device 125M Transformer Model for Real-Time Piano AutocompleteHACKERNEWS · simedw · simedw.com · Aug 20, 12:04 PM16d
10Does Persistent Memory Without Weight Updates Count as Recursive Self-Improvement?REDDIT · /u/derspenti · reddit.com · Aug 18, 02:10 PM18d
11Ant Group Open-Sources Ling-3.0-tiny MoE Model for Local DeploymentRSS · IT HOME · ithome.com · Aug 11, 09:57 AM26d
12inclusionAI Releases Ling-3.0-tiny, a Fast Local MoE ModelREDDIT · /u/-Cubie- · reddit.com · Aug 10, 05:11 PM26d
13Nathan Lambert Releases New Textbook on LLM Post-TrainingRSS · Interconnects · interconnects.ai · Aug 10, 01:02 PM26d
14Collaborative AI Research Project Announced by PKU, UCLA, and UMDTWITTER · arena · x.com · Aug 08, 04:47 PM28d
15Local AI Enthusiasts Discuss Training Models From Scratch to Test Research PapersREDDIT · /u/Sadge404 · reddit.com · Aug 06, 03:30 AM31d
16Running MiniMax-H3 Omni-Modal Model Locally on Apple Silicon via MLXRSS · Simon Willison · simonwillison.net · Aug 04, 07:10 PM32d
17Liquid AI Releases LFM 2.5, a 2.6-Billion Parameter ModelREDDIT · /u/Alarming_Positive_59 · reddit.com · Aug 04, 05:30 PM32d
18Cursor Open-Sources Mixture-of-Kittens (MoK) for Faster MoE Training on NVL72TWITTER · cursor_ai · x.com · Aug 04, 04:00 PM32d
19Cursor Deploys 'MoK' to Boost GPU Training Throughput by 1.41xTWITTER · cursor_ai · x.com · Aug 04, 04:00 PM32d
20Guide Outline for LLM Decoding Strategies and Output ControlRSS · Machine Learning Mastery · machinelearningmastery.com · Aug 03, 02:36 PM33d