GLM 5.2 (744B) on 25 GB RAM consumer machineJul 11 ⋅ Least-Tangerine-8402 ⋅ r/LocalLLM ⋅ #744B #25 #2 #5
Tested glm-5 after ignoring the hype for weeks. ok I get it nowMar 13 ⋅ Weird_Perception1728 ⋅ r/LocalLLM ⋅ #5
GLM 5.2 model — 744 billion parameters / 384 GB — running on a laptop 😏Aug 12 ⋅ lucyferorg ⋅ r/LocalLLM ⋅ #384 #744 #2 #5
Are we at a tipping point for local AI? Qwen3.5 might just be.Mar 5 ⋅ Far_Noise_5886 ⋅ r/LocalLLM ⋅ #5
Running GLM-4.7 (355B MoE) in Q8 at ~5 Tokens/s on 2015 CPU-Only Hardware – Full Optimization GuideDec 2025 ⋅ at0mi ⋅ r/LocalLLM ⋅ #355B-MoE #2015 #4 #5 #7
🚀Pocket LLM v1.5.0 is out: offline Android LLM chat with voice, image input, OCR, and camera captureApr 26 ⋅ 100daggers_ ⋅ r/LocalLLM ⋅ #0 #5
Glm 5.2 weights hit hf today under MIT, frontier-level open source is actually happeningJun 16 ⋅ Exact-Literature-395 ⋅ r/LocalLLM ⋅ #2 #5
Glm-5.1 claims near opus level coding performance: Marketing hype or real? I ran my own testsApr 8 ⋅ Yssssssh ⋅ r/LocalLLM ⋅ #1 #5
2B Qwen model beats Gemini 3.5 Flash on a basic addition questionMay 21 ⋅ hurn2k ⋅ r/LocalLLM ⋅ #3 #5
GLM-5.2 753B (IQ1_S) fully local across 2×M5 Max over one TB5 cable — ~16 tok/s, llama.cpp RPC [video]Jun 29 ⋅ AiLocalGuy ⋅ r/LocalLLM ⋅ #video #IQ1_S #16 #2 #5
MiniMax-M3-uncensored NVFP4 update: 795.5 GiB to 242.4 GiB, now practical on one 4x 96 GB Blackwell nodeJul 16 ⋅ rressl ⋅ r/LocalLLM ⋅ #242 #795 #96 #4 #5
Anthropic says U.S. export controls on Claude Fable 5 & Mythos 5 have been liftedJul 1 ⋅ star_Light570 ⋅ r/LocalLLM ⋅ #5
How to use Llama-swap, Open WebUI, Semantic Router Filter, and Qwen3.5 to its fullestMar 8 ⋅ andy2na ⋅ r/LocalLLM ⋅ #5
DeepSeek-V4-Flash-0731(284B) on a 64G M5 Pro Mac — 6.5 tok/s decode with ~30 GB memory usage in CodexAug 9 ⋅ Fantastic_Honey1470 ⋅ r/LocalLLM ⋅ #284B #0731 #30 #5 #6
While losers still use Gemma 4 or Qwen, gigachads already test Gemma 5Aug 18 ⋅ Additional_Hope_2031 ⋅ r/LocalLLM ⋅ #4 #5
Mistral Medium 3.5: A reliability first open model from Europe.Apr 30 ⋅ Much_Ask3471 ⋅ r/LocalLLM ⋅ #3 #5
Built little pixel pets for your desktop running Qwen 3.5 0.8B!Jul 19 ⋅ AntiqueFeedback7447 ⋅ r/LocalLLM ⋅ #0 #3 #5
Qwen 3.6-35B-A3B: Reddit Asked, So I Tested If the 3.5 Tool Calling Fixes Carry OverApr 20 ⋅ Expensive-Register-5 ⋅ r/LocalLLM ⋅ #3 #5 #6
Qwen3.8 27B is matching DeepSeek V4 Pro and GPT 5.6 Luna on Artificial AnalysisAug 18 ⋅ gargetisha ⋅ r/LocalLLM ⋅ #5 #6 #8
Qwen-3.8-27B, Nemotron-3.5-Lightning-30B-A3B, Ornith-1.5-35B-A3B, Muse-Glimmer-30B oQ8e comparison5d ⋅ DerTomsn ⋅ r/LocalLLM ⋅ #1 #3 #5 #8
I got Qwen3.5 35B A3B (~21 GB / 35B MoE) running on an RTX 2050 with just 4 GB VRAM and 16gb ram. Can token generation be improved further?Jul 11 ⋅ ImBadGuyInEveryStory ⋅ r/LocalLLM ⋅ #2050 #21 #4 #5
I re-ran Qwen3.8 27b browsing benchmarks after messing up my config. It's now on par with GPT 5.6 Luna (xhigh)Aug 19 ⋅ pierreb5 ⋅ r/LocalLLM ⋅ #xhigh #5 #6 #8
No OS, just a UEFI app (x86-64 + Raspberry Pi 5)Jul 28 ⋅ centoslinux ⋅ r/LocalLLM ⋅ #Raspberry-Pi-5 #x86-64-+-Raspberry-Pi-5 #x86-64
I benchmarked N-gram, MTP, EAGLE3, and DFlash speculative decoding on Qwen3.5-122B on Single DGX SparkJul 18 ⋅ kristiyanstoyanovAI ⋅ r/LocalLLM ⋅ #5
Qwen3.5 9b gets stuck in a seemingly infinite loop after I ask what year it thinks it isJun 27 ⋅ BaliFlipperfrenzy ⋅ r/LocalLLM ⋅ #5
GPT-5.5 vs Claude Fable 5 vs Local Qwen: 3 AI Agents, 1 TaskJul 4 ⋅ Acceptable-Object390 ⋅ r/LocalLLM ⋅ #1 #3 #5
753B model (GLM-5.2) wrote Pac-Man and is playing its own game — 2× M5 Max, ~18 tok/s [video]Jun 30 ⋅ AiLocalGuy ⋅ r/LocalLLM ⋅ #video #18 #2 #5
I compared 4 of the 120b range with a 5 question test. There's a clear winner.Mar 25 ⋅ TheRiddler79 ⋅ r/LocalLLM ⋅ #4 #5
Running Qwen 3.5 VL 2B locally on my phone + the character feature is actually pretty funMar 5 ⋅ EthanJohnson01 ⋅ r/LocalLLM ⋅ #3 #5
Running DeepSeek V4 Flash on 48GB M5 Pro MacBook Pro ~5 tok/sJul 4 ⋅ Financial-Yoghurt946 ⋅ r/LocalLLM ⋅ #5
672 GB VRAM on 7x RTX PRO 6000 Blackwell. Kimi K3 wants 1.5 TB. More GPUs, or 1 TB of system RAM?Aug 5 ⋅ rressl ⋅ r/LocalLLM ⋅ #6000 #672 #1 #5