Fine-Tuning

LoRA, RLHF, GRPO, model adaptation, training techniques

385 articles across 124 editions

Articles

  1. Simon Willison: Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things -- 2026-08-19
  2. Ollama's MTP variant of Qwen3.8-27B is 2x slower than the non-MTP one, measured, with a negative control -- 2026-08-19
  3. Taking Qwen3.5-9B quants to SOTA. New lineup incoming :) -- 2026-08-19
  4. [Paper] Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning -- 2026-08-19
  5. dealignai/Gemma-4-31B-JANG_4M-CRACK -- 2026-08-19
  6. Meta releases open weights for Muse Glimmer-30B -- 2026-08-13
  7. Qwen3.6 35B (2 min) vs Muse Glimmer 30B (4 min) on custom Llama.cpp build (RTX 5080) -- 2026-08-13
  8. Qwen 35B-A3B MoE vs 27B dense in local coding tests: ~4× faster, much smaller quality gap than I expected -- 2026-08-13
  9. I asked DeepSeek-V4-Flash to work with Muse-Glimmer for Vision ability in PI agent and it produced this -- 2026-08-13
  10. enabling PCI-E p2p for consumer Nvidia cards will yield you more than you think -- 2026-08-12
  11. Rumored 50-series Super refresh bumps everything +50% VRAM -- 2026-08-12
  12. Muse Glimmer 30B + DFlash speculative decoding on vLLM: 6 patches needed, 25 → 57 tok/s. Dockerfile and numbers inside. -- 2026-08-11
  13. Apple Silicon and macOS VMs: 11–16× Faster LLM Inference with Llama.cpp -- 2026-08-11
  14. Echo Dot 2 can run 28M LLM at decent speed -- 2026-08-11
  15. Chunked KL loss for running Knowledge Distillation locally (<6GB VRAM at 32K context length) -- 2026-08-11
  16. Sonic Pi v5 -- 2026-08-11
  17. Harness Engineering for Self-Improvement -- 2026-08-04
  18. Dual-Dimensional Consistency: Balancing Budget and Quality in Adaptive Inference-Time Scaling -- 2026-08-04
  19. [Paper] Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression -- 2026-08-04
  20. Pydantic AI 2.0: The New Best Way to Build AI Agents is Composing Capabilities -- 2026-07-29
  21. [Editorial] The Gauntlet Loop -- 2026-07-29
  22. The Blind Curator: How a Biased Judge Silently Disables Skill Retirement in Self-Evolving Agents -- 2026-07-29
  23. [Editorial] Anatomy of a Frontier Lab Agent Intrusion: Full Technical Timeline -- 2026-07-29
  24. [Editorial] HuggingFace Agent Intrusion: Interactive Replay -- 2026-07-29
  25. Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident -- 2026-07-29
  26. Document-borne AI worms can self-propagate through Copilot for Word -- 2026-07-29
  27. Kimi-K3 Releases on HuggingFace 6/27 -- 2026-07-28
  28. Ling-3.0-flash weights: SGLang says day-0, vLLM says when they land, llama.cpp closed the 2.6 request as not_planned -- 2026-07-28
  29. Sanctions on Open Source. hope they don't do anything stupid here. -- 2026-07-28
  30. [Editorial] -- 2026-07-28
  31. [Editorial] -- 2026-07-28
  32. I distilled an 8B teacher into a 0.6B student on my Mac (MLX). The 0.6B went from 36% to 100% on the task, but few-shot prompting actively made it worse. -- 2026-07-28
  33. OpenAI and Anthropic unite against open-weight AI risks to their bottom line -- 2026-07-27
  34. [Editorial] -- 2026-07-27
  35. [Editorial] -- 2026-07-27
  36. What do we know about the "AI Accelerators" used to train LongCat-2? -- 2026-07-24
  37. Running a 13M ASR conformer on a microcontroller -- 2026-07-24
  38. Running Qwen 3.6 35B MoE (Q4_K_M) on a Zeus (Xiaomi 12 Pro, 12GB RAM) -- 2026-07-24
  39. Squeeze-Release: Iterative Pruning with Exact Structural Minimization -- 2026-07-24
  40. [Editorial] -- 2026-07-23
  41. CEO of Hugging face: Heading to San Francisco to have a little chat with that "rogue agent" -- 2026-07-23
  42. [Editorial] -- 2026-07-21
  43. [Editorial] -- 2026-07-21
  44. Kimi K3 Agentic Benchmark -- 2026-07-17
  45. Kimi k3 is 2.8t! Will need to have an aggressive iQ2_XXS or IQ1.8! -- 2026-07-17
  46. Governments, companies, nonprofits should invest in free, open source AI [pdf] -- 2026-07-17
  47. Zig Creator Calls Spade a Spade, Anthropic Blows Smoke -- 2026-07-15
  48. Apple sues OpenAI alleging trade secret theft, says scheme was 'at every level' -- 2026-07-15
  49. China's DeepSeek developing its own AI chip, sources say -- 2026-07-15
  50. Data centers have hiked electricity prices on the public by $23B -- 2026-07-15
  51. [Editorial] CMMC Baseline Did Not Move — The Floor Did -- 2026-07-14
  52. Anthropic warns that AI will soon be able to improve itself without human intervention -- 2026-07-13
  53. Zhipu founder backs open-source AI as global security debate intensifies -- 2026-07-13
  54. [Editorial] -- 2026-07-13
  55. Kode Dot Programmable pocket device for makers, pentesters and geeks -- 2026-07-13
  56. stabilityai/stable-audio-3-medium -- 2026-07-13
  57. LiquidAI/LFM2.5-230M -- 2026-07-13
  58. Beyond Attack-Success Rate: Action-Graded Severity Scale for Tool-Using AI Agents -- 2026-07-10
  59. [Editorial] -- 2026-07-10
  60. Getting Someone Else's Chat -- 2026-07-10
  61. [Editorial] -- 2026-07-10
  62. rheodev/cpa-plugin-privacyfilter -- 2026-07-10
  63. The Underhanded C Contest -- 2026-07-03
  64. Senior SWE-Bench: open-source benchmark that assesses agents as senior engineers -- 2026-07-02
  65. New bench designed for smaller models: ObviousBench.com -- 2026-07-02
  66. I built an autonomous dev pipeline and ran the same project head to head: a 27B local on a modded 4090, then again on cheap cloud LLMs -- 2026-07-02
  67. Stdlib or Third-Party? Empirical Performance and Correctness of LLM-Assisted Zero-Dependency Python Libraries -- 2026-07-02
  68. DeepSeek V4 official version launching mid-July -- 2026-06-30
  69. OpenPangu-2.0-Flash: 92B MoE (6B active) on Ascend with 512K context -- 2026-06-30
  70. Anthropic's Amodei: "Open Source models [could take us to] a very dangerous place." -- 2026-06-30
  71. Even Google still believes in small models for coding — Gemma 4 31B hackathon at 1500 tok/s -- 2026-06-30
  72. Previewing GPT‑5.6 Sol: a next-generation model -- 2026-06-29
  73. [Editorial] -- 2026-06-29
  74. [Editorial] -- 2026-06-29
  75. We built a calibration-aware Q4_K_M quant of Qwen3.5 0.8B that recovers 96.5% of the BF16 gap vs pure llama.cpp Q4_K_M (SpectralQuant) -- 2026-06-29
  76. Update: First Manual Results from Testing Procedural Skill Transfer in Small Models -- 2026-06-29
  77. Multi Tier MoE Caching -- 2026-06-29
  78. GLM-5.2 is a step change for open agents -- 2026-06-25
  79. poolside/Laguna-M.1 · Hugging Face - 225B-A23B -- 2026-06-25
  80. Mimo 2.5 is _fast_ at large context (dual RTX Pro 6000) -- 2026-06-25
  81. The Eagle(3) has landed (for Qwen) -- 2026-06-25
  82. CPU-only TTS benchmark: Kokoro 82M vs Supertonic 3 vs Inflect-Nano-v1 (4.6M params), with UTMOS scoring on every sample -- 2026-06-25
  83. [Editorial] Video Content -- 2026-06-16
  84. [Editorial] Video Content -- 2026-06-16
  85. [Editorial] Video Content -- 2026-06-16
  86. [Editorial] StandardAgents Arrow-JS — JavaScript Agent Framework -- 2026-06-16
  87. archex: Local-First Deterministic Code-Context for AI Agents — No API Key, No Telemetry (Apache 2.0) -- 2026-06-16
  88. Ironsmith: Open Source macOS App That Creates macOS Apps From Prompts — Works With Local Models -- 2026-06-16
  89. zengxiao-he/tessera -- 2026-06-11
  90. Tencent-Hunyuan/UniRL -- 2026-06-11
  91. SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning -- 2026-06-11
  92. Physics-Grounded Multi-Agent Architecture for Traceable, Risk-Aware Human-AI Decision Support in Manufacturing -- 2026-06-11
  93. Config Files That Run Code: Supply Chain Security Blindspot -- 2026-06-10
  94. Surveillance is not safety: A statement on the UK's latest threat to privacy -- 2026-06-10
  95. FrontierCode -- 2026-06-10
  96. Introducing North Mini Code: Cohere's First Model For Developers -- 2026-06-10
  97. [Editorial] jedArden/ARMOR -- 2026-06-05
  98. m-sec-org/wafkiller -- 2026-06-05
  99. zmn-hamid/sni-spoofing-scanner -- 2026-06-05
  100. Everything at Every Scale: Scale-Invariant Diffusion with Continuous Super-Resolution -- 2026-06-04
  101. Gaussian Point Splatting -- 2026-06-04
  102. DaVinci Resolve 21 -- 2026-06-04
  103. Profiling in PyTorch (Part 1): A Beginner's Guide to torch.profiler -- 2026-06-04
  104. microsoft/harrier-oss-v1-270m -- 2026-06-04
  105. CohereLabs/tiny-aya-global -- 2026-06-04
  106. Use your Nvidia GPU's VRAM as swap space on Linux -- 2026-06-03
  107. [Editorial] chipotlai-max -- 2026-06-03
  108. [Editorial] Video Submission -- 2026-06-03
  109. It is an amazing time for programmers -- 2026-06-03
  110. A 10 year old Xeon is all you need -- 2026-06-02
  111. Inferencing at 10.33 t/s on Qwen 3.5 35B on a $300 laptop -- 2026-06-02
  112. OSC: Hardware Efficient W4A4 Quantization via Outlier Separation in Channel Dimension -- 2026-06-02
  113. Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation -- 2026-06-02
  114. Qwen3.6 huge quality gain from Q4 to Q6 for coding agent -- 2026-06-02
  115. 1-Bit Bonsai Image 4B Image Generation for Local Devices -- 2026-06-01
  116. Old Mac Pro still proving its worth -- 2026-06-01
  117. Heterogeneous GPU Weighting & Layer Splitting -- 2026-06-01
  118. I finally put my NPU (Intel Arrow Lake) to use doing ASR for my smart home -- 2026-06-01
  119. OpenMOSS-Team/MOSS-TTS-v1.5 · Hugging Face -- 2026-06-01
  120. [Editorial] -- 2026-06-01
  121. [Editorial] -- 2026-06-01
  122. [Editorial] -- 2026-06-01
  123. [Editorial] -- 2026-06-01
  124. [Editorial] Barracuda Nightmare Eclipse Zero-Days -- 2026-05-29
  125. [Editorial] Video -- 2026-05-29
  126. Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL -- 2026-05-28
  127. PACZero: PAC-Private Fine-Tuning of Language Models via Sign Quantization -- 2026-05-28
  128. If AI writes your code, why use Python? -- 2026-05-12
  129. [Editorial] Lex Fridman — DeepSeek Deep Dive with Dylan Patel & Nathan Lambert -- 2026-05-12
  130. [Editorial] Reuven Cohen on AI Industry Developments -- 2026-05-12
  131. [Editorial] When Helpfulness Becomes Sycophancy -- 2026-05-08
  132. [Editorial] Vibe-Cast JEPA — Joint Embedding Predictive Architecture Exploration -- 2026-05-08
  133. [Editorial] -- 2026-05-07
  134. ProgramBench: Can we really rebuild huge binaries from scratch? (doesn't look like it) -- 2026-05-07
  135. Adding Benchmaxxer Repellant to the Open ASR Leaderboard -- 2026-05-07
  136. Does the "6 months gap" still hold? -- 2026-05-07
  137. Claude Code @ Opus 4.7 vs OpenCode @ qwen3.6:27b. Both shipped a playable cozy roguelite. -- 2026-05-07
  138. Fine-tuned Qwen3.6-35B-A3B DeltaNet experiment -- 2026-05-07
  139. Zyphra/ZAYA1-8B -- 2026-05-07
  140. Anthropic ships Claude for Creative Work with nine MCP-native connectors -- 2026-05-05
  141. n8n Just Got a New Tool (and it can SUPERCHARGE Claude Automations) -- 2026-05-05
  142. [Editorial] How to Use ADRs in Ruflo -- 2026-05-05
  143. Microsoft and OpenAI end their exclusive and revenue-sharing deal -- 2026-05-01
  144. [Editorial] Video: AI Development Insights -- 2026-05-01
  145. Talkie: a 13B vintage language model from 1930 -- 2026-05-01
  146. Your phone is about to stop being yours -- 2026-04-29
  147. OpenAI almost banned me because I tried to automate YouTube download -- 2026-04-29
  148. [Editorial] Video editorial submission -- 2026-04-29
  149. GPT-5.5 is out -- 2026-04-28
  150. OpenAI CEO's Identity Verification Company Announced Fake Bruno Mars Partnership -- 2026-04-28
  151. Trump picked a fight with Anthropic. Now the administration is backing off. -- 2026-04-28
  152. [Editorial] -- 2026-04-28
  153. Generalization at the Edge of Stability -- 2026-04-28
  154. [Editorial] -- 2026-04-28
  155. [Editorial] NIST Cybersecurity MLX Pipeline -- 2026-04-27
  156. [Editorial] LLM Fine-Tuning Guide -- 2026-04-27
  157. [Editorial] Open Generative AI — Curated Resource List -- 2026-04-27
  158. [Editorial] SecWest: AMD VVI — Hardware-Level Vulnerability Research -- 2026-04-20
  159. [Editorial] Phenoelit/Halvar — Legendary Security Research -- 2026-04-20
  160. Lean proved this program correct; then I found a bug -- 2026-04-14
  161. Multi-Agentic Software Development Is a Distributed Systems Problem -- 2026-04-14
  162. [Editorial] The Room That Quoted Back -- 2026-04-14
  163. [Editorial] -- 2026-04-13
  164. Qwen3.5-397B is shockingly useful at Q2 -- 2026-04-13
  165. [Editorial] -- 2026-04-13
  166. Liquid AI releases LFM2.5-VL-450M - structured visual understanding at 240ms -- 2026-04-13
  167. LLM Novice Uplift on Dual-Use Biology Tasks — 4x Accuracy Boost Bypasses Safeguards -- 2026-04-10
  168. [Editorial] Your AI Is Developing Capabilities Nobody Tested -- 2026-04-10
  169. The current state of the Chinese LLMs scene -- 2026-03-26
  170. Alibaba confirms they are committed to continuously open-sourcing new Qwen and Wan models -- 2026-03-26
  171. Cursor's Composer 2 apparently built on Kimi K2.5 without attribution -- 2026-03-26
  172. Nemotron Cascade 2 30B A3B -- 2026-03-26
  173. NVIDIA 2026 Conference LIVE. New Base model coming! -- 2026-03-20
  174. Nemotron 3 Nano 4B: A Compact Hybrid Model for Efficient Local AI -- 2026-03-20
  175. Granite 4.0 1B Speech: Compact, Multilingual, and Built for the Edge -- 2026-03-20
  176. [New Model & Agent] LocoTrainer-4B: A Claude Code-style local agent designed specifically to master the MS-SWIFT framework (4B, 32K, GGUF) -- 2026-03-20
  177. shallowdream204/BitDance-14B-16x -- 2026-03-20
  178. [Editorial] Karpathy autoresearch -- 2026-03-10
  179. [Editorial] Doc-to-LoRA -- 2026-03-10
  180. GPT-5.4 -- 2026-03-09
  181. [Editorial] -- 2026-03-09
  182. YuanLabAI/Yuan3.0-Ultra: 1010B MoE, fully open weights -- 2026-03-05
  183. We could be hours (or less than a week) away from true NVFP4 support in Llama.cpp GGUF format -- 2026-03-05
  184. Step-3.5-Flash-Base & Midtrain (in case you missed them) -- 2026-03-05
  185. Qwen3.5-9B Uncensored Aggressive Release (GGUF) -- 2026-03-05
  186. unknown -- 2026-03-05
  187. [Editorial] David Maynor Security Gist -- 2026-03-04
  188. [Editorial] arXiv:2602.23093 -- 2026-03-04
  189. unpromptedcon.org -- 2026-03-04
  190. Inside the M4 Apple Neural Engine, Part 1: Reverse Engineering -- 2026-03-03
  191. Hydroph0bia – fixed SecureBoot bypass for UEFI firmware from Insyde H2O (2025) -- 2026-03-03
  192. [Editorial] -- 2026-02-26
  193. [Editorial] -- 2026-02-26
  194. [Editorial] -- 2026-02-26
  195. Free ASIC Llama 3.1 8B inference at 16,000 tok/s - no, not a joke -- 2026-02-25
  196. [Editorial] Cognitum -- 2026-02-25
  197. Hetzner Prices increase 30-40% -- 2026-02-25
  198. [Editorial] -- 2026-02-24
  199. Anthropic Accuses DeepSeek, Moonshot AI, and MiniMax of Creating 24,000 Fake Claude Accounts -- 2026-02-24
  200. Gemini 3.1 Pro -- 2026-02-23
  201. 15 years of FP64 segmentation, and why the Blackwell Ultra breaks the pattern -- 2026-02-20
  202. Nvidia and OpenAI abandon unfinished $100B deal in favour of $30B investment -- 2026-02-20
  203. Google releases Gemini 3.1 Pro with Benchmarks -- 2026-02-20
  204. praetorian-inc/brutus -- 2026-02-19
  205. [Editorial] Unicornscan Getting Started -- 2026-02-19
  206. [Editorial] Unicornscan Alicorn -- 2026-02-19
  207. What Your Bluetooth Devices Reveal About You -- 2026-02-19
  208. [Editorial] When everyone can build software, who learns well? -- 2026-02-19
  209. Sonnet 4.6 feels like Opus 4.5 at Sonnet pricing -- 2026-02-19
  210. Anthropic Raises $30,000,000,000 As Run-Rate Revenue Grew 10x Annually Over Three Years -- 2026-02-19
  211. REASONING AUGMENTED RETRIEVAL (RAR) is the production-grade successor to single-pass RAG -- 2026-02-19
  212. Qwen Released Qwen 3.5 397B and Qwen 3.5 Plus! -- 2026-02-17
  213. Qwen3.5 NVFP4 (Blackwell) is up! -- 2026-02-17
  214. Running Gemma 3n E2B natively on Android via LiteRT -- 2026-02-17
  215. Deploying Open WebUI + vLLM on Amazon EKS -- 2026-02-17
  216. [Editorial] https://forge-quality.dev/articles/case-of-passing-tests-investigation -- 2026-02-02
  217. [Editorial] https://www.linkedin.com/posts/ownyourai_deepseek-just-released-the-first-vision-ai-activity-7421818927657385987-V1yo -- 2026-01-27
  218. Unsloth announces support for finetuning embedding models -- 2026-01-27
  219. matrixhub-ai/matrixhub -- 2026-01-14
  220. HM-RunningHub/ComfyUI_RH_DreamID-V -- 2026-01-14
  221. YouTube has removed the ability to search by upload date -- 2026-01-14
  222. Tried this open-source framework for LLM fine-tuning over UI -- 2025-12-12
  223. Golang optimizations for high‑volume services -- 2025-12-12
  224. [Editorial] https://www.linkedin.com/posts/stuart-winter-tear_assessing-llms-for-serendipity-discovery-activity-7396596796938153984-JY9u -- 2025-12-12
  225. Deprecations via warnings don't work for Python libraries -- 2025-12-11
  226. The "Confident Idiot" Problem: Why LLM-as-a-Judge fails in production. -- 2025-12-10
  227. Toyota unintended acceleration and the big bowl of "spaghetti" code (2013) -- 2025-12-09
  228. Free yourself from the Spotify desktop client with spotifyd -- 2025-12-04
  229. Llamacpp Parameters Tuning -- 2025-12-02
  230. [Editorial] https://huggingface.co/blog/grimjim/norm-preserving-biprojected-abliteration -- 2025-12-02
  231. Pong Gets the Boot -- 2025-11-26
  232. Building the largest known Kubernetes cluster, with 130k nodes -- 2025-11-26
  233. The Qtile Window Manager: A Python-Powered Tiling Experience -- 2025-11-25
  234. Most Stable Raspberry Pi? Better NTP with Thermal Management -- 2025-11-25
  235. Quantum physicists have shrunk and "de-censored" DeepSeek R1 -- 2025-11-20
  236. Gain 60% performance on RDNA 4 using this fix -- 2025-11-19
  237. Scale-out is the silent killer of LLM applications. Are we solving the wrong problem? -- 2025-11-19
  238. [Editorial] https://brianhorakh.medium.com/just-mcp-to-reduce-context-waste-in-spec-driven-development-3935922da5cf -- 2025-11-18
  239. [AutoBE] Qwen3-80B suddenly wrote doomsday AI mythology while generating a TODO app -- 2025-11-18
  240. My trick for better Claude Code collaboration: CLAUDE.md with conditional loading -- 2025-11-18
  241. A proper way to connect a local LLM to iMessage? -- 2025-11-13
  242. How do I level up from normie to normie pro with Claude -- 2025-11-13
  243. POC: Model Context Protocol integration for native Ollama app -- 2025-11-12
  244. Skills are in a weird middle ground between RAG and Custom GPTs, and I think that's why they feel so awkward -- 2025-11-12
  245. Native LLM Router Integration with Cost Transparency for OpenWebUI -- 2025-11-12
  246. Last week in Multimodal AI - Local Edition -- 2025-11-12
  247. DeepSeek-OCR GGUF model runs great locally - simple and fast -- 2025-11-12
  248. Qwen3-VL works really good with Zoom-in Tool -- 2025-11-12
  249. lightonai/LightOnOCR-1B-1025 -- 2025-11-12
  250. Qwen/Qwen3-VL-2B-Thinking -- 2025-11-12
  251. [Editorial] https://www.linkedin.com/posts/daniel-cuthbert0x_a-month-ago-gadi-evron-and-i-set-about-building-ugcPost-7393643597729845248-TSTD -- 2025-11-11
  252. Breakdown of New RunC Vulnerabilities -- 2025-11-11
  253. [Editorial] https://www.linkedin.com/posts/andriyburkov_this-paper-shows-a-27-million-parameter-model-activity-7393432619365052416-SFLO -- 2025-11-10
  254. Trajectory Distillation for Foundation Models -- 2025-11-10
  255. sail-sg/Precision-RL -- 2025-11-10
  256. inclusionAI/LLaDA2.0-flash-preview -- 2025-11-10
  257. Finetuning DeepSeek 671B locally with only 80GB VRAM and Server CPU -- 2025-11-07
  258. IPEX-LLM llama.cpp portable GPU and NPU working really well on laptop -- 2025-11-07
  259. Building a PV Solar-Powered Quadcopter -- 2025-11-07
  260. Adding a RTX 5080 into a 2U server with OcuLink -- 2025-11-06
  261. Why does Image Recognition work in llama-server but not through Open WebUI? -- 2025-11-06
  262. 2025 Component Abuse Challenge: A Piezo Disk Powers A Transmitter -- 2025-11-06
  263. [Editorial] Does the EU know that there are many countries outside of the EU that do not care at all about their -- 2025-11-03
  264. Ilya Sustkever's deposition reveals previously unknown details [pdf] -- 2025-11-03
  265. CISA and NSA share tips on securing Microsoft Exchange servers -- 2025-11-02
  266. The Smol Training Playbook: The Secrets to Building World-Class LLMs -- 2025-11-02
  267. Latest Update from Anthropic's new model - Neptune V6 -- 2025-11-02
  268. AI "Phone Farm" Startup Gets Funding from Marc Andreessen to Flood Social Media With Spam -- 2025-11-02
  269. Minimax-M2 cracks top 10 overall LLMs (production LLM performance gap shrinking: 7 points from GPT-5 in Artificial Analysis benchmark) -- 2025-11-01
  270. 🚨 OpenAI Gives Microsoft 27% Stake, Completes For-Profit Shift -- 2025-11-01
  271. FlashPack: High-throughput tensor loading for PyTorch -- 2025-11-01
  272. Kafka is Fast – I'll use Postgres -- 2025-11-01
  273. Analog Surround Sound Was Everywhere, But You Probably Didn’t Notice -- 2025-11-01
  274. Optimizing gpt-oss-120B on AMD RX 6900 XT 16GB: Achieving 19 tokens/sec -- 2025-10-31
  275. Flamingo 3 released in safetensors -- 2025-10-31
  276. Jeep Issues Emergency Recall for OTA-Bricked Wrangler 4xes -- 2025-10-29
  277. queenkiley/AI-Art-Generator -- 2025-10-26
  278. Unlock the power of images with AI Sheets -- 2025-10-26
  279. Open WebUI Context Menu -- 2025-10-26
  280. [Editorial] Browsers you can socially engineer -- 2025-10-24
  281. Update on Plans for Privacy Sandbox Technologies -- 2025-10-24
  282. PlayDiffusion finetune for audio inpainting non-verbal tags -- 2025-10-21
  283. Nvidia has produced the first Blackwell wafer on US soil -- 2025-10-21
  284. Show HN: Syna – Minimal ML and RL Framework Built from Scratch with NumPy -- 2025-10-21
  285. C Project Turns Into Full-Fledged OS -- 2025-10-21
  286. [Editorial] Agentic Orchestration -- 2025-10-19
  287. Chicken Squisher 3000: Squish-Proof Security -- 2025-10-19
  288. [Editorial] Claude Skills are awesome, maybe a bigger deal than MCP -- 2025-10-18
  289. [Editorial] Claude Skills -- 2025-10-18
  290. Copy-and-Patch: A Copy-and-Patch Tutorial -- 2025-10-17
  291. We built 3B and 8B models that rival GPT-5 at HTML extraction while costing 40-80x less - fully open source -- 2025-10-17
  292. Comparing Popular AI Evaluation Platforms for 2025 -- 2025-10-17
  293. State of AI Report 2025 -- 2025-10-17
  294. Nvidia breakthrough gives 4-bit pretraining technique the accuracy of FP8 -- 2025-10-15
  295. AI assisted suite - Doubt about n_gpu layer test -- 2025-10-15
  296. ibm-granite/granite-4.0-h-micro -- 2025-10-15
  297. Qwen/Qwen3-VL-235B-A22B-Instruct -- 2025-10-15
  298. Get your VLM running in 3 simple steps on Intel CPUs -- 2025-10-15
  299. A 5-minute, no-BS way to pick a local model for your real task -- 2025-10-14
  300. [Update] CodeLens.AI - Crowdsourced AI Leaderboard 3 Days Later: Blind Voting and What We Learned -- 2025-10-14
  301. ZephrFish/OmniProx -- 2025-10-14
  302. Preference optimization with ORPO and LoRA -- 2025-10-12
  303. [Show] SpiralTorch: A Rust-based PyTorch-style autograd engine (Python 3.14-ready) -- 2025-10-12
  304. 2G Gone? Bring It Back Yourself! -- 2025-10-12
  305. [Editorial] https://www.anthropic.com/research/small-samples-poison -- 2025-10-11
  306. [Editorial] https://www.linkedin.com/pulse/from-chatbot-operating-system-what-openais-next-move-means-leimer-ju18c -- 2025-10-11
  307. Rubygems.org AWS Root Access Event – September 2025 -- 2025-10-11
  308. Stop flexing Pass@N — show Pass-all-N -- 2025-10-11
  309. Architecting a project for optimal AI coding, any tips? -- 2025-10-11
  310. Basekick-Labs/arc -- 2025-10-11
  311. ServiceNow-AI/Apriel-1.5-15b-Thinker -- 2025-10-11
  312. meituan-longcat/LongCat-Flash-Chat -- 2025-10-11
  313. Did anyone try out GLM-4.5-Air-GLM-4.6-Distill ? -- 2025-10-10
  314. Thank you Anthropic &amp; this community! Our little side project just hit 1M visits and even made it on National TV! -- 2025-10-10
  315. Sharing my free tool for easy handwritten fine-tuning datasets! -- 2025-10-09
  316. vllm setup for nvidia (can use llama) -- 2025-10-05
  317. Full-fine tuning doesn't require much vRAM with gradient checkpointing... -- 2025-10-05
  318. Qwen/Qwen3-Omni-30B-A3B-Thinking -- 2025-10-05
  319. inclusionAI/Ring-mini-linear-2.0 -- 2025-10-05
  320. llama.cpp: Quantizing from bf16 vs f16 -- 2025-10-05
  321. GLM 4.6 is nice -- 2025-10-04
  322. NVFP4 or MXFP4 MOE on sm120 (RTX 5900 RTX 6000 PRO) -- 2025-10-04
  323. Ask Hackaday: How Do You Distro Hop? -- 2025-09-30
  324. Uncensor Qwen3 models without retraining -- 2025-09-20
  325. Depth upscaling? -- 2025-09-20
  326. Definitive proof openai/gpt-oss-20b is dumb as hell -- 2025-09-19
  327. Free 10%+ Speedup for CPU/Hybrid Inference on Intel CPUs with Efficiency Cores -- 2025-09-17
  328. Claude Performance Report with Workarounds - September 7 to September 14 -- 2025-09-16
  329. PSA/RFC: KV Cache quantization forces excess processing onto CPU in llama.cpp -- 2025-09-15
  330. native tool calling support for DeepSeek V3.1 just merged in llama.cpp -- 2025-09-15
  331. model : add grok-2 support by CISC · Pull Request #15539 · ggml-org/llama.cpp -- 2025-09-15
  332. An Afternoon at the Recursive Café: Two Threads Interleaving -- 2025-09-14
  333. blacktop/go-hypervisor -- 2025-09-14
  334. TSYJ-He/AutoEnvForge -- 2025-09-14
  335. Hackaday Links: September 7, 2025 -- 2025-09-14
  336. Any idea how to use ollama (debian) with 2x GPUs to load larger models? -- 2025-09-14
  337. Rails on SQLite: new ways to cause outages -- 2025-09-14
  338. Qwen3-Coder-480B Q2_K_XL same speed as Qwen3-235b-instruct Q3_K_XL WHY? -- 2025-09-09
  339. Renting GPUs is hilariously cheap -- 2025-09-09
  340. Ex-Miner Turned Local LLM Enthusiast, now I have a Dilemma -- 2025-09-09
  341. Tencent-Hunyuan/HunyuanWorld-Voyager -- 2025-09-09
  342. How the “Kim” dump exposed North Korea's credential theft playbook -- 2025-09-09
  343. Further Adventures in Colorimeter Hacking -- 2025-09-09
  344. 🌟Introducing Art-0-8B: Reasoning the way you want it to with Adaptive Thinking🌟 -- 2025-09-02
  345. Fine Tune Model for Home Assistant? -- 2025-09-02
  346. [Editorial] Claude Code - massive issues -- 2025-09-01
  347. TheDrummer is on fire!!! -- 2025-09-01
  348. NVIDIA-Nemotron-Nano-12B-v2 -- 2025-09-01
  349. Gpt-oss Fine-tuning - now with 60K context length and fits on &lt;13GB VRAM -- 2025-08-30
  350. A PLL For Perfect Pitch -- 2025-08-27
  351. Why are users still able to edit system prompts or memories even after disabling it? -- 2025-08-20
  352. Making your prompts better with GEPA-Lite using Ollama! -- 2025-08-15
  353. Optimizing OpenWebUI's speed through indexing (using PostgreSQL as a back-end) -- 2025-08-11
  354. PSA: DuckDuckGo search in OWUI routes to non-privacy friendly providers like Bing, Google, and Yahoo. -- 2025-08-11
  355. Open-webui Tools for Firewalla -- 2025-08-11
  356. Local RAG with 97% smaller index and Claude Code–compatible semantic search -- 2025-08-10
  357. Trump Announces 100% Tariff on Semiconductors, unless made in US -- 2025-08-08
  358. The Tape Speed Keyboard -- 2025-08-08
  359. My first finetune: Gemma 3 4B unslop via GRPO -- 2025-08-02
  360. Supervised Fine Tuning on Curated Data is Reinforcement Learning -- 2025-08-02
  361. Debugging the Pixel 8 kernel via KGDB -- 2025-07-31
  362. Ollama + Open WebUI -- is there a way for the same query to run through the same model multiple times (could be 3 times, could be 100 times), then gather all the answers together to summarise/count? -- 2025-07-25
  363. Localllama’s (first?) IFTA - I’ll Fine-Tune Anything -- 2025-07-20
  364. fsndzomga/metadspy -- 2025-07-10
  365. osmosis-ai/Osmosis-Apply-1.7B -- 2025-07-10
  366. IntervitensInc/pangu-pro-moe-model -- 2025-07-10
  367. Continual Gradient Low-Rank Projection Fine-Tuning for LLMs -- 2025-07-10
  368. Creating custom kernels for the AMD MI300 -- 2025-07-10
  369. i made a commit message generator that can be used offline and for free -- 2025-07-05
  370. THUDM/GLM-4.1V-9B-Thinking -- 2025-07-05
  371. baidu/ERNIE-4.5-VL-424B-A47B-Base-Paddle -- 2025-07-05
  372. LoRA Fine-Tuning Without GPUs: A CPU-Efficient Meta-Generation Framework for LLMs -- 2025-07-05
  373. trycua/cua -- 2025-06-30
  374. ml0-1337/claude-gate -- 2025-06-30
  375. mstrYoda/go-arctest -- 2025-06-30
  376. Accelerating Docker Builds by Halving EC2 Boot Time -- 2025-06-30
  377. Advanced Time Manipulation with GDB -- 2025-06-21
  378. Practical SDR: Getting started with software-defined radio -- 2025-06-21
  379. liaotxcn/Probabilistic-Filters -- 2025-06-19
  380. flohoss/gocron -- 2025-06-19
  381. Lessons from Mixing Rust and Java: Fast, Safe, and Practical -- 2025-06-19
  382. 100 prisoners and a lightbulb -- looking back -- 2025-06-19
  383. Paper2Poster/Paper2Poster -- 2025-06-10
  384. Octoberfest7/zip_smuggling -- 2025-05-30
  385. Silencing Firefox's Chattiness for Web App Testing -- 2025-05-30