AI Agents

Agent frameworks, autonomy, MCP, tool use, multi-agent orchestration

1793 articles across 327 editions

Articles

  1. Sleeper Agents and How to Tame Them (TNG Technology Consulting) -- 2026-10-08
  2. Seeing Red, Thinking Bad: Color Bias in Vision Language Models (Stealth Visual Prompts) -- 2026-10-08
  3. reddit.com -- 2026-10-08
  4. The Agent Said It Was Done. The Database Disagreed. -- 2026-10-07
  5. [Editorial] polytoken.dev -- 2026-10-07
  6. [Editorial] YouTube: 0otdTJa9IUc -- 2026-10-07
  7. [Editorial] Sierra introduces the Personal Agent Protocol -- 2026-10-07
  8. [Editorial] YouTube: 0j8wy4jUlsw -- 2026-10-07
  9. [Editorial] RuOS by Cognitum: Rust/WASM agentic OS for swarm and edge agents -- 2026-10-07
  10. [Editorial] Vulnerability in agents from Google and others exposes structural flaw in MCP (Ars Technica) -- 2026-10-06
  11. [Editorial] YouTube video (I3iRsti-Lmw) -- 2026-10-06
  12. [Editorial] LinkedIn post (lnkd.in/p/djg-UhfE) -- 2026-10-06
  13. [Editorial] getpaseo/paseo on GitHub -- 2026-10-06
  14. I got tired of hunting through AI coding sessions, so I put them on a physical control deck -- 2026-10-06
  15. My agent could do everything except get past a login page, so I built it a browser that hands the login to my phone -- 2026-10-05
  16. [Editorial] REAmon (xdCloudy/REAmon) on GitHub -- 2026-10-05
  17. Soniavasseur/Wallet-Risk-Scanner -- 2026-10-05
  18. [Editorial] Reuven Cohen: Use ChatGPT and Claude for unpublished research -- 2026-10-05
  19. IOActive: LLM-Assisted Vulnerability Research — Finding Real Bugs with Code-Reasoning Models -- 2026-10-02
  20. TROOPERS26 Talk: AI and Offensive Security (Conference Session) -- 2026-10-02
  21. vphone-cli: Virtual iPhone from the Command Line -- 2026-10-02
  22. Editorial Video Pick (YouTube) -- 2026-10-02
  23. Cyberspace Administration of China: AI Regulatory Document (PDF) -- 2026-10-02
  24. Bradley Leimer: Blood Again (LinkedIn) -- 2026-10-02
  25. mvanhorn/agent-tincan -- 2026-10-01
  26. [Editorial] -- 2026-10-01
  27. [Editorial] -- 2026-10-01
  28. reddit.com -- 2026-10-01
  29. SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response -- 2026-10-01
  30. [Editorial] -- 2026-10-01
  31. [Editorial] How AI agents exposed developer screenshots from leading tech companies -- 2026-09-30
  32. [Editorial] Meta's Muse AI assistant has a zero-day that can turn it into a Mac backdoor -- 2026-09-30
  33. Grok answered me in my own voice , twice and in two languages, then denied it. -- 2026-09-30
  34. Holo4: powering generalist computer-use agents -- 2026-09-29
  35. Imp is a full port of DSPy to the BEAM -- 2026-09-29
  36. Cloudflare: Rust Workers via the Emscripten target -- 2026-09-29
  37. [Editorial] -- 2026-09-28
  38. [Editorial] -- 2026-09-28
  39. The agent exited cleanly with status 0, did nothing, and reported success -- 2026-09-28
  40. [Editorial] paperclipai/paperclip: agent orchestration -- 2026-09-25
  41. [Editorial] Editor's video pick -- 2026-09-25
  42. [Editorial] Asymptote Labs: agent-beacon -- 2026-09-25
  43. AI coding has made CI a bottleneck, so we reworked ours to keep up -- 2026-09-25
  44. [Editorial] ThreatDown: Carbonato malware analysis -- 2026-09-25
  45. iQingshan/Toshell: single-binary C2 framework with team server, web console and multi-platform implants -- 2026-09-25
  46. [Editorial] Bypassing EDR with local AI -- 2026-09-25
  47. Heretic removes restrictions from language models -- 2026-09-25
  48. [Editorial] Transluce: Agent Activity (transluce.org) -- 2026-09-24
  49. Exfiltrate Your Weights -- 2026-09-24
  50. an (unsafe) command shell mcp server -- 2026-09-24
  51. [Editorial] Google AI Overview Jailbreaking Itself (aifails.substack.com) -- 2026-09-24
  52. Autonomous AI agents are hitting online retailers (Gambit Security) -- 2026-09-23
  53. What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection under Browser Automation -- 2026-09-23
  54. agent-substrate/substrate: infrastructure for autonomous agents -- 2026-09-23
  55. okf-memory/okf-agent-memory -- 2026-09-23
  56. [Editorial] botsbench: GLM-5.3 on the CyBT CTF benchmark -- 2026-09-22
  57. [Editorial] cyberkimi-benchmarks ExploitBench: CVE-2024-6100 -- 2026-09-22
  58. [Editorial] YouTube: 1ADD60wyrbg -- 2026-09-22
  59. [Editorial] Clemens865/Workspace-OS -- 2026-09-22
  60. [Editorial] arXiv 2609.21032 -- 2026-09-22
  61. [Editorial] lsp-client/LSAP -- 2026-09-22
  62. [Editorial] iwe-org/iwe -- 2026-09-22
  63. The Proxy That Made No Sense (Phrack 73) -- 2026-09-21
  64. Editorial video pick (YouTube, 5vbl5FL-nsI) -- 2026-09-21
  65. [Editorial] Anthropic — Mythos 5 Incident Transcript -- 2026-09-17
  66. [Editorial] Video Feature (OhOmLqR5nN4) -- 2026-09-17
  67. [Editorial] Video Feature (98syxABbUPk) -- 2026-09-17
  68. reddit.com -- 2026-09-17
  69. [Editorial] erudenko on LinkedIn: AI developer productivity with Claude -- 2026-09-16
  70. [Editorial] typesafe.ai -- 2026-09-16
  71. [Editorial] Meko Data (mekodata.ai) -- 2026-09-16
  72. What is Missing from AI Post-Training AI: An Empirical Analysis -- 2026-09-15
  73. Semantic Bandits: In-Context Exploration-Exploitation is Biased by Semantic Priors -- 2026-09-15
  74. Your Agent Aced the Task. Will It Do It Again? -- 2026-09-15
  75. Hot take: the agentic workflow is deeply wrong -- 2026-09-15
  76. OpenAI is building Codex Replay, a tool that invites Claude Code users to put Codex head-to-head on their own work — rerunning imported tasks and comparing the results -- 2026-09-15
  77. My NEW FAVORITE Skill - Claude Code Drives My Whole Computer (Better Computer Use) -- 2026-09-15
  78. How well do agents use test/verification techniques? -- 2026-09-10
  79. [Editorial] google/mantis (GitHub) -- 2026-09-10
  80. [Editorial] YouTube video (SGodxQHnVxc) -- 2026-09-10
  81. Pi Agent Users - Nvidia Released Sol-Pi - A Pi-Extension based on AutoResearch loops to make the Harness more efficient -- 2026-09-10
  82. Multi-Agents LLM Financial Trading Framework -- 2026-09-10
  83. [Editorial] Daniel Miessler: The Socrates Agent -- 2026-09-10
  84. [Editorial] From Prompting to Autonomy: The Evolution of Adversarial AI (Google Threat Intelligence) -- 2026-09-09
  85. The OpenAI Huggingface incident from an agents POV -- 2026-09-09
  86. [Editorial] -- 2026-09-07
  87. browser-use/macos-harness -- 2026-09-07
  88. [Editorial] -- 2026-09-07
  89. [Editorial] -- 2026-09-07
  90. [Editorial] -- 2026-09-07
  91. [Editorial] -- 2026-09-07
  92. LG smart TVs caught logging audio with screen off and snooping on local devices -- 2026-09-07
  93. Headlong: An open source agent microharness featuring persistent agency and recursive LLMs -- 2026-08-31
  94. ApodexAI/FrontierAgent -- 2026-08-31
  95. What I Learned About AI Trust from Reconciling over 100B Transactions -- 2026-08-31
  96. Creepy Crawlies -- 2026-08-31
  97. Mechanical Turk shutting down September 30 -- 2026-08-31
  98. [Editorial] Visa open-sources a vulnerability-hunting agentic harness -- 2026-08-28
  99. How Much Memory Does Your Agent Actually Need? -- 2026-08-28
  100. Grok Build Max -- 2026-08-28
  101. [Editorial] Fences, Not Sandboxes -- 2026-08-25
  102. [Editorial] lionagi -- 2026-08-25
  103. [Editorial] Video Feature II -- 2026-08-25
  104. GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge -- 2026-08-25
  105. FrontisAI/OpenRSI -- 2026-08-25
  106. Lynricsy/OneSSH -- 2026-08-21
  107. What sandbox are you all using for AI agents? -- 2026-08-21
  108. [Editorial] LinkedIn feature -- 2026-08-21
  109. coleam00/skills -- 2026-08-21
  110. [Editorial] browser-use/browsercode -- 2026-08-21
  111. Agent Zero v2.10: the Browser now works with Gmail and sites that used to block it, plus @ tagging for everything and ACP support for external editor -- 2026-08-21
  112. NVIDIA dropped an NVIDIA-hosted CUDA MCP for AI-assisted CUDA operations, such as searching official, up-to-date documentation, writing optimized GPU code, and analyzing performance data -- 2026-08-21
  113. [Editorial] Reuven Cohen: how I build so much -- 2026-08-20
  114. [Editorial] ruvnet/LatentMesh -- 2026-08-20
  115. fx: Tiny, open, native coding agent -- 2026-08-20
  116. [Editorial] DreamLab AI Workshops -- 2026-08-20
  117. Estonia wants to give AI agents their own digital IDs -- 2026-08-20
  118. [Editorial] X Trending Topic -- 2026-08-20
  119. [Editorial] CSA Research Note: AI-Driven Threat Actor Autonomous Exploitation -- 2026-08-20
  120. Researchers created "mind viruses" that spread between AI agents by convincing one agent to adopt an idea then transmit it onwards to other agents. -- 2026-08-20
  121. Malicious Rust crate Arrayref runs a build-time payload -- 2026-08-20
  122. Israel creates fake think tank in likely attempt to dupe AI chatbots -- 2026-08-19
  123. The AI Credit Resale Economy -- 2026-08-19
  124. FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial Workflows -- 2026-08-18
  125. [Editorial] Editor's Pick (LinkedIn) -- 2026-08-18
  126. [Editorial] Editor's Pick (Video) -- 2026-08-18
  127. AI agents lie, cheat and steal. That is putting off users -- 2026-08-18
  128. Florida man told ChatGPT he'd murder his ex. OpenAI alerted the FBI -- 2026-08-18
  129. Token Budget Saturation and Mechanistic Early Detection of Reasoning Non-Convergence in Chain-of-Thought Models -- 2026-08-18
  130. Decoding Claude's DNA: Comparing System Prompts Across Fable 5, Opus 5/4.8/4.6, Sonnet 5 & Haiku 4.5 -- 2026-08-18
  131. [Editorial] Be The Adversary: DGX Spark Red Team Bench -- 2026-08-17
  132. Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents -- 2026-08-17
  133. [Editorial] nicobailon/pi-web-access -- 2026-08-17
  134. [Editorial] CodeWiki: ruvnet/ruvector -- 2026-08-17
  135. [Editorial] VisionFlow -- 2026-08-17
  136. wanmol/goal-flow -- 2026-08-17
  137. [Editorial] ruvnet/llm-stream-reformat -- 2026-08-17
  138. sv-number/skills — Give your AI agent a phone number and read SMS verification codes over API -- 2026-08-14
  139. Nvidia Nemo Switchyard -- 2026-08-14
  140. [Editorial] ruvnet/dream-machine (GitHub) -- 2026-08-14
  141. Adaptive and Explicit safe: Triggering Latent Safety Awareness in Large Reasoning Models -- 2026-08-14
  142. [Editorial] LinkedIn Post (gPEg3ZXB) -- 2026-08-14
  143. [Editorial] Infisical Agent Vault -- 2026-08-13
  144. [Editorial] jedarden/seam -- 2026-08-13
  145. [Editorial] LinkedIn feature -- 2026-08-13
  146. My Agent Setup -- 2026-08-12
  147. [Editorial] -- 2026-08-12
  148. AmazingAng/old-coder -- 2026-08-12
  149. oil-oil/codex-deepseek-subagent -- 2026-08-12
  150. AML-memory/agent-memory-leaderboard -- 2026-08-12
  151. Understanding the Rejection of Fixes Generated by Agentic Pull Requests -- Insights from the AIDev Dataset -- 2026-08-10
  152. [Editorial] -- 2026-08-10
  153. [Editorial] -- 2026-08-10
  154. I gave a Claude Fable 5 agent a domain and $90 it can't spend without me. It named itself Cairn. -- 2026-08-07
  155. [Editorial] Agentic Coding Ladder -- 2026-08-07
  156. Poirot — Deep Research Agent Kernel -- 2026-08-07
  157. [Editorial] Agent Plugins -- 2026-08-07
  158. TIME Is Serving AI Bots a Different Website, with Ads Built In -- 2026-08-07
  159. [Editorial] Stop Letting AI Flatter You — Prompt That Gets the Truth -- 2026-08-07
  160. Building an open, challengeable evidence portfolio with multi-agent AI -- 2026-08-07
  161. NVIDIA-NeMo/labs-OO-Agents -- 2026-08-06
  162. [Editorial] -- 2026-08-06
  163. Gemini Spark can now tap into Google Chrome's auto browse feature to take care of complex errands for you online -- 2026-08-06
  164. Persistent background agents may matter more than another coding benchmark -- 2026-08-06
  165. NEURA for Browser — Open WebUI agent mode in every browser tab -- 2026-08-05
  166. SimpleEnglish — Agent skill for ASD-STE100 Simplified Technical English -- 2026-08-05
  167. [Editorial] Awesome AI Tokenomics — Curated AI Cost & Pricing Reference -- 2026-08-05
  168. [Editorial] YouTube: AI Technical Content -- 2026-08-05
  169. [Editorial] AI-Assisted Vulnerability Management -- 2026-08-03
  170. Show HN: Nightcrawler – A local AI pentesting agent running on a smartphone -- 2026-08-03
  171. Perplexity Numbat – Visibility into AI agent activity on endpoints -- 2026-08-03
  172. [Editorial] Video Content -- 2026-08-03
  173. [Editorial] Arxiv Research Paper -- 2026-08-03
  174. [Editorial] QM by YC Software -- 2026-08-03
  175. [Editorial] Agentic Workflow Guide -- 2026-08-03
  176. [Editorial] Claude Code Artifact -- 2026-08-03
  177. [Editorial] RPS Interactive Demo -- 2026-08-03
  178. You can't solve computer use by ignoring the interface -- 2026-07-31
  179. GraphFlow: Formally Verifiable Visual Workflows for Reliable Agentic AI -- 2026-07-31
  180. MarbleOS — What should the GUI for AI agents look like? -- 2026-07-31
  181. [Editorial] -- 2026-07-30
  182. [Editorial] -- 2026-07-30
  183. [Editorial] -- 2026-07-30
  184. [Editorial] -- 2026-07-30
  185. [Editorial] -- 2026-07-30
  186. [Editorial] -- 2026-07-30
  187. Claude thought I could be having a stroke. I was. -- 2026-07-30
  188. [Editorial] NanoNets/Graft -- 2026-07-29
  189. [Editorial] jcode -- 2026-07-29
  190. agentacct: Local-first agent work intelligence for coding agents -- 2026-07-29
  191. pocketdev: One command to run AI coding CLIs on a remote Hetzner box via Tailscale -- 2026-07-29
  192. handdraw-story-video: Turn hand-drawn illustrations into line-reveal videos -- 2026-07-29
  193. The first documented case of an end-to-end ransomware operation executed autonomously by an LLM has successfully performed extortion without a human operator. -- 2026-07-28
  194. [Editorial] -- 2026-07-28
  195. Anthropic's stance on cybersecurity is completely backwards -- 2026-07-28
  196. An0nUD4Y/Offensive-COM -- 2026-07-28
  197. [Editorial] -- 2026-07-28
  198. Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems -- 2026-07-28
  199. Don't ask an LLM for a confidence score -- 2026-07-28
  200. GigaToken: ~1000x faster Language model tokenization -- 2026-07-27
  201. Llama.cpp now has full MCP support! -- 2026-07-27
  202. Introducing DWARF-55M-Base -- 2026-07-27
  203. New Model: Nanbeige4.2-3B (Looped Transformer, outperforms 4x size) -- 2026-07-27
  204. I built a Triton backend for Falcon3-10B-1.58bit: 97.5 tok/s decode on an RTX 5070 -- 2026-07-27
  205. Extended garlic to run Qwen3.5 35B A3B float8 at 55 tok/s on RTX 5060 Ti -- 2026-07-27
  206. [Editorial] -- 2026-07-27
  207. Show HN: Palmier Pro – Open-source macOS video editor built for AI -- 2026-07-27
  208. [Editorial] -- 2026-07-27
  209. [Editorial] -- 2026-07-27
  210. Towards Agentic Investigation of Security Alerts -- 2026-07-27
  211. [Editorial] -- 2026-07-27
  212. [Editorial] -- 2026-07-27
  213. [Editorial] -- 2026-07-27
  214. [Editorial] -- 2026-07-27
  215. Agent swarms and the new model economics -- 2026-07-24
  216. Nanako0129/pilotfish -- 2026-07-24
  217. OpenWorker by Andrew Ng -- 2026-07-24
  218. Agentic Kit -- 2026-07-24
  219. RuVector Explainer -- 2026-07-24
  220. ANSI escape injection in MCP servers: Hidden from humans, visible to AI -- 2026-07-23
  221. Stop Using OpenCode -- 2026-07-23
  222. [Editorial] -- 2026-07-23
  223. [Editorial] -- 2026-07-23
  224. [Editorial] -- 2026-07-23
  225. Code mode yields a 99.2% cost reduction in our systems -- 2026-07-23
  226. Jack Dorsey launches Buzz to combine team chat, AI agents and Git hosting -- 2026-07-23
  227. [Editorial] -- 2026-07-23
  228. [Editorial] -- 2026-07-22
  229. [Editorial] -- 2026-07-22
  230. Laguna S 2.1 -- 2026-07-22
  231. Sahir619/fable-method -- 2026-07-22
  232. [Editorial] -- 2026-07-22
  233. I Love the Karpathy LLM Wiki but it Doesn't Scale. Here's What Does. -- 2026-07-20
  234. [Editorial] Incident-to-Eval Synthesis — AI Pattern Book -- 2026-07-20
  235. [Editorial] HuggingFace Security Incident — July 2026 -- 2026-07-20
  236. Hacker wipes Romania's land registry database -- 2026-07-20
  237. [Editorial] CapitalOne VulnHunter -- 2026-07-20
  238. ClaudeBrain — Karpathy LLM Harness for Pentesting/Bugbounty -- 2026-07-20
  239. [Editorial] The Sharpest Questions About AI -- 2026-07-20
  240. [Editorial] CommandCode.ai -- 2026-07-20
  241. [Editorial] Kimchi.dev Coding -- 2026-07-20
  242. Kastor — Terraform-Style Source-of-Truth Layer for AI Agents -- 2026-07-20
  243. What building Shippy taught us about building agents -- 2026-07-17
  244. [2607.07508] Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning -- 2026-07-17
  245. [Editorial] -- 2026-07-17
  246. A read-only triage subagent wrote its own jailbreak on turn 1 (no poisoned input anywhere) -- 2026-07-17
  247. Codex starts encrypting prompts, uses ciphertext for inference instead -- 2026-07-17
  248. The human-in-the-loop is tired -- 2026-07-17
  249. Who Owns This Agent? Tracing AI Agents Back to Their Owners -- 2026-07-16
  250. [Editorial] -- 2026-07-16
  251. Show HN: Juggler – an open-source GUI coding agent, by the creator of JUCE -- 2026-07-15
  252. [Editorial] -- 2026-07-15
  253. nagisanzenin/engram -- 2026-07-15
  254. [Editorial] -- 2026-07-15
  255. Two Claude Code agents, two worktrees, one port: parallel agents don't collide on code, they collide on runtime -- 2026-07-15
  256. The Economics of Recursive Self-Improvement -- 2026-07-14
  257. What Will Be Left for Us to Work On? -- 2026-07-14
  258. [Editorial] Video Submission -- 2026-07-14
  259. Apple's New SpeechAnalyzer API, Benchmarked Against Whisper -- 2026-07-14
  260. YouTube Guitar Tab Parser — Claude Vision Extracts Tabs from Video -- 2026-07-14
  261. Blender + Seedance Workflows for AI Filmmaking -- 2026-07-14
  262. OpenEnvision/WorldFoundry: Unified World Model Inference & Evaluation -- 2026-07-14
  263. [Editorial] oomol-lab/open-connector -- 2026-07-14
  264. [Editorial] OakLab AI Mission -- 2026-07-14
  265. [Editorial] -- 2026-07-13
  266. Xingyu-Zheng/MrFlow -- 2026-07-13
  267. Every day my local agent reads all new arXiv papers in my field and tells me which ones matter to MY research -- 2026-07-13
  268. Control the Ideas, Not the Code -- 2026-07-13
  269. [Editorial] -- 2026-07-13
  270. [Editorial] -- 2026-07-13
  271. [Editorial] -- 2026-07-13
  272. [Editorial] -- 2026-07-10
  273. [Editorial] -- 2026-07-10
  274. Koder: browser UI based harness for coding and computer use -- 2026-07-10
  275. Using "applications" to make a smaller model more effective at bigger tasks. -- 2026-07-10
  276. SAGA: Workflow-Atomic Scheduling for AI Agent Inference on GPU Clusters -- 2026-07-09
  277. OfficeCLI: Office suite for AI agents to read and edit Microsoft Office files -- 2026-07-09
  278. Toolport: Use as many MCP servers as you want without the token tax -- 2026-07-09
  279. GPT-5.6 Sol Ultra will be in Codex -- 2026-07-09
  280. Qwen 3.6 27B absolutely fails at agentic work -- 2026-07-09
  281. I tested Anthropic's new Jacobian Lens on open models, then it turned into a local-model hallucination router -- 2026-07-09
  282. Complete local model asset generation pipeline -- 2026-07-09
  283. krea-ai/krea-2 -- 2026-07-09
  284. I made a tool that chains a small local model into a big coding model and auto-unloads VRAM between them -- 2026-07-09
  285. Ternlight – 7 MB embedding model that runs in browser (WASM) -- 2026-07-09
  286. I tested freshly merged DFlash in llama.cpp on Qwen 3.6 27B Local AI win. 4.44x faster at 36K context. Here are my findings RTX 6000 PRO. -- 2026-07-08
  287. Ollama 0.31: Faster Gemma 4 on Apple Silicon with MTP. Here is my test showing a 56% boost on M1 Pro 16GB (2021) -- 2026-07-08
  288. AMD Ryzen AI Halo – $4k AI Dev Kit -- 2026-07-08
  289. So... anyone copped one of these? -- 2026-07-08
  290. Uh.. Honey, how do you feel about takeout? -- 2026-07-08
  291. COMAP: Co-Evolving World Models and Agent Policies for LLM Agents -- 2026-07-08
  292. NVlabs/SpatialClaw -- 2026-07-08
  293. LeRobot v0.6.0: Imagine, Evaluate, Improve -- 2026-07-08
  294. AI-Builder-Club/skills -- 2026-07-08
  295. My Agentic Workbench -- 2026-07-08
  296. [Editorial] -- 2026-07-08
  297. Feeling dumb day by day after using claude code -- 2026-07-08
  298. SentinelMCP: Open-Source MCP Firewall & Security Gateway for AI Agents -- 2026-07-07
  299. Codex CLI Jailbreak Guide — Customizing System Prompt via model_instructions_file -- 2026-07-07
  300. [Editorial] Video Feature -- 2026-07-07
  301. Waveloom: Terminal Coding Agent Optimized for DeepSeek Prefix Caching — 95-99% Cache Hit, 1/50th Cost -- 2026-07-07
  302. Qwen3.6-27B Vibecodes A* Pathfinding in Java Game — 12 Hours of Autonomous Testing -- 2026-07-07
  303. Codex Builds Pikachu Volleyball in UmLang — an Obscure Korean Esoteric Programming Language -- 2026-07-07
  304. [Editorial] @metaharness/flywheel -- 2026-07-07
  305. [Editorial] Beijing Looking to Curb Overseas Access to China's Top AI Models -- 2026-07-07
  306. Alibaba to ban Claude Code in workplace over alleged backdoor risks, source says -- 2026-07-07
  307. Virginia bans sale of geolocation data -- 2026-07-07
  308. [Editorial] Open-Source AI Coding Agents Shell Injection Vulnerability -- 2026-07-06
  309. [Editorial] arXiv Research Paper 2606.24496 -- 2026-07-06
  310. Is it ever possible to have a malicious LLM with a backdoor -- 2026-07-06
  311. Android Developer Verification: Threat masquerading as Protection -- 2026-07-06
  312. [Editorial] Video Content -- 2026-07-06
  313. [Editorial] The Future of Agentics Is Not the Model -- 2026-07-06
  314. [Editorial] Loop Came Home -- 2026-07-06
  315. [Editorial] ruvnet Technical Gist -- 2026-07-06
  316. [Editorial] ruvnet Technical Gist -- 2026-07-06
  317. [Editorial] Video Content -- 2026-07-06
  318. [Editorial] Video Content -- 2026-07-06
  319. [Editorial] Video Content -- 2026-07-06
  320. Does code cleanliness affect coding agents? A controlled minimal-pair study -- 2026-07-06
  321. The Safari MCP server for web developers -- 2026-07-06
  322. cellebrite-labs/ghidra-rpc — Agentic Reverse Engineering Skill for Ghidra -- 2026-07-03
  323. raiyanyahya/recall — Durable Offline Memory for Claude Code -- 2026-07-03
  324. [Editorial] Agentic Rust Optimizer -- 2026-07-03
  325. [Editorial] Nexu Open Design -- 2026-07-03
  326. [Editorial] Agents Optimizing Their Own Behavior -- 2026-07-03
  327. Forsy-AI/agent-apprenticeship -- 2026-07-03
  328. Scalable Inference Architectures for Compound AI Systems: A Production Deployment Study -- 2026-07-03
  329. OmniAct: Advancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical Autonomy -- 2026-07-03
  330. [Editorial] OpenAI Proposes US Government Own 5% Stake -- 2026-07-03
  331. OpenAI: In early talks to give 5% stake to US Government -- 2026-07-03
  332. The number 1 public enemy of open-source. -- 2026-07-03
  333. Model Registry: Torrents for open models using Hugging Face as a fallback web seed. -- 2026-07-03
  334. Anatomy of a Failed (Nation-State?) Attack -- 2026-07-01
  335. projectdiscovery/depx -- 2026-07-01
  336. Claude Code suddenly tried to open a Remote Desktop connection on my PC. This seriously scared me. -- 2026-07-01
  337. Google's masterclass on agentic engineering patterns -- 2026-06-30
  338. [Editorial] AgentBBS — Bulletin Board System for AI Agents -- 2026-06-30
  339. OpenKnowledge: Open source AI-first alternative to Obsidian/Notion with Claude/Codex integration -- 2026-06-30
  340. Bingo - AI-powered Red Team Terminal (DeepSeek/Claude/GPT/GLM) -- 2026-06-30
  341. [Editorial] REcon Conference for Reversers and Security Researchers -- 2026-06-30
  342. Enhancing X11 Application Security with LXC -- 2026-06-30
  343. TxBench-PP: Analyzing AI Agent Performance on Small-Molecule Preclinical Pharmacology -- 2026-06-29
  344. Herdr: Agent multiplexer that lives in your terminal -- 2026-06-29
  345. [Editorial] -- 2026-06-27
  346. [Editorial] -- 2026-06-27
  347. OpenAI Codex has a bug that could kill your SSD in under a year -- 2026-06-27
  348. [Editorial] -- 2026-06-27
  349. [Editorial] -- 2026-06-27
  350. [Editorial] -- 2026-06-27
  351. [Editorial] -- 2026-06-27
  352. Unlimited-OCR is now on ModelScope! A 3.3B multilingual OCR model for one-shot parsing across single images, multi-page documents, and PDFs. License: MIT -- 2026-06-27
  353. [Editorial] Xiaomi HarnessX — self-rewriting AI scaffolding -- 2026-06-26
  354. yzfly/TokenCode -- 2026-06-26
  355. How I'm handling per-agent isolation and environment lifecycle in a harness-agnostic orchestration library -- 2026-06-26
  356. Andrezi: a local-first memory governance layer for Claude Code (honest writeup, MIT) -- 2026-06-26
  357. [Editorial] Agentic QE v3.11.1 -- 2026-06-26
  358. Computer Use in Gemini 3.5 Flash -- 2026-06-25
  359. gemini-web2api: Convert Gemini Web to OpenAI-Compatible API — Zero Auth, Single File -- 2026-06-25
  360. [Editorial] Mistral OCR 4 -- 2026-06-25
  361. AADvark: Agent-Aided Design for Dynamic CAD Models with Moving Parts -- 2026-06-25
  362. [Editorial] Agentic Context Engine -- 2026-06-19
  363. [Editorial] The Most Important Idea in AI Today — Reuven Cohen -- 2026-06-19
  364. [Editorial] Video Pick 2 -- 2026-06-19
  365. Managing entire business banking through Claude MCP -- 2026-06-19
  366. [Editorial] The Flat Curve Society — Steve Yegge -- 2026-06-19
  367. Peopleless economy? Not technically impossible -- 2026-06-19
  368. Local coding agents are good now, but only if you babysit them -- 2026-06-18
  369. STOP Using Claude Code Without This Fable 5 Agentic OS -- 2026-06-18
  370. [Editorial] -- 2026-06-18
  371. openclaw/agent-skills -- 2026-06-18
  372. Agentic Resource Discovery: Let agents search -- 2026-06-18
  373. henliveira/av-curator -- 2026-06-18
  374. [Editorial] Cursor: Agent Autonomy & Auto-Review -- 2026-06-17
  375. paradigmxyz/centaur — Multiplayer, Self-Hosted, Secure Agents -- 2026-06-17
  376. tastyeffectco/sandboxes — Self-Hosted Dev Sandboxes -- 2026-06-17
  377. Autonomous LLM-Guided Disease Forecasting Matches CDC Expert Ensembles in Prospective Evaluation -- 2026-06-16
  378. AI Giants Score Below 25% in UC Berkeley-Led Test of Real-World Application Across 50+ Industries -- 2026-06-16
  379. [Editorial] Research Paper -- 2026-06-16
  380. [Editorial] Claude Fable 5 Made This Entire Video -- 2026-06-15
  381. [Editorial] Agent Harness Generator -- 2026-06-15
  382. [Editorial] AI Demo/Showcase Video -- 2026-06-15
  383. [Editorial] Ponytail — Open Source Tool -- 2026-06-15
  384. [Editorial] CISO Perspective on Cyber + AI Convergence -- 2026-06-15
  385. [Editorial] ISO 27001 Meets Agent Security -- 2026-06-15
  386. AzureRedOps — Offensive Security Toolkit for Microsoft Entra ID -- 2026-06-15
  387. [Editorial] AI Tooling Walkthrough Video -- 2026-06-15
  388. [Editorial] Gadi Evron on Forcing Agents to Find -- 2026-06-12
  389. townsendmerino/ken -- 2026-06-12
  390. [Editorial] OB1 Project -- 2026-06-12
  391. [Editorial] Vimeo Feature -- 2026-06-12
  392. Google's Agents CLI: The CLI + Skills Combination to Ship AI Agents EASILY -- 2026-06-12
  393. ultracode is the most powerful claude code feature in months -- 2026-06-12
  394. Top 3 Underrated Open Source Repos Nobody Talks About -- 2026-06-12
  395. [Editorial] -- 2026-06-11
  396. Harness Engineering: What Separates Top Agentic Engineers Right Now -- 2026-06-11
  397. Claude Plans, Gemini Designs: One Workflow for Beautiful Frontends (LIVE) -- 2026-06-11
  398. [Editorial] Cole Medin: Most Powerful AI Coding Setup -- 2026-06-11
  399. Codex Remote is a GAME CHANGER -- 2026-06-11
  400. This Claude Code + Obsidian Command Center is INSANE -- 2026-06-11
  401. Top 5 Web Design Plugins for Claude Code -- 2026-06-11
  402. How an Agent Built a 3D Paris Gallery by Chaining Two Hugging Face Spaces -- 2026-06-09
  403. Leap in DNA synthesis slashes time to build new genetic sequences -- 2026-06-09
  404. Thi.ng – open-source building blocks for computational design and art -- 2026-06-09
  405. Exploring the Potential of Probabilistic Transformer for Time Series Modeling: A Report on the ST-PT Framework -- 2026-06-09
  406. Efficient-Large-Model/SANA-WM_bidirectional -- 2026-06-09
  407. [Editorial] -- 2026-06-09
  408. ajsai47/backdoor -- 2026-06-09
  409. This Open Source Repo Just Solved Claude Code's #1 Problem -- 2026-06-09
  410. KyrieCheungYep/ky-design-to-html-skill -- 2026-06-09
  411. Show HN: Gitdot – A better GitHub. Open-source, written in Rust -- 2026-06-09
  412. [Editorial] -- 2026-06-08
  413. The Most Powerful Claude Code Feature In Months Dropped & Nobody is Talking About It -- 2026-06-08
  414. ultracode is INSANE and nobody is talking about it -- 2026-06-08
  415. ethanhq/cc-fleet -- 2026-06-08
  416. [Editorial] -- 2026-06-08
  417. [Editorial] -- 2026-06-08
  418. [Editorial] -- 2026-06-08
  419. [Editorial] -- 2026-06-08
  420. perplexityai/bumblebee -- 2026-06-08
  421. wexaai/cognodb -- 2026-06-08
  422. [Editorial] -- 2026-06-08
  423. [Editorial] -- 2026-06-08
  424. MisoLabs/MisoTTS -- 2026-06-08
  425. Anthropic Just Dropped a Masterclass on Building Agent Harnesses (for Large Codebases) -- 2026-06-05
  426. [Editorial] chopratejas/headroom -- 2026-06-05
  427. Is It Time To Switch to Codex? -- 2026-06-05
  428. The Storyboard Trick That Stops AI Slop Code -- 2026-06-05
  429. rulyone/Simple-ReAct-Agent -- 2026-06-05
  430. AEM: Adaptive Entropy Modulation for Multi-Turn Agentic Reinforcement Learning -- 2026-06-05
  431. Deciphering Shortcut Learning from an Evolutionary Game Theory Perspective -- 2026-06-05
  432. DynaTree: Dynamic Agentic Retrieval Tree for Time-Sensitive News Retrieval -- 2026-06-04
  433. Property-Guided LLM Program Synthesis for Planning -- 2026-06-04
  434. Conversational Demand Response: Bidirectional Aggregator-Prosumer Coordination through Agentic AI -- 2026-06-04
  435. mims-harvard/AutoScientists -- 2026-06-04
  436. [Editorial] -- 2026-06-04
  437. [Editorial] -- 2026-06-04
  438. [Editorial] -- 2026-06-04
  439. You Don't Understand the Power of a Claude Code Agentic OS -- 2026-06-04
  440. [Editorial] Agentic Tracebit — AI-Powered Security Deception -- 2026-06-03
  441. puck-security/puck-scout -- 2026-06-03
  442. V0id-v2/Void-Tools-v2.0 -- 2026-06-03
  443. ClaudioDrews/memory-os -- 2026-06-03
  444. Gograph: AST-Based MCP Server That Cuts Claude Code Token Use by 95% in Go Repos -- 2026-06-03
  445. [Editorial] The Future of Software Development -- 2026-06-03
  446. I ran 8 open-weight models as agents in a persistent MMO for 10 days. Here's the 93k event dataset and some things that I learned -- 2026-06-01
  447. How Qwen3.6-35B-A3B fails differently as a sub agent compared to solo -- 2026-06-01
  448. I built a computer use sandbox framework for codex on headless linux. GPU passthrough, computer use, and sudo access for codex all work. -- 2026-06-01
  449. ZJU-REAL/SDAR -- 2026-06-01
  450. Claude Opus 4.8 -- 2026-06-01
  451. [Editorial] -- 2026-06-01
  452. New DeepSWE benchmark finds Claude Opus cheats -- 2026-05-29
  453. ITBench-AA: Frontier Models Score Below 50% on Enterprise IT Tasks — by Artificial Analysis and IBM -- 2026-05-29
  454. Context, Reasoning, and Hierarchy: Cost-Performance Study of Compound LLM Agent Design in an Adversarial POMDP -- 2026-05-29
  455. 10 years of AI robustness tricks (PGD, RLHF, Data Augmentation) are actually computing the same hidden matrix -- 2026-05-29
  456. [Editorial] Dynamic Workflows in Claude Code -- 2026-05-29
  457. Using AI to write better code more slowly -- 2026-05-29
  458. Overnight autonomous coding with Claude Code -- 2026-05-29
  459. [Editorial] Harness Engineering for AI Coding -- 2026-05-28
  460. [Editorial] pacphi Gist -- 2026-05-28
  461. Patdolitse/piia-engram -- 2026-05-28
  462. yliust/Tactile: accessibility-first operating layer for agents -- 2026-05-28
  463. hadriansecurity/OpenHack -- 2026-05-28
  464. OpenAI cofounder Karpathy joins Anthropic to teach Claude to improve itself without humans -- 2026-05-28
  465. China Clamps Down on Overseas Travel for AI Talent at Alibaba, DeepSeek -- 2026-05-28
  466. MOSS: Self-Evolution through Source-Level Rewriting in Autonomous Agent Systems -- 2026-05-27
  467. Bian Que: An Agentic Framework with Flexible Skill Arrangement for Online System Operations -- 2026-05-27
  468. Same task in github-copilot, pi, claude-code, and opencode with Qwen3.6 27B -- 2026-05-27
  469. [Editorial] Millions of AI Agents Imperiled by Critical Vulnerability in Open-Source Package -- 2026-05-27
  470. ekomsSavior/Centipede — Self-replicating Linux worm with multi-layer C2 -- 2026-05-27
  471. [Editorial] Editor's Pick (Video) -- 2026-05-27
  472. Dead Weights, Live Signals: Feedforward Graphs of Frozen Language Models -- 2026-05-27
  473. [Editorial] leochlon/ntkmirror -- 2026-05-27
  474. BBuf/kernel-pilot -- 2026-05-27
  475. A vector index can't tell if today's "Karpathy" is the same one it saw yesterday. Here's the fix -- 2026-05-26
  476. [Editorial] -- 2026-05-26
  477. [Editorial] -- 2026-05-26
  478. DanOps-1/Gpt-Agreement-Payment -- 2026-05-26
  479. A Workflow-Oriented Framework for Asynchronous Human-AI Collaboration in Hybrid and Compute-Intensive HPC Environments -- 2026-05-22
  480. Simple Multi-Agent Architecture Running Across Our Entire Org. Keeping everything in Loop. -- 2026-05-22
  481. eight-acres-lab/openmelon -- 2026-05-22
  482. GPT 5.5 (Codex) leading the future prediction race -- 2026-05-22
  483. Agentic Multi-Agent Architecture for Cybersecurity Risk Management -- 2026-05-21
  484. AiSOC: Open-Source AI-Powered Security Operations Center -- 2026-05-21
  485. Cuocuo: Encrypted Tunnel Relay (XChaCha20-Poly1305 + Protobuf) -- 2026-05-21
  486. Trojan's Whisper: Stealthy Manipulation of Coding Agents via Injected Guidance -- 2026-05-21
  487. Agent issued rm -rf / to test its own command blocking -- 2026-05-21
  488. OpenSquilla: Token-Efficient AI Agent -- 2026-05-21
  489. SmallCode: Open-source agentic coding tool stabilized after 90+ bug fixes -- 2026-05-21
  490. Qwen3-Coder-Next lands on HuggingFace -- 2026-05-21
  491. Open Relay v4.1–4.3: Terminal, Performance, and Code Block Overhaul for Open WebUI -- 2026-05-21
  492. No Slop Grenade -- 2026-05-21
  493. Gemini 3.5 Flash -- 2026-05-20
  494. Cursor Introduces Composer 2.5 -- 2026-05-20
  495. [Editorial] -- 2026-05-20
  496. We let AIs run radio stations -- 2026-05-19
  497. Agora-1: The Multi-Agent World Model -- 2026-05-19
  498. Running agents 2x might be the simplest way to improve performance -- 2026-05-19
  499. [Editorial] -- 2026-05-19
  500. Open Source vs frontier models on a single-file HTML canvas driving animation - results -- 2026-05-19
  501. Every Claude Code User NEEDS To Watch This -- 2026-05-19
  502. The reason for the new limits is that SpacexAI is renting its servers to Anthropic. -- 2026-05-19
  503. [Editorial] -- 2026-05-19
  504. [Editorial] -- 2026-05-18
  505. [Editorial] -- 2026-05-18
  506. Claude Mythos Speeds Up macOS Security Exploit Research -- 2026-05-18
  507. tiangolo/library-skills — Library Agent Skills -- 2026-05-15
  508. taito — A package manager for local AI skill/agent bundles -- 2026-05-15
  509. Let's build Claude Code from scratch — nanoclaude -- 2026-05-15
  510. [Editorial] The Graveyard Folder -- 2026-05-15
  511. Claude for Small Business -- 2026-05-15
  512. [Editorial] -- 2026-05-15
  513. Reimagining the mouse pointer for the AI era -- 2026-05-15
  514. Twin Brothers Wipe 96 Government Databases Minutes After Being Fired -- 2026-05-14
  515. [Editorial] Ghost SIM Attack — Black Hat SecTor Research -- 2026-05-14
  516. From Skills to Talent: Organising Heterogeneous Agents as a Real-World Company -- 2026-05-13
  517. GammaLabTechnologies/harmonist -- 2026-05-13
  518. How OpenAI runs its Codex coding agent safely at scale -- 2026-05-13
  519. amanning3390/deepswarm -- 2026-05-13
  520. [Editorial] -- 2026-05-13
  521. [Editorial] -- 2026-05-13
  522. Rant: The realization that most of what ive been calling "evals" has been vibe checks. -- 2026-05-13
  523. [Editorial] -- 2026-05-13
  524. [Editorial] -- 2026-05-13
  525. FaultLine - LLM memory with a bouncer at the door -- 2026-05-13
  526. regent-vcs/re_gent — Version Control for AI Coding Agents -- 2026-05-12
  527. Show HN: adamsreview – better multi-agent PR reviews for Claude Code -- 2026-05-12
  528. nateherkai/token-dashboard — Claude Code Token Cost Analytics -- 2026-05-12
  529. Open WebUI v0.9.3 (and v0.9.4) is out — massive performance wins, message editing finally fixed -- 2026-05-12
  530. We built and open-sourced Caliby: An embedded, high-performance vector database for AI Agents (Beats pgvector by 4x) -- 2026-05-12
  531. [Editorial] Arxiv Research Paper -- 2026-05-12
  532. QKVShare: Quantized KV-Cache Handoff for Multi-Agent On-Device LLMs -- 2026-05-12
  533. MachinaCheck: Building a Multi-Agent CNC Manufacturability System on AMD MI300X -- 2026-05-12
  534. Agents need control flow, not more prompts -- 2026-05-11
  535. [Editorial] NEEDLE — AI Search & Retrieval Tool -- 2026-05-11
  536. [Editorial] Mesh — Agent Mesh Framework -- 2026-05-11
  537. [Editorial] NEEDLE Getting Started Guide -- 2026-05-11
  538. bschoepke/ableton-live-mcp — MCP Bridge for Ableton Live -- 2026-05-11
  539. Recursive Agent Optimization — RL for Recursive Agent Spawning -- 2026-05-08
  540. [Editorial] Anthropic Introducing Dreaming for Agents -- 2026-05-08
  541. Qwen WebWorld 32B/14B/8B — Open Web World Model for Agent Training -- 2026-05-08
  542. Tilde.run — Agent Sandbox with Transactional Versioned Filesystem -- 2026-05-08
  543. [Editorial] -- 2026-05-07
  544. -- 2026-05-07
  545. context-labs/HALO -- 2026-05-07
  546. [Editorial] -- 2026-05-07
  547. GPT 5.5 just leaked its chain of thought to me in codex, and it looks like an idea from 5 months ago in this sub. -- 2026-05-07
  548. Agents can now create Cloudflare accounts, buy domains, and deploy -- 2026-05-07
  549. hacktivist123/agent-session-resume -- 2026-05-07
  550. When everyone has AI and the company still learns nothing -- 2026-05-07
  551. [Editorial] When AI Writes the AI Strategy -- 2026-05-06
  552. [Editorial] Video -- 2026-05-06
  553. Project Deal: Anthropic created a marketplace for their employees & tasked Claude with buying, selling and negotiating on employees behalf. -- 2026-05-06
  554. Here's 45 seconds of Facebook telling me the White House shooter was a former staffer of literally almost every major sports team -- 2026-05-06
  555. Lessons for Agentic Coding: What should we do when code is cheap? -- 2026-05-06
  556. [Editorial] Claude's Multi-Stage Multi-Level Agentic -- 2026-05-06
  557. What's new in CC 2.1.124 (+166 tokens) and 2.1.126 (-87 tokens) system prompt -- 2026-05-06
  558. github.com -- 2026-05-06
  559. Humanoid Robot Actuators -- 2026-05-05
  560. [Editorial] Four Levels of Agentic Software Development -- 2026-05-04
  561. [Editorial] Most Companies Aren't Ready for AI -- 2026-05-04
  562. DeepClaude — Claude Code Agent Loop with DeepSeek V4 Pro -- 2026-05-04
  563. Two Claude Code Agents Collaborating in a Shared Chat Room -- 2026-05-04
  564. paradigm-memory: Local Cognitive Memory MCP for AI Coding Agents -- 2026-05-04
  565. [Editorial] OIA Agentics — Open Interoperability for Agentic AI -- 2026-05-01
  566. Agentic Microphysics: A Manifesto for Generative AI Safety -- 2026-05-01
  567. [Editorial] Agentic AI: Lessons from the Trenches -- 2026-05-01
  568. [Editorial] Human Work Time Allocation in the Hybrid AI Era -- 2026-05-01
  569. Preference-Aligned LoRA Merging: Preserving Subspace Coverage and Addressing Directional Anisotropy -- 2026-04-30
  570. [Editorial] Autonomous Knowledge Graph Exploration -- 2026-04-30
  571. Caveman – Claude Code skill that cuts 75% of tokens by talking like caveman -- 2026-04-29
  572. [Editorial] Matt Pocock's Claude Code Skills -- 2026-04-29
  573. Opencode-power-pack – Claude Code skills ported to OpenCode -- 2026-04-29
  574. [Editorial] ChatGPT Images 2 + Claude Design Guide -- 2026-04-29
  575. Show HN: OSS Agent I built topped the TerminalBench on Gemini-3-flash-preview -- 2026-04-28
  576. yzhao062/anywhere-agents -- 2026-04-28
  577. run-llama/ParseBench -- 2026-04-28
  578. An update on recent Claude Code quality reports -- 2026-04-27
  579. MeshCore development team splits over trademark dispute and AI-generated code -- 2026-04-27
  580. MemPalace: The highest-scoring AI memory system ever benchmarked -- 2026-04-27
  581. [Editorial] AgentBox — Sandboxed Agent Execution -- 2026-04-27
  582. [Editorial] Design Council -- 2026-04-27
  583. [Editorial] Claude Code Game Studios -- 2026-04-27
  584. S. Korea police arrest man over AI image of runaway wolf that misled authorities -- 2026-04-24
  585. jkeatn/Rainmaker — Autonomous Weather Prediction Agent (73% Win Rate on Polymarket) -- 2026-04-24
  586. 0x0funky/agent-sprite-forge — AI Agent Skill for 2D Sprite Sheet Generation -- 2026-04-24
  587. Over-editing refers to a model modifying code beyond what is necessary -- 2026-04-23
  588. Scoring Show HN submissions for AI design patterns -- 2026-04-23
  589. [Editorial] Video Feature -- 2026-04-23
  590. [Editorial] BankerToolBench — Evaluating AI Agents -- 2026-04-23
  591. Kimi vendor verifier – verify accuracy of inference providers -- 2026-04-23
  592. [Editorial] AI Industry Perspective -- 2026-04-23
  593. R2RAG: Routing-to-RAG — Award-Winning Dynamic RAG Architecture -- 2026-04-22
  594. [Editorial] Agent Observability: Required but We're Not There Yet -- 2026-04-22
  595. [Editorial] ArXiv Research Paper -- 2026-04-22
  596. [Editorial] Video Content -- 2026-04-22
  597. [Editorial] Lean Island: Castaneda's Philosophy as System Design -- 2026-04-22
  598. Benchmarked 4 agent memory systems: Mem0 scores 49% recall (worse than a coin flip), Zep uses 340x more tokens for 15 points improvement. Here's what's actually going on. -- 2026-04-21
  599. Open-sourced my OpenWebUI router — semantic routing, citation verification, and per-chat memory for any LLM. -- 2026-04-21
  600. THIS SHOULD NOT BE POSSIBLE IN OPEN WEBUI: LIVE VISUALIZATION RENDERING - Inline Visualizer v2 is HERE! -- 2026-04-21
  601. Is Your Site Agent-Ready? (By Cloudflare) -- 2026-04-21
  602. Codex for almost everything -- 2026-04-21
  603. I'm Building an AI Dark Factory That Ships Its Own Code (Public Experiment) -- 2026-04-21
  604. The PR you would have opened yourself -- 2026-04-21
  605. Anthropic says OpenClaw-style Claude CLI usage is allowed again -- 2026-04-21
  606. [Editorial] -- 2026-04-21
  607. Guy builds AI driven hardware hacker arm from duct tape, old cam and CNC machine -- 2026-04-21
  608. mliu98/awesome-human-distillation -- 2026-04-21
  609. [Editorial] -- 2026-04-21
  610. [Editorial] arxiv:2603.19461 — AI Research Paper -- 2026-04-20
  611. [Editorial] NousResearch Hermes Agent — Open Agentic Framework -- 2026-04-20
  612. [Editorial] Agentic AI with Local LLMs (NousResearch) -- 2026-04-20
  613. [Editorial] Evo — Evolutionary AI Framework -- 2026-04-20
  614. linkedin.com -- 2026-04-20
  615. [Editorial] IETF Agent Authentication Protocol Draft -- 2026-04-17
  616. [Editorial] Video Submission -- 2026-04-17
  617. [Editorial] Hermes Security Upgraded with ClawSec Skill -- 2026-04-17
  618. [Editorial] GitHub Gist Submission -- 2026-04-17
  619. [Editorial] Claude Opus 4.7 Launch -- 2026-04-17
  620. [Editorial] Cole Medin on Opus 4.7 -- 2026-04-17
  621. Qwen3.6-35B-A3B: Agentic coding power, now open to all -- 2026-04-17
  622. [Editorial] Unsloth Qwen3.6 Model Docs -- 2026-04-17
  623. Major drop in intelligence across most major models -- 2026-04-17
  624. [Editorial] My Team Built 30 AI Agents with Claude — Ignored All of Them -- 2026-04-16
  625. [Editorial] Video Content -- 2026-04-16
  626. [Editorial] Your Agent Needs a SOUL.md -- 2026-04-16
  627. [Editorial] Prompt Engineering Is Dead, Long Live Prompt Engineering -- 2026-04-16
  628. InfoSeeker: A Scalable Hierarchical Parallel Agent Framework for Web Information Seeking -- 2026-04-16
  629. [Editorial] Arxiv Research Paper -- 2026-04-16
  630. [Editorial] Claude Code Quiet but Important Update -- 2026-04-16
  631. [Editorial] duh — Developer Utility Harness -- 2026-04-16
  632. [Editorial] Yet Another Harness — Here's Why -- 2026-04-16
  633. [Editorial] Vibe Coding: Build a Website with AI -- 2026-04-16
  634. [Editorial] ASI-01 Agent Goal Hijack: A Practical Security Guide -- 2026-04-15
  635. [Editorial] AGHAST: Open Source Security Tool Release -- 2026-04-15
  636. Ransomware Is Growing Three Times Faster Than the Spending Meant to Stop It -- 2026-04-15
  637. Offensive Security Professional Blocked by Claude — Cyber Use Case Form Ignored -- 2026-04-15
  638. [Editorial] Archon: AI Agent Framework -- 2026-04-15
  639. CoreCoder: Minimal AI Coding Agent in ~950 Lines of Python -- 2026-04-15
  640. HitCC: Complete Reverse-Engineering of Claude Code CLI v2.1.84 -- 2026-04-15
  641. [Editorial] Pipecat Announcements -- 2026-04-15
  642. [Editorial] LiteParse Samples by Jerry Liu -- 2026-04-15
  643. [Editorial] exploraX Update -- 2026-04-15
  644. Zed Industries zeta-2: Editor-Native AI Model -- 2026-04-15
  645. Supply-Chain Poisoning Attacks Against LLM Coding Agent Skill Ecosystems -- 2026-04-14
  646. elastic/supply-chain-monitor -- 2026-04-14
  647. [Editorial] UK AI Security Institute + Claude -- 2026-04-14
  648. Pro Max 5x quota exhausted in 1.5 hours despite moderate usage -- 2026-04-14
  649. OpenAI releases new $100 Pro tier, rebalances Codex usage -- 2026-04-14
  650. Claude used to push back, now it just agrees with everything -- 2026-04-14
  651. Claude Thinking Blocks Are Being Summarized By A Second Agent -- 2026-04-14
  652. SafeRL-Lab/nano-claude-code — Python reimplementation supporting any model -- 2026-04-14
  653. LiteCode — free, open-source CLI coding agent for 8k-context LLMs -- 2026-04-14
  654. GitHub Stacked PRs -- 2026-04-14
  655. [Editorial] -- 2026-04-14
  656. I still prefer MCP over skills -- 2026-04-13
  657. [Editorial] -- 2026-04-13
  658. [Editorial] -- 2026-04-13
  659. [Editorial] -- 2026-04-13
  660. [Editorial] -- 2026-04-13
  661. [Editorial] -- 2026-04-13
  662. [Editorial] -- 2026-04-13
  663. Research-Driven Agents: When an agent reads before it codes -- 2026-04-13
  664. aaronjmars/MiroShark -- 2026-04-13
  665. [Editorial] -- 2026-04-13
  666. [Editorial] Four Layers of Sandboxing LLM Agents -- 2026-04-10
  667. Freestyle – Sandboxes for Coding Agents -- 2026-04-10
  668. [Editorial] Arxiv Research Paper -- 2026-04-10
  669. [Editorial] Provable Assurance for Agentic Systems -- 2026-04-10
  670. [Editorial] Anthropic's New Managed Agents -- 2026-04-10
  671. [Model Release] 9B Agentic Data Analyst LoRA — 89% Autonomous Workflow Completion -- 2026-04-10
  672. Continuous Batching for Agent Swarms — 42 Minutes to 70 Seconds -- 2026-04-10
  673. [Editorial] Karpathy Gist -- 2026-04-10
  674. [Editorial] Arxiv Research Paper -- 2026-04-10
  675. Box Maze: Process-Control Architecture for Reliable LLM Reasoning -- 2026-04-10
  676. Major Cache Reuse Bug Traced to Qwen 3.5's Chat Template -- 2026-04-10
  677. Anthropic's Claude Managed Agents Public Beta — Production Agent Infrastructure -- 2026-04-09
  678. botctl — Process Manager for Autonomous AI Agents -- 2026-04-09
  679. Feynman — AI Learning Companion -- 2026-04-09
  680. Claude Code Video Toolkit -- 2026-04-09
  681. [Editorial] -- 2026-04-08
  682. Show HN: Hippo, biologically inspired memory for AI agents -- 2026-04-08
  683. [Editorial] -- 2026-04-08
  684. honeybadge-labs/virtui -- 2026-04-08
  685. OpenWebUI integration, code intelligence for 248 languages, and more in Kreuzberg v4.7.0 -- 2026-04-08
  686. CLI-Anything Just Brought Claude Code Into The Future -- 2026-04-08
  687. Claude Code + Codex = AI GOD -- 2026-04-08
  688. Any Custom Frontend with Gradio's Backend -- 2026-04-08
  689. GLM-5.1: Towards Long-Horizon Tasks -- 2026-04-08
  690. MegaTrain: Full Precision Training of 100B+ Parameter LLMs on a Single GPU -- 2026-04-08
  691. kevinrgu/autoagent — autonomous harness engineering -- 2026-04-07
  692. remorses/usecomputer — Fast computer automation CLI for AI agents -- 2026-04-07
  693. Claude Code + LightRAG = UNSTOPPABLE -- 2026-04-07
  694. [Editorial] Career-Ops -- 2026-04-07
  695. Claude Code is unusable for complex engineering tasks -- 2026-04-07
  696. Eight years of wanting, three months of building with AI -- 2026-04-07
  697. 'Addictive' agentic coding has developers losing sleep -- 2026-04-07
  698. [Editorial] Video Feature -- 2026-04-06
  699. [Editorial] Everything Claude Code -- 2026-04-06
  700. [Editorial] Steve Yegge: Gas Town — From Clown Show to V1.0 -- 2026-04-06
  701. Karpathy's Obsidian RAG + Claude Code = CHEAT CODE -- 2026-04-06
  702. [Editorial] Elastic Open-Sources Their AI Tool -- 2026-04-06
  703. [Editorial] CVE-2026-22738 Proof of Concept -- 2026-04-06
  704. [Editorial] Linux Kernel — The Clearest Example -- 2026-04-06
  705. [Editorial] FindEvil — Security Tooling Hackathon -- 2026-04-06
  706. Coding agents could make free software matter again -- 2026-04-02
  707. GitHub backs down, kills Copilot pull-request ads after backlash -- 2026-04-02
  708. How do you know your AI audit tool actually checked everything? I was fairly confident that my skill suite did. It didn't. -- 2026-04-02
  709. Slop is not necessarily the future -- 2026-04-02
  710. [Editorial] -- 2026-04-02
  711. Holo3: Breaking the Computer Use Frontier -- 2026-04-02
  712. openyak/openyak -- 2026-04-02
  713. [Editorial] -- 2026-04-02
  714. Stanford, Harvard and MIT spent two weeks watching AI agents run loose. The paper is unsettling. -- 2026-04-01
  715. [Editorial] Agent Responsibly — Vercel's Guide -- 2026-04-01
  716. [Editorial] Slightly Safer Vibecoding by Adopting Better Practices -- 2026-04-01
  717. [Editorial] When AI Writes the AI Strategy -- 2026-04-01
  718. [Editorial] AI Development Video Tutorial -- 2026-04-01
  719. [Editorial] Learn Claude Code — Engineering Guide -- 2026-04-01
  720. [Editorial] Ask RuvNet — AI Assistant Demo -- 2026-04-01
  721. Claude Wrote a Full FreeBSD Remote Kernel RCE with Root Shell (CVE-2026-4747) -- 2026-04-01
  722. [Editorial] Claude Mythos Cracked the Linux Kernel -- 2026-04-01
  723. [Editorial] AI-Powered Pentesting in Practice -- 2026-04-01
  724. [Editorial] Meta-Harness: Learning to Harness AI Agents -- 2026-03-31
  725. [Editorial] Video: AI Development Deep Dive -- 2026-03-31
  726. Planner Agent V3 with SubAgents for Open WebUI -- 2026-03-31
  727. Benchmarked 31 STT Models on Medical Audio: VibeVoice 9B Is the New Open-Source Leader -- 2026-03-31
  728. The Missing Piece of Voxtral TTS: Enabling Voice Cloning -- 2026-03-31
  729. [Editorial] NoFxAiOS/nofx -- 2026-03-31
  730. We Rewrote JSONata with AI in a Day, Saved $500k/year -- 2026-03-31
  731. ChatGPT won't let you type until Cloudflare reads your React state -- 2026-03-30
  732. ClawShield: Security proxy for AI agents -- 2026-03-30
  733. [Editorial] NanoClaw Milestones -- 2026-03-30
  734. [Editorial] Chasing the Perfect Agent -- 2026-03-30
  735. AI agent on a $7/month VPS with IRC as its transport layer -- 2026-03-30
  736. [Editorial] Oh My Claude Code -- 2026-03-30
  737. [Editorial] Eyes for Claude -- 2026-03-30
  738. [Editorial] NotebookLM Python -- 2026-03-30
  739. [Editorial] -- 2026-03-28
  740. [Editorial] -- 2026-03-28
  741. [Editorial] -- 2026-03-28
  742. [Editorial] -- 2026-03-28
  743. [Editorial] -- 2026-03-28
  744. [Editorial] -- 2026-03-28
  745. [Editorial] -- 2026-03-28
  746. [Editorial] -- 2026-03-28
  747. You can now enable Claude to use your computer to complete tasks ! -- 2026-03-27
  748. Schedule tasks on the web -- 2026-03-27
  749. Anthropic CEO predicts AI could handle end-to-end software development in 6–12 months -- 2026-03-27
  750. [Editorial] Cron Jobs, Not Agents -- 2026-03-26
  751. [Editorial] HuggingFace hf-mount -- 2026-03-26
  752. [Editorial] YouTube Editorial Pick -- 2026-03-26
  753. Superpowers for Open WebUI — brainstorm → spec → plan → execute workflow for local LLMs -- 2026-03-26
  754. drivelineresearch/autoresearch-claude-code -- 2026-03-26
  755. mgechev/skills-best-practices -- 2026-03-26
  756. [Editorial] YouTube Editorial Pick -- 2026-03-26
  757. [Editorial] YouTube Editorial Pick -- 2026-03-26
  758. Netryx: Open-Source Street-Level Geolocation -- 2026-03-26
  759. [Editorial] -- 2026-03-25
  760. [Editorial] -- 2026-03-25
  761. [Editorial] Second Brain with Pi -- 2026-03-25
  762. [Editorial] ByteDance DeerFlow 2 Agent Runtime -- 2026-03-24
  763. AgencyCLI: Lightweight CLI for Self-Managing AI Agent Teams -- 2026-03-24
  764. SmarterRouter 2.2.1 — Self-Hosted AI Model Router (MoE Proxy) -- 2026-03-24
  765. Anthropic Launches Claude Dispatch — Control Desktop AI Tasks from Your Phone -- 2026-03-24
  766. You're Hardly Using What Claude Code Has to Offer (ColeMedin) -- 2026-03-24
  767. [Editorial] Claude Code Deep Dive — ColeMedin -- 2026-03-24
  768. [Editorial] Swictation v0.7.30 Release -- 2026-03-24
  769. Walmart: ChatGPT Checkout Converted 3x Worse Than Website -- 2026-03-24
  770. Agents of Chaos -- 2026-03-23
  771. [Editorial] How Vulnerable Are AI Agents to Indirect Prompt Injection -- 2026-03-23
  772. [Editorial] Autonomy Scales Exposure Before It Scales Value -- 2026-03-23
  773. [Editorial] arxiv:2603.15371 -- 2026-03-23
  774. leo-lilinxiao/codex-autoresearch -- 2026-03-23
  775. [Editorial] The Book That Talked Back -- 2026-03-23
  776. Show HN: Sub-millisecond VM sandboxes using CoW memory forking -- 2026-03-20
  777. Mistral AI Releases Forge -- 2026-03-20
  778. I built an open-source AI that lets you talk to your database — ask questions in plain English and get graphical insights instantly -- 2026-03-20
  779. [Editorial] You Are Not Deploying Agents You... -- 2026-03-19
  780. Build an Agent That Thinks Like a Data Scientist: How We Hit #1 on DABStep with Reusable Tool Generation -- 2026-03-19
  781. HKUDS/CLI-Anything -- 2026-03-19
  782. nextlevelbuilder/goclaw -- 2026-03-19
  783. Create Browser Swarms with Claude Code + Playwright CLI -- 2026-03-19
  784. [Editorial] Video Submission -- 2026-03-19
  785. [Editorial] RTK AI Toolkit -- 2026-03-18
  786. epiral/agent-clip -- 2026-03-18
  787. [Editorial] Octobot -- 2026-03-18
  788. Stripe's Coding Agents Ship 1,300 PRs EVERY Week - Here's How They Do It -- 2026-03-18
  789. Leanstral: Open-source agent for trustworthy coding and formal proof engineering -- 2026-03-18
  790. [Editorial] Hamilton Carter on AI Insights -- 2026-03-18
  791. [Editorial] Pwning AWS AgentCore Code Interpreter -- 2026-03-18
  792. [Editorial] xBow Raises $120M to Scale -- 2026-03-18
  793. [Editorial] AI Cyber Magazine Winter 2026 -- 2026-03-18
  794. I was backend lead at Manus. After building agents for 2 years, I stopped using function calling entirely. -- 2026-03-17
  795. dennisonbertram/agentic-hosting -- 2026-03-17
  796. [Editorial] Paperclip.ing -- 2026-03-17
  797. Beyond Semantic Similarity: Introducing NVIDIA NeMo Retriever's Generalizable Agentic Retrieval Pipeline -- 2026-03-17
  798. Holotron-12B - High Throughput Computer Use Agent -- 2026-03-17
  799. LocoreMind/LocoOperator-4B -- 2026-03-17
  800. [Editorial] MiroFish Demo -- 2026-03-17
  801. [Editorial] MiroFish GitHub -- 2026-03-17
  802. sstklen/trump-code -- 2026-03-17
  803. [Editorial] Which Countries Use Claude AI the Most -- 2026-03-17
  804. [Editorial] How My Agentic Workflow Actually Works (March 2026) -- 2026-03-16
  805. Spine Swarm (YC S23) — AI Agents That Collaborate on a Visual Canvas -- 2026-03-16
  806. [Editorial] The Developer Productivity Trap -- 2026-03-16
  807. [Editorial] Clarity Was Always the Bottleneck -- 2026-03-16
  808. [Editorial] The Hammer Problem -- 2026-03-16
  809. Structured Distillation for Personalized Agent Memory: 11x Token Reduction -- 2026-03-16
  810. Self-Flow by Black Forest Labs -- 2026-03-16
  811. GATED_DELTA_NET for Vulkan Merged in llama.cpp -- 2026-03-16
  812. [Editorial] AWS Security Agent -- 2026-03-16
  813. [Editorial] Caido AI Hunting Platform -- 2026-03-16
  814. ADPulse — Active Directory Security Pulse Tool -- 2026-03-16
  815. [Editorial] Hackers Gonna Hack — Be Prepped -- 2026-03-16
  816. 1B Identity Records Exposed in ID Verification Data Leak -- 2026-03-16
  817. goclaw: Self-hosted AI agent gateway written in Go -- 2026-03-14
  818. axon: Graph-powered code intelligence engine for AI agents via MCP -- 2026-03-14
  819. [Editorial] AAuth Full Demo — Authentication for Agentic Systems -- 2026-03-14
  820. Trace your LLM API and MCP calls with zero code changes (eBPF, Linux) -- 2026-03-14
  821. The Death of MCPs & The Rise of CLIs -- 2026-03-14
  822. [Editorial] Decentralized Self-Improving AI System That Builds Itself -- 2026-03-14
  823. [Editorial] How AI Agents Complete Two Months of Architecture Work in One Sprint -- 2026-03-14
  824. [Editorial] Video Submission -- 2026-03-14
  825. [Editorial] Everyone Reading This Works in a Profession That Didn't Exist in 1998 -- 2026-03-14
  826. [Editorial] Video Submission -- 2026-03-14
  827. [Editorial] Video Submission -- 2026-03-14
  828. New Model: LeVo 2 (SongGeneration 2), an open-source music foundation model -- 2026-03-14
  829. [Editorial] BinaryDefense NightBeacon -- 2026-03-13
  830. [Editorial] Root Evidence -- 2026-03-13
  831. [Editorial] tl;dr sec #319 -- 2026-03-13
  832. [Editorial] NSA Ghidra 12.0.4 Release -- 2026-03-13
  833. OmniCoder-9B: 9B coding agent fine-tuned on 425K agentic trajectories -- 2026-03-13
  834. [Editorial] Context Maturity for AI Coding Teams -- 2026-03-13
  835. Rudel: Analyzed 1,573 Claude Code Sessions to See How AI Agents Work -- 2026-03-13
  836. [Editorial] OpenClaw -- 2026-03-13
  837. Prompt-caching: Auto-Injects Anthropic Cache Breakpoints (90% Token Savings) -- 2026-03-13
  838. [Editorial] Video Submission -- 2026-03-13
  839. [Editorial] OpenAI: Designing Agents to Resist Prompt Injection -- 2026-03-13
  840. [Editorial] Anthropic Research Paper -- 2026-03-13
  841. [Editorial] Guardian: Mounting Concern Over Rogue AI Agents -- 2026-03-13
  842. [Editorial] Security in the Age of Agents -- 2026-03-13
  843. [Editorial] YousifAstar Post -- 2026-03-13
  844. Sandboxing local agents: Zero-trust CrewAI running entirely on Local Qwen 2.5 7B via Ollama -- 2026-03-13
  845. [Editorial] -- 2026-03-12
  846. [Editorial] -- 2026-03-12
  847. Whistleblower: DOGE member took Social Security data to new job -- 2026-03-12
  848. [Editorial] McKinsey AI Chatbot Hacked -- 2026-03-11
  849. AI Agent Hacks McKinsey -- 2026-03-11
  850. [Editorial] Red Amon — Faster and Cheaper Recon -- 2026-03-11
  851. [Editorial] The Agentic Coding Security Report -- 2026-03-11
  852. [Editorial] Gas Town by Kilo -- 2026-03-11
  853. [Editorial] Archive Feature -- 2026-03-11
  854. [Editorial] Your AI Says Whatever You Want to Hear — Here's How to Measure It -- 2026-03-11
  855. Agents That Run While I Sleep -- 2026-03-11
  856. [Editorial] Claude Code Review Economics -- 2026-03-11
  857. [Editorial] How AI Assistants are Moving the Security Goalposts -- 2026-03-10
  858. [Editorial] Ai owasp -- 2026-03-10
  859. [Editorial] SANS AI security -- 2026-03-10
  860. 89luca89/clampdown -- 2026-03-10
  861. [Editorial] Sovereign Shield -- 2026-03-10
  862. [Editorial] Openfang -- 2026-03-10
  863. [Editorial] Claude Code Deep Dive: The SDK Strikes Back -- 2026-03-10
  864. Upload files to PYODIDE code interpreter! MANY Open Terminal improvements AND MASSIVE PERFORMANCE GAINS - 0.8.9 is here! -- 2026-03-10
  865. [Editorial] Turbo Flow -- 2026-03-10
  866. [Editorial] Vibium browser automation -- 2026-03-10
  867. [NEWS] White House Preparing Executive Order to Ban Anthropic AI From Federal Operations -- 2026-03-10
  868. [Editorial] -- 2026-03-09
  869. [Editorial] -- 2026-03-09
  870. [Editorial] -- 2026-03-09
  871. [Editorial] -- 2026-03-09
  872. [Editorial] -- 2026-03-09
  873. [Editorial] -- 2026-03-09
  874. [Editorial] -- 2026-03-09
  875. [Editorial] -- 2026-03-09
  876. [Editorial] You Don't Give Agents Credentials, You Grant Them Power -- 2026-03-07
  877. [Editorial] From Discovery to Drift: Securing -- 2026-03-07
  878. [Editorial] Agents Change the Proof Standard -- 2026-03-07
  879. [Editorial] Aegis -- 2026-03-07
  880. AlexsJones/sympozium -- 2026-03-07
  881. [Editorial] Open-Sourcing git-stint -- 2026-03-07
  882. You can now train LLMs in VS Code for free via Google Colab & unsloth! -- 2026-03-07
  883. [Editorial] Clinejection: When Your AI Tool Installs Another -- 2026-03-07
  884. [Editorial] Memories Are All We Are: What the Road to AGI Is Missing -- 2026-03-06
  885. [Editorial] Open Trajectory Gym for AI Agents -- 2026-03-06
  886. [Editorial] Research Paper (arXiv 2603.03251) -- 2026-03-06
  887. Wave-Field LLM: O(n log n) Language Model via Wave Equation Dynamics -- 2026-03-06
  888. [Editorial] No-Cloud Tool-Calling Agents on Consumer Hardware (LFM2-24B-A2B) -- 2026-03-06
  889. [Editorial] Claude Cowork: Collaborative AI Coding -- 2026-03-06
  890. [Editorial] Google Workspace CLI -- 2026-03-06
  891. [Editorial] Nitpicker: AI Code Review Tool -- 2026-03-06
  892. [Editorial] Ramping Up on AI Development -- 2026-03-06
  893. [Editorial] Video Pick -- 2026-03-05
  894. [Editorial] SOFAI Workshop -- 2026-03-05
  895. Show HN: Sub-500ms latency voice agent from scratch -- 2026-03-05
  896. Agentic Engineering Patterns -- 2026-03-05
  897. [Editorial] Claude Code and AI Developer Tools -- 2026-03-05
  898. Open WebUI v0.8.6: Terminal integration, performance overhaul, security fixes -- 2026-03-05
  899. [Editorial] World Intel MCP -- 2026-03-05
  900. [Editorial] Public APIs Collection -- 2026-03-05
  901. Reverse CAPTCHA: We tested whether invisible Unicode characters can hijack LLM agents: 8,308 outputs across 5 models -- 2026-03-04
  902. [Editorial] Provos: Iron Curtain for AI Agents -- 2026-03-04
  903. [Editorial] Niels Provos on InfoSec, AI Agents & LLM Security -- 2026-03-04
  904. Catching an AI Red Teamer in the Wild: Using Reverse Prompt Injection as a Honeypot Detection Mechanism -- 2026-03-04
  905. [Editorial] Steve Yegge: Welcome to the Wasteland — A Thousand Gas Towns -- 2026-03-04
  906. [Editorial] manaflow-ai/cmux -- 2026-03-04
  907. Claude Code with subagents inside subagents cooked for 3 days — Delivered 3D renderer that draws with terminal symbols -- 2026-03-04
  908. [Editorial] obra/superpowers -- 2026-03-04
  909. [Editorial] Daniel Miessler: Personal AI Infrastructure -- 2026-03-04
  910. [Editorial] Ferricula -- 2026-03-04
  911. bcurts/agentchattr -- 2026-03-03
  912. [Editorial] ComposioHQ Agent Orchestrator -- 2026-03-03
  913. Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies -- 2026-03-03
  914. [Editorial] System Design Meets AI Reverse Engineering -- 2026-03-02
  915. [Editorial] AI Agent Patterns & Implementation -- 2026-03-02
  916. [Editorial] Advanced Agent Orchestration Techniques -- 2026-03-02
  917. [Editorial] ArXiv Research — Novel AI Methods -- 2026-03-02
  918. [Editorial] The AI Agent Security Gap Nobody Is Talking About -- 2026-03-02
  919. [Editorial] Systematic Jailbreak Attack Surface Mapping -- 2026-03-02
  920. [Editorial] Spec-Driven Development with Claude Code -- 2026-03-02
  921. [Editorial] Claude Code on Your Phone -- 2026-03-02
  922. If AI writes code, should the session be part of the commit? -- 2026-03-02
  923. [Editorial] AI Development Deep Dive -- 2026-03-02
  924. [Editorial] Portable Orchestra — AI Music Generation -- 2026-03-02
  925. Pi – A minimal terminal coding harness -- 2026-02-28
  926. [Editorial] -- 2026-02-28
  927. [Editorial] -- 2026-02-28
  928. [Editorial] -- 2026-02-28
  929. [Editorial] Momentum building for ruvector, rvf, etc -- 2026-02-28
  930. [Editorial] -- 2026-02-28
  931. [Editorial] -- 2026-02-28
  932. [Editorial] -- 2026-02-28
  933. 40,000+ AI Agents Exposed to the Internet with Full System Access -- 2026-02-28
  934. [Editorial] -- 2026-02-28
  935. We ran 56K multi-agent simulations - 1 misaligned agent collapses cooperation in a group of 5 -- 2026-02-28
  936. [Editorial] -- 2026-02-28
  937. [Editorial] -- 2026-02-28
  938. [Editorial] Introducing Perplexity Computer -- 2026-02-27
  939. [Editorial] Perplexity Computer Complete Guide -- 2026-02-27
  940. [Editorial] Claude Cowork Might Be the Most Consequential -- 2026-02-27
  941. [Editorial] Reid Hoffman: We're All Becoming Gamers -- 2026-02-27
  942. What Claude Code chooses -- 2026-02-27
  943. [Editorial] Video: AI Development Insights -- 2026-02-27
  944. [Editorial] Claude Skills Collection -- 2026-02-27
  945. [Editorial] AI Remediation Developers Actually Want to Use -- 2026-02-27
  946. github.com -- 2026-02-27
  947. [Editorial] AI Industry Commentary -- 2026-02-27
  948. [Editorial] -- 2026-02-26
  949. [Editorial] -- 2026-02-26
  950. [Editorial] -- 2026-02-26
  951. [Editorial] -- 2026-02-26
  952. [Editorial] -- 2026-02-26
  953. In the long run, everything will be local -- 2026-02-26
  954. [Editorial] -- 2026-02-26
  955. [Editorial] -- 2026-02-26
  956. Void-Box: Capability-Bound Agent Runtime -- 2026-02-26
  957. Show HN: enveil – hide your .env secrets from prAIng eyes -- 2026-02-26
  958. [Editorial] -- 2026-02-26
  959. [Editorial] -- 2026-02-26
  960. [Editorial] -- 2026-02-26
  961. The Age Verification Trap: Verifying age undermines everyone's data protection -- 2026-02-26
  962. [Editorial] Clawker -- 2026-02-25
  963. Built a honeypot token library for AI agents — detects prompt injection the moment it succeeds -- 2026-02-25
  964. [Editorial] AppSec, CVE, and Open Source Security -- 2026-02-25
  965. I Verified My LinkedIn Identity. Here's What I Handed Over -- 2026-02-25
  966. [Editorial] How I Came to Understand the 100x Claim -- 2026-02-25
  967. [Editorial] Claude Code for Live Structured Data -- 2026-02-25
  968. [Release] LocalAgent v0.1.1: Local-first agent runtime (LM Studio / Ollama / llama.cpp + Playwright MCP + eval/replay) -- 2026-02-25
  969. [Editorial] ContextGraph -- 2026-02-25
  970. [Editorial] Starlog -- 2026-02-25
  971. [Editorial] -- 2026-02-24
  972. [Editorial] -- 2026-02-24
  973. [Editorial] -- 2026-02-24
  974. The Missing Semester of Your CS Education – Revised for 2026 -- 2026-02-24
  975. [Editorial] -- 2026-02-24
  976. [Editorial] -- 2026-02-24
  977. BakeLens/crust -- 2026-02-24
  978. hazcod/claudleak -- 2026-02-24
  979. klawsh/klaw.sh -- 2026-02-24
  980. [Editorial] -- 2026-02-24
  981. [Editorial] -- 2026-02-24
  982. [Editorial] Bugcrowd Guide to Prompt Injection -- 2026-02-23
  983. [Editorial] arXiv Research -- 2026-02-23
  984. [Editorial] Exploitation Validator -- 2026-02-23
  985. What Breaks Embodied AI Security: LLM Vulnerabilities, CPS Flaws, or Something Else? -- 2026-02-23
  986. [Editorial] The AI Automation Ceiling -- 2026-02-23
  987. [Editorial] Faramesh — Research Paper -- 2026-02-23
  988. [Editorial] Faramesh — Core Repository -- 2026-02-23
  989. [Editorial] Faramesh — Video Introduction -- 2026-02-23
  990. Charlotte: Open Source Browser MCP Server — 136x More Token-Efficient for Agents -- 2026-02-23
  991. Kilntainers: Give Every Agent an Ephemeral Linux Sandbox via MCP [Open Source] -- 2026-02-23
  992. [Editorial] Run-Agent -- 2026-02-23
  993. [Editorial] Manifold -- 2026-02-23
  994. [Editorial] arXiv Research -- 2026-02-23
  995. [Editorial] Introducing AgentDB v3 -- 2026-02-23
  996. [Editorial] Agentic Quality Engineering -- 2026-02-23
  997. Zero-day CSS: CVE-2026-2441 exists in the wild -- 2026-02-21
  998. Microsoft says bug causes Copilot to summarize confidential emails -- 2026-02-21
  999. [Editorial] WebMCP — MCP for the Web -- 2026-02-21
  1000. [Editorial] Video: AI & Security Perspectives -- 2026-02-21
  1001. [Editorial] Why Probabilistic Engineering Breaks Deterministic Systems -- 2026-02-21
  1002. IBM and UC Berkeley Diagnose Why Enterprise Agents Fail Using IT-Bench and MAST -- 2026-02-20
  1003. Forensic audit on local AI assistant: 40.8% of tasks were fabricated -- 2026-02-20
  1004. [Editorial] OpenAI Practical Guide to Building Agents -- 2026-02-20
  1005. [Editorial] Generalized Hill Climbing Runtime -- 2026-02-20
  1006. [Editorial] Build Quality Skill: How I Ship Software 10x Faster -- 2026-02-20
  1007. AI45Lab/TrinityGuard: A Unified Framework for Safeguarding Multi-Agent System Safety -- 2026-02-20
  1008. HackMyClaw — Adversarial Security Challenge for AI Agents -- 2026-02-20
  1009. [Editorial] Video Feature -- 2026-02-20
  1010. Study: Self-generated Agent Skills are useless -- 2026-02-19
  1011. [Editorial] Claude Code RAG with Local Vector Database -- 2026-02-19
  1012. Ibrahim-3d/conductor-orchestrator-superpowers -- 2026-02-19
  1013. agno-agi/dash -- 2026-02-19
  1014. ST-EVO: Towards Generative Spatio-Temporal Evolution of Multi-Agent Communication Topologies -- 2026-02-18
  1015. Google Deepmind has released their take on multi-agent orchestration they're calling Intelligent AI Delegation -- 2026-02-18
  1016. [Editorial] BeadHub — AI Creative Tool -- 2026-02-18
  1017. I built a local AI coding agent with an 8-layer security sandbox — then had ChatGPT try to break it for 240+ rounds -- 2026-02-18
  1018. [Editorial] How to Sandbox Claude Code with Nono -- 2026-02-18
  1019. tomascupr/sandstorm — One API call. Full Claude agent. Completely sandboxed. -- 2026-02-18
  1020. [Editorial] AI Agent Security Strategy -- 2026-02-18
  1021. [Editorial] WebMCP and Enhanced Page Protocol -- 2026-02-17
  1022. WebMPC, has anyone used it? -- 2026-02-17
  1023. [Editorial] Voice-Controlled UI Agent Design -- 2026-02-17
  1024. [Editorial] The Agentic Operating System -- 2026-02-17
  1025. Forked OpenClaw to run fully air-gapped (no cloud deps) -- 2026-02-17
  1026. Anthropic still won't give the Pentagon unrestricted access to its AI models -- 2026-02-17
  1027. OpenAI uses internal version of ChatGPT to identify staffers who leak information: report -- 2026-02-17
  1028. FormalTask: Open-source declarative orchestration for Claude Code agents -- 2026-02-17
  1029. [Editorial] Get Shit Done -- 2026-02-17
  1030. [Editorial] Karpathy Gist -- 2026-02-17
  1031. [Editorial] AI DevOps and Developer Productivity -- 2026-02-16
  1032. OpenClaw Skill for Cost-Optimized Model Routing Based on Task Complexity -- 2026-02-16
  1033. [Editorial] O16G Platform -- 2026-02-16
  1034. [Editorial] GrubCrawler — Web Crawling Tool -- 2026-02-16
  1035. [Editorial] Storybook — UI Component Development -- 2026-02-16
  1036. [Editorial] Video Content -- 2026-02-16
  1037. [Editorial] ACM Research Paper -- 2026-02-16
  1038. [Editorial] https://mrinal.com/articles/agent-identities -- 2026-02-13
  1039. [Editorial] https://labs.zenity.io/p/perplexity-comet-a-reversing-story -- 2026-02-13
  1040. Jasonzzt/ComfyUI-CacheDiT -- 2026-02-12
  1041. ysharma3501/LuxTTS -- 2026-02-12
  1042. [Editorial] https://www.linkedin.com/pulse/when-brain-os-meets-real-operating-systems-rafael-knuth-4hcsf -- 2026-02-11
  1043. [Editorial] https://docs.entire.io/core-concepts -- 2026-02-11
  1044. Why System Prompts are failing your local agent builds (and why you need a Logic Floor) -- 2026-02-11
  1045. I built an MCP server that syncs Cursor, Claude Desktop, and Windsurf with one brain [Open Source] -- 2026-02-11
  1046. [Editorial] https://forge-quality.dev/articles/orchestra-learns-to-tune-itself -- 2026-02-10
  1047. I built an embodied agent in Minetest using Llama 3.2 + Vector Memory. Tonight, she passed the "Turing Test" by refusing to work because she was "tired. -- 2026-02-10
  1048. PlanDrop - Chrome extension to drop prompts from browser to AI coding agents on remote servers -- 2026-02-10
  1049. [Editorial] https://github.com/ikennaokpala/forge -- 2026-02-09
  1050. [Editorial] https://github.com/ruvnet/claude-flow/issues/1098 -- 2026-02-09
  1051. [Editorial] https://factory.strongdm.ai/ -- 2026-02-09
  1052. [Editorial] https://www.linkedin.com/posts/reuvencohen_both-the-new-codex-parallel-agents-and-the-activity-7425697703445196800-xCjI -- 2026-02-09
  1053. [Editorial] https://www.linkedin.com/posts/reuvencohen_most-intelligent-systems-fail-because-they-activity-7425306022862344192-TPtE -- 2026-02-06
  1054. [Editorial] https://www.linkedin.com/pulse/continuous-behavioral-verification-ongoing-path-done-ikenna-okpala-k9kme -- 2026-02-06
  1055. Need Help: AI Model for Local PDF & Image Extraction on Win11 (32GB RAM + RTX 2090) -- 2026-02-06
  1056. kmizu/embodied-claude -- 2026-02-06
  1057. benjiyaya/HeartMuLa_ComfyUI -- 2026-02-06
  1058. adithya-s-k/manim_skill -- 2026-02-04
  1059. Running DOOM and Super Mario 64 Inside a PDF File -- 2026-02-04
  1060. Nemotron-Personas-Brazil: Co-Designed Data for Sovereign AI -- 2026-02-04
  1061. [Editorial] https://www.linkedin.com/posts/patrickdebois_github-jedi4everaddt-run-ai-coding-agents-activity-7424653736788099072-7Aov -- 2026-02-04
  1062. MCP + Ghidra for AI-powered binary analysis — 110 tools, cross-version function matching via normalized hashing -- 2026-02-04
  1063. Arguably, the best AI code review MCP server (with Serena integration) -- 2026-02-04
  1064. EPYC 8124P (Siena) Build for Agentic Coding -- 2026-02-04
  1065. The 80% Problem in Agentic Coding – Addy Osmani -- 2026-02-04
  1066. [Editorial] https://www.linkedin.com/posts/dragan-spiridonov_agenticqe-agenticsfoundation-qualityengineering-ugcPost-7424143676773277696-EikW -- 2026-02-03
  1067. rodydavis/agent-skills-generator -- 2026-02-03
  1068. An Event Badge Re-Imagined As A Cyberdeck -- 2026-02-03
  1069. Open-Vocabulary Functional 3D Human-Scene Interaction Generation -- 2026-02-03
  1070. [Editorial] https://unhypedai.substack.com/p/the-knowledge-we-never-had-to-explain -- 2026-02-02
  1071. [Editorial] https://www.linkedin.com/posts/reuvencohen_i-keep-coming-back-to-this-realization-and-activity-7415150024868892672-E4rE -- 2026-02-02
  1072. [Editorial] https://humanemulator.co/ -- 2026-01-30
  1073. Generating skills for api+local CUAs via noVNC demonstration recording MCP -- 2026-01-30
  1074. Our Agent Rebuilt Itself in 26 Hours. AMA👀 -- 2026-01-30
  1075. I built a multi-agent orchestration layer for Claude Code - sharing in case it's useful to anyone -- 2026-01-30
  1076. I got tired of my AI agents overwriting each other's code, so I built a conflict manager for them -- 2026-01-27
  1077. Skill.md: An open standard for agent skills -- 2026-01-27
  1078. Unlocking Agentic RL Training for GPT-OSS: A Practical Retrospective -- 2026-01-27
  1079. [Editorial] https://github.com/ruvnet/ruvector/blob/claude/clawdbot-ruvector-setup-RHW3a/npm/packages/ruvbot/docs/FEATURE_COMPARISON.md -- 2026-01-27
  1080. Can companies "hack" ChatGPT to promote them? -- 2026-01-27
  1081. [Editorial] Vercel Labs' agent-browser + claude flow -- 2026-01-26
  1082. I wrote a URI scheme for agent identity that doesn't break when you move things -- 2026-01-26
  1083. [Open Sourse] I built a tool that forces 5 AIs to debate and cross-check facts before answering you -- 2026-01-26
  1084. An underrated way to turn AI code into real AI agents -- 2026-01-26
  1085. [Editorial] https://github.com/Combat-Drones-Detection-AI/Icarus -- 2026-01-26
  1086. [Editorial] https://unhypedai.substack.com/p/the-ai-operating-model-moment -- 2026-01-26
  1087. devstral small 2 vs glm 4.7 flash for agentic coding -- 2026-01-23
  1088. HeartMuLa/HeartMuLa-oss-3B -- 2026-01-23
  1089. [Editorial] agentic qe -- 2026-01-22
  1090. Demo: On-device browser agent (Qwen) running locally in Chrome -- 2026-01-22
  1091. Am I the only one in to enjoy the latest remote code sessions on Claude.ai with my full agentic config? Anyone else had some breakthrough with it? -- 2026-01-22
  1092. [Editorial] https://www.linkedin.com/posts/reuvencohen_llms-are-a-dead-end-not-because-they-are-activity-7419916372274470912-_5Lc -- 2026-01-22
  1093. All major AI stupid again, alternatives? -- 2026-01-22
  1094. [Resource] AI Guardrails: Open-source middleware to add PII Redaction & Injection Defense to local LLMs -- 2026-01-21
  1095. Jailbreak Challenge: Can You Break My Agent??? -- 2026-01-21
  1096. Do AI agents need TLS-style identities and ‘certificates’? -- 2026-01-21
  1097. Demo: On-device browser agent (Qwen) running locally in Chrome -- 2026-01-20
  1098. Agent observability is way different from regular app monitoring - maintainer's pov -- 2026-01-20
  1099. charIesding/agent-dashboard -- 2026-01-20
  1100. Learning Latency-Aware Orchestration for Parallel Multi-Agent Systems -- 2026-01-20
  1101. [Editorial] https://www.linkedin.com/posts/cole-medin-727752184_ive-been-testing-vercels-agent-browser-activity-7418832504754872320-PCA0 -- 2026-01-19
  1102. [Editorial] https://addyosmani.com/blog/good-spec -- 2026-01-19
  1103. AgentStudio: A VLA-based Kiosk Automation Agent using Gemini 3 and LangGraph -- 2026-01-19
  1104. Claude Skills Magic -- 2026-01-19
  1105. 7x Longer Context Reinforcement Learning in Unsloth -- 2026-01-19
  1106. openbmb/AgentCPM-Explore -- 2026-01-19
  1107. black-forest-labs/FLUX.2-klein-4B -- 2026-01-19
  1108. [Editorial] https://www.linkedin.com/posts/reuvencohen_announcing-claude-flow-v3-a-full-rebuild-activity-7417928335160262656-NYqJ -- 2026-01-16
  1109. [Editorial] https://www.linkedin.com/posts/sandstream_i-just-shipped-ralph-inferno-10-to-npm-activity-7417606358654406657-zBPY -- 2026-01-16
  1110. [Editorial] https://www.linkedin.com/posts/rasmuswiding_parallel-ai-agents-the-complete-infrastructure-activity-7417646422436777984-D1Zw -- 2026-01-16
  1111. Ralph Loop inspired me to build this - AI decides what Claude Code does next orchestrating claude code until task is done -- 2026-01-16
  1112. [Editorial] https://www.linkedin.com/posts/calebsima_due-to-popular-demand-here-is-my-%F0%9D%97%96%F0%9D%97%BC%F0%9D%97%B1%F0%9D%97%B6-activity-7417371887598514176-J6eg -- 2026-01-15
  1113. [Editorial] https://www.linkedin.com/posts/cole-medin-727752184_ralph-wiggum-is-everywhere-in-ai-right-now-activity-7417369954963910656-PQ3c -- 2026-01-15
  1114. [Editorial] https://www.linkedin.com/posts/craigmcluckie_coding-agents-are-crippling-oss-communities-activity-7417250625391915009-pcbA -- 2026-01-15
  1115. Agent reliability testing is harder than we thought it would be -- 2026-01-15
  1116. The Ralph Loop Made Easy -- 2026-01-15
  1117. [Editorial] https://github.com/pnocera/skilld -- 2026-01-15
  1118. [Editorial] https://www.linkedin.com/posts/hiltch_today-we-are-launching-openwork-an-open-source-ugcPost-7417259004294488064-KvyW -- 2026-01-15
  1119. [Editorial] https://www.linkedin.com/posts/claudio-stamile_if-youre-building-agents-youve-probably-activity-7416401402438205440-t9V_ -- 2026-01-13
  1120. [Editorial] https://www.linkedin.com/posts/matthewrwadams_threatmodeling-agenticai-aiagents-ugcPost-7416389760795176960-Ytut -- 2026-01-13
  1121. The hidden memory problem in coding agents -- 2026-01-13
  1122. I gave Claude Code a single instruction file and let it autonomously solve Advent of Code 2025. It succeeded on 20/22 challenges without me writing a single line of code. -- 2026-01-13
  1123. CloudAI-X/claude-workflow -- 2026-01-13
  1124. [Editorial] https://www.linkedin.com/posts/reuvencohen_a-year-ago-deepseek-landed-and-everyone-argued-activity-7416833905653329921-Xt9R -- 2026-01-13
  1125. Qwen3 235 VL hallucinates Tool calls -- 2026-01-13
  1126. [Editorial] https://www.sciencedirect.com/science/article/abs/pii/S1084804511000774 -- 2026-01-13
  1127. AgentSense: LLMs Empower Generalizable and Explainable Web-Based Participatory Urban Sensing -- 2026-01-13
  1128. [Editorial] https://github.com/leochlon/pythea/tree/main/strawberry -- 2026-01-12
  1129. [Editorial] https://arxiv.org/abs/2509.11208 -- 2026-01-12
  1130. One cargo install gives your AI 142 tools to perceive and control your machine - rmcp-presence -- 2026-01-09
  1131. AI agents for searching and reasoning over internal documents -- 2026-01-09
  1132. I built Plano - a framework-friendly data plane with orchestration for agents -- 2026-01-09
  1133. I built a TUI to manage multiple Claude Code agents in devcontainers (works great on mobile too) -- 2026-01-09
  1134. System: Control your Mac from anywhere using natural language -- 2026-01-09
  1135. Connect any LLM to all your knowledge sources and chat with it -- 2026-01-08
  1136. Have claude code interact with another claude code session interactively to test a plugin im building -- 2026-01-08
  1137. Semantic geometry for visual grounding -- 2026-01-08
  1138. zai-org/AutoGLM-Phone-9B -- 2026-01-08
  1139. facebook/sam-audio-large -- 2026-01-08
  1140. [Editorial] https://www.linkedin.com/posts/andriyburkov_a-major-breakthrough-in-reinforcement-learning-activity-7414543177648472064-_omq -- 2026-01-08
  1141. AskUserQuestionTool: if I have another kid, I know what I am going to name them. -- 2026-01-07
  1142. [Editorial] https://www.linkedin.com/posts/reuvencohen_ralph-wiggum-as-people-are-talking-about-activity-7414663704081981440-54bK -- 2026-01-07
  1143. [Editorial] https://joshclemm.com/writing/ralph-wiggum-future-of-coding -- 2026-01-07
  1144. [Editorial] https://ghuntley.com/ralph -- 2026-01-07
  1145. [Editorial] https://github.com/coleam00/Linear-Coding-Agent-Harness -- 2026-01-05
  1146. MCP Chat Studio v2: Workspace mode, workflows, contracts, mocks, and more -- 2026-01-05
  1147. Way to build powerful agents using natural language and code -- 2026-01-05
  1148. GLM-4.7 running full agentic workflows in Claude Code for 15 min straight - no failures -- 2026-01-05
  1149. I (almost) built an open-source, self-hosted runtime for AI agents in TypeScript... -- 2026-01-02
  1150. How to get started with automated workflows? -- 2026-01-02
  1151. Safe, Untrusted, "Proof-Carrying" AI Agents: toward the agentic lakehouse -- 2026-01-02
  1152. I built HMLR, an open source (full MIT) memory layer for your agent -- 2025-12-31
  1153. I built a "Recursive Swarm" engine inside a VS Code fork. It forces the LLM to explore 10,000 logic branches (System 2) before committing to code—trading 20 minutes of compute for accuracy. -- 2025-12-31
  1154. BOAD: Discovering Hierarchical Software Engineering Agents via Bandit Optimization -- 2025-12-31
  1155. Built an MCP Server for Andrej Karpathy's LLM Council -- 2025-12-31
  1156. Bounded autonomy: how the "is it an agent?" question changed my QA bot design -- 2025-12-31
  1157. eliasjudin/oai-skills -- 2025-12-31
  1158. zai-org/GLM-ASR -- 2025-12-31
  1159. AI Video Generation Made Easier with Wan 2.6 -- 2025-12-31
  1160. HKUDS/MCPNext -- 2025-12-29
  1161. virtual pet / life simulation using Ollama and Unity 6 -- 2025-12-23
  1162. YatharthS/MiraTTS -- 2025-12-23
  1163. stepfun-ai/Step-Audio-R1 -- 2025-12-23
  1164. [Editorial] https://x.ai/news/grok-voice-agent-api -- 2025-12-19
  1165. [Editorial] https://www.linkedin.com/posts/reuvencohen_sitting-on-a-beach-in-playa-del-carmen-activity-7407460969188163584-HQup -- 2025-12-19
  1166. [Editorial] https://www.linkedin.com/posts/yotam-perkal_comparing-ai-agents-to-cybersecurity-professionals-activity-7407076565357887488-KI5M -- 2025-12-18
  1167. Building an event-driven alternative to LangGraph because single-threaded loops are killing me. Roast my architecture. -- 2025-12-18
  1168. Intent vectors for AI search + knowledge graphs for AI analytics -- 2025-12-17
  1169. Cracking a 25-Year-Old Password with Claude Code -- 2025-12-17
  1170. Weird Email Appliance Becomes AI Terminal -- 2025-12-17
  1171. Zero-Shot Vehicle Model Recognition via Text-Based Retrieval-Augmented Generation -- 2025-12-17
  1172. stepfun-ai/Step-Audio-EditX -- 2025-12-17
  1173. AIDC-AI/Ovis-Image-7B -- 2025-12-17
  1174. [Editorial] https://www.linkedin.com/posts/resilientcyber_levels-of-autonomy-for-ai-agents-activity-7406679623167803392-OFJK -- 2025-12-16
  1175. ManiAgent: An Agentic Framework for General Robotic Manipulation -- 2025-12-16
  1176. CUGA on Hugging Face: Democratizing Configurable AI Agents -- 2025-12-16
  1177. AI Agent from scratch: Django + Ollama + Pydantic AI - A Step-by-Step Guide -- 2025-12-12
  1178. [Editorial] https://github.com/humanlayer/humanlayer -- 2025-12-12
  1179. Large update: 12 new frontier models added to the Step Game social reasoning benchmark. -- 2025-12-11
  1180. DeepMath: A lightweight math reasoning Agent with SmolAgents -- 2025-12-11
  1181. Nanbeige4-3B: Lightweight with strong reasoning capabilities -- 2025-12-10
  1182. mistralai/Devstral-2-123B-Instruct-2512 -- 2025-12-10
  1183. Can codex create multiple outputs, I check which is best? -- 2025-12-10
  1184. stepfun-ai/GELab-Zero-4B-preview -- 2025-12-10
  1185. Need opinion/help on my Memory System for LLM -- 2025-12-09
  1186. FlowCoder: Visual agentic workflow customization for Claude Code and Codex -- 2025-12-09
  1187. I built a CLI tool to manage AI configs across repos (aipaca) 🦙 -- 2025-12-09
  1188. Counterfactual-based Agent Influence Ranker for Agentic AI Workflows -- 2025-12-08
  1189. Run Any Model Provider on OpenWebUI immediately by discovering AI services on your LAN -- 2025-12-08
  1190. We gave 5 LLMs $100K to trade stocks for 8 months -- 2025-12-08
  1191. DevCrew agent swarm for accelerating your software development -- 2025-12-08
  1192. Connect and use Nova 2 Lite with Claude Code -- 2025-12-08
  1193. The security risks of "Emoji Smuggling" and Hidden Prompts for Local Agents -- 2025-12-08
  1194. We were tired of guessing which local model to use for which query. built a speculative execution lib that figures it out (github) -- 2025-12-05
  1195. Claude vs Codex: Claude won again 🏅 -- 2025-12-04
  1196. NornicDB - API compatible with neo4j - MIT - GPU accelerated vector embeddings -- 2025-12-04
  1197. gregorydickson/memory-graph -- 2025-12-04
  1198. Building Deep Research: How we Achieved State of the Art -- 2025-12-03
  1199. Claude launched 3 'explore agents' by itself -- 2025-12-02
  1200. OpenAI realtime API opensource alternative -- 2025-12-02
  1201. Built a Modular Agentic RAG System – Zero Boilerplate, Full Customization -- 2025-12-02
  1202. [Editorial] https://www.linkedin.com/posts/ownyourai_i-just-finished-testing-the-new-metas-omnilingual-activity-7400801588635836416-gpo- -- 2025-12-01
  1203. Xthebuilder/JRVS -- 2025-12-01
  1204. tigillo/githubmodels-go -- 2025-12-01
  1205. A Bird Watching Assistant -- 2025-12-01
  1206. InteractComp: Evaluating Search Agents With Ambiguous Queries -- 2025-11-28
  1207. Agent framework chaos? > Better Agents CLI -- 2025-11-28
  1208. Communication and Verification in LLM Agents towards Collaboration under Information Asymmetry -- 2025-11-26
  1209. Sibyl: an open source orchestration layer for LLM workflows -- 2025-11-25
  1210. Looking for 10 early testers building with agents, need brutally honest feedback👋 -- 2025-11-25
  1211. Claud Agent Dashboard -- 2025-11-25
  1212. [Editorial] https://github.com/punkpeye/awesome-mcp-servers -- 2025-11-24
  1213. Cornserve: Microservices Architecture for Serving Any-to-Any Models like Qwen Omni! -- 2025-11-24
  1214. How I’m Building Declarative, Shareable AI Agents With Docker cagent -- 2025-11-24
  1215. An open-source "Slack" for AI Agents to orchestrate n8n, Flowise, and OpenAI agents in one place -- 2025-11-24
  1216. modelscope/AgentEvolver -- 2025-11-24
  1217. [Editorial] https://www.linkedin.com/posts/stuart-winter-tear_ibm-the-2025-chief-data-officer-study-activity-7397614050433462272-0GmF -- 2025-11-21
  1218. Do you sandbox MCPs / Claude Code / Opencode on Linux? How ? -- 2025-11-21
  1219. Verifying hardware quality of rented gpus -- 2025-11-21
  1220. Ollama signin docker compose -- 2025-11-21
  1221. What's your Claude Code workflow setup? -- 2025-11-21
  1222. Measuring political bias in Claude -- 2025-11-21
  1223. Looking for feedback - I built Socratic, an open source knowledge base builder where YOU stay in control -- 2025-11-21
  1224. [Editorial] https://www.linkedin.com/posts/quanta-magazine_the-awful-consequence-of-an-observer-free-activity-7396969815078236160-Vo5G?u -- 2025-11-20
  1225. [Editorial] https://www.linkedin.com/posts/stuart-winter-tear_harmful-traits-of-ai-companions-activity-7397309575928131584-8H4J -- 2025-11-20
  1226. [Editorial] https://www.linkedin.com/posts/stuart-winter-tear_realist-and-pluralist-conceptions-of-intelligence-activity-7397231918871703554-FmSP?utm_source=social_share_send&utm_medium=member_desktop_web&rcm=ACoAAAAEV6YBBmyIQkYRxMIFJ7EWVq99NXg4qV4 -- 2025-11-20
  1227. How are you all orchestrating multi-agent workflows (beyond one-shot prompt chaining)? -- 2025-11-20
  1228. deliveryhero/asya -- 2025-11-20
  1229. [Editorial] https://www.linkedin.com/posts/stuart-winter-tear_aws-a-more-realistic-evaluation-activity-7396951453182967808-_H_c -- 2025-11-19
  1230. Should Spec-Driven-Development have a procedural orchestrator, or an LLM? -- 2025-11-19
  1231. Where are the gaps in Claude's "reasoning" capabilities? -- 2025-11-19
  1232. Smart Bandage Leverages AI Model For Healing Purposes -- 2025-11-19
  1233. miromind-ai/MiroThinker-v1.0-72B -- 2025-11-19
  1234. GPT-5-pro is likely a universal agentic gateway / Large Agentic Model -- 2025-11-19
  1235. BSD MAC LLM UI: Minimal, Auditable LLM Front End for Secure Environments -- 2025-11-18
  1236. easy-oidc/easy-oidc -- 2025-11-18
  1237. Disrupting the first reported AI-orchestrated cyber espionage campaign -- 2025-11-18
  1238. The Challenge of Large File Checksums -- 2025-11-18
  1239. Building A Smart Speaker Outside The Corporate Cloud -- 2025-11-18
  1240. [Editorial] https://www.linkedin.com/posts/stuart-winter-tear_i-saved-forty-ai-research-papers-recently-activity-7395547917983580160-BuOX -- 2025-11-17
  1241. [Editorial] https://www.linkedin.com/posts/reuvencohen_i-just-finished-rebuilding-dspyts-on-top-activity-7395872853092495360-OFb8 -- 2025-11-17
  1242. [Editorial] https://www.marktechpost.com/2025/11/08/how-to-build-an-agentic-voice-ai-assistant-that-understands-reasons-plans-and-responds-through-autonomous-multi-step-intelligence/ -- 2025-11-17
  1243. Local-First LLM That Safely Runs Real System Tasks — Looking for Engineering Feedback -- 2025-11-17
  1244. [MCP] Open-sourced a CSV-to-PostgreSQL loader server (vibe-coded with Claude) -- 2025-11-17
  1245. MCP Server for Industrial IoT - Built for PolyMCP Agent Orchestration -- 2025-11-17
  1246. Mimir - Parallel Agent task orchestration - Drag and drop UI (preview) -- 2025-11-17
  1247. Claude helped me make a multi agent ecosystem where models interact with each other autonomously -- 2025-11-17
  1248. AnythingLLM MCP Bridge & Prompt Injector -- 2025-11-14
  1249. Katakate/k7 -- 2025-11-14
  1250. Dicklesworthstone/mcp_agent_mail -- 2025-11-14
  1251. [Editorial] https://www.linkedin.com/posts/ivandj_as-ai-agents-multiply-across-tools-and-protocols-activity-7394057385872556032-SlAQ -- 2025-11-13
  1252. [Editorial] https://www.linkedin.com/posts/henrikgothberg_anthropic-building-effective-ai-agents-ugcPost-7394348623796350977-tcq1 -- 2025-11-13
  1253. [Editorial] https://www.linkedin.com/posts/reuvencohen_the-latest-mcp-spec-feels-like-the-moment-activity-7394373616471072768-okAg -- 2025-11-13
  1254. How to link an AI to a code execution environment? -- 2025-11-13
  1255. [Editorial] https://www.linkedin.com/posts/emollick_we-need-more-papers-like-this-one-which-examines-ugcPost-7392918095805222912-YjvU?utm_source=social_share_send&utm_medium=member_desktop_web&rcm=ACoAAAAEV6YBBmyIQkYRxMIFJ7EWVq99NXg4qV4 -- 2025-11-12
  1256. Agent failures in production pushed me to simulation-based testing -- 2025-11-12
  1257. Building agents that work like a band, not a factory line - anyone experimenting with emergent multi-agent coordination? -- 2025-11-12
  1258. Qwen3-VL works really good with Zoom-in Tool -- 2025-11-12
  1259. [Update] mlx-knife 2.0 stable — MLX model manager for Apple Silicon -- 2025-11-12
  1260. Vascura BAT - configuration Tool for Llama.Cpp Server via simple BAT files. -- 2025-11-12
  1261. Beelzebub MCP: Securing AI Agents with Honeypot Functions, Prompt Injection Detection -- 2025-11-11
  1262. Problem Uploading PDFs in Self hosted AI -- 2025-11-11
  1263. openai/gpt-oss-safeguard-20b -- 2025-11-11
  1264. Dexmal/dexbotic -- 2025-11-11
  1265. Blender 5.1 -- 2025-11-11
  1266. Qwen/Qwen3-VL-2B-Instruct -- 2025-11-11
  1267. [Editorial] https://www.linkedin.com/posts/stuart-winter-tear_my-company-is-forcing-me-to-become-ai-agent-activity-7393927479004135424-hI8p -- 2025-11-11
  1268. Hephaestus: AI workflows that discover and create their own tasks as they work -- 2025-11-11
  1269. Built my own IDE -- 2025-11-11
  1270. Roo Code 3.30.3 Release Updates | kimi‑k2‑thinking support | UI improvements | Bug fixes -- 2025-11-11
  1271. Claude-Bumper-Lanes - Vibe Code with Review Discipline -- 2025-11-11
  1272. We just released a multi-agent framework. Please break it. -- 2025-11-10
  1273. ⚡️ I scaled Coding-Agent RL to 32x H100s. Achieving 160% improvement on Stanford's TerminalBench. All open source! -- 2025-11-10
  1274. Agent Learning via Early Experience -- 2025-11-10
  1275. [Editorial] https://www.linkedin.com/posts/reuvencohen_claude-code-web-is-amazing-its-my-primary-activity-7393649498251644928-rAc8 -- 2025-11-10
  1276. CodeWiki: Research-Grade Repository Documentation at Scale [Open Source] -- 2025-11-10
  1277. Website builder powered by Claude AI - generating full websites in minutes -- 2025-11-10
  1278. “AI, Make Me A Degree Certificate” -- 2025-11-10
  1279. Self-hosted platform for running third-party AI agents with Ollama support (Apache-2.0) -- 2025-11-07
  1280. Open Source Alternative to NotebookLM/Perplexity -- 2025-11-07
  1281. Decade-qiu/Multi-Source-Media-MCP-Server -- 2025-11-07
  1282. v0.2.0 - GenFilesMCP -- 2025-11-07
  1283. ⚡️ Scaling Coding-Agent RL to 32x H100s. Achieving 160% improvement on Stanford's TerminalBench -- 2025-11-06
  1284. Bifrost: A High-Performance Gateway for LLM-Powered AI Agents (50x Faster than LiteLLM) -- 2025-11-06
  1285. Stop fighting with AI to build your project -- 2025-11-06
  1286. OpenSkills - a open sourced and completely private Claude Skills -- 2025-11-05
  1287. I used Llama + Droidrun to create a self-running Twitter bot -- 2025-11-05
  1288. Thread vs. Session based short-term memory -- 2025-11-05
  1289. kayba-ai/agentic-context-engine -- 2025-11-05
  1290. [Editorial] Collaboration gap -- 2025-11-05
  1291. Looking for advanced workflow tips: How are power-users integrating Claude (and other LLMs) into high-volume legal practice? -- 2025-11-05
  1292. Lessons from interviews on deploying AI Agents in production -- 2025-11-05
  1293. [Open Source] We deployed numerous agents in production and ended up building our own GenAI framework -- 2025-11-04
  1294. First LangFlow Flow Official Release - Elephant v1.0 -- 2025-11-04
  1295. zeusftk/FTK_CANVAS_AGENT_for_Comfyui -- 2025-11-04
  1296. Qwen3-VL-32B Q8 speeds in llama.cpp vs vLLM FP8 on a RTX PRO 6000 -- 2025-11-03
  1297. thu-coai/Glyph -- 2025-11-03
  1298. AndroidControl-Curated: Revealing the True Potential of GUI Agents through Benchmark Purification -- 2025-11-03
  1299. [Editorial] Agentic Flow -- 2025-11-03
  1300. I'm making an AI similar to a vtuber using ollama, here's what I have so far! (looking for advice on anything, really) -- 2025-11-03
  1301. Remember that simple online PDF bank converter tool making $40k/month? I did the exact same workflow with my general AI agent (only 1 prompt needed!) -- 2025-11-03
  1302. [Editorial] Agent limits -- 2025-11-02
  1303. [Editorial] Agent Identity -- 2025-11-02
  1304. [Editorial] AI Defense -- 2025-11-02
  1305. I built a privacy focused AI assistant for WearOS that supports locally hosted LLMs -- 2025-11-02
  1306. VellumForge2 - A high performance, very configurable and really easy to use DPO dataset generation tool, create high quality datasets for completely free -- 2025-11-01
  1307. PokeeAI/pokee_research_7b -- 2025-11-01
  1308. [Editorial] https://itrevolution.com/articles/from-line-cookto-head-chef-orchestrating-ai-teams/ -- 2025-11-01
  1309. [Editorial] Cursor 2.0 -- 2025-11-01
  1310. Open Source Lovable with Custom Agents, Full Stack Support, and Local Models -- 2025-11-01
  1311. A highly adaptable toolkit to build APIs and agents, with friendly interfaces for streaming and multimodality -- 2025-11-01
  1312. Is it possible to enable mcp server on for specific sub agent? -- 2025-11-01
  1313. [Editorial] AGI Defined. -- 2025-11-01
  1314. [Editorial] Know Your Agent (KYA) -- 2025-10-30
  1315. Spent the last few weeks falling down the Claude Agent SDK rabbit hole... built AgCluster.dev (open source) -- 2025-10-30
  1316. Found a faster way to build Claude Skills -- 2025-10-30
  1317. Agentic AI for Financial Crime Compliance -- 2025-10-30
  1318. GraphScout: Intelligent Routing for Local LLM Agent Workflows -- 2025-10-30
  1319. Show HN: Butter – A Behavior Cache for LLMs -- 2025-10-30
  1320. katanemo/Arch-Router-1.5B -- 2025-10-30
  1321. QAgent: A modular Search Agent with Interactive Query Understanding -- 2025-10-30
  1322. [Open Source] We deployed numerous agents in production and ended up building our own GenAI framework -- 2025-10-29
  1323. Claude Skills but running locally in Apple container -- 2025-10-29
  1324. OpenSkills CLI - Use Claude Code Skills with ANY coding agent -- 2025-10-29
  1325. Prompts avoiding Yes Men moments? -- 2025-10-29
  1326. severity1/claude-code-prompt-improver -- 2025-10-29
  1327. Built Coyote — An AI Agent That Feels Like Texting a Friend and released first model supporting native Async Tools -- 2025-10-29
  1328. Distil NPC: Family of SLMs responsing as NPCs -- 2025-10-29
  1329. nvidia/audio-flamingo-3-hf -- 2025-10-29
  1330. microsoft/UserLM-8b -- 2025-10-29
  1331. Test-Time Scaling Strategies for Generative Retrieval in Multimodal Conversational Recommendations -- 2025-10-29
  1332. [Editorial] New calculus of coding -- 2025-10-28
  1333. we had 2 weeks to build 5 microservices with 3 devs, tried running multiple AI agents in parallel -- 2025-10-28
  1334. Claude Code 2.0.27 -- 2025-10-28
  1335. steveyegge/vc -- 2025-10-28
  1336. [Editorial] MCP Scanner, security -- 2025-10-28
  1337. [Editorial] Data provenance -- 2025-10-28
  1338. Who is Introducing the Failure? Automatically Attributing Failures of Multi-Agent Systems via Spectrum Analysis -- 2025-10-28
  1339. [Editorial] Virtual false positive, physical problems -- 2025-10-28
  1340. Show HN: A fast, privacy-first image converter that runs in browser -- 2025-10-28
  1341. Microsoft Releases AI Call Center Stack with Voice, SMS, and Memory -- 2025-10-28
  1342. Robot Phone Home…Or Else -- 2025-10-28
  1343. vngrs-ai/Kumru-2B -- 2025-10-27
  1344. Training Gemma 3n for Transcription and Translation -- 2025-10-27
  1345. Agentic Exploration of Physics Models -- 2025-10-27
  1346. [Editorial] For the vibes -- 2025-10-27
  1347. Best way to implement a detailed plan in an MD file? -- 2025-10-27
  1348. sci-m-wang/ACE-open -- 2025-10-27
  1349. StepWiser: Stepwise Generative Judges for Wiser Reasoning -- 2025-10-27
  1350. DeepAnalyze: Agentic Large Language Models for Autonomous Data Science -- 2025-10-26
  1351. Claude for Computer Use using Sonnet 4.5 -- 2025-10-26
  1352. Any way to have sub-agent's keep context between invocations? -- 2025-10-26
  1353. Learning to Steer: Input-dependent Steering for Multimodal LLMs -- 2025-10-26
  1354. [Editorial] Promethean Fire -- 2025-10-26
  1355. Google AI falsely named an innocent journalist as a notorious child murderer -- 2025-10-26
  1356. Built my own MCP server for my app and was pleasantly shocked by how good it is -- 2025-10-25
  1357. facebook/cwm -- 2025-10-25
  1358. When Judgment Becomes Noise: How Design Failures in LLM Judge Benchmarks Silently Undermine Validity -- 2025-10-25
  1359. Building the Open Agent Ecosystem Together: Introducing OpenEnv -- 2025-10-25
  1360. [Editorial] Leading AI Agent Swarms: The Agentic QE 1.2.0 Journey -- 2025-10-24
  1361. lupantech/AgentFlow -- 2025-10-24
  1362. jaguarliuu/xunlong -- 2025-10-24
  1363. [Editorial] Browsers you can socially engineer -- 2025-10-24
  1364. [Editorial] share terminal sessions using Claude Code for web -- 2025-10-24
  1365. [Project] VT Code — Rust coding agent now with Ollama (gpt-oss) support for local + cloud models -- 2025-10-24
  1366. How path-based pattern matching helps AI code follow your team's coding best practice -- 2025-10-24
  1367. Show HN: FlowLens – MCP server for debugging with Claude Code -- 2025-10-24
  1368. We built ContextAgent — a context-centric take on multi-agent systems (rethinking what an “agent” is) -- 2025-10-23
  1369. Claude Haiku 4.5 for Computer Use -- 2025-10-23
  1370. Sonnet 4.5 subagent Haiku question -- 2025-10-23
  1371. disler/big-3-super-agent -- 2025-10-23
  1372. usieye/flowma -- 2025-10-23
  1373. [Editorial] https://github.com/jingyaogong/minimind/blob/master/README_en.md -- 2025-10-23
  1374. After treating RL training like an SRE project, I see why they chose CISPO -- 2025-10-23
  1375. Chatgpt or Claude for web coding assitant -- 2025-10-22
  1376. Does Claude Desktop support MCP Server Notifications? -- 2025-10-22
  1377. Ollama Cloud API Tool usage -- 2025-10-22
  1378. [Editorial] https://www.linkedin.com/posts/mavlevin_aisecurity-zeroday-cybersecurity-activity-7386478715813330944-P9OP -- 2025-10-22
  1379. Linux Capabilities Revisited -- 2025-10-22
  1380. I got fed up with Open WebUI/LibreChat for local LLMs so I made an open source tool to turn my GPU server into an always-on assistant -- 2025-10-21
  1381. This is how I track usage and improve my AI assistant without exposing sensitive data -- 2025-10-21
  1382. Roadmap for building scalable AI agents! -- 2025-10-21
  1383. My TypeScript MCP server template `mcp-ts-template` just hit v2.3.7. Declarative tool definitions. Pluggable Storage. Edge-native (Cloudflare Workers). Optional OpenTelemetry. OAuth with Scope Enforcement, etc. -- 2025-10-21
  1384. virattt/dexter -- 2025-10-21
  1385. UI-AGILE: Advancing GUI Agents with Effective Reinforcement Learning and Precise Inference-Time Grounding -- 2025-10-21
  1386. [Editorial] https://www.linkedin.com/posts/gadievron_another-day-another-attack-on-ai-coding-activity-7386382494117466112-tXuF -- 2025-10-21
  1387. [Editorial] https://www.linkedin.com/posts/reuvencohen_ive-seen-the-future-of-coding-and-it-activity-7386187612597714944-jXQn -- 2025-10-21
  1388. Qwen3-vl:235b-cloud Ollama model error -- 2025-10-21
  1389. Expose MCP at the LLM server level? -- 2025-10-20
  1390. I got tired of copy-pasting NotebookLM answers into Claude, so I built an MCP server for it -- 2025-10-20
  1391. Use n8n in Open WebUI without maintaining pipe functions -- 2025-10-20
  1392. Slack sync into OpenWebUI Knowledge -- 2025-10-20
  1393. [Editorial] Chart a path -- 2025-10-18
  1394. [Editorial] Agentic Flow -- 2025-10-18
  1395. [Editorial] Turbo Flow -- 2025-10-18
  1396. Claudiomiro: How to Achieve 100% Autonomous (Complex) Coding -- 2025-10-18
  1397. Flowchart vs handoff: two paradigms for building AI agents -- 2025-10-18
  1398. Compare Claude Code and Codex from one prompt -- 2025-10-18
  1399. Claude Agent SDK + Cloudflare Containers is the perfect agent platform -- 2025-10-18
  1400. [Editorial] Getting more out of Claude Code SDK -- 2025-10-17
  1401. [Editorial] Agentic Flow - AI Agent Framework That Gets Smarter AND Faster Every Time It Runs -- 2025-10-17
  1402. oracle/agent-spec -- 2025-10-17
  1403. Holy Marketplaces, Batman! -- 2025-10-16
  1404. Show HN: Metorial (YC F25) – Vercel for MCP -- 2025-10-16
  1405. Enhancing Generative Auto-bidding with Offline Reward Evaluation and Policy Search -- 2025-10-15
  1406. How to handle long running tools in realtime conversations. -- 2025-10-15
  1407. Anyone else having reasoning parser issue with Qwen-cli + GLM4.6 combo in vllm? -- 2025-10-15
  1408. Plan mode coming to Codex CLI -- 2025-10-15
  1409. Something is wrong with Sonnet 4.5 -- 2025-10-15
  1410. Xrvitd/MeshMosaic -- 2025-10-15
  1411. Qwen/Qwen3-VL-235B-A22B-Instruct -- 2025-10-15
  1412. Scene Graph-Guided Proactive Replanning for Failure-Resilient Embodied Agent -- 2025-10-15
  1413. Vibe Coding and the Popularization of CLI Interfaces: Why Don’t Big Companies Use Millions of Users as Contributors to Improve Models? -- 2025-10-14
  1414. rexleimo/agno-Go -- 2025-10-14
  1415. The Silent Scientist: When Software Research Fails to Reach Its Audience -- 2025-10-14
  1416. OpenAI’s AgentKit makes building AI agents way easier, design, chat, test, and connect everything in one place! -- 2025-10-14
  1417. NewtonBench: Benchmarking Generalizable Scientific Law Discovery in LLM Agents -- 2025-10-14
  1418. A list of models released or updated this week on this sub, in case you missed any (10 Oct). -- 2025-10-14
  1419. Alibaba-NLP/Tongyi-DeepResearch-30B-A3B -- 2025-10-14
  1420. ibm-granite/granite-4.0-micro -- 2025-10-14
  1421. BasedBase/GLM-4.5-Air-GLM-4.6-Distill -- 2025-10-14
  1422. A 5-minute, no-BS way to pick a local model for your real task -- 2025-10-14
  1423. [Update] CodeLens.AI - Crowdsourced AI Leaderboard 3 Days Later: Blind Voting and What We Learned -- 2025-10-14
  1424. How to re-create OpenAI Assistants locally? -- 2025-10-14
  1425. M2 Max 96GB - llama.cpp with codex and gpt-oss 120b to edit files and github upload -- 2025-10-14
  1426. Why You Should Build AI Agents with Ollama First -- 2025-10-14
  1427. OpenWebUI en Docker no detecta modelo LLaMA3 instalado con Ollama en Linux -- 2025-10-14
  1428. [Editorial] The Reality of Agentic Development -- 2025-10-13
  1429. [AutoBE] achieved 100% compilation success of backend generation with "qwen3-next-80b-a3b-instruct" -- 2025-10-13
  1430. What ACTUALLY works after testing every AI coding tool for 6 months -- 2025-10-13
  1431. Issue with long parameter values when using tool calling with Anthropic API -- 2025-10-13
  1432. Moondream3 and Salesforce GTA-1 for UI grounding in computer-use agents -- 2025-10-12
  1433. vdpiya/batchi -- 2025-10-12
  1434. Agentic generative AI for media content discovery at the national football league -- 2025-10-12
  1435. demo: my open-source local LLM platform for developers -- 2025-10-10
  1436. Modelfile. Do I need these tags PER prompt? -- 2025-10-10
  1437. Script to install a bunch of AI or Dev tools automatically.. what can I add to it or improve? -- 2025-10-10
  1438. Claude Code compaction fails with “Conversation too long” even when context is below 75% -- 2025-10-10
  1439. Show HN: FleetCode – Open-source UI for running multiple coding agents -- 2025-10-10
  1440. Local Terminal Access -- 2025-10-10
  1441. xcLee001/SonicVale -- 2025-10-10
  1442. InternRobotics/VLAC -- 2025-10-10
  1443. meituan-longcat/LongCat-Flash-Thinking -- 2025-10-10
  1444. LiquidAI/LFM2-1.2B-Tool -- 2025-10-10
  1445. Hcompany/Holo1.5-7B -- 2025-10-09
  1446. [Editorial] Agentics Newsletter -- 2025-10-09
  1447. [Editorial] Latest batch from rUv. -- 2025-10-09
  1448. TheAgentArk/Toucan -- 2025-10-09
  1449. [Editorial] Increased edit speed, reduced LLM cost -- 2025-10-08
  1450. AI agents face off -- 2025-10-08
  1451. How to make Claude Code work for you at night? -- 2025-10-08
  1452. tfriedel/claude-office-skills -- 2025-10-08
  1453. What happens if AI agents start trusting everything they read? (I ran a test.) -- 2025-10-06
  1454. High-performance mice can be used as a microphone to spy on users -- 2025-10-06
  1455. How can I test bad behavior in model APIs without getting banned? -- 2025-10-06
  1456. Framework or custom for local rag/agentic systems -- 2025-10-05
  1457. Test your MCP server against Llama, no key required -- 2025-10-05
  1458. aiprodcoder/MIXAPI -- 2025-10-05
  1459. williavs/AGENTDL -- 2025-10-05
  1460. Ally finally got RAG – everything runs local now -- 2025-10-05
  1461. RawdodReverend/TermNet -- 2025-10-05
  1462. [Editorial] https://www.linkedin.com/posts/albertochierici_lol-i-cant-stop-thinking-about-this-we-activity-7379840898626502656-bUYZ -- 2025-10-03
  1463. Vyzer9/Valkan -- 2025-10-03
  1464. Bypassing TLS Certificate Validation with Ld_preload -- 2025-10-03
  1465. I built Solveig, it turns any LLM into an agentic assistant in your terminal that can safely use your computer -- 2025-10-02
  1466. # 🥔 Meet Tater Totterson — The Local AI Assistant That Doesn’t Need MCP Servers -- 2025-10-02
  1467. Do I need to run /init on a repo if I already have AGENTS.md? -- 2025-10-02
  1468. sshllm/sshai -- 2025-10-02
  1469. [Editorial] System prompts are getting outdated! -- 2025-10-02
  1470. [Editorial] https://github.com/emcie-co/parlant -- 2025-10-02
  1471. Microsoft Agent Framework (Preview): Making AI Agents Simple for Every Developer -- 2025-10-02
  1472. Codex is mind blowing -- 2025-09-29
  1473. LongLLaDA: Unlocking Long Context Capabilities in Diffusion LLMs -- 2025-09-29
  1474. Apple called out every major AI company for fake reasoning and Anthropic's response proves their point -- 2025-09-29
  1475. Help with running Ai models with internet connectivity -- 2025-09-28
  1476. AWS announces EC2 instance attestation -- 2025-09-28
  1477. The Perplexity Search API -- 2025-09-28
  1478. Reinforcement Learning with Rubric Anchors -- 2025-09-28
  1479. PHM-Bench: A Domain-Specific Benchmarking Framework for Systematic Evaluation of Large Models in Prognostics and Health Management -- 2025-09-28
  1480. Roo Code 3.28.6 Release Notes - GPT-5-Codex IS HERE!! -- 2025-09-28
  1481. Main thing I use claude for is to prevent Codex from gaslighting me -- 2025-09-28
  1482. Model answers include raw <br> tags when generating tables – how to fix in Open WebUI? -- 2025-09-28
  1483. How to embed images in responses? -- 2025-09-28
  1484. New Agent benchmark from Meta Super Intelligence Lab and Hugging Face -- 2025-09-27
  1485. evalops/dspy-micro-agent -- 2025-09-27
  1486. nvidia/NVIDIA-Nemotron-Nano-9B-v2 -- 2025-09-27
  1487. inclusionAI/Ling-flash-2.0 -- 2025-09-27
  1488. 1K+ schemas of agentic projects visualized -- 2025-09-26
  1489. what AI agent framework is actually production viable and/or least problematic? -- 2025-09-26
  1490. Squeeze the Soaked Sponge: Efficient Off-policy Reinforcement Finetuning for Large Language Model -- 2025-09-26
  1491. CogniSQL-R1-Zero: Lightweight Reinforced Reasoning for Efficient SQL Generation -- 2025-09-26
  1492. Link a git repo to llama.cpp server? -- 2025-09-24
  1493. oxbshw/LLM-Agents-Ecosystem-Handbook -- 2025-09-24
  1494. Native MCP (streamable HTTP) may be on the way -- 2025-09-24
  1495. nvidia/NVIDIA-Nemotron-Nano-12B-v2 -- 2025-09-23
  1496. Gaia2 and ARE: Empowering the community to study agents -- 2025-09-23
  1497. Gaia2 and ARE: Empowering the community to study agents -- 2025-09-23
  1498. Open sourced my AI video generation project -- 2025-09-23
  1499. Zen, many Code CLI instances (/commands) for peaceful parallel task execution. -- 2025-09-23
  1500. twiggy-tools/Twiggy -- 2025-09-23
  1501. KubeAgentic-Community/KubeAgentic -- 2025-09-23
  1502. MyLocalAI - Enhanced Local AI Chat Interface (vibe coded first project!) -- 2025-09-23
  1503. Tesslate/WEBGEN-4B-Preview -- 2025-09-23
  1504. tencent/SRPO -- 2025-09-23
  1505. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning -- 2025-09-22
  1506. Atom-Searcher: Enhancing Agentic Deep Research via Fine-Grained Atomic Thought Reward -- 2025-09-22
  1507. [Editorial] A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks -- 2025-09-21
  1508. Claude Code native subagents vs. Claude Flow vs. BMAD -- 2025-09-21
  1509. Hallucination in LLM-Based Code Generation: An Automotive Case Study -- 2025-09-21
  1510. Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes -- 2025-09-20
  1511. GGUF security concerns -- 2025-09-20
  1512. Democratizing AI Safety with RiskRubric.ai -- 2025-09-20
  1513. VoxCPM 0.5B : Tokenizer-Free TTS and Voice Cloning -- 2025-09-18
  1514. Alibaba-NLP/Tongyi-DeepResearch-30B-A3B · Hugging Face -- 2025-09-18
  1515. NexaAI/OmniNeural-4B -- 2025-09-18
  1516. MobileLLM-R1-950M meets Apple Silicon -- 2025-09-18
  1517. VS Code Chat: Introducing auto model selection (preview) -- 2025-09-18
  1518. ircfspace/masque-plus -- 2025-09-18
  1519. First AI Agent for DevOps/SRE and Platform Engineering -- 2025-09-17
  1520. This AI assistant became our go-to Unity co-pilot (not just another LLM) -- 2025-09-17
  1521. Runtime intelligence in games -- 2025-09-17
  1522. [Editorial] Villager -- 2025-09-16
  1523. Update: we got our revenge and now beat Deepmind, Microsoft, Zhipu AI and Alibaba -- 2025-09-16
  1524. Building Ai Agent from Scratch (Python) -- 2025-09-15
  1525. Siddhant-K-code/tokenvm -- 2025-09-15
  1526. Taming Uncertainty via Automation: Observing, Analyzing, and Optimizing Agentic AI Systems -- 2025-09-15
  1527. Qwen3-Next-80B-A3B - a big step up may be the best open source reasoning model so far -- 2025-09-14
  1528. Qwen/Qwen3-Next-80B-A3B-Thinking -- 2025-09-14
  1529. Nothing concrete to show yet, I just wanted to celebrate getting a remote MCP server\connector with oAuth working :) -- 2025-09-09
  1530. The Dark Side of LLMs Agent-based Attacks for Complete Computer Takeover -- 2025-09-09
  1531. An LLM-powered Natural-to-Robotic Language Translation Framework with Correctness Guarantees -- 2025-09-09
  1532. [Editorial] Why Language Models Hallucinate -- 2025-09-09
  1533. [Editorial] Compression Failures in LLMs -- 2025-09-09
  1534. [Editorial] Active Inference AI -- 2025-09-09
  1535. I built a Graph RAG pipeline (VeritasGraph) that runs entirely locally with Ollama (Llama 3.1) and has full source attribution. -- 2025-09-09
  1536. Environments Hub walkthrough: Your Language Model needs better (open) environments to learn -- 2025-09-08
  1537. The Landscape of Agentic Reinforcement Learning for LLMs -- 2025-09-08
  1538. What are your struggles with tool-calling and local models? -- 2025-09-08
  1539. [Project Update] From Brittle Scripts to a Resilient, Self-Auditing Architecture: The Evolution of MeganX 3.0 -- 2025-09-07
  1540. I accidentally beat Claude Code this weekend - multi-agent-coder now #12 on Stanford's TerminalBench 😅 -- 2025-09-07
  1541. Open-source tool to let Claude Code control your computer -- 2025-09-07
  1542. Trustworthy Agents for Electronic Health Records through Confidence Estimation -- 2025-09-07
  1543. Context Reasoning Benchmarks: GPT-5, Claude, Gemini, Grok on Real Tasks -- 2025-09-05
  1544. The CLAUDE.md Framework: A Guide to Structured AI-Assisted Work (prompts included) -- 2025-09-05
  1545. Team-intN18-SoybeanSeclab/Typhon -- 2025-09-05
  1546. DatarusAI/Datarus-R1-14B-preview -- 2025-09-05
  1547. Are there any SDKs that offer native tool calling functionality that can be used with any LLMs -- 2025-09-04
  1548. Open source wrapper around AugmentCode -- 2025-09-04
  1549. Producer Pal: control Ableton Live and make music with Claude -- 2025-09-04
  1550. ChatGPT on the Road: Leveraging Large Language Model-Powered In-vehicle Conversational Agents for Safer and More Enjoyable Driving Experience -- 2025-09-04
  1551. Jupyter Agent Dataset -- 2025-09-04
  1552. Training & Querying 3 Ollama Models with Zer00logy: Symbolic Cognition Framework and Void-Math OS -- 2025-09-04
  1553. unsloth/Qwen3-Coder-480B-A35B-Instruct-GGUF -- 2025-09-03
  1554. HeroBench: A Benchmark for Long-Horizon Planning and Structured Reasoning in Virtual Worlds -- 2025-09-03
  1555. Achieving 80% task completion: Training LLMs to actually USE tools -- 2025-09-03
  1556. githubnext/gh-aw -- 2025-09-03
  1557. Toad: Universal TUI for Agentinc Coding from Will McGugan (Rich/Textual) -- 2025-09-03
  1558. How do you do RL 100% locally without a NVIDIA GPU? -- 2025-08-31
  1559. NiceWebRL: a Python library for human subject experiments with reinforcement learning environments -- 2025-08-31
  1560. Coquette Mobile - Android App, Ollama with Agentic Properties - desktop control. -- 2025-08-30
  1561. Testers for Seed-OSS tool calling wanted! -- 2025-08-29
  1562. Codebase to Knowledge Graph generator -- 2025-08-29
  1563. GaohaoZhou-ops/Tello-LLM-ROS -- 2025-08-29
  1564. Exploring Autonomous Agents: A Closer Look at Why They Fail When Completing Tasks -- 2025-08-29
  1565. Built an AI Agent Orchestration Platform - Handles 70% of Our Dev Tasks -- 2025-08-29
  1566. Hobbyist project : enabling smaller language models to interact with large code bases -- 2025-08-28
  1567. Evaluate any computer-use agent with HUD + OSWorld-Verified -- 2025-08-28
  1568. The outer loop vs. the inner loop of agents. A simple mental model to evolve the agent stack quickly and push to production faster. -- 2025-08-28
  1569. AgentCheck: Local AI-powered code review agents for Claude Code -- 2025-08-28
  1570. [Editorial] The Complete Guide to BuildingAI Agents -- 2025-08-27
  1571. Tencent/Youtu-agent -- 2025-08-27
  1572. AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications -- 2025-08-27
  1573. CNCF Webinar–AI Model Packaging with KitOps -- 2025-08-27
  1574. [Editorial] Sense of Self and Time in Borderline Personality -- 2025-08-27
  1575. [Editorial] AI and security tools. -- 2025-08-27
  1576. MetaAgent: Automatically Constructing Multi-Agent Systems Based on Finite State Machines -- 2025-08-26
  1577. Models to complement GPT-5? -- 2025-08-26
  1578. Why claude.md fails and How CORE Fixes Memory in Claude Code -- 2025-08-26
  1579. Free Preview of Qoder: The Future of Agentic Coding? -- 2025-08-25
  1580. What MCP Servers are You Using -- 2025-08-25
  1581. I built real-time course correction for Claude Code... and it's also a Tamagotchi -- 2025-08-25
  1582. Not a model, but Open Source Memory framework claims to beat Mem0 on public benchmarks -- 2025-08-24
  1583. ComputerRL: Scaling End-to-End Online Reinforcement Learning for Computer Use Agents -- 2025-08-24
  1584. CausalPlan: Empowering Efficient LLM Multi-Agent Collaboration Through Causality-Driven Planning -- 2025-08-24
  1585. Jules is already making excuses like a senior dev trying to explain why they pushed to main on a Friday. -- 2025-08-23
  1586. Build a Local AI Agent with MCP Tools Using GPT-OSS, LangChain & Streamlit -- 2025-08-23
  1587. Codanna Adds TypeScript Parsing and Modular Language Registry. Context-First Coding. -- 2025-08-23
  1588. Presenton now supports presentation generation via MCP -- 2025-08-23
  1589. In 44 lines of code, we have an actually useful agent that runs entirely locally, powered by Qwen3 30B A3B Instruct -- 2025-08-20
  1590. Web Agent Memory Protocol (WAMP): Building a Shared Memory Layer for the Web -- 2025-08-20
  1591. Learning from building my first saas using claude code -- 2025-08-20
  1592. Generate Images with Claude and Hugging Face -- 2025-08-20
  1593. [Editorial] AI agents are rendering GitHub's human-centric collaboration tools obsolete -- 2025-08-18
  1594. dongguanting/ARPO -- 2025-08-18
  1595. MCP for Research: How to Connect AI to Research Tools -- 2025-08-18
  1596. Tencent-Hunyuan/HunyuanWorld-1.0 -- 2025-08-16
  1597. Rediscovering Microsoft’s Oddball Music Generator From The 1990s -- 2025-08-16
  1598. Trying to decide between Kilocode, Cline and Roo code -- 2025-08-15
  1599. GPT-5 vs Claude Opus 4.1: Which New AI Model Wins? -- 2025-08-15
  1600. bosonai/higgs-audio-v2-generation-3B-base -- 2025-08-14
  1601. Chain-GPT/Solidity-LLM -- 2025-08-14
  1602. Bottom-up Domain-specific Superintelligence: A Reliable Knowledge Graph is What We Need -- 2025-08-14
  1603. 🇵🇭 FilBench - Can LLMs Understand and Generate Filipino? -- 2025-08-14
  1604. Miro ODR: Another Deep Research Agent model just went open source -- 2025-08-14
  1605. Is the Aider polyglot coding leaderboard still being updated? GPT-5? -- 2025-08-14
  1606. Claude going crazy on extended thinking? -- 2025-08-14
  1607. Building a self-hosted AI support agent (using GPT-OSS) that can both guide users and perform real actions – looking for feedback -- 2025-08-12
  1608. Local model recommendations for lightweight, repeated screenshot analysis on macOS? -- 2025-08-12
  1609. A free goldmine of tutorials for the components you need to create production-level agents Extensive open source resource with tutorials for creating robust AI agents -- 2025-08-12
  1610. declare-lab/jamify -- 2025-08-12
  1611. Intelligent-Internet/II-Search-4B -- 2025-08-12
  1612. THUDM/GLM-4.1V-9B-Thinking -- 2025-08-12
  1613. QwenLM/Qwen-Image -- 2025-08-12
  1614. A specific asynchronous workflow pattern -- 2025-08-11
  1615. mozilla-ai/any-llm -- 2025-08-11
  1616. SunzeY/SEAgent -- 2025-08-11
  1617. [Editorial] Three Things I Learned About Voice Agents from Kwindla Kramer -- 2025-08-09
  1618. NVIDIA AI-Q Achieves Top Score for Open, Portable AI Deep Research (LLM with Search Category) -- 2025-08-09
  1619. Vibe Coding an AI article generator using Onuro 🔥 -- 2025-08-09
  1620. Claude Code v1.0.71 - Background Commands -- 2025-08-09
  1621. Doriandarko/make-it-heavy -- 2025-08-09
  1622. universal-tool-calling-protocol/go-utcp -- 2025-08-09
  1623. [Editorial] Open source GUI for Claude Code -- 2025-08-08
  1624. DoubleAgents: Fine-tuning LLMs for Covert Malicious Tool Calls -- 2025-08-08
  1625. Hey folks, I’m one of the contributors to Bifrost, and we just launched it on Product Hunt -- 2025-08-08
  1626. Funny but annoying time bug -- 2025-08-08
  1627. A free goldmine of tutorials for the components you need to create production-level agents Extensive open source resource with tutorials for creating robust AI agents -- 2025-08-08
  1628. OS Agents: A Survey on MLLM-based Agents for General Computing Devices Use -- 2025-08-08
  1629. So multi agents.. and context.. how does that work -- 2025-08-07
  1630. Can you import chats in JSON? How? -- 2025-08-07
  1631. MAESTRO, a deep research assistant/RAG pipeline that runs on your local LLMs -- 2025-08-07
  1632. Quantize your own GGUFs the same way as your fav Unsloth Dynamic GGUFs -- 2025-08-07
  1633. Read your code -- 2025-08-07
  1634. weixin-omni/omni-bot-sdk-oss -- 2025-08-06
  1635. Kart – Distributed version-control for geospatial and tabular data -- 2025-08-06
  1636. [Editorial] Turn-Taking model for Voice AI Agents -- 2025-08-06
  1637. [Editorial] a more mature phase of the AI cycle. -- 2025-08-05
  1638. disler/claude-code-hooks-multi-agent-observability -- 2025-08-05
  1639. The Parallel Lives of an AI Engineer -- 2025-08-05
  1640. Any toolkits or predefined subagents for claude code that you think are a game changer? -- 2025-08-04
  1641. ramakay/claude-self-reflect -- 2025-08-04
  1642. [Editorial] Agentic Web: Weaving the Next Web with AI Agents -- 2025-08-03
  1643. [Editorial] Gemini Flow -- 2025-08-03
  1644. Pwn2Own Contestants hold on to Ollama exploits due to its rapid update cycle -- 2025-08-02
  1645. Claude Code sub agents not working as expected -- 2025-08-02
  1646. syou6162/cchook -- 2025-08-02
  1647. I need a tutorial for coding with any model (but currently trying with DeepSeek coder) -- 2025-08-02
  1648. The tradeoff between human and AI context -- 2025-08-02
  1649. Building a custom LLM trained on luciform prompts + ShadeOS daemon dialogues – seeking help -- 2025-08-01
  1650. I built a zsh plugin that turns natural language into shell commands using locally hosted Ollama -- 2025-08-01
  1651. Some thoughts on vibe / ai-driven coding -- 2025-08-01
  1652. [Editorial] AI in hostile environments... -- 2025-08-01
  1653. leesh3288/CVE-2025-32023 -- 2025-08-01
  1654. In search of riches, hackers plant 4G-enabled Raspberry Pi in bank network -- 2025-08-01
  1655. [Editorial] PRP, google cli fork -- 2025-07-31
  1656. [Editorial] Alternative to claude code cli -- 2025-07-31
  1657. Why I Forked Qwen Code -- 2025-07-31
  1658. Unwanted and unrelated changes to my code: my biggest gripe with ChatGPT -- 2025-07-31
  1659. How to Stop Claude from Being a Yes-Man? (Anchoring Bias Problem) -- 2025-07-31
  1660. We just open sourced NeuralAgent: The AI Agent That Lives On Your Desktop and Uses It Like You Do! -- 2025-07-30
  1661. Help with UnifyAI – Setting Up Local LLMs and UI Integration -- 2025-07-30
  1662. Show HN: Terminal-Bench-RL: Training Long-Horizon Terminal Agents with RL -- 2025-07-30
  1663. Show HN: Flyde 1.0 – Like n8n, but in your codebase -- 2025-07-30
  1664. Reachy The Robot Gets a Mini (Kit) Version -- 2025-07-30
  1665. Hey AI, Generate Me a Hardware Code! Agentic AI-based Hardware Design & Verification -- 2025-07-30
  1666. 100 lines of Python is all you need: A radically minimal coding agent that scores 65% on SWE-bench (near SotA!) [Princeton/Stanford NLP group] -- 2025-07-30
  1667. [Editorial] laude Code Videos and Demos by Ruv (claude-swarm fame) -- 2025-07-29
  1668. [Editorial] It was fun while it lasted... bring on the $1000/mo max plan. -- 2025-07-29
  1669. Claude Code Best Practices/Tips/Tricks -- 2025-07-29
  1670. Everything I've Learned so far About OpenAI's Agents -- 2025-07-29
  1671. Why isn't this already a standard in robotics? -- 2025-07-28
  1672. The 14 Pains of Billing for AI Agents -- 2025-07-28
  1673. [Editorial] Product Requirement Prompts (PRP) -- 2025-07-28
  1674. Red flag phrases -- 2025-07-28
  1675. Say hello to `hf`: a faster, friendlier Hugging Face CLI ✨ -- 2025-07-28
  1676. [Editorial] Local voice AI, 235B LLM -- 2025-07-28
  1677. I stopped typing. Now I just use a hotkey. I built Agent-CLI to make it possible. -- 2025-07-28
  1678. Local cross-platform speech-to-speech and real-time captioning with OpenAI Whisper, Vulkan GPU acceleration and more -- 2025-07-28
  1679. Devstral & Magistral as adapters of Mistral -- 2025-07-28
  1680. [Editorial] Intersection of Product Management and Development -- 2025-07-27
  1681. 🔓 I built Hearth-UI — A fully-featured desktop app for chatting with local LLMs (Ollama-ready, attachments, themes, markdown, and more) -- 2025-07-27
  1682. UIGEN-X 8B supports React Headless, Flutter, React Native, Static Site Generators, Tauri, Vue, Gradio/Python, Tailwind, and prompt-based design. GGUF/GPTQ/MLX Available -- 2025-07-27
  1683. Realtime codebase indexing for coding agents with ~ 50 lines of Python (open source) -- 2025-07-27
  1684. Freigeist - The new Vibe Coding Platform -- 2025-07-27
  1685. What are some unique uses of OpenWebUI that you can't get otherwise? -- 2025-07-27
  1686. Claude Code finally told me the truth about agents :) -- 2025-07-26
  1687. Airfare Discrimination as a Service: Airlines' Favorite New Pricing Trick -- 2025-07-25
  1688. would this make an ai dev's life easier? -- 2025-07-25
  1689. Let’s sync on CLI agents! What’s actually working for you? -- 2025-07-25
  1690. Security Issue - Recent Claude Code behavior favoring fast/easy/simple took an API key and hardcoded it as a default value -- 2025-07-25
  1691. What is the best agent framework for Qwen3? -- 2025-07-24
  1692. Qwen/Qwen3-Coder-480B-A35B-Instruct -- 2025-07-24
  1693. Tool calling or not, I will use anyway -- 2025-07-24
  1694. Do you give your LLM terminal and code execution access? -- 2025-07-24
  1695. Built Ollamaton - Universal MCP Client for Ollama (CLI/API/GUI) -- 2025-07-23
  1696. What models/ai-code editors don't train on my codebase? -- 2025-07-23
  1697. Can someone PLEASE ELI5 MCPs, Connectors, and Extensions for me? -- 2025-07-23
  1698. Made My Own Auto Tool System and Enhanced Web Search Tool + Questions -- 2025-07-23
  1699. omar-haris/cursor-buddy-mcp -- 2025-07-22
  1700. EU is being left behinde and it sucks! -- 2025-07-22
  1701. We built Explainable AI with pinpointed citations & reasoning — works across PDFs, Excel, CSV, Docs & more -- 2025-07-20
  1702. Ready to go multi agent workflow on github? -- 2025-07-20
  1703. How do we secure AI agents that act on their own? -- 2025-07-19
  1704. Migrating a semantically-anchored assistant from OpenAI to local environment (Domina): any successful examples of memory-aware agent migration? -- 2025-07-19
  1705. Trying to get my Ollama model to run faster, is my solution a good one? -- 2025-07-19
  1706. GitHub - boneylizard/Eloquent: A local front-end for open-weight LLMs with memory, RAG, TTS/STT, Elo ratings, and dynamic research tools. Built with React and FastAPI. -- 2025-07-18
  1707. A free goldmine of tutorials for the components you need to create production-level agents Extensive open source resource with tutorials for creating robust AI agents -- 2025-07-18
  1708. Five Big Improvements to Gradio MCP Servers -- 2025-07-18
  1709. Migrating a semantically-anchored assistant from OpenAI to local environment (Domina): any successful examples of memory-aware agent migration? -- 2025-07-18
  1710. ARGO - A Local-First, Offline AI Agent That Puts You in Control -- 2025-07-17
  1711. Why LangGraph overcomplicates AI agents (and my Go alternative) -- 2025-07-17
  1712. new MCP alt. just dropped -- 2025-07-17
  1713. pydantic/fasta2a -- 2025-07-17
  1714. Share your MCP servers and experiments! -- 2025-07-17
  1715. OPENCODE - Like Claude Code or Gemini CLI, but works with local models and/or paid ones as well -- 2025-07-15
  1716. I built a Deep Researcher agent and exposed it as an MCP server! -- 2025-07-15
  1717. awwaiid/gremllm -- 2025-07-15
  1718. Ollama calling tools -- 2025-07-15
  1719. 🪝 Claude-Flow@Alpha v2: We've implemented the new Claude Code Hooks in the latest Claude Flow alpha release combining hive style swarms, neural pattern recognition, and 87 MCP tools (install using: npx claude-flow@alpha) -- 2025-07-14
  1720. k2-fsa/ZipVoice -- 2025-07-13
  1721. K-intelligence/Midm-2.0-Base-Instruct -- 2025-07-13
  1722. AutoTester.dev: First AI-Driven Automatic Test Tool for Web Apps -- 2025-07-13
  1723. eiondb/eion -- 2025-07-13
  1724. mistralai/Devstral-Small-2507 -- 2025-07-13
  1725. What product or extension is great at autocomplete and predictive typescript/javascript and kotlin code. Cursor is out because I'm not going to pay even $1 on a greedy and scammy product, and Windsurf performs moderately well -- 2025-07-11
  1726. trufflesecurity/force-push-scanner -- 2025-07-11
  1727. LEGO/kube-tf-reconciler -- 2025-07-11
  1728. agentica-org/DeepSWE-Preview -- 2025-07-11
  1729. Thanks to you, I built an open-source website that can watch your screen and trigger actions. It runs 100% locally and was inspired by all of you! -- 2025-07-11
  1730. Preceptor – A Local AI Focus App That Nudges You Back on Track | Waitlist + Suggestions needed -- 2025-07-11
  1731. AGI is not multimodal -- 2025-07-09
  1732. How Do Vision-Language Models Process Conflicting Information Across Modalities? -- 2025-07-09
  1733. Building a Potato-based GLaDOS as an Introduction to AI -- 2025-07-07
  1734. We built runtime API discovery for LLM agents using a simple agents.json -- 2025-07-06
  1735. OWUI 0.6.15 OpenTelemetry (Experimental) -- 2025-07-06
  1736. [Open Source] Moondream MCP - Vision for AI Agents -- 2025-07-05
  1737. Kyutai's STT with semantic VAD now opensource -- 2025-07-05
  1738. brizzai/auto-mcp -- 2025-07-05
  1739. Lifailon/openrouter-bot -- 2025-07-05
  1740. Augment Code?? -- 2025-07-04
  1741. Simple-Efficient/RL-Factory -- 2025-07-04
  1742. Ratler/airuler -- 2025-07-04
  1743. [Setup discussion] AMD RX 7900 XTX workstation for local LLMs — Linux or Windows as host OS? -- 2025-07-04
  1744. 🧠💬 Introducing AI Dialogue Duo – A Two-AI Conversational Roleplay System (Open Source) -- 2025-07-04
  1745. Qwen 2.5 32B or Similar Models -- 2025-07-04
  1746. Extending Minds with Generative AI -- 2025-07-04
  1747. Trying to Make Llama Extract Smarter with a Schema-Building AI Agent -- 2025-07-02
  1748. Want help in retrieving links from DB -- 2025-07-02
  1749. Ingesting docs for context -- 2025-07-02
  1750. Agents via OpenWebUI Functions -- 2025-07-02
  1751. pfnet/plamo-2-translate -- 2025-06-30
  1752. Self-Adapting Language Models -- 2025-06-27
  1753. tencent/Hunyuan-A13B-Instruct -- 2025-06-27
  1754. maya-research/Veena -- 2025-06-27
  1755. jennyzzt/dgm -- 2025-06-26
  1756. Looking to build a local AI assistant - Where do I start? -- 2025-06-24
  1757. Real-time conversational AI running 100% locally in-browser on WebGPU -- 2025-06-24
  1758. UI + RAG solution for 5000 documents possible? -- 2025-06-24
  1759. Good stable voice cloning and TTS with NOT much complicated installation? -- 2025-06-24
  1760. 🚀 I built a lightweight web UI for Ollama – great for local LLMs! -- 2025-06-24
  1761. How to train a VLM with a dataset that has text and images? -- 2025-06-24
  1762. Top open-source AI Agent in both SWE-bench Verified and Lite -- 2025-06-24
  1763. AllTracker: Efficient Dense Point Tracking at High Resolution -- 2025-06-24
  1764. I Read All of Cloudflare's Claude-Generated Commits -- 2025-06-24
  1765. Show HN: I created an tool that creates interactive product demos in 2 minutes -- 2025-06-24
  1766. I’m the Maintainer (and Team) behind Open WebUI – AMA 2025 Q2 -- 2025-06-24
  1767. Eleven v3 -- 2025-06-22
  1768. SAGA Update: Now with Autonomous Knowledge Graph Healing & A More Robust Core! -- 2025-06-21
  1769. A free goldmine of tutorials for the components you need to create production-level agents -- 2025-06-21
  1770. Build a full on-device rag app using qwen3 embedding and qwen3 llm -- 2025-06-21
  1771. Build LLM from Scratch | Mega Playlist of 43 videos -- 2025-06-21
  1772. Running an LLM on a PS Vita -- 2025-06-21
  1773. haiku.rag a local sqlite RAG library -- 2025-06-21
  1774. LLMs Fine-Tuning -- 2025-06-21
  1775. Do you still use GPT APIs for demo apps? I'm leaning towards open models. -- 2025-06-21
  1776. Guidelines on how to be a scientific sleuth released -- 2025-06-21
  1777. Which models are you able to use with MCP servers? -- 2025-06-21
  1778. Rig upgraded to 8x3090 -- 2025-06-21
  1779. moonshotai/Kimi-Dev-72B -- 2025-06-21
  1780. Show HN: DaedalOS – Desktop Environment in the Browser -- 2025-06-19
  1781. tencent/SongGeneration -- 2025-06-19
  1782. haasonsaas/ocode -- 2025-06-19
  1783. dagger/container-use -- 2025-06-13
  1784. lerobot/smolvla_base -- 2025-06-10
  1785. brendanhogan/picoDeepResearch -- 2025-06-08
  1786. sarvamai/sarvam-m -- 2025-06-07
  1787. Qwen/Qwen3-Reranker-0.6B -- 2025-06-07
  1788. Hcompany/Holo1-7B -- 2025-06-06
  1789. huggingface/smolagents -- 2025-06-05
  1790. openpubkey/opkssh -- 2025-06-05
  1791. hashicorp/terraform -- 2025-06-05
  1792. NousResearch/atropos -- 2025-06-04
  1793. google/A2A -- 2025-06-03
  1794. hydropix/TranslateBookWithLLM -- 2025-05-31
  1795. sisig-ai/doctor -- 2025-05-31