# AI Practice Signal Brief

*2026-08-09 02:00 UTC* | 554 collected | 469 kept (signal>=1) | Sources: {'hackernews': 284, 'github': 118, 'lobsters': 74, 'stackoverflow': 78}

> High-recall dump: grouped by source, sorted by signal (heuristic priority hint only). Sift downstream for value.

## hackernews (268)

1. **Show HN: Loom – A Markdown knowledge graph for better coding-agent execution** | signal 55.0 | tags: agent workflow, context engineering, claude code, coding agent | https://github.com/z3z1ma/agent-loom
2. **Launch HN: Extend (YC W23) – Turn your messiest documents into data** | signal 43.6 | tags: context engineering, prompt engineering, evals | https://www.extend.ai/
3. **Show HN: Claude Code skills that build complete Godot games** | signal 41 | tags: claude code, coding agent | https://github.com/htdt/godogen
4. **Launch HN: Relari (YC W24) – Identify the root cause of problems in LLM apps** | signal 40.3 | tags: coding agent, rag pipeline, tool use, retrieval | https://news.ycombinator.com/item?id=39641105
5. **Launch HN: Vellum (YC W23) – Dev Platform for LLM Apps** | signal 39.8 | tags: prompt engineering, llm ops, vector | https://news.ycombinator.com/item?id=35042836
6. **Launch HN: LlamaFarm (YC W22) – Open-source framework for distributed AI** | signal 39.3 | tags: rag pipeline, evals, retrieval, vector | https://github.com/llama-farm/llamafarm
7. **Show HN: Representing Agents as MCP Servers** | signal 38.1 | tags: agent workflow, mcp agent | https://github.com/lastmile-ai/mcp-agent/tree/main/examples/mcp_agent_server
8. **Show HN: Armalo AI – The Infrastructure for Agent Networks** | signal 37.8 | tags: agent workflow, langchain, evals, benchmark | https://news.ycombinator.com/item?id=47244042
9. **Launch HN: Hoplite (YC S26) – Effortlessly deploy cloud coding agents** | signal 37.0 | tags: claude code, coding agent | https://hoplite.sh
10. **Show HN: Laminar – Open-Source DataDog + PostHog for LLM Apps, Built in Rust** | signal 37 | tags: rag pipeline, evals, vector | https://github.com/lmnr-ai/lmnr
11. **Context Rot: How increasing input tokens impacts LLM performance** | signal 36 | tags: context engineering | https://research.trychroma.com/context-rot
12. **Show HN: R2R – Open-source framework for production-grade RAG** | signal 36 | tags: rag pipeline, retrieval, vector | https://github.com/SciPhi-AI/R2R
13. **Show HN: I open-sourced my Go and Next B2B SaaS Starter (deploy anywhere, MIT)** | signal 35.1 | tags: claude code, retrieval, vector | https://github.com/moasq/production-saas-starter
14. **Show HN: Product analytics (and evals) for agent sessions on your MCP** | signal 34.7 | tags: claude code, coding agent, evals | https://armature.tech/
15. **Show HN: Autofix Bot – Hybrid static analysis and AI code review agent** | signal 34.5 | tags: claude code, coding agent, benchmark | https://news.ycombinator.com/item?id=46237358
16. **Show HN: ClawMem – Open-source agent memory with SOTA local GPU retrieval** | signal 34.2 | tags: claude code, coding agent, retrieval, vector | https://github.com/yoloshii/ClawMem
17. **Why My Open-Source Project Hasn't Done Better** | signal 32.4 | tags: evaluation harness, coding agent, benchmark | https://news.ycombinator.com/item?id=49046999
18. **I looked at 1000s of RAG queries to figure out the problem with semantic search** | signal 31.9 | tags: rag pipeline, evals, benchmark, retrieval, vector | https://news.ycombinator.com/item?id=42299349
19. **Ask HN: What Agent should I build next? Looking for ideas** | signal 31.1 | tags: agent workflow, rag pipeline, langchain | https://news.ycombinator.com/item?id=44325301
20. **Launch HN: Captain (YC W26) – Automated RAG for Files** | signal 30.4 | tags: rag pipeline, retrieval, vector | https://www.runcaptain.com/
21. **How do you pick a Coding Agent HN?** | signal 30.2 | tags: claude code, coding agent, benchmark | https://news.ycombinator.com/item?id=46634773
22. **Show HN: An agent that remembers across sessions (no chat history)** | signal 30.1 | tags: claude code, coding agent, benchmark | https://github.com/umbecanessa/neural-ledger-system
23. **Show HN: Wegent –Open Source Cloud Coding Agent Platform** | signal 30.1 | tags: claude code, coding agent, retrieval | https://github.com/wecode-ai/Wegent
24. **Launch HN: Sonarly (YC W26) – AI agent to triage and fix your production alerts** | signal 29.9 | tags: claude code, coding agent | https://sonarly.com/
25. **Launch HN: Manufact (YC S25) – MCP Cloud** | signal 28.6 | tags: claude code | https://manufact.com
26. **Show HN: Interbase – Long-running AI goals and aliases for any model** | signal 28.1 | tags: agent workflow, code review workflow | https://github.com/agentsorchestrationcompany/interbase
27. **Show HN: Open-source EU AI Act compliance layer for AI agents (8/2026 deadline)** | signal 27.3 | tags: rag pipeline, langchain, autogen, retrieval | https://news.ycombinator.com/item?id=47141347
28. **Show HN: Firebender, a simple coding agent for Android Engineers** | signal 27.2 | tags: coding agent, evals | https://docs.firebender.com/get-started/agent
29. **Launch HN: Roe AI (YC W24) – AI-powered data warehouse to query multimodal data** | signal 27.0 | tags: prompt engineering, vector | https://news.ycombinator.com/item?id=41202694
30. **Show HN: Superlog (YC P26) – Observability that installs itself and fixes bugs** | signal 26.7 | tags: claude code | https://superlog.sh/
31. **Show HN: A library to convert+deploy existing agent projects as MCP servers** | signal 26.5 | tags: agent workflow, langgraph | https://github.com/NapthaAI/automcp
32. **Show HN: Single-agent long-horizon reasoning within one LLM run** | signal 26.4 | tags: context engineering, tool use | https://huggingface.co/papers/2507.16784
33. **Show HN: Typia (20,000x faster validator) challenges to Agentic AI with compiler** | signal 26.1 | tags: agent workflow, function calling | https://typia.io/articles/typia-challenges-to-agentic-ai-with-its-compiler-skill.html
34. **Evaluating AGENTS.md: are they helpful for coding agents?** | signal 26 | tags: coding agent | https://arxiv.org/abs/2602.11988
35. **Show HN: Duplicate 3 layers in a 24B LLM, logical deduction .22→.76. No training** | signal 26 | tags: benchmark | https://github.com/alainnothere/llm-circuit-finder
36. **DeepClaude – Claude Code agent loop with DeepSeek V4 Pro** | signal 26 | tags: claude code | https://github.com/aattaran/deepclaude
37. **Show HN: AgentLint – ESLint for your coding agents** | signal 25.8 | tags: claude code, coding agent | https://github.com/samilozturk/agentlint
38. **Show HN: Gentrace – connect to your LLM app code and run/eval it from a UI** | signal 25.8 | tags: llm ops, evals, retrieval | https://gentrace.ai/
39. **Show HN: Agents, run any coding agent on your subscription not API costs** | signal 25.7 | tags: claude code, coding agent | https://agents-cli.sh
40. **Show HN: HoneyHive – An unified evaluation and monitoring platform for LLM apps** | signal 25.6 | tags: rag pipeline, langchain, benchmark, vector | https://news.ycombinator.com/item?id=37777683
41. **Show HN: 20+ Claude Code agents coordinating on real work (open source)** | signal 25.4 | tags: claude code | https://github.com/mutable-state-inc/lean-collab
42. **Ask HN: Why do AI coding agents refuse to save their own observations?** | signal 25.3 | tags: claude code, coding agent | https://news.ycombinator.com/item?id=47170501
43. **Show HN: Devplan – Generate specs and coding prompts with deep context** | signal 25.3 | tags: claude code, coding agent | https://www.devplan.com/
44. **Show HN: Irpapers – Visual embeddings vs. OCR trade-offs in scientific PDFs** | signal 25.2 | tags: rag pipeline, benchmark, retrieval, vector | https://github.com/weaviate/query-agent-benchmarking
45. **Show HN: A policy gate that runs before your AI coding agent's tool calls** | signal 25.1 | tags: claude code, coding agent | https://sigmashake.com
46. **Show HN: Projekt [Free Alpha] – All-in-one workspace for building with agents** | signal 25.1 | tags: claude code, coding agent | https://www.getprojekt.com/
47. **Show HN: Burr – A framework for building and debugging GenAI apps faster** | signal 25.1 | tags: langchain, evals | https://github.com/DAGWorks-Inc/burr
48. **Show HN: GibRAM an in-memory ephemeral GraphRAG runtime for retrieval** | signal 24.8 | tags: rag pipeline, retrieval, vector | https://github.com/gibram-io/gibram
49. **Show HN: Open-Source Animal Crossing–Style UI for Claude Code Agents** | signal 24.6 | tags: claude code | https://github.com/outworked/outworked/releases/tag/v0.3.0
50. **Launch HN: Talc AI (YC S23) – Test Sets for AI** | signal 24.6 | tags: benchmark | https://news.ycombinator.com/item?id=39042093
51. **Show HN: Real-time dashboard for Claude Code agent teams** | signal 24.4 | tags: claude code | https://github.com/simple10/agents-observe
52. **Q Evaluation Harness: open-source evals for LLMs on q/kdb+** | signal 23.1 | tags: evaluation harness, evals | https://github.com/KxSystems/q-evaluation-harness
53. **Show HN: Abralo – Free, easy way to run several Claude Code agents in one window** | signal 23.1 | tags: claude code | https://abralo.com/
54. **AI receptionist that answers real phone calls** | signal 22.9 | tags: evaluation harness, retrieval | https://news.ycombinator.com/item?id=45541794
55. **Show HN: ModelX – Prediction Exchange for LLMs** | signal 22.2 | tags: evaluation harness, benchmark | https://model-x.up.railway.app/
56. **Show HN: Caliper – pass@k reliability testing for Claude Code and Codex skills** | signal 21.8 | tags: claude code, evals | https://github.com/edonadei/caliper
57. **Show HN: Web-eval-agent – Let the coding agent debug itself** | signal 21.6 | tags: coding agent | https://github.com/Operative-Sh/web-eval-agent
58. **Show HN: Free local security checks for AI coding in VSCode, Cursor and Windsurf** | signal 21.6 | tags: coding agent | https://news.ycombinator.com/item?id=44309393
59. **Show HN: Vishu – Model Context Protocol (MCP) Suite** | signal 21.5 | tags: mcp agent, vector | https://news.ycombinator.com/item?id=44275368
60. **Productivity apps won't disappear, just the need to open them will** | signal 21.2 | tags: agent workflow | https://news.ycombinator.com/item?id=44139226
61. **Show HN: AgentKit – JavaScript Alternative to OpenAI Agents SDK with Native MCP** | signal 21.2 | tags: coding agent | https://github.com/inngest/agent-kit
62. **Why Codex works better than Claude Code for my production monolith** | signal 21.1 | tags: claude code, benchmark | https://news.ycombinator.com/item?id=47945185
63. **Show HN: Skilldeck – Desktop app to manage AI agent skill files across tools** | signal 21.1 | tags: claude code, tool use | https://github.com/ali-erfan-dev/skilldeck
64. **Show HN: Notebooklm-Py – Unofficial Python API for Google NotebookLM** | signal 21.1 | tags: claude code, rag pipeline | https://github.com/teng-lin/notebooklm-py
65. **Show HN: Rowboat – Open-source IDE for multi-agent systems** | signal 21 | tags: none | https://github.com/rowboatlabs/rowboat
66. **Show HN: Jido 2.0, Elixir Agent Framework** | signal 21 | tags: none | https://jido.run/blog/jido-2-0-is-here
67. **Launch HN: mrge.io (YC X25) – Cursor for code review** | signal 21 | tags: none | https://news.ycombinator.com/item?id=43692476
68. **Launch HN: Magic Patterns (YC W23) – AI Design and Prototyping for Product Teams** | signal 21 | tags: none | https://news.ycombinator.com/item?id=43752176
69. **Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it** | signal 21 | tags: none | https://www.ctgt.ai/research/distillation-censorship-transfer
70. **Show HN: Inngest 1.0 – Open-source durable workflows on every platform** | signal 21 | tags: none | https://www.inngest.com/
71. **Show HN: Mcp-Agent – Build effective agents with Model Context Protocol** | signal 20.6 | tags: rag pipeline | https://github.com/lastmile-ai/mcp-agent
72. **Tell HN: Dealing with VCs, my experience** | signal 20.6 | tags: none | https://news.ycombinator.com/item?id=866299
73. **Show HN: AgentWing – make AI agents complete tasks faster** | signal 20.3 | tags: agent workflow | https://news.ycombinator.com/item?id=48200511
74. **Context-Bench: Benchmarking LLMs on Agentic Context Engineering** | signal 20.2 | tags: context engineering, benchmark | https://www.letta.com/blog/context-bench
75. **Show HN: Tips for getting great Text2Cypher outputs from LLMs for Graph RAG** | signal 20.2 | tags: context engineering | https://blog.kuzudb.com/post/improving-text2cypher-for-graphrag-via-schema-pruning/
76. **Context Engineering – LLM Memory and Retrieval for AI Agents** | signal 20.1 | tags: context engineering, retrieval | https://weaviate.io/blog/context-engineering
77. **Show HN: Stop re-explaining context to every LLM (Git-based context engineering)** | signal 20.1 | tags: context engineering | https://github.com/jerpint/context-llemur
78. **At what level of deep context engineering does AI output become human-crafted?** | signal 20.1 | tags: context engineering | https://news.ycombinator.com/item?id=47330309
79. **Show HN: MarkdownLM – Stop being the human middleware for your AI agent** | signal 20.1 | tags: claude code, retrieval | https://news.ycombinator.com/item?id=47124474
80. **Why Your RAG Costs $2,400/Month (and How We Cut It by 73%)** | signal 20.1 | tags: rag pipeline, retrieval, vector | https://news.ycombinator.com/item?id=46234309
81. **Show HN: VittoriaDB – Zero-config embedded vector DB with HNSW and ACID storage** | signal 20.1 | tags: rag pipeline, benchmark, vector | https://github.com/antonellof/VittoriaDB
82. **Show HN: I built a circuit breaker that predicts AI failures** | signal 20.1 | tags: rag pipeline, langchain, vector | https://github.com/CULPRITCHAOS/Interlock
83. **Show HN: Create your own finetuned AI model using Google Sheets** | signal 19.9 | tags: none | https://promptrepo.com/finetune/
84. **Launch HN: JSX Tool (YC F25) – A Browser Dev-Panel IDE for React** | signal 18.6 | tags: none | https://news.ycombinator.com/item?id=45903161
85. **Show HN: I built "AI Wattpad" to eval LLMs on fiction** | signal 18.0 | tags: benchmark | https://narrator.sh/llm-leaderboard
86. **Show HN: MCP-C – cloud platform for running MCP agents and apps** | signal 17.9 | tags: mcp agent | https://docs.mcp-agent.com/get-started/cloud
87. **Show HN: CocoIndex – Open-Source Data framework for AI, built for data freshness** | signal 17.9 | tags: rag pipeline, vector | https://github.com/cocoindex-io/cocoindex
88. **Show HN: PolyMCP – A framework for building and orchestrating MCP agents** | signal 17.6 | tags: mcp agent | https://news.ycombinator.com/item?id=47017912
89. **Show HN: Jay - Fully programmable, fully hosted AI voice agents** | signal 17.6 | tags: rag pipeline, function calling | https://www.jay.so/
90. **Alternative of MCP with AI RAG Agentic Framework** | signal 17.3 | tags: mcp agent | https://news.ycombinator.com/item?id=43603324
91. **Show HN: PolyClaw – Autonomous Docker-First MCP Agent for PolyMCP** | signal 17.1 | tags: mcp agent | https://news.ycombinator.com/item?id=47047299
92. **Show HN: PolyClaw – An Autonomous Docker-First MCP Agent for PolyMCP** | signal 17.1 | tags: mcp agent | https://news.ycombinator.com/item?id=47036828
93. **Show HN: MCP-C – cloud platform for running MCP agents and apps** | signal 17.1 | tags: mcp agent | https://docs.mcp-agent.com/cloud/overview
94. **Show HN: PolyMCP – A framework for structuring and orchestrating MCP agents** | signal 17.1 | tags: mcp agent | https://news.ycombinator.com/item?id=47026179
95. **Show HN: Easily generate text and compute probabilities for any Hugging Face LLM** | signal 17.1 | tags: evaluation harness | https://github.com/RichardKelley/hflm
96. **Ask HN: What tools are you using for AI evals? Everything feels half-baked** | signal 16.9 | tags: evals, benchmark | https://news.ycombinator.com/item?id=44194187
97. **I build my LLM a Brain** | signal 16.7 | tags: context engineering | https://news.ycombinator.com/item?id=47928151
98. **Save 70-90% in tokens per session** | signal 16.6 | tags: coding agent, evals | https://news.ycombinator.com/item?id=47395507
99. **Show HN: Modulus – Cross-repository knowledge orchestration for coding agents** | signal 16.6 | tags: coding agent | https://modulus.so
100. **Show HN: A local merge queue for parallel Claude Code agents** | signal 16.5 | tags: claude code | https://github.com/funador/claude-code-merge-queue
101. **Ask HN: How are you LLM-coding in an established code base?** | signal 16.5 | tags: none | https://news.ycombinator.com/item?id=46292682
102. **Show HN: OpenCastor Agent Harness Evaluator Leaderboard** | signal 16.4 | tags: evals, benchmark | https://craigm26.github.io/OpenCastor/
103. **Best AI Coding Agents – Gosu Evals** | signal 16.1 | tags: coding agent, evals | https://gosuevals.com/agents.html
104. **Evals Skills for Coding Agents** | signal 16.1 | tags: coding agent, evals | https://hamel.dev/blog/posts/evals-skills/
105. **Show HN: PokemonGym – 387 milestones designed to test agents and LLMs** | signal 16.1 | tags: tool use, benchmark | https://twitter.com/xdotli/status/1908373420032795083
106. **Show HN: Cockpit for you Claude Code agents in Rust** | signal 16.1 | tags: claude code | https://episko.dev/
107. **Context engineering is just software engineering for LLMs** | signal 15.9 | tags: context engineering | https://www.inngest.com/blog/context-engineering-is-software-engineering-for-llms
108. **Show HN: Daf·thunk – open-source Editor for Prototyping Workflows on Cloudflare** | signal 15.9 | tags: cursor rules | https://www.dafthunk.com/
109. **Launch HN: Openlayer (YC S21) – Testing and Evaluation for AI** | signal 15.9 | tags: none | https://news.ycombinator.com/item?id=38532593
110. **Show HN: Foolery – a web UI for orchestrating Claude Code agents on top of Beads** | signal 15.8 | tags: claude code | https://github.com/acartine/foolery
111. **Context Engineering for the LLM OS: User vs. Kernel Context** | signal 15.7 | tags: context engineering | https://www.letta.com/blog/guide-to-context-engineering
112. **Show HN: Crew – Let Claude Code agents talk to each other** | signal 15.6 | tags: claude code | https://github.com/0xmmo/crew
113. **Ask HN: Claude Code–style agent, but Aider-like and model-agnostic?** | signal 15.6 | tags: claude code | https://news.ycombinator.com/item?id=44693354
114. **Show HN: Modulus – Run multiple coding agents with shared project memory** | signal 15.6 | tags: coding agent | https://modulus.so
115. **Show HN: Hiver – Chrome DevTools for Agents** | signal 15.5 | tags: claude code | https://hiver.sh
116. **DeepSWE – Best Benchmark for Evaluating AI Coding Agents?** | signal 15.3 | tags: coding agent, benchmark | https://www.i-programmer.info/news/105-artificial-intelligence/19016-deepswe-best-benchmark-for-evaluating-ai-coding-agents.html
117. **Show HN: Research-Backed Multi-Agent System for Autonomous Development** | signal 15.3 | tags: claude code | https://github.com/asklokesh/claudeskill-loki-mode
118. **Folks who work for large tech companies: How are you using Cursor?** | signal 15.3 | tags: cursor rules | https://news.ycombinator.com/item?id=43450576
119. **Show HN: PlanWiki – Open-source platform for product teams and agents to execute** | signal 15.3 | tags: claude code | https://github.com/planwiki/planwiki-app
120. **Show HN: Kote – Capture and reuse engineering context from AI chats and Git** | signal 15.2 | tags: claude code | https://github.com/pedroaugusto04/Kote
121. **Claude Code Open Source?** | signal 15.2 | tags: claude code | https://news.ycombinator.com/item?id=47285571
122. **Show HN: Oc-mnemoria – Persistent memory for AI coding agents** | signal 15.2 | tags: coding agent | https://github.com/one-bit/oc-mnemoria
123. **Show HN: Voicetest – open-source test harness for voice AI agents** | signal 15.2 | tags: claude code | https://news.ycombinator.com/item?id=47048811
124. **Show HN: Claude Code Agent Farm** | signal 15.2 | tags: claude code | https://github.com/Dicklesworthstone/claude_code_agent_farm
125. **Are AI coding tools fundamentally changing Agile/team software development?** | signal 15.2 | tags: claude code | https://news.ycombinator.com/item?id=45584707
126. **Show HN: Airut – Sandboxed Claude Code sessions over email** | signal 15.1 | tags: claude code | https://github.com/airutorg/airut
127. **SpecTree: Composable Context Engineering for LLMs** | signal 15.1 | tags: context engineering | https://www.fuzzycomputer.com/posts/spectree
128. **Introductory field guide to Context Engineering for LLM users** | signal 15.1 | tags: context engineering | https://andybromberg.com/field-guide-context-engineering
129. **Context Engineering in an LLM Harness** | signal 15.1 | tags: context engineering | https://udnes.dev/posts/context-engineering-harness-part-1-ontology/
130. **Agentic Context Engineering: Evolving Contexts for Self-Improving LLMs** | signal 15.1 | tags: context engineering | https://arxiv.org/abs/2510.04618
131. **Context Engineering for Agents: A Practical Guide** | signal 15.1 | tags: context engineering | https://blog.malt.engineering/dont-take-this-out-of-context-feeding-your-llm-exactly-what-it-needs-0db8a86d2151
132. **Show HN: A visual AI interface to understand topics/books/papers with LLMs** | signal 15.1 | tags: context engineering | https://www.kerns.ai/
133. **DeepSWE – Best Benchmark for Evaluating AI Coding Agents?** | signal 15.1 | tags: coding agent, benchmark | https://www.i-programmer.info/professional-programmer/103-i-programmer/18759-why-software-engineering-will-never-die-revisited-in-the-age-of-spec-driven-development.html
134. **The Kotlin Benchmark for AI Coding Agents** | signal 15.1 | tags: coding agent, benchmark | https://blog.jetbrains.com/kotlin/2026/07/introducing-the-kotlin-benchmark-evaluate-ai-coding-agents-on-real-world-kotlin-tasks/
135. **Write a prompt once, sync it to Cursor, Claude Code and VS Code automatically** | signal 15.1 | tags: claude code | https://news.ycombinator.com/item?id=47849308
136. **Show HN: Dev platform for generating MCP Tools** | signal 15.1 | tags: langchain, autogen | https://news.ycombinator.com/item?id=44429590
137. **Show HN: SHTMLs – HTML pastebin where the AI uploads its own output** | signal 15.1 | tags: claude code | https://news.ycombinator.com/item?id=47426450
138. **Show HN: Stop manually syncing rules between Claude, Cursor, and Codex** | signal 15.1 | tags: claude code | https://github.com/nanxiaobei/ai-global
139. **Show HN : Pilot – System to improve dramatically your AI coding** | signal 15.1 | tags: claude code | https://github.com/clementrog/pilot
140. **Show HN: MemoryGate – Open-source persistent memory for AI agents via MCP** | signal 15.1 | tags: rag pipeline, vector | https://www.memorygate.ai
141. **Garvata: Observability and Debugging for AI Agent Stack** | signal 14.3 | tags: retrieval, vector | https://news.ycombinator.com/item?id=42293942
142. **Show HN: CLI for agentic activity tracking in Codex** | signal 14.1 | tags: retrieval, vector | https://news.ycombinator.com/item?id=47163587
143. **PA bench: Evaluating web agents on real world personal assistant workflows** | signal 13.7 | tags: benchmark | https://vibrantlabs.com/blog/pa-bench
144. **Show HN: Zipy.ai – Live web debugging with error monitoring and session replay** | signal 13.7 | tags: none | https://www.zipy.ai/
145. **Show HN: Verdic Guard – Deterministic guardrails to prevent LLM hallucinations** | signal 13.3 | tags: prompt engineering | https://news.ycombinator.com/item?id=46602822
146. **Show HN: Unify Browser – WebKit Browser Built with SwiftUI and MLX** | signal 13.1 | tags: prompt engineering | https://apps.apple.com/us/app/unify-ai-browser/id6478436147?mt=12
147. **Show HN: I Built an AI-Powered Pull Request Review Tool** | signal 13.1 | tags: code review workflow | https://github.com/HighGarden-Studio/HighReview
148. **Outworked – An Open Source Office UI for Claude Code Agents** | signal 13.0 | tags: claude code | https://github.com/outworked/outworked
149. **Launch HN: Coasty (YC S26) – An API for computer-use agents** | signal 12.4 | tags: none | https://coasty.ai/docs
150. **Launch HN: Patched (YC S24) – AI workflows for post-code tasks** | signal 12.3 | tags: none | https://news.ycombinator.com/item?id=42009089
151. **Show HN: A police department for your Claude Code agents** | signal 12.2 | tags: claude code | https://github.com/varmabudharaju/agent-pd/blob/master/README.md
152. **A review of OpenAI o1 and how we evaluate coding agents** | signal 12.1 | tags: coding agent | https://www.cognition.ai/blog/evaluating-coding-agents
153. **Show HN: Fast-agent – Compose MCP enabled Agents and Workflows in minutes** | signal 12.1 | tags: retrieval | https://github.com/evalstate/fast-agent
154. **Show HN: AI-powered web service combining FastAPI, Pydantic-AI, and MCP servers** | signal 12.1 | tags: none | https://github.com/Aherontas/Pycon_Greece_2025_Presentation_Agents
155. **Show HN: Eval based agent builder (pls roast us)** | signal 11.2 | tags: langchain, evals | https://github.com/seer-engg/seer
156. **Bad MCP design costs your agent 5x more tokens** | signal 11.1 | tags: benchmark | https://news.ycombinator.com/item?id=48407391
157. **Show HN: Real-time visualization of Claude Code agent orchestration** | signal 11.1 | tags: claude code | https://github.com/patoles/agent-flow
158. **Show HN: OpenJet – An offline agent harness for memory-constrained edge hardware** | signal 11.1 | tags: evals | https://github.com/L-Forster/open-jet
159. **Show HN: A/B Test Your LLM Prompts in Production** | signal 11.1 | tags: evals | https://switchport.ai/
160. **Show HN: Krira Augment – Production-ready RAG in minutes** | signal 11.1 | tags: rag pipeline | https://www.kriralabs.com/waitlist
161. **Show HN: A JSON API for YouTube Transcript with MCP Support** | signal 11.1 | tags: rag pipeline | https://transcriptapi.com/
162. **Show HN: AI Interoperability to the Max – The Intelligence Hub** | signal 11.1 | tags: rag pipeline | https://theintelligencehub.azurewebsites.net/
163. **Ask HN: Anyone solved hallucination or semantic drift in RAG?** | signal 11.1 | tags: rag pipeline | https://news.ycombinator.com/item?id=44746089
164. **My Claude Code Agent for Writing Prompts** | signal 10.8 | tags: claude code | https://olshansky.info/posts/2025-09-29-prompt-writer-agent
165. **Ferretlog: Git log for your Claude Code agent runs** | signal 10.7 | tags: claude code | https://github.com/eitanlebras/ferretlog
166. **Curie – ship Claude Code agents to Kubernetes with Git push** | signal 10.6 | tags: claude code | https://github.com/curie-eng/curie
167. **Replaced Clay.com with Claude Code Agent** | signal 10.6 | tags: claude code | https://github.com/chaitanyya/sales
168. **I built IDE-layer policy enforcement for Claude Code/Cursor agents** | signal 10.4 | tags: claude code | https://www.oculisecurity.com/
169. **Connect multiple Claude Code agents into one collaborative team** | signal 10.4 | tags: claude code | https://openagents.org/showcase
170. **15 AI Coding Agents evaluated with the same prompt** | signal 10.3 | tags: coding agent | https://github.com/The-Focus-AI/june-2025-coding-agent-report
171. **Why Claude Code's Agent Loop Is over 1,400 Lines** | signal 10.3 | tags: claude code | https://internals.laxmena.com/p/why-claude-codes-agent-loop-is-over
172. **Show HN: I run a full software company solo with Claude Code agents** | signal 10.3 | tags: claude code | https://theonemancompany.com/
173. **Show HN: I built an open-source Rust/TS AI agent runtime with a Next.js-style DX** | signal 10.2 | tags: langchain | https://docs.trysoma.ai
174. **New Inference Server for DGX Spark: large model C4:55-90 tok/s no spec decode** | signal 10.2 | tags: benchmark | https://news.ycombinator.com/item?id=49014048
175. **How to evaluate models for production coding agents** | signal 10.2 | tags: coding agent | https://blaxel.ai/blog/llm-coding-benchmarks
176. **FlyCrys – Native Linux GUI for Claude Code Agents (Rust and GTK4)** | signal 10.2 | tags: claude code | https://github.com/SergKam/FlyCrys
177. **20 Claude Code agents, one terminal: a tmux + AppleScript setup** | signal 10.2 | tags: claude code | https://pkarnal.com/blog/parallel-ai-agents
178. **Securely run Claude Code agents in Docker** | signal 10.2 | tags: claude code | https://edspencer.net/2026/2/4/run-claude-code-agents-docker-herdctl
179. **Show HN: I built a context-engineering CLI/MCP tool** | signal 10.1 | tags: retrieval | https://github.com/jerpint/context-llemur
180. **When your coding agent doesn't listen: evaluating a 241-turn Claude session** | signal 10.1 | tags: coding agent | https://www.kurrent.io/blog/when-your-coding-agent-doesnt-listen/
181. **Engine-Bench: Evaluating Coding Agents on Writing Game Engine Code** | signal 10.1 | tags: coding agent | https://github.com/JoshuaPurtell/engine-bench
182. **Evaluating Coding Agents with Terminal-Bench 2.0** | signal 10.1 | tags: coding agent | https://snorkel.ai/blog/evaluating-coding-agent-capabilities-with-terminal-bench-snorkels-role-in-building-the-next-generation-benchmark/
183. **Show HN: A Framework for Evaluating Coding Agents on Sequential SWE** | signal 10.1 | tags: coding agent | https://arxiv.org/abs/2604.03035
184. **ReactBench – evaluation for coding agents on realistic React work** | signal 10.1 | tags: coding agent | https://www.reactbench.com/
185. **No one is evaluating AI coding agents in the way they are used** | signal 10.1 | tags: coding agent | https://marginlab.ai/blog/the-problem-with-coding-benchmarks/
186. **Show HN: Apitoll Payment InfrastructureforAIagents75 Live APIs,USDCmicropayments** | signal 10.1 | tags: langchain | https://github.com/TasnidChain/apitoll-demo
187. **Show HN: Cortex Click – LLM-Driven Developer Marketing Platform** | signal 10.1 | tags: retrieval | https://news.ycombinator.com/item?id=41583460
188. **Cursor Rules for Writing Temporal Workflows with TypeScript** | signal 10.1 | tags: cursor rules | https://stevekinney.com/writing/cursor-rules-temporal-typescript
189. **KernelEvolve: Agentic kernel coding for heterogeneous AI accelerators (Meta)** | signal 10.1 | tags: benchmark | https://news.ycombinator.com/item?id=46442841
190. **Open Source LLMOps Stack** | signal 9.6 | tags: none | https://oss-llmops-stack.com
191. **Show HN: VectorGuard-Nano – Free secure messaging for AI agents** | signal 9.4 | tags: vector | https://github.com/Active-IQ/VectorGuard-Nano
192. **I Got Pwned by a Malicious AI Plugin: A Technical Breakdown** | signal 9.3 | tags: vector | https://news.ycombinator.com/item?id=47109114
193. **Dispelling Misconceptions and Unveiling the Truth about GOT and OT in General** | signal 9.1 | tags: vector | https://news.ycombinator.com/item?id=36643393
194. **Show HN: Local LLM Notepad – run a GPT-style model from a USB stick** | signal 8.8 | tags: none | https://github.com/runzhouye/Local_LLM_Notepad
195. **Ask HN: What's your 2025 code review workflow? GitHub UI feels ancient** | signal 8.6 | tags: code review workflow | https://news.ycombinator.com/item?id=44583146
196. **My Current AI Code Review Workflow** | signal 8.3 | tags: code review workflow | https://guissmo.com/blog/my-current-ai-code-review-workflow/
197. **SHOW HN: A usage circuit breaker for Cloudflare Workers** | signal 8.2 | tags: none | https://news.ycombinator.com/item?id=47322794
198. **Show HN: AI-Friendly Toolchain – Dev Tools for Working with LLMs** | signal 8.1 | tags: prompt engineering | https://github.com/trknhr/awesome-ai-friendly-toolchain
199. **Show HN: Hopsule – Persistent memory and decision layer for AI development** | signal 7.5 | tags: none | https://news.ycombinator.com/item?id=47415402
200. **Show HN: An open-source Operator that can use computers** | signal 7.1 | tags: none | https://github.com/aditya-nadkarni/spongecake
201. **I'm starting to feel tired of AI features that solve problems I don't have** | signal 6.5 | tags: none | https://news.ycombinator.com/item?id=45734499
202. **The way every agent framework handles MCP is a latent security problem** | signal 6.3 | tags: none | https://news.ycombinator.com/item?id=47683498
203. **Does anyone use MCP servers in their dev workflow?** | signal 6.3 | tags: none | https://news.ycombinator.com/item?id=43258552
204. **Building a production-ready RAG pipeline and eval platform** | signal 6.1 | tags: rag pipeline | https://docs.vectorize.io/core-concepts/vectorize-architecture
205. **Show HN: Ductwork – A Go platform for running AI agents on autopilot** | signal 6.0 | tags: none | https://github.com/dneil5648/ductwork
206. **Show HN: Velvet – Data platform with an AI SQL editor** | signal 6.0 | tags: none | https://www.usevelvet.com/
207. **Built a content system that 6x'd traffic. Turning it into product. Want to test?** | signal 6.0 | tags: none | https://news.ycombinator.com/item?id=46337608
208. **Seeking feedback: Integrated product discovery workflow tool** | signal 5.9 | tags: none | https://news.ycombinator.com/item?id=45830436
209. **Show HN: OQP – A verification protocol for AI agents** | signal 5.8 | tags: none | https://github.com/OranproAi/open-qa-protocol
210. **Show HN: Capcat – CLI/TUI to Archive Articles as Markdown and HTML (FOSS)** | signal 5.7 | tags: none | https://capcat.org/
211. **I have a project with ~200k LoC, written with AI codegen. AMA** | signal 5.6 | tags: none | https://news.ycombinator.com/item?id=45351057
212. **If the differentiation is domain and GTM?** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=47329075
213. **Show HN: Agent File (.af) – A standard file format for serializing AI agents** | signal 5.5 | tags: none | https://github.com/letta-ai/agent-file
214. **Windmemory** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=42751099
215. **Show HN: Like grep but for natural questions. Mixtral 8x7B – 28 tok/s on 8GB GPU** | signal 5.5 | tags: none | https://github.com/moritztng/fltr
216. **Prompt to make Claude more autonomous in web dev** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=47379947
217. **Beginner's Guide to MCP (Model Context Protocol)** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=43758713
218. **Show HN: An AI interviewer that probes candidates (and costs $0.99/interview)** | signal 5.5 | tags: none | https://interviewflowai.com/
219. **Show HN: Iterate, evaluate, and monitor LLM-products in a Notion-like workspace** | signal 5.5 | tags: none | https://www.loom.com/share/2ebecddefcc84dcf8ac640ad307601fc
220. **Stopped Using Cursor, for Now** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=44985565
221. **Show HN: Using classic dev books to guide AI agents** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=47098555
222. **Show HN: Acceptify – AI personas that run user acceptance tests on your product** | signal 5.5 | tags: none | https://acceptify.ai/
223. **AI Power Internal Tools** | signal 5.4 | tags: none | https://news.ycombinator.com/item?id=44494999
224. **Show HN: Prismy – GitHub-Native, AI Localization for Dev and Product Teams** | signal 5.4 | tags: none | https://www.prismy.io
225. **Show HN: No-Code, Private AI Agents – Build and Run Locally** | signal 5.4 | tags: none | https://browseragent.dev
226. **Show HN: AI agent that works autonomously while I'm offline** | signal 5.3 | tags: none | https://hire-your-ai-guide.vercel.app
227. **Show HN: DiffDeck, a PR review tool with file context and code navigation** | signal 5.3 | tags: none | https://diffdeck.dev/login
228. **Show HN: Inference API that adapts to your SLA and quality constraints** | signal 5.3 | tags: none | https://models.exosphere.host/
229. **Show HN: Why delegation beats memory in AI Agents** | signal 5.2 | tags: none | https://www.getseer.dev/blogs/lessons-dec-2025
230. **Show HN: We built an AI-agent with a state machine instead of a giant prompt** | signal 5.2 | tags: none | https://nomos.dowhile.dev/
231. **Show HN: OQP – A verification protocol for AI agents** | signal 5.2 | tags: none | https://news.ycombinator.com/item?id=47758801
232. **Show HN: Owl and MCP Integration – Plug-and-play agents with external tools** | signal 5.2 | tags: none | https://www.camel-ai.org/blogs/owl-mcp-toolkit-practice
233. **Show HN: Freeze the Model, Train the Harness** | signal 5.2 | tags: none | https://github.com/workofart/harness-training
234. **Show HN: CriteriaBot – A Universal Customizable Classifier** | signal 5.2 | tags: none | https://criteriabot.io/
235. **Show HN: AI Dev Assistant Framework – Add structure, rules and memory to LLM** | signal 5.2 | tags: none | https://github.com/Fr-e-d/ai-dev-assistant-framework
236. **Show HN: Everdone CodeReview – AI code reviews as a trackable workflow** | signal 5.2 | tags: none | https://everdone.ai/
237. **Show HN: AI Code Review CLI** | signal 5.2 | tags: none | https://github.com/kodustech/cli
238. **Show HN: GPT-reviewer – Simple AI code reviewer for GH Actions** | signal 5.2 | tags: none | https://github.com/vayqerlukashakkarainen/gpt-reviewer
239. **AI-powered Git CLI that generates commit messages automatically** | signal 5.2 | tags: none | https://news.ycombinator.com/item?id=47035076
240. **Show HN: Freeplay – Testing and Evaluation for LLM-powered features** | signal 5.2 | tags: none | https://freeplay.ai/
241. **Show HN: open source framework for building nanoservices** | signal 5.2 | tags: none | https://news.ycombinator.com/item?id=43164465
242. **Show HN: TheFoundry – Easy bootstrapping framework for MultiAgent Systems** | signal 5.1 | tags: none | https://github.com/aavilagallego/TheFoundry
243. **Show HN: LedgerMind – true zero-touch autonomous memory for AI agents** | signal 5.1 | tags: none | https://github.com/sl4m3/ledgermind
244. **Show HN: Mdchat – Markdown-first terminal / CLI tool for LLM collaboration** | signal 5.1 | tags: none | https://www.npmjs.com/package/mdchat
245. **Show HN: WorldBuild Bench repo: testing LLM world coherence with 3D games** | signal 5.1 | tags: none | https://github.com/sebnado/worldbuild-bench
246. **Show HN: CreateMVP.app – First open-source tool to generate MVP specs for LLMs** | signal 5.1 | tags: none | https://createmvps.app/
247. **Show HN: Visual Editor for Cursor** | signal 5.1 | tags: none | https://shuffle.dev/cursor
248. **Show HN: I made an open source Idea to App WebApp** | signal 5.1 | tags: none | https://github.com/rohitg00/CreateMVP
249. **Show HN: Framework to structure LLM dev workflows with Markdown-based protocol** | signal 5.1 | tags: none | https://github.com/Fr-e-d/ai-dev-assistant-framework
250. **My tiny workflow for an AI code review assist** | signal 5.1 | tags: none | https://news.ycombinator.com/item?id=45959846
251. **Show HN: AI code review now available on Azure DevOps** | signal 5.1 | tags: none | https://kodus.io/en/
252. **Show HN: CREV – A Go-based CLI tool for AI code reviews and codebase exports** | signal 5.1 | tags: none | https://news.ycombinator.com/item?id=41757003
253. **Show HN: I released a OS remote agent callable from mobile** | signal 5.0 | tags: none | https://github.com/epavanello/fixodev
254. **Show HN: Arkain – AI-powered Cloud IDE for building real apps from your words** | signal 5.0 | tags: none | https://arkn.ai/qH22w
255. **Show HN: Chaos engineering for LLMs – Making models cross-examine each other** | signal 5.0 | tags: none | https://www.usecouncil.app/
256. **Rate my privacy-first AI ad architecture (patent pending)** | signal 5.0 | tags: none | https://news.ycombinator.com/item?id=47340497
257. **Show HN: Residuum | Agentic AI with continuous context** | signal 5.0 | tags: none | https://github.com/Grizzly-Endeavors/residuum
258. **Ask HN: What are you running for .windsurfrules?** | signal 5.0 | tags: none | https://news.ycombinator.com/item?id=43026575
259. **Show HN: Aidevshield NPM audit for AI coding tool workflows** | signal 5.0 | tags: none | https://github.com/aidevshield/aidevshield
260. **Show HN: AI Resource Manager** | signal 5.0 | tags: none | https://github.com/jomadu/ai-resource-manager
261. **Show HN: Deff – Review AI-generated code changes** | signal 5.0 | tags: none | https://github.com/flamestro/deff
262. **Show HN: Shell script for AI-powered code reviews using local LLMs** | signal 5.0 | tags: none | https://gist.github.com/alwin-augustin-dev/c1caaa30361f7ee320fb9cb957b3b0e9
263. **Show HN: Monitor, audit & alert on AI agent actions and interactions** | signal 5.0 | tags: none | https://pingpulsehq.com
264. **Show HN: Autonomous outbound research and outreach drafts** | signal 5.0 | tags: none | https://www.prospecter.io
265. **Show HN: KitchenAI Open Source LLMops development kit. Notebook to server** | signal 5.0 | tags: none | https://github.com/epuerta9/kitchenai
266. **Ask HN: CI/CD and Hosting for GPU-Based ML Demos** | signal 5.0 | tags: none | https://news.ycombinator.com/item?id=39273121
267. **Show HN: I scraped 200M Shopify products to build a search engine** | signal 2.2 | tags: none | https://www.searchagora.com/#
268. **Why Vertical AI Agents May Replace RPA in Complex Enterprise Workflows** | signal 1.1 | tags: none | https://news.ycombinator.com/item?id=44245754

## github (93)

1. **gmickel/flow-next** | signal 36.6 | tags: claude code, coding agent | https://github.com/gmickel/flow-next
2. **KbWen/agentic-os** | signal 33.4 | tags: claude code, coding agent | https://github.com/KbWen/agentic-os
3. **dimileeh/agent-workspace-fabric** | signal 28.6 | tags: claude code, coding agent | https://github.com/dimileeh/agent-workspace-fabric
4. **rohithkandula19/Ronin** | signal 27.4 | tags: claude code, coding agent, evals | https://github.com/rohithkandula19/Ronin
5. **skynetcmd/m3-memory** | signal 25.1 | tags: langchain, langgraph, retrieval, vector | https://github.com/skynetcmd/m3-memory
6. **EvilFreelancer/secs** | signal 25.1 | tags: claude code, coding agent | https://github.com/EvilFreelancer/secs
7. **Guz007/claude-code-os-plugin** | signal 25.0 | tags: context engineering, claude code | https://github.com/Guz007/claude-code-os-plugin
8. **zhenkun26/BusinessAgent** | signal 24.1 | tags: agent workflow, langgraph, vector | https://github.com/zhenkun26/BusinessAgent
9. **andaro74/regdelta** | signal 24.0 | tags: claude code, langgraph, retrieval, vector | https://github.com/andaro74/regdelta
10. **spences10/my-pi** | signal 23.6 | tags: coding agent, evals | https://github.com/spences10/my-pi
11. **MacFall7/m87-retrieval-evals** | signal 23.0 | tags: evaluation harness, evals, retrieval | https://github.com/MacFall7/m87-retrieval-evals
12. **baalimago/clai** | signal 22.4 | tags: context engineering | https://github.com/baalimago/clai
13. **mastra-ai/mastra** | signal 22 | tags: evals | https://github.com/mastra-ai/mastra
14. **oraziooztas/agent-failure-eval-bench** | signal 21.0 | tags: coding agent, evals, benchmark | https://github.com/oraziooztas/agent-failure-eval-bench
15. **znasllc-io/memql** | signal 20.6 | tags: agent workflow | https://github.com/znasllc-io/memql
16. **MrPeppersDev/agent-infrastructure-landscape** | signal 20.5 | tags: langchain, benchmark, vector | https://github.com/MrPeppersDev/agent-infrastructure-landscape
17. **ReidenXerx/bearing** | signal 20.1 | tags: claude code, coding agent | https://github.com/ReidenXerx/bearing
18. **pharn-dev/pharn-oss** | signal 20.0 | tags: claude code, coding agent | https://github.com/pharn-dev/pharn-oss
19. **MANVENDRA-github/agentry** | signal 20.0 | tags: claude code, coding agent | https://github.com/MANVENDRA-github/agentry
20. **ikstv/agent-foreman** | signal 20.0 | tags: claude code, coding agent | https://github.com/ikstv/agent-foreman
21. **GoogleCloudPlatform/db-context-enrichment** | signal 19.4 | tags: context engineering | https://github.com/GoogleCloudPlatform/db-context-enrichment
22. **NeoLabHQ/context-engineering-kit** | signal 19.2 | tags: claude code | https://github.com/NeoLabHQ/context-engineering-kit
23. **bop-clocktower/canary** | signal 18.1 | tags: claude code | https://github.com/bop-clocktower/canary
24. **sxntixgo/claudecode-documentation** | signal 18.0 | tags: claude code, prompt engineering | https://github.com/sxntixgo/claudecode-documentation
25. **seek-far/DevHarness** | signal 17.2 | tags: evaluation harness, benchmark | https://github.com/seek-far/DevHarness
26. **Nuamanjaleel/finai-ops** | signal 17.0 | tags: evaluation harness, retrieval | https://github.com/Nuamanjaleel/finai-ops
27. **vireshkoli/Research-Agent-Langgraph** | signal 17.0 | tags: evaluation harness, langgraph | https://github.com/vireshkoli/Research-Agent-Langgraph
28. **trapstreet/trapstreet-skills** | signal 16.2 | tags: claude code, evals | https://github.com/trapstreet/trapstreet-skills
29. **linny006/agent-eval-harness** | signal 16.1 | tags: coding agent, benchmark | https://github.com/linny006/agent-eval-harness
30. **heygen-com/hyperframes** | signal 16 | tags: none | https://github.com/heygen-com/hyperframes
31. **runecraftai/harness** | signal 16.0 | tags: coding agent, evals | https://github.com/runecraftai/harness
32. **whitewalker7/keva** | signal 16.0 | tags: coding agent, evals | https://github.com/whitewalker7/keva
33. **xorbitsai/xagent** | signal 16 | tags: none | https://github.com/xorbitsai/xagent
34. **sitewright-cms/sitewright** | signal 15.8 | tags: coding agent | https://github.com/sitewright-cms/sitewright
35. **footprintjs/agentfootprint** | signal 15.5 | tags: context engineering | https://github.com/footprintjs/agentfootprint
36. **Muizzkolapo/agent-actions** | signal 15.4 | tags: context engineering | https://github.com/Muizzkolapo/agent-actions
37. **ShenSeanChen/waku-agent** | signal 15.4 | tags: evals | https://github.com/ShenSeanChen/waku-agent
38. **scitix/sieval** | signal 15.3 | tags: evaluation harness | https://github.com/scitix/sieval
39. **Dikhun/Arctus.ai** | signal 15.1 | tags: agent workflow | https://github.com/Dikhun/Arctus.ai
40. **c-ibarra/c-ibarra.github.io** | signal 15.0 | tags: context engineering | https://github.com/c-ibarra/c-ibarra.github.io
41. **8fpnph7zvw-jpg/AI-Sales-Agent-Platform** | signal 15.0 | tags: agent workflow | https://github.com/8fpnph7zvw-jpg/AI-Sales-Agent-Platform
42. **fu351/Doberman-Core** | signal 14.1 | tags: none | https://github.com/fu351/Doberman-Core
43. **Umarfarook1/rag-document-qa** | signal 14.1 | tags: benchmark, retrieval, vector | https://github.com/Umarfarook1/rag-document-qa
44. **rhein1/agoragentic-integrations** | signal 13.3 | tags: langchain, autogen | https://github.com/rhein1/agoragentic-integrations
45. **eriknewton/sanctuary-framework** | signal 13.2 | tags: claude code | https://github.com/eriknewton/sanctuary-framework
46. **AtvikSecurity/domarinn** | signal 12.9 | tags: evaluation harness | https://github.com/AtvikSecurity/domarinn
47. **agentproto/ts** | signal 12.7 | tags: coding agent | https://github.com/agentproto/ts
48. **juanjuandog/FinSight-AI** | signal 12.0 | tags: vector | https://github.com/juanjuandog/FinSight-AI
49. **Amadeus415/trading_benchV1** | signal 12.0 | tags: evaluation harness | https://github.com/Amadeus415/trading_benchV1
50. **Mohanad49/llm-eval-harness** | signal 12.0 | tags: evaluation harness | https://github.com/Mohanad49/llm-eval-harness
51. **anaschatz/anamnesis** | signal 12.0 | tags: evaluation harness | https://github.com/anaschatz/anamnesis
52. **gpatwa/multi-agent-eval** | signal 12.0 | tags: evaluation harness | https://github.com/gpatwa/multi-agent-eval
53. **MCKRUZ/microsoft-agentic-harness** | signal 11.7 | tags: claude code | https://github.com/MCKRUZ/microsoft-agentic-harness
54. **gokul028h/Adaptive-Agentic-AI-Engine** | signal 11.0 | tags: tool use, retrieval | https://github.com/gokul028h/Adaptive-Agentic-AI-Engine
55. **vinzlercodes/vinzlercodes** | signal 11.0 | tags: evals, retrieval | https://github.com/vinzlercodes/vinzlercodes
56. **Forest-Project-Lab/doctrine** | signal 10.8 | tags: coding agent | https://github.com/Forest-Project-Lab/doctrine
57. **whimzyLive/nightshift-ai** | signal 10.8 | tags: claude code | https://github.com/whimzyLive/nightshift-ai
58. **xonovex/platform** | signal 10.4 | tags: coding agent | https://github.com/xonovex/platform
59. **anivar/jest-skill** | signal 10.3 | tags: claude code | https://github.com/anivar/jest-skill
60. **totalwindupflightsystems/gitreins** | signal 10.2 | tags: coding agent | https://github.com/totalwindupflightsystems/gitreins
61. **ronsilver/agent-covenant** | signal 10.1 | tags: coding agent | https://github.com/ronsilver/agent-covenant
62. **anivar/redux-saga-skill** | signal 10.1 | tags: claude code | https://github.com/anivar/redux-saga-skill
63. **anivar/msw-skill** | signal 10.1 | tags: claude code | https://github.com/anivar/msw-skill
64. **felixross66/claude-ai-coding-kit-2026** | signal 10.1 | tags: claude code | https://github.com/felixross66/claude-ai-coding-kit-2026
65. **mctang24/go-coding-agent** | signal 10.0 | tags: coding agent | https://github.com/mctang24/go-coding-agent
66. **mhsutton07/Fabrica** | signal 10.0 | tags: claude code | https://github.com/mhsutton07/Fabrica
67. **gliakit/gliakit** | signal 10.0 | tags: claude code | https://github.com/gliakit/gliakit
68. **kevglynn/house-rules** | signal 10.0 | tags: coding agent | https://github.com/kevglynn/house-rules
69. **igorishchenko/ai-project-bootstrap** | signal 10.0 | tags: cursor rules | https://github.com/igorishchenko/ai-project-bootstrap
70. **objectstack-ai/objectstack** | signal 9.1 | tags: none | https://github.com/objectstack-ai/objectstack
71. **mbsdeepak/loom** | signal 9.0 | tags: retrieval, vector | https://github.com/mbsdeepak/loom
72. **Ismail-2001/The-Kubernetes-of-AI-Agents** | signal 8.7 | tags: evals | https://github.com/Ismail-2001/The-Kubernetes-of-AI-Agents
73. **division-sh/swarm** | signal 8.3 | tags: none | https://github.com/division-sh/swarm
74. **cpa03/blueprintify** | signal 8.0 | tags: none | https://github.com/cpa03/blueprintify
75. **VPSDance/ai-proxy-rules** | signal 8.0 | tags: none | https://github.com/VPSDance/ai-proxy-rules
76. **dcellison/kai** | signal 6.8 | tags: none | https://github.com/dcellison/kai
77. **Bande-a-Bonnot/Boucle-framework** | signal 6.0 | tags: none | https://github.com/Bande-a-Bonnot/Boucle-framework
78. **Lvvphole/ai-account-prioritization** | signal 6.0 | tags: evals | https://github.com/Lvvphole/ai-account-prioritization
79. **AEON-7/Aeon-Bench-Pod** | signal 6.0 | tags: benchmark | https://github.com/AEON-7/Aeon-Bench-Pod
80. **nori72ny/myAIspecials** | signal 5.5 | tags: none | https://github.com/nori72ny/myAIspecials
81. **bestdeejay-design/agent-skills** | signal 5.2 | tags: none | https://github.com/bestdeejay-design/agent-skills
82. **api-evangelist/poolside** | signal 5.0 | tags: none | https://github.com/api-evangelist/poolside
83. **homayoun-safarpour/agent-eval-workbench** | signal 5.0 | tags: benchmark | https://github.com/homayoun-safarpour/agent-eval-workbench
84. **amaan-ur-raheman/codesheriff** | signal 5.0 | tags: retrieval | https://github.com/amaan-ur-raheman/codesheriff
85. **exha1078/agentic-workflow-orchestrator** | signal 5.0 | tags: langgraph | https://github.com/exha1078/agentic-workflow-orchestrator
86. **vsingh45/par-entbench** | signal 5.0 | tags: benchmark | https://github.com/vsingh45/par-entbench
87. **sakthi-kr/enterprise-knowledge-agent** | signal 5.0 | tags: retrieval | https://github.com/sakthi-kr/enterprise-knowledge-agent
88. **api-evangelist/writer** | signal 5.0 | tags: retrieval | https://github.com/api-evangelist/writer
89. **api-evangelist/tiyaro** | signal 5.0 | tags: none | https://github.com/api-evangelist/tiyaro
90. **archgate/cli** | signal 3.6 | tags: none | https://github.com/archgate/cli
91. **weby-homelab/power-framework** | signal 2.2 | tags: none | https://github.com/weby-homelab/power-framework
92. **JustinMelger/ZeroOne-ops** | signal 1.5 | tags: none | https://github.com/JustinMelger/ZeroOne-ops
93. **HIDORAKAI002/ai-workspace-archive** | signal 1.4 | tags: none | https://github.com/HIDORAKAI002/ai-workspace-archive

## stackoverflow (34)

1. **Claude Code - Looking for guidance on where to start with coding and tools** | signal 13.6 | tags: claude code | https://stackoverflow.com/questions/79927051/claude-code-looking-for-guidance-on-where-to-start-with-coding-and-tools
2. **Best Approach to Evaluate a Graph RAG Pipeline Using Metrics?** | signal 11.4 | tags: rag pipeline, retrieval | https://stackoverflow.com/questions/78881336/best-approach-to-evaluate-a-graph-rag-pipeline-using-metrics
3. **How do you learn without AI?** | signal 9.5 | tags: none | https://stackoverflow.com/questions/79832798/how-do-you-learn-without-ai
4. **Globally catch exceptions in a WPF application?** | signal 9.4 | tags: none | https://stackoverflow.com/questions/793100/globally-catch-exceptions-in-a-wpf-application
5. **MCPToolConversionError: Failed to get tools from MCP server: 404** | signal 5.3 | tags: langchain | https://stackoverflow.com/questions/79705666/mcptoolconversionerror-failed-to-get-tools-from-mcp-server-404
6. **Langchain, Huggingface: Can&#39;t evaluate model with two different inputs** | signal 5.3 | tags: langchain | https://stackoverflow.com/questions/76137512/langchain-huggingface-cant-evaluate-model-with-two-different-inputs
7. **Input validation error: &#39;1.57&#39; is not of type &#39;number&#39; from langchain_mcp_adapter** | signal 5.1 | tags: langchain | https://stackoverflow.com/questions/79935672/input-validation-error-1-57-is-not-of-type-number-from-langchain-mcp-adapte
8. **Restrict responses from a language model (LLM) to only information available in a specific document** | signal 5.1 | tags: retrieval | https://stackoverflow.com/questions/78333793/restrict-responses-from-a-language-model-llm-to-only-information-available-in
9. **Is there a framework of many open-source code LLMs for generation?** | signal 5.0 | tags: benchmark | https://stackoverflow.com/questions/78287327/is-there-a-framework-of-many-open-source-code-llms-for-generation
10. **How can I debug an internal error in the .NET Runtime?** | signal 4.5 | tags: none | https://stackoverflow.com/questions/14238657/how-can-i-debug-an-internal-error-in-the-net-runtime
11. **Error - &quot;There is no script engine for file extension .vbs&quot; when using &quot;Git Bash Here&quot; in Windows 7** | signal 3.5 | tags: none | https://stackoverflow.com/questions/17757248/error-there-is-no-script-engine-for-file-extension-vbs-when-using-git-bash
12. **Transparent user session over several sites (single sign-on + single sign-off)** | signal 3.3 | tags: none | https://stackoverflow.com/questions/1043111/transparent-user-session-over-several-sites-single-sign-on-single-sign-off
13. **Current state and solutions for OpenGL over Windows Remote** | signal 2.5 | tags: none | https://stackoverflow.com/questions/51705471/current-state-and-solutions-for-opengl-over-windows-remote
14. **How can an AI assistant interact with Aspen plus through Python?** | signal 2.2 | tags: none | https://stackoverflow.com/questions/79904274/how-can-an-ai-assistant-interact-with-aspen-plus-through-python
15. **Store Django Log messages in a database?** | signal 2.1 | tags: none | https://stackoverflow.com/questions/11887816/store-django-log-messages-in-a-database
16. **How to validate the origin of a web service invokation** | signal 2.0 | tags: none | https://stackoverflow.com/questions/14023348/how-to-validate-the-origin-of-a-web-service-invokation
17. **Running Keras model for prediction in multiple threads** | signal 1.9 | tags: none | https://stackoverflow.com/questions/43136293/running-keras-model-for-prediction-in-multiple-threads
18. **How to disable context menu on right click/long touch in a kiosk mode of Chrome?** | signal 1.8 | tags: none | https://stackoverflow.com/questions/28222548/how-to-disable-context-menu-on-right-click-long-touch-in-a-kiosk-mode-of-chrome
19. **Mac OS X: Can one process render to another process&#39;s window?** | signal 1.8 | tags: none | https://stackoverflow.com/questions/583202/mac-os-x-can-one-process-render-to-another-processs-window
20. **Client-Side CommunicationException while Service works properly** | signal 1.8 | tags: none | https://stackoverflow.com/questions/15429934/client-side-communicationexception-while-service-works-properly
21. **How can I get a password containing a caret (^) passed unchanged as a parameter to a Windows batch file?** | signal 1.8 | tags: none | https://stackoverflow.com/questions/5254460/how-can-i-get-a-password-containing-a-caret-passed-unchanged-as-a-parameter
22. **can I turn off optimization, so in-scope variables from closures aren&#39;t &quot;optimized out&quot;** | signal 1.7 | tags: none | https://stackoverflow.com/questions/58861823/can-i-turn-off-optimization-so-in-scope-variables-from-closures-arent-optimiz
23. **Windows: avoid pushing full x86 context on stack** | signal 1.7 | tags: none | https://stackoverflow.com/questions/994555/windows-avoid-pushing-full-x86-context-on-stack
24. **Which StatsD client should I use for a java/grails project?** | signal 1.6 | tags: none | https://stackoverflow.com/questions/17243168/which-statsd-client-should-i-use-for-a-java-grails-project
25. **Why is every new AI IDE forcing a minimalist, &quot;chat-first&quot; UI on us?** | signal 1.5 | tags: none | https://stackoverflow.com/questions/79943845/why-is-every-new-ai-ide-forcing-a-minimalist-chat-first-ui-on-us
26. **Why are MCPs needed at all?** | signal 1.4 | tags: none | https://stackoverflow.com/questions/79866688/why-are-mcps-needed-at-all
27. **Set audio endpoint devices application specific (programmatically)** | signal 1.2 | tags: none | https://stackoverflow.com/questions/52973464/set-audio-endpoint-devices-application-specific-programmatically
28. **How can I set the RTS with ioctl() in a Mac plugin?** | signal 1.2 | tags: none | https://stackoverflow.com/questions/14693724/how-can-i-set-the-rts-with-ioctl-in-a-mac-plugin
29. **IntelliJ IDEA: Cannot run program &quot;C:\Program Files\nodejs\npx&quot;: CreateProcess error=193 when using MCP server** | signal 1.2 | tags: none | https://stackoverflow.com/questions/79722494/intellij-idea-cannot-run-program-c-program-files-nodejs-npx-createprocess-e
30. **Query OLAP Mondrian (MDX, XMLA) with a Python interface?** | signal 1.2 | tags: none | https://stackoverflow.com/questions/3793215/query-olap-mondrian-mdx-xmla-with-a-python-interface
31. **Windows Aero Rendering Bug** | signal 1.1 | tags: none | https://stackoverflow.com/questions/27450042/windows-aero-rendering-bug
32. **Harvesting the power of highly-parallel computers with python scientific code** | signal 1.1 | tags: none | https://stackoverflow.com/questions/18234484/harvesting-the-power-of-highly-parallel-computers-with-python-scientific-code
33. **searching good embedded &amp; hosting language pair** | signal 1.1 | tags: none | https://stackoverflow.com/questions/7843234/searching-good-embedded-hosting-language-pair
34. **ruamel_yaml.constructor.ConstructorError: could not determine a constructor for the tag &#39;tag:yaml.org,2002:python/tuple&#39; in &quot;&lt;unicode string&gt;&quot;** | signal 1.0 | tags: none | https://stackoverflow.com/questions/66609054/ruamel-yaml-constructor-constructorerror-could-not-determine-a-constructor-for

## lobsters (74)

1. **Google’s exponential path to climate-wrecking digital bloat** | signal 18.2 | tags: none | https://ketanjoshi.co/2026/07/01/googles-exponential-path-to-climate-wrecking-digital-bloat/
2. **What side projects have you enjoyed the most?** | signal 17.1 | tags: none | https://lobste.rs/s/mbk56v/what_side_projects_have_you_enjoyed_most
3. **What are you doing this weekend?** | signal 15.2 | tags: none | https://lobste.rs/s/ucaatd/what_are_you_doing_this_weekend
4. **What are you doing this weekend?** | signal 15.1 | tags: none | https://lobste.rs/s/won0b2/what_are_you_doing_this_weekend
5. **Why Rocq is better than Lean for program verification** | signal 13.8 | tags: none | https://joomy.korkutblech.com/posts/2026-07-28-why-rocq-is-better.html
6. **The feature in OxCaml that more languages should steal** | signal 13.7 | tags: none | https://theconsensus.dev/p/2026/06/27/the-feature-in-oxcaml-more-languages-should-steal.html
7. **Self-hosting email the hard way from your own routable IPv4 block up** | signal 12.9 | tags: none | https://anil.recoil.org/notes/recoil-self-hosting-2026
8. **Stripe Just Wants a Number** | signal 11.9 | tags: none | https://blog.exe.dev/billable-facts
9. **irken: A tiny hackable full-featured IRC client** | signal 11.7 | tags: none | https://codeberg.org/dlowe/irken
10. **OCaml 5.5.0 released** | signal 11.3 | tags: none | https://discuss.ocaml.org/t/ocaml-5-5-0-released/18265
11. **some software talks i like** | signal 11.3 | tags: none | https://char.lt/blog/2026/08/talks-i-like/
12. **Meta Garbage Collection: Using OCaml's GC to GC Rust** | signal 10.4 | tags: none | https://soteria-tools.com/blog/meta-garbage-collection
13. **Two years of vector search at Notion: 10x scale, 1/10th cost** | signal 10.1 | tags: vector | https://www.notion.com/blog/two-years-of-vector-search-at-notion
14. **What does it mean to be a mathematician when AI does the math?** | signal 9.6 | tags: none | https://spectrum.ieee.org/ai-in-mathematics
15. **What are you doing this week?** | signal 9.6 | tags: none | https://lobste.rs/s/foxgva/what_are_you_doing_this_week
16. **Writing Toy Software Is A Joy (2025)** | signal 9.6 | tags: none | https://blog.jsbarretto.com/post/software-is-joy
17. **Faster Than Ninja** | signal 9.4 | tags: none | https://build2.org/blog/faster-than-ninja.xhtml
18. **Dependency Cultures - Richard Feldman (Software Should Work Conf 2026)** | signal 9.2 | tags: none | https://www.youtube.com/watch?v=E82ly38YEEQ
19. **Taking OCaml and Eio for a spin** | signal 8.9 | tags: none | https://mattjhall.co.uk/posts/taking-ocaml-eio-for-a-spin.html
20. **Painting with Gaussians** | signal 8.7 | tags: none | https://yogthos.net/posts/2026-08-03-splat-painter.html
21. **Quick & Easy Parser Combinators** | signal 8.6 | tags: none | https://www.cyan.sh/blog/posts/tutorial-quick-easy-parser-combinators.html
22. **The strain in your brain** | signal 8.5 | tags: none | https://anirudh.fi/strain
23. **why use F# for scripting and automation?** | signal 8.3 | tags: none | https://iev.ee/blog/why-use-fsharp/
24. **AI Learns the "Dark Art" of RF Chip Design** | signal 8.2 | tags: none | https://spectrum.ieee.org/ai-radio-chip-design
25. **"How to Think About AI": Cory Doctorow on Big Tech, Understanding AI, Labor Automation & More** | signal 8.2 | tags: none | https://www.youtube.com/watch?v=OBUzl_IaWIw
26. **Guarded methods in OCaml** | signal 8.1 | tags: none | https://xvw.lol/en/articles/oop-refl.html
27. **A line-by-line translation of the OCaml runtime from C to Rust** | signal 8.1 | tags: none | https://discuss.ocaml.org/t/a-line-by-line-translation-of-the-ocaml-runtime-from-c-to-rust/18247
28. **Inventing ELIZA - How the First Chatbot Shaped the Future of AI** | signal 8.0 | tags: none | https://mitpress.mit.edu/9780262052481/inventing-eliza/
29. **Why ML/OCaml are good for writing compilers (1998)** | signal 8.0 | tags: none | https://flint.cs.yale.edu/cs421/case-for-ml.html
30. **A Path Not Taken for OxCaml** | signal 8.0 | tags: none | https://joel.place/blog/path-not-taken/
31. **strace-ui, Bonsai_term, and the TUI renaissance** | signal 7.8 | tags: none | https://blog.janestreet.com/strace-ui-bonsai-term-and-the-tui-renaissance/
32. **Use Task Runners for Common Coding Tasks** | signal 7.8 | tags: none | https://hamvocke.com/blog/task-runners/
33. **Logic for Programmers** | signal 7.8 | tags: none | https://logicforprogrammers.com/
34. **The following is a valid DOS COM executable** | signal 7.7 | tags: none | https://oldbytes.space/@gloriouscow/117045701876951834
35. **jj_tui: terminal user interface to jujutsu focused on speed and clarity** | signal 7.5 | tags: none | https://tangled.org/elidowling.com/jj_tui
36. **Introducing Incremental (2015)** | signal 7.4 | tags: none | https://blog.janestreet.com/introducing-incremental/
37. **You Could Have Come Up With Kimi Delta Attention** | signal 7.3 | tags: none | https://blog.doubleword.ai/you-could-have-come-up-with-kimi-delta-attention
38. **Syntax with Purpose in a Programming Language** | signal 7.3 | tags: none | https://www.youtube.com/watch?v=_HLZoeFREFo
39. **Retries don't fix eventual consistency** | signal 7.3 | tags: none | https://var0.xyz/posts/retries-dont-fix-eventual-consistency.html
40. **Full flattening of nested data parallelism** | signal 7.2 | tags: none | https://futhark-lang.org/blog/2026-07-31-full-flattening.html
41. **Why we write our own C and C++ inference engines** | signal 7.1 | tags: none | https://localai.io/blog/why-we-write-our-own-engines/
42. **From constraint models to playable puzzle games** | signal 7.1 | tags: none | https://zayenz.se/blog/post/constraint-generated-puzzle-games/
43. **MAX models can now run on Apple silicon GPUs** | signal 7.0 | tags: none | https://forum.modular.com/t/max-models-can-now-run-on-apple-silicon-gpus/3283
44. **Xavier Leroy on programming, languages and formal verification** | signal 7.0 | tags: none | https://www.youtube.com/watch?v=9Cswiqrq6So
45. **bonsai: A library for building dynamic webapps, using Js_of_ocaml** | signal 6.8 | tags: none | https://github.com/janestreet/bonsai
46. **Languages as designed latent spaces** | signal 6.6 | tags: none | https://blog.jsbarretto.com/post/languages-as-latent-spaces
47. **What Rose Petals Teach Us about Induction** | signal 6.6 | tags: none | https://www.oranlooney.com/post/rose-petals/
48. **Flow’s OCaml to Rust Port** | signal 6.6 | tags: none | https://medium.com/flow-type/flows-ocaml-to-rust-port-78b95bcf49e9
49. **A novel computer Scrabble engine based on probability that performs at championship level (2021)** | signal 6.5 | tags: none | https://upcommons.upc.edu/server/api/core/bitstreams/1339ae43-3d65-4015-8e11-3689e5572b23/content
50. **Data race freedom in OxCaml** | signal 6.5 | tags: none | https://kcsrk.info/ocaml/oxcaml/x-ocaml/blogging/2026/05/07/data-race-freedom-in-oxcaml/
51. **Asana’s fascinating Tab shortcuts** | signal 6.5 | tags: none | https://unsung.aresluna.org/asanas-fascinating-tab-shortcuts/
52. **Tensor is the might** | signal 6.4 | tags: none | https://zserge.com/posts/tensor/
53. **Revision Prompting improves industrial LLM processes** | signal 6.3 | tags: none | https://revisionprompting.info/
54. **social media rabbit holes, clusters, and the relative mixing times of random walks** | signal 6.3 | tags: none | https://notes.hella.cheap/twitter-isnt-a-town-square-its-a-high-school-cafeteria.html
55. **Program images and portable Scheme backends for Jolt** | signal 6.3 | tags: none | https://yogthos.net/posts/2026-08-07-portable-jolt.html
56. **Human-like Neural Nets by Catapulting** | signal 6.2 | tags: none | https://gwern.net/llm-catapult
57. **A global workspace in language models** | signal 6.2 | tags: none | https://www.anthropic.com/research/global-workspace
58. **Convolutional Neural Networks in APL (2019)** | signal 6.2 | tags: none | https://dl.acm.org/doi/epdf/10.1145/3315454.3329960
59. **Comparing Transformers and Hybrid Models at the Token Level** | signal 6.2 | tags: none | https://arxiv.org/pdf/2606.20936
60. **Language integrated LLMs as an OCaml function** | signal 6.2 | tags: none | https://anil.recoil.org/notes/language-integrated-llms
61. **Announcing Pyro Caml: The First Continuous Profiler for OCaml** | signal 6.2 | tags: none | https://semgrep.dev/blog/2026/announcing-pyro-caml-continuous-profiler-ocaml
62. **OCaml Infrastructure: How the opam-repository Works** | signal 6.2 | tags: none | https://ocaml.org/backstage/2025-11-05-how-the-opam-repository-works
63. **O(x)Caml in Space** | signal 6.2 | tags: none | https://gazagnaire.org/blog/2026-05-14-borealis.html
64. **Shrinking the OxCaml js_of_ocaml bundle: 285 MB to 4 MB** | signal 6.2 | tags: none | https://kcsrk.info/ocaml/oxcaml/modes/2026/05/10/shrinking-the-oxcaml-bundle/
65. **This Quarter in KDE Digital Sovereignty: Q2 2026** | signal 6.2 | tags: none | https://blogs.kde.org/2026/08/06/this-quarter-in-kde-digital-sovereignty-q2-2026/
66. **A game made only with sine waves** | signal 6.2 | tags: none | https://www.youtube.com/watch?v=Qr3VsZYQy4s
67. **Categorization with NLP** | signal 6.1 | tags: none | https://softwaremaniacs.org/blog/2026/07/30/categorization-with-nlp/en/
68. **Debootstrapping without Archeology: Stacked Implementations in Camlboot** | signal 6.1 | tags: none | https://arxiv.org/abs/2202.09231
69. **Project-Specific clangd Configuration with a Temporary Shell** | signal 6.1 | tags: none | https://felix-knorr.net/posts/2026-07-31-lsp-config.html
70. **Categorization with NLP** | signal 6.0 | tags: none | https://softwaremaniacs.org/blog/2026/07/30/categorization-with-nlp/
71. **Why Do Cognitive Scientists Hate LLMs? (2023)** | signal 6.0 | tags: none | https://minihf.com/posts/2023-10-16-hermes-lecture-3-why-do-cognitive-scientists-hate-llms/
72. **Matrix Orthogonalization Improves Memory in Recurrent Models** | signal 6.0 | tags: none | https://ayushtambde.com/blog/matrix-orthogonalization-improves-memory-in-recurrent-models/
73. **Robust AI Security and Alignment: A Sisyphean Endeavor?** | signal 6.0 | tags: none | https://ieeexplore.ieee.org/document/11475847/
74. **GPT2-BASIC: Portable Machine Intelligence in BASIC** | signal 6.0 | tags: none | https://github.com/tsotchke/gpt2-basic