# AI Practice Signal Brief

*2026-08-07 02:00 UTC* | 551 collected | 451 kept (signal>=1) | Sources: {'hackernews': 281, 'github': 118, 'lobsters': 74, 'stackoverflow': 78}

> High-recall dump: grouped by source, sorted by signal (heuristic priority hint only). Sift downstream for value.

## hackernews (265)

1. **Launch HN: Extend (YC W23) – Turn your messiest documents into data** | signal 43.6 | tags: context engineering, prompt engineering, evals | https://www.extend.ai/
2. **Show HN: Claude Code skills that build complete Godot games** | signal 41 | tags: claude code, coding agent | https://github.com/htdt/godogen
3. **Launch HN: Relari (YC W24) – Identify the root cause of problems in LLM apps** | signal 40.3 | tags: coding agent, rag pipeline, tool use, retrieval | https://news.ycombinator.com/item?id=39641105
4. **Launch HN: Vellum (YC W23) – Dev Platform for LLM Apps** | signal 39.8 | tags: prompt engineering, llm ops, vector | https://news.ycombinator.com/item?id=35042836
5. **Launch HN: LlamaFarm (YC W22) – Open-source framework for distributed AI** | signal 39.3 | tags: rag pipeline, evals, retrieval, vector | https://github.com/llama-farm/llamafarm
6. **Show HN: Representing Agents as MCP Servers** | signal 38.1 | tags: agent workflow, mcp agent | https://github.com/lastmile-ai/mcp-agent/tree/main/examples/mcp_agent_server
7. **Show HN: Armalo AI – The Infrastructure for Agent Networks** | signal 37.8 | tags: agent workflow, langchain, evals, benchmark | https://news.ycombinator.com/item?id=47244042
8. **Launch HN: Hoplite (YC S26) – Effortlessly deploy cloud coding agents** | signal 37.0 | tags: claude code, coding agent | https://hoplite.sh
9. **Show HN: Laminar – Open-Source DataDog + PostHog for LLM Apps, Built in Rust** | signal 37 | tags: rag pipeline, evals, vector | https://github.com/lmnr-ai/lmnr
10. **Context Rot: How increasing input tokens impacts LLM performance** | signal 36 | tags: context engineering | https://research.trychroma.com/context-rot
11. **Show HN: R2R – Open-source framework for production-grade RAG** | signal 36 | tags: rag pipeline, retrieval, vector | https://github.com/SciPhi-AI/R2R
12. **Show HN: I open-sourced my Go and Next B2B SaaS Starter (deploy anywhere, MIT)** | signal 35.1 | tags: claude code, retrieval, vector | https://github.com/moasq/production-saas-starter
13. **Show HN: Product analytics (and evals) for agent sessions on your MCP** | signal 34.7 | tags: claude code, coding agent, evals | https://armature.tech/
14. **Show HN: Autofix Bot – Hybrid static analysis and AI code review agent** | signal 34.5 | tags: claude code, coding agent, benchmark | https://news.ycombinator.com/item?id=46237358
15. **Show HN: ClawMem – Open-source agent memory with SOTA local GPU retrieval** | signal 34.2 | tags: claude code, coding agent, retrieval, vector | https://github.com/yoloshii/ClawMem
16. **Why My Open-Source Project Hasn't Done Better** | signal 32.4 | tags: evaluation harness, coding agent, benchmark | https://news.ycombinator.com/item?id=49046999
17. **I looked at 1000s of RAG queries to figure out the problem with semantic search** | signal 31.9 | tags: rag pipeline, evals, benchmark, retrieval, vector | https://news.ycombinator.com/item?id=42299349
18. **Launch HN: Captain (YC W26) – Automated RAG for Files** | signal 30.4 | tags: rag pipeline, retrieval, vector | https://www.runcaptain.com/
19. **Show HN: Cognikernel- Local Memory for AI Coding Assistants** | signal 30.3 | tags: claude code, coding agent, retrieval | https://github.com/KanishkNoir/cognikernel
20. **How do you pick a Coding Agent HN?** | signal 30.2 | tags: claude code, coding agent, benchmark | https://news.ycombinator.com/item?id=46634773
21. **Show HN: Wegent –Open Source Cloud Coding Agent Platform** | signal 30.1 | tags: claude code, coding agent, retrieval | https://github.com/wecode-ai/Wegent
22. **Launch HN: Sonarly (YC W26) – AI agent to triage and fix your production alerts** | signal 29.9 | tags: claude code, coding agent | https://sonarly.com/
23. **Launch HN: Manufact (YC S25) – MCP Cloud** | signal 28.6 | tags: claude code | https://manufact.com
24. **Show HN: Interbase – Long-running AI goals and aliases for any model** | signal 28.1 | tags: agent workflow, code review workflow | https://github.com/agentsorchestrationcompany/interbase
25. **Show HN: Open-source EU AI Act compliance layer for AI agents (8/2026 deadline)** | signal 27.3 | tags: rag pipeline, langchain, autogen, retrieval | https://news.ycombinator.com/item?id=47141347
26. **Show HN: Firebender, a simple coding agent for Android Engineers** | signal 27.2 | tags: coding agent, evals | https://docs.firebender.com/get-started/agent
27. **Launch HN: Roe AI (YC W24) – AI-powered data warehouse to query multimodal data** | signal 27.0 | tags: prompt engineering, vector | https://news.ycombinator.com/item?id=41202694
28. **Show HN: Superlog (YC P26) – Observability that installs itself and fixes bugs** | signal 26.7 | tags: claude code | https://superlog.sh/
29. **Show HN: A library to convert+deploy existing agent projects as MCP servers** | signal 26.5 | tags: agent workflow, langgraph | https://github.com/NapthaAI/automcp
30. **Show HN: Single-agent long-horizon reasoning within one LLM run** | signal 26.4 | tags: context engineering, tool use | https://huggingface.co/papers/2507.16784
31. **Show HN: Typia (20,000x faster validator) challenges to Agentic AI with compiler** | signal 26.1 | tags: agent workflow, function calling | https://typia.io/articles/typia-challenges-to-agentic-ai-with-its-compiler-skill.html
32. **Evaluating AGENTS.md: are they helpful for coding agents?** | signal 26 | tags: coding agent | https://arxiv.org/abs/2602.11988
33. **Show HN: Duplicate 3 layers in a 24B LLM, logical deduction .22→.76. No training** | signal 26 | tags: benchmark | https://github.com/alainnothere/llm-circuit-finder
34. **DeepClaude – Claude Code agent loop with DeepSeek V4 Pro** | signal 26 | tags: claude code | https://github.com/aattaran/deepclaude
35. **Show HN: AgentLint – ESLint for your coding agents** | signal 25.8 | tags: claude code, coding agent | https://github.com/samilozturk/agentlint
36. **Show HN: Gentrace – connect to your LLM app code and run/eval it from a UI** | signal 25.8 | tags: llm ops, evals, retrieval | https://gentrace.ai/
37. **Show HN: Agents, run any coding agent on your subscription not API costs** | signal 25.7 | tags: claude code, coding agent | https://agents-cli.sh
38. **Show HN: HoneyHive – An unified evaluation and monitoring platform for LLM apps** | signal 25.6 | tags: rag pipeline, langchain, benchmark, vector | https://news.ycombinator.com/item?id=37777683
39. **Show HN: 20+ Claude Code agents coordinating on real work (open source)** | signal 25.4 | tags: claude code | https://github.com/mutable-state-inc/lean-collab
40. **Ask HN: Why do AI coding agents refuse to save their own observations?** | signal 25.3 | tags: claude code, coding agent | https://news.ycombinator.com/item?id=47170501
41. **Show HN: Devplan – Generate specs and coding prompts with deep context** | signal 25.3 | tags: claude code, coding agent | https://www.devplan.com/
42. **Show HN: Irpapers – Visual embeddings vs. OCR trade-offs in scientific PDFs** | signal 25.2 | tags: rag pipeline, benchmark, retrieval, vector | https://github.com/weaviate/query-agent-benchmarking
43. **Show HN: Agents Council – Connect Claude, Codex, and Local Agents via MCP** | signal 25.1 | tags: claude code, coding agent | https://github.com/MrLesk/agents-council
44. **Show HN: Projekt [Free Alpha] – All-in-one workspace for building with agents** | signal 25.1 | tags: claude code, coding agent | https://www.getprojekt.com/
45. **Show HN: Burr – A framework for building and debugging GenAI apps faster** | signal 25.1 | tags: langchain, evals | https://github.com/DAGWorks-Inc/burr
46. **Show HN: GibRAM an in-memory ephemeral GraphRAG runtime for retrieval** | signal 24.8 | tags: rag pipeline, retrieval, vector | https://github.com/gibram-io/gibram
47. **Show HN: Open-Source Animal Crossing–Style UI for Claude Code Agents** | signal 24.6 | tags: claude code | https://github.com/outworked/outworked/releases/tag/v0.3.0
48. **Launch HN: Talc AI (YC S23) – Test Sets for AI** | signal 24.6 | tags: benchmark | https://news.ycombinator.com/item?id=39042093
49. **Show HN: Real-time dashboard for Claude Code agent teams** | signal 24.4 | tags: claude code | https://github.com/simple10/agents-observe
50. **Q Evaluation Harness: open-source evals for LLMs on q/kdb+** | signal 23.1 | tags: evaluation harness, evals | https://github.com/KxSystems/q-evaluation-harness
51. **Show HN: Abralo – Free, easy way to run several Claude Code agents in one window** | signal 23.1 | tags: claude code | https://abralo.com/
52. **AI receptionist that answers real phone calls** | signal 22.9 | tags: evaluation harness, retrieval | https://news.ycombinator.com/item?id=45541794
53. **Show HN: ModelX – Prediction Exchange for LLMs** | signal 22.2 | tags: evaluation harness, benchmark | https://model-x.up.railway.app/
54. **Show HN: Caliper – pass@k reliability testing for Claude Code and Codex skills** | signal 21.8 | tags: claude code, evals | https://github.com/edonadei/caliper
55. **Show HN: Web-eval-agent – Let the coding agent debug itself** | signal 21.6 | tags: coding agent | https://github.com/Operative-Sh/web-eval-agent
56. **Show HN: Free local security checks for AI coding in VSCode, Cursor and Windsurf** | signal 21.6 | tags: coding agent | https://news.ycombinator.com/item?id=44309393
57. **Productivity apps won't disappear, just the need to open them will** | signal 21.2 | tags: agent workflow | https://news.ycombinator.com/item?id=44139226
58. **Show HN: AgentKit – JavaScript Alternative to OpenAI Agents SDK with Native MCP** | signal 21.2 | tags: coding agent | https://github.com/inngest/agent-kit
59. **Why Codex works better than Claude Code for my production monolith** | signal 21.1 | tags: claude code, benchmark | https://news.ycombinator.com/item?id=47945185
60. **Show HN: Skilldeck – Desktop app to manage AI agent skill files across tools** | signal 21.1 | tags: claude code, tool use | https://github.com/ali-erfan-dev/skilldeck
61. **Show HN: Notebooklm-Py – Unofficial Python API for Google NotebookLM** | signal 21.1 | tags: claude code, rag pipeline | https://github.com/teng-lin/notebooklm-py
62. **Show HN: Rowboat – Open-source IDE for multi-agent systems** | signal 21 | tags: none | https://github.com/rowboatlabs/rowboat
63. **Show HN: Jido 2.0, Elixir Agent Framework** | signal 21 | tags: none | https://jido.run/blog/jido-2-0-is-here
64. **Launch HN: mrge.io (YC X25) – Cursor for code review** | signal 21 | tags: none | https://news.ycombinator.com/item?id=43692476
65. **Launch HN: Magic Patterns (YC W23) – AI Design and Prototyping for Product Teams** | signal 21 | tags: none | https://news.ycombinator.com/item?id=43752176
66. **Show HN: Inngest 1.0 – Open-source durable workflows on every platform** | signal 21 | tags: none | https://www.inngest.com/
67. **Show HN: Mcp-Agent – Build effective agents with Model Context Protocol** | signal 20.6 | tags: rag pipeline | https://github.com/lastmile-ai/mcp-agent
68. **Tell HN: Dealing with VCs, my experience** | signal 20.6 | tags: none | https://news.ycombinator.com/item?id=866299
69. **Show HN: AgentWing – make AI agents complete tasks faster** | signal 20.3 | tags: agent workflow | https://news.ycombinator.com/item?id=48200511
70. **Context-Bench: Benchmarking LLMs on Agentic Context Engineering** | signal 20.2 | tags: context engineering, benchmark | https://www.letta.com/blog/context-bench
71. **Show HN: Tips for getting great Text2Cypher outputs from LLMs for Graph RAG** | signal 20.2 | tags: context engineering | https://blog.kuzudb.com/post/improving-text2cypher-for-graphrag-via-schema-pruning/
72. **Context Engineering – LLM Memory and Retrieval for AI Agents** | signal 20.1 | tags: context engineering, retrieval | https://weaviate.io/blog/context-engineering
73. **Show HN: Stop re-explaining context to every LLM (Git-based context engineering)** | signal 20.1 | tags: context engineering | https://github.com/jerpint/context-llemur
74. **At what level of deep context engineering does AI output become human-crafted?** | signal 20.1 | tags: context engineering | https://news.ycombinator.com/item?id=47330309
75. **Show HN: OneManCompany The first AI company with real corporate org structure** | signal 20.1 | tags: claude code, langgraph | https://one-man-company.com/
76. **Show HN: MarkdownLM – Stop being the human middleware for your AI agent** | signal 20.1 | tags: claude code, retrieval | https://news.ycombinator.com/item?id=47124474
77. **Why Your RAG Costs $2,400/Month (and How We Cut It by 73%)** | signal 20.1 | tags: rag pipeline, retrieval, vector | https://news.ycombinator.com/item?id=46234309
78. **Show HN: VittoriaDB – Zero-config embedded vector DB with HNSW and ACID storage** | signal 20.1 | tags: rag pipeline, benchmark, vector | https://github.com/antonellof/VittoriaDB
79. **Show HN: I built a circuit breaker that predicts AI failures** | signal 20.1 | tags: rag pipeline, langchain, vector | https://github.com/CULPRITCHAOS/Interlock
80. **Show HN: Create your own finetuned AI model using Google Sheets** | signal 19.9 | tags: none | https://promptrepo.com/finetune/
81. **Show HN: Plexe – ML Models from a Prompt** | signal 19.5 | tags: none | https://github.com/plexe-ai/plexe
82. **Launch HN: JSX Tool (YC F25) – A Browser Dev-Panel IDE for React** | signal 18.6 | tags: none | https://news.ycombinator.com/item?id=45903161
83. **Show HN: I built "AI Wattpad" to eval LLMs on fiction** | signal 18.0 | tags: benchmark | https://narrator.sh/llm-leaderboard
84. **Show HN: MCP-C – cloud platform for running MCP agents and apps** | signal 17.9 | tags: mcp agent | https://docs.mcp-agent.com/get-started/cloud
85. **Show HN: CocoIndex – Open-Source Data framework for AI, built for data freshness** | signal 17.9 | tags: rag pipeline, vector | https://github.com/cocoindex-io/cocoindex
86. **Show HN: PolyMCP – A framework for building and orchestrating MCP agents** | signal 17.6 | tags: mcp agent | https://news.ycombinator.com/item?id=47017912
87. **Show HN: Jay - Fully programmable, fully hosted AI voice agents** | signal 17.6 | tags: rag pipeline, function calling | https://www.jay.so/
88. **Alternative of MCP with AI RAG Agentic Framework** | signal 17.3 | tags: mcp agent | https://news.ycombinator.com/item?id=43603324
89. **Show HN: PolyClaw – Autonomous Docker-First MCP Agent for PolyMCP** | signal 17.1 | tags: mcp agent | https://news.ycombinator.com/item?id=47047299
90. **Show HN: PolyClaw – An Autonomous Docker-First MCP Agent for PolyMCP** | signal 17.1 | tags: mcp agent | https://news.ycombinator.com/item?id=47036828
91. **Show HN: Easily generate text and compute probabilities for any Hugging Face LLM** | signal 17.1 | tags: evaluation harness | https://github.com/RichardKelley/hflm
92. **Ask HN: What tools are you using for AI evals? Everything feels half-baked** | signal 16.9 | tags: evals, benchmark | https://news.ycombinator.com/item?id=44194187
93. **I build my LLM a Brain** | signal 16.7 | tags: context engineering | https://news.ycombinator.com/item?id=47928151
94. **Save 70-90% in tokens per session** | signal 16.6 | tags: coding agent, evals | https://news.ycombinator.com/item?id=47395507
95. **Show HN: Modulus – Cross-repository knowledge orchestration for coding agents** | signal 16.6 | tags: coding agent | https://modulus.so
96. **Show HN: RunAgent; Multi-Framework Agent Deployment and Rust,Go,JS SDKs(+others)** | signal 16.5 | tags: langchain, langgraph | https://github.com/runagent-dev/runagent
97. **Show HN: A local merge queue for parallel Claude Code agents** | signal 16.5 | tags: claude code | https://github.com/funador/claude-code-merge-queue
98. **Ask HN: How are you LLM-coding in an established code base?** | signal 16.5 | tags: none | https://news.ycombinator.com/item?id=46292682
99. **Best AI Coding Agents – Gosu Evals** | signal 16.1 | tags: coding agent, evals | https://gosuevals.com/agents.html
100. **Evals Skills for Coding Agents** | signal 16.1 | tags: coding agent, evals | https://hamel.dev/blog/posts/evals-skills/
101. **Show HN: PokemonGym – 387 milestones designed to test agents and LLMs** | signal 16.1 | tags: tool use, benchmark | https://twitter.com/xdotli/status/1908373420032795083
102. **Show HN: Cockpit for you Claude Code agents in Rust** | signal 16.1 | tags: claude code | https://episko.dev/
103. **Context engineering is just software engineering for LLMs** | signal 15.9 | tags: context engineering | https://www.inngest.com/blog/context-engineering-is-software-engineering-for-llms
104. **Show HN: Daf·thunk – open-source Editor for Prototyping Workflows on Cloudflare** | signal 15.9 | tags: cursor rules | https://www.dafthunk.com/
105. **Launch HN: Openlayer (YC S21) – Testing and Evaluation for AI** | signal 15.9 | tags: none | https://news.ycombinator.com/item?id=38532593
106. **Show HN: Foolery – a web UI for orchestrating Claude Code agents on top of Beads** | signal 15.8 | tags: claude code | https://github.com/acartine/foolery
107. **Context Engineering for the LLM OS: User vs. Kernel Context** | signal 15.7 | tags: context engineering | https://www.letta.com/blog/guide-to-context-engineering
108. **Show HN: Crew – Let Claude Code agents talk to each other** | signal 15.6 | tags: claude code | https://github.com/0xmmo/crew
109. **Ask HN: Claude Code–style agent, but Aider-like and model-agnostic?** | signal 15.6 | tags: claude code | https://news.ycombinator.com/item?id=44693354
110. **Show HN: Modulus – Run multiple coding agents with shared project memory** | signal 15.6 | tags: coding agent | https://modulus.so
111. **Show HN: Hiver – Chrome DevTools for Agents** | signal 15.5 | tags: claude code | https://hiver.sh
112. **DeepSWE – Best Benchmark for Evaluating AI Coding Agents?** | signal 15.3 | tags: coding agent, benchmark | https://www.i-programmer.info/news/105-artificial-intelligence/19016-deepswe-best-benchmark-for-evaluating-ai-coding-agents.html
113. **Folks who work for large tech companies: How are you using Cursor?** | signal 15.3 | tags: cursor rules | https://news.ycombinator.com/item?id=43450576
114. **Show HN: PlanWiki – Open-source platform for product teams and agents to execute** | signal 15.3 | tags: claude code | https://github.com/planwiki/planwiki-app
115. **Show HN: Kote – Capture and reuse engineering context from AI chats and Git** | signal 15.2 | tags: claude code | https://github.com/pedroaugusto04/Kote
116. **Claude Code Open Source?** | signal 15.2 | tags: claude code | https://news.ycombinator.com/item?id=47285571
117. **Show HN: Oc-mnemoria – Persistent memory for AI coding agents** | signal 15.2 | tags: coding agent | https://github.com/one-bit/oc-mnemoria
118. **Show HN: Voicetest – open-source test harness for voice AI agents** | signal 15.2 | tags: claude code | https://news.ycombinator.com/item?id=47048811
119. **Show HN: Claude Code Agent Farm** | signal 15.2 | tags: claude code | https://github.com/Dicklesworthstone/claude_code_agent_farm
120. **Are AI coding tools fundamentally changing Agile/team software development?** | signal 15.2 | tags: claude code | https://news.ycombinator.com/item?id=45584707
121. **Show HN: Airut – Sandboxed Claude Code sessions over email** | signal 15.1 | tags: claude code | https://github.com/airutorg/airut
122. **SpecTree: Composable Context Engineering for LLMs** | signal 15.1 | tags: context engineering | https://www.fuzzycomputer.com/posts/spectree
123. **Introductory field guide to Context Engineering for LLM users** | signal 15.1 | tags: context engineering | https://andybromberg.com/field-guide-context-engineering
124. **Context Engineering in an LLM Harness** | signal 15.1 | tags: context engineering | https://udnes.dev/posts/context-engineering-harness-part-1-ontology/
125. **Agentic Context Engineering: Evolving Contexts for Self-Improving LLMs** | signal 15.1 | tags: context engineering | https://arxiv.org/abs/2510.04618
126. **Context Engineering for Agents: A Practical Guide** | signal 15.1 | tags: context engineering | https://blog.malt.engineering/dont-take-this-out-of-context-feeding-your-llm-exactly-what-it-needs-0db8a86d2151
127. **Show HN: A visual AI interface to understand topics/books/papers with LLMs** | signal 15.1 | tags: context engineering | https://www.kerns.ai/
128. **DeepSWE – Best Benchmark for Evaluating AI Coding Agents?** | signal 15.1 | tags: coding agent, benchmark | https://www.i-programmer.info/professional-programmer/103-i-programmer/18759-why-software-engineering-will-never-die-revisited-in-the-age-of-spec-driven-development.html
129. **The Kotlin Benchmark for AI Coding Agents** | signal 15.1 | tags: coding agent, benchmark | https://blog.jetbrains.com/kotlin/2026/07/introducing-the-kotlin-benchmark-evaluate-ai-coding-agents-on-real-world-kotlin-tasks/
130. **Write a prompt once, sync it to Cursor, Claude Code and VS Code automatically** | signal 15.1 | tags: claude code | https://news.ycombinator.com/item?id=47849308
131. **Show HN: Open-source desktop agent that uses a local folder as its memory** | signal 15.1 | tags: claude code | https://github.com/zqiren/Orbital
132. **Show HN: SHTMLs – HTML pastebin where the AI uploads its own output** | signal 15.1 | tags: claude code | https://news.ycombinator.com/item?id=47426450
133. **Show HN: Stop manually syncing rules between Claude, Cursor, and Codex** | signal 15.1 | tags: claude code | https://github.com/nanxiaobei/ai-global
134. **Show HN : Pilot – System to improve dramatically your AI coding** | signal 15.1 | tags: claude code | https://github.com/clementrog/pilot
135. **Show HN: MemoryGate – Open-source persistent memory for AI agents via MCP** | signal 15.1 | tags: rag pipeline, vector | https://www.memorygate.ai
136. **Show HN: Polyfire – Javascript SDK to build AI apps without a backend** | signal 14.6 | tags: langchain, vector | https://github.com/polyfire-ai/polyfire-js
137. **Garvata: Observability and Debugging for AI Agent Stack** | signal 14.3 | tags: retrieval, vector | https://news.ycombinator.com/item?id=42293942
138. **PA bench: Evaluating web agents on real world personal assistant workflows** | signal 13.7 | tags: benchmark | https://vibrantlabs.com/blog/pa-bench
139. **Show HN: Zipy.ai – Live web debugging with error monitoring and session replay** | signal 13.7 | tags: none | https://www.zipy.ai/
140. **Show HN: Verdic Guard – Deterministic guardrails to prevent LLM hallucinations** | signal 13.3 | tags: prompt engineering | https://news.ycombinator.com/item?id=46602822
141. **Show HN: Unify Browser – WebKit Browser Built with SwiftUI and MLX** | signal 13.1 | tags: prompt engineering | https://apps.apple.com/us/app/unify-ai-browser/id6478436147?mt=12
142. **Show HN: I Built an AI-Powered Pull Request Review Tool** | signal 13.1 | tags: code review workflow | https://github.com/HighGarden-Studio/HighReview
143. **Outworked – An Open Source Office UI for Claude Code Agents** | signal 13.0 | tags: claude code | https://github.com/outworked/outworked
144. **Launch HN: Coasty (YC S26) – An API for computer-use agents** | signal 12.4 | tags: none | https://coasty.ai/docs
145. **Launch HN: Patched (YC S24) – AI workflows for post-code tasks** | signal 12.3 | tags: none | https://news.ycombinator.com/item?id=42009089
146. **Show HN: A police department for your Claude Code agents** | signal 12.2 | tags: claude code | https://github.com/varmabudharaju/agent-pd/blob/master/README.md
147. **A review of OpenAI o1 and how we evaluate coding agents** | signal 12.1 | tags: coding agent | https://www.cognition.ai/blog/evaluating-coding-agents
148. **Show HN: Fast-agent – Compose MCP enabled Agents and Workflows in minutes** | signal 12.1 | tags: retrieval | https://github.com/evalstate/fast-agent
149. **Show HN: AI-powered web service combining FastAPI, Pydantic-AI, and MCP servers** | signal 12.1 | tags: none | https://github.com/Aherontas/Pycon_Greece_2025_Presentation_Agents
150. **Show HN: ContextVault – Shared memory layer for your AI and your team** | signal 11.8 | tags: vector | https://www.contextvault.dev/
151. **Show HN: Eval based agent builder (pls roast us)** | signal 11.2 | tags: langchain, evals | https://github.com/seer-engg/seer
152. **Bad MCP design costs your agent 5x more tokens** | signal 11.1 | tags: benchmark | https://news.ycombinator.com/item?id=48407391
153. **Show HN: Real-time visualization of Claude Code agent orchestration** | signal 11.1 | tags: claude code | https://github.com/patoles/agent-flow
154. **Show HN: OpenJet – An offline agent harness for memory-constrained edge hardware** | signal 11.1 | tags: evals | https://github.com/L-Forster/open-jet
155. **Show HN: A/B Test Your LLM Prompts in Production** | signal 11.1 | tags: evals | https://switchport.ai/
156. **Show HN: Krira Augment – Production-ready RAG in minutes** | signal 11.1 | tags: rag pipeline | https://www.kriralabs.com/waitlist
157. **Show HN: A JSON API for YouTube Transcript with MCP Support** | signal 11.1 | tags: rag pipeline | https://transcriptapi.com/
158. **Show HN: AI Interoperability to the Max – The Intelligence Hub** | signal 11.1 | tags: rag pipeline | https://theintelligencehub.azurewebsites.net/
159. **Ask HN: Anyone solved hallucination or semantic drift in RAG?** | signal 11.1 | tags: rag pipeline | https://news.ycombinator.com/item?id=44746089
160. **My Claude Code Agent for Writing Prompts** | signal 10.8 | tags: claude code | https://olshansky.info/posts/2025-09-29-prompt-writer-agent
161. **Ferretlog: Git log for your Claude Code agent runs** | signal 10.7 | tags: claude code | https://github.com/eitanlebras/ferretlog
162. **Curie – ship Claude Code agents to Kubernetes with Git push** | signal 10.6 | tags: claude code | https://github.com/curie-eng/curie
163. **Replaced Clay.com with Claude Code Agent** | signal 10.6 | tags: claude code | https://github.com/chaitanyya/sales
164. **I built IDE-layer policy enforcement for Claude Code/Cursor agents** | signal 10.4 | tags: claude code | https://www.oculisecurity.com/
165. **Connect multiple Claude Code agents into one collaborative team** | signal 10.4 | tags: claude code | https://openagents.org/showcase
166. **15 AI Coding Agents evaluated with the same prompt** | signal 10.3 | tags: coding agent | https://github.com/The-Focus-AI/june-2025-coding-agent-report
167. **Why Claude Code's Agent Loop Is over 1,400 Lines** | signal 10.3 | tags: claude code | https://internals.laxmena.com/p/why-claude-codes-agent-loop-is-over
168. **Show HN: I run a full software company solo with Claude Code agents** | signal 10.3 | tags: claude code | https://theonemancompany.com/
169. **Show HN: I built an open-source Rust/TS AI agent runtime with a Next.js-style DX** | signal 10.2 | tags: langchain | https://docs.trysoma.ai
170. **New Inference Server for DGX Spark: large model C4:55-90 tok/s no spec decode** | signal 10.2 | tags: benchmark | https://news.ycombinator.com/item?id=49014048
171. **How to evaluate models for production coding agents** | signal 10.2 | tags: coding agent | https://blaxel.ai/blog/llm-coding-benchmarks
172. **FlyCrys – Native Linux GUI for Claude Code Agents (Rust and GTK4)** | signal 10.2 | tags: claude code | https://github.com/SergKam/FlyCrys
173. **20 Claude Code agents, one terminal: a tmux + AppleScript setup** | signal 10.2 | tags: claude code | https://pkarnal.com/blog/parallel-ai-agents
174. **Securely run Claude Code agents in Docker** | signal 10.2 | tags: claude code | https://edspencer.net/2026/2/4/run-claude-code-agents-docker-herdctl
175. **Show HN: I built a context-engineering CLI/MCP tool** | signal 10.1 | tags: retrieval | https://github.com/jerpint/context-llemur
176. **When your coding agent doesn't listen: evaluating a 241-turn Claude session** | signal 10.1 | tags: coding agent | https://www.kurrent.io/blog/when-your-coding-agent-doesnt-listen/
177. **Engine-Bench: Evaluating Coding Agents on Writing Game Engine Code** | signal 10.1 | tags: coding agent | https://github.com/JoshuaPurtell/engine-bench
178. **Evaluating Coding Agents with Terminal-Bench 2.0** | signal 10.1 | tags: coding agent | https://snorkel.ai/blog/evaluating-coding-agent-capabilities-with-terminal-bench-snorkels-role-in-building-the-next-generation-benchmark/
179. **Show HN: A Framework for Evaluating Coding Agents on Sequential SWE** | signal 10.1 | tags: coding agent | https://arxiv.org/abs/2604.03035
180. **ReactBench – evaluation for coding agents on realistic React work** | signal 10.1 | tags: coding agent | https://www.reactbench.com/
181. **No one is evaluating AI coding agents in the way they are used** | signal 10.1 | tags: coding agent | https://marginlab.ai/blog/the-problem-with-coding-benchmarks/
182. **Show HN: Apitoll Payment InfrastructureforAIagents75 Live APIs,USDCmicropayments** | signal 10.1 | tags: langchain | https://github.com/TasnidChain/apitoll-demo
183. **Show HN: Cortex Click – LLM-Driven Developer Marketing Platform** | signal 10.1 | tags: retrieval | https://news.ycombinator.com/item?id=41583460
184. **Cursor Rules for Writing Temporal Workflows with TypeScript** | signal 10.1 | tags: cursor rules | https://stevekinney.com/writing/cursor-rules-temporal-typescript
185. **KernelEvolve: Agentic kernel coding for heterogeneous AI accelerators (Meta)** | signal 10.1 | tags: benchmark | https://news.ycombinator.com/item?id=46442841
186. **Open Source LLMOps Stack** | signal 9.6 | tags: none | https://oss-llmops-stack.com
187. **Show HN: VectorGuard-Nano – Free secure messaging for AI agents** | signal 9.4 | tags: vector | https://github.com/Active-IQ/VectorGuard-Nano
188. **I Got Pwned by a Malicious AI Plugin: A Technical Breakdown** | signal 9.3 | tags: vector | https://news.ycombinator.com/item?id=47109114
189. **Show HN: AI-gent Workflows – locally reasoning AI Agents** | signal 9.3 | tags: vector | https://ai-gents.work
190. **Show HN: Boucle – A self-dogfooding autonomous AI agent framework in Rus** | signal 9.2 | tags: vector | https://github.com/Bande-a-Bonnot/Boucle-framework
191. **Dispelling Misconceptions and Unveiling the Truth about GOT and OT in General** | signal 9.1 | tags: vector | https://news.ycombinator.com/item?id=36643393
192. **Show HN: Local LLM Notepad – run a GPT-style model from a USB stick** | signal 8.8 | tags: none | https://github.com/runzhouye/Local_LLM_Notepad
193. **Ask HN: What's your 2025 code review workflow? GitHub UI feels ancient** | signal 8.6 | tags: code review workflow | https://news.ycombinator.com/item?id=44583146
194. **My Current AI Code Review Workflow** | signal 8.3 | tags: code review workflow | https://guissmo.com/blog/my-current-ai-code-review-workflow/
195. **SHOW HN: A usage circuit breaker for Cloudflare Workers** | signal 8.2 | tags: none | https://news.ycombinator.com/item?id=47322794
196. **Show HN: AI-Friendly Toolchain – Dev Tools for Working with LLMs** | signal 8.1 | tags: prompt engineering | https://github.com/trknhr/awesome-ai-friendly-toolchain
197. **Show HN: Hopsule – Persistent memory and decision layer for AI development** | signal 7.5 | tags: none | https://news.ycombinator.com/item?id=47415402
198. **Show HN: An open-source Operator that can use computers** | signal 7.1 | tags: none | https://github.com/aditya-nadkarni/spongecake
199. **Show HN: Upsonic: An AI agent framework with client-server architecture** | signal 6.5 | tags: none | https://github.com/Upsonic/Upsonic
200. **I'm starting to feel tired of AI features that solve problems I don't have** | signal 6.5 | tags: none | https://news.ycombinator.com/item?id=45734499
201. **Does anyone use MCP servers in their dev workflow?** | signal 6.3 | tags: none | https://news.ycombinator.com/item?id=43258552
202. **Ask HN: Is anyone HOPEFUL about our robot overlords?** | signal 6.2 | tags: none | https://news.ycombinator.com/item?id=34096780
203. **Building a production-ready RAG pipeline and eval platform** | signal 6.1 | tags: rag pipeline | https://docs.vectorize.io/core-concepts/vectorize-architecture
204. **Show HN: Ductwork – A Go platform for running AI agents on autopilot** | signal 6.0 | tags: none | https://github.com/dneil5648/ductwork
205. **Show HN: Velvet – Data platform with an AI SQL editor** | signal 6.0 | tags: none | https://www.usevelvet.com/
206. **Built a content system that 6x'd traffic. Turning it into product. Want to test?** | signal 6.0 | tags: none | https://news.ycombinator.com/item?id=46337608
207. **Seeking feedback: Integrated product discovery workflow tool** | signal 5.9 | tags: none | https://news.ycombinator.com/item?id=45830436
208. **Show HN: OQP – A verification protocol for AI agents** | signal 5.8 | tags: none | https://github.com/OranproAi/open-qa-protocol
209. **Show HN: Capcat – CLI/TUI to Archive Articles as Markdown and HTML (FOSS)** | signal 5.7 | tags: none | https://capcat.org/
210. **Show HN: Rucat – Cat for Prompt Engineers** | signal 5.7 | tags: none | https://github.com/brianredbeard/rucat
211. **I have a project with ~200k LoC, written with AI codegen. AMA** | signal 5.6 | tags: none | https://news.ycombinator.com/item?id=45351057
212. **If the differentiation is domain and GTM?** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=47329075
213. **Show HN: Agent File (.af) – A standard file format for serializing AI agents** | signal 5.5 | tags: none | https://github.com/letta-ai/agent-file
214. **Windmemory** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=42751099
215. **Show HN: Like grep but for natural questions. Mixtral 8x7B – 28 tok/s on 8GB GPU** | signal 5.5 | tags: none | https://github.com/moritztng/fltr
216. **Prompt to make Claude more autonomous in web dev** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=47379947
217. **Beginner's Guide to MCP (Model Context Protocol)** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=43758713
218. **Show HN: An AI interviewer that probes candidates (and costs $0.99/interview)** | signal 5.5 | tags: none | https://interviewflowai.com/
219. **Stopped Using Cursor, for Now** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=44985565
220. **Show HN: Using classic dev books to guide AI agents** | signal 5.5 | tags: none | https://news.ycombinator.com/item?id=47098555
221. **Show HN: Acceptify – AI personas that run user acceptance tests on your product** | signal 5.5 | tags: none | https://acceptify.ai/
222. **AI Power Internal Tools** | signal 5.4 | tags: none | https://news.ycombinator.com/item?id=44494999
223. **Show HN: Prismy – GitHub-Native, AI Localization for Dev and Product Teams** | signal 5.4 | tags: none | https://www.prismy.io
224. **Show HN: No-Code, Private AI Agents – Build and Run Locally** | signal 5.4 | tags: none | https://browseragent.dev
225. **Show HN: AI agent that works autonomously while I'm offline** | signal 5.3 | tags: none | https://hire-your-ai-guide.vercel.app
226. **Show HN: DiffDeck, a PR review tool with file context and code navigation** | signal 5.3 | tags: none | https://diffdeck.dev/login
227. **Show HN: Inference API that adapts to your SLA and quality constraints** | signal 5.3 | tags: none | https://models.exosphere.host/
228. **Show HN: Why delegation beats memory in AI Agents** | signal 5.2 | tags: none | https://www.getseer.dev/blogs/lessons-dec-2025
229. **Show HN: We built an AI-agent with a state machine instead of a giant prompt** | signal 5.2 | tags: none | https://nomos.dowhile.dev/
230. **Show HN: Owl and MCP Integration – Plug-and-play agents with external tools** | signal 5.2 | tags: none | https://www.camel-ai.org/blogs/owl-mcp-toolkit-practice
231. **Show HN: Notte – Full-stack web-agent framework (open-source)** | signal 5.2 | tags: none | https://github.com/nottelabs/notte
232. **Show HN: Freeze the Model, Train the Harness** | signal 5.2 | tags: none | https://github.com/workofart/harness-training
233. **Show HN: CriteriaBot – A Universal Customizable Classifier** | signal 5.2 | tags: none | https://criteriabot.io/
234. **Show HN: AI Dev Assistant Framework – Add structure, rules and memory to LLM** | signal 5.2 | tags: none | https://github.com/Fr-e-d/ai-dev-assistant-framework
235. **Show HN: Everdone CodeReview – AI code reviews as a trackable workflow** | signal 5.2 | tags: none | https://everdone.ai/
236. **Show HN: AI Code Review CLI** | signal 5.2 | tags: none | https://github.com/kodustech/cli
237. **Show HN: GPT-reviewer – Simple AI code reviewer for GH Actions** | signal 5.2 | tags: none | https://github.com/vayqerlukashakkarainen/gpt-reviewer
238. **AI-powered Git CLI that generates commit messages automatically** | signal 5.2 | tags: none | https://news.ycombinator.com/item?id=47035076
239. **Show HN: Freeplay – Testing and Evaluation for LLM-powered features** | signal 5.2 | tags: none | https://freeplay.ai/
240. **Show HN: open source framework for building nanoservices** | signal 5.2 | tags: none | https://news.ycombinator.com/item?id=43164465
241. **Show HN: TheFoundry – Easy bootstrapping framework for MultiAgent Systems** | signal 5.1 | tags: none | https://github.com/aavilagallego/TheFoundry
242. **Show HN: LedgerMind – true zero-touch autonomous memory for AI agents** | signal 5.1 | tags: none | https://github.com/sl4m3/ledgermind
243. **Show HN: Mdchat – Markdown-first terminal / CLI tool for LLM collaboration** | signal 5.1 | tags: none | https://www.npmjs.com/package/mdchat
244. **Show HN: WorldBuild Bench repo: testing LLM world coherence with 3D games** | signal 5.1 | tags: none | https://github.com/sebnado/worldbuild-bench
245. **Show HN: CreateMVP.app – First open-source tool to generate MVP specs for LLMs** | signal 5.1 | tags: none | https://createmvps.app/
246. **Show HN: Visual Editor for Cursor** | signal 5.1 | tags: none | https://shuffle.dev/cursor
247. **Show HN: I made an open source Idea to App WebApp** | signal 5.1 | tags: none | https://github.com/rohitg00/CreateMVP
248. **Show HN: Framework to structure LLM dev workflows with Markdown-based protocol** | signal 5.1 | tags: none | https://github.com/Fr-e-d/ai-dev-assistant-framework
249. **My tiny workflow for an AI code review assist** | signal 5.1 | tags: none | https://news.ycombinator.com/item?id=45959846
250. **Show HN: AI code review now available on Azure DevOps** | signal 5.1 | tags: none | https://kodus.io/en/
251. **Show HN: CREV – A Go-based CLI tool for AI code reviews and codebase exports** | signal 5.1 | tags: none | https://news.ycombinator.com/item?id=41757003
252. **Show HN: I released a OS remote agent callable from mobile** | signal 5.0 | tags: none | https://github.com/epavanello/fixodev
253. **Show HN: Arkain – AI-powered Cloud IDE for building real apps from your words** | signal 5.0 | tags: none | https://arkn.ai/qH22w
254. **Show HN: Chaos engineering for LLMs – Making models cross-examine each other** | signal 5.0 | tags: none | https://www.usecouncil.app/
255. **Show HN: Aidevshield NPM audit for AI coding tool workflows** | signal 5.0 | tags: none | https://github.com/aidevshield/aidevshield
256. **Show HN: AI Resource Manager** | signal 5.0 | tags: none | https://github.com/jomadu/ai-resource-manager
257. **Show HN: Deff – Review AI-generated code changes** | signal 5.0 | tags: none | https://github.com/flamestro/deff
258. **Show HN: Shell script for AI-powered code reviews using local LLMs** | signal 5.0 | tags: none | https://gist.github.com/alwin-augustin-dev/c1caaa30361f7ee320fb9cb957b3b0e9
259. **Show HN: Monitor, audit & alert on AI agent actions and interactions** | signal 5.0 | tags: none | https://pingpulsehq.com
260. **Show HN: Autonomous outbound research and outreach drafts** | signal 5.0 | tags: none | https://www.prospecter.io
261. **Show HN: KitchenAI Open Source LLMops development kit. Notebook to server** | signal 5.0 | tags: none | https://github.com/epuerta9/kitchenai
262. **Ask HN: CI/CD and Hosting for GPU-Based ML Demos** | signal 5.0 | tags: none | https://news.ycombinator.com/item?id=39273121
263. **Show HN: TrustVector – Trust evaluations for AI models, agents, & MCP** | signal 4.3 | tags: benchmark, vector | https://github.com/guard0-ai/TrustVector
264. **Show HN: I scraped 200M Shopify products to build a search engine** | signal 2.2 | tags: none | https://www.searchagora.com/#
265. **Why Vertical AI Agents May Replace RPA in Complex Enterprise Workflows** | signal 1.1 | tags: none | https://news.ycombinator.com/item?id=44245754

## github (78)

1. **ratel-ai/ratel** | signal 35.2 | tags: context engineering, retrieval, vector | https://github.com/ratel-ai/ratel
2. **kris-hansen/comanda** | signal 33.6 | tags: agent workflow, claude code | https://github.com/kris-hansen/comanda
3. **NVIDIA/skills** | signal 30.6 | tags: claude code, coding agent | https://github.com/NVIDIA/skills
4. **sceneview/sceneview** | signal 26 | tags: cursor rules | https://github.com/sceneview/sceneview
5. **proliferate-ai/proliferate** | signal 24.9 | tags: claude code | https://github.com/proliferate-ai/proliferate
6. **JasonColapietro/suede-creator-skills** | signal 23.6 | tags: claude code, evals | https://github.com/JasonColapietro/suede-creator-skills
7. **baalimago/clai** | signal 22.4 | tags: context engineering | https://github.com/baalimago/clai
8. **langgenius/dify** | signal 22 | tags: rag pipeline | https://github.com/langgenius/dify
9. **mastra-ai/mastra** | signal 22 | tags: evals | https://github.com/mastra-ai/mastra
10. **artokun/comfyui-mcp** | signal 21 | tags: none | https://github.com/artokun/comfyui-mcp
11. **wellpreserved-sarcoptidae182/auto-harness** | signal 21.0 | tags: coding agent, evals, benchmark | https://github.com/wellpreserved-sarcoptidae182/auto-harness
12. **MrPeppersDev/agent-infrastructure-landscape** | signal 20.5 | tags: langchain, benchmark, vector | https://github.com/MrPeppersDev/agent-infrastructure-landscape
13. **Tonser974/autonome-framework** | signal 20.2 | tags: agent workflow, langchain | https://github.com/Tonser974/autonome-framework
14. **Uky0Yang/agent-rules-lint** | signal 20.1 | tags: cursor rules, coding agent | https://github.com/Uky0Yang/agent-rules-lint
15. **pharn-dev/pharn-oss** | signal 20.0 | tags: claude code, coding agent | https://github.com/pharn-dev/pharn-oss
16. **GoogleCloudPlatform/db-context-enrichment** | signal 19.6 | tags: context engineering | https://github.com/GoogleCloudPlatform/db-context-enrichment
17. **NeoLabHQ/context-engineering-kit** | signal 19.2 | tags: claude code | https://github.com/NeoLabHQ/context-engineering-kit
18. **outerlayerai/outerlayer** | signal 17.1 | tags: coding agent, evals | https://github.com/outerlayerai/outerlayer
19. **aaddii09/llm-eval-harness** | signal 17.1 | tags: evaluation harness, benchmark | https://github.com/aaddii09/llm-eval-harness
20. **linny006/agent-eval-harness** | signal 16.1 | tags: coding agent, benchmark | https://github.com/linny006/agent-eval-harness
21. **heygen-com/hyperframes** | signal 16 | tags: none | https://github.com/heygen-com/hyperframes
22. **dataelement/bisheng** | signal 16 | tags: none | https://github.com/dataelement/bisheng
23. **xorbitsai/xagent** | signal 16 | tags: none | https://github.com/xorbitsai/xagent
24. **bop-clocktower/canary** | signal 15.8 | tags: claude code | https://github.com/bop-clocktower/canary
25. **footprintjs/agentfootprint** | signal 15.5 | tags: context engineering | https://github.com/footprintjs/agentfootprint
26. **Tommieoxidative416/video-evaluator** | signal 15.1 | tags: coding agent, benchmark | https://github.com/Tommieoxidative416/video-evaluator
27. **Dikhun/Arctus.ai** | signal 15.1 | tags: agent workflow | https://github.com/Dikhun/Arctus.ai
28. **MValentimTech/grounded-context** | signal 15.0 | tags: context engineering | https://github.com/MValentimTech/grounded-context
29. **atliliw/langchainrust** | signal 14.7 | tags: langchain, langgraph, vector | https://github.com/atliliw/langchainrust
30. **Umarfarook1/rag-document-qa** | signal 14.1 | tags: benchmark, retrieval, vector | https://github.com/Umarfarook1/rag-document-qa
31. **eriknewton/sanctuary-framework** | signal 13.2 | tags: claude code | https://github.com/eriknewton/sanctuary-framework
32. **AtvikSecurity/domarinn** | signal 12.9 | tags: evaluation harness | https://github.com/AtvikSecurity/domarinn
33. **prime-radiant-inc/superpowers-evals** | signal 12.5 | tags: evals | https://github.com/prime-radiant-inc/superpowers-evals
34. **soppressata/OpenHarness** | signal 12.2 | tags: evaluation harness | https://github.com/soppressata/OpenHarness
35. **liminalarc/litmus-ai** | signal 12.1 | tags: evaluation harness | https://github.com/liminalarc/litmus-ai
36. **adityamhaske/agent-arena** | signal 12.0 | tags: evaluation harness | https://github.com/adityamhaske/agent-arena
37. **krzysztofdudek/Yggdrasil** | signal 11.8 | tags: coding agent | https://github.com/krzysztofdudek/Yggdrasil
38. **ainova-systems/operator-autopilot** | signal 11.1 | tags: coding agent | https://github.com/ainova-systems/operator-autopilot
39. **anvmn/agentic-delivery-evals** | signal 11.0 | tags: evals, benchmark | https://github.com/anvmn/agentic-delivery-evals
40. **Forest-Project-Lab/doctrine** | signal 10.8 | tags: coding agent | https://github.com/Forest-Project-Lab/doctrine
41. **stevenfackley/opencode-amplifier** | signal 10.2 | tags: coding agent | https://github.com/stevenfackley/opencode-amplifier
42. **totalwindupflightsystems/gitreins** | signal 10.2 | tags: coding agent | https://github.com/totalwindupflightsystems/gitreins
43. **baksohyeon/mogui-agent-harness** | signal 10.1 | tags: claude code | https://github.com/baksohyeon/mogui-agent-harness
44. **imbflool/cc-plugin-eval** | signal 10.1 | tags: claude code | https://github.com/imbflool/cc-plugin-eval
45. **genuschristellaoverhang846/agent-rules-kit** | signal 10.1 | tags: cursor rules | https://github.com/genuschristellaoverhang846/agent-rules-kit
46. **felixross66/claude-ai-coding-kit-2026** | signal 10.1 | tags: claude code | https://github.com/felixross66/claude-ai-coding-kit-2026
47. **mctang24/go-coding-agent** | signal 10.0 | tags: coding agent | https://github.com/mctang24/go-coding-agent
48. **imagin5786/ases-ai-scrum-system** | signal 10.0 | tags: claude code | https://github.com/imagin5786/ases-ai-scrum-system
49. **Thebaultsemirigid251/GlideGrail** | signal 10.0 | tags: coding agent | https://github.com/Thebaultsemirigid251/GlideGrail
50. **jkhines/ai-rules** | signal 10.0 | tags: claude code | https://github.com/jkhines/ai-rules
51. **rcrdk/agent-kit** | signal 10.0 | tags: cursor rules | https://github.com/rcrdk/agent-kit
52. **LeoBergmiller/rag-evaluation** | signal 10.0 | tags: benchmark, retrieval | https://github.com/LeoBergmiller/rag-evaluation
53. **maee-co/cc-autoship** | signal 10.0 | tags: claude code | https://github.com/maee-co/cc-autoship
54. **objectstack-ai/objectstack** | signal 8.9 | tags: none | https://github.com/objectstack-ai/objectstack
55. **contextforge-org/cpex** | signal 8.6 | tags: none | https://github.com/contextforge-org/cpex
56. **Aryansingh009/awesome-llm-knowledge-systems** | signal 8.0 | tags: prompt engineering | https://github.com/Aryansingh009/awesome-llm-knowledge-systems
57. **VPSDance/ai-proxy-rules** | signal 8.0 | tags: none | https://github.com/VPSDance/ai-proxy-rules
58. **InfyEdge/system-prompts-and-models-of-ai-tools-chinese** | signal 8.0 | tags: none | https://github.com/InfyEdge/system-prompts-and-models-of-ai-tools-chinese
59. **PradeepaRW/project-nova** | signal 6.8 | tags: none | https://github.com/PradeepaRW/project-nova
60. **vivekmidas/enterprise-llm-gateway** | signal 6.5 | tags: langgraph | https://github.com/vivekmidas/enterprise-llm-gateway
61. **Lvvphole/ai-account-prioritization** | signal 6.2 | tags: evals | https://github.com/Lvvphole/ai-account-prioritization
62. **builtbycyun/journeyman-agents** | signal 6.2 | tags: evals | https://github.com/builtbycyun/journeyman-agents
63. **Bande-a-Bonnot/Boucle-framework** | signal 6.0 | tags: none | https://github.com/Bande-a-Bonnot/Boucle-framework
64. **eliasfeitan-pixel/llm-eval-framework** | signal 6.0 | tags: rag pipeline | https://github.com/eliasfeitan-pixel/llm-eval-framework
65. **ChiaLungChuang/dam-agentic** | signal 5.2 | tags: langgraph | https://github.com/ChiaLungChuang/dam-agentic
66. **PurpleBlossomAI/instar** | signal 5.1 | tags: benchmark | https://github.com/PurpleBlossomAI/instar
67. **eusoro-stack/model-gauntlet** | signal 5.0 | tags: benchmark | https://github.com/eusoro-stack/model-gauntlet
68. **exha1078/agentic-workflow-orchestrator** | signal 5.0 | tags: langgraph | https://github.com/exha1078/agentic-workflow-orchestrator
69. **cupidnavaz/ai-agent-platform** | signal 5.0 | tags: retrieval | https://github.com/cupidnavaz/ai-agent-platform
70. **sxmimhd/knowledgeforge** | signal 5.0 | tags: retrieval | https://github.com/sxmimhd/knowledgeforge
71. **dhyansraj/mcp-mesh** | signal 4.7 | tags: none | https://github.com/dhyansraj/mcp-mesh
72. **asiraky/harnesst** | signal 4.2 | tags: none | https://github.com/asiraky/harnesst
73. **phnx-labs/agents-cli** | signal 3.3 | tags: none | https://github.com/phnx-labs/agents-cli
74. **GEMISIS/leviath** | signal 3.2 | tags: none | https://github.com/GEMISIS/leviath
75. **cerredz/Vidbyte-SDK** | signal 2.4 | tags: none | https://github.com/cerredz/Vidbyte-SDK
76. **h4vzz/awesome-ai-agent-skills** | signal 2.0 | tags: none | https://github.com/h4vzz/awesome-ai-agent-skills
77. **virastack/ai** | signal 1.1 | tags: none | https://github.com/virastack/ai
78. **jacksonanstee/agent-harness-JA** | signal 1.0 | tags: none | https://github.com/jacksonanstee/agent-harness-JA

## stackoverflow (34)

1. **Claude Code - Looking for guidance on where to start with coding and tools** | signal 13.6 | tags: claude code | https://stackoverflow.com/questions/79927051/claude-code-looking-for-guidance-on-where-to-start-with-coding-and-tools
2. **Best Approach to Evaluate a Graph RAG Pipeline Using Metrics?** | signal 11.4 | tags: rag pipeline, retrieval | https://stackoverflow.com/questions/78881336/best-approach-to-evaluate-a-graph-rag-pipeline-using-metrics
3. **How do you learn without AI?** | signal 9.5 | tags: none | https://stackoverflow.com/questions/79832798/how-do-you-learn-without-ai
4. **Globally catch exceptions in a WPF application?** | signal 9.4 | tags: none | https://stackoverflow.com/questions/793100/globally-catch-exceptions-in-a-wpf-application
5. **MCPToolConversionError: Failed to get tools from MCP server: 404** | signal 5.3 | tags: langchain | https://stackoverflow.com/questions/79705666/mcptoolconversionerror-failed-to-get-tools-from-mcp-server-404
6. **Langchain, Huggingface: Can&#39;t evaluate model with two different inputs** | signal 5.3 | tags: langchain | https://stackoverflow.com/questions/76137512/langchain-huggingface-cant-evaluate-model-with-two-different-inputs
7. **Input validation error: &#39;1.57&#39; is not of type &#39;number&#39; from langchain_mcp_adapter** | signal 5.1 | tags: langchain | https://stackoverflow.com/questions/79935672/input-validation-error-1-57-is-not-of-type-number-from-langchain-mcp-adapte
8. **Restrict responses from a language model (LLM) to only information available in a specific document** | signal 5.1 | tags: retrieval | https://stackoverflow.com/questions/78333793/restrict-responses-from-a-language-model-llm-to-only-information-available-in
9. **Is there a framework of many open-source code LLMs for generation?** | signal 5.0 | tags: benchmark | https://stackoverflow.com/questions/78287327/is-there-a-framework-of-many-open-source-code-llms-for-generation
10. **How can I debug an internal error in the .NET Runtime?** | signal 4.5 | tags: none | https://stackoverflow.com/questions/14238657/how-can-i-debug-an-internal-error-in-the-net-runtime
11. **Error - &quot;There is no script engine for file extension .vbs&quot; when using &quot;Git Bash Here&quot; in Windows 7** | signal 3.5 | tags: none | https://stackoverflow.com/questions/17757248/error-there-is-no-script-engine-for-file-extension-vbs-when-using-git-bash
12. **Transparent user session over several sites (single sign-on + single sign-off)** | signal 3.3 | tags: none | https://stackoverflow.com/questions/1043111/transparent-user-session-over-several-sites-single-sign-on-single-sign-off
13. **Current state and solutions for OpenGL over Windows Remote** | signal 2.5 | tags: none | https://stackoverflow.com/questions/51705471/current-state-and-solutions-for-opengl-over-windows-remote
14. **How can an AI assistant interact with Aspen plus through Python?** | signal 2.2 | tags: none | https://stackoverflow.com/questions/79904274/how-can-an-ai-assistant-interact-with-aspen-plus-through-python
15. **Store Django Log messages in a database?** | signal 2.1 | tags: none | https://stackoverflow.com/questions/11887816/store-django-log-messages-in-a-database
16. **How to validate the origin of a web service invokation** | signal 2.0 | tags: none | https://stackoverflow.com/questions/14023348/how-to-validate-the-origin-of-a-web-service-invokation
17. **Running Keras model for prediction in multiple threads** | signal 1.9 | tags: none | https://stackoverflow.com/questions/43136293/running-keras-model-for-prediction-in-multiple-threads
18. **How to disable context menu on right click/long touch in a kiosk mode of Chrome?** | signal 1.8 | tags: none | https://stackoverflow.com/questions/28222548/how-to-disable-context-menu-on-right-click-long-touch-in-a-kiosk-mode-of-chrome
19. **Mac OS X: Can one process render to another process&#39;s window?** | signal 1.8 | tags: none | https://stackoverflow.com/questions/583202/mac-os-x-can-one-process-render-to-another-processs-window
20. **Client-Side CommunicationException while Service works properly** | signal 1.8 | tags: none | https://stackoverflow.com/questions/15429934/client-side-communicationexception-while-service-works-properly
21. **How can I get a password containing a caret (^) passed unchanged as a parameter to a Windows batch file?** | signal 1.8 | tags: none | https://stackoverflow.com/questions/5254460/how-can-i-get-a-password-containing-a-caret-passed-unchanged-as-a-parameter
22. **can I turn off optimization, so in-scope variables from closures aren&#39;t &quot;optimized out&quot;** | signal 1.7 | tags: none | https://stackoverflow.com/questions/58861823/can-i-turn-off-optimization-so-in-scope-variables-from-closures-arent-optimiz
23. **Windows: avoid pushing full x86 context on stack** | signal 1.7 | tags: none | https://stackoverflow.com/questions/994555/windows-avoid-pushing-full-x86-context-on-stack
24. **Which StatsD client should I use for a java/grails project?** | signal 1.6 | tags: none | https://stackoverflow.com/questions/17243168/which-statsd-client-should-i-use-for-a-java-grails-project
25. **Why is every new AI IDE forcing a minimalist, &quot;chat-first&quot; UI on us?** | signal 1.5 | tags: none | https://stackoverflow.com/questions/79943845/why-is-every-new-ai-ide-forcing-a-minimalist-chat-first-ui-on-us
26. **Why are MCPs needed at all?** | signal 1.4 | tags: none | https://stackoverflow.com/questions/79866688/why-are-mcps-needed-at-all
27. **Set audio endpoint devices application specific (programmatically)** | signal 1.2 | tags: none | https://stackoverflow.com/questions/52973464/set-audio-endpoint-devices-application-specific-programmatically
28. **How can I set the RTS with ioctl() in a Mac plugin?** | signal 1.2 | tags: none | https://stackoverflow.com/questions/14693724/how-can-i-set-the-rts-with-ioctl-in-a-mac-plugin
29. **IntelliJ IDEA: Cannot run program &quot;C:\Program Files\nodejs\npx&quot;: CreateProcess error=193 when using MCP server** | signal 1.2 | tags: none | https://stackoverflow.com/questions/79722494/intellij-idea-cannot-run-program-c-program-files-nodejs-npx-createprocess-e
30. **Query OLAP Mondrian (MDX, XMLA) with a Python interface?** | signal 1.2 | tags: none | https://stackoverflow.com/questions/3793215/query-olap-mondrian-mdx-xmla-with-a-python-interface
31. **Windows Aero Rendering Bug** | signal 1.1 | tags: none | https://stackoverflow.com/questions/27450042/windows-aero-rendering-bug
32. **Harvesting the power of highly-parallel computers with python scientific code** | signal 1.1 | tags: none | https://stackoverflow.com/questions/18234484/harvesting-the-power-of-highly-parallel-computers-with-python-scientific-code
33. **searching good embedded &amp; hosting language pair** | signal 1.1 | tags: none | https://stackoverflow.com/questions/7843234/searching-good-embedded-hosting-language-pair
34. **ruamel_yaml.constructor.ConstructorError: could not determine a constructor for the tag &#39;tag:yaml.org,2002:python/tuple&#39; in &quot;&lt;unicode string&gt;&quot;** | signal 1.0 | tags: none | https://stackoverflow.com/questions/66609054/ruamel-yaml-constructor-constructorerror-could-not-determine-a-constructor-for

## lobsters (74)

1. **Google’s exponential path to climate-wrecking digital bloat** | signal 18.2 | tags: none | https://ketanjoshi.co/2026/07/01/googles-exponential-path-to-climate-wrecking-digital-bloat/
2. **What side projects have you enjoyed the most?** | signal 17.1 | tags: none | https://lobste.rs/s/mbk56v/what_side_projects_have_you_enjoyed_most
3. **What are you doing this weekend?** | signal 15.0 | tags: none | https://lobste.rs/s/won0b2/what_are_you_doing_this_weekend
4. **Why Rocq is better than Lean for program verification** | signal 13.8 | tags: none | https://joomy.korkutblech.com/posts/2026-07-28-why-rocq-is-better.html
5. **The feature in OxCaml that more languages should steal** | signal 13.7 | tags: none | https://theconsensus.dev/p/2026/06/27/the-feature-in-oxcaml-more-languages-should-steal.html
6. **The Productivity Mirage** | signal 13.1 | tags: none | https://frantic.im/mirage
7. **What are you doing this week?** | signal 13.1 | tags: none | https://lobste.rs/s/r7zjlm/what_are_you_doing_this_week
8. **Self-hosting email the hard way from your own routable IPv4 block up** | signal 12.9 | tags: none | https://anil.recoil.org/notes/recoil-self-hosting-2026
9. **Stripe Just Wants a Number** | signal 11.9 | tags: none | https://blog.exe.dev/billable-facts
10. **irken: A tiny hackable full-featured IRC client** | signal 11.7 | tags: none | https://codeberg.org/dlowe/irken
11. **OCaml 5.5.0 released** | signal 11.3 | tags: none | https://discuss.ocaml.org/t/ocaml-5-5-0-released/18265
12. **Meta Garbage Collection: Using OCaml's GC to GC Rust** | signal 10.4 | tags: none | https://soteria-tools.com/blog/meta-garbage-collection
13. **Two years of vector search at Notion: 10x scale, 1/10th cost** | signal 10.1 | tags: vector | https://www.notion.com/blog/two-years-of-vector-search-at-notion
14. **What does it mean to be a mathematician when AI does the math?** | signal 9.6 | tags: none | https://spectrum.ieee.org/ai-in-mathematics
15. **What are you doing this week?** | signal 9.6 | tags: none | https://lobste.rs/s/foxgva/what_are_you_doing_this_week
16. **Writing Toy Software Is A Joy (2025)** | signal 9.6 | tags: none | https://blog.jsbarretto.com/post/software-is-joy
17. **Faster Than Ninja** | signal 9.3 | tags: none | https://build2.org/blog/faster-than-ninja.xhtml
18. **Dependency Cultures - Richard Feldman (Software Should Work Conf 2026)** | signal 9.2 | tags: none | https://www.youtube.com/watch?v=E82ly38YEEQ
19. **Taking OCaml and Eio for a spin** | signal 8.9 | tags: none | https://mattjhall.co.uk/posts/taking-ocaml-eio-for-a-spin.html
20. **Painting with Gaussians** | signal 8.6 | tags: none | https://yogthos.net/posts/2026-08-03-splat-painter.html
21. **Quick & Easy Parser Combinators** | signal 8.6 | tags: none | https://www.cyan.sh/blog/posts/tutorial-quick-easy-parser-combinators.html
22. **The strain in your brain** | signal 8.5 | tags: none | https://anirudh.fi/strain
23. **why use F# for scripting and automation?** | signal 8.3 | tags: none | https://iev.ee/blog/why-use-fsharp/
24. **AI Learns the "Dark Art" of RF Chip Design** | signal 8.2 | tags: none | https://spectrum.ieee.org/ai-radio-chip-design
25. **"How to Think About AI": Cory Doctorow on Big Tech, Understanding AI, Labor Automation & More** | signal 8.2 | tags: none | https://www.youtube.com/watch?v=OBUzl_IaWIw
26. **Guarded methods in OCaml** | signal 8.1 | tags: none | https://xvw.lol/en/articles/oop-refl.html
27. **A line-by-line translation of the OCaml runtime from C to Rust** | signal 8.1 | tags: none | https://discuss.ocaml.org/t/a-line-by-line-translation-of-the-ocaml-runtime-from-c-to-rust/18247
28. **Inventing ELIZA - How the First Chatbot Shaped the Future of AI** | signal 8.0 | tags: none | https://mitpress.mit.edu/9780262052481/inventing-eliza/
29. **Why ML/OCaml are good for writing compilers (1998)** | signal 8.0 | tags: none | https://flint.cs.yale.edu/cs421/case-for-ml.html
30. **A Path Not Taken for OxCaml** | signal 8.0 | tags: none | https://joel.place/blog/path-not-taken/
31. **strace-ui, Bonsai_term, and the TUI renaissance** | signal 7.8 | tags: none | https://blog.janestreet.com/strace-ui-bonsai-term-and-the-tui-renaissance/
32. **Use Task Runners for Common Coding Tasks** | signal 7.8 | tags: none | https://hamvocke.com/blog/task-runners/
33. **Logic for Programmers** | signal 7.8 | tags: none | https://logicforprogrammers.com/
34. **N-body gravity simulation in O(N)** | signal 7.7 | tags: none | https://www.youtube.com/watch?v=FhMftauQZqU
35. **jj_tui: terminal user interface to jujutsu focused on speed and clarity** | signal 7.5 | tags: none | https://tangled.org/elidowling.com/jj_tui
36. **Introducing Incremental (2015)** | signal 7.4 | tags: none | https://blog.janestreet.com/introducing-incremental/
37. **You Could Have Come Up With Kimi Delta Attention** | signal 7.3 | tags: none | https://blog.doubleword.ai/you-could-have-come-up-with-kimi-delta-attention
38. **Syntax with Purpose in a Programming Language** | signal 7.3 | tags: none | https://www.youtube.com/watch?v=_HLZoeFREFo
39. **The following is a valid DOS COM executable** | signal 7.3 | tags: none | https://oldbytes.space/@gloriouscow/117045701876951834
40. **Chatbots vs Ozone** | signal 7.2 | tags: none | https://blog.dshr.org/2026/05/chatbots-vs-ozone.html
41. **Retries don't fix eventual consistency** | signal 7.2 | tags: none | https://var0.xyz/posts/retries-dont-fix-eventual-consistency.html
42. **Full flattening of nested data parallelism** | signal 7.2 | tags: none | https://futhark-lang.org/blog/2026-07-31-full-flattening.html
43. **Functional programming from first principles, part 1 – motivation** | signal 7.2 | tags: none | https://www.endoflineblog.com/functional-programming-from-first-principles-part-1-motivation
44. **Why we write our own C and C++ inference engines** | signal 7.1 | tags: none | https://localai.io/blog/why-we-write-our-own-engines/
45. **MAX models can now run on Apple silicon GPUs** | signal 7.0 | tags: none | https://forum.modular.com/t/max-models-can-now-run-on-apple-silicon-gpus/3283
46. **Xavier Leroy on programming, languages and formal verification** | signal 7.0 | tags: none | https://www.youtube.com/watch?v=9Cswiqrq6So
47. **bonsai: A library for building dynamic webapps, using Js_of_ocaml** | signal 6.8 | tags: none | https://github.com/janestreet/bonsai
48. **I wrote a music player (2022)** | signal 6.8 | tags: none | https://www.omarpolo.com/post/amused.html
49. **Languages as designed latent spaces** | signal 6.6 | tags: none | https://blog.jsbarretto.com/post/languages-as-latent-spaces
50. **What Rose Petals Teach Us about Induction** | signal 6.6 | tags: none | https://www.oranlooney.com/post/rose-petals/
51. **Flow’s OCaml to Rust Port** | signal 6.6 | tags: none | https://medium.com/flow-type/flows-ocaml-to-rust-port-78b95bcf49e9
52. **A novel computer Scrabble engine based on probability that performs at championship level (2021)** | signal 6.5 | tags: none | https://upcommons.upc.edu/server/api/core/bitstreams/1339ae43-3d65-4015-8e11-3689e5572b23/content
53. **Data race freedom in OxCaml** | signal 6.5 | tags: none | https://kcsrk.info/ocaml/oxcaml/x-ocaml/blogging/2026/05/07/data-race-freedom-in-oxcaml/
54. **Asana’s fascinating Tab shortcuts** | signal 6.5 | tags: none | https://unsung.aresluna.org/asanas-fascinating-tab-shortcuts/
55. **Tensor is the might** | signal 6.4 | tags: none | https://zserge.com/posts/tensor/
56. **Human-like Neural Nets by Catapulting** | signal 6.2 | tags: none | https://gwern.net/llm-catapult
57. **A global workspace in language models** | signal 6.2 | tags: none | https://www.anthropic.com/research/global-workspace
58. **Convolutional Neural Networks in APL (2019)** | signal 6.2 | tags: none | https://dl.acm.org/doi/epdf/10.1145/3315454.3329960
59. **Comparing Transformers and Hybrid Models at the Token Level** | signal 6.2 | tags: none | https://arxiv.org/pdf/2606.20936
60. **AI Agents Enable Adaptive Computer Worms** | signal 6.2 | tags: none | https://cleverhans.io/worm.html
61. **Language integrated LLMs as an OCaml function** | signal 6.2 | tags: none | https://anil.recoil.org/notes/language-integrated-llms
62. **Announcing Pyro Caml: The First Continuous Profiler for OCaml** | signal 6.2 | tags: none | https://semgrep.dev/blog/2026/announcing-pyro-caml-continuous-profiler-ocaml
63. **OCaml Infrastructure: How the opam-repository Works** | signal 6.2 | tags: none | https://ocaml.org/backstage/2025-11-05-how-the-opam-repository-works
64. **O(x)Caml in Space** | signal 6.2 | tags: none | https://gazagnaire.org/blog/2026-05-14-borealis.html
65. **Shrinking the OxCaml js_of_ocaml bundle: 285 MB to 4 MB** | signal 6.2 | tags: none | https://kcsrk.info/ocaml/oxcaml/modes/2026/05/10/shrinking-the-oxcaml-bundle/
66. **Categorization with NLP** | signal 6.1 | tags: none | https://softwaremaniacs.org/blog/2026/07/30/categorization-with-nlp/en/
67. **Debootstrapping without Archeology: Stacked Implementations in Camlboot** | signal 6.1 | tags: none | https://arxiv.org/abs/2202.09231
68. **A game made only with sine waves** | signal 6.1 | tags: none | https://www.youtube.com/watch?v=Qr3VsZYQy4s
69. **Project-Specific clangd Configuration with a Temporary Shell** | signal 6.1 | tags: none | https://felix-knorr.net/posts/2026-07-31-lsp-config.html
70. **Categorization with NLP** | signal 6.0 | tags: none | https://softwaremaniacs.org/blog/2026/07/30/categorization-with-nlp/
71. **Why Do Cognitive Scientists Hate LLMs? (2023)** | signal 6.0 | tags: none | https://minihf.com/posts/2023-10-16-hermes-lecture-3-why-do-cognitive-scientists-hate-llms/
72. **Matrix Orthogonalization Improves Memory in Recurrent Models** | signal 6.0 | tags: none | https://ayushtambde.com/blog/matrix-orthogonalization-improves-memory-in-recurrent-models/
73. **Robust AI Security and Alignment: A Sisyphean Endeavor?** | signal 6.0 | tags: none | https://ieeexplore.ieee.org/document/11475847/
74. **GPT2-BASIC: Portable Machine Intelligence in BASIC** | signal 6.0 | tags: none | https://github.com/tsotchke/gpt2-basic