Tools Bench.

Product launches and open-source repos with enough signal to earn a second look.

Last Brew Time: Jun 6, 2026, 7:25 AM PT

Insight

People have quietly shifted to selling the workflow around the model, not the model itself

Featured

GitHub219.3K

An agentic skills framework & software development methodology that works.

Market Signal

Why It Has Market Pull

Superpowers is the most ambitious open-source skill library for Claude Code right now — a Socratic-design → TDD → review pipeline packaged as ~15 composable, auto-triggering skills. Built by Jesse Vincent (the same obra behind Request Tracker, K-9 Mail, and Keyboardio), it ships through Anthropic's official plugin marketplace, was amplified by Simon Willison on HN, and has crossed 219k stars.

  • 219,449 GitHub stars and 19,530 forks since launch in October 2025; 5 tagged releases, last push June 2026.
  • Distributed via Anthropic's official plugin marketplace — install with `/plugin install superpowers@claude-plugins-official`.
  • Also packaged for Codex CLI, Factory Droid, Gemini CLI, OpenCode, Cursor, and Copilot CLI; not Claude-Code-locked.
  • Author Jesse Vincent has a decades-long open-source track record (RT, K-9 Mail, Keyboardio), which explains the unusual reception speed.
  • Termdock review measured a 14% token cut and quality gains on most coding tasks, with the caveat that some tasks regress.

feedbacks

What People Are Saying

  • "I can't recommend this post strongly enough. The way Jesse is using these tools is wildly more ambitious than most other people."HN comment

  • "The brainstorming skill is great. It helps flesh out a rough early idea."HN comment

  • "I personally don't like superpowers very much — Claude makes more mistakes when using superpowers than when not."HN comment

  • "Instead what we're getting is straight up voodoo nonsense."HN comment

  • "Cut tokens by 14% and boosted code quality — but not for every task."Termdock review

  • "Mandatory rather than suggested workflows — that's the whole point and also why it irritates some users."HN comment

  • "The systematic-debugging and using-git-worktrees skills are the ones I actually keep."Reddit r/ClaudeAI

GitHub33.0K

The Frontend Stack for Agents & Generative UI. React, Angular, Mobile, Slack, and more. Makers of the AG-UI Protocol

Market Signal

Why It Has Market Pull

CopilotKit is the breakout React layer for in-app AI agents, with the AG-UI Protocol now adopted by Google, Microsoft, AWS, Oracle, LangChain, and Mastra. A fresh $27M Series A in May 2026 and a 33k-star monorepo shipping multiple releases per week put it in the top tier of agent-frontend tooling worth real workflow testing.

  • 33,038 GitHub stars and 4,222 forks on the main repo; 14,073 stars on the sister AG-UI Protocol repo — both pushed today.
  • $27M Series A (May 2026) led by Glilot Capital, NFX, and SignalFire after a $7M previously-unannounced seed.
  • AG-UI Protocol shipped as an official integration in Microsoft's Agent Framework, with Google, AWS, Oracle, LangChain, and PydanticAI also onboard.
  • 5 releases pushed in the past 8 days (v1.59.1 → v1.59.5); founders report ~7M weekly agent-user interactions across deployments.
  • Named enterprise users include Deutsche Telekom, Docusign, and Cisco; main React surface is mature, with Angular and mobile in beta.

feedbacks

What People Are Saying

  • "The way Jesse is using these tools is wildly more ambitious than most other people."HN comment

  • "Self-host only with no proprietary inference endpoints — finally usable for strict-data orgs."HN comment

  • "Plug-and-play components like CopilotPortal and CopilotTextarea make it much easier to connect existing apps to agent backends."Product Hunt comment

  • "Saved us a ton of effort! Kudos."Product Hunt comment

  • "Using 'Copilot' is going to be a searchability and trademark nightmare given Microsoft's branding."HN comment

  • "React-only requirement wasn't clearly labeled upfront."HN comment

  • "De facto standard to embed AI copilots in modern React applications."Comparateur-IA review

GitHub26.4K

An Open Source implementation of Notebook LM with more flexibility and features

Market Signal

Why It Has Market Pull

Open Notebook is the canonical self-hosted NotebookLM clone with real practitioner momentum: 26.4k stars, weekly releases, and broad provider coverage (OpenAI, Anthropic, Gemini, Groq, DeepSeek, plus local Ollama and LM Studio). Strongest fit for privacy-conscious teams and Ollama-heavy hobbyists; output quality on defaults still lags Google's free NotebookLM in head-to-head user tests.

  • 26,453 GitHub stars and 3,027 forks; v1.9.0 shipped 2026-06-02 with 30+ tagged releases since launch.
  • Supports 18+ model providers including OpenAI, Anthropic, Gemini, Groq, Mistral, Ollama, LM Studio, Perplexity, DeepSeek, OpenRouter, Qwen.
  • Differentiates against NotebookLM with multi-speaker (1–4 voice) podcasts, custom transformations, REST API, and reasoning-model support (DeepSeek-R1, Qwen3).
  • Two XDA Developers reviews and KDnuggets coverage; sole maintainer (Luis Novo, São Paulo) responsive in discussions.
  • Honest weak point: head-to-head users report shorter, lower-quality podcast output on defaults vs Google's NotebookLM free tier.

feedbacks

What People Are Saying

  • "Output from NotebookLM: 926 words. Output from Open-notebook: 321 words of lesser quality. I cannot see myself using this as a sufficient substitute."GitHub discussion

  • "Hi developer, really great app. Thank you for making it open source."GitHub discussion

  • "Trying to make my own database of knowledge, but the lack of folder organization makes this not a tool I am able to use due to the number of files."GitHub discussion

  • "The privacy benefits are there and the workflow customization is plentiful compared to NotebookLM."XDA Developers article

  • "A fantastic self-hosted alternative to NotebookLM — rather than being restricted to Gemini, Open Notebook can be paired with most of the popular AI platforms."XDA Developers article

  • "Deploying an instance of Open Notebook can be a pain, as you'll have to jump through certain API hoops."XDA Developers article

  • "From creating the directory to accessing the interface, Open Notebook can be up and running in under two minutes."KDnuggets review

Product Hunt161

A 550B MoE frontier-intelligence open model built for long-running agents. It delivers 5x faster inference and lowers the cost of complex agentic tasks by up to 30% versus other open frontier models.Ultra excels at complex tasks like coding and deep research. Long-running agents spend their time planning, using tools, recovering from failures, and deciding what to do next.

Market Signal

Why It Has Market Pull

Nemotron 3 Ultra is NVIDIA's most consequential open-weight release of the year — a 550B-total / 55B-active hybrid Mamba-2 + MoE model with a 1M-token context, OpenMDW-1.1 commercial-friendly license, and day-0 deployment on HuggingFace, OpenRouter, NIM, DeepInfra, and SGLang. Artificial Analysis ranks it as the highest-scoring US-developed open-weight model today, and NVIDIA's own benchmarks show roughly 5x faster agentic inference on Blackwell-class hardware.

  • Announced at Computex 2026 and released June 4, 2026 on HuggingFace as `nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16` with NVFP4 and GenRM variants.
  • Artificial Analysis Intelligence Index of 48 — highest of any US-developed open-weight model (Kimi K2.6 leads globally at 54).
  • 550B total parameters / 55B active; hybrid Mamba-2 + MoE + attention layers with multi-token prediction for throughput.
  • Day-0 inference paths on OpenRouter, NVIDIA NIM, DeepInfra, and SGLang; an `unsloth/...-GGUF` community quant is already published.
  • NVIDIA reports up to 5x faster agentic inference and ~30% lower cost on SWE-bench and Terminal-Bench 2.0 vs prior open-weight peers.

feedbacks

What People Are Saying

  • "Completes tasks at a much faster pace than peers due to its high inference speed."Artificial Analysis on X

  • "Highest score of any US-developed open-weight model — but still behind Kimi K2.6 at 54 vs 48."ChatForest builders log

  • "Day-0 inference support — Ultra, ASR, and Content Safety all live."Eigen AI blog

  • "Resource requirements are real — even quantized this isn't a single-GPU local run."Reddit r/LocalLLaMA

  • "Is NVIDIA becoming a model company now? The hybrid architecture suggests yes."HN comment

  • "OpenMDW-1.1 license actually allows commercial use without the usual research-only carve-outs."Reddit r/LocalLLaMA

  • "Speed advantage is real on Blackwell — open question whether it holds on H100 / MI300 stacks."HN comment

Hacker News35 pts

We launched Infracost on HN five years ago ( https://news.ycombinator.com/item?id=26064588 ) where our CLI generated cost estimates for infra-as-code, e.g. "this Terraform PR adds $400/mo". The idea was to shift cloud costs (FinOps) left, so engineers get visibility of costs before deployment and make better decisions. Earlier this year we started seeing agent traffic in our logs and it looked like coding agents were calling our CLI. But that CLI wasn't designed with coding agents in mind. We we... (35 points, 22 comments).

Market Signal

Why It Has Market Pull

Cost.dev is the agent-facing surface of Infracost — YC W21, freshly funded with a $15M Series A in November 2025, and already used inside 10% of the Fortune 500. The headline is misleading: this is not an LLM cost router but a way to keep coding agents from racking up cloud infrastructure bills, with a side benefit of cutting Claude output tokens 79% by redesigning the CLI for agent callers.

  • $15M Series A in Nov 2025 led by Pruven Capital with Sequoia, Y Combinator, Mango Capital, TIAA Ventures, plus angels from Supabase and Essence VC.
  • 12.3k GitHub stars, 117 releases on infracost/infracost; ships three Claude/Codex/Gemini skills via infracost/agent-skills.
  • Reports 79% Claude output token reduction and 67% API cost reduction vs a bare-Claude baseline after redesigning the CLI output for agent callers, not humans.
  • Adopted by 3,500+ companies including ~10% of the Fortune 500 per Series A press materials.
  • Install line is one command — `claude plugins marketplace add infracost/agent-skills` — making it concretely steal-able into a workflow today.

feedbacks

What People Are Saying

  • "People still underestimate how quickly token limits and context bloat become the bottleneck when you start running agents in production loops."HN comment

  • "The interesting bit is making cloud cost a first-class constraint for the agent loop, not just a post-hoc report."HN comment

  • "Agents writing IaC in prod is rare today but not zero — the token-optimization pattern generalizes well beyond infra CLIs."HN comment

  • "Tokens can get expended very quickly driving costs up substantially — for a large company with a huge AI bill, the savings could make up for the cost of the service."HN comment

  • "I don't know how they can justify a $250/month bill, let alone $1000/month."HN comment

  • "Why would anyone need 10,000 runs a month? Do people modify their infrastructure 10,000 times a month?"HN comment

  • "Not really seeing the point. I just use openrouter if I'm penny pinching."HN comment

Sources

GitHub

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

AI-powered job search system built on Claude Code. 14 skill modes, Go dashboard, PDF generation, batch processing.

AI agent skill that researches any topic across Reddit, X, YouTube, HN, Polymarket, and the web - then synthesizes a grounded summary

Give your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili, XiaoHongShu — one CLI, zero API fees.

Product Hunt

Running even one online store is a full-time job. SellerClaw is a team of AI agents that runs it for you: specialized agents for product sourcing, store management, and advertising, coordinated by a supervisor you direct. Tell it what to sell — the agents build listings, manage ads and pricing, and handle fulfillment and support across Shopify, eBay, and more. You stay in control: every action is visible and approvable, and you set how much runs on its own. Free to start.

Every great Claude response starts with context. Minimi listens across your Mac - docs, calls, messages, tabs - and gives Claude the full picture. No prompting. All on-device and private.

381

Leni is the most accurate and verifiable AI for serious investment work. Built on 21,000+ decision traces and processing 100M+ rows daily, it delivers finance-grade outputs with full auditability through source links, timestamps, and grounded comps. Leni outperforms GPT, Claude, and Manus on independent benchmarks for accuracy, modeling, and valuation while giving teams the trust they need when millions are on the line. Leni is part of Google Startups and a serious machine for investors.

Ideogram 4.0 is an open-weight text-to-image model trained from scratch, with bounding-box layout control, multilingual text rendering, and native 2K output. For developers and enterprises building on visual AI.

Most AI benchmarks test models in controlled environments. Agent Mode tests them on complex tasks to get more work done. Run autonomous agents that browse, research, code, use files, and complete multi-step workflows from a single prompt. Then watch each workflow unfold step by step. Every run contributes to the Agent Arena Leaderboard, ranking frontier models by real-world agentic performance.

LocalClicky is a Mac menubar app that lets you have a real conversation with your computer - completely offline. Say "Computer" to start a session. It stays listening. You chain commands back to back. Say "goodbye" when you're done. Everything runs on your machine: voice transcription, LLM multi models, VAD, macOS say No API keys. No subscription. No data leaving your Mac. MIT licensed.

Hacker News

Hi HN, not sure if anyone would be interested, but just wanted to share that I've been maintaining my small tool called 'lowfat' that helps me filters some of my verbose CLI output. It's a single binary, works as an agent hook or a shell wrapper. It has a plugin system to customize filters per command. The idea is pretty simple: agents don't need the full kubectl get -o yaml or any 10k-line dump to make decisions. So that lowfat sits in between, strips the noise, and passes through what matters.... (134 points, 68 comments).

To my knowledge, this is the first formally verified implementation of an intersection algorithm for polygons. The experience of working with AI agents on this project changed a lot with recent model releases, as I describe in the readme. Opus 4.8 is able to provide algorithm implementation with formal proof in one shot, whereas previous models required me to provide proof strategies in multiple steps. Trust in the correctness comes entirely from the Lean checker and human review of a small spec... (79 points, 17 comments).

At my work they provided a single Claude subscription for everyone on the team. To be honest I like kiro better as it provides a way better SDD management. But the company can't provide it and I can't afford it yet. Turns out I had the skill creator skill in my claude instance so I made use of it to create this Skill. I made it fully by using Claude but I wanted to make it open source, so I asked it to help me make tests and preparations for it, even a CI to run python tests. Well, we got this r... (40 points, 17 comments).

I’ve spent the last seven months building a tool I wish I’d had in my previous roles. MimicScribe is a macOS menu bar app that fits the "AI notetaker" category. It has accurate on-device speaker identification (a first possibly?), real-time meeting talking points for discovery calls, and a fully keyboard- and voice-driven interface. I believe the accuracy of the speaker ID system is its biggest strength. I used fluid audio’s port of ( https://github.com/fluidInference/FluidAudio ) Pyannote's com... (21 points, 5 comments).

Hey I'm Will from the Prisma team, engineering manager and also the lead developer on Prisma Next. I'd like to introduce you all to the next version of Prisma: a full rewrite in TypeScript that builds on the established patterns in Prisma and comes with a family of skills that integrate it into whatever AI tooling you're using in 2026. (Read the announcement on our blog here: https://pris.ly/pn-ea ) The three topics in the title are brand new concepts in Prisma Next so let me give you a quick ru... (13 points, 2 comments).

HF Spaces

Track, rank and evaluate open LLMs and chatbots Modern React interface for comparing Large Language Models (LLMs) in an open and reproducible way. 📊 Interactive table with advanced sorting and filtering 🔍 Semantic model search 📌 Pin models for comparison 📱 Responsive and modern interface 🎨 Dark/Light mode ⚡️ Optimized performance with virtualization The project is split into two main parts: React Material-UI TanStack Table & Virtual Express.js FastAPI Hugging Face API Docker The application is containerized using Docker and can be run using: Open LLM Leaderboard is a Hugging Face Space tagged with docker, leaderboard, modality:text, submission:automatic, test:public. It has 14005 likes on Hu...

11.1K likes

Create your own AI comic with a single prompt Last release: AI Comic Factory 1.2 The AI Comic Factory has an official website: aicomicfactory.app For more information about my other projects please check linktr.ee/FLNGR. If you like the AI Comic Factory, let me know! I am always creating new spaces and exploring new ideas for demos, meaning I don't have much time to take care of all of them (I wish I could clone myself or ask robots to do it). If you appreciate the AI Comic Factory and would like to leave a tip, that would be very kind 🫶 First, I would like to highlight that everything is open-source (see here, here, here, here). However the project isn't a monolithic Space that can be dupli...

Kolors Virtual Try-On is a Hugging Face Space tagged with gradio, region:us. It has 10100 likes on Hugging Face.

5.1K likes

Wan2.2 Animate is a Hugging Face Space tagged with gradio, region:us. It has 5118 likes on Hugging Face.

5.1K likes

AudioCraft is a PyTorch library for deep learning research on audio generation. AudioCraft contains inference and training code for two state-of-the-art AI generative models producing high-quality audio: AudioGen and MusicGen. Installation AudioCraft requires Python 3.9, PyTorch 2.0.0. To install AudioCraft, you can run the following: We also recommend having ffmpeg installed, either through your system or Anaconda: At the moment, AudioCraft contains the training code and inference code for: MusicGen: A state-of-the-art controllable text-to-music model. AudioGen: A state-of-the-art text-to-sound model. EnCodec: A state-of-the-art high fidelity neural audio codec. Multi Band Diffusion: An EnC...

Arena Leaderboard is a Hugging Face Space tagged with static, leaderboard, region:us. It has 4905 likes on Hugging Face.

AI Tools — June 6, 2026 Edition | Agentic Brew