Tools Bench.

Product launches and open-source repos with enough signal to earn a second look.

Last Brew Time: Sep 16, 2026, 11:38 AM PT

Insight

The AI tools breaking through right now share one instinct: push the model onto your own machine, not a cloud API

Featured

GitHub24.2K

Open source repository of plugins primarily intended for knowledge workers to use in Claude Cowork

Market Signal

Why It Has Market Pull

This is directly backed by Anthropic and tied to Claude Cowork, a strategic product push that has already moved markets — coverage from CNBC, VentureBeat, and Fortune shows this isn't a niche experiment but part of Anthropic's core roadmap for knowledge-work automation, with the plugin repo itself seeing steady real usage growth.

  • 24,252 GitHub stars, 2,919 forks, 31 contributors as of September 2026, launched January 23, 2026
  • Direct Anthropic product: plugins for Claude Cowork, which launched as a research preview in January 2026 and is now being merged into the main Claude chat product (announced September 16, 2026)
  • Press coverage from CNBC, VentureBeat, Fortune, and PYMNTS — one Gizmodo piece described the Cowork plugin launch as unsettling software-stock markets
  • 77 open issues, proportionate for an actively used official repo
  • Installable directly from within Cowork via a built-in plugin marketplace, lowering adoption friction versus a typical GitHub project

feedbacks

What People Are Saying

  • "Plugins tailor Claude to specific job functions"PYMNTS coverage

  • "Claude Cowork's industry-specific plug-in freaking out investors"Gizmodo coverage

  • "spent nine days testing Claude Cowork and compiled 56 practical tips based on workflows they rebuilt from scratch"HackerNoon writeup

  • "Anthropic brings Claude Cowork to mobile and web as usage data shows most users aren't coding"VentureBeat headline

  • "each plugin bundles the skills, connectors, slash commands, and sub-agents for a specific job function"project README

  • "companies able to connect Claude Cowork to Google Drive, Gmail, DocuSign and FactSet"press coverage

HF Spaces3.5K likes

generate a video from an image with a text prompt Wan2.2 14B Fast is a Hugging Face Space tagged with gradio, mcp-server, region:us. It has 3517 likes on Hugging Face.

Market Signal

Why It Has Market Pull

By far the strongest signal in this set: an official Hugging Face infrastructure team project (ZeroGPU + Ahead-of-Time Inductor optimization), not a hobbyist fork, that has sustained very high engagement for over a year. The scale of likes and the depth/duration of real troubleshooting activity is a clear step above every other Space evaluated here.

  • 3,517+ likes — by a wide margin the highest of any Space reviewed in this set.
  • Built by zerogpu-aoti, Hugging Face's own team demonstrating FP8 quantization + Ahead-of-Time Inductor compilation techniques on the Wan2.2 14B video model — an official reference implementation, not an individual's side project.
  • 42+ active discussion threads spanning over a year, including real technical troubleshooting ('Can I run this model on my computer?', 'Can't run the space on H100', video generation timing questions) — sustained hands-on usage, not a one-time spike.
  • Tagged mcp-server, making it directly callable from agent tooling.
  • Part of a broader zerogpu-aoti collection of accelerated Wan2.2 demos, suggesting an ongoing, maintained initiative rather than a one-off release.

feedbacks

What People Are Saying

  • "How much time it takes to generate 20 sec video"HF Space discussion (umair894)

  • "If I can't run it on my computer then where can I run it?"HF Space discussion (AsiGPorat)

  • "Can I run this model on my computer?"HF Space discussion (Tukis)

  • "Why is the space down at the moment"HF Space discussion (Goku619)

  • "Can't run the space on H100"HF Space discussion (Bear-ai)

  • "Support aspect-ratio-preserving output (avoid forced 832x480)"HF Space discussion (feature request)

HF Spaces540 likes

Real trained RL policies for the Microduck robot, running fully in the browser: MuJoCo compiled to WebAssembly steps the physics, onnxruntime-web runs the policy network at 50 Hz. No server, no backend. Two locomotion variants of the same robot are included: legs (walking, the default) and rollers (the wheeled skating variant). Press M (or hold D-pad up ~1 s on a gamepad) to switch; the roller model, meshes and policies are lazy-loaded on the first switch. | Mode | Checkpoint | What it does | |--------|-----------|--------------| | Run (legs) | BESTalphawalking.onnx | Velocity-tracking locomotion (arrows / WASD to steer) | | Sit | BESTalphasitstand.onnx | Sits down on its hull, stands back u...

Market Signal

Why It Has Market Pull

This is backed by a real, currently thriving robotics business rather than a solo hobbyist project. Pollen Robotics (now part of Hugging Face) opened pre-orders for its physical $399 Microduck robot and sold over $2.5 million worth in the first 24 hours, passing $5 million and 10,000+ units within days — independently reported demand strong enough to push delivery timelines out and trigger reseller scalping. The browser simulator is the free on-ramp for that same ecosystem, and it's already attracting outside builders.

  • 540+ likes and 8 active community discussion threads, including outside developers building on top of it (e.g. a 'Microduck Racer' modification and requests to hit the robot's real API from the simulator).
  • The physical Microduck robot sold $2.5M+ in pre-orders within 24 hours and surpassed $5M / 10,000+ units within days — independently confirmed by Engadget and CNBC coverage.
  • Strong enough demand to push new-order lead times from a Christmas 2026 promise out to 4-6 months, and to trigger visible reseller scalping activity.
  • Runs real trained RL policies fully in-browser (MuJoCo compiled to WebAssembly + onnxruntime-web at 50Hz, no server/backend) — a genuine technical differentiator, not just a marketing demo.
  • Backed by Pollen Robotics, a real robotics company now under the Hugging Face umbrella, with the full SDK/sim/training stack open-sourced under Apache-2.0.

feedbacks

What People Are Saying

  • "Try Microduck routines in the simulator using the robot's API?"HF Space discussion (ramzi-abbyad)

  • "Microduck Racer - Simulator modified for racing"HF Space discussion (Nirav-Madhani)

  • "Hugging Face's new duck robot is selling fast"CNBC headline

  • "Over $2,500,000 worth of Microducks were sold in the first 24 hours"press coverage (Engadget/CNBC roundup)

  • "sales surpassing $5 million"press coverage

  • "Why Resellers Are Buying Pollen Robotics' Microducks"Resell Calendar (independent market coverage)

Product Hunt482

Run Cloud Agents, control VS Code and CLI sessions, review Pull Requests, and respond to coding agents from your iPhone, iPad, or Android.

Market Signal

Why It Has Market Pull

This is a mobile companion for an already well-established open-source AI coding agent with a large, active developer base -- among the stronger candidates in this set, worth a closer look.

  • Parent project Kilo Code has 27.3k GitHub stars and 3.2k forks with 31,000+ commits and 135 open PRs
  • Company claims 5M+ users and 10T+ tokens processed per month across the platform
  • Mobile launch ranked #3 Product of the Day with 482 upvotes
  • Maker actively responded to feature requests during the launch thread
  • Active Reddit community (r/kilocode) discussing the product favorably

feedbacks

What People Are Saying

  • "Everything you can do at your desk, I want you to be able to do from your phone too...The app is open source, and I'll gladly consider your request."Product Hunt comment (maker)

  • "The idea was to build an app that can cover the whole workflow...I usually use it for new landing pages, including research and implementation, and reviewing PRs."Product Hunt comment (maker)

  • "One small thing that will help is a biometric authentication feature on the app...It will be safe for me to use this even out of the house if there is face ID."Product Hunt comment

  • "We have already Unlock with biometric. You can find it in Settings -> Preferences -> General."Product Hunt comment (maker reply)

  • "Cursor but open and free"Reddit r/kilocode

  • "Coding from your phone is real now, and we tried it with Kilo"Medium article

Product Hunt186

OpenAI's managed version of the Codex harness to build cloud agents with one API call instead of your own orchestration. Handles long sessions, smart tool use, and subagents working in parallel. Runs on OpenAI or partner sandboxes. No extra fees beyond token and tool usage. Open source harness, public beta now.

Market Signal

Why It Has Market Pull

This is a first-party managed agent platform from OpenAI itself, backed by broad partner-sandbox integrations and heavy press and developer-community coverage -- among the strongest signals in this set, worth a closer look.

  • Released into public beta on September 10, 2026, directly from OpenAI
  • Sandbox partner ecosystem includes Vercel, Cloudflare, DigitalOcean, E2B, Modal, Oracle, Runloop, and Blaxel
  • Covered by TechCrunch, InfoQ, and multiple independent technical blogs within days of launch
  • Generated an active Hacker News discussion thread comparing it to Anthropic's equivalent offerings
  • Pricing is usage-based on model tokens (e.g. $10/M input, $50/M output) with no separate platform fee

feedbacks

What People Are Saying

  • "this agent-as-a-service approach lets developers plug in the tools they need while allowing OpenAI to encapsulate and iterate on core components like memory and context management"Hacker News comment

  • "constrained to using their own proprietary frontier models"Hacker News comment

  • "The Agents API: OpenAI Just Made the Harness a Product"Independent blog (CellCog)

  • "OpenAI's Agents API Is Really a Hosted Codex Harness"Independent blog (AgentConn)

  • "OpenAI's new Agents API gives developers the infrastructure behind Codex and ChatGPT"Tech press (the-decoder.com)

  • "OpenAI Agents API + NEW Codex Harness = Cloud Coding Agents INSTANTLY"YouTube video title

Sources

GitHub

260.0K

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.

68.2K

Autonomous coding agent as an SDK, IDE extension, or CLI assistant.

Product Hunt

One system for auth, payments, customer data, and analytics. One command to install. Launch a paid product the same day you start building.

Voiskey starts from what you meant, not just what you said. Speak a rough thought and it comes back shaped for where it's going and who's reading it: casual with a friend, composed with a colleague, technical with an AI. It arrives 5x faster than typing, cleaned up and ready to send, while you still sound like you. Voiskey is now available on iOS, macOS, Android, and Windows, in over 100 languages. It's free to start. Join during launch and get a free month of Pro.

Building a business with AI can quickly become a mess of chats, tools, ideas, and conflicting advice. siift turns that noise into a living map of your business, connecting strategy, evidence, decisions, and results in one place. From validation and build strategy to GTM and growth, our AI uses your evolving business context and proven processes to challenge assumptions, identify risks, and help you focus on what matters next.

Traditional research is deep but slow. Social listening is fast but shallow. LLMs are fluent but culturally blind. Anthropologic reads the whole internet through our 'Human Context Protocol' to uncover consumer, category and cultural truths. Nine workflows, 239 markets, 100+ languages. Research, innovation and foresight answers in minutes.

Your Axari AI twin understands your security world, works across your tools and teams, and keeps work moving until it is actually done. Assign it a goal, let it work proactively, or give it recurring responsibilities – all from within Slack or MS Teams. Attackers aren't going to use less AI, neither should you. You + your AI twin = indefatigable.

Narrative brings video editing, custom motion graphics, and reference-video style matching into one interface. Upload your footage, describe your edit, and keep refining through chat. Add music, sound effects, and transitions, or get inspiration from reference videos. From podcast clips to launch videos, create without learning Premiere or After Effects. We handle rendering and storage. Our launch video was made entirely in Narrative.

YC Launch

Papaya finds what's hurting your agent's quality, cost, and latency, and fixes it for you. Papaya · Fall 2026 · B2B Tags: AIOps, B2B, Enterprise Software, Infrastructure, AI. Website: https://papaya.fyi/

Hacker News

Hey HN, Toby from Nari Labs here. We've been working on making OSS speech models super-fast. Last year, we built Dia, the first OSS text-to-speech model capable of doing natural dialogue. Since then, so many more great speech models have been released to the public. But the market is still dominated by closed source models. We think that's an inference problem. Existing systems such as vLLM / SGLang are not well suited for multimodal inference. To prove this, we built an inference engine special... (90 points, 31 comments).

Hi HN - long-time lurker (since 2012!), first time poster. Pizza Bot is a self-hosted desktop app for Mac, Windows, and Linux that runs AI agents in the background and exposes them through an email-like UI. Finished work shows up in Unread, and anything waiting on your approval shows up in Action. It's Apache 2.0-licensed, there's no signup and no telemetry, and you bring your own model provider: Anthropic, Amazon Bedrock, Google Gemini, OpenAI, OpenRouter, or a local model through Ollama. There... (55 points, 33 comments).

Hi everyone, Been working on Otis, an open-source ai agent that gives you one minimal experience across local and hosted open-weight models, privacy-focused by design. On setup it recommends a local model based on the hardware Otis is running on, downloads it and runs it through llama.cpp for you. Ollama, LM Studio and Nvidia PAIR are supported too. Excited for everyone to try it and all feedback is welcome! (19 points, 4 comments).

Firefox Extension/Userscript and API to get Pangram scores for all articles on the hackernews frontpage. The extension allows you to hide articles with a high score. This is about detecting posts written by LLMs, not posts about AI. Feel free to use the API to build your own tooling/readers. Big thanks to https://news.ycombinator.com/user?id=salahadawi for providing the data :). (16 points, 3 comments).

We’re opening the private beta for AgentDrive, a persistent file/workspace layer for AI agents. AgentDrive gives agents: - Versioned artifacts and folders - Hosted MCP access for coding agents - Direct upload sessions for larger files - Share links and controlled public publishing - Workspace-scoped authorization and auditability - Many supported file formats such as MD, HTML, JSON, images, videos, and even spreadsheets The MCP surface is intentionally narrow and reviewed rather than exposing th... (6 points, 12 comments).

Hi, I built Chat-Man because I wanted a cheap way to give my agents access to WhatsApp without integrating a WhatsApp library separately in every project. You get WhatsApp MCP server, so you can connect WhatsApp to an agent and programmatically read, search, extract and send messages - but also have a web UI for non-techies. You can also receive webhooks for incoming messages for starred conversations. What you can do with the MCP? -Read WhatsApp messages and turn them into CRM records -Summaris... (14 points, 31 comments).

HF Spaces

Blind A/B ranking of MiniMax-H3 acceleration variants Human-judged ranking of ~26 MiniMax-H3 acceleration variants over a 200-prompt corpus, from blind pairwise votes on pre-generated clips, with confidence intervals, cost and slice breakdowns. The design and its reasoning are in arena/DESIGN.md; the app's own notes are in arena/README.md. This Space is private and must stay private until deliberately flipped. It streams ~3,700 clips out of the private dataset multimodalart/h3-pre-gen-arena. See Going public below. hfoauth: true above creates the OAuth app and injects OAUTHCLIENTID, OAUTHCLIENTSECRET, OAUTHSCOPES and OPENIDPROVIDERURL. arena/space_auth.py implements the flow by hand (this is...

A coding agent running entirely in your browser Pi browser workspace · MiniCPM5-2B Run Pi's real agent core, just-bash, and MiniCPM5-2B entirely in your browser. Chat with Pi, inspect and edit files, and execute shell commands in a persistent virtual workspace. A browser worker runs the agent loop, local WebGPU inference, and every tool. No inference API, API key, or shell server is involved. Send your first message. If the complete verified model is already cached in this browser, Pi restores it and sends your message automatically, without a confirmation dialog. Otherwise, a Load model & send dialog explains the download. Confirming closes the dialog and shows progress in the top bar, then...

Demo of the Collection of Qwen Image Edit LoRAs Prompt-based image editing on Qwen-Image-Edit-2511, with identity preservation (change pose/clothes/scene while keeping the person's face), HD output, and an anti-deformed-hands option. The Space exposes one endpoint: /edit. Base URL / Space id: kulkas2pintu/QWENEDITIMAGE. | name | type | default | meaning | |---|---|---|---| | image | image (filepath/URL) | — | required. The image to edit (the person/subject whose identity is kept). | | prompt | str | — | required. What to change, e.g. "change her outfit to a red dress; turn to look left". | | loraadapter | str | "" | Optional style LoRA name (see list below). "" = no style / base edit. | | im...

Video generation with a synchronized soundtrack MiniMax-H3 — unquantized, split across two Spaces Joint video and soundtrack out of a single denoising pass, at bfloat16 with no quantization anywhere. This Space is the denoising half: the 61.73 GiB transformer and the two autoencoders. The 62.14 GiB Qwen3-VL conditioner runs in qwen3vl-conditioner, which this Space calls over the gradio API for every request. The weights are the public MiniMaxAI/MiniMax-H3 diffusers checkpoint. MiniMax-H3 is 195.9 GiB in bfloat16 and a ZeroGPU Space is evicted at 150 GB of storage. An unquantized single Space is therefore impossible, which is why quantized demos of it run NVFP4 or float8 weights. Cut the Mini...

Demo of the Collection of Qwen Image Edit LoRAs QIE-2511 Rapid-AIO LoRAs Fast (Experimental) is a Hugging Face Space tagged with gradio, mcp-server, region:us. It has 213 likes on Hugging Face.

Simulate a fruit fly in your browser using WebGPU kernels A standalone fruit fly connectome demo: paint neurons, stimulate the network, and watch an articulated Three.js fly respond. Vanilla JavaScript, Vite, Three.js, and @huggingface/kernels. Simulated neural activity drives crafted walking, turning, and flight animations. The movements are illustrative, not validated predictions of fly behavior. Open the address printed by Vite, then click Download & start. All neural weights, fly meshes, fonts, and kernel templates are included in public/; setup loads them into the browser and caches verified weight chunks. No API key or external model service is required. Deploy the generated dist/ dire...