Tools Bench.

Product launches and open-source repos with enough signal to earn a second look.

Last Brew Time: Oct 8, 2026, 11:18 AM PT

Insight

AI builders keep reaching for the same playbook before writing a line of new model code, settling into a shared skills-and-plugins layer for agents

Featured

GitHub280.8K

Skills for Real Engineers. Straight from my .agents directory.

Market Signal

Why It Has Market Pull

This is a real phenomenon, not hype -- a personal agent-skills collection from a well-known TypeScript educator that became one of the fastest-growing repos on all of GitHub, now anchoring a dedicated product page built around it.

  • 280,000+ GitHub stars and 23,500+ forks as of today, with commits pushed within the last day
  • Logged the largest single-day star gain on all of GitHub trending -- 5,551 stars in one day, per independent trend trackers
  • Spun into a dedicated product page (aihero.dev/skills) with its own social account posting regular updates
  • 38 distinct skills covering TDD loops, git guardrails, and the widely-cited 'grill-me' interrogation skill
  • MIT-licensed, nearly 1,500 watchers, active release cadence (v1.3.1 and ongoing)

feedbacks

What People Are Saying

  • "the most useful skill I've written, and I use it even outside of coding"X post

  • "an absolute banger"X reply

  • "went through the tools and found them really interesting"X reply

  • "a healthier pattern for agent-assisted engineering"Tech blog

  • "reads like a working developer's drawer rather than a course"Tech blog

  • "still the top of GitHub trending"X post

GitHub27.4K

Open source repository of plugins primarily intended for knowledge workers to use in Claude Cowork

Market Signal

Why It Has Market Pull

This is a real, highly credible release: Anthropic's own open-source plugin library for Claude Cowork, officially maintained and shipped by the company itself. It has exploded in adoption, has clear ongoing builder momentum, and has already moved public markets, making it one of the strongest traction signals in this category.

  • 27,450+ GitHub stars and 3,190 forks, with commits still landing the same day as this review (Oct 8, 2026)
  • Hit #1 on GitHub Trending after its January 30, 2026 official launch
  • 11 official plugins (sales, legal, finance, marketing, data analysis, and more) with active ongoing development: 135 open issues, 59 open pull requests
  • Reported to have triggered a two-day, $285B stock-value rout across legal-software incumbents (Thomson Reuters -18%, RELX -14%, Wolters Kluwer -13%) after the legal plugin shipped
  • Covered by TechCrunch, Axios, CNBC, PYMNTS and InfoQ as a flagship enterprise AI release

feedbacks

What People Are Saying

  • "README invites external PRs that CI closes automatically"GitHub issue

  • "the single most impressive plugin in the library"Tech blog review

  • "absurdly good"Tech blog review

  • "slash commands are the real differentiator from standalone Skills"Tech blog review

  • "sales has been particularly successful for both direct sales people and those sales-adjacent roles"Press coverage

  • "a $285 billion rout across software, legal tech, financial services, and asset management sectors"Press coverage

GitHub21.7K

Reverse engineer anything with agents, from app behavior down to native binaries.

Market Signal

Why It Has Market Pull

rea is a genuine, explosively-trending open-source project, an MCP server that gives coding agents (Claude Code, Cursor, Codex, Gemini CLI, and more) the ability to reverse-engineer native binaries, APKs, Electron apps, and .NET assemblies via Ghidra/Hopper integration. Star growth is verifiably fast and sustained rather than a one-day spike, and it has real community contribution activity plus multiple independent review write-ups within its first week.

  • Star count climbed from roughly 10,400 to 13,000 to 23,926 (as of this review) within days, sustained growth, not a single-day pop
  • Hit #1 on GitHub Trending; gained about 2,956 stars in a single 24-hour period per independent coverage
  • 2,648 forks and 93 open issues, with external contributors actively shipping new analysis features
  • Integrates with 10+ coding agents (Claude Code, Codex, Cursor, Gemini CLI, Windsurf, Devin, Copilot CLI) through one MCP server
  • Covered independently by multiple review sites within its first week of going viral

feedbacks

What People Are Saying

  • "REA removes the tool-switching step and gives the agent a persistent evidence trail, which is the part that is genuinely hard to do by hand."independent review

  • "When a provider behaves unexpectedly, you are debugging REA's bridge rather than your own script."independent review

  • "demonstrates thoughtful design that prioritizes evidence-based findings over speculative conclusions"independent review

  • "inspect historical web network captures"GitHub issue

  • "inspect offline ELF layout through pwntools"GitHub issue

  • "Roadmap: clarify investigation ownership and modularize REA capability integration"GitHub issue

Sources

GitHub

Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More

Editorial diagram design for Claude Code, Codex, GitHub Copilot, Factory Droid, and Pi. 42 diagram types. Self-contained HTML + SVG. No shadows. No Mermaid slop.

ArtCraft is an intentional crafting engine for artists, designers, and filmmakers

Tool for automatic PS5 executables porting to Linux and Windows

Product Hunt

IrisGo helps solopreneurs automate recurring work. Get daily briefs, triage email, track to-dos, and turn demonstrated tasks into reusable workflows with Watch & Learn. Spend less time on admin and more time building your business. Free during beta.

Your marketing campaigns are personalized. Your landing pages should be too. GenPage learns your brand, builds dedicated pages for every ad, keyword, or target account, then continuously A-Z tests them on Autopilot - turning more traffic into pipeline and revenue.

Your team chats in one app, meets in another, and every AI works alone. Velozity brings it all into one place. Chat, audio and video calls with transcripts, tasks, calendar and files, plus AI agents (Claude, Codex, etc.) on the subscriptions you already pay for. Every agent shares the same context, so work moves between people and agents without having to re-explain. For startups, teams in dev, product, sales, and ops, or solo users juggling several AIs. No second AI bill. Free to start.

Pegasus 1.6 is the TwelveLabs model that understands video from a first-person point of view, built for the egocentric and teleoperated footage robotics and physical AI teams already collect. It recognizes entities more accurately, and produces persistent metadata across segment types. Point it at raw teleoperation or wearable footage and get back labeled, timestamped, robot-ready data. No manual labeling pass required.

Figma's agent lives right on the Design canvas and works with your real components, variables and team files, not flat images. Generate layout directions, bulk-edit across screens, turn comments into changes and build prototypes from a prompt. Pull context from Notion, Slack, GitHub and Linear via MCP, build your own plugins and shaders, and save repeatable workflows as skills your team runs with "/". Stay in flow from idea to review.

Mistral Large 4 (aka Le Chonk) is Mistral's largest model yet: a natively multimodal mixture of experts model with 1T total and 49B active parameters and a 1M token context window. Mistral says it is the strongest open weight model from the US or Europe, and it ranks top 5 on the Artificial Analysis Cyber Index. It was trained in Mistral's own European datacenters. The preview API is live today on Mistral Studio, with open weights due by the end of October.

YC Launch

AI Chemistry Foundry - from a molecule on a screen to a compound in a vial b12 Labs · Summer 2025 · Healthcare Tags: AI-powered Drug Discovery, Robotics. Website: https://b12-labs.com/

Voyager runs on your computer and lets you use models like Astra, Opus, and DeepSeek to steer 100+ AI image/video models and creative tools. Voyager · Fall 2026 · B2B Tags: Design, Video, Marketing. Website: https://voyager.so

Hacker News

Hi HN! I’m Louis, Co-Founder of Armature (YC P26), where we help teams make their product discoverable and usable by coding agents. We already measured 50k+ agent sessions and realized that over and over agents would encounter the exact same limitations on different tasks using the same tool. So we wondered why these weren’t fixed. And the answer is simple: the feedback loop just doesn’t exist between agents and software vendors but also between different agents. Humans can share their experienc... (64 points, 47 comments).

Hi HN, we're Thomas and Olivier from Terse ( https://www.useterse.ai/ ) We've built Durable Actors, an open-source alternative to Cloudflare's Durable Objects. A Durable Object/Actor is a tiny server that handles one request at a time and has its own SQLite database. There's exactly one of each in the world and it is addressed by name. This is the perfect primitive for deploying multiplayer agents. Each agent can have its own Durable Actor, and each user can connect to that Actor via websocket.... (42 points, 25 comments).

Hello HN, I'm Ajo and I built Strata. I spent 4 years at Netflix solving self service for non-technical business users. I think I cracked it with my unique approach to semantic layer design. The key challenge is balancing expressiveness with ease of use for our non technical colleagues. It just so happens that focus made it work pretty well with LLMs too. Strata is a full stack solution. It includes a semantic layer, dashboards, subscriptions, and google sheets exports. All of it can be done vie... (25 points, 17 comments).

Hi HN, this is Yarik and Vlad from VOYGR - we are building the tools for agents and apps to engage with local businesses. It all started with our own pain point at VOYGR: calling businesses to verify if they are open. We are both from Google (Maps and Search) and even there, the merchants and venues don’t keep this info updated. So we built an API and started using it in-house. On July 4th, we were driving through Portland looking for a place to eat. Google Maps was saying “Holiday hours may var... (16 points, 4 comments).

Hi HN, I'm Justin. Breadcrumb records everything you do on your Mac (screen + meetings + AI transcripts + what you and your AI decided) and turns it into memory your AI can search. It's local and encrypted. You can also teach it rules by talking to it and it makes sure the right rules turn up in the right context. Works with Claude Code / Codex / Cursor / opencode. All of this is exposed to your AI as 30+ MCP tools (here's the definitions): https://innerloop.works/breadcrumb/mcp I started it in... (49 points, 9 comments).

HF Spaces

Benchmarks and news on various repros of TypeSafe's Jev Who is rebuilding TypeSafe's Jev (System One / RLCD) in the open? This static Space opens on the Decision Index leaderboard; the News tab tracks the artifacts in one combined grid, color-coded by kind: Decoding: parallel constrained decoding on stock models (inference technique, no new weights) Diffusion: text diffusion models run in a "Jev mode" Trained: Jev-like scoring heads and fine-tunes, weights often on the Hub, promised models listed last Prior art: "this already exists" claims Explainers: architecture speculation, explainers, benchmarks and roundups Cards sort by a trending score: ♥ likes on X + 5 × GitHub stars + 8 × Hub likes...

6-step Qwen-Image-2.1, T2I + editing, vs-base comparison Viggle Turbo v0.3 — 6-step Qwen-Image-2.1 A distilled Qwen-Image-2.1 that generates and edits images in 6 steps with no classifier-free guidance, about 5× faster than the 40-step base model. On most prompts it is hard to tell apart from the base model; small, dense text and complicated edits (multi-reference composition, face swaps, identity-preserving edits) can still fall short of it. v0.3 (2026-09-29): at 6 steps, less grain than v0.2.1 and a little softer on fine texture. We think 6 steps is close to its capacity: every further gain we found cost something elsewhere. The new 9-step setting runs 7 turbo steps and lets the base model...

Train open models with RL inside real agent harnesses A research article built with research-article-template. Source lives in FineEnvs under content/articles/multi-harness-rl/. | Path | What | | --- | --- | | app/src/content/article.mdx | Frontmatter and the chapter registry — the explicit import list is the running order | | app/src/content/chapters/ | One .mdx per section | | app/src/content/embeds/ | Standalone HTML/D3 visualizations, one file each | | app/src/content/assets/image/ | Images | | app/src/content/assets/data/ | Data files, served at /data/ | | app/src/content/bibliography.bib | References, cited as [@key] | From the repo root, over the Hub HTTP endpoint (no git remote, no n...

141 likes

Hugging Face model download stats, history & trends Daily download history, likes and derivative families for every model on the Hugging Face Hub, going back to July 2024. Look up any model: ?model=org/name Compare up to five models on one chart See how much of a model's reach comes from its quantizations, fine-tunes, adapters and merges Galaxy: every model built on a base model, drawn as a galaxy (?view=galaxy) Wrapped: any author's last 12 months on the Hub in six cards (?view=wrapped) Weekly rankings: most downloaded, fastest growing, new breakouts, biggest families, organizations A README badge with monthly downloads and a sparkline, updated daily Data comes from daily snapshots of cfahl...

Play Mario, Rubik's Cube and Tetris with JEV-27B Launch a live JEV-27B game run in the game arena. Mario: original NES World 1-1, with movement and jump decisions. 3D Rubik's Cube: a 25-turn scramble; select a seed or create a new scramble. Tetris: smooth gravity acceleration, continuing at maximum speed until 20 lines. The right panel shows the selected action, option probabilities, and current / average individual model inference duration. Mario movement and jump are separate calls. Start and stop runs yourself; one run per game executes at a time, with a short queue. Recent runs reconnect when you reload the page. The game worker runs on AutoTrust's existing B300 deployment. Model inputs...

Video generation with a synchronized soundtrack MiniMax-H3 — unquantized, split across two Spaces Joint video and soundtrack out of a single denoising pass, at bfloat16 with no quantization anywhere. This Space is the denoising half: the 61.73 GiB transformer and the two autoencoders. The 62.14 GiB Qwen3-VL conditioner runs in qwen3vl-conditioner, which this Space calls over the gradio API for every request. The weights are the public MiniMaxAI/MiniMax-H3 diffusers checkpoint. MiniMax-H3 is 195.9 GiB in bfloat16 and a ZeroGPU Space is evicted at 150 GB of storage. An unquantized single Space is therefore impossible, which is why quantized demos of it run NVFP4 or float8 weights. Cut the Mini...