From Product Hunt

Another agent-shaped day: a desktop pet that feeds on your Claude Code tokens, an agent that actually runs your repo to reproduce bugs, plus local dictation and a browser-side architecture x-ray.

Tamamon — a desktop pet that grows as you code with Claude Code A macOS desktop companion that hatches from an egg and evolves based on your local Claude Code token activity — 20 collectible species via a weekly gacha, weather/time-reactive behavior (it wanders home to rest at night or in the rain), plus a HUD showing today/this-week coding and live CPU/memory. Everything stays on your Mac: no account, no sign-in, nothing uploaded. Free (Ko-fi), macOS 15+ Apple Silicon, signed & notarized.
▲ 254 · tamamons.com
Osloq — an AI agent that reproduces GitHub issues for you Instead of reading your code and guessing, Osloq runs it: connect GitHub, pick an issue, and an agent spins up a fresh sandbox, clones and runs your repo, and tries to reproduce the bug the way a developer would — returning a verdict backed by real evidence (logs, screenshots, the exact code path) or admitting it can't. Read-only least-privilege GitHub App, sandbox destroyed after each run, no training on your source. JS/TS, Python and Go today.
▲ 208 · osloq.com
Archify — understand any web app from your browser A browser extension that reverse-engineers live web apps: hover any element to see its framework, component type, library and the APIs it fires (each with a confidence score); click the toolbar for a whole-page profile — tech stack, where it's hosted (Vercel/Cloudflare/Netlify), and a client-side security roll-up flagging which third-party scripts can read sensitive fields like password or card number. 100% local, Apache-2.0 open source, no account.
▲ 188 · archify.salahxd.dev
Notta Desktop — Privacy Mode, fully local meeting transcription Bot-free recording that captures system + mic audio from Zoom, Meet, Teams, Slack or Webex without adding a bot to the call. Privacy Mode keeps audio, transcripts and notes on-device via a downloadable offline model (auto-detect plus English, Japanese, Korean, Cantonese, Simplified Chinese) — no cloud round-trip for sensitive conversations. Notta Desktop 1.2 for Mac.
▲ 55 · notta.ai
LetMeCheck.ai — a "blood test" for AI-generated codebases A health checkup for your code: run a diagnostic and get a report catching hidden bugs, vulnerabilities and quality issues in minutes. The angle for agent workflows is that it hands your coding agent skills, not one-off fixes — a review layer aimed squarely at vibe-coded projects.
▲ 17 · letmecheck.ai
Glaze by Raycast — now open to everyone Update: after last week's beta, Raycast opened Glaze to all. Describe an app in plain English and it builds a real native Mac app that lives in your dock, launches instantly, works offline, and taps real OS access (keyboard shortcuts, menu bar, filesystem, background processes) — then you iterate by chatting. Publish to a public or private store. Free to start, macOS Tahoe + Apple Silicon, more platforms to come.
▲ 488 · glaze.app
From Reddit

Mistral shipped a theorem-prover that finds real bugs, LongCat's giant MoE finally opened its weights, local DeepSeek kept outrunning Sonnet on wall-clock, and Fable's "inner voice" turned out to be adorable.

DeepSeek V4 Flash lands ~Sonnet 5 quality at ~3x the speed, locally A follow-up to last week's llama.cpp patch, now an actual indie coding bench on 2x RTX Pro 6000 (vLLM): V4 Flash finishes real coding tasks in ~2 min versus Sonnet 5's ~6 min — roughly 3x faster wall-clock — at around Sonnet-level quality. Opus and Fable still take the best single diff by a clear margin, but for local, fast, "good enough" agentic coding this is a strong data point.
LongCat-2.0 weights are finally public Meituan published INT8 and FP8 weights for LongCat-2.0 — the 1.6T-param MoE (~48B active) that had been the stealth "owl-alpha" on OpenRouter, trained and served entirely on AI-ASIC superpods with its own sparse-attention + 3-step MTP speculative decoding. Last week's blog reveal withheld the weights; now they're on Hugging Face for the community to benchmark.
Fable 5's "inner voice" is just muttering to itself Screenshots of Fable 5's raw reasoning trace went viral for reading like a creature grumbling and muttering the whole time it works. Framed as a "leak," but it's actually documented in the Fable 5 system card (pg 107-108) — and the top comment decodes the "muttering" as competitive-programming graph-proof checkpoint markers, not feelings. A charming footnote to Fable 5's contentious relaunch.

Quick hits — workflow shifts and hardware.

Andrew Ng: in 3-6 months, everyone's on self-improving loops Ng argues the shift away from one-shot prompting toward agentic evaluate-and-iterate loops is imminent — "no more prompting."
GLM-5.2 on 5x RTX Pro 6000 + a 5090 — an "expensive journey" A near-endgame local rig build log: Threadripper Pro 9975WX, WRX90 Sage SE, 4x48GB DDR5-6400, chasing full PCIe 5.0 x16 across every slot to run GLM-5.2. Relatable rabbit-hole reading for anyone with a multi-Pro-6000 box.
From Reddit

Control kept maturing on the Krea 2 open weights — a depth ControlNet and a training-free style reference — plus a pure-C++ audio engine that does 10-minute music locally, and a story-to-comic workflow.

Krea 2 Depth ControlNet for ComfyUI A depth ControlNet plus ComfyUI nodes for the shadow-dropped Krea 2 open weights, adding real structural control — pin composition and geometry from a depth map instead of fighting the prompt. Part of the fast-maturing Krea 2 workflow ecosystem (weights + a matching depth model are on Hugging Face).
audio.cpp — local music, SFX and stem separation in pure C++/GGML The all-in-one ggml audio engine (no Python) added a big media-generation batch: ACE-Step 1.5 Turbo/Base, Stable Audio 3 Small/Medium, HeartMuLa for music, plus Mel-Band RoFormer and HTDemucs for source separation. HeartMuLa now generates ~10 minutes of music in one run, and ACE-Step Turbo does 600s of music in ~60s wall-time (~10x real-time) — all through one native framework path.
Instant story-to-comic generator (ComfyUI, workflow included) A ComfyUI workflow that turns a plain-text story into a multi-panel comic — no LoRAs, no reference images required. A practical asset-generation pipeline rather than another one-off showcase.
ComfyUI-Krea2-StyleTransfer — training-free style reference Apply a reference image's style on Krea 2 with low content leakage — no LoRA training required.
A hands-on Ideogram V4 LoRA-training walkthrough Workflow-included write-up of training a LoRA and generating on the open-weight Ideogram V4 — the "how", warts and all.
Street View, but for historical events (built with GPT images) A browsable "time-travel atlas" that uses GPT image generation to render historical moments — a novel, if Eurocentric, take on generative-image as a product.
From Reddit

A native Raycast-alternative command center, open-source local dictation with 11 engines, an MCP gateway that stops tools from eating your context, a self-hosted YouTube player, and a bidirectional Figma agent.

Vehla — a native Mac command center with AI built in An Alfred/Raycast-style keyboard palette that folds in local + cloud AI models, RAG chat over your notes and clipboard, file/app search, Shortcuts, port and process tools, brew, custom AI actions and workflows — one keystroke away. Privacy-first with local model support, sold as a lifetime license for two Macs rather than a subscription. A serious Raycast alternative if you want native Mac actions plus your own model choice.
TypeWhisper 1.5 — open-source local dictation with 11 engines Free, open-source macOS dictation that runs fully local by default: press a hotkey, speak, and text is inserted app-aware into whatever you're using. Eleven engines including WhisperKit, Parakeet TDT v3, Apple SpeechAnalyzer and MLX-based Granite/Qwen3/Voxtral, plus optional Groq/OpenAI/xAI cloud — all behind a plugin system with Workflows. Exactly the kind of configurable, local-first tool for heavy transcription.
Toolport — run many MCP servers without the token tax An MIT, local-first MCP gateway: put 15+ servers behind one endpoint with "lazy discovery" — instead of dumping every tool list into context each turn, it exposes just three meta-tools the agent searches on demand (up to ~91% fewer tokens at the same task success). Secrets live in the OS keychain and are injected at call time, and it flags tool-poisoning, rug-pulls and agentjacking. Import your Claude MCP config once and reuse it across 20+ agents.
YT-DLP Web Player — a self-hosted YouTube front-end A self-hosted web player/downloader built on yt-dlp, pitched as a Revanced / YT Premium alternative: background play, downloads and no ads, running on your own box. The top of r/selfhosted today.
Figwright — a local, bidirectional Figma agent for MCP clients A free, open-source alternative to Figma's Dev Mode MCP that runs locally (no API token, no paid seat). It reads your project's components, design tokens, icons and coding conventions so generated code matches your codebase — and writes back to the canvas (frames, auto-layout, styles, variables). 92 tools, Activity + Debug panels, and it works with Claude Code and Cursor.
Pookify — Claude Code status in a Dynamic Island A macOS app that surfaces your live Claude Code session status in a Dynamic-Island-style pill at the top of the screen.
SnapGrep — grep for your screen A menu-bar utility that fuzzy-searches any visible on-screen text and lets you act on it — find and grab text the OS won't normally select.
Serverless P2P File Transfer — WebRTC + Rust WASM Direct browser-to-browser file transfer with zero servers, zero-knowledge and no size limit.
From Twitter

On Twitter today, open models became a click inside real coding tools, Claude skills turned into a package ecosystem, AI video jumped to 4K inside CapCut, and self-hosted apps kept closing the gap with paid software.