Radar

What I'm watching

Curated open source and AI projects worth a founder's time — scannable, opinionated, no noise.

115 projectsPage 1 / 3
Diagram Design

Oct 10, 2026

Diagram Design

Worth testing
Demo

Agent skill that draws branded editorial diagrams as one self-contained HTML+SVG file — 44 types, no Mermaid slop.

Why it matters+

Founders still paste Mermaid into docs or burn an hour in Figma for a deck slide. Diagram Design gives your coding agent taste rules, type-specific layouts, and brand colors so the diagram you publish looks intentional on the first pass.

agent-skillsdiagramssvgarchitecturedesign

Similar · Archify, json-render, Taste Skill

X
NiubiGEO

Oct 9, 2026

NiubiGEO

★5,471Worth testing
Demo

Self-hosted GEO lab: ask models about your product, keep the raw answers, and watch who else shows up in AI search.

Why it matters+

Founders still buy black-box “AI visibility” scores or guess whether ChatGPT recommends them. NiubiGEO puts the evidence on your machine — domain + keyword runs across models, competitors named in the answers, citations you can open, and history you can re-run when positioning changes.

geoai-visibilitybrandself-hostedopen-source

Similar · OpenSEO, Agent Reach, HyperResearch

X
PhotoCraft

Oct 8, 2026

PhotoCraft

★19,817Must watch
MCP

Open-source Photoshop in pure Rust — real PSD, GPU compositor, and every edit is a command agents can run.

Why it matters+

Founders still rent Adobe or glue brittle scripts onto web editors. PhotoCraft is a native offline engine with layered PSD fidelity plus CLI, JSON control, and MCP — so a coding agent can open, grade, and batch real design files without a SaaS middleman.

image-editingrustagentsmcppsd

Similar · Open Design, Impeccable, HyperFrames

X
AutoHarness

Oct 7, 2026

AutoHarness

★9,164Must watch
MCP

Self-learning skill layer for Claude Code — distills real sessions into skills, merges duplicates, and prunes what stops getting used.

Why it matters+

Static skill packs rot the week after you install them. Founders running Claude Code all day need the harness to keep its own playbooks current from real work — not another dump of SKILL.md files and not a separate labeling job. AutoHarness is that missing lifecycle: learn, fold, graduate, archive.

agentsskillsclaude-codeharnessmcp

Similar · Skill Recorder, SkillOpt, Impeccable

X
e2e

Oct 6, 2026

e2e

★5,184Worth testing

Natural-language e2e for web and mobile: the agent drives the app, then later runs replay the same path with no model calls until the UI changes.

Why it matters+

Most teams still choose between brittle selectors and expensive always-on agent runs. e2e gives founders both: agents write the hard paths once, CI stays cheap on replay, and failures stay readable when the product actually changed. That is the missing layer between Playwright scripts and chat-driven browser demos.

testinge2eagentswebmobileplaywright

Similar · chrome-devtools-mcp, page-agent, Firecrawl

X
REA

Oct 5, 2026

REA

★3,072Must watch
MCP

Reverse engineer any app with your coding agent — one CLI and MCP server from binary behavior down to native code.

Why it matters+

Founders copy product ideas by guessing from the UI. REA lets the agent inspect a real app — even without source — explain how a feature works with evidence, then rebuild a version for your stack in the same session. That is a capability gap no generic coding harness fills: investigation tools under the agent, not another chat wrapper.

agentsmcpreverse-engineeringdevtoolscli

Similar · chrome-devtools-mcp, meat, CLI-Anything

X
Impeccable

Oct 4, 2026

Impeccable

★75,559Must watch
Demo

Design language for AI coding agents — one skill, 24 commands, live browser iteration, and 61 deterministic anti-slop detectors.

Why it matters+

Every model defaults to the same SaaS look: Inter, purple gradients, nested cards, gray on color. Thin “frontend skill” packs tell the agent what to avoid; Impeccable turns that into a repeatable workflow with product context, named commands, and rules that run without an API key. Founders shipping agent-built UI get fewer generic screens and a way to keep taste consistent across the team’s tools.

agentsdesignfrontendskillsdevtools

Similar · shadcn/lint, Taste Skill, Hallmark

X
Monid

Oct 3, 2026

Monid

★1,386Must watch
API

OpenRouter for agent tools — one key, discover across 2,000+ endpoints, pay only when a call actually runs.

Why it matters+

Agent stacks already glue five SDKs and five keys just to search, enrich, and scrape. Monid collapses that surface: discover ranks endpoints by job fit with live health and latency, then run settles usage on the raw response so empty or failed vendor results do not get billed like success. Founders shipping agents care about that more than another model router.

agentstoolsapi-gatewaymcpdevtools

Similar · OmniRoute, Weave Router

X
Context Mode

Oct 2, 2026

Context Mode

★24,856Must watch
MCP

MCP layer that keeps fat tool dumps out of the coding-agent context window — sandbox the work, return only the answer.

Why it matters+

Context rot is the daily tax of shipping with coding agents. Most stacks either compress tokens after the damage or beg you to start a new chat. Context Mode fixes it at the source: keep bulk data off the wire into the model, keep session memory local, and stretch a useful session from tens of minutes into hours without a cloud middleman.

agentsmcpcontextcoding-agentsdevtools

Similar · Paritok, codebase-memory-mcp, cost-xray

X
OpenAPPA

Oct 1, 2026

OpenAPPA

★925Must watch
API

Deterministic guardrails between an agent and its tools — every tool call is checked against where sensitive data is allowed to go.

Why it matters+

Most agent safety today is soft: classifiers, PII detectors, or hope the model behaves. Founders shipping agents that touch customer data need a hard gate that does not flake. OpenAPPA is that gate — policy you can replay in CI, not a vibe check at runtime.

agentssecurityguardrailspermissionsdevtools

Similar · SkillSpector, OpenShell, Deepsec

X
magpie

Sep 30, 2026

magpie

★3,191Must watch
API

Menu-bar control plane for every coding agent on your machine — one place to set Claude Code, Codex, OpenCode, and friends on any model.

Why it matters+

If you run more than one coding agent, model config is a tax. magpie kills the dance of editing settings.json, config.toml, and opencode files by hand, and it lets one subscription feed the rest of your stack. That is daily leverage for anyone shipping with agents, not another harness.

agentscoding-agentsmodelsgatewaydevtools
X
PageIndex

Sep 29, 2026

PageIndex

★36,544Must watch
APIMCP

Vectorless RAG that builds a tree over long PDFs and lets the model reason to the right section — no chunking, no vector DB.

Why it matters+

If your product answers questions over long PDFs, vector RAG still misses the section that matters and returns lookalike paragraphs. PageIndex is a different retrieval job: structure first, then reason. That is the unlock founders need for reliable doc agents without babysitting chunk sizes.

ragretrievaldocumentsagents
X
HyperFrames

Sep 28, 2026

HyperFrames

★53,723Must watch
DemoAPI

Write HTML and CSS, get a deterministic MP4 — a video framework built so coding agents can ship real motion, not slide dumps.

Why it matters+

Founders still hand video work to editors or fight Remotion timelines while agents write the rest of the product. HyperFrames closes that gap: agents author seekable HTML compositions, lint them, preview, and render stable MP4s with the same tools they already use for code.

videoagentshtmlrenderingskillsdevtools

Similar · OpenMontage, OpenCut, json-render

X
Paperclip

Sep 27, 2026

Paperclip

★87,926Must watch
Demo

Open-source control plane for agent companies — org charts, goals, budgets, and heartbeats so you manage the business, not twenty terminals.

Why it matters+

Founders already run fleets of coding agents and lose the thread: which tab owns which task, who spent the budget, what survives a reboot. Paperclip turns that mess into a task manager with company structure — review gates, monthly budgets that stop runaway loops, and portable org templates you can export without leaking secrets.

agentsorchestrationopsorg-chartbudgets
X
cost-xray

Sep 26, 2026

cost-xray

★2,163Must watch

Wire-level X-ray for Claude Code and Codex — see the real API request and what each part costs, not just a session total.

Why it matters+

Log-based spend tools answer how much you spent. They miss the request-time prefix that never lands in the transcript: injected schemas, MCP catalogs, reminders. That invisible half of the context is often what bloats the bill and crowds the window. Founders running heavy agent days need to know which MCP is dead weight and which tool call owns the dollars.

agentsobservabilitycostclaude-codecodex
X
Treg

Sep 25, 2026

Treg

★3,243Must watch
APIMCP

OpenRouter for agent tools — one token, 3,000+ endpoints across 60+ providers, priced per call, no vendor signup.

Why it matters+

Real agent work dies on tool access. Semrush, Apollo, Crunchbase and friends cost monthly seats nobody buys for a single run, and sharing keys across Claude Code, Codex, and CI is a mess. Model routers fix which LLM you call. Treg fixes which capability the agent can touch without ten signups or leaking credentials into every session.

agentstoolsmcpregistrycredentials
X
AX

Sep 24, 2026

AX

★9,423Must watch
API

Google's kubectl-shaped runtime for agent workloads — declare tasks, sandboxes, workspaces, and network fences like cluster jobs.

Why it matters+

Founders shipping multi-agent products hit the same wall: untrusted loops that burn money, leak network access, and cannot pause cleanly. Coding harnesses and in-process SDKs do not solve fleet isolation at scale. AX treats agents like Kubernetes treated containers — declarative specs, sandboxes, gateways, and ops verbs you already know. That is infrastructure, not another terminal coding agent.

agentsorchestrationkubernetessandboxgoogle
X
Univer

Sep 23, 2026

Univer

★15,810Must watch
DemoAPI

Open office runtime for people and agents — sheets, docs, slides, and tables with one Facade API in the browser and on Node.

Why it matters+

Founders building finance, ops, and BI products still glue Excel parsers to chat agents and hope the grid stays coherent. Univer is the missing surface: embed real editable office UIs, let agents hit the same APIs server-side, then let people review worktree drafts before merge. That is product infrastructure, not another coding TUI.

office-sdkspreadsheetsagentsheadlessdocuments

Similar · json-render, Firecrawl

X
Laya

Sep 22, 2026

Laya

★12,353Must watch
DemoAPI

Non-autoregressive System 1 engine — typed choice/score/yes-no answers over tickets and JSON in ~33 ms, no free-form text to parse.

Why it matters+

Most agent stacks still burn a full LLM call to classify a ticket, then hope the JSON comes back clean. Founders shipping support, ops, and risk pipelines need fast, calibrated labels without hallucinated free text. Laya is that layer: milliseconds, typed outputs, multilingual out of the box — the cheap System 1 beside your slower planner.

decision-modelssystem-1agentsmultilingualinference

Similar · RTK, Paritok

X
json-render

Sep 21, 2026

json-render

★17,581Must watch
Demo

Vercel’s Generative UI framework — the model only emits JSON against your component catalog, so AI-built screens stay schema-safe and streamable.

Why it matters+

Most AI UI demos still dump raw JSX or markdown and hope nothing breaks. Founders shipping agent dashboards, copilots, and personalized product surfaces need the opposite: dynamic screens that stay on-brand and type-checked. json-render is the practical middle path — AI fills the tree, you own the design system — and it is already packaged for the stacks product teams actually ship.

generative-uireactagentsuivercel

Similar · shadcn/lint

X
OpenShell

Sep 20, 2026

OpenShell

★8,709Must watch
API

NVIDIA’s sandboxed runtime for autonomous agents — YAML policies gate files, network, and credentials so a coding agent can’t exfiltrate your machine.

Why it matters+

Teams are handing coding agents shell access and API keys with almost no boundary. SkillSpector checks the skills you install; OpenShell is where those agents actually run — filesystem locks, process limits, provider credentials injected only to allowed endpoints, and denials logged instead of silent leaks. That is the missing production gate for founder stacks that already live on Claude Code or Codex.

securitysandboxagentsruntimepolicy

Similar · SkillSpector, Arcbox, Deepsec

X
SkillSpector

Sep 19, 2026

SkillSpector

★17,791Must watch
MCP

NVIDIA’s scanner that checks AI agent skills and MCP packs for prompt injection, exfil, and supply-chain risk before you install them.

Why it matters+

Founders are installing skill dumps and MCP servers into Claude Code, Codex, and Cursor every week. One bad skill can leak keys or widen agency. Deepsec-style repo audits catch bugs in your app; this catches the plugins you bolt onto the agent before they touch production.

securityagent-skillsmcpsupply-chain
X
Claude-Mem

Sep 18, 2026

Claude-Mem

★94,150Must watch
DemoAPIMCP

Persistent session memory for coding agents — it captures what the agent did, compresses it, and feeds the right bits into the next session.

Why it matters+

Most agent setups still wake up blank every time you open a new chat. Claude-Mem closes that gap for Claude Code, Codex, OpenCode, Hermes, Copilot, and more: tool use and decisions stick around as searchable memory, so the agent stops relearning your repo politics and unfinished threads from scratch.

memoryagentsclaude-codemcpdevtoolscontext

Similar · codebase-memory-mcp, OpenWiki, RTK

X
OpenResearch

Sep 17, 2026

OpenResearch

★4,648Must watch
Demo

Local-first workspace that turns Claude Code, Codex, OpenCode, or Cursor into research agents with parallel runs and git-native experiments.

Why it matters+

Most “research agent” tools are one long chat that forgets the last failed idea. OpenResearch treats research like software: each direction gets its own agent session and worktree, every run is archived to a commit, and logs, diffs, and papers stay tied to the work that produced them — so founders and builders can actually iterate on hypotheses instead of pasting another vague prompt.

researchagentslocal-firstclaude-codeexperimentsdevtools

Similar · Hyperresearch, Agent Reach, PRAXIST

X
RTK

Sep 16, 2026

RTK

★80,635Must watch

Rust CLI that sits in front of bash for coding agents and shrinks noisy command output before it hits the model context.

Why it matters+

Most agent token burn is not the prompt — it is fat git diffs, full test logs, and long ls dumps the model barely needs. RTK cuts that shell noise at the source so Claude Code, Codex, Cursor, and friends stay sharper and cheaper without a new model or a compression gateway.

agentsclitokenscodingrustdevtools

Similar · Paritok, Gigatoken, FreeToken

X
shadcn/lint

Sep 15, 2026

shadcn/lint

★1,210Must watch

Agent-first linter for Tailwind design systems: you define what components may change, agents get errors that name the fix from your variants and theme.

Why it matters+

Coding agents still restyle buttons with random padding and off-theme colors. Soft “taste” skills help, but they do not fail CI. shadcn/lint turns design contracts into Oxlint or ESLint rules so the same loop that fixes TypeScript also pulls UI back onto your system — cheaper correction rounds in their agent evals.

design-systemtailwindagentslintshadcnfrontend

Similar · Taste Skill, Hallmark, UI Skills

X
Open Code Review

Sep 14, 2026

Open Code Review

★23,888Must watch
DemoMCP

Alibaba’s open code-review CLI: deterministic pipelines plus an LLM agent, line-level findings, multi-language rules — built for real PR volume, not a generic “review this” skill.

Why it matters+

Coding agents ship more diffs than humans can review. Pure agent skills miss files, drift line numbers, and burn tokens. Open Code Review hard-codes selection, bundling, and positioning so the model only does judgment — higher precision at roughly one-ninth the tokens versus a general agent on their bench.

code-reviewagentsclicideveloper-toolsalibaba

Similar · meat, Code Review Graph, Deepsec

X
Worktrunk

Sep 13, 2026

Worktrunk

★7,335Must watch
Demo

Git worktrees made as easy as branches — built for running five or ten coding agents in parallel without them stomping the same checkout.

Why it matters+

Once agents can hold a long task, the bottleneck is the repo: one dirty tree, path fights, and half your day spent on git worktree add. Worktrunk turns that into short branch-named commands so founders can fan out Claude Code, Codex, or any CLI agent and still merge clean.

gitworktreesagentscliparalleldeveloper-tools

Similar · Orca, qm, OpenWorker

X
Hyperresearch

Sep 12, 2026

Hyperresearch

★2,788Must watch

Turns Claude Code into a deep research agent: a 16-step pipeline, cite-checked report, and a vault that keeps every source for the next run.

Why it matters+

Founders still paste a vague ask into chat and get a pretty report with soft citations. Hyperresearch treats research like a product: width sweep, contradiction graphs, adversarial critics, and a hard cite-check before anything ships. Sources land in a searchable vault so the next session starts smarter instead of re-fetching the same web.

deep-researchclaude-codeagentsknowledge-basecitationsresearch

Similar · Agent Reach, OpenWiki, Spec Kit

X
Spec Kit

Sep 11, 2026

Spec Kit

★135,358Must watch
Demo

GitHub’s open toolkit so you write the spec first, then any coding agent implements against it — specify, plan, tasks, implement, converge.

Why it matters+

Most agent runs still start from a vague chat and hope the code lands. Spec Kit flips that: the spec becomes the source of truth, agents fill the gaps, and converge checks the build against what you actually asked for. If you ship with Copilot, Claude Code, Codex, or Cursor, this is the missing process layer — not another chat wrapper.

spec-driven-developmentai-codinggithubcliagent-workflowplanning

Similar · ECC, OpenWorker, Harness Engineering

X
ARTEMIS

Sep 10, 2026

ARTEMIS

★1,040Must watch
DemoAPIMCP

Google’s open stack so coding agents drive a real Android phone — natural language in, taps and Logcat out, MCP into your IDE.

Why it matters+

Web agents already click browsers. Mobile still breaks most agent loops. If you ship Android, ARTEMIS lets Claude Code, Codex, or Antigravity install a build, walk the UI, grab screenshots and logs, and report failures without a hand-written Appium suite.

androidmobilemcptest-automationai-agentsgoogle
X
Atlas

Sep 4, 2026

Atlas

★3,144Must watch
DemoMCP

Source control for coding agents — every commit linked back to the session, prompts, tool calls, and reasoning that made it.

Why it matters+

Agents write a huge share of your code and leave almost none of the why. When you run Claude Code next to Codex, review becomes archaeology. Atlas keeps the session next to the commit graph so you can ask what changed, who did it, and why — months later.

source-controlcoding-agentscheckpointsmulti-agentclaude-codecodexlocal-firstmacos

Similar · meat, loopx, qm

X
Commerce Agents

Sep 3, 2026

Commerce Agents

★793Must watch
DemoAPIMCP

Anthropic’s open reference for two commerce roles on Claude: a shopping agent customers use in your app, and a merchant agent staff use in the back office.

Why it matters+

Most agent demos stop at chat over a product catalog. Real commerce needs search, cart, policy answers, inventory alerts, pricing drafts, and hard gates so writes never hit production without a human. This repo is the blueprint founders can fork instead of inventing safety and skill contracts from scratch.

shopping-agentmerchant-agentanthropicclaudeecommerceagent-sdkmcp

Similar · trycrm, Open Connector, eve

X
CodeBurn

Sep 2, 2026

CodeBurn

★10,522Must watch
DemoMCP

Local-first ledger for AI coding spend — tokens and dollars across Claude Code, Cursor, Codex, Gemini, and 40+ tools, by model, project, and task.

Why it matters+

Provider bills only show a total. Founders running three coding agents never see which model burned the budget on chat instead of code, or which project is eating the month. CodeBurn turns session files already on disk into that breakdown so you can cut waste before the next invoice.

cost-trackingobservabilityclicoding-agentslocal-firstmcp

Similar · Tracely, Paritok, FreeToken

X
Utopia

Sep 1, 2026

Utopia

★1,280Must watch
DemoAPI

Open-source enterprise world model — bitemporal knowledge graph plus ontology so agents reason on governed facts, not a flat vector dump.

Why it matters+

Most agent stacks still treat company knowledge as chunks in a vector store. Utopia puts time and ontology in the base layer: facts carry validity windows, conflicts surface for review, and decisions leave an audit trail on hardware you control.

world-modelknowledge-graphontologyagentsself-hostedrust

Similar · OpenViking, codebase-memory-mcp, Graphify

X
ArcBox

Aug 31, 2026

ArcBox

★1,489Must watch
API

Open-source OrbStack-class runtime for Mac — drop-in Docker plus Firecracker microVMs so coding agents run with a real kernel boundary, not hope and allowlists.

Why it matters+

Founders still hand Claude Code and friends the whole laptop, or they rent brittle cloud sandboxes. ArcBox makes isolation the default on the machine you already use: disposable microVMs with their own kernel, filesystem, and network, while still speaking Docker for everyday containers.

sandboxmicrovmagentsdockermacosrust

Similar · OpenWorker, Cloudflare OS, Destructive Command Guard

X
Weave Router

Aug 30, 2026

Weave Router

★2,891Worth testing
DemoAPI

Drop-in model router for agent stacks — scores each action and sends it to the right model in under 50ms so you cut spend without rewriting tools.

Why it matters+

Most coding agents burn one expensive default model on every tool call. Weave Router sits in front of Claude Code, Codex, Cursor, and friends, picks the model per action with a local embedder, and keeps provider keys on your box. Founders get continuity plus a real cost lever, not another key-rotation dashboard.

model-routerai-gatewayagentsclaude-codecost

Similar · OmniRoute, Paritok, OpenWorker

X
OpenSEO

Aug 29, 2026

OpenSEO

★14,303Worth testing
DemoMCP

Open-source Semrush/Ahrefs-style SEO suite you control — keyword research, ranks, backlinks, audits, AI visibility — with MCP and agent skills baked in.

Why it matters+

Founders still rent bloated SEO suites or bounce between half-finished scripts. OpenSEO puts the core workflows in one app you can self-host or pay-as-you-go, then hands the same data to Claude Code, Hermes, and friends over MCP so research stops living only in a browser tab.

seomcpagent-skillsself-hostedmarketing

Similar · Firecrawl, Career Ops, Agent Reach

X
Praxist

Aug 28, 2026

Praxist

★906Must watch
Demo

Autonomous R&D loop for projects that already run and score: parallel research peers, durable evidence, and generation-to-generation synthesis until the budget ends.

Why it matters+

Most agent demos chat once and forget. Praxist treats research as a multi-generation process with evaluators, evidence lanes, and a planning panel — so long campaigns stop re-learning the same dead ends and you can inspect why a claim stuck.

autonomous-researchr-and-dagentsevalscodex

Similar · LoopX, OpenWorker, qm

X
Taste Skill

Aug 27, 2026

Taste Skill

★81,157Must watch
Demo

Anti-slop agent skills for frontend: layout, type, motion, and spacing rules plus image-gen boards so coding agents stop shipping generic UI.

Why it matters+

Agent UIs still default to the same safe look. Taste Skill is a portable skill pack (not a component library) that forces a design language, hard checks, and redesign audits before code lands — install once, use in Claude Code, Codex, or Cursor.

agent-skillsfrontenddesignclaude-codecodex
X
ThreeUI

Aug 26, 2026

ThreeUI

★4,019Worth testing
Demo

Open Three.js + React UI catalog you can browse live, copy Community source, and drop into a product without a login wall.

Why it matters+

Most WebGL component kits hide the good stuff behind accounts. ThreeUI Community ships the same shell, live previews, variants, and full free source — so founders can evaluate motion/UI building blocks before paying for Pro.

threejsreactwebglui-componentsshaders
X
FreeToken

Aug 25, 2026

FreeToken

★6,139Worth testing
DemoAPI

Run frontier-scale MoE models on a gaming PC — edge-native serving with agent-aware cache so tool edits do not recompute the whole context.

Why it matters+

Coding agents burn tokens rewriting long contexts after every tool call. FreeToken is built for that loop: it serves huge open MoE weights on consumer GPUs and keeps semantic checkpoints so agentic edits can reuse work instead of restarting the prefill tax.

local-llmmoeinferenceedgeagents

Similar · Colibri, Needle, Paritok

X
Cumora

Aug 24, 2026

Cumora

★2,987Worth testing
Demo

Team chat where AI agents sit on the same roster as people — DMs, rooms, Kanban, calendar, and real email.

Why it matters+

Most agent tools still feel like a private chat with one bot. Cumora treats agents as teammates who claim work, coordinate without stomping each other, and can run on your cloud or your own Claude Code / Codex machine without giving the server your API keys.

agentsteam-chatmulti-agentbyoa

Similar · Rakazo, OpenWorker, LoopX

X
Hister

Aug 23, 2026

Hister

★2,213Must watch
DemoMCP

Your own search engine for pages you visit and files you keep — full text, optional semantics, and MCP for agents.

Why it matters+

Founders lose hours re-finding half-remembered docs, tabs, and PDFs. Cloud bookmark tools are shallow and leaky. Hister keeps the index on machines you control and lets coding agents query it the same way you do, so personal research becomes something the agent can actually use.

searchprivacymcppersonal-knowledgeself-hostedai-agents

Similar · Agent Reach, OpenViking

X
Tracely

Aug 22, 2026

Tracely

★924Must watch
DemoAPIMCP

Trace-native CI for agents: production failures freeze into hermetic cases that block the PR for $0 replay.

Why it matters+

Most agent eval stacks make you invent datasets. Real breakage already shows up in production traces with the exact tools and model turns. Founders shipping agents need those failures to become gates, not another dashboard tab you check after the customer complains.

ai-agentsobservabilityci-cdevalsllmopsmcp

Similar · Deepsec, meat

X
watermarks-remover

Aug 21, 2026

watermarks-remover

★16,207Must watch
API

Agent skill + local HTTP service that strips multi-vendor AI provenance marks from text and files you own.

Why it matters+

Founders ship drafts, assets, and decks that accumulate invisible Unicode, C2PA bags, and model metadata they never asked for. You need a hygiene layer that agents can call, not a one-off hex edit. This turns provenance cleanup into inspect and clean endpoints with a real skill path.

aiprovenancec2paagent-skillprivacycontent

Similar · no-ai-slop, Hallmark

X
Portless

Aug 20, 2026

Portless

★10,662Must watch
Demo

Swap localhost:3000 chaos for stable named URLs like https://myapp.localhost — built for humans and agents.

Why it matters+

Every multi-app founder still juggles random ports, broken deep links, and agents that forget which service is on which number. Portless makes local URLs durable so demos, worktrees, and agent tools keep the same hostnames after restarts.

developer-toolslocal-devagentsdxproxy

Similar · CLI-Anything, OmniRoute

X
OpenViking

Aug 19, 2026

OpenViking

★29,654Must watch
DemoAPI

A context database for agents: memories, docs, and skills live under viking:// so the model browses context like files, not a black-box vector dump.

Why it matters+

Agent memory is usually a vague retrieval bag that burns tokens and is hard to debug. Founders shipping multi-session agents need a real store they can inspect, tier, and evolve. OpenViking turns context into something you can ls, tree, and audit with a path history.

ai-agentscontextmemoryragskills

Similar · codebase-memory-mcp, Paritok

X
career-ops

Aug 18, 2026

career-ops

★65,021Must watch
Demo

Open-source AI job search that runs inside your coding CLI: score listings, tailor CVs, track applications locally.

Why it matters+

Companies already filter candidates with AI. Founders and builders still chase roles with spreadsheets and spray-and-pray apps. career-ops turns Claude Code, Codex, OpenCode, and peers into a local filter that ranks fit, kills ghost jobs, and only spends energy on offers that clear a real bar.

ai-agentsjob-searchclicareerautomation

Similar · CLI-Anything, OpenWorker

X
Needle

Aug 17, 2026

Needle

★6,815Must watch
Demo

A 14MB foundation model that does tool calling and structured extraction on phones, wearables, and robots.

Why it matters+

Most on-device models still chat. Founders shipping voice, home, and robot products need tiny models that call tools with JSON they can trust, stay offline, and fit in tens of megabytes of RAM. Needle turns that into a single pip package and a single binary engine.

on-device-aitool-callingedgelocal-llmstructured-output

Similar · Bonsai, Colibri

X