Developer-Tools

39 tools reviewed

7.6
agent-memory Review 2026 — A File-Truth Long-Term Memory Runtime for AI Agents: Rebuildable Indexes, Sleep-Time Manage, and a 52.9% LongMemEval-S Result

agent-memory is a local-first, agent-agnostic long-term memory runtime (created 2026-09-01, 160+ stars, 71 commits in its first week) built on an unusual thesis: files are the truth and every index is a rebuildable cache — memories are plain markdown files in one store, the SQLite index beside them can be deleted at any time, and rm -rf .index && mem rebuild loses zero knowledge, enforced by a test. Writes are triggered at conversation boundaries, not at the agent's discretion, and an independent sleep-time Manage pass consolidates, ages and forgets by value while deletion only ever arrives as a proposal you confirm. There is no LLM client inside the library — zero API keys, judgment borrowed from the host agent's own CLI (Claude Code, Codex CLI or Hermes share one store). In its own bounded-haystack LongMemEval-S experiment with 120 episodes it scored 127/240 (52.9%) versus 86/240 (35.8%) for MemCore and 7/120 (5.8%) with no memory. This review covers the file-truth storage model, the Manage layer, the three recall tracks, the MCP/CLI/hooks wiring, the benchmark's honest framing, and the notable absence of a license.

7
Choruz Review 2026 — A Local-First Slack for Humans and AI Coding Agents, Where Each Agent Runs a Real CLI in Its Own Workspace

Choruz (inclusionAI, created on GitHub 2026-09-02 with 305+ stars in five days, MIT license, v0.1.0 developer preview) is a local-first collaboration space where humans and AI agents work together in a Slack-like interface — direct chats, groups, mentions, threads and channel task boards — while every agent runs a real CLI (Claude Code, Codex, Pi, Grok, OpenCode, or a webhook agent) in its own workspace directory or git worktree. Agents are not simulated in a sandbox: the platform spawns the actual terminal or headless CLI on your machine (or an SSH runtime host), talks to it through a documented agent protocol — a [choruz-incoming] envelope, a $CHORUZ_SEND helper, a Maildir-style outbox under .choruz-outbox/new/ and CLAUDE.md/AGENTS.md instruction files — and routes work between humans and agents via an event-sourced Postgres pipeline with CDC intake, leased command dispatch, idempotent writes keyed by client_msg_id and turn_id, and per-device sync cursors. Built as a Rust modular monolith (Cargo workspace crates, a choruz-api-gateway Rust service, a choruz-pipeline worker and a Next.js web client), it adds an AI Manager agent that tracks workflow state, cron-scheduled agent jobs, Slack/Telegram bridges, a kanban for channel tasks, remote-control over SSH or a Cloudflare Worker relay, and plugins. This review covers the agent protocol, the architecture, the developer-preview friction of a four-process local stack, and how Choruz compares with hosted agent workspaces and plain terminal multi-agent setups.

7.4
codenotch Review 2026 — A macOS Notch That Shows How Much of Your Claude Code, Cursor and Codex Limits You've Burned

codenotch is an open-source macOS app (created 2026-09-05, 679+ stars and 90 forks in its first two days, MIT license) that pins a small black notch to any screen edge showing how much of each AI coding assistant's usage limit you have burned — Claude Code, Cursor, Codex, Antigravity and GLM — with a spinning arc when a session is busy and a pulsing amber ring when one is blocked waiting on you. Instead of signing in anywhere, every provider adapter borrows the credential or session the owning tool already holds: Claude Code's OAuth token against the same endpoint its own /usage panel reads, Cursor's signed-in session from its local SQLite state, Codex's app server asked live for rate limits, Antigravity's local language server, and GLM's Z.ai Coding Plan monitor endpoint. Each adapter declares a Fidelity tier (official, derived or manual) so the UI never presents a guess as a vendor-published number, polling backs off against 429s with a persisted deadline, and every failure degrades to a visible stale/needsAuth/error state instead of an invented percentage. This review covers the provider matrix, the notch UI and its state bands, the honest caveat that no vendor publishes a clean usage API, and how codenotch compares with alternatives like Honey for Devs and manual /usage checks.

8
Composio Review 2026 — 1000+ Toolkits for AI Agents

Hands-on Composio review 2026 — tested connecting AI agents to 1000+ tools, real integration benchmarks, pricing breakdown ($20/mo to enterprise), and how it compares to native MCP servers and Zapier Central.

7.4
Kitter Review 2026 — A Rust Desktop App That Keeps One Agent-Skill Library and Links Only What Each Project Needs

Kitter (what1f, created September 2, 2026, 287 stars, Apache-2.0) is a Rust and GPUI desktop application and a matching CLI for managing Agent Skills across projects: one maintained library, per-project linked installations, a live view of the skills every agent actually discovers (including ones Kitter did not install), and a per-agent token-cost estimate. This review covers the library/install/project model, the managed-links approach that avoids update drift, the CLI surface, where skills are stored on each platform, the unsigned macOS build and the unvalidated Linux desktop, and the honest limits of a three-week-old project whose token numbers are estimates.

7.6
Munder Difflin Review 2026 — Run an Office of AI Clones on Your Own Laptop

Munder Difflin is a free, open-source multi-agent harness that runs 'an office of your clones' on your laptop, wrapping Claude Code, Codex, and 10 other CLI agents. We review the 3,651-star project, its deterministic office simulation, E2E-encrypted clone messaging, and the Office-IP controversy from its 110-comment Hacker News thread.

7.4
okf-agent-memory Review 2026 — Git-Native Agent Memory That Costs Nothing to Query

okf-agent-memory (created September 5, 2026, MIT, 549 stars) is a git-native persistent memory layer for AI coding agents built on Google's Open Knowledge Format v0.2. This review covers the knowledge/ bundle of Markdown plus strict YAML frontmatter, the zero-dependency Go CLI and its embedded stdio MCP server, the sub-300-microsecond in-memory BM25 search that replaces embedding API calls, progressive disclosure and its 80 percent token-reduction claim, the ten-point agent convention, the trust tiers that separate generated from verified knowledge, the adversarial security hardening in v0.1.4 and v0.1.5, and the honest limits: it is a convention and tooling kit, not an auto-capturing memory, and lexical BM25 is not semantic recall.

7.4
skill-cabinet Review 2026 — A Local Catalog for Every Agent Skill Installed on Your Machine

skill-cabinet is a free MIT-licensed local catalog for the agent skills installed on your machine: it scans user-level drawers like .agents, .claude, .codex and .cursor (including plugins), lets you read each skill's body and YAML frontmatter, filter by drawer, risk, symlink status or copies, follow GitHub origins, and delete skill folders from disk. One command — npx skill-cabinet — starts a localhost server (port 3781) and opens a browser UI with search, j/k keyboard navigation, theme switching and static-risk analysis. This review covers how the drawer model works, the safety model around deletion, the honest limitations (single-machine scope, no cloud sync, deletion is permanent), and how it compares to eyeballing ~/.claude/skills in a file manager.

7.3
tokentab Review 2026 — The Local-Only CLI That Prices Your AI Coding Sessions

tokentab (created September 7, 2026, MIT, 852 stars) is a local-only CLI and web dashboard that reads the session logs Claude Code, Codex and Gemini CLI already write to disk and turns them into token counts and dollar costs by model, project, day and kind of work. This review covers the three log formats it parses, the hand-kept pricing table and its fuzzy model matching, the caching fixes that stop double-counting, the standard-library web dashboard on localhost:4747, the heuristic activity classifier, and the honest limits: a single-commit project, unfinished Cursor support, best-effort prices and a 'guess, not gospel' activity signal.

7.9
x64dbg-MCP Server Review 2026 — Agentic Reverse Engineering in Zig, 84 MCP Tools, Zero Dependencies

x64dbg-MCP Server is a native MCP plugin for the x64dbg debugger written in Zig: 84 MCP tools, 22 event callbacks, dual Streamable HTTP + SSE transport, mandatory bearer auth, and a single cross-compiled binary with zero dependencies. It rocketed to 1,209 stars in three days. We review what it can do — breakpoints, memory patching, OEP detection, tracing — and where agentic reverse engineering still hurts.