580 tools and products — trending open source, and what gets used in AI and other work.
DevAssure is an AI-powered software testing platform whose O2 agent reads code changes, maps potentially affected application flows, and validates them automatically, including through browser interactions based on plain-English expected behavior. It generates scriptless, self-healing acceptance tests and reports results for pull requests.
Arcade is a runtime for the Model Context Protocol (MCP) that helps AI agents take actions in real systems. It provides authentication and permissions management, reliable tool calls, and audit trails.
Muse Code is an AI-powered terminal coding agent for inspecting repositories, investigating code issues, generating HTML reports, modifying code, spawning parallel sub-agents, and auditing pull requests.
An evaluation toolkit for Strands AI agents that measures off-script behavior, violations of behavioral guardrails, and other incorrect behavior against test and live data. It is associated with Strands Agents, an open-source AI agent SDK for Python and TypeScript.
Vois is a local AI voice-generation and audio-processing studio for writing, casting, voice selection and cloning, multi-speaker scripts, editing, and mastering. It provides local text-to-speech, more than 100 natural voices, audio tools for podcasts, audiobooks, and video, and integrations with AI agents; its Pro plan adds Omni, more than 600 languages, and Voice Design.
Vapi is a developer platform for building, testing, and deploying conversational voice AI agents. It provides a cascaded voice-agent architecture and supports customer-supplied text-to-speech servers for specialized applications.
Pipecat is an open-source Python framework maintained by Daily and the community for building real-time voice and multimodal conversational agents. It orchestrates audio and video, transports such as WebSockets and WebRTC, speech recognition and text-to-speech services, speech preprocessing, turn detection, model calls, tool calls, and streamed audio through composable conversation pipelines. Pipelines can operate as single agents or as multi-agent systems whose specialists hand off, fan out in parallel, or coordinate over a shared bus locally or across processes and machines. The project also provides a CLI for scaffolding, monitoring, and deploying agents, along with client SDKs and related tools for structured conversations and pipeline debugging.
Smallest AI is a voice-AI platform offering text-to-speech, speech-to-text, speech-to-speech, and orchestrated voice-agent models for real-time deployments.
Deepgram is a speech AI platform providing Speech-to-Text, Text-to-Speech, and Voice Agent APIs. Its services support real-time transcription and voice applications, including the speech transcription component in deployment architectures.
Gradient is a framework for training research agents with reinforcement learning. It exposes the agents’ searches, citations, tool calls, and evaluation results.
Ambient Context is a macOS menu bar app that records the text of the focused window for use by an LLM or other agent. Through the macOS accessibility tree, it reads window text every few seconds and appends deduplicated blocks to one plain Markdown file per day, including the document path or URL and an AGENTS.md file describing the format. It does not use screenshots or video and, in the current build, makes no network calls, accounts, servers, telemetry, or bundled model; password fields, password-manager and private-browsing windows are excluded, and credentials, API keys, and card-shaped numbers are scrubbed before writing. The project is early and unsigned, requires macOS 14 or later on Apple Silicon, and currently must be built from source with Node, Rust, and Xcode Command Line Tools.
System Atlas is an agent skill from Inkboard that turns architecture discussions into an explorable isometric atlas. A single data file serves as the source for an interactive map and a generated SYSTEM.md text view, keeping structures, execution flows, design decisions, and open questions synchronized. The map supports hoverable structures, pinned views, drill-down into execution steps, pan and zoom, progressive-disclosure chapters, and inspectable data packets showing routes and representative JSON payloads. The generated SYSTEM.md includes a decisions table with ADR links, structure descriptions and steps, flow tables, and an index of questions tracked by stable IDs and states such as open, resolved, or routed. The skill can be installed with `npx skills add inkboard/system-atlas`. It generates a self-contained HTML map with no build step or runtime dependencies.
agenttrail is a local, open-source observability layer for AI coding agents. It watches plans, tool calls, file changes, and progress from Claude Code, OpenAI Codex, Cursor, or other agents that edit files, then presents them as a live project map with repository components, dependency arrows, progress, and working, blocked, or completed states. It compares declared agent intent with observed filesystem activity, including renewed changes to supposedly finished work, and displays the current run, task list, streaming tool activity, elapsed time, recent calls, session plans, and a live repository tree. The tool runs locally through commands such as `npx agenttrail`, with no account, global installation, or telemetry. `npx agenttrail init` adds the agenttrail convention to CLAUDE.md and AGENTS.md, creates a starter PLAN.md, and installs additive local Claude Code hooks; the resulting plan and file activity are used to generate a map of roughly 5–9 repository components with dependencies and verifiable statuses. It supports one daemon per repository, a shared board tab switcher, restart via `npx agenttrail up`, and login autostart through `npx agenttrail autostart`.
backpass is a local-first command-line tool that analyzes coding-agent session transcripts and proposes evidence-backed edits to an agent memory file such as AGENTS.md or CLAUDE.md, plus project skills. It treats the memory file as weights, agent sessions as forward passes, and their on-disk transcripts as a loss signal, then collects sessions for the current repository, distills failures, aggregates repeated instruction changes, and produces diffs and skill extractions. The tool reads transcript stores from seven agent harnesses directly from disk, requires evidence from at least two independent sessions for a new instruction, and limits each run to at most five edits. Proposed changes include verbatim session quotes; analysis never writes files, while `backpass apply` presents each edit for human acceptance or rejection before applying it. It is distributed through npm or npx, requires Node.js 22.5 or later and `acpx`, uses no API keys of its own, and sends no transcripts to a separate service; obvious secrets are redacted before model calls through an already authenticated harness.
Sentio is an email inbox API and multi-tenant mail server for AI agents, developed by Truespar. It gives each agent a real email address, authenticates and scans inbound SMTP mail, scores and routes it, then delivers the message as a structured webhook; agents send threaded replies through a REST API, which signs outbound mail with DKIM, queues it, and delivers it over SMTP. Tenant isolation covers domains, mailboxes, API keys, rate limits, suppression lists, sending reputation, and spam profiles. The Rust service implements inbound and outbound mail infrastructure including DKIM, SPF, DMARC, ARC, MTA-STS, DANE, and three-tier anti-spam, and includes an MCP server for exposing email as native agent tools. The repository provides Docker deployment with PostgreSQL, Redis, NATS/JetStream, MinIO, ClamAV, and rspamd, plus an API reference and testing UI.
OwnMem is a repository-owned memory system for AI coding agents. It stores project knowledge as reviewable Markdown in a .ownmem/ directory, shared through Git so it can be cloned, reviewed, and rolled back across Claude Code, Codex, Cursor, Gemini CLI, Grok CLI, and other hosts. Its local recall pipeline compiles schema, graph, lifecycle, and evidence checks into an immutable, content-addressed snapshot. Five deterministic candidate lanes—exact matching, BM25F, n-gram, fuzzy, and graph retrieval—are fused locally; embeddings are optional and remain disabled by default until local evaluation supports them. Four delivery gates assess relevance, epistemic validity, task applicability, and action risk, allowing normal delivery, advisory output, quarantine, or abstention. OwnMem separates repository memory from trust receipts and uses quotas, duplicate gates, lifecycle rules, audits, and reversible changes to bound unattended evolution. Its rules allow automation to promote only replay-proven, quota-bounded R0 retrieval metadata, while prose, policy, and higher-risk changes become review material. The repository describes the project as local, deterministic, git-native, evidence-governed, and licensed under Apache-2.0.
MonoCode is a desktop GUI for running coding-agent command-line interfaces in tabbed sessions. Each tab represents a session, and a shared composer provides the prompt input while the selected CLI uses the user's own logged-in subscription; MonoCode does not sell tokens or act as a model provider. It supports Claude Code, Codex, Cursor CLI, OpenCode, Pi, omp, and fx when they are installed and authenticated. The application supports macOS and Linux, with an Apple Silicon macOS disk-image distribution and source builds requiring Node.js 20 or later and a current stable Rust toolchain. It is licensed under the MIT License and is described by its repository as an early project that may contain bugs.
open-sheet is a spreadsheet framework for coding agents that represents workbook models as React and TypeScript source. Instead of writing fragile A1-style cell addresses, models can use named references such as `ref('pl').column('revenue')`, which open-sheet resolves at compile time while handling cell addressing, formula references, recalculation, and formula-tree previews. It exports live formulas to XLSX, as well as CSV, HTML, and PDF formats, and reports cells it cannot compute as `#NOT_EVALUATED` rather than emitting a plausible number. The project provides an `npx @open-sheet/cli` initializer, is distributed under the MIT license, and is hosted at open-sheet.dev.
Halofy is an open-source governance layer between organizational knowledge and AI agents. It resolves actor identity, roles, namespaces, and source scope on the server; enforces access policies; and records provenance, supersedence, append-only audit history, policy decisions, and signed erasure certificates for reads and writes. The repository exposes governed memory and context operations over MCP and HTTP, including writing, reading, searching, assembling, faulting, statistics, forgetting, policy, sharing, pinning, export, and manifest operations. Postgres with pgvector is the operational authority, while retrieval engines connect through an ACL-scoped read-only driver interface; encrypted, Git-versioned knowledge provides a cold tier, and embedded PGlite is the zero-setup local default. It includes connectors for filesystem, Postgres, Obsidian, and manual CSV import, plus an offline demo and hermetic test lane using deterministic stub components.
Kubernetes-sigs Agent Sandbox creates isolated, disposable environments for running untrusted or agent-generated code. It is intended to separate such code execution from other workloads.
Gemini Live API is a Google Gemini API for building voice agents that listen and respond through live audio-to-audio interactions. It supports real-time responses, interruption handling, audio streaming, and calls to application tools.
live-dj is an open-source voice-agent demo in which users talk to Mira, a late-night radio DJ, ask her to play music, and interrupt her while she speaks. It uses the Gemini Live API through the raw `google-genai` SDK rather than an agent framework. The browser handles microphone capture, 16 kHz audio input, 24 kHz playback, client-side barge-in, and music ducking. The server maintains one asynchronous Live API session per browser and runs the core loop of opening a session, sending microphone audio, receiving streamed voice responses, and playing them. The full application adds Mira's persona, transcripts, music-tool dispatch, and controls for playing playlists or tracks, skipping, and pausing; the tools return immediately so the voice response does not stall. A minimal backend exposes the 39-line voice-only primitive, while the full server demonstrates the complete DJ application. The repository also documents a per-turn `session.receive()` behavior that requires an outer loop for continuing conversation, and shows that microphone audio must be sent through `send_realtime_input` rather than `send_client_content`. It runs locally with `uv`, Uvicorn, and a Gemini Developer API key. The repository includes a browser client, four dream-pop tracks, persona assets, and examples of the relevant implementation pitfalls.
Free Claude Code is an independent open-source local proxy for routing Claude Code, Codex, Pi, OpenCode, and other coding-agent requests to selected free, paid, subscription, or local model providers while preserving their existing APIs. It provides a searchable model catalog and an Admin UI for configuring providers, supports automatic fallback to another configured model after provider retries are exhausted, and can be launched from a terminal, desktop app, IDE, Discord, Telegram, or phone. Optional RTK filtering reduces common terminal-output tokens, while voice input can use local Whisper or NVIDIA NIM transcription. The project states that it is not affiliated with or endorsed by Anthropic, and that provider free-tier availability and limits may change.
Agent Plugins for AWS is an AWS Labs collection of plugins for AI coding agents, helping them architect, deploy, and operate on AWS. Supported agents include Claude Code, Codex, and Cursor. Each plugin can package agent skills—structured workflows and best-practice playbooks—alongside MCP servers that provide access to live documentation, pricing data, and other APIs; hooks that validate changes or trigger workflows; and references containing documentation and configuration defaults. The repository describes the plugins as reusable, versioned capabilities intended to reduce prompt context and standardize agent behavior. It warns that generative AI can make mistakes and recommends reviewing generated code, costs, security, and credentials. The repository also identifies Agent Toolkit for AWS as the successor for production use, while stating that this project continues to work and accept contributions.
Droids are Factory’s autonomous software-development agents, designed to carry out engineering work for enterprise teams. They are an AI software-development product rather than a general-purpose category.
OpenHuman is a local-first personal AI system from tinyhumansai that combines persistent memory, agent orchestration, and research tools. It stores a user's data as scored Markdown trees in SQLite on the local machine and mirrors the result to an editable Obsidian vault; its TokenJuice component compresses tool output before it reaches the language model. Its orchestration layer runs checkpointed, durable agent workflows and worker fleets on graphs, with triggers, approvals, steering, halting, and replay support. The system also provides web search, scraping, coding tools, a browser, native voice through in-process Whisper, model routing, messaging integrations, and support for provider keys or fully local Ollama models. The repository describes it as an early beta under active development.
Markdown-based agent skill that rewrites AI-sounding text to read human-written without changing what it says — works with any agent that supports skills. Rewrites against the 35 patterns from Wikipedia's 'Signs of AI writing' (WikiProject AI Cleanup): a first pass free to restructure, then a check of the draft against those patterns and the original claims before rewriting what still sounds artificial. Its rules bar invented facts — names, numbers, dates and quotes must come from the source or the writer — and preserve a writer's personal style (or follow a provided sample). The skill shows its work: the first rewrite and a critique of what still sounds artificial precede the final version.
TURNR is an AI marketing agent for home service businesses. It learns from customer calls, the business website, and existing content, then creates social media posts, video hooks, SEO blog posts, and ad copy, with a weekly email summarizing the generated materials.
MemoraX Code is a memory plugin and shared memory layer for AI coding agents, developed by MemoraX. It integrates with Codex, Claude Code, CodeBuddy/WorkBuddy, DeepSeek Harness, and OpenCode to retrieve relevant context for new tasks and capture reusable knowledge from completed work. Its memory is divided into Coding Memory for engineering lessons and design decisions, Repo Memory for repository structure and history evidence, Personal Memory for user preferences, and Procedure Memory for reusable steps and validation gates. Background writeback extracts selected knowledge from trusted workspace turns, while the bundled skill and CLI support explicit search and memory operations; repository and personal/procedure content are maintained under the documented local storage boundaries. The package is distributed through npm and requires Node.js 20 or later, with Python 3 required for Repo Memory operations. Cloud-backed search and storage require a MemoraX account, while guest mode is available for a limited period. Local trace capture is enabled by default for supported clients and may retain prompts, responses, recalled memory, reminder text, and local paths; the project documents settings for metadata-only capture or disabling traces. The repository is licensed under the MIT License.
sepia is a portable Agent Skill for Claude Code, Codex, Grok Build, and Antigravity that revises AI-assisted fiction and professional prose at the narrative-architecture and discourse levels, rather than only changing word choice. For fiction, its three-pass protocol addresses narrative architecture, discourse flow, and surface style; its rules cover issues such as overly tidy causality, explained themes, linear time, sparse character networks, uniform emotional rendering, templated paragraph flow, and predictable endings. Professional-writing rules are matched to document venues including release notes, pull-request and issue replies, postmortems, tickets, and technical articles. The package provides write, review, refactor, and recreate operations, plus a general router. Review diagnoses without editing, refactor makes minimal in-place changes, and recreate rewrites from source facts and intent. It includes a 30-feature diagnosis rubric, model-specific fingerprint corrections, shared professional-prose checks, and research references. The repository distributes the skill as a plugin package with native installation paths for the supported tools and an alternative Skills CLI installation. It is licensed under the MIT License.
ask is a command-line tool for asking AI questions from a terminal without giving an agent control over the project. It runs an already installed and authenticated Codex, Claude Code, Pi, or OpenCode CLI in read-only mode in the current directory, prints answers to stdout, and sends prompts and errors to stderr. One-shot questions use `ask [QUESTION...]`; interactive sessions can be continued with `ask -c`, with turns, agent choices, model settings, reasoning settings, and the underlying agent session ID saved for later use in the folder. It can also reopen saved sessions, configure defaults, and upgrade itself. The installer supports Intel and ARM Macs and x86_64 and ARM64 Linux, and the project is licensed under MIT.
Goldie is an app-store screenshot and preview-video generator for iOS apps, designed for coding agents and human users. It uses Argent flows to replay app interactions in an iOS simulator, captures the results, adds device bezels, backgrounds, headlines and other design elements, joins the clips into preview videos, and checks the output against Apple's upload rules. It is framework agnostic and can drive SwiftUI, UIKit, Flutter, React Native and Kotlin Multiplatform apps through the simulator. The CLI provides commands to check tools, simulators and flows, run the capture-and-render pipeline, and open a browser-based studio for editing backgrounds, templates, bezels, fonts and per-tile copy. Design settings are saved in goldie.design.json, while generated assets are written to an output directory per locale. Goldie requires macOS with iOS simulators, Node.js 20 or newer and ffmpeg; its previews must be 15 to 30 seconds long. The project is sponsored by Software Mansion, the creator of Argent.
Open Pstack is an unofficial, plugin-based workflow kit that brings Lauren Tan's pstack to Claude Code and Codex. It gives coding agents engineering rules, task-specific workflows, focused skills, and small local tools rather than providing a new model or hosted service. Its main poteto-mode workflow reads a task, selects an appropriate workflow, learns how the existing system works, compares designs when needed, favors small changes, and can use multiple models to challenge important decisions. It runs the code and checks behavior as a user would instead of stopping at passing tests, then can continue through review and continuous integration to prepare a pull request. Additional skills cover system explanation, architectural decisions, competing implementations, design interrogation, verification-skill creation and maintenance, pull-request supervision, and reflection. The repository supports installation as a Claude Code or Codex plugin and shares the same skill set between those applications. It is distributed under the MIT license and tracks the upstream pstack project while adapting its skills for Claude Code and Codex.
Whip is a Go-based coding-agent harness distributed as a single binary with no runtime dependency. It runs an LLM tool-use loop for bash commands, reading, writing, editing, and subagents, alongside an interactive Bubble Tea terminal interface. The harness supports parallel tool calls, streaming, background subagents, MCP servers, and provider-routable models with live discovery from provider catalogs; any OpenAI-compatible endpoint can be used as a provider. Prebuilt checksum-verified binaries are available for Linux and macOS on x64 and arm64, and it can also be installed from source with Go.
Open Steps is an open-source pack of agent skills that translates coding-agent output into plain-language reports, verdicts, questions, and next steps. Its skills include done-or-not reports, step-by-step instructions for nontechnical users, simple explanations of agent questions, premortems for hard-to-reverse decisions, verification of another session's claims, next-task recommendations, and plain-language rewrites. The pack is built and measured primarily for Claude Code, where it installs as a plugin with two shell hooks. The session-start hook injects the routing table and the latest report, while the stop hook checks for completed work and requests a report when appropriate. Reports are stored outside project repositories. Skills and routing instructions can also be installed for Codex, Cursor, and Gemini CLI, although the repository notes that hook support differs across those tools. Open Steps separates measured results from assumptions and marks unchecked information as "not checked." It is free software under the MIT license and is developed by Pavlo Kharmanskyi.
headcount is an agent organization for Claude Code, structured as a company with independently installable departmental plugins containing named skills for engineering, business, and operational work. Projects install only the departments they need, address skills as department:skill to avoid name collisions, and can invoke skills directly or have them load when a request matches their territory. Each department includes an agent charter for delegation as a subagent with an exclusive write surface. The repository organizes agents by ownership boundaries rather than topic, and documents cross-department workflows, decision logs, surface ownership, and an interactive searchable organization chart. Reviewer-class security and legal-risk departments can block work under review. headcount is distributed as a Claude Code plugin marketplace and is licensed under the MIT License. The repository identifies Chris Brock as its builder and includes a validation script and CI checks for skill front matter, manifests, department references, license text, and ownership-surface consistency.
Lemmalog is a Datalog engine for LLM-agent memory, distributed as a Rust crate with an MCP server, REPL, and agent skill. It treats extracted statements as provenance-tracked base facts, then uses runtime-parsed stratified Datalog to derive temporal projections, closures, contradiction candidates, relevance relationships, and aggregates. The engine supports seminaive incremental fixpoint evaluation, stratified negation with negative-cycle rejection, bi-temporal facts, confidence and provenance annotations, proof trees through why(), scoped retraction recomputation, demand-driven ask_deep queries using magic sets, and persistence of episodes, rules, and base facts while rebuilding derived relations on load. Its AgentMemory facade connects an extractor to deterministic ADD, UPDATE, NOOP, and escalation policies, and assembles budgeted contexts from distilled facts and their source episodes. The MCP server exposes the engine to agent harnesses such as Claude Code and Kimi CLI through tools for observing facts, querying, explaining proofs, retrieving context, saving state, installing rules, and running hypotheticals.
Epic Infographics is an open-source skill for AI agents that generates data-driven infographics as HTML/CSS scenes rather than stock dashboard templates. It guides the agent through audience and story-angle selection, visual-metaphor and design-language selection, scene composition, chart construction, rendering, and self-review. Charts are computed from explicit arithmetic, palettes undergo color-blind-safety checks, and a headless preflight script checks text collisions, clipping, canvas boundaries, and readable sizes. The skill can render still PNGs and animate the same HTML through CSS keyframes into MP4 or GIF output using Playwright, headless Chromium, and ffmpeg; its motion workflow treats animation as a layer on an approved still and reviews contact sheets of the result. It includes named design-language specifications, composition and chart references, HTML templates, render, and animation workflows.
ACRYL is an agent-agnostic Agentic Development Environment and continuity layer for software work. It provides a persistent project workspace with a canonical event stream, durable tasks and artifacts, agent identities and sessions, context projections, structured handoffs, workspaces, and checkpoints, allowing replaceable coding agents to work on the same project context. ACRYL is built on the Cordis meta-framework, which supplies lifecycle-managed plugins, named services and replaceable providers, reactive dependency injection, typed events, reversible effects, scoped composition, and configuration-driven application profiles. Agents, models, memory systems, code graphs, tools, workflows, terminals, and user-interface surfaces are treated as composable capabilities. Its Development Canvas can host PTY terminals, coding-agent sessions, files and editors, browser tabs, and other capability-provided views. The project is distributed as separate Desktop GUI, terminal TUI, and local web installations; the web surface runs locally rather than as a hosted cloud service. ACRYL is in active early development, is licensed under the MIT License, and is developed independently of the DeepSeek Harness project while retaining architectural influences from it.
LiveKit Agents is an open-source Python framework for building programmable, real-time multimodal voice agents that run as server-side participants. It combines speech-to-text, large language models, text-to-speech, and realtime APIs with LiveKit's WebRTC clients, telephony stack, RPC and data APIs, and MCP tool support. The framework provides agent sessions, server-side job scheduling and dispatch, semantic turn detection based on a transformer model, multi-agent handoffs, and tools for voice, text-only, transcription, vision, and video-avatar applications. Agents can be tested with native test integration, including event assertions and LLM-based judges, and can run locally in console mode, in development mode with hot reloading, or in production mode. The framework can use LiveKit Cloud or a self-hosted LiveKit server and is distributed as a Python package with plugins for model providers. The Agents framework is licensed under Apache-2.0; LiveKit turn-detection models use the LiveKit Model License.
Paseo is a self-hosted platform for orchestrating multiple coding agents, including Claude Code, Codex, GitHub Copilot, OpenCode, and Pi. Its local daemon manages agent processes on users' machines, while desktop, mobile, web, and CLI clients connect to run agents in parallel, stream output, send follow-up tasks, and work in specified directories or worktrees. Paseo supports voice task dictation and control, agent handoffs, advisor and committee workflows, local tools, configurations, skills, and development environments, and provides an MCP server, WebSocket API, and TypeScript SDK for integrations, dashboards, and orchestration services. Remote connections use an end-to-end encrypted relay, TCP, Tailscale, or another VPN. Paseo can run as an installed application, headlessly, or as a Dockerized daemon with a self-hosted web UI. The repository is licensed under Apache-2.0 and states that Paseo has no telemetry, tracking, or forced log-ins.
A personal AI learning system distributed as a pi configuration. It encodes a teaching philosophy and learning process in skills, including teaching and diagram-based visualization, and adds extensions for question popups, graded quizzes, Markdown session logs, and visualization tools. The configuration delegates research and visual creation to researcher, SVG-maker, and Mermaid-maker subagents; it can also run without subagents, with those delegation-based capabilities omitted. It is intended for one learner and is shared as-is under the repository's stated installation instructions.
LiveKit is an open-source framework and developer platform for building, testing, deploying, scaling, and observing real-time voice, video, and physical AI agents. Its agent pipeline streams user speech from an app, browser, or phone call to an agent, which applies custom business logic and returns a response; the platform supports automatic turn detection and interruption handling, speech-to-text, language-model, and text-to-speech providers, web and mobile applications, and telephony through phone numbers and SIP integrations. LiveKit Cloud provides deployment and scaling on LiveKit's real-time infrastructure, alongside an inference gateway and full-stack observability for agent sessions.
A downloadable web-development project bundle from GreatStack for building a full-stack AI website builder with MongoDB, Express.js, React.js, and Node.js. The project includes starter assets and source files for a React website generator that accepts text prompts, builds websites step by step, exposes generation progress, and supports manual source-code editing, follow-up AI prompts, exporting, and publishing. Its tutorial project includes user authentication, REST APIs, an OpenRouter model integration, and an agent chat API for updating generated projects.
Crawl4AI is an open-source Python web crawler and scraper that converts web pages into structured, LLM-ready Markdown for retrieval-augmented generation, agents, and data pipelines. Its asynchronous Playwright-based crawler supports Chromium, Firefox, and WebKit, dynamic JavaScript pages, sessions, persistent browser profiles, cookies, headers, proxies, screenshots, media, iframes, lazy loading, full-page scanning, caching, and deep crawling with BFS, DFS, and best-first strategies. For extraction, it provides heuristic Markdown filtering including BM25-based relevance filtering, CSS- and XPath-based schema extraction, chunking and cosine-similarity strategies, and optional LLM-driven structured JSON extraction. It also includes adaptive crawling, link analysis, URL seeding, virtual-scroll handling, anti-bot and proxy escalation features, and customizable hooks. Crawl4AI can be installed with pip and used through Python or its command-line interface. It is also distributed as a Dockerized FastAPI server with JWT authentication, browser pooling, monitoring dashboards, a playground, and endpoints for crawling, HTML extraction, screenshots, PDF generation, and JavaScript execution. The repository states that it is licensed under Apache License 2.0.
GitNexus, developed by Akon Labs, is a code-intelligence engine that indexes repositories into a knowledge graph for code exploration and AI-agent context. Its indexing pipeline walks the file tree, parses source with Tree-sitter, resolves imports, calls, inheritance, constructor-inferred receiver types and other relationships, groups symbols into functional communities, traces execution processes, and builds BM25-plus-semantic hybrid search indexes backed by LadybugDB. The resulting graph supports MCP tools and CLI commands for process-grouped search, symbol context, call-path tracing, blast-radius and Git-diff impact analysis, structural checks, coordinated renaming, API and route mapping, taint and dependence queries, and Cypher access; repository groups can link contracts and impacts across multiple repositories. The CLI runs locally and can connect editors such as Claude Code, Cursor, Codex and others through MCP, skills and selected hooks. GitNexus also provides a browser-based WebAssembly UI with an interactive graph explorer and AI chat, plus a local HTTP server and Docker deployment mode for accessing indexed repositories through a backend. The web-only mode keeps repository processing in the browser and is constrained by browser memory, while the native CLI stores indexes locally in each repository's .gitnexus directory and uses a global registry for multi-repository access.
FuXi is a self-contained terminal AI coding agent developed by FUXI. Built in Go and distributed as a static binary, it uses a Think → Act → Verify loop to read and edit code, run shell commands, drive tools, connect to MCP servers, and route requests across multiple LLM providers with automatic failover and cost-aware settings. Its built-in capabilities include file operations, shell execution, code search, web fetching, LSP diagnostics, Jupyter, browser use, background tasks, and parallel sub-agents. Shell commands pass an AST-based safety classifier, while permissions and audit logs govern autonomous actions. Sessions persist to disk, with checkpoints for resuming, rolling back, or forking; the TUI also supports memory consolidation and automatic context compaction. FuXi supports provider API keys or FuXi OAuth, configurable OpenAI-compatible endpoints, MCP clients, hooks, skills, plugins, and slash commands. The repository contains documentation, installers, release information, and issue-tracking materials; it states that the product source is proprietary and not published.
MyContext is a local-first desktop app from openTrinity that builds a private personal work-context layer from sources such as instant-messaging conversations, documents, and meeting records. It stores local copies, indexes, source references, and derived context in an on-disk SQLite vault, then organizes them into a personal context graph linking people, projects, topics, events, conversations, and supporting facts. Its search and answer workflow combines local full-text search, semantic retrieval, and graph queries, with agents assembling answers from traceable source material and falling back to ranked local results when the agent runtime is unavailable. A digital-self workflow recalls relationship-specific context and communication history to draft replies, while sending, deletion, and other consequential actions require explicit user confirmation. The repository describes an Electron and React desktop architecture with source connectors, incremental ingestion, context processing, retrieval, knowledge-graph, persona, and isolated agent-runtime layers. MyContext is in developer preview and under active development; its README warns of compatibility-breaking changes and migrations that may require recollection. It is licensed under the Elastic License 2.0, which permits use, modification, and self-hosting but restricts offering it as a hosted or managed service to third parties.
screenpipe is a local AI-agent memory layer that continuously captures computer history on macOS, Windows, and Linux. It records screen frames with OCR, accessibility data, microphone and system audio with transcripts, and application activity, storing the underlying history locally. Agents can search the history through a local REST API, database, or MCP server and use it for tasks such as meeting summaries, follow-ups, and workflow automations. Users can exclude apps, windows, URLs, or time periods and redact sensitive fields on the device; the page describes the project as source-available.
OpenClaude is an open-source, terminal-first coding-agent CLI for cloud and local model providers. It connects to OpenAI-compatible APIs, Gemini, GitHub Models, Codex, Ollama, Atomic Chat, and other supported backends, providing prompts, streaming output, Bash and file tools, grep, glob, agents, tasks, MCP, slash commands, web search and fetch, and image inputs for compatible providers. The CLI supports guided provider setup with saved profiles, conversation continuation and forking, detached local background sessions, model-specific agent routing, repository maps based on PageRank-ranked code structure, and a headless bidirectional-streaming gRPC server for integrations, CI/CD pipelines, and custom interfaces. A bundled VS Code extension provides launch integration, in-editor chat, provider-aware controls, and theme support. OpenClaude runs on Node.js 22 or newer, is distributed through npm and an Arch Linux AUR package, and can use local inference or remote APIs. Its repository is licensed MIT for the project's modifications and states that it is an independent community project derived from and substantially modified from the Claude Code codebase, without Anthropic affiliation.
Skill Cabinet is a local catalog for agent skills installed on a machine. It scans user-level skill locations such as .agents, .claude, .codex, .cursor plugins, Hermes profiles, and other ~/.* /skills folders, then lets users filter skills by drawer, metadata, risk, invocation mode, and status; inspect rendered or source bodies, YAML frontmatter, extra files, symlinks, duplicates, origins, and broken links; and review disk usage. It runs with Node 20 or later through npx skill-cabinet, starting a server bound to 127.0.0.1 and opening the catalog in a browser. Users can quarantine skills to ~/.skill-cabinet/quarantine and restore them, or delete individual skills or groups; deletion removes folders, files, or symlinks from the scanned locations, while a symlink's target is retained. The project is distributed under the MIT license.
SkillRadar is open-source discovery, security, ranking, and routing infrastructure for Agent Skills and Codex. It discovers public SKILL.md files, parses their contents, performs conservative static checks for capabilities such as shell commands, dynamic execution, secret access, networking, package installation, filesystem writes, and deployment tooling, then classifies and ranks candidates in a safety-gated registry. D and Blocked candidates are kept audit-only and excluded from automatic routing. Its Codex plugin provides task-to-Top-3 routing, skill search, provenance and safety inspection, and a read-only Skill Budget Doctor. Routing can use a bundled offline registry and returns relevance, SkillRadar score, security grade, provenance, reasons, and match details without executing candidate repositories or depending on live GitHub discovery. The repository includes radar data, a matching system, a router-quality benchmark, a local registry UI, and daily bot-refreshed generated data; it is released under the MIT license.
Procedura is an open-source agentic 3D-modeling tool from SpatiaOS that converts text prompts into editable procedural assemblies rather than point clouds or triangle meshes. It generates OpenSCAD source with named parts and typed mates, plans and builds parts incrementally, and can use reference images and optional Blender render feedback during refinement. Optional passes assign per-part PBR materials and plan articulation, exporting motion to OpenUSD and URDF with Isaac-based validation. It runs locally as a Bun/TypeScript pipeline using a configurable OpenAI-compatible, Gemini, or local model endpoint; it does not provide hosted inference or API keys. The pipeline uses a Manifold-capable OpenSCAD build to compile the generated programs and Blender for renders, and includes a web Studio for composing runs and inspecting intermediate artifacts. The repository is MIT-licensed.
CDAF is an open sidecar format and toolkit for video that stores a timestamped plain-text description beside the corresponding video file. Its Python library and CLI can generate, parse, validate, read, and report the status of `.cdaf` files, while an agent skill teaches video agents to check for a matching sidecar before processing footage. Each sidecar contains a minimal versioned header with the video filename, SHA-256 hash, byte size, duration, generator, and creation time, followed by sections such as summary, timestamped segments, transcript, on-screen text, and tags. Conforming tools verify the video's freshness and refuse to use a stale sidecar after the video changes; the format is model-agnostic even though the included generator uses the Gemini Files API, with optional local-model support. The repository includes a normative specification, reproducible sidecar-versus-direct-video benchmarks, an agent skill installable with `npx cdaf-skill`, and CLI commands for generation, validation, reading, and status checks. The core validation functions require only the Python standard library; generation requires Python 3.10 or later and a user-supplied Gemini API key. The project is licensed under MIT.
Claude 5.1 is described in the supplied video evidence as an Anthropic AI model for programming, scientific research, and agentic tasks. The video claims that it can work continuously for dozens of hours on codebases and multistep research, with separate lower pricing for cached, ordinary, and complex agent tasks.
DBOS is a database-oriented operating-system project and application environment that stores important system state in a database. Its practical focus includes durable, recoverable workflows, particularly workflows used by agentic AI systems.
Academic Research Skills for Claude Code is an open-source suite of Claude Code skills for academic research and publication workflows, maintained by Cheng-I Wu. It provides separate deep-research, academic-paper, academic-paper-reviewer, and academic-pipeline skills for literature reviews, systematic reviews, guided research, drafting, citation conversion, revision, peer review, rebuttal auditing, methodology review, and re-review. The suite uses staged, human-in-the-loop workflows with multi-agent orchestration, Socratic checkpoints, style calibration, writing-quality checks, citation formatting, and outputs in Markdown, DOCX when Pandoc is available, and LaTeX/PDF through tectonic. Its pipeline orchestrator connects activities through ten stages, user-confirmation checkpoints, Material Passport handoffs, claim and citation verification, integrity gates, and final process summaries. Citation verification can cross-check references against Semantic Scholar, OpenAlex, Crossref, and arXiv when available.
Claude-Mem is an open-source persistent-memory plugin and service for coding agents. It captures agent activity through lifecycle hooks, stores sessions, observations, and summaries in SQLite, and uses hybrid full-text and Chroma vector search to retrieve relevant context across sessions. Its MCP search workflow uses progressive disclosure: the agent first searches a compact index, then reviews a timeline, and finally fetches full observations for selected result IDs. A local worker service, managed by Bun, provides the HTTP API, search endpoints, and web viewer; the project also supports integrations with Claude Code and other listed agent environments, configurable context injection, private-content exclusion tags, and optional cloud synchronization. The repository states that it is distributed under the Apache License 2.0 and requires Node.js 20 or later, with Bun, uv, and SQLite used by the runtime. It can be installed through its npx installer or Claude Code's plugin marketplace.
CodeBurn is a free, open-source, local-first tool by AgentSeal that reads session files written by AI coding tools and reports token usage and estimated cost by provider, model, project, task, and activity. It provides a terminal dashboard and reports, a localhost web dashboard, desktop and tray or menubar views, exports, model comparisons, subscription-plan tracking, and optional cross-device aggregation. Its deterministic analyzers classify work into task categories from tool usage and message keywords, calculate token costs using LiteLLM pricing cached locally, and scan coding-agent sessions for patterns such as repeated file reads, low read-to-edit ratios, uncapped shell output, unused MCP servers, bloated configuration, and retry-heavy work. The optimize command produces estimated savings and fixes; applicable configuration changes are backed up and journaled so they can be undone and later compared with observed usage. The yield command heuristically correlates sessions with Git commits to classify spend as productive, reverted, abandoned, or ambiguous. The guard feature installs opt-in Claude Code hooks for soft and hard session-spending caps, checkpoints, and status-line reporting. CodeBurn also exposes local usage and savings through an MCP server over stdio. The CLI reads data from the local machine without wrappers, proxies, API keys, or uploads; the optional desktop applications can send anonymous bucketed telemetry after consent. It requires Node.js 22.13 or newer, and the repository is licensed under MIT.
here.now is an agent-oriented hosting and publishing service for publishing files and folders—including websites, documents, dashboards, presentations, prototypes, games, and media—to the web and receiving a live URL. Any AI agent that can make HTTP requests can publish to it; no account is required, but unauthenticated sites expire after 24 hours, while registered accounts can keep sites permanently. Sites are public by default with randomly generated URLs, and can be protected with passwords or restricted to invited email addresses or domains. The service also supports custom domains and team workspaces with member-only visibility and workspace subdomains.