567 tools and products — trending open source, and what gets used in AI and other work.
Claude Code is Anthropic's agentic coding tool for the terminal, IDEs, and GitHub. It uses natural-language commands to understand a codebase, create and read files, execute commands, run tests, explain code, manage Git workflows, and handle routine development tasks. It can also load persistent project context, run custom slash commands, use plugins with custom commands and agents, and operate with configurable autonomy while leaving actions such as final pull-request merging to a human. The official repository documents installation for macOS, Linux, and Windows, and identifies npm installation as deprecated.
OpenAI Codex is an AI coding agent from OpenAI available as a command-line tool (Codex CLI) that helps developers produce software. It can be used alongside Gemini for adversarial audits of software requirements and implementation plans, listed as a supported coding-agent or model option in several projects, and its logs can be joined with task and test evidence. The Codex CLI can also receive and answer requests from the Penako canvas.
Cursor is an AI-powered coding agent and integrated development environment designed to accelerate software development by handing off coding tasks to AI. It evolved from an email client into a multimodel development tool, supporting broader developer workflows. The platform also offers MCP-connected capabilities for tasks like searching and editing notes.
OpenClaw is a self-hosted AI assistant and agent platform developed by the OpenClaw Foundation. It runs on macOS, Linux, Windows, or WSL2 devices and connects hosted or local model providers, tools, skills, plugins, messaging channels, and optional companion apps through a Gateway. The same Gateway architecture supports a personal assistant on one device or a trusted shared-team deployment. The Gateway is the local control plane for sessions, tools, events, and channel connections; the Control UI, CLI, and terminal UI connect to it. Channels include WhatsApp, Telegram, Slack, Discord, Google Chat, Signal, and iMessage, while companion apps and nodes can provide voice, Canvas, camera, screen, and device-local actions on supported platforms. OpenClaw’s security model treats inbound messages as untrusted input, pairs unknown senders by default on direct-message-capable channels, and runs tools on the host for the main session unless sandboxing is configured. The project is distributed under the MIT license.
Higgsfield AI is an AI-native creative platform for generating and editing images and videos from text prompts and reference images, with tools for cinematic video creation, visual effects, AI-assisted content production, animating photos, creating shorts, generating voiceovers, cloning voices, upscaling, reframing, removing backgrounds, and automating creative workflows with an AI agent. It supports complete faceless-video production with visuals, voice-over, subtitles, music, and sound effects from an approved script and storyboard. The platform connects through an MCP server and ChatGPT plugin to MCP-compatible clients such as ChatGPT, Claude, Claude Code, Codex, and other agents; the ChatGPT integration uses an existing Higgsfield account and its credit system, while the MCP page states that no API key is required. It is available on the web and mobile.
Model Context Protocol (MCP) is an open, standardized protocol layer for connecting large language models and AI agents with external data sources, hosted infrastructure tools, and other contextual data. It defines formats, metadata, protocol schemas, and APIs for sharing information and tools, attaching, referencing, and validating context such as documents, embeddings, and provenance, while supporting scoped authentication and permissions. MCP translates JSON requests from an agent into calls to service APIs, including CRM, container-management, and GKE capabilities, and can provide design-system context and related tools for generating consistent applications. The official project publishes its specification, documentation, and protocol schema; the schema is defined first in TypeScript and also provided as JSON Schema for broader compatibility. The protocol was created by David Soria Parra and Justin Spahr-Summers, is hosted at modelcontextprotocol.io, and is licensed under the MIT License.
OpenAI is an AI model provider whose language models are available through an API and can serve as one provider in multi-model routing architectures. The videos describe its models being used for AI replies and extraction, Amazon-review analysis and customer-service email generation, legal workflows, complex software development and debugging, and experimental or auxiliary agent tasks.
T3 Code is an open-source agent-harness control surface for coding agents running on a developer's computer. It provides iOS and Android mobile apps, a web app, and an Electron-based desktop app for controlling locally configured Codex, Claude Code, Cursor, Grok Build, and OpenCode agents while using the user's existing provider subscriptions. The project can be launched with `npx t3@latest`, which starts its backend and local web interface on the machine. It also distributes desktop builds for Windows, macOS, and Arch Linux, and supports remote access from a phone or another machine. The repository describes the project as early-stage and warns that bugs are expected.
Fable orchestrator is a small, local-first routing skill for Codex that separates planning and adjudication from code implementation. Fable 5.1 plans tasks and makes final decisions without writing code or owning the workspace; Codex validates the bounded graph it returns, runs ready implementation nodes in parallel when useful, collects evidence, verifies the result, and requests further adjudication when needed. Implementation work is delegated through callable OpenCode Go agents, with GPT-5.6 Luna assigned to normal implementation and DeepSeek V4 Flash to loops and repeated high-throughput execution. The repository contains an installable skill, agent configuration, an invocation script, an installer, and shell-based tests. It does not provide a proxy, dashboard, model catalog, credential store, or API-key management. The installer supports dry runs, copying into a Codex skills direct
An agent-facing collection of design and engineering skills by Emil Kowalski for building and reviewing user interfaces. Its guidance covers animation curves, durations, properties, gestures, haptics, screen transitions, motion placement, borders, shadows, typography, Apple interface principles distilled from WWDC talks, UI-library selection, modern Swift, prototyping, and product-interface details. The skills can create animations from scratch, review animations against defined rules, audit a codebase and produce prioritized implementation plans, identify worthwhile opportunities for motion while flagging what should not be animated, and build multiple UI variants navigated through a switcher. A React Native and Expo skill covers gestures, sheets, haptics, screen transitions, and keeping motion off the JavaScript thread; other skills guide use of the Sonner toast library and library selection. The collection is installed with `npx skills@latest add emilkowalski/skills` and is based on Kowalski’s stated experience at Vercel and Linear.
Hugging Face is an AI and machine-learning collaboration platform that hosts and distributes models, datasets, and applications. Its Hub supports public model, dataset, and application repositories across text, image, video, audio, and 3D workloads, while its open-source stack includes Transformers, Diffusers, Safetensors, Tokenizers, TRL, Transformers.js, smolagents, and PEFT. The platform provides a unified API for accessing models from AI providers, GPU-based compute and Inference Endpoints, and paid team and enterprise features such as private datasets, access controls, single sign-on, audit logs, and dedicated support.
Hermes Agent is a self-improving AI agent developed by Nous Research. It provides a CLI and terminal interface, messaging gateway, and desktop interface for interacting with language models, operating terminals and browsers, and maintaining continuity across sessions. It supports model providers including Nous Portal, OpenRouter, OpenAI, and custom endpoints, along with terminal backends including local execution, Docker, SSH, Singularity, Modal, Daytona, and Vercel Sandbox. Its learning loop creates and improves skills from task experience, stores persistent memories, searches previous sessions with FTS5 and LLM summarization, and builds user profiles through Honcho. It can delegate parallel work to isolated subagents, run Python tool-calling scripts through RPC, schedule unattended jobs with a natural-language cron scheduler, and deliver results through Telegram, Discord, Slack, WhatsApp, Signal, email, and the CLI. The terminal interface includes multiline editing, slash-command completion, conversation history, interrupt-and-redirect controls, and streaming tool output. It also supports voice-memo transcription and cross-platform conversation continuity.
Perplexity is an AI-powered answer engine and search assistant that retrieves current information, verifies facts, and supports search operators such as site:reddit.com to narrow results. It is also offered as a cloud-based model and agent platform that can run as an AI assistant or computer agent. It is a commercial product developed by Perplexity AI.
GitHub Spec Kit is an open-source, community-driven toolkit from GitHub for spec-driven software development with AI coding agents. It provides the Specify CLI to initialize projects and create feature branches, establishes project principles in a constitution, defines specifications, plans implementations, breaks plans into actionable tasks, implements them through an agent, and checks the result against the specification, plan, and tasks. The toolkit supports integrations with AI coding agents and includes extensions, presets, role-based bundles, and an opt-in bug-fixing extension with an assess–fix–test workflow.
ElevenLabs is an AI voice and audio platform offering text-to-speech, voice generation and cloning, speech-to-text transcription, dubbing, music generation, and voice agents. Its transcription capability produces timing information for video cuts, subtitles, and synchronized animations. It provides more than 5,000 voices across more than 70 languages, supports custom voice clones, and offers APIs and SDKs through ElevenAPI. ElevenCreative supports content creation, while ElevenAgents supports customer-experience applications.
Apify is a marketplace and platform for web-data extraction and AI automation tools called Actors. Actors collect data from sources such as TikTok, Instagram, Google Maps, and websites; their results can be exported, accessed through an API, scheduled, monitored, or integrated with applications, workflows, and AI agents. Apify’s Website Content Crawler cleans HTML, extracts website text in Markdown and other rich formats, downloads files, and supplies content to AI models, LLM applications, vector databases, and retrieval-augmented generation pipelines, with integrations including LangChain and LlamaIndex.
LangChain is an open-source framework and agent engineering platform for building agents and applications powered by large language models. It chains interoperable components and third-party integrations, providing standard interfaces for models, embeddings, vector stores, retrievers, tools, and other data sources. It supports sequential agentic workflows, context and tool-call orchestration, model substitution, and application development primarily through Python; the project also provides a separate JavaScript/TypeScript library. LangChain can be used standalone or with related tools for agent orchestration, evaluation, observability, debugging, and deployment.
An agent-orchestration framework for Python that models complex, multilayered AI workflows as directed agent graphs. It manages context and tool calls, provides state checkpoints and persistence, and supports human-in-the-loop approval gates for production deployments. The framework is reported to allow existing agents to be exported/imported into IBM watsonx Orchestrate.
An AI software engineering agent developed by Cognition that runs isolated cloud coding sessions, writes and modifies code, and can test applications through a browser. Devin is designed as an end-to-end software-development environment for planning and implementing coding tasks. Vals evaluated it internally and increased its adoption after finding it to be token-efficient.
Ollama is a software platform for running and serving open models, including as a local model runtime for tasks such as resume evaluation, tagging, and summarization. Its website describes integrations with coding agents and other workflows, allowing users to launch tools such as Claude Code, Codex, OpenCode, and VS Code while switching models without changing the workflow. Local runs remain on the user's machine, while Ollama also offers hosted cloud models. The service states that prompts are not tracked or used for training by providers, and that its open-source software and open model support are intended to keep data private while automating work.
Anti Slop is an open-source set of rules and agent skills for AI coding agents that filters generic AI-generated user interfaces, copy, and code without prescribing colors, fonts, layouts, or other visual styles. Its core contains 38 mandatory rules across Hard Gate, Purpose-Gate, and Quality Locks tiers; a Liveliness Toolkit with ENERGY, RHYTHM, and MOTION controls; a Design Read; and a mandatory four-part PASS/FAIL Delivery Gate. Task-specific skills are organized as separate SKILL.md folders and loaded additively, while a project’s DESIGN.md supplies the visual direction. The repository provides installation through the npx antislop-ai picker, the skills.sh-compatible npx skills add command, or a Claude Code marketplace plugin, with support for agents including Claude Code, Codex, Antigravity, OpenCode, Cursor, Gemini CLI, and Hermes.
Zapier is an automation and AI orchestration platform for connecting business applications, data, AI models, and agents. Its workflows use triggers and actions to move information between tools and perform tasks such as formatting new Asana projects, connecting incoming email to OpenAI for a customer-service workflow, scheduling recurring scrapes, and sending notifications when qualifying data appears. The Zapier site says it supports workflows and agents across more than 9,000 app integrations, with governed access, model management, and guardrails. It also provides Zapier MCP for connecting AI clients such as Claude, ChatGPT, Cursor, and OpenClaw to configured tools, plus a beta Zapier SDK and SDK CLI for building integrations.
Framer is a web-based design and website-building platform that combines visual design, prototyping, and site creation with built-in hosting, CMS, analytics, security, and SEO. Its AI design agent generates page designs and websites from textual prompts on a visual canvas; the generated designs remain editable and can be deployed through Framer’s hosted SaaS.
Archify is an open-source agent skill and Node.js rendering and validation system that turns a codebase or system description into searchable, shareable interactive diagrams. It supports architecture, workflow, sequence, data-flow, and lifecycle diagrams and can be used with Cursor, Claude Code, Codex CLI, and OpenCode. Agents produce typed JSON intermediate representations, which Archify deterministically validates and compiles into self-contained HTML and SVG artifacts. The generated viewer supports node search, optional revision-verified source inspection, upstream and downstream reach tracing, route inspection, semantic-role comparisons, guided stories, themes, presets, and finite motion. Archify can compare two validated snapshots as Before, Delta, and After views, identifying added, removed, changed, moved, and rerouted facts. Its exports include PNG, SVG, WebM, and 1200×630 share cards; it can be installed as a global skill with `npx skills add tt-a1i/archify -g`.
OpenCode is an open-source AI coding agent developed by anomalyco. It is used as a terminal-based coding agent and external coding harness, with a beta desktop application also available. The CLI includes a full-access build agent for development work and a read-only plan agent for code analysis and exploration; the plan agent denies file edits by default and asks permission before running shell commands. OpenCode also provides a general subagent for complex searches and multistep tasks, and can be installed through its shell installer, package managers, or platform-specific desktop packages.
Orca is a free, open-source agent development environment for running multiple coding agents in parallel on desktop, mobile, or a VPS. It runs CLI agents such as Codex, Claude Code, OpenCode, and Pi side by side, placing each agent in an isolated Git worktree so results can be compared and merged. The environment provides split terminals, a VS Code editor, diff annotation and review, browser-based Design Mode that sends selected HTML, CSS, and screenshots to an agent, native GitHub and Linear views, SSH worktrees with port forwarding, and a mobile companion for monitoring agents and sending follow-ups. Its CLI can script workflows such as creating worktrees, taking snapshots, clicking, and filling fields. Orca is designed to work with any coding agent that runs in a terminal and uses the user's own agent subscriptions.
Agent Skills is an open-source collection of engineering workflows for AI coding agents, developed by Addy Osmani. It packages senior-engineering practices into skills covering the development lifecycle: defining specifications, planning, incremental implementation, testing and debugging, constraints, code review, web-performance auditing, code simplification, and shipping. Slash commands such as /spec, /plan, /build, /test, /review, and /ship activate the corresponding workflows, while skills can also activate automatically based on the work being performed. The /build workflow can generate a plan and implement its tasks in an approved pass, while retaining test-driven verification, individual commits, and pauses for failures or risky steps. It can be installed with the skills CLI or integrated into supported coding agents including Claude Code, Cursor, Codex, Copilot, and Cline.
Open Code Review is an open-source, AI-powered command-line code-review tool developed from Alibaba Group’s internal review assistant. It reads Git diffs and sends changed files to a configurable LLM through an agent with tool-use capabilities; the agent can read complete files, search the repository, and inspect other changed files for context before producing structured comments tied to source lines. Its hybrid architecture combines deterministic engineering with dynamic agent decisions: deterministic stages select and bundle files, match rules to file characteristics, and apply independent comment-positioning and reflection modules, while scenario-specific prompts and tools guide context retrieval and review. The `ocr scan` command audits complete files or directories without requiring a meaningful diff, and the project provides multilingual rules for issues including null-pointer exceptions, thread safety, cross-site scripting, and SQL injection. It supports configurable model endpoints, including OpenAI- and Anthropic-compatible endpoints, and requires Git 2.41 or newer.
Reverse Skill is an open-source, client-neutral cybersecurity skills router for AI coding agents performing reverse engineering, authorized penetration testing, security research, and CTF work. It processes a task through structured routing rules and a primary master-routing workflow, initializes authorization and network scope before action, selects a scenario-specific skill, checks available tools, MCP servers, and scripts, and records a timeline, evidence-to-finding path, report, and field journal rather than relying on guessed commands. Supported scenarios include APK and mobile analysis, binary and .NET reverse engineering, frontend JavaScript and encrypted-parameter analysis, DSL-VM reverse engineering, HTTP capture and request replay, malware and YARA analysis, penetration testing, attack-chain orchestration, case review, CTFs, firmware and IoT security, patch-diff analysis, exploit development, EDR bypass research, API and GraphQL security, supply-chain and SBOM security, and LLM security. Its documented tool ecosystem includes JADX, Apktool, Frida, IDA Pro, radare2, Ghidra, BurpSuite, and YARA. The repository provides platform-specific setup and tool-index refresh scripts for Windows, Linux, macOS, and Kali Linux, with prerequisites including a JDK, Node.js, and Python. It supports Claude Code, Codex, Cursor, OpenCode, and other compatible clients while keeping client adapters separate from the routing core, and includes cross-platform CI, routing regression tests, structure and supply-chain checks, and generated skill-navigation indexes.
Lightfield is an AI-native customer relationship management (CRM) platform for early-stage teams that captures, analyzes, and surfaces customer information. It updates itself from customer interactions, allowing its agents to run outbound activities, flag deals at risk, and identify where teams should focus next. The platform includes records for accounts, opportunities, contacts, tasks, meetings, notes, lists, signals, sequences, automations, and chats.
n8n is a fair-code workflow automation platform developed by n8n GmbH for building and deploying AI agents and multi-step workflows. It combines a visual canvas with JavaScript, Python, and npm packages to integrate services, automate tasks, run custom code, and connect user data, models, tools, logic, human approvals, and observability features. It supports no-code/low-code workflows, more than 1,500 integrations, workflow templates, and models from OpenAI, Anthropic, Google, and open-source providers. n8n can be self-hosted or run in the cloud, including with Docker, and is distributed under the Sustainable Use License and n8n Enterprise License; enterprise plans add features and support.
Firecrawl is an open-source web context API and hosted service for AI agents and applications. It searches the web, scrapes individual pages, crawls websites, maps site URLs, and batch-processes large URL sets, returning content as clean Markdown, HTML, screenshots, structured JSON, and other extracted data. It handles JavaScript-heavy pages, rotating proxies, orchestration, rate limits, and blocked content, and can parse web-hosted PDFs and DOCX files. Its interaction endpoint lets users or agents click, scroll, write, wait, and press on a page before extracting content; its agent endpoint gathers web data from a natural-language request without requiring URLs. Firecrawl provides Python and Node.js SDKs, cURL and CLI interfaces, and an MCP connection for AI agents and applications.
gbrain is a hosted, cloud-based AI agent with persistent, long-term memory that adapts to a user’s preferences and way of thinking. It provides built-in skills for personal use and shared “brains” that support collaboration within teams and organizations.
World Monitor is an open-source global intelligence dashboard developed in the koala73/worldmonitor project. It aggregates curated global and regional news, geopolitical information, infrastructure signals, and financial data into a unified situational-awareness interface, producing AI-synthesized briefs and tracking military, economic, disaster, and escalation signals. The application provides both a 3D globe and a WebGL flat map using a shared map-layer catalog, a server-authoritative Country Instability Index, and finance views for stock exchanges, commodities, crypto, and composite market signals. It supports local AI through Ollama without API keys, as well as other documented AI providers, and exposes an MCP server for programmatic access by agents and scripts. The repository supports multiple site variants from one codebase, including world, tech, finance, commodity, energy, and happy views. It also distributes a Tauri desktop application for macOS, Windows, and Linux, and documents self-hosting through Vercel, Docker, or static deployment.
Cloudflare OS is an open-source AI productivity environment developed by Cloudflare and built on Cloudflare Workers. It provides an agent chat interface preloaded with company-specific knowledge, sandboxed development of shareable applications called gadgets, and a security framework called Gatekeepers for controlling agents and apps. It is intended to be customized into an organization's own company operating environment rather than adopted unchanged as a traditional computer operating system. Each user receives a private, separately sandboxed instance of a productivity application, such as a slide-deck tool, whiteboard, game, or dashboard. Agents can create and modify these gadgets and perform tasks involving configured company systems and integrations. Gatekeepers are service-specific Workers that expose a Cap'n Web API, handle authorization such as OAuth, restrict access to the intended resource, log gadget and agent actions, and support human approval for side effects. Instead of stopping synchronously for approval, a Gatekeeper can simulate the outcome locally, let the agent continue and queue actions, then allow the user to approve or reject those actions later in bulk or individually. The full stack can be run locally with pnpm, Wrangler, and workerd using `pnpm run-local`, or deployed to a Cloudflare account. The repository describes the project as an early-access release under heavy development; the local setup is intended for evaluation rather than production use.
DeepTutor is an open-source, self-hostable AI tutoring platform developed by HKUDS for lifelong personalized tutoring. It provides a web application and command-line interface for interactive learning, problem solving, quiz generation, deep research, visualization, and mastery practice, with persistent memory and learning state shared across knowledge bases, books, notebooks, tutor personas, and other activities. The platform supports multi-agent problem solving, retrieval-augmented learning with source-traced and page-level citations, math animations, interactive visualizations, and tutor bots built from personal study materials. Its knowledge sources include personal documents, EPUBs and books with annotations, GitHub repositories, web search, and connected libraries; documented retrieval and ingestion options include GraphRAG, PageIndex, LightRAG, linked knowledge bases, Obsidian, and configurable parsing and vector backends. DeepTutor also includes courses, research workflows, an ecosystem of MCP services and tool or capability plugins, and integrations with connected coding agents and other partners. It can run locally or through Docker.
Tines is a software company that provides a secure, governed workflow automation platform for IT, security, and other teams. It supports building and operating AI agents, applications, and automated workflows, with visibility, access control, and administrative governance for maintaining control over automated processes.
Vercel AI Gateway is a managed service from Vercel that routes application requests to language-model providers, centralizes API key management, and provides observability, caching, and rate-limiting for LLM usage. It is intended to simplify integration of multiple LLMs and enforce access controls for apps and autonomous agents.
book-to-skill is an open-source agent skill that converts technical books, document folders, or collections of source files into a unified skill for GitHub Copilot CLI, Amp, or Claude Code. It accepts PDFs, EPUBs, office documents, plain text, folders, globs, and file lists; the videos also describe OCR support for scanned PDFs. The tool distills source material into structured knowledge rather than a single summary, extracting frameworks, decision rules, anti-patterns, and per-chapter Markdown files. It also generates a core SKILL.md with a chapter index, plus glossary, patterns, and cheatsheet files. Agents load the relevant chapter and supporting files on demand, allowing answers to be grounded in the source content without placing the entire book in the context.
jcode is an open-source, Rust-based terminal harness for AI coding agents and LLM workflows, designed for interactive development across multiple sessions. It supports resumable sessions, session search, multi-agent swarm and helper-agent workflows, compatible model-provider integrations, MCP, browser automation, and a self-development mode; installation scripts target macOS, Linux, and Windows. For automatic memory, jcode embeds turns and responses as semantic vectors, searches a graph of stored memories using cosine similarity, and feeds relevant results into the conversation. A memory sideagent can verify and expand retrieval, while periodic extraction stores new memories and ambient consolidation reorganizes entries and checks for staleness and conflicts. Explicit memory tools and traditional retrieval over previous sessions are also available. Its terminal UI includes side panels, diff views, inline Mermaid diagrams, information widgets, custom scrolling, and real-time rendering. The project supplies a Mermaid renderer without browser or TypeScript dependencies and publishes benchmarks covering RAM consumption, startup and input-readiness times, and memory scaling across multiple active sessions.
Moli is an open-source headless browser engine built in Rust for AI agents. It fetches and extracts web pages, searches the web, and automates browser tasks while providing JavaScript, DOM, CSS, networking, storage, and browser automation capabilities. Moli treats the DOM as the source of truth and performs layout and rendering on demand, so DOM-first operations can avoid visual layout and paint. Layout enables geometry queries, coordinate input, screenshots, screencasts, and PDF output. It can be used through its CLI, CDP, WebDriver Classic, and WebDriver BiDi interfaces, including direct connections from Playwright over CDP, and supports Linux, macOS, and Windows.
Skills for Real Engineers is an open-source collection of small, composable agent skills by Matt Pocock for disciplined software development with Claude Code, Codex, and other coding agents. The skills support requirements clarification, project-language setup, test-driven development, debugging, code review, specification, planning, triage, and codebase surveys. The collection includes skills such as `/grill-me` and `/grill-with-docs`, which question the user about a proposed change before implementation, and a setup skill that configures an issue tracker, ticket labels, and documentation storage for a repository. It can be installed as a managed Claude Code plugin or copied into a project as editable files through the `skills` installer; the skills are intended to work with any model and can be adapted by the user.
DeepSeek Harness (dsh) is an open-source agent harness developed by DeepSeek AI. It uses an architecture in which everything is a plugin and is powered by Cordis, whose design is described in “A Programming Paradigm for Spatiotemporal Composability.” The npm package can launch a local Web UI with `npx @deepseek-ai/dsh web`, serving at `http://127.0.0.1:3080` by default and opening the default browser; it can also be built and run from a repository checkout. The project is in developer preview and warns that compatibility-breaking changes may occur. It is licensed under the MIT License, with third-party dependencies and their licenses documented separately.
AgentENV (AENV) is an open-source distributed platform developed by kvcache-ai for running agent environments at scale, including environments used for agentic reinforcement-learning training such as Kimi K3. It runs Firecracker microVM sandboxes across machines, loads OCI-compatible images on demand through overlaybd, and uses local disks as bounded caches for image and snapshot data. The aggregate image and snapshot footprint can exceed local disk capacity because cold data is evicted from the cache. Snapshot-backed environments can boot or resume in under 50 ms and pause in under 100 ms. AENV incrementally snapshots memory and filesystem changes, stores snapshots in S3-compatible object storage or a shared distributed filesystem, and can fork a running environment into independent sandboxes for parallel workflows. It uses ublk for I/O and memory ballooning to return reclaimable guest memory to hosts; the documentation reports a 9.6× memory overcommit ratio in production. The distribution includes an AENV server and the `aenv` command-line client for pulling OCI images as templates, starting, attaching to, executing commands in, pausing, resuming, and deleting sandboxes. It exposes an E2B-compatible HTTP API that works with the standard E2B Python and TypeScript SDKs. It requires Linux kernel 6.8 or later and `/dev/kvm` access for Firecracker execution. API requests are authenticated, but traffic is not encrypted; the documentation recommends a trusted network or HTTPS termination at a reverse proxy or load balancer.
An AI software-development agent from Replit that builds, modifies, publishes, and deploys applications from natural-language prompts. It can generate user interfaces and databases, implement application workflows such as electronic signing and checkout, connect custom domains, and manage deployment.
Cumora is a cross-platform team-chat platform in which AI agents participate alongside human teammates. It provides shared rosters, direct messages, group conversations, a Kanban board, and a calendar; agents can maintain personas and memory, claim work, coordinate, and send and receive email. Agents can run in Cumora's cloud in per-agent Kubernetes pods, using a multi-hop tool-calling loop on the OpenAI Responses API, or through its BYOA mode, where a local daemon connects the platform to agent CLIs such as Claude Code, Codex, Grok Build, Cursor Agent, OpenCode, or pi. The frontend uses React, Vite, TypeScript, and Tailwind across web, desktop, and mobile shells; the backend is a stateless Node service using Express and WebSockets, with PostgreSQL as the source of truth and Redis for pub/sub and presence. Agent coordination uses a freshness gate that holds stale replies for reconsideration, atomic claims on work units, and a triage gate intended to reduce unnecessary large-model calls.
herdr is a terminal multiplexer and persistent runtime for coding agents, developed as a Rust binary. It runs a background server that keeps agent terminals and sessions available across terminal disconnections, SSH reconnections, network loss, lid closure, and machine restarts; users can reattach from another terminal and use tmux-style keyboard controls or mouse interactions to split, move, and manage panes. Each pane is marked working, blocked, or idle, and agents can control herdr through its CLI and socket API to spawn panes, prompt other agents, and wait for an agent that is blocked. It hosts existing tools such as Claude Code, Codex, Cursor, OpenCode, and Grok without wrapping or replacing them, and supports plugins for extending panes and workflows. The project provides installation scripts and package-manager installation, documents remote use and session state, and is licensed under the Apache License 2.0.
OpenViking is an open-source context database for AI agents developed by Volcengine. It unifies agent memories, resources for knowledge retrieval, and skills in a virtual filesystem accessed through the viking:// protocol, allowing agents to browse context with filesystem-style operations such as ls, tree, and find instead of querying an opaque vector store. Content is processed into three loading tiers: L0 abstracts for relevance checks, L1 overviews for planning, and L2 full details loaded on demand. Recursive retrieval first locates a relevant directory through vector search and then drills down through its hierarchy, preserving the surrounding context and recording the browsing trajectory for inspection and debugging. After a session is committed, OpenViking asynchronously extracts user preferences and agent experience into long-term memory. The project also provides an OpenViking Studio browser playground, documentation, and a live demo.
Pi is an open-source AI agent harness and toolkit from earendil-works for building and running coding agents. Its packages provide a unified multi-provider LLM API, an agent runtime with tool calling and state management, an interactive coding-agent CLI, a terminal UI library with differential rendering, and vendor-neutral telemetry contracts and adapters. Slack and chat automation are provided through a separate package project. Pi does not provide built-in restrictions for filesystem, process, network, or credential access; it runs with the permissions of its launching user and process. The project documents containerization and sandboxing approaches, including a local Linux micro-VM, Docker, and policy-controlled sandboxing, for stronger isolation.
fx is an open-source coding-agent harness and command-line interface written in Zig by Vercel Labs. It is designed as a compact Unix-like alternative to a terminal IDE, with interactive and one-shot requests for inspecting and modifying repository code, running shell commands, and saving or resuming sessions. The agent supports skills, MCP tools, plugins, subagents, permission rules, and headless requests, and can be embedded natively or through WebAssembly. It is model-agnostic, supports local and cloud inference, is distributed under the Apache-2.0 license, and is marked experimental by its repository.
Apache Maka (Incubating) is a local-first agent workspace developed under the Apache Software Foundation. It inspects projects and runs tools through a shared Runtime Host within a sandbox boundary; tools that leave the sandbox require approval. Model messages, tool calls, tool results, and how a turn ended are stored as recoverable execution facts on the user's machine, while older tool output can be omitted from later prompts without deleting the saved history. Model connections can use cloud APIs, local models, or compatible gateways, with sessions, settings, and run records kept local by default. Maka provides an Electron and React desktop application with streaming sessions, tool timelines, branching, search, recovery, artifacts, and model and sandbox settings. Its TUI/CLI supports work in a project directory and non-interactive turns, while its evaluation surface runs declarative multi-arm experiments across Maka and external subjects. Built-in tools include Read, Write, Edit, Bash, Glob, and Grep; Computer Use and catalog skills are optional. The project is under active development and Apache incubation; the README describes the macOS Apple Silicon desktop build as an early public release and notes that data formats, CLI commands, and experimental capabilities may change.
OpenBot is an open-source AI agent platform from CopilotKit that gives each agent its own computer with a browser, logins, files, tools, and a dedicated interaction channel. It runs in the user's infrastructure and accepts agents that speak the AG-UI protocol, including agents built with LangGraph, Mastra, CrewAI, Pydantic AI, Google ADK, or custom code. A single gateway mediates actions involving the computer, files, MCP servers, and UI components: it decides whether an action is permitted before execution and records it afterward. Agents can operate their own browser screen, hand control to a human for restricted actions, and return component-based responses. Docker Compose runs the application components and PostgreSQL, while the administrator supplies the model credentials; the repository describes the project as alpha software under active development.
Agent Deck is an open-source terminal dashboard for running and monitoring multiple AI coding agents in parallel. It displays each session’s status, active tool, working directory, and last prompt in real time, with Vim-style keyboard navigation for creating, focusing, closing, and renaming panes. It supports Claude Code and OpenCode through automatically installed hooks and runs as the single-binary `dot-agent-deck`, with native embedded terminal panes and no external terminal multiplexer required. The dashboard runs inside terminals including Ghostty, iTerm2, Alacritty, Kitty, and WezTerm, while preserving the agent clients’ existing shortcuts, skills, and configurations. Per-project TOML configuration can pair an agent with side panes for test runs, log tails, or `kubectl` watches. macOS and Linux installation is available through Homebrew, Nix, prebuilt binaries, or source builds; Windows use is currently through WSL. The project is distributed under the MIT license.
Macroscope is an AI-powered code review tool that analyzes pull-request changes and provides feedback on them, including changes generated by coding agents.
Grok Bot is an AI agent application from xAI that acts as an AI teammate on a persistent cloud computer. Its agents can perform research, monitoring, file-based workflows, automations, administrative tasks, and coding fixes through a computer interface, with documented system behavior and bot security boundaries. Subscription plans, usage limits, and billing are managed through Cursor.
Hyperagent is an AI agent platform for assigning work to a fleet of agents that research, build, deliver and maintain outputs using an organization's data, tools and working style. Its agents operate through real browsers and shells, can search, code, make decisions and generate artifacts such as live websites, videos, presentations, documents and dashboards. Agents run in individual computing environments, expose their execution as it happens, and can learn new memories and skills over time. Work is delivered through Slack, email, Telegram, webhooks or scheduled runs, and agents are deployed and managed from a central command center.
Praxist is an autonomous research system by Sapient Intelligence for computer-executable projects with measurable objectives. It turns an existing runnable project into a persistent research run: parallel research agents develop competing implementations or hypotheses, a task-defined evaluator converts their results into structured evidence, and a planning panel uses that evidence to set the agenda for later generations. Successful candidates and supporting evidence can move through incubator, frontier, and Gems retention lanes, while multi-metric evaluation, optional Quality-Diversity and Deep Innovation Gate allocation, resource scheduling, replay, and monitoring support longer runs. The task project remains responsible for its code, evaluator, metrics, baselines, prompts, roles, data, and domain constraints. Praxist can be installed as a Python package and operated through Codex, Claude Code, or its direct CLI; it requires CPython 3.11+ and a runnable project with measurable evaluation. The source is publicly available under the Fair Source License Agreement 1.0, with commercial terms described in the repository's license summary.
Munder Difflin is a free, open-source desktop multi-agent harness that runs multiple terminal-based AI coding agents as coordinated local agents. It wraps agents including Claude Code, OpenAI Codex, Gemini CLI, Qwen, OpenCode, GitHub Copilot CLI, and custom commands, working with users' existing subscriptions and hourly usage limits. Each agent runs as a real node-pty process and is rendered through xterm.js, while an Electron and React interface displays sessions as avatars on a Pixi.js office floor. A GOD orchestrator called Michael assigns and routes work, adjudicates agent messages, and escalates spending, destructive operations, scope changes, and other configured approvals. Agents coordinate through a local Git-backed hive of plain-file memory, atomic mailboxes, a shared blackboard, and an append-only event log; a markdown-first semantic memory layer supports recall across sessions. Optional Git worktrees isolate parallel agents.
OpenMAIC (Open Multi-Agent Interactive Classroom) is an open-source AI learning platform from THU-MAIC that turns topics or uploaded documents into interactive classrooms. It generates slides, quizzes, HTML simulations, and project-based learning activities, with AI teachers and classmates that can speak, draw on a whiteboard, conduct discussions, and respond to learners in real time. Its classic generation pipeline has two stages: an AI-generated lesson outline followed by scene generation for each outline item. The platform also provides a database-backed agent workbench that plans, builds, and revises courses through validated tools, with resumable sessions, follow-up steering, uploaded or web-retrieved materials, reusable skills, and support for importing PPTX files. Multi-agent orchestration uses a LangGraph director graph, while the playback and action engines handle classroom state and actions such as speech, whiteboard drawing, spotlights, and laser effects. OpenMAIC supports browser-only storage by default and can use PostgreSQL or S3-backed storage through its swappable storage packages. It accepts document, image, audio, and video materials through configured extraction providers, supports multiple LLM, media, speech, search, and local-provider configurations, and exports editable PPTX slides, interactive HTML, or classroom ZIP files. The repository is licensed under the MIT License, with separate terms for bundled components including an LGPL-licensed MathML-to-Office-Math package.
ECC is an MIT-licensed open-source agent-harness performance system maintained by affaan-m. It packages reusable skills, specialized agents, project rules, commands, hooks, memory, and security tooling for Claude Code as its primary target, with supported or limited adapters for Codex, Cursor, OpenCode, Gemini, Zed, GitHub Copilot, and other coding harnesses. Its core workflow turns plan, test, implement, review, verify, remember, and improve into reusable agent workflows. Skills are loaded for tasks such as test-driven development, research, security review, end-to-end testing, documentation, and refactoring; agents isolate planning, implementation, and review; rules provide always-loaded project or language standards; and hooks run event-triggered checks and session automation outside the model context. The optional Memory Vault stores inspectable Markdown handoffs and session context, while AgentShield scans agent files, hooks, MCP configurations, permissions, prompts, and secrets for security risks. ECC can be installed from its repository or through the ecc@ecc Claude Code plugin and ecc-universal package. The repository also provides selective installers, native or project-local integrations for several harnesses, a desktop dashboard, and the optional ecc-agentshield security-auditing package. Feature parity varies by harness: GitHub Copilot receives instructions and reusable prompts but not ECC hooks or agent delegation, while Codex has a native marketplace plugin with a narrower hook model.