1,018 tools and products — trending open source, and what gets used in AI and other work.
Lindy is an AI-powered assistant marketed as an "AI teammate" for workplace productivity. The product, presented at lindy.ai, is positioned to automate routine tasks and help users manage work and communications.
Firecrawl is an open-source web context API and hosted service for AI agents and applications. It searches the web, scrapes individual pages, crawls websites, maps site URLs, and batch-processes large URL sets, returning content as clean Markdown, HTML, screenshots, structured JSON, and other extracted data. It handles JavaScript-heavy pages, rotating proxies, orchestration, rate limits, and blocked content, and can parse web-hosted PDFs and DOCX files. Its interaction endpoint lets users or agents click, scroll, write, wait, and press on a page before extracting content; its agent endpoint gathers web data from a natural-language request without requiring URLs. Firecrawl provides Python and Node.js SDKs, cURL and CLI interfaces, and an MCP connection for AI agents and applications.
Anthropic API is the cloud API provided by Anthropic, the AI research company, offering programmatic access to its Claude family of large language models for text generation, chat, and reasoning. The API is intended for developers to integrate Anthropic's models into applications and services.
Tensor Processing Unit (TPU) is a family of application-specific integrated circuits (ASICs) and managed services developed and operated by Google to accelerate machine learning workloads. TPUs are available as on-premises hardware and as a managed Cloud TPU service for training and inference of neural networks and deep learning models. First announced in 2016, TPUs are used inside Google's datacenters and offered to external customers via Google Cloud.
gbrain is a hosted, cloud-based AI agent with persistent, long-term memory that adapts to a user’s preferences and way of thinking. It provides built-in skills for personal use and shared “brains” that support collaboration within teams and organizations.
gstack is an open-source skill and tooling framework for Claude Code, developed by Garry Tan. It turns Claude Code into a virtual engineering team through Markdown-based slash commands and specialist roles covering product review, engineering management, design review, code review, browser-based QA, security auditing, release management, documentation, and other software-delivery tasks. Its tools can drive a real browser and support unit-test, end-to-end-test, and black-box verification loops. The repository describes 23 specialist tools and eight power tools, and distributes the project under the MIT license. It requires Claude Code, Git, and Bun; Node.js is additionally required on Windows.
Gamma is a web-based presentation and document platform that uses generative AI to help create and edit slides and visual documents. It is developed and offered under the Gamma brand as an online productivity tool for creating presentations and shareable documents.
Cerebras Systems (branded Cerebras) is a company that designs AI‑optimized compute hardware and software for large‑scale model training and inference. Its products include the Wafer‑Scale Engine (WSE) and CS‑class systems, which provide high‑performance, large‑memory accelerators for deep learning workloads.
World Monitor is an open-source global intelligence dashboard developed in the koala73/worldmonitor project. It aggregates curated global and regional news, geopolitical information, infrastructure signals, and financial data into a unified situational-awareness interface, producing AI-synthesized briefs and tracking military, economic, disaster, and escalation signals. The application provides both a 3D globe and a WebGL flat map using a shared map-layer catalog, a server-authoritative Country Instability Index, and finance views for stock exchanges, commodities, crypto, and composite market signals. It supports local AI through Ollama without API keys, as well as other documented AI providers, and exposes an MCP server for programmatic access by agents and scripts. The repository supports multiple site variants from one codebase, including world, tech, finance, commodity, energy, and happy views. It also distributes a Tauri desktop application for macOS, Windows, and Linux, and documents self-hosting through Vercel, Docker, or static deployment.
Cloudflare OS is an open-source AI productivity environment developed by Cloudflare and built on Cloudflare Workers. It provides an agent chat interface preloaded with company-specific knowledge, sandboxed development of shareable applications called gadgets, and a security framework called Gatekeepers for controlling agents and apps. It is intended to be customized into an organization's own company operating environment rather than adopted unchanged as a traditional computer operating system. Each user receives a private, separately sandboxed instance of a productivity application, such as a slide-deck tool, whiteboard, game, or dashboard. Agents can create and modify these gadgets and perform tasks involving configured company systems and integrations. Gatekeepers are service-specific Workers that expose a Cap'n Web API, handle authorization such as OAuth, restrict access to the intended resource, log gadget and agent actions, and support human approval for side effects. Instead of stopping synchronously for approval, a Gatekeeper can simulate the outcome locally, let the agent continue and queue actions, then allow the user to approve or reject those actions later in bulk or individually. The full stack can be run locally with pnpm, Wrangler, and workerd using `pnpm run-local`, or deployed to a Cloudflare account. The repository describes the project as an early-access release under heavy development; the local setup is intended for evaluation rather than production use.
Semantica is an open-source Python infrastructure layer for building context graphs and knowledge graphs for AI systems. It ingests enterprise and other multi-source data, extracts entities and relationships, flags conflicting facts, merges duplicates, and supports ontology management, knowledge modeling, graph analytics, and causal reasoning. The system records provenance and execution trails for the context supplied to an AI system and the decisions produced from it, making those relationships and decisions queryable. Its repository describes deterministic graph construction, reasoning, and provenance that do not require an LLM; it explains the data and policies outside an LLM rather than exposing the model's internal reasoning. Semantica supports RDF and labeled-property-graph storage, W3C standards, self-hosted deployment, and installation with pip.
DeepTutor is an open-source, self-hostable AI tutoring platform developed by HKUDS for lifelong personalized tutoring. It provides a web application and command-line interface for interactive learning, problem solving, quiz generation, deep research, visualization, and mastery practice, with persistent memory and learning state shared across knowledge bases, books, notebooks, tutor personas, and other activities. The platform supports multi-agent problem solving, retrieval-augmented learning with source-traced and page-level citations, math animations, interactive visualizations, and tutor bots built from personal study materials. Its knowledge sources include personal documents, EPUBs and books with annotations, GitHub repositories, web search, and connected libraries; documented retrieval and ingestion options include GraphRAG, PageIndex, LightRAG, linked knowledge bases, Obsidian, and configurable parsing and vector backends. DeepTutor also includes courses, research workflows, an ecosystem of MCP services and tool or capability plugins, and integrations with connected coding agents and other partners. It can run locally or through Docker.
Tines is a software company that provides a secure, governed workflow automation platform for IT, security, and other teams. It supports building and operating AI agents, applications, and automated workflows, with visibility, access control, and administrative governance for maintaining control over automated processes.
quill is a fully local, ultra-minimalist macOS meeting recorder and transcriber developed by digimata. It runs as a single Swift menu-bar binary and uses macOS Core Audio process taps to capture microphone and system audio as separate CAF tracks without a virtual audio device or kernel extension. After recording stops, quill transcribes each track on the device, shifts the results by their start offsets, and merges the timestamped segments into a speaker-tagged transcript. The separate microphone and system tracks provide two-party diarization without a speaker-identification model. Each session produces the audio tracks, metadata, canonical JSON and Markdown transcripts, and a transcription log in a recordings directory. The default on-device transcription engine is Parakeet TDT 0.6B v2 through FluidAudio's Core ML port. It requires macOS 15 or later, supports optional launch-at-login operation, and can resume unfinished transcription jobs after relaunch. A WhisperKit fallback or re-transcription engine is planned.
OmniRoute is an open-source, self-hosted AI gateway developed by diegosouzapw. It exposes model providers through a single OpenAI-compatible endpoint and routes requests according to provider availability, quotas, cost, and latency, with automatic fallback and model chaining. The gateway is designed for local-first operation, keeping provider keys locally and supporting encrypted key storage. Its documented features include quota-aware scheduling, context compression through the RTK and Caveman mechanisms, and integrations for MCP and A2A workflows. The repository describes compatibility with coding clients such as Claude Code, Codex, Cursor, OpenCode, Cline, and Copilot, along with desktop and progressive web app interfaces. It is distributed under the MIT license and supports hundreds of providers and more than a thousand model identifiers.
Vercel AI Gateway is a managed service from Vercel that routes application requests to language-model providers, centralizes API key management, and provides observability, caching, and rate-limiting for LLM usage. It is intended to simplify integration of multiple LLMs and enforce access controls for apps and autonomous agents.
book-to-skill is an open-source agent skill that converts technical books, document folders, or collections of source files into a unified skill for GitHub Copilot CLI, Amp, or Claude Code. It accepts PDFs, EPUBs, office documents, plain text, folders, globs, and file lists; the videos also describe OCR support for scanned PDFs. The tool distills source material into structured knowledge rather than a single summary, extracting frameworks, decision rules, anti-patterns, and per-chapter Markdown files. It also generates a core SKILL.md with a chapter index, plus glossary, patterns, and cheatsheet files. Agents load the relevant chapter and supporting files on demand, allowing answers to be grounded in the source content without placing the entire book in the context.
jcode is an open-source, Rust-based terminal harness for AI coding agents and LLM workflows, designed for interactive development across multiple sessions. It supports resumable sessions, session search, multi-agent swarm and helper-agent workflows, compatible model-provider integrations, MCP, browser automation, and a self-development mode; installation scripts target macOS, Linux, and Windows. For automatic memory, jcode embeds turns and responses as semantic vectors, searches a graph of stored memories using cosine similarity, and feeds relevant results into the conversation. A memory sideagent can verify and expand retrieval, while periodic extraction stores new memories and ambient consolidation reorganizes entries and checks for staleness and conflicts. Explicit memory tools and traditional retrieval over previous sessions are also available. Its terminal UI includes side panels, diff views, inline Mermaid diagrams, information widgets, custom scrolling, and real-time rendering. The project supplies a Mermaid renderer without browser or TypeScript dependencies and publishes benchmarks covering RAM consumption, startup and input-readiness times, and memory scaling across multiple active sessions.
A minimal chatbot starter built with Next.js, the AI SDK, shadcn/ui, shadcn/react, shadcn/typeset, and the Vercel AI Gateway. It provides streaming chat with Markdown rendering, provider-native web search, tool calling, and a human-in-the-loop questionnaire in which the model can ask clarifying questions. The chat route streams responses with `streamText`, while the interface renders assistant messages as typed parts for text, GitHub repository lookups, web searches, questionnaires, and source URLs. Tool definitions are kept in separate files and composed through a central index; tool-part components display progress, results, errors, answered questions, and deduplicated search citations. Vercel deployments authenticate to the AI Gateway through OIDC, while local development uses an AI Gateway API key. The public unauthenticated chat route includes request validation, model restrictions, output and step limits, and client-disconnect handling, but the README recommends adding rate limiting, spending limits, and authentication before exposing it to real traffic.
Moli is an open-source headless browser engine built in Rust for AI agents. It fetches and extracts web pages, searches the web, and automates browser tasks while providing JavaScript, DOM, CSS, networking, storage, and browser automation capabilities. Moli treats the DOM as the source of truth and performs layout and rendering on demand, so DOM-first operations can avoid visual layout and paint. Layout enables geometry queries, coordinate input, screenshots, screencasts, and PDF output. It can be used through its CLI, CDP, WebDriver Classic, and WebDriver BiDi interfaces, including direct connections from Playwright over CDP, and supports Linux, macOS, and Windows.
Skills for Real Engineers is an open-source collection of small, composable agent skills by Matt Pocock for disciplined software development with Claude Code, Codex, and other coding agents. The skills support requirements clarification, project-language setup, test-driven development, debugging, code review, specification, planning, triage, and codebase surveys. The collection includes skills such as `/grill-me` and `/grill-with-docs`, which question the user about a proposed change before implementation, and a setup skill that configures an issue tracker, ticket labels, and documentation storage for a repository. It can be installed as a managed Claude Code plugin or copied into a project as editable files through the `skills` installer; the skills are intended to work with any model and can be adapted by the user.
DeepSeek Harness (dsh) is an open-source agent harness developed by DeepSeek AI. It uses an architecture in which everything is a plugin and is powered by Cordis, whose design is described in “A Programming Paradigm for Spatiotemporal Composability.” The npm package can launch a local Web UI with `npx @deepseek-ai/dsh web`, serving at `http://127.0.0.1:3080` by default and opening the default browser; it can also be built and run from a repository checkout. The project is in developer preview and warns that compatibility-breaking changes may occur. It is licensed under the MIT License, with third-party dependencies and their licenses documented separately.
Needle 2 is an open-source, 45-million-parameter foundation model for tool calling, device use, and structured extraction on small devices. Its weights are compressed into a single 14 MB binary and integrated with its inference engine; the repository reports that a full session uses about 28 MB of RAM and performs inference without network access after setup. The model converts text and tool descriptions into schema-constrained JSON calls using a byte-level grammar, attaches a learned confidence score for threshold-based escalation, and can retrieve the top five tools from a larger catalogue. It uses a 256-token sliding window with tools pinned as key-value sinks to bound memory, and its Simple Attention Network combines a Hadamard MLP, grouped-query attention, engram key-value memory, and multi-lane hyper-connections. The Python package supports inference, LoRA fine-tuning, and checkpoint export; the inference engine is fetched from Hugging Face and cached, while the training stack is an optional installation.
AgentENV (AENV) is an open-source distributed platform developed by kvcache-ai for running agent environments at scale, including environments used for agentic reinforcement-learning training such as Kimi K3. It runs Firecracker microVM sandboxes across machines, loads OCI-compatible images on demand through overlaybd, and uses local disks as bounded caches for image and snapshot data. The aggregate image and snapshot footprint can exceed local disk capacity because cold data is evicted from the cache. Snapshot-backed environments can boot or resume in under 50 ms and pause in under 100 ms. AENV incrementally snapshots memory and filesystem changes, stores snapshots in S3-compatible object storage or a shared distributed filesystem, and can fork a running environment into independent sandboxes for parallel workflows. It uses ublk for I/O and memory ballooning to return reclaimable guest memory to hosts; the documentation reports a 9.6× memory overcommit ratio in production. The distribution includes an AENV server and the `aenv` command-line client for pulling OCI images as templates, starting, attaching to, executing commands in, pausing, resuming, and deleting sandboxes. It exposes an E2B-compatible HTTP API that works with the standard E2B Python and TypeScript SDKs. It requires Linux kernel 6.8 or later and `/dev/kvm` access for Firecracker execution. API requests are authenticated, but traffic is not encrypted; the documentation recommends a trusted network or HTTPS termination at a reverse proxy or load balancer.
Diagram Design is a diagram-generation skill for Claude Code, Codex, Factory Droid, and Pi. It produces self-contained HTML and SVG diagrams through 39 editorial visual types, including architecture, flowchart, sequence, state-machine, data-flow, timeline, Sankey, Wardley map, kanban, user journey, deployment, dependency graph, UML class, story map, and database-schema diagrams. It can read a website to match its brand, and can redraw draw.io or Mermaid sources at a selected format, size, and detail level. Its semantic system patterns describe behavior separately from layout, so queues, policy traces, and trust boundaries can use an existing diagram type without adding new layout types. The repository supplies each type in minimal-light, minimal-dark, and full-editorial variants. Static HTML is the default; optional accessible motion is available for ordered explanations. The diagrams require no build step, JavaScript, or external image dependency.
Simile is a company and simulation platform that builds behavioral foundation models and synthetic "digital twins" of people to simulate populations and human decision‑making. The platform accepts scenario specifications, surveys and experimental designs as inputs and produces simulated populations, individual-level decision trajectories and aggregated outcome statistics that can be queried for policy analysis, product testing, forecasting and similar use cases. Simile is published from the vendor website and presents itself as a simulation platform for human behavior.
An AI software-development agent from Replit that builds, modifies, publishes, and deploys applications from natural-language prompts. It can generate user interfaces and databases, implement application workflows such as electronic signing and checkout, connect custom domains, and manage deployment.
GLM-5.3 is a version in the GLM series of large language models released by Zhipu AI (branded z.ai). It is presented on the Z.ai blog as a model release in the GLM family for natural language tasks.
Cumora is a cross-platform team-chat platform in which AI agents participate alongside human teammates. It provides shared rosters, direct messages, group conversations, a Kanban board, and a calendar; agents can maintain personas and memory, claim work, coordinate, and send and receive email. Agents can run in Cumora's cloud in per-agent Kubernetes pods, using a multi-hop tool-calling loop on the OpenAI Responses API, or through its BYOA mode, where a local daemon connects the platform to agent CLIs such as Claude Code, Codex, Grok Build, Cursor Agent, OpenCode, or pi. The frontend uses React, Vite, TypeScript, and Tailwind across web, desktop, and mobile shells; the backend is a stateless Node service using Express and WebSockets, with PostgreSQL as the source of truth and Redis for pub/sub and presence. Agent coordination uses a freshness gate that holds stale replies for reconsideration, atomic claims on work units, and a triage gate intended to reduce unnecessary large-model calls.
herdr is a terminal multiplexer and persistent runtime for coding agents, developed as a Rust binary. It runs a background server that keeps agent terminals and sessions available across terminal disconnections, SSH reconnections, network loss, lid closure, and machine restarts; users can reattach from another terminal and use tmux-style keyboard controls or mouse interactions to split, move, and manage panes. Each pane is marked working, blocked, or idle, and agents can control herdr through its CLI and socket API to spawn panes, prompt other agents, and wait for an agent that is blocked. It hosts existing tools such as Claude Code, Codex, Cursor, OpenCode, and Grok without wrapping or replacing them, and supports plugins for extending panes and workflows. The project provides installation scripts and package-manager installation, documents remote use and session state, and is licensed under the Apache License 2.0.
h3-metal is a native plain-C and Metal inference engine for the MiniMax H3 video model on Apple Silicon Macs. It provides a command-line interface and Iris-style interactive session for prompt-to-video and prompt-to-audio generation, including synchronized stereo audio, first- and last-frame conditioning, and ordered Ref2VA image, video, and audio references. The CLI can inspect model layout and the selected Metal device, run prompt-to-media jobs, and retain exact BF16 prompt conditioning, the prepared diffusion transformer, and video decoder in memory so repeated prompts with different seeds avoid reloading and re-encoding them. Configurable controls include denoising passes, transition reuse, transformer layers, frame dimensions, seeds, and output files. The project is developed through incremental vertical slices covering deterministic host and model metadata, portable Metal block parity, prompt encoding, media generation, conditioning, and H3-specific Metal performance and memory optimization.
OpenViking is an open-source context database for AI agents developed by Volcengine. It unifies agent memories, resources for knowledge retrieval, and skills in a virtual filesystem accessed through the viking:// protocol, allowing agents to browse context with filesystem-style operations such as ls, tree, and find instead of querying an opaque vector store. Content is processed into three loading tiers: L0 abstracts for relevance checks, L1 overviews for planning, and L2 full details loaded on demand. Recursive retrieval first locates a relevant directory through vector search and then drills down through its hierarchy, preserving the surrounding context and recording the browsing trajectory for inspection and debugging. After a session is committed, OpenViking asynchronously extracts user preferences and agent experience into long-term memory. The project also provides an OpenViking Studio browser playground, documentation, and a live demo.
Pi is an open-source AI agent harness and toolkit from earendil-works for building and running coding agents. Its packages provide a unified multi-provider LLM API, an agent runtime with tool calling and state management, an interactive coding-agent CLI, a terminal UI library with differential rendering, and vendor-neutral telemetry contracts and adapters. Slack and chat automation are provided through a separate package project. Pi does not provide built-in restrictions for filesystem, process, network, or credential access; it runs with the permissions of its launching user and process. The project documents containerization and sandboxing approaches, including a local Linux micro-VM, Docker, and policy-controlled sandboxing, for stronger isolation.
fx is an open-source coding-agent harness and command-line interface written in Zig by Vercel Labs. It is designed as a compact Unix-like alternative to a terminal IDE, with interactive and one-shot requests for inspecting and modifying repository code, running shell commands, and saving or resuming sessions. The agent supports skills, MCP tools, plugins, subagents, permission rules, and headless requests, and can be embedded natively or through WebAssembly. It is model-agnostic, supports local and cloud inference, is distributed under the Apache-2.0 license, and is marked experimental by its repository.
Harvey is an AI legal technology company that provides software for legal research, document analysis, drafting, and other professional legal workflows. It competes with legal-software platforms such as Legora.
Apache Maka (Incubating) is a local-first agent workspace developed under the Apache Software Foundation. It inspects projects and runs tools through a shared Runtime Host within a sandbox boundary; tools that leave the sandbox require approval. Model messages, tool calls, tool results, and how a turn ended are stored as recoverable execution facts on the user's machine, while older tool output can be omitted from later prompts without deleting the saved history. Model connections can use cloud APIs, local models, or compatible gateways, with sessions, settings, and run records kept local by default. Maka provides an Electron and React desktop application with streaming sessions, tool timelines, branching, search, recovery, artifacts, and model and sandbox settings. Its TUI/CLI supports work in a project directory and non-interactive turns, while its evaluation surface runs declarative multi-arm experiments across Maka and external subjects. Built-in tools include Read, Write, Edit, Bash, Glob, and Grep; Computer Use and catalog skills are optional. The project is under active development and Apache incubation; the README describes the macOS Apple Silicon desktop build as an early public release and notes that data formats, CLI commands, and experimental capabilities may change.
OpenBot is an open-source AI agent platform from CopilotKit that gives each agent its own computer with a browser, logins, files, tools, and a dedicated interaction channel. It runs in the user's infrastructure and accepts agents that speak the AG-UI protocol, including agents built with LangGraph, Mastra, CrewAI, Pydantic AI, Google ADK, or custom code. A single gateway mediates actions involving the computer, files, MCP servers, and UI components: it decides whether an action is permitted before execution and records it afterward. Agents can operate their own browser screen, hand control to a human for restricted actions, and return component-based responses. Docker Compose runs the application components and PostgreSQL, while the administrator supplies the model credentials; the repository describes the project as alpha software under active development.
Agent Deck is an open-source terminal dashboard for running and monitoring multiple AI coding agents in parallel. It displays each session’s status, active tool, working directory, and last prompt in real time, with Vim-style keyboard navigation for creating, focusing, closing, and renaming panes. It supports Claude Code and OpenCode through automatically installed hooks and runs as the single-binary `dot-agent-deck`, with native embedded terminal panes and no external terminal multiplexer required. The dashboard runs inside terminals including Ghostty, iTerm2, Alacritty, Kitty, and WezTerm, while preserving the agent clients’ existing shortcuts, skills, and configurations. Per-project TOML configuration can pair an agent with side panes for test runs, log tails, or `kubectl` watches. macOS and Linux installation is available through Homebrew, Nix, prebuilt binaries, or source builds; Windows use is currently through WSL. The project is distributed under the MIT license.
Macroscope is an AI-powered code review tool that analyzes pull-request changes and provides feedback on them, including changes generated by coding agents.
Grok Bot is an AI agent application from xAI that acts as an AI teammate on a persistent cloud computer. Its agents can perform research, monitoring, file-based workflows, automations, administrative tasks, and coding fixes through a computer interface, with documented system behavior and bot security boundaries. Subscription plans, usage limits, and billing are managed through Cursor.
JAX is a Python library for accelerator-oriented array computation and program transformation, designed for numerical computing and large-scale machine learning. It transforms Python and NumPy-style functions through automatic differentiation with `jax.grad`, automatic batching and vectorization with `jax.vmap`, and just-in-time compilation with `jax.jit`. It uses XLA to compile computations for GPUs, TPUs, and other accelerators. JAX supports compiler-driven multi-device execution: `Mesh`, `PartitionSpec`, and `NamedSharding` declare array placement across device meshes, allowing distributed training code to include automatically managed parallel work such as gradient synchronization. Its compilation tools also include ahead-of-time compilation, portable artifact export with `jax.export`, and conversion to TensorFlow SavedModels through `jax2tf` for serving integrations. The project describes itself as a research project rather than an official Google product.
TensorFlow is an open-source, end-to-end machine-learning platform originally developed by Google Brain. It provides tools, libraries, community resources, and stable Python and C++ APIs for building, training, and deploying machine-learning applications, with support for CPU and GPU execution and additional device plugins. In TensorFlow-based deployment stacks, it provides serving infrastructure and the SavedModel format; JAX graphs can be converted to SavedModel for integration with TensorFlow serving. The project is distributed through PyPI packages, Docker containers, and source builds, including CPU-only and nightly packages.
The Modular Platform is an open-source platform for AI development and deployment, comprising the MAX Framework and the Mojo programming language. Its repository includes the Mojo compiler and standard library, MAX accelerator libraries and kernels, Python-based model pipelines, code examples, and an OpenAI-compatible MAX inference server. The platform supports model serving and AI development while retaining Python-based workflows. The repository is licensed under the Apache License 2.0 with LLVM Exceptions, while MAX usage and distribution are licensed under the Modular Community License.
Hyperagent is an AI agent platform for assigning work to a fleet of agents that research, build, deliver and maintain outputs using an organization's data, tools and working style. Its agents operate through real browsers and shells, can search, code, make decisions and generate artifacts such as live websites, videos, presentations, documents and dashboards. Agents run in individual computing environments, expose their execution as it happens, and can learn new memories and skills over time. Work is delivered through Slack, email, Telegram, webhooks or scheduled runs, and agents are deployed and managed from a central command center.
Praxist is an autonomous research system by Sapient Intelligence for computer-executable projects with measurable objectives. It turns an existing runnable project into a persistent research run: parallel research agents develop competing implementations or hypotheses, a task-defined evaluator converts their results into structured evidence, and a planning panel uses that evidence to set the agenda for later generations. Successful candidates and supporting evidence can move through incubator, frontier, and Gems retention lanes, while multi-metric evaluation, optional Quality-Diversity and Deep Innovation Gate allocation, resource scheduling, replay, and monitoring support longer runs. The task project remains responsible for its code, evaluator, metrics, baselines, prompts, roles, data, and domain constraints. Praxist can be installed as a Python package and operated through Codex, Claude Code, or its direct CLI; it requires CPython 3.11+ and a runnable project with measurable evaluation. The source is publicly available under the Fair Source License Agreement 1.0, with commercial terms described in the repository's license summary.
Munder Difflin is a free, open-source desktop multi-agent harness that runs multiple terminal-based AI coding agents as coordinated local agents. It wraps agents including Claude Code, OpenAI Codex, Gemini CLI, Qwen, OpenCode, GitHub Copilot CLI, and custom commands, working with users' existing subscriptions and hourly usage limits. Each agent runs as a real node-pty process and is rendered through xterm.js, while an Electron and React interface displays sessions as avatars on a Pixi.js office floor. A GOD orchestrator called Michael assigns and routes work, adjudicates agent messages, and escalates spending, destructive operations, scope changes, and other configured approvals. Agents coordinate through a local Git-backed hive of plain-file memory, atomic mailboxes, a shared blackboard, and an append-only event log; a markdown-first semantic memory layer supports recall across sessions. Optional Git worktrees isolate parallel agents.
OpenMAIC (Open Multi-Agent Interactive Classroom) is an open-source AI learning platform from THU-MAIC that turns topics or uploaded documents into interactive classrooms. It generates slides, quizzes, HTML simulations, and project-based learning activities, with AI teachers and classmates that can speak, draw on a whiteboard, conduct discussions, and respond to learners in real time. Its classic generation pipeline has two stages: an AI-generated lesson outline followed by scene generation for each outline item. The platform also provides a database-backed agent workbench that plans, builds, and revises courses through validated tools, with resumable sessions, follow-up steering, uploaded or web-retrieved materials, reusable skills, and support for importing PPTX files. Multi-agent orchestration uses a LangGraph director graph, while the playback and action engines handle classroom state and actions such as speech, whiteboard drawing, spotlights, and laser effects. OpenMAIC supports browser-only storage by default and can use PostgreSQL or S3-backed storage through its swappable storage packages. It accepts document, image, audio, and video materials through configured extraction providers, supports multiple LLM, media, speech, search, and local-provider configurations, and exports editable PPTX slides, interactive HTML, or classroom ZIP files. The repository is licensed under the MIT License, with separate terms for bundled components including an LGPL-licensed MathML-to-Office-Math package.
ECC is an MIT-licensed open-source agent-harness performance system maintained by affaan-m. It packages reusable skills, specialized agents, project rules, commands, hooks, memory, and security tooling for Claude Code as its primary target, with supported or limited adapters for Codex, Cursor, OpenCode, Gemini, Zed, GitHub Copilot, and other coding harnesses. Its core workflow turns plan, test, implement, review, verify, remember, and improve into reusable agent workflows. Skills are loaded for tasks such as test-driven development, research, security review, end-to-end testing, documentation, and refactoring; agents isolate planning, implementation, and review; rules provide always-loaded project or language standards; and hooks run event-triggered checks and session automation outside the model context. The optional Memory Vault stores inspectable Markdown handoffs and session context, while AgentShield scans agent files, hooks, MCP configurations, permissions, prompts, and secrets for security risks. ECC can be installed from its repository or through the ecc@ecc Claude Code plugin and ecc-universal package. The repository also provides selective installers, native or project-local integrations for several harnesses, a desktop dashboard, and the optional ecc-agentshield security-auditing package. Feature parity varies by harness: GitHub Copilot receives instructions and reusable prompts but not ECC hooks or agent delegation, while Codex has a native marketplace plugin with a narrower hook model.
Comp AI CRM is an open-source, self-hostable customer relationship management system for AI agents, developed by Comp AI and hosted by Try Comp AI. Its agent runs as an independent durable deployment on its own schedule and database-backed work queue rather than waiting for browser requests: it selects records to investigate, researches contacts and companies, spends a research budget, schedules follow-ups and rechecks, records observed evidence, and sends weak or ambiguous matches to a human instead of writing them directly to records. It supports durable sessions, contact and company records, an Agent tab for viewing work and answering questions, authored tools for reading CRM history, searching records, identifying contacts, researching people, enriching companies, recording facts and scheduling rechecks, and versioned Markdown skills. The agent uses file-based tools and Markdown skills on Vercel's eve durable-agent framework, while its queue leases due tasks with PostgreSQL row locking so concurrent workers claim disjoint work and expired leases release tasks from failed runs. A restricted shell sandbox provides bash, grep, glob and a workspace without network egress, database credentials or direct database access. Optional sources include mailbox history, company brand data, LinkedIn and Perplexity web research.
An open-source Python tool from Pathway that generates original ARC-AGI-1-style grid puzzles with a distribution matched to the public evaluation set. It produces fresh tasks for private model evaluation, filters near-duplicates and malformed outputs, and writes a tasks.json file in the standard ARC format with train and test sections. The generated tasks are compatible with existing ARC evaluation harnesses and accompany Pathway's BDH-CQ model evaluation work.
IP as Logo Skill is a compact Agent Skill in the open Agent Skills format that guides compatible AI agents in turning a product brief into highly simplified, rounded mascot-character concepts. It gathers product context, proposes three design directions, and after approval generates six independent full-resolution square candidates using one dominant silhouette of roughly four to seven large shapes, two mascot or IP colors, a named solid background color, thick rounded forms, and lower-left or lower-right corner emergence. Its default batch uses two variants per direction with a three-left, three-right composition split rather than a contact sheet. Familiar animals are the default subjects, while objects, machines, fantasy artifacts, and other unusual subjects require a clear product-related reason. The skill provides instructions and prompts rather than an image generator, can be installed with the Agent Skills CLI, and is intended for agents such as Codex, Coze, Doubao, YouMind, Manus, Gemini Apps, and Replit Agent when paired with a supported image model.
Sol Advisor is a Codex-only orchestration workflow and plugin for capability-routed software delivery. Its primary Sol / High session receives the goal and constraints, declares a risk-gated route before using task tools, and owns planning, implementation or delegation, verification, and acceptance. The workflow supports solo, delegate, audit, and full routes: solo is the default; delegation assigns bounded or higher-risk implementation to a Luna or Terra auxiliary role while the root session verifies the result; audit and full routes add a fresh, read-only Sol / High review, and any requested fix requires another review. Auxiliary work substitutes for root work rather than duplicating it. The project limits auxiliary work to one by default and requires an explicit exception for the full route. It is installed as a Codex plugin with a companion script that installs and verifies the role files in a fail-closed manner.
VoiceStudio is an open-source, fully local desktop application for voice cloning and design, text-to-speech, transcription, dictation, video dubbing, and long-form audio production. It runs on macOS Apple Silicon, Windows, Linux, or Docker and keeps voices, projects, settings, and generated files on the machine by default. The application combines a Tauri desktop shell, a React and Vite interface, and a local FastAPI backend with registries for multiple TTS and ASR engines. Its workflows can transcribe and translate video, preserve speaker assignments, synthesize replacement speech, render audiobook chapters, separate vocals, diarize speakers, process batch jobs, and route work to CPU, CUDA, Apple Silicon MPS/MLX, ROCm, or optional remote workers. It also exposes local REST, SSE, WebSocket, OpenAI-compatible audio, and MCP interfaces, including speech synthesis and transcription endpoints. VoiceStudio is licensed under AGPL-3.0; downloaded models and tokenizers retain their own licenses. The local workflow requires no account, API key, subscription, or usage meter, although remote workers, configured external ASR endpoints, analytics, and other network-backed functions are opt-in. The repository describes the software as an active beta and distributes packaged desktop releases as well as source code.
fal is a generative media platform for developers that provides a unified API and SDKs for running image, video, audio, 3D, and code-generation models, including open models, user-provided LoRAs, and custom model endpoints. It offers a gallery of production-ready models and runs inference on a globally distributed serverless GPU infrastructure that can scale deployments from zero to large numbers of GPUs. fal also provides on-demand GPUs, compute for training and fine-tuning, private model endpoints, dedicated clusters for custom workloads, and observability tools for monitoring deployments and usage. The platform develops and post-trains open-weight models such as MiniMax H3 to improve generation speed, cost, and control for real-time video applications.
An MCP integration that connects compatible AI clients, including Claude and ChatGPT, to Higgsfield AI’s creative-generation workflows through a custom connector URL. Higgsfield generates and edits images, videos, and voice content from text prompts or references, supporting workflows for games, motion graphics, and interactive 3D experiences on the web and mobile.
Composio is a developer platform for orchestrating just-in-time tool calls, secure delegated authentication, sandboxed execution environments, and parallel execution across integrations with 1,000+ applications. It is designed to let applications and AI agents invoke external tools safely and at scale.
Amazon Bedrock is a managed AWS service that provides access to foundation models and APIs for building generative AI applications. It is offered by Amazon Web Services and enables use of Amazon's and third-party foundation models with managed infrastructure and tooling.
Enrich Directory is a web tool that extracts structured directory data from Google Maps reviews. It analyzes review text and matches evidence from reviews to populate missing business-directory fields, such as amenities, hours, and services.
LangFlow is an open-source, low-code visual builder for creating and running language-model pipelines, agent workflows, and retrieval-augmented generation (RAG) applications. The project is developed and maintained by the open-source community led by the GitHub user logaretm and integrates with language-model providers and tools such as LangChain.
An AI image-generation and image-editing model from Google, also known as Gemini 3 Pro Image. It generates images from text prompts and can create or modify specific visual compositions, such as depicting a model holding a supplied object.