703 tools and products — trending open source, and what gets used in AI and other work.
An open-source collection of Markdown output styles for Claude Code, Codex, and other coding agents. It changes how an agent communicates rather than how it codes: responses are answer-first, written in plain English, and formatted for skimming. The collection includes Attention-kind, which spaces points apart, uses arrows and bold emphasis, and expands only when useful; Spartan, a terse style with no warmth; and Rundown, a brief-briefing style. Each style is a single Markdown file that can be added and switched on independently.
OpenMausBot is an open-source, local-first chat application for managing a team of AI bots as separate contacts. Each bot can have its own personality, model, conversation thread, computer, and connected applications, using agents launched through locally installed Claude, Codex, or Grok command-line tools and existing user logins. A local harness server on 127.0.0.1 manages agent processes, transcripts, keys, and events in the user's local OpenMausBot data directory. Bots can work through a cloud Linux desktop, an isolated local virtual machine, or—where supported—the host computer. Shell commands, file edits, and questions are routed through a permission broker and shown as approval or response cards. The application includes model selection, bot and conversation management, browser-accessible desktop takeover, and connected applications through Composio, including services such as Gmail, Slack, GitHub, Notion, and Linear. The repository provides desktop distributions for macOS, Windows, and Ubuntu, with Ubuntu Wayland host control disabled according to the README while a stated issue is resolved.
Rakazo is an open-source platform for running persistent AI teammates. Each bot has conversations, memory, routines, and history, and can access a browser, terminal, files, and a graphical desktop. Users can access bots through the web app, Electron desktop app, or Expo mobile app, and can take control of the graphical desktop when needed. The platform supports shared team computers and isolated private computers, bot delegation to peer bots or short-lived subagents, voice interaction, and user-provided model credentials through Pi. It can use Docker, E2B, Daytona, Box, or trusted local computers as computer providers, and supports integrations through Composio, Pipedream Connect, MCP, OpenAPI, and Treg sources. Rakazo can be self-hosted with Docker and is also available as a managed application. The repository describes it as being in beta and identifies a TypeScript stack using React, Vite, Tailwind CSS, Electron, Expo, Hono, oRPC, PostgreSQL, Prisma, Better Auth, and Graphile Worker.
pgbot is a read-only PostgreSQL observability and health-diagnosis tool for AI agents, applications, and operators. A static binary connects to PostgreSQL, reads the database's own statistics views, and produces deterministic, findings-first health reports, including a health score, critical/warning/note findings, verified healthy subsystems, and change information from local baselines. It provides focused commands for indexes, queries, tables, and vacuum, along with versioned JSON output for agents and scripts and an MCP mode. Its optional ask and explain commands add an AI interpretation of the same findings; inspection and database queries otherwise require no model key or external service. The project is in beta and uses a read-only pg_monitor role with session-level read-only, statement, and lock timeouts as additional safeguards.
Dockhand is a Docker management web application by Finsys for managing containers, Compose stacks, and multiple local or remote Docker hosts. It provides real-time container status and resource monitoring, visual Compose editing, Git-based deployment with webhooks and auto-sync, interactive shell access, log streaming, file and volume browsing, image and network management, vulnerability scanning with Grype and Trivy, schedules for pruning and updates, notifications, and activity auditing. Environments can connect through a socket, agent, or direct TCP. The application uses direct Docker API calls, a SvelteKit frontend, a Bun-based backend, and SQLite or PostgreSQL through Drizzle ORM. It is licensed under the Business Source License 1.1, with the repository stating that it converts to Apache 2.0 on January 1, 2029; personal use, internal business use, non-profits, education, and evaluation are permitted, while offering it as a commercial hosted service is not.
img2threejs is an agent skill that reconstructs an object or character from a reference image as a code-only, procedural Three.js model. It generates a TypeScript THREE.Group factory using primitives, procedural shaders, and generated geometry rather than photogrammetry, mesh extraction, or downloaded art assets. The generated scene includes runtime structures such as pivots, sockets, and colliders for animation, and the project documents separate hard-surface and anatomy-aware reconstruction paths. It runs under Claude Code, Codex, or OpenCode and provides live browser demos whose models can be inspected as generated source.
Personal Model is an open-source, local-first AI memory runtime from Intuition Lab for giving coding agents evidence-linked context about a user. After macOS permissions are granted, it captures focused activity from applications on the user's Mac and builds a portable HUMAN.md-style personal context model, storing the data locally and allowing it to be inspected, corrected, exported, or deleted. The runtime progressively organizes sourced observations and events into relationships and changes, supported patterns, higher-order structures, and an integrated current model. It retains source receipts for important claims so later evidence can strengthen, revise, or overturn inferences, then exposes this context through MCP to trusted clients including Claude Code, Codex, Cursor Agent, Claude Desktop, and other compatible tools.
AGENTS.md is an open format for guiding AI coding agents through project-specific context and instructions. It provides a dedicated, predictable file—similar to a README for agents—where projects can document development-environment commands, testing and linting procedures, package or project conventions, and pull-request requirements. The format is documented by the agentsmd project, which also maintains a basic Next.js website with examples.
An Apache-2.0-licensed open-source Go SDK from Grafana for building AI-powered backends and agents. It provides GenerateText and StreamText APIs for calling language models, returning complete or streamed responses, executing Go functions as tools, generating schema-validated structured output, and running multi-step tool workflows, with approval for consequential tool actions. The SDK supports Anthropic, Amazon Bedrock, OpenAI, OpenAI-compatible APIs, and Grafana's hosted endpoint, along with retries, configurable timeouts, fallbacks, logging, Prometheus metrics, and Agent Observability. It follows the design of Vercel's AI SDK and is wire-compatible with its TypeScript React frontend hooks, including useChat, useCompletion, and useObject, allowing Go backends to stream Server-Sent Events directly to an existing React frontend without a protocol adapter.
Zeron is a local control layer for coding agents, including Claude Code, Codex, Cursor, Grok, Hermes, and Pi. Each device runs a small engine that stores agent sessions locally by default, without requiring an account or network connection. Optional synchronization lets users start an agent on one device and follow or control it from another; an always-on machine such as a VPS can keep agents running after a laptop is closed. The Linux installer starts a daemon that persists across reboots, while macOS users can install the desktop release or build from source. Zeron is licensed under the MIT License. The videos refer to the project as Comet.
icm-architect is a Claude skill that designs processes, ideas, and problems as ICM (Interpretable Context Methodology) workspaces, using folder structure as agent architecture, or restructures an existing folder, repository, or vault. Numbered folders encode sequencing, hierarchy scopes context, and plain Markdown files store state, allowing an agent to orient and act by reading the relevant files. In build mode, it extracts stages, human approval points, and stable versus per-run elements, then scaffolds a workspace using one of six forms: Pipeline, Umbrella, Record library, Knowledge bundle, Context map, or System map. In restructure mode, it audits files as catalog, contract, factory, product, or dead, proposes a migration map for approval, then migrates and validates the result. The walk test checks whether an agent with no prior memory can orient, act, and report status from the workspace alone. It can be installed as a Claude Code skill or uploaded to Claude applications, and is MIT licensed.
unlazy is an agent skill for enforcing completion discipline in substantial AI-agent tasks. It requires an acceptance ledger to be written first, with runnable CHECK commands and EXPECT conditions; its gate checker reviews commands, executes approved gates, records evidence, and can reverify completed work. The skill also supports a Depth Tree method that splits tasks into fresh-context subtasks with integration checks, giving each leaf the full task time budget. It can be installed for supported agents with the skills CLI or manually for Claude Code and Codex CLI; its checker and optional hook require Node.js 16 or newer and no third-party runtime packages.
Endoplexity is a Chrome MV3 side panel and local bridge for agentic control of the browser a person is already using. It lets Claude or Cursor agent CLIs operate existing tabs and logged-in sessions by clicking, typing, filling forms, navigating, switching tabs, uploading files, and reading pages, without sending requests directly to a model or requiring a separate metered API key. The side panel owns the Chrome DevTools Protocol connection and communicates over an origin-pinned WebSocket with a bridge on localhost; the bridge exposes browser tools over token-gated MCP/HTTP and enforces the safety policy, including approval gates and autonomy modes. Pages are provided to the agent as accessibility-tree snapshots rather than raw HTML, and actions return the resulting page state so subsequent turns can resend only changed lines. Its file-reading tool extracts text from PDFs, DOCX, XLSX, PPTX, CSV, JSON, Markdown, and other text files.
NorthCinder is an open-source, local-first MCP server for AI shopping agents. It compares products from user-selected sources against a buyer's criteria, explains rankings and exclusions, shows source facts and unverified details, and normally returns up to three useful options rather than claiming one universal best product. Its research workflow provides separate product and seller guides, creates a research plan, and uses controlled research tools; results remain provisional when sources conflict or do not identify the exact item or seller. Purchasing is a separate human-in-the-loop decision: each checkout requires fresh approval for one exact offer and unit, with the merchant, variant, price, known total, and spending cap signed into a single-use approval. It rejects raw card details, using an opaque payment token for supported automated checkout or handing the buyer a cart link. NorthCinder runs alongside an MCP-capable AI application. Its local setup requires Node.js 20 or later; the init command stores configuration locally and starts the MCP server and search engine in one process. The repository states that there is no hosted NorthCinder account or cloud service, and that order outcomes remain local.
TrueForge is an open-source, vendor-neutral agent harness developed by TrueFoundry. It runs the agent execution loop, including model calls, MCP tool use, skills, sandboxed code and file execution, approvals, context management, and session state. Agents can use configured model providers, remote MCP servers, git-backed SKILL.md instruction packs, and on-demand sandboxes; context features include subagents, deferred tool loading, Code Mode, large-result offloading, and compaction. TrueForge exposes agents through a bundled chat UI, an HTTP API with a TypeScript SDK, and an embeddable UI SDK. It supports local mode with SQLite and hosted deployments using Postgres and Redis through Docker Compose or Helm; the repository warns that local mode is intended for personal use on localhost rather than production or internet-facing deployment.
sloptrim is a local prose linter and coding-agent plugin that detects patterns associated with AI-generated writing when an agent saves a document. It extracts prose from supported text and office formats, scores it on a 0–100 scale, and reports suspicious spans for revision; it analyzes writing patterns rather than determining authorship and does not process code. The command-line tool uses Python's standard library, requires no model or network connection, and keeps text on the local machine. It integrates with Claude Code and can be initialized for other agents through an agent contract and editor-rule files.
J-Space Cognition Suite is a model-agnostic, inference-time control system packaged as a cross-platform AI-agent Skill for deep reasoning, long-horizon work, tool use, verification, and recovery. It leaves model weights and training unchanged, organizing an agent's working representations through one entry point and nine selectively loaded modules supported by four references. Its fast, full, and loop operating modes use selective workspace loading, a shared broadcast hub for constraints and values, dense internal reasoning traces, explicit intermediate steps before conclusions, metacognitive routing of confidence and failure signals, and bounded empirical verification. An optional standard-library controller records durable task state, including goals, next actions, checkpoints, open questions, seams, and recovery state. The suite is intended for low-friction integration with AI hosts that support Skills or equivalent system/developer-instruction mechanisms.
Docker Sandboxes is a Docker security feature for running AI coding agents and tools inside isolated micro virtual machines. Each sandbox has its own kernel, filesystem, and network, allowing agents such as Claude Code to run with reduced access to the host machine.
Iris is a Rust command-line tool and local stdio MCP server for capturing screenshots of live websites with an installed Chrome-family browser such as Chrome, Chromium, Edge, or Brave. It can capture viewport or full-page images, emulate mobile sizes and user agents, apply dark color schemes, frame the first matching element with optional padding, wait for selectors or page content, and process URL batches concurrently or from standard input. The same binary exposes an MCP `capture` tool that returns an inline image with structured metadata and can optionally write an output file. It is distributed as the `iris-screenshot` crates.io package or through an install script; building from source requires Rust 1.88 or newer.
Wake is a native desktop application that gathers coding-agent sessions from local data directories into a single searchable library. Built with Rust and GPUI, it reads supported agent histories read-only, groups them by agent and project, watches files for incremental updates, and indexes transcripts with SQLite FTS5 trigram search, including code substrings and CJK text. Its transcript view renders messages, collapsible tool-call clusters, thinking summaries, and syntax-highlighted code. Wake can resume supported conversations in a terminal at the original project directory, and also provides session management, Markdown export, deletion with tombstones, and library statistics. Data remains local; the application does not make background network requests and only contacts GitHub when an update is explicitly checked. It is macOS-first, with experimental Linux support and experimental Windows support.
Amagine3D is an open-source 3D capability layer for hardware creation, developed by Amagine. Its parametric CAD workflow turns natural-language hardware requirements, reference images, and key dimensions into editable enclosures and assembly structures around internal components, producing Python and build123d source code alongside STEP, STL, or color-aware 3MF exports. A 3D-native agent first organizes the requirements into a design brief, then runs the generated source in a browser geometry runtime. It uses the resulting model state and check results for dimensions, part connectivity, interference, and motion clearance to revise or accept the design. The workflow supports multipart structures such as covers, hinges, and latches, including assembly clearances and printing tolerances. Generated parameters can be adjusted in the workbench and written back to the source without regenerating the model. The system can preview, measure, modify, and save complete designs and their check reports.
Google Cloud Shell is a browser-accessible command-line environment hosted by Google Cloud for managing projects, creating files, configuring permissions, and deploying resources. It provides a temporary virtual machine with common cloud-development tools and a persistent home directory. In the cited video context, it is used to run commands for an AI agent and execute Kubernetes commands.
macOS Harness is an experimental, MIT-licensed, macOS-only Python harness from browser-use that gives an LLM or coding agent raw primitives for controlling a Mac in one persistent Python process, without app-specific tools, recipes, or framework rails. Its six core primitives—`see`, `key`, `type`, `click`, `ax`, and `script`—capture native and Electron app windows, send keyboard and coordinate input directly to an application process, expose Apple Accessibility and Apple Events, execute AppleScript, and draw a click-through pointer without moving the physical cursor. The agent can write missing task logic in ordinary Python during execution. The same process provides access to a real, logged-in Chrome browser through Browser Harness, as well as the local filesystem and shell commands through Python, `Path`, and `subprocess`. It can capture background application windows without bringing them to the foreground and does not activate or raise the target app. The project includes installation and skill-registration workflows, a `doctor` command for checking required macOS permissions, and verification instructions. Anonymous telemetry is enabled by default and records the CLI command category, success, duration, package version, operating system and architecture, and detected agent client; the project states that it does not record prompts, app names, screenshots, UI text, scripts, paths, or window titles, and provides `macos-harness telemetry disable` to turn telemetry off.
OpenSpec is an open-source, tool-agnostic spec-driven development framework from Fission AI for AI coding assistants. It stores requirements and development artifacts as plain Markdown files in a repository, using an `openspec/` structure containing specifications and proposed changes. Its workflow supports exploratory discussions, `/opsx:propose` for generating a proposal, requirements and scenarios, a technical design, and implementation tasks, followed by `/opsx:apply` to implement the tasks and `/opsx:archive` to update the specifications. Coding agents can interact with it through slash commands or an MCP server. The project is designed for iterative, brownfield development and supports shared or cross-repository requirements through beta Stores. The CLI requires Node.js 20.19.0 or higher and is installed with npm.
Jenkins is an open-source automation server for continuous integration and continuous delivery (CI/CD). It automates software builds, tests, and deployments through Pipeline definitions and a large plugin ecosystem, and supports distributed build agents across multiple operating systems.
LatticeDB is an embedded, single-file property-graph database written in Zig for local applications. It combines relationship traversal, HNSW vector similarity search, and BM25 full-text search in one query language and engine, allowing graph, semantic, and textual queries over the same dataset. It also provides durable named event streams and a graph changefeed through the same transaction and write-ahead-log path as graph writes. The database operates without a server or configuration and is designed for one owning process on one machine with WAL-backed durability. It is distributed through a CLI and bindings or packages for Python, TypeScript/Node.js, Java, and Go. Graph RAG, agent memory, and local knowledge tools are documented as example workloads rather than the engine's definition.
CarWatch is an open-source, offline in-car AI agent from ThinkOffApp that runs on a Raspberry Pi 5 with a locally hosted language model. It joins chat rooms as a vehicle agent, answers questions from the car's owner's manual using lexical retrieval-augmented generation and page citations, and reads live vehicle and system data while stating when information cannot be sensed. A continuous energy-based voice listener and whisper.cpp provide local speech input without a wake word, with responses played through the car's speakers. The system can monitor OBD and engine data, send departure and arrival messages, trip summaries, and dashcam clips, and expose a phone dashboard for approvals, replies, model selection, and maintenance. The repository describes systemd-managed services for the model server, room agent, voice listener, dashboard, and engine watcher, with Python standard-library code and no cloud subscription required.
Factory is a reference software factory for Claude Code and Codex that installs a repeatable, version-controlled software-delivery workflow into an existing GitHub project. GitHub Issues serve as the work queue, while committed policies, skills, labels, handoff comments, pull requests, and run records preserve state between fresh agent sessions. Scheduled agents triage issues, route them to implementation, specification, questions, or blockers, claim bounded work, implement it on a branch, run configured type, lint, test, build, audit, and architecture checks, obtain independent verification, and open draft pull requests. An independent verifier reads the diff and checks that the new test fails without the implementation. Humans retain responsibility for ambiguous requirements, system design, significant changes, and merging pull requests. The repository has no custom orchestrator or queue service. Claude Code routines provide the default scheduling and compute, with a thin Codex adapter using the same policies, gates, and evidence files; GitHub events or an optional API-triggered Action can also start runs. It is configured through files such as a human-owned charter and gates script, and includes a /factory control room backed by live issues, pull requests, and run records.
FrontierAgent is an open-source agent runtime, terminal product, and evaluation suite for long-horizon research and file-based work. Its native terminal TUI provides a ReAct workflow in which one stateful agent researches, reads files, writes deliverables, runs commands, and iterates in a task-scoped sandbox, and an Agent Team workflow in which a coordinator maintains a task board, delegates bounded independent assignments to parallel sub-agents, collects structured reports, and synthesizes the result. Shell and file tools use a shared sandbox with read-only /inputs, working-state /workspace, and persistent /outputs directories. The TUI supports queued instructions during execution, approval-required diffs for mutating operations, local action traces, checkpointed sessions, resumption, and reversion. The same workflow engine powers a subprocess benchmark runner with deterministic artifact collection, concurrency, progress inspection, and reruns of individual failures. The repository also documents connection to the OpenAI-compatible Apodex-1.1 endpoint through the Apodex API.
ai-memory is an open-source local server for long-term memory and session handoff among AI coding agents. It captures prompts, tool calls, and agent context through MCP integrations and lifecycle hooks, storing the material in a searchable, Git-backed Markdown wiki. When a session ends or is explicitly finalized, it generates handoff context for a subsequent session, including the project architecture, failed approaches, and open questions, so work can continue across agent vendors without restating the context; clients without a true session-end hook use explicit finalization commands. The project documents integrations for Claude Code, Codex, Command Code, Devin CLI, OpenCode, Cursor, Gemini CLI, Oh My Pi, Pi, and Crush, with agent-specific configuration, generated plugins or extensions, and capture exclusions. It runs on Linux, macOS, and Windows through WSL2, with experimental native Windows support, and is distributed through Docker images and native release binaries.
AI Engineering from Scratch is a free, open-source curriculum developed by rohitg00 for learning to build and ship AI systems end to end. Its 20-phase sequence covers development tooling, mathematical and machine-learning foundations, deep learning, NLP, computer vision, generative AI, LLM applications, agent engineering, Model Context Protocol (MCP), agent skills, reinforcement learning, and related topics, with implementations in Python, TypeScript, Rust, and Julia. The repository describes 511 lessons and approximately 329 hours of material; each lesson produces a reusable artifact such as a prompt, skill, agent, or MCP server. Learners are instructed to read the lesson documentation, type and run the code from the repository root, record command evidence and outputs, and make a small change before continuing. It provides routes for complete foundations, mathematics and machine learning, production LLM applications, agent engineering, MCP, agent skills, and Claude certification preparation. The repository is distributed under the MIT license and is accompanied by a website containing the same lesson content.
Bizee is a business-formation and compliance service for creating LLCs and other business entities. Its services include registered-agent management, filing-deadline support, EIN filings, operating agreements, and related business documents.
Experiential is an open-source gateway and router for agent workflows from Experiential Labs. Its `exp` CLI provides hosted, bring-your-own-key, local, and custom models through OpenAI-compatible and Anthropic Messages APIs, with model aliases, access controls, use-case restrictions, and spending limits for users and agents. It can persist provider connections and configuration locally, expose API routes on loopback, and load a fitted project router as an OpenAI client through its Python interface. It collects OpenTelemetry traces from agent traffic to build simulations and optimize routing for quality, speed, and cost, and supports harness optimization, endpoint serving, model distillation, and fine-tuning an owned open-source model through Tinker. The hosted gateway is available at `api.experientiallabs.ai`; anonymous aggregate PostHog telemetry is enabled by default locally and can be disabled.
Doop is an open-source multiplayer design canvas where people and AI agents create and edit designs together. Each canvas contains frames that render real HTML in sandboxed iframes; humans edit in the browser, while agents connect through the built-in Model Context Protocol (MCP) server and can create frames, stream HTML in chunks, inspect screenshots, and revise designs. The app synchronizes cursors, presence, frame edits, agent status, comments, tasks, and activity over WebSocket rooms, and includes a server-side Doop Agent that can process queued cards, mentions, and feedback through specialist roles. Canvases are private by default, with email invitations or optional link sharing, and agents inherit the access of the user who authenticates them. Doop also provides design-memory features for exemplar frames, decisions, and proposed style rules, and can be self-hosted with Docker Compose or Bun using embedded Postgres through PGlite.
ego lite is a macOS browser from Citro Labs designed for AI-agent browser automation alongside a user's normal browsing. It gives each agent or task an isolated Space within the same browser, allowing multiple tasks to run in parallel without taking over the user's tabs; users can observe, take over, or stop an agent's Space. Through the ego-browser skill, agents such as Claude Code, Codex, Cursor, or custom agents can call in-page JavaScript tools including snapshot, fill, click, wait, navigate, and capture; an agent can compose several operations into one JavaScript execution rather than repeatedly issuing CLI commands. During setup, optional Chrome-data migration can transfer existing logins, cookies, extensions, and bookmarks, while browsing data remains on the user's device. The browser is distributed as a separate free download, the repository is MIT-licensed, and Windows and Linux support are listed as planned.
SwarmForge is a local, tmux-based orchestration platform for coordinating multiple AI agents across Git worktrees. It reads a project-local configuration that assigns roles, agent backends, worktrees, task or batch handling, and handoff propagation; launches each role in an isolated tmux session; and provides a browser-based pack cockpit for projects, tasks, approvals, clarifications, live status, agent-pane inspection, and teardown. Agents communicate through a daemon-delivered handoff protocol and helper scripts rather than direct tmux messages. The repository supplies two-pack, four-pack, and six-pack workflow templates with role prompts and constitution articles, while projects can define their own swarm topology. It runs locally with zsh, Git, tmux, Babashka, and at least one supported agent backend such as Claude, Codex, Copilot, or Grok; swarm state is kept in the working director
Rome is an agentic operating environment from Rome OS for collaboration between humans and AI agents. It provides manifests, typed actions, agents, skills, hooks, tools, workflows, memory, interfaces, policies, and persistent database-backed state in a guardrailed environment where agents can build applications, define standard operating procedures, and orchestrate workflows under human guidance. Rome supports one-off tasks, scheduled work, and long-running follow-through. Rome Apps combine a purpose-built web interface, agent reasoning, agent-owned collaborators, reusable workflows and skills, lifecycle hooks, typed actions, HTTP APIs, and persistent app-private databases or files into installable products. Apps are defined with an app.yaml manifest and use the @rome-os/app-runtime and @rome-os/app-web-sdk packages; Rome can generate and iterate on app or workflow source code from a plain-language request. The project can run locally through Docker or be accessed through Rome Cloud, which the repository describes as a private preview environment.
Proliferate is an open-source AI IDE for running coding agents such as Claude Code, Codex, OpenCode, Cursor, and Grok through their native harnesses. It runs agents in parallel within one workspace, giving each task an isolated Git branch and worktree along with its own terminal, conversation, and review state. Agents can delegate scoped work to subagents, while MCPs, skills, Computer Use, Browser Use, and custom tools can be configured once and shared across agents. Its workflow system supports recurring and event-driven runs such as review passes, alert triage, and dependency updates. The desktop application can use a local runtime or connect to a control plane that runs locally or in the cloud. The control plane is self-hostable through Docker, AWS, GCP, Azure, Kubernetes, or air-gapped deployments. The repository lists Rust, Node.js, and pnpm as source-build requirements and is licensed under AGPL-3.0.
MiniMind is an open-source tiny large-language-model project and end-to-end training tutorial developed by Jingyao Gong. Its main dense model is approximately 64M parameters, and the repository provides the model architecture, tokenizer, datasets, inference code, and training pipeline. The pipeline covers pretraining, supervised fine-tuning, hand-written LoRA, DPO, PPO, GRPO, CISPO, model distillation, tool calling, adaptive thinking, and agentic reinforcement learning. Core algorithms are implemented directly with native PyTorch rather than relying on high-level abstractions from third-party training libraries. Its agentic reinforcement-learning path performs multi-turn rollouts, executes generated tool calls, appends tool observations to the context, calculates trajectory-level rewards, and updates the policy; rollout can use local PyTorch generation or an SGLang server. MiniMind supports dense and mixture-of-experts variants, single- and multi-GPU training, YaRN-based RoPE length extrapolation, Transformers-format models, and inference through llama.cpp, vLLM, or Ollama. The repository also includes a Streamlit chat interface and a lightweight OpenAI-compatible API server with tool-call and reasoning fields. It is released under the Apache License 2.0.
Codewhale is an open-source coding agent for the terminal, built in Rust and independently maintained. It reads repositories, edits files, runs commands, inspects results, and works toward user-defined goals. It connects to hosted providers or local models through Ollama, vLLM, or SGLang, supports switching providers and models during a session, and can run interactively in a terminal UI or non-interactively with the `exec` command. Its control mechanisms include read-only planning, Ask, Auto-Review, and Full Access approval modes; `/undo` reverts the last turn and `/restore` returns the workspace to an earlier snapshot. It also supports saved sessions, durable goals, reviewable workflows, agent coordination, MCP servers, skills, hooks, and configurable agent roles. The project runs on the user's machine with the access granted to it, with optional operating-system sandboxing where supported, and is distributed under the MIT license.
TrustMeBro is a Go command-interception tool for controlled red-team testing of coding agents such as Codex, Claude Code, and pi. It uses PATH shims to intercept configured command-line tools without requiring a plugin, hook, or MCP integration; rules can spoof generated or fixed output, rewrite stdout from a real command while preserving stderr and its exit status, pass through to the real binary, or reject the call. Each decision is recorded in a timestamped JSONL audit log. Rules match command names, domains, DNS record types, argument globs, and regular expressions. The bundled DNS generators support dig, nslookup, and host, while custom shims can use fixed output, exit codes, and standard streams. On Linux, lab mode uses Bubblewrap to shadow PATH lookups and discovered absolute paths so an agent cannot bypass interception merely by invoking a resolved system binary. TrustMeBro is distributed as a single MIT-licensed Go binary with installation, configuration validation, status, rule-listing, and uninstall commands. Lab mode requires Linux and Bubblewrap and is explicitly an interception namespace rather than a security sandbox: it reuses the host filesystem, workspace, network, environment, and agent credentials. Outside lab mode, absolute paths, changed PATH environments, and in-process DNS clients can bypass command shims.
HexStellar is an agent-first computational platform whose Cortex service executes structured optimization, decision, scientific-computing, and verification requests through a Python CLI and API. An agent submits a JSON formulation for problems such as QUBO/Ising optimization, maximum cut, traveling-salesperson routing, facility assignment, mixed-integer optimization, selection, ranking, scheduling, coloring, and business-rule feasibility; the managed service returns a structured result with execution metadata, a receipt, and an assurance label distinguishing certified optima, heuristics, operations, and abstentions. The CLI also provides free validation, estimation, service re-checks, local witness recomputation for supported families, reproducible seeds and model versions, batch and compressed-binary transport, and an MCP server over standard input/output, with read-only mode for free analysis and verification tools. The public package is a zero-dependency Python thin client: the proprietary solver runs on HexStellar-managed infrastructure rather than inside the package. A separate enterprise runtime is licensed for compatible customer-controlled compute paths under NDA; it has no public runtime download in version 1.0.
Editable Visual Design is an open-source, coding-agent-driven toolkit for creating editable visual artifacts from prompts. A persistent coding agent interprets a brief, plans a composition, generates assets, implements the design in semantic HTML, renders and observes the result, and applies repairs. An image model supplies visual direction for composition, hierarchy, color, and spatial relationships, while the delivered artifact rebuilds typography and layout in HTML rather than shipping reference pixels. The resulting artifact contains real text, independent assets, selectable and movable layers, a visual editor, rendered PNG output, an animated layer breakdown, and a replayable creation process. Deterministic checks cover the canvas, fonts, layer contracts, rendering, and editor round trips. The repository distributes the workflow as two Codex skills, including editable-design for fixed-canvas designs and a separate html-to-pptx skill for converting compatible HTML designs into editable presentations.
Magnitude is an open-source local inference server and CLI for running language models with AI agent harnesses. It profiles a machine's chip, memory, and bandwidth, recommends compatible models, downloads the selected models, tunes inference settings such as speculative decoding and concurrency, and runs models on demand. Models are loaded when needed and unloaded when idle or when memory is constrained; prompts, files, and models remain on the local machine, allowing offline operation after setup. It integrates with Pi, OpenCode, Hermes, OpenClaw, Codex, Claude Code, Oh My Pi, and Cline, and also provides a built-in harness. Magnitude supports macOS and Linux, with Windows support through WSL, and is licensed under Apache 2.0.
Chrome DevTools for agents (chrome-devtools-mcp) is an MCP server from the Chrome DevTools project that lets coding agents control and inspect a live Google Chrome or Chrome for Testing browser. It uses Puppeteer for browser automation, including navigation and actions that wait for results, and exposes DevTools capabilities for recording and inspecting performance traces and insights, analyzing network requests, taking screenshots, and inspecting browser console messages with source-mapped stack traces. A CLI is also provided for use without MCP, with configuration options including headless, isolated, slim, concurrent-session, persistent-profile, WebSocket, and Android-debugging modes. It requires Node.js LTS, a current stable Chrome version or newer, and npm. Browser contents are exposed to MCP clients; performance tools may query the Google CrUX API, and usage statistics and update checks are enabled by default with an opt-out flag.
Monocolor Editorial Print is an agent skill for generating one-ink or controlled two-ink editorial images for posters, zines, portraits, packaging, and visual field notes. It accepts a theme, phrase, object, article idea, or supplied photograph and produces a raster image when image-generation tools are available, the exact production prompt, and a short recipe describing the print mode, palette, layout, typography, process, and originality changes. The skill identifies the subject and intent, selects a layout family, assigns one or two ink plates, composes the page with visible paper and an asymmetric grid, and generates and inspects the result. Its visual system uses adaptive white, gray, or beige substrates; halftone, risograph grain, cyanotype exposure, or photocopy breakup; responsive typography; and controlled negative space. Without image generation, it returns the prompt and states the limitation.
Sierra is an AI agent platform for businesses. Its agents use tool calling to handle customer support, sales, and other pre-sale and post-sale workflows, with an outcome-based pricing model.
HyperFrames is an open-source framework for turning HTML, CSS, media, and seekable animations into deterministic MP4 videos. It can be used locally through a CLI, by AI coding agents through skills, or as a rendering core for hosted authoring workflows. Compositions are defined as HTML with data attributes for timing, tracks, media, and sub-compositions. The renderer seeks each frame in headless Chrome, uses animation adapters such as GSAP, CSS, Lottie, Three.js, Anime.js, or WAAPI, and encodes the result with FFmpeg; its toolchain includes preview, linting, inspection, local rendering, browser-based editing, reusable catalog components, and AWS Lambda rendering. The project requires Node.js 22 or newer and FFmpeg, does not require React or a build step, and is licensed under Apache 2.0. Its agent skills support workflows including product videos, explainers, pull-request videos, captions, motion graphics, slideshows, music-driven videos, and general compositions.
Ponytail is a rule set and plugin for AI coding agents that directs them to implement the smallest working solution. Its decision ladder skips speculative requirements, reuses code already in the repository, prefers the standard library and native platform features, avoids adding dependencies, and then chooses the shortest implementation that works. The project says that validation, error handling, security, and accessibility are not simplified away. It provides chat commands for changing enforcement intensity, reviewing the current diff for over-engineering, auditing an entire repository for bloat, recording deferred shortcuts in a debt ledger, and viewing benchmark results. It can be installed across multiple coding-agent environments, including Claude Code, Codex, Copilot CLI, Gemini CLI, and Pi; the page also lists support for other agents. The project reports lower code volume, token use, cost, and execution time in benchmarked Claude Code sessions editing a FastAPI and React repository, while noting that results vary by task and model.
DeskcommCRM is an open-source, self-hosted CRM and AI sales operating system for businesses that sell through WhatsApp and other chat channels. It combines a multi-tenant CRM and sales pipeline with inbox, Kanban pipelines, contacts, team governance, audit logs, LGPD features, webhooks, automations, scheduling, and human takeover capabilities, and supports WhatsApp through WAHA QR connections or the official Meta Cloud API. Its AI agents use tenant-specific RAG, organizational memory, executable skills, intent routing, sentiment analysis, follow-ups, and audited handoff to human attendants. Agents can operate CRM records such as leads and funnel stages, while the CRM exposes an internal MCP interface for agent operations. Tenants can receive leads through public webhook endpoints and process event-driven QUANDO/SE/ENTÃO rules through an event-log queue drained by scheduled workers. The application is built with Next.js, TypeScript, and Supabase/Postgres.
PR Lens is a pull-request visualization tool by Coldtea that generates animated architecture and data-flow diagrams and posts them inside GitHub pull requests. It maps a change's components and call paths, uses green for new elements, amber for changed elements, and red for removed elements, and provides step-by-step walkthroughs, nested diagrams, and interactive canvases with pan, zoom, and light or dark themes. It is available as a GitHub App, GitHub Action, CLI, or coding-agent skill. The CLI analyzes a diff against a merge base, produces a graph document, renders theme-paired SVGs, validates documents against a schema, and can compose the pull-request comment. The GitHub Action can use Gemini, OpenAI, or an OpenAI-compatible endpoint; the agent skill lets a coding agent generate and attach the diagrams. Repository configuration in .github/pr-lens.yml supports renames, exclusions, lane pins, and groupings. The repository documents Node 20.11+ and pnpm 10 requirements and is licensed under the MIT License.
OKF Agent Memory is a Git-native persistent memory layer for AI coding agents, developed as a pure Go CLI and library. It stores project knowledge as plain Markdown files with YAML front matter in a repository, following the vendor-neutral Open Knowledge Format (OKF) v0.2 rather than relying on an external vector database. The tooling parses and validates OKF bundles, builds an in-memory BM25 index for local lexical search, and provides progressive disclosure through hierarchical index files and link graphs. Its search-before-write workflow queries existing concepts before creating new ones; concepts can include provenance, trust tiers, lifecycle metadata, and code references for linking knowledge to source files. The standalone executable also supports creating and updating concepts, bootstrapping the memory structure into a project, and running an embedded Model Context Protocol (MCP) server over stdio for agent platforms. The repository reports sub-300-microsecond concept search and approximately 4-millisecond graph validation for its benchmark cases, with no external databases or API costs for retrieval. It is distributed under the MIT License.
Bot Crossing is a local web application that visualizes coding-agent threads as an astronaut colony. It reads session records from installed agent harnesses on the user's machine, maps repositories to persistent hex zones, and represents sessions as astronauts whose behavior reflects states such as running, waiting, errored, merged, archived, or inactive. Selecting an astronaut or repository opens its associated thread through the owning harness, while the application can also start sessions, reveal folders, mark threads viewed, and archive them. The application uses per-harness adapters to scan local session files and merge them into a common thread model. It currently documents support for Claude Code, Codex, and Cursor, with adapters for other harnesses able to be added under `server/harnesses/`. Astronaut navigation uses a rasterized navigation grid, A* routing, string-pulling, collision handling, and crowd separation. The browser renders the colony with Three.js techniques including instanced GPU-skinned astronauts, shader-based construction progress, procedural terrain and surfaces, and configurable planets, lighting, and quality settings. Bot Crossing runs locally through a Vite development server or a built Node server, binds to loopback by default, and stores the colony layout in `data/colony.json`. It does not upload session data or use an account; it reads harness records and writes only the colony state plus the documented archive field. The repository requires Node 22.13 or newer, supports macOS, Linux, and Windows, and is licensed under MIT. Bundled Kay Lousberg assets are separately covered by CC0, while bundled Material Design Icons use Apache-2.0.
anything2explainer is a Claude Code and Codex skill for producing narrated explainer videos from a topic or document. It outputs a 1280×720 H.264 MP4 in Chinese or English, with synchronized TTS voiceover, word-aligned subtitles, chapter cards, a top HUD, and a chapter progress bar. Every frame is drawn as code with Remotion, React, and TypeScript rather than generated video or stock footage. Its pipeline researches the subject and records sources, writes the narration, generates voiceover and a frame-accurate timeline, storyboards each shot, dispatches parallel agents to build Remotion components, renders the film, and runs quantitative frame metrics plus chapter-level quality-control and fix passes. The repository includes a compilable Remotion template, visual primitives, style and motion specifications, scripts for voiceover, storyboarding, rendering and QC, and a complete reference film with its research and production records. It is not a standalone CLI; it is a skill and production method intended for AI coding agents. The toolkit is licensed under PolyForm Noncommercial 1.0.0: noncommercial use is free, while commercial use requires prior authorization. The videos produced with it are described as belonging to their creators. It renders through headless Chromium on the CPU, supports landscape output rather than 9:16 video, and provides two backdrops within its specified visual style.
WeKnora is an open-source, LLM-powered knowledge platform from Tencent for turning documents into queryable knowledge bases, retrieval-augmented Q&A, autonomous reasoning workflows, and self-maintaining Markdown wikis. It supports RAG-based quick Q&A and a ReAct agent that orchestrates retrieval, MCP tools, sandbox skills, and web search for multi-step tasks. Its Wiki Mode distills source documents into structured, interlinked Markdown pages with an interactive knowledge graph, manual editing, revision history, line-level diffs, and rollback. The platform ingests formats including PDF, Word, Markdown, HTML, images, spreadsheets, presentations, and XMind, and can synchronize sources such as Feishu, GitLab, Tencent IMA, Notion, Yuque, and RSS. Its modular pipeline supports interchangeable parsers, LLMs, embedding providers, vector databases, and storage backends, with dense, sparse, hybrid, parent-child, and graph-based retrieval strategies. WeKnora provides a web interface, REST API, command-line client, MCP server, website embed widget, and integrations with messaging channels. It can be deployed locally, with Docker, or on Kubernetes, including private and offline deployments. The repository states that it is licensed under the MIT License and supports workspace RBAC, scoped API keys, audit logs, and Langfuse-based observability.
SoL-Pi is an open-source standalone extension for the Pi coding agent, developed and maintained by NVIDIA. It packages four opt-in efficiency mechanisms: Action Fusion runs a follow-up validation command in the same edit or write call; ObservationPack replaces repeated large tool results with stable handles that support exact paged recall; the Evidence-Preserving Reducer condenses diagnostic logs into receipts only when retained quotations match archived source material; and Online Context Compact marks completed plan steps for Pi's native compaction when economic and context-window checks allow it. The extension operates through Pi's public APIs without patching or vendoring Pi, leaves the mechanisms disabled by default, and preserves original observations locally. It requires Node.js 22.19 or newer and is released under the MIT License.
An open-source Codex configuration that uses Astra as the root orchestrator and independent reviewer, with Luna-based subagents for exploration, implementation, testing, and research. It provides separate Pro and Plus profiles, TOML role configurations, an Astra orchestrator skill, project-level AGENTS.md instructions, and shell and PowerShell installers that copy the selected configuration into a target repository. The profiles set model, reasoning-effort, approval, sandbox, and concurrent-thread settings, while named role files can override the defaults. The repository also includes guides for orchestration patterns, plan-specific setup, iteration, and token-usage reporting from Codex session logs. It is licensed under Apache License 2.0.
Birdview is an open-source developer tool that makes AI coding agents map a project's architecture before editing. Its Birdview workflow defines modules, responsibilities, file ownership, evidence, relationships, and layout in an architecture JSON file; an agent then declares planned task scope, targets, files, lifecycle phases, and verification records against that map in an activity JSONL stream. Validators check the schemas and cross-record rules before a renderer produces a standalone interactive HTML view with architecture, changes, comparison, module evidence, and activity-history views. Birdview supports automatic or on-demand activation through project-level agent guidance and includes Node.js scripts, JSON Schemas, example maps, and tests. The generated output has no server or network dependency, but the v0.1 implementation is file-based: activity is agent-declared rather than automatically observed, and updates require regenerating the HTML. It is MIT-licensed and is not published to npm.
Rune is a GPU-rendered, keyboard-driven IDE for power users, developed by Unstable Build. It combines code editing, terminals, command-line tools, language intelligence, debugging, file and workspace management, and AI agents in a composable multi-workspace environment with a character-grid interface. Rune Agent is provided separately as an extension in the same repository rather than being part of the core editor. Rune is an application written in Go and built with standard Go tooling; its GPU renderer uses cgo and requires platform-specific graphics and development libraries. The application is licensed under the GNU General Public License version 3 or later, while the separate rune-go-sdk extension module is licensed under Apache-2.0.
Coder is a self-hosted platform for cloud development environments and AI coding agents. It defines workspaces with Terraform and can provision them on EC2 virtual machines, Kubernetes pods, Docker containers, and other infrastructure, connecting users through a secure WireGuard tunnel and automatically shutting down idle resources. Coder Agents runs a native AI coding-agent loop in the control plane on the operator's infrastructure. It supports models from Anthropic, OpenAI, Google, Amazon Bedrock, and self-hosted providers without placing LLM credentials in workspaces, while providing centralized model governance, cost tracking, identity on actions, and audit logging. Workspaces can be accessed through existing IDEs including VS Code and JetBrains products.