Agent Skills is an open-source collection of engineering workflows for AI coding agents, developed by Addy Osmani. It packages senior-engineering practices into skills covering the development lifecycle: defining specifications, planning, incremental implementation, testing and debugging, constraints, code review, web-performance auditing, code simplification, and shipping. Slash commands such as /spec, /plan, /build, /test, /review, and /ship activate the corresponding workflows, while skills can also activate automatically based on the work being performed. The /build workflow can generate a plan and implement its tasks in an approved pass, while retaining test-driven verification, individual commits, and pauses for failures or risky steps. It can be installed with the skills CLI or integrated into supported coding agents including Claude Code, Cursor, Codex, Copilot, and Cline.
Caddy is an extensible, cross-platform web server and server platform written in Go. It supports HTTP/1.1, HTTP/2, and HTTP/3, uses TLS by default, and can automatically obtain and manage certificates through ZeroSSL and Let's Encrypt for public names or a local CA for internal names and IP addresses. Caddy uses JSON as its native configuration format and exposes a JSON API for dynamic configuration; the Caddyfile and other adapters can convert alternative formats such as YAML, TOML, and NGINX configuration into JSON. Its modular architecture treats server capabilities as Go-based apps and modules, with standard HTTP and TLS apps, plugin support, and configuration changes that can be applied without restarting the server. It can be downloaded as a standalone executable, built from source, or customized with plugins through the xcaddy builder. The project is maintained in the caddyserver/caddy repository and provides documentation and support resources through its website and community forum.
Claude Code Templates is a CLI tool and web catalog for configuring and monitoring Anthropic's Claude Code. It provides installable agents, custom slash commands, settings, hooks, skills, MCP integrations, and project templates, which can be selected interactively or installed with commands such as `npx claude-code-templates@latest --agent ...`. The CLI also includes analytics for monitoring development sessions, a conversation monitor with optional Cloudflare Tunnel access, health checks for Claude Code installations, and a plugin dashboard for viewing marketplaces, installed plugins, and permissions. The project aggregates components from Anthropic and community sources, while retaining their stated licenses and attributions; the project itself is licensed under the MIT License.
Claude-Mem is an open-source persistent-memory plugin and service for coding agents. It captures agent activity through lifecycle hooks, stores sessions, observations, and summaries in SQLite, and uses hybrid full-text and Chroma vector search to retrieve relevant context across sessions. Its MCP search workflow uses progressive disclosure: the agent first searches a compact index, then reviews a timeline, and finally fetches full observations for selected result IDs. A local worker service, managed by Bun, provides the HTTP API, search endpoints, and web viewer; the project also supports integrations with Claude Code and other listed agent environments, configurable context injection, private-content exclusion tags, and optional cloud synchronization. The repository states that it is distributed under the Apache License 2.0 and requires Node.js 20 or later, with Bun, uv, and SQLite used by the runtime. It can be installed through its npx installer or Claude Code's plugin marketplace.
Context Mode is an MCP server and plugin for AI coding agents that reduces context-window usage by routing large tool outputs through sandboxed subprocesses. Its execution tools run code in supported languages, process files, fetch and analyze URLs, and return selected stdout or search results instead of exposing raw logs, snapshots, API responses, or file contents to the conversation. The project reports up to 98% context reduction in its benchmarks. It also provides session continuity across supported agents. Hooks capture tool calls, edits, prompts, decisions, errors, and other session events in SQLite; before compaction, the system builds a prioritized resume snapshot, and after compaction or session resumption it retrieves relevant events through SQLite FTS5 search with BM25 ranking. Content indexing uses heading-aware chunking, stemming, trigram matching, reciprocal-rank fusion, proximity reranking, fuzzy correction, and smart snippets. The MCP interface includes execution, batching, indexing, search, fetching, statistics, diagnostics, upgrade, and purge tools. Context Mode supports multiple coding-agent platforms through MCP servers, native plugins, and platform-specific hooks, with automatic routing enforcement where hooks are available and instruction files for platforms without them. The repository states that processing and SQLite storage are local, with no account or telemetry requirement. It is licensed under the Elastic License 2.0, which permits use, modification, and distribution but restricts offering the software as a hosted or managed service and removing licensing notices.
DSCODE is a macOS terminal coding agent and CLI built on DeepSeek Harness. Its persistent shell retains the working directory, environment, background jobs, edits, searches, patches, and test runs, while the TUI, CLI, and scripts share one session runtime. Sessions can hand tasks to other sessions, run side questions in read-only child sessions, delegate work to child agents in isolated Git worktrees, and use an independent read-only model to review changes and approve permissions. The `dscode exec` command supports non-interactive scripts and CI, with output streamed to stdout and tool activity and session identifiers sent to stderr. It also supports MCP servers, triggers, cross-session memory, multiple model providers, and a workspace-write sandbox with human-approval and automatic-review modes. Distribution is available through npm, Homebrew, GitHub release tarballs, and source; the project requires macOS 14 or later and is licensed under MIT. It is an independent community project, not an official DeepSeek product.
Effect is a TypeScript library for building robust, type-safe applications. It provides typed errors, dependency injection, structured concurrency, scheduling, tracing, observability integrations, and unified schema validation, with platform, SQL, testing, OpenAPI, and AI provider packages in the same monorepo. It is distributed through npm, requires strict TypeScript checking, supports Node.js 18 or newer for general use, and is licensed under the MIT License.
gstack is an open-source skill and tooling framework for Claude Code, developed by Garry Tan. It turns Claude Code into a virtual engineering team through Markdown-based slash commands and specialist roles covering product review, engineering management, design review, code review, browser-based QA, security auditing, release management, documentation, and other software-delivery tasks. Its tools can drive a real browser and support unit-test, end-to-end-test, and black-box verification loops. The repository describes 23 specialist tools and eight power tools, and distributes the project under the MIT license. It requires Claude Code, Git, and Bun; Node.js is additionally required on Windows.
iCode is a lightweight, extensible, fully offline terminal development platform and toolkit for orchestrating AI agents and workflows, developed as an open-source project by openJiuwen-ai. Its terminal interface shows the active agent, model, approval mode and session; renders Markdown and diagrams; and provides workflow execution views with per-node models, tool calls and token spending. Agents are configured rather than coded and can include instructions, tools, sub-agents, skills, MCP servers and memory. Model profiles from different providers can be switched during a conversation, while file diffs, turn-based rollback, token and cache usage, model/tool timelines, and an embedded shell expose and control execution. iCode does not include models, telemetry, analytics or crash reporting; it uses only the providers, MCP servers and hooks configured by the user. The repository describes the project as a work in progress, requires uv to run, and distributes it under the Apache License 2.0 alongside separately licensed third-party components.
Marketing Skills is a collection of Markdown-based skills for AI coding agents, built by Corey Haines. It provides specialized workflows for conversion optimization, copywriting, SEO, analytics, advertising, growth engineering, sales, and related marketing tasks, and supports Claude Code, OpenAI Codex, Cursor, Windsurf, and agents implementing the Agent Skills specification. The skills share context and cross-reference one another. The product-marketing skill serves as the foundation, with other skills consulting it for product, audience, and positioning information before handling tasks such as CRO, copywriting, SEO audits, analytics setup, email flows, pricing, or sales enablement. Skills can be installed individually or as a collection through the Skills CLI, Claude Code's plugin system, a repository clone or submodule, or SkillKit. The repository is built and maintained through community contributions, with issues and pull requests welcomed. It is free and MIT-licensed.
OpenMontage is an open-source, agent-driven video production system that turns plain-language instructions into video projects through an AI coding assistant. Its workflow covers research, scripting, production planning, asset generation, editing, and final composition, with production pipelines, approval gates, cost estimates, and post-render checks. The system can produce image-based videos as well as videos assembled from real motion footage: its agent can build a corpus from free stock footage and open archives, retrieve clips, edit them into a timeline, and render the result. It runs locally with FFmpeg and Remotion and includes 12 production pipelines, more than 100 tools, and hundreds of agent-skill and production-knowledge files.
Pi is an open-source AI agent harness and toolkit from earendil-works for building and running coding agents. Its packages provide a unified multi-provider LLM API, an agent runtime with tool calling and state management, an interactive coding-agent CLI, a terminal UI library with differential rendering, and vendor-neutral telemetry contracts and adapters. Slack and chat automation are provided through a separate package project. Pi does not provide built-in restrictions for filesystem, process, network, or credential access; it runs with the permissions of its launching user and process. The project documents containerization and sandboxing approaches, including a local Linux micro-VM, Docker, and policy-controlled sandboxing, for stronger isolation.
Ponytail is an MIT-licensed agent skill and ruleset from Dietrich Gebert that instructs AI coding agents to implement only what a task needs while retaining validation, error handling, security, accessibility, and tests for risky logic. Its decision ladder has the agent read the affected code and trace the real data flow, then reuse existing code, prefer the standard library and native platform features, avoid dependencies, and choose the shortest working implementation. It records deliberate shortcuts and reports skipped checks and risks rather than simplifying away important safeguards. It provides intensity modes through `/ponytail [lite | full | ultra | off]`, plus `/ponytail-review` for reviewing changed code, `/ponytail-audit` for checking an entire repository, and `/ponytail-debt` for collecting shortcut comments into a debt ledger. Review and audit examine bugs, security, real-load behavior, missing tests, performance, and unnecessary code. Ponytail is distributed for multiple coding-agent environments, including Claude Code, Codex, Copilot CLI, Gemini CLI, Pi, and Devin CLI, and the project lists benchmark results.
pstack-claude is an MIT-licensed port of Lauren Tan’s pstack agent workflow stack for Claude Code, Codex, Pi, OpenCode, Gemini CLI, and other agent harnesses. Its skills and routing instructions select workflows for tasks such as bug fixing, planning, feature work, refactoring, performance investigations, PR maintenance, and shipping; the bug-fixing workflow reproduces the failure, investigates how and why it occurs, delegates the fix, and reruns the failing case, involving an architect when the change crosses a function boundary. The package provides Claude Code and Codex plugins, a Pi extension, slash commands, subagent and question tools, wake-up tools, and setup for model defaults and reasoning effort. It runs scripts locally and has no server or telemetry; content read by its skills is sent to the configured model provider, and pull-request tools use the local GitHub CLI login.
Sentry is a developer-focused application performance monitoring, error-tracking, and debugging platform. It helps identify and trace issues in real-world applications through official SDKs for JavaScript, Python, Ruby, PHP, Go, Rust, Java/Kotlin, C#/F#, C/C++, Dart/Flutter, and other platforms and languages. Its agent-focused features and Sentry MCP expose traces, request and token costs, timelines, chat transcripts, errors, and the full request pipeline so large-language-model agents can investigate and debug production issues. The project is developed by Sentry and is available as an open-source repository with a fair-source topic designation.
T3 Code is an open-source agent-harness control surface for coding agents running on a developer's computer. It provides iOS and Android mobile apps, a web app, and an Electron-based desktop app for controlling locally configured Codex, Claude Code, Cursor, Grok Build, and OpenCode agents while using the user's existing provider subscriptions. The project can be launched with `npx t3@latest`, which starts its backend and local web interface on the machine. It also distributes desktop builds for Windows, macOS, and Arch Linux, and supports remote access from a phone or another machine. The repository describes the project as early-stage and warns that bugs are expected.
text-to-cad is a library of agent skills for creating, inspecting, sourcing, slicing, previewing, and handing off CAD, CAE, CAM, robotics, and hardware-design artifacts from local project files. Its skills generate and edit CAD models from plain-language or image requests, primarily producing STEP files with optional STL, 3MF, and GLB exports; preview local CAD and robot files in a browser; source off-the-shelf STEP parts; and create DXF drawings from Python sources or CAD geometry. Additional skills write URDF robot structures, SRDF/MoveIt2 planning and collision data, and SDF simulator models and worlds; check DXF and STEP files for SendCutSend; measure mesh printability by wall thickness, overhangs, support volume, and build orientation; slice supported meshes into validated, printer-profiled FDM G-code using slicer command-line tools; and dry-run, upload, and cautiously start local Bambu Lab print jobs. An experimental implicit-CAD skill creates browser-native models with GLSL signed-distance fields and CAD Viewer raymarch rendering. The library is installed with the Skills CLI, including `npx skills add earthtojake/text-to-cad`, or through provider-native plugins for Codex, Claude Code, and Grok Build.
Three.js Particle Fluids is a GPU particle-physics library for Three.js that simulates and renders liquids, soft bodies, cloth, and smoke with WebGPU. Its API provides a Simulation class, particle rendering, fluid and liquid-surface solvers, soft-body shape matching and mesh skinning, cloth constraints with wind, smoke tracers, colliders, neighbor grids, and mesh-to-distance-field baking. It includes demo presets such as water drops, dam breaks, viscous honey, jelly bodies, cloth, and buoyant smoke, along with documentation and an API reference. It is distributed as the npm package threejs-particle-fluids under the MIT license and requires Three.js r184 and a browser with WebGPU; there is no WebGL fallback.
Searchable transcript of Top Dev Tool Projects : Sentry, Caddy, Effect, claude-mem, OpenMontage & iCode — ManuAGI - AutoGPT Tutorials (14:50). Search for a phrase, then click its timestamp to jump straight to that moment in the video.
Captions sourced from the original video on YouTube, published by ManuAGI - AutoGPT Tutorials. The video, its captions and all related intellectual property remain the property of their respective owners; AINotes claims no ownership. Provided for research, accessibility and search — see the Transcript Notice and Copyright Policy.
00:00 Every week, developers release powerful open-source dev tools, and this weekly DevTool project update video brings them together in one place. Top trending open-source and best DevTool projects this week. Cover coding agents, agent memory, web servers, error tracking, and natural language testing. You'll discover useful and trending developer tools you can start using right away.
00:22 Explained fast and to the point. Without wasting time, let's get started. Before we jump into today's project updates, here's a quick announcement for everyone, we've launched a brand new YouTube channel called AI Agent Studio, dedicated entirely to AI Agent projects, tutorials, and tools. So, if you're interested in staying up to date with the latest AI Agent open source projects, learning how to build your own agents, or exploring cuttingedge agent frameworks, make sure to check it out.
00:52 Subscribe now to get weekly videos, in-depth guides, and real-time project breakdowns. The link is right there in the description. Don't miss it. All right, let's get into today's video. >> Project number one, context mode. Stop AI coding agents from wasting context. Context mode is an MCP server that protects an AI coding agents context window, which fills fast as tool calls dump raw data into it.
01:16 It sandboxes tool output so results are computed in an isolated subprocess and only the answer enters context reporting large reductions and it indexes every file edit task and decision into a searchable store so work survives compaction. It pushes agents to write scripts rather than read many files and works across 17 platforms via MCP and hooks. Node source available.
01:41 Install it and reclaim your context. Project number two, FI agent harness. A minimal coding agent you make your own. FI is a minimal extensible AI agent harness and toolkit built to be adapted to your workflows rather than the reverse. It ships a terminal coding agent with strong defaults, but deliberately leaves out features like sub aents and plan mode, so you add what you need through extensions, skills, prompt templates, and themes.
02:06 Bundled as sharable packages, it includes a unified multi-provider LLM API, an agent runtime, and a terminal UI library, and runs interactively in print or JSON mode over RPC or via a TypeScript SDK, MIT licensed. Install it and give it a task. Project number three, BE Claude Code Templates. Readymade components and monitoring for Cloud Code. Claude Code templates is a command line tool and browsable catalog that drops readytouse configurations into claude code without manual setup from a web interface or the terminal
02:42 you install components AI agents for specific domains custom/comands MCP integrations for services like GitHub and Postgresque settings automation hooks and skills it also bundles realtime session analytics a mobile friendly conversation monitor a health check and a plug-in dashboard MIT licensed Next, install a component and set up your agent. Project number four, me and DScode, DeepSseek terminal coding agent whose sessions talk.
03:08 DSCode is a terminal coding agent for Mac OS built on the DeepSseek harness designed so sessions on your machine can see and hand work to each other. A persistent shell reads and writes code and runs tests shared by the terminal UI CLI and your scripts and one session can pass a task to another which reads the transcript replies and hands the result back.
03:32 It adds an independent review model for approvals and diffs cross- session memory sub aents and non-interactive runs node MIT licensed independent install it and give it a task. Project number five, any PS5 port PS5 executables to run on PC. Any PS5 is a tool that automatically ports PS5 executables to run natively on Linux and Windows without emulation or a separate runtime process.
03:58 A relinker converts the executable into the target systems native format and the project supplies its own implementations of the PS5 system PRX libraries for dynamic linking. So, the game runs directly on your machine. It includes a shader recompiler that turns PS5 shaders into Vulcan Spear V. Supports SDL game controllers and configurable keyboard and mouse input and runs some tested games smoothly.
04:22 It ships no copyrighted firmware or keys. C++ GPL 2.0 build it and port an executable. Project number six, MANS E2E natural language endtoend testing for web and mobile. E2E is an endto-end testing framework for web and mobile apps where you describe a goal in plain language and an agent interacts with the app to complete it. While locators and assertions in the same test check exact results, an agent step that a later assertion verifies records its actions.
04:52 So the next run replays them with no model calls until the app changes. And tests without agent steps need no model at all. You bring your own model and it has browser and mobile engines. Apache 2.0. Initialize it and write your first test. Project number seven, marketing skills for AI agents. Give your agent marketing expertise. Marketing skills for AI agents is a collection of agent skills that give coding agents specialized marketing knowledge and workflows built for technical marketers and founders.
05:20 It spans about 50 skills across conversion optimization, copywriting, SEO and AI search, analytics, AB testing, paid ads, retention, growth, and sales. Each a markdown file the agent applies when it recognizes the task. A foundational product marketing skill is read first so every other skill knows your product and audience. Works with claude code, codeex, and cursor.
05:44 MIT licensed. Install the skills and optimize your page. Project number eight. Ponytail. Make your coding agent write less code. Ponytail is an agent skill that makes an AI coding agent build only what a task needs. modeled on the senior dev who replaces 50 lines with one. Before writing code, it climbs a ladder. Skip what isn't needed. Reuse what's in the codebase.
06:07 Use the standard library or a native platform feature. Then a single line. Then the minimum that works while never cutting validation, error handling, security or accessibility. It has intensity levels and review commands and works across about 20 agents. MIT licensed. Install it and write less code. Project number nine, text to CAD design hardware through your coding agent.
06:31 Text to CAD, published as CAD skills, is a library of agent skills that lets an AI coding agent generate, inspect, source, slice, and handoff CAD, robotics, and hardware files from your project. You describe a part in plain language or from an image, and it produces real geometry with step as the main output plus STL, 3MF, and GB and a local viewer.
06:54 It also finds off-the-shelf parts, draws 2D DXFs, writes robot description files, validates for laser cutting and slices to G-code, works with codecs and claw code, Python, MIT licensed. Install it and describe your first part. Project number 10, Vanshei Sentry developer first error tracking and performance monitoring. Sentry is a debugging platform that helps developers detect, trace, and fix issues in their applications.
07:21 You add one of its SDKs available for languages and frameworks including JavaScript, Python, Ruby, PHP, Go, Rust, Java, and more. And it captures errors and crashes with the context to diagnose them alongside performance monitoring, tracing, session replay, logs, and uptime checks. This repository is the main platform built with Python and TypeScript, usable as a hosted service or self-hosted.
07:46 Add an SDK and start catching errors. Project number 11, Vanit Gojand. Open Montage. Turn your coding agent into a video studio. Open Montage is an open-source agentic video production system that turns an AI coding assistant into a full production studio. You describe a video in plain language or paste a reference clip and the agent runs the whole pipeline, research, scripting, asset generation, editing, and final composition.
08:15 It ships many production pipelines, over a 100 tools, and many provider integrations spanning image, video, voice, and music with free local options, and makes real footage videos from open archives, not just animated stills. Human approval gates, scored provider selection, and budget caps throughout. Python AGPL-3.0. Install it and describe your video.
08:39 Project number 12, DANC3 code. Drive your coding agents from your phone. T3 code is a control surface for the AI coding agents already installed on your machine, letting you start, monitor, and steer them from a polished set of apps instead of replacing them. It detects and takes control of your authenticated setups for cloud code, codecs, cursor, grock, build, and open code.
09:00 And you reach it through mobile, web, and desktop apps. Because the backend runs on your own machine, it's remote ready. So you can direct a longunning agent from your phone. MIT licensed early run it and connect your agents. Project number 13. Caddy web server with automatic HTTPS. Caddy is a fast extensible web server that serves every site over HTTPS by default, obtaining and renewing TLS certificates automatically from Let's Encrypt and zero SSL for public names and a managed local CA for internal ones.
09:33 You configure it with the simple caddy file or its native JSON API. And it supports HTTP 1.1, HTTP2, and HTTP/3. Reverse proxying and static file serving. Built-in Go with no external dependencies and a modular plug-in system. It runs anywhere. Apache 2.0. Download it and serve your site over HTTPS. Project number 14. Bstack. Rigorous agent workflows ported beyond cursor.
10:03 Bstack is a port of Laurent Tan's opinionated cursor skill stack to Claude Code, Codeex, PI, Open Code, Gemini, and Prime Agent, translating its cursor primitives for other harnesses to improve agent outcomes. You tell its potato mode your goal, and it picks the right workflow. For a bug, it reproduces the failure, investigates, delegates the fix, and reruns the failing case, bringing in an architect skill when a change crosses boundaries, and hands back passing and failing evidence.
10:32 Playbooks cover planning, features, refactoring and shipping. No telemetry, MIT licensed. Install it and give potato mode a goal. Project number 15, agent skills. 25 engineering workflows for coding agents. Agent skills is a pack of 25 reusable skills that make an AI coding agent follow a disciplined engineering process across the whole life cycle from defining and planning through building, testing, reviewing, and shipping.
10:58 You reach them through nine slashcomands or let the agent trigger them by task. Each skill is a step-by-step workflow with verification gates and rebuttals to the shortcuts agents take and it installs into clawed code, cursor, codecs, and many more. MIT licensed install it and start from a spec project number 16. Glaude mem persistent memory across your agents sessions.
11:22 Glaude mem is a memory system that gives AI coding agents persistent context across sessions. So, an agent remembers your project after a session ends or restarts instead of making you reexplain it. It captures the agents tool use as it works, compresses those observations into semantic summaries with AI and injects the relevant ones back into future sessions automatically.
11:44 It runs a local worker with a SQLite and vector store exposes token efficient search over MCP and works with clawed code, codecs, Gemini, and more. Typescript Apache 2.0. Install it and let your agent remember. Project number 17, GStack. Gary Tan's Claude Code setup as a virtual team. GStack is Gary Tan's own Claude Code setup, a collection of 23 opinionated tools and eight power tools that turn a coding agent into a virtual engineering team.
12:13 Slash Command Skills play distinct roles. A CEO who rethinks the product. An ing manager who locks architecture. A designer who catches AI slop. A reviewer. a QA lead who drives a real browser, a security officer, and a release engineer run in sprint order from think to ship. It works across 10 agents MIT licensed install it and run office hours. Project number 18, Effect, a TypeScript library for production grade apps.
12:41 Effect is a TypeScript library for building robust, maintainable, typesafe applications, giving you one toolkit for the hard problems that show up at scale. It brings typed errors, dependency injection, structured concurrency, scheduling, retries, tracing, and unified schema validation into a composable core. So failures and dependencies are tracked in the type system rather than left to convention.
13:04 A large monorrepo adds platform, ACQL, AI provider, and open telemetry integrations. Requires strict TypeScript, MIT licensed. Install it and build your first effect. Project number 19, 3JS Particle Fluids, GPU particle fluid simulations in the browser. 3JS Particle Fluids is a position-based fluids library and interactive demo for 3JS running realtime particle simulations on the GPU with web GPU and the 3JS shading language.
13:34 It ships curated presets exploring fluid, soft body, cloth, and smoke behavior with live parameter controls, adjustable particle budgets, and performance diagnostics. The engine is source you import covering particle buffers, solvers, collisions and screen space surface rendering. Requires a web GPU desktop browser and GPU with no webgl. MIT licensed run it and explore a simulation.
14:00 Project number 20. Banshee offline terminal platform for agents and workflows. ICode is a lightweight, extensible, fully offline development platform and agent toolkit with a terminal interface giving you complete control over your data, agents, and workflows. From the TUI, you chat with an agent, switch models, and approval modes mid-con conversation, and run multi-step workflows node by node.
14:23 An agent is configuration, not code, instructions, tools, sub aents, skills, MCP servers, and memory edited in one place. It adds diffs, roll back, trajectory breakdowns, and a real embedded shell. It ships no models. You connect your own, and it has no telemetry. Python Apache 2.0. Run it and add a model. Thanks for watching. See you in the next update.