2,435 tools and products — trending open source, and what gets used in AI and other work.
LiveKit Agents is an open-source Python framework for building programmable, real-time multimodal voice agents that run as server-side participants. It combines speech-to-text, large language models, text-to-speech, and realtime APIs with LiveKit's WebRTC clients, telephony stack, RPC and data APIs, and MCP tool support. The framework provides agent sessions, server-side job scheduling and dispatch, semantic turn detection based on a transformer model, multi-agent handoffs, and tools for voice, text-only, transcription, vision, and video-avatar applications. Agents can be tested with native test integration, including event assertions and LLM-based judges, and can run locally in console mode, in development mode with hot reloading, or in production mode. The framework can use LiveKit Cloud or a self-hosted LiveKit server and is distributed as a Python package with plugins for model providers. The Agents framework is licensed under Apache-2.0; LiveKit turn-detection models use the LiveKit Model License.
Paseo is a self-hosted platform for orchestrating multiple coding agents, including Claude Code, Codex, GitHub Copilot, OpenCode, and Pi. Its local daemon manages agent processes on users' machines, while desktop, mobile, web, and CLI clients connect to run agents in parallel, stream output, send follow-up tasks, and work in specified directories or worktrees. Paseo supports voice task dictation and control, agent handoffs, advisor and committee workflows, local tools, configurations, skills, and development environments, and provides an MCP server, WebSocket API, and TypeScript SDK for integrations, dashboards, and orchestration services. Remote connections use an end-to-end encrypted relay, TCP, Tailscale, or another VPN. Paseo can run as an installed application, headlessly, or as a Dockerized daemon with a self-hosted web UI. The repository is licensed under Apache-2.0 and states that Paseo has no telemetry, tracking, or forced log-ins.
A personal AI learning system distributed as a pi configuration. It encodes a teaching philosophy and learning process in skills, including teaching and diagram-based visualization, and adds extensions for question popups, graded quizzes, Markdown session logs, and visualization tools. The configuration delegates research and visual creation to researcher, SVG-maker, and Mermaid-maker subagents; it can also run without subagents, with those delegation-based capabilities omitted. It is intended for one learner and is shared as-is under the repository's stated installation instructions.
Google 3D Tiles is a Google service that provides 3D geographic tiles covering the globe.
An OpenAI API for real-time interactions, powering voice interaction with a 3D globe.
LiveKit is an open-source framework and developer platform for building, testing, deploying, scaling, and observing real-time voice, video, and physical AI agents. Its agent pipeline streams user speech from an app, browser, or phone call to an agent, which applies custom business logic and returns a response; the platform supports automatic turn detection and interruption handling, speech-to-text, language-model, and text-to-speech providers, web and mobile applications, and telephony through phone numbers and SIP integrations. LiveKit Cloud provides deployment and scaling on LiveKit's real-time infrastructure, alongside an inference gateway and full-stack observability for agent sessions.
An AI framework for defining model architectures and supporting model serving.
Tesseract OCR is software for performing offline optical character recognition on captured document regions. Its official documentation includes a user manual and source-code documentation.
Tailcat is an open-source Go CLI and library from Tailscale that provides Netcat-style connections over Tailscale's encrypted data plane without using the Tailscale control plane or requiring a Tailscale account. A server generates a connection token containing WireGuard and relay metadata; a client uses the token to establish an end-to-end encrypted tunnel, initially through a DERP relay, while magicsock attempts NAT traversal and a direct peer-to-peer UDP connection. Its userspace networking stack does not alter routing tables, DNS settings, or require root access. The CLI can pipe standard input and output between machines, forward local TCP ports, provide an unauthenticated SSH server, run commands through a SOCKS5 proxy, act as an exit node, ping and inspect connection tokens, and publish tokens through DNS TXT records. The Go library exposes server and client APIs for dialing TCP ports through the tunnel. Tailcat can use Tailscale's public rate-limited DERP relays or user-operated relays, and an experimental WebAssembly browser demo supports file and text transfer. The project states that its CLI, API, and wire format have no stability guarantees.
Screendrop is a native, local-first macOS menu bar application for capturing, annotating, recording, editing, compressing, and sharing screenshots and screen recordings. It captures displays, windows, or selected areas; provides configurable hotkeys and after-capture workflows; and includes a floating preview stack, offline History, non-destructive annotation with redaction and mockup effects, and a recording studio with separate screen, camera, microphone, system-audio, cursor, click, and keystroke data. Recording projects support timeline edits, zooms, reconstructed cursors, camera layouts, captions, transcript-based cuts, social aspect ratios, and reusable style presets. On-device speech recognition provides transcription, karaoke-style captions, filler-word removal, silence trimming, and a notch-native teleprompter. Optional self-hosted sharing uses a Cloudflare Worker, R2, and D1: the Worker authenticates uploads, stores media and metadata, and returns share links for screenshot viewers or Loom-style recording pages with transcripts, comments, playback controls, and scrub previews. Captures, projects, annotation sidecars, and transcripts remain on the Mac unless explicitly uploaded, and there is no central Screendrop account or hosted server. The project is distributed as a DMG or Homebrew cask, requires macOS 26.4 or newer according to its README, and is dedicated to the public domain under CC0 1.0 Universal.
React Router DOM provides browser routing for React applications, including routes, links, outlets, and navigation used to connect authentication and application pages. It is part of React Router, which the official documentation describes as a standards-focused, multi-strategy router that can be deployed in different environments.
Axios is a promise-based HTTP client for browser and Node.js applications. It creates API instances and sends requests such as authentication, session, logout, and project operations. Its interceptor system can modify request, response, and error handling, and it includes TypeScript support.
Lucide React is a React icon package from the Lucide community icon toolkit. It provides customizable, lightweight SVG interface icons for controls such as loading, authentication, prompts, uploads, microphones, passwords, and form submission. Icons can be adjusted by color, size, and stroke width, and the package supports tree shaking so applications can import only the icons they use.
React Hot Toast is a lightweight React notification library for adding toast messages to React applications. In the cited project, it displays success and error notifications for authentication and project actions.
Moment.js is a JavaScript date and time formatting library. It formats dates and times, calculates relative and calendar time, supports date arithmetic, and provides localized formatting across many languages. The project is in maintenance mode and can be installed through package managers including npm, Yarn, spm, and Meteor.
Sandpack is a React component toolkit for embedding live-running code editing and preview experiences in web applications. It runs JavaScript and Node.js applications in the browser and can be used for interactive documentation, coding environments, low-code tools, generated website previews, and embedded playgrounds. Sandpack provides templates containing project files and dependencies, custom setup for changing dependencies or file structure, configurable built-in components, themes, and composable components and hooks for custom interfaces. It uses CodeMirror for editor functionality, supports npm dependencies, hot module reloading, error overlays, caching, and Node.js execution through its Nodebox runtime. The Sandpack Client package exposes a framework-agnostic bundler protocol, and the toolkit supports frameworks including Next.js, Remix, Vite, and Astro.
JSZip is a JavaScript library for creating, reading, and editing ZIP archives. Its API lets applications add generated content to files and folders, then asynchronously produce a ZIP archive such as a browser Blob for download; it also supports base64 data and use with Node.js. The project is distributed under either the MIT or GPLv3 license and can be installed with npm or included manually in a web page.
A downloadable web-development project bundle from GreatStack for building a full-stack AI website builder with MongoDB, Express.js, React.js, and Node.js. The project includes starter assets and source files for a React website generator that accepts text prompts, builds websites step by step, exposes generation progress, and supports manual source-code editing, follow-up AI prompts, exporting, and publishing. Its tutorial project includes user authentication, REST APIs, an OpenRouter model integration, and an agent chat API for updating generated projects.
Heretic is a command-line tool for automatically removing safety alignment from transformer-based language models without post-training. It combines directional ablation (abliteration) with a Tree-structured Parzen Estimator optimizer powered by Optuna, co-minimizing refusal counts and KL divergence from the original model to select ablation parameters automatically. For supported transformer components, currently attention output projections and MLP down-projections, Heretic computes per-layer residual directions from the difference between first-token hidden states for harmful and harmless prompts, then orthogonalizes the associated weight matrices against those directions. Its optimizer can interpolate between residual directions and select separate, flexible layer-weight kernels for different components. The tool supports most dense models, many multimodal models, several mixture-of-experts architectures, and some hybrid architectures; pure state-space models and certain research architectures are not supported out of the box. Heretic runs in a Python 3.10+ environment with PyTorch 2.2 or later, supports optional bitsandbytes 4-bit quantization, and can save or upload generated models, launch a chat evaluation, and run standard benchmarks. An optional research installation provides residual-vector visualization using PaCMAP and residual-geometry analysis. The project is distributed under the GNU Affero General Public License version 3 or later.
Crawl4AI is an open-source Python web crawler and scraper that converts web pages into structured, LLM-ready Markdown for retrieval-augmented generation, agents, and data pipelines. Its asynchronous Playwright-based crawler supports Chromium, Firefox, and WebKit, dynamic JavaScript pages, sessions, persistent browser profiles, cookies, headers, proxies, screenshots, media, iframes, lazy loading, full-page scanning, caching, and deep crawling with BFS, DFS, and best-first strategies. For extraction, it provides heuristic Markdown filtering including BM25-based relevance filtering, CSS- and XPath-based schema extraction, chunking and cosine-similarity strategies, and optional LLM-driven structured JSON extraction. It also includes adaptive crawling, link analysis, URL seeding, virtual-scroll handling, anti-bot and proxy escalation features, and customizable hooks. Crawl4AI can be installed with pip and used through Python or its command-line interface. It is also distributed as a Dockerized FastAPI server with JWT authentication, browser pooling, monitoring dashboards, a playground, and endpoints for crawling, HTML extraction, screenshots, PDF generation, and JavaScript execution. The repository states that it is licensed under Apache License 2.0.
IPATool is a command-line tool for searching the App Store and downloading app packages (IPA files) for iOS, iPadOS, tvOS, and visionOS. It authenticates with an Apple ID, searches by term and platform, obtains app licenses, lists owned apps and available versions, retrieves version metadata, and downloads encrypted packages to a specified path. It runs on Windows, Linux, and macOS, can be installed manually from GitHub releases or with Homebrew on macOS, and is compiled with the Go toolchain. IPATool is released under the MIT license.
Zod is a TypeScript-first schema validation library by Colin McDonnell (@colinhacks) for defining schemas, parsing untrusted data, and producing strongly typed validated results. It works with TypeScript and plain JavaScript in Node.js and modern browsers, has no external dependencies, and supports immutable schema composition, JSON Schema conversion, synchronous and asynchronous parsing, detailed ZodError issues, and safeParse result objects. Zod infers static types from schemas through utilities such as z.infer, z.input, and z.output; transforms can produce different input and output types. For hot validation paths, z.compile(schema) can create an ahead-of-time compiled fast path using new Function, while unsupported or asynchronous schemas use the regular parser.
GitNexus, developed by Akon Labs, is a code-intelligence engine that indexes repositories into a knowledge graph for code exploration and AI-agent context. Its indexing pipeline walks the file tree, parses source with Tree-sitter, resolves imports, calls, inheritance, constructor-inferred receiver types and other relationships, groups symbols into functional communities, traces execution processes, and builds BM25-plus-semantic hybrid search indexes backed by LadybugDB. The resulting graph supports MCP tools and CLI commands for process-grouped search, symbol context, call-path tracing, blast-radius and Git-diff impact analysis, structural checks, coordinated renaming, API and route mapping, taint and dependence queries, and Cypher access; repository groups can link contracts and impacts across multiple repositories. The CLI runs locally and can connect editors such as Claude Code, Cursor, Codex and others through MCP, skills and selected hooks. GitNexus also provides a browser-based WebAssembly UI with an interactive graph explorer and AI chat, plus a local HTTP server and Docker deployment mode for accessing indexed repositories through a backend. The web-only mode keeps repository processing in the browser and is constrained by browser memory, while the native CLI stores indexes locally in each repository's .gitnexus directory and uses a global registry for multi-repository access.
ClickHouse is an open-source, column-oriented database management system for online analytical processing (OLAP) and real-time analytics. It stores data by column and uses optimized compression and vectorized query execution to process analytical SQL queries while using available CPU resources efficiently. It supports large-scale workloads such as real-time analytics, observability, data warehousing, and machine learning or generative AI applications, including vector search and aggregations. ClickHouse can be self-managed on a machine or deployed through ClickHouse Cloud on AWS, Google Cloud, and Microsoft Azure. ClickHouse Local runs queries on local files such as CSV, TSV, and Parquet without a server. The project provides installers for macOS, Linux, FreeBSD, and Windows, as well as Docker and playground options; the self-managed database is open source.
Co-Invest is Liquid's AI trading assistant for researching markets, sizing positions, and placing trades through Claude, ChatGPT, iMessage, a Chrome extension, or the Liquid account. It uses market data such as positioning, funding, liquidation maps, on-chain flows, news, macroeconomic events, earnings, and ETF flows to produce trade ideas with a direction, size, named catalyst, and source links. It supports crypto, stocks, ETFs, indices, commodities, foreign exchange, and other markets available through Liquid, including long and short positions, multipliers, stop-losses, and take-profits. Trades require explicit user confirmation: the assistant returns a confirmation card containing the symbol, direction, size, multiplier, and rationale, and the user taps to confirm or cancel. According to the product page, Co-Invest can read a portfolio and propose trades but cannot transfer funds, change account settings, or submit an order without confirmation. The page distinguishes this assistant from Co-Invest Computer, which it identifies as Liquid's option for scheduled, hands-off execution; the video description presents scheduled trading routines as part of Liquid's AI trading offering. Co-Invest itself is described as free, with Liquid's standard trading fees and applicable perpetual-futures funding rates. It also provides a paper-trading mode using simulated balances against live market data.
Lighter is a perpetual-exchange venue integrated into Liquid's aggregation approach. It is used in the context of trading on-chain derivatives and perpetual contracts through Liquid.
Polymarket is a prediction-market platform for trading on the outcomes of future events across topics such as politics, sports, crypto, and technology. Its markets present outcome shares with prices that represent the platform’s displayed probabilities, and users can use the market data to assess events such as the passage of legislation.
Tangem is a self-custody cryptocurrency hardware wallet and companion app for storing and managing digital assets. Its NFC-powered cards contain a secure element that generates and keeps private keys inside the chip; multiple cards can be linked as backups, while a recovery seed phrase remains optional. Users tap a card to a phone to sign transactions and manage assets in the app, including sending, buying, selling, swapping, staking, and spending crypto across supported blockchain networks. The wallet requires no battery, cable, or charging, and the site describes an EAL6+-certified chip, audited firmware, and a 25-year limited hardware warranty.
Suno is an AI music generator and web-based generative audio workstation from Suno, Inc. It creates complete songs from text prompts or user-supplied melodies, lyrics, audio, moods, genres, and themes, including vocals, instrumentation, and production. The service can be used to create customized local-business jingles, including versions tailored to particular markets. Users can regenerate, extend, edit, remix, reorder sections, rewrite lyrics, record or upload audio, and control aspects such as voices, vocal gender, exclusions, weirdness, and style. Suno Studio combines AI music creation with digital-audio-workstation functionality; generated tracks can be separated into time-aligned WAV stems for use in other audio software. Free accounts provide limited daily song creation, while paid plans add commercial rights and expanded creation, editing, stem-separation, and Studio features.
Local Internet Presence is a local marketing service operated by Mike Stewart for small businesses. It offers help with websites, content, Google Business Profiles, advertising, and local customer acquisition; its website invites business owners to book a Zoom session to discuss being found online and converting visitors into customers. The accompanying description says Stewart also creates jingles for businesses such as plumbers, roofers, and pest-control companies and runs them as skippable YouTube ads targeted to their ZIP codes, using tools including Suno and ElevenLabs.
The Koerner Office is a business and entrepreneurship podcast hosted by Chris Koerner. Its episodes examine small-business ideas, growth strategies, side hustles, real-estate investing, and stories from people who have made money in surprising ways. Episodes alternate between broad discussions of several ideas with a guest or business partner and detailed examinations of a single business or topic; the podcast is released three times a week and is available through major podcast platforms.
FuXi is a self-contained terminal AI coding agent developed by FUXI. Built in Go and distributed as a static binary, it uses a Think → Act → Verify loop to read and edit code, run shell commands, drive tools, connect to MCP servers, and route requests across multiple LLM providers with automatic failover and cost-aware settings. Its built-in capabilities include file operations, shell execution, code search, web fetching, LSP diagnostics, Jupyter, browser use, background tasks, and parallel sub-agents. Shell commands pass an AST-based safety classifier, while permissions and audit logs govern autonomous actions. Sessions persist to disk, with checkpoints for resuming, rolling back, or forking; the TUI also supports memory consolidation and automatic context compaction. FuXi supports provider API keys or FuXi OAuth, configurable OpenAI-compatible endpoints, MCP clients, hooks, skills, plugins, and slash commands. The repository contains documentation, installers, release information, and issue-tracking materials; it states that the product source is proprietary and not published.
MyContext is a local-first desktop app from openTrinity that builds a private personal work-context layer from sources such as instant-messaging conversations, documents, and meeting records. It stores local copies, indexes, source references, and derived context in an on-disk SQLite vault, then organizes them into a personal context graph linking people, projects, topics, events, conversations, and supporting facts. Its search and answer workflow combines local full-text search, semantic retrieval, and graph queries, with agents assembling answers from traceable source material and falling back to ranked local results when the agent runtime is unavailable. A digital-self workflow recalls relationship-specific context and communication history to draft replies, while sending, deletion, and other consequential actions require explicit user confirmation. The repository describes an Electron and React desktop architecture with source connectors, incremental ingestion, context processing, retrieval, knowledge-graph, persona, and isolated agent-runtime layers. MyContext is in developer preview and under active development; its README warns of compatibility-breaking changes and migrations that may require recollection. It is licensed under the Elastic License 2.0, which permits use, modification, and self-hosting but restricts offering it as a hosted or managed service to third parties.
JoyAI-Video-Edit is an open-source, instruction-guided system for editing live camera streams or uploaded videos as frames arrive. It processes frames causally without waiting for the complete video, requiring a predefined sequence length, or revisiting future frames, and supports subject and local edits, background replacement, style and motion changes, and reference-guided editing. Its autoregressive diffusion architecture combines an MLLM-based condition encoder, a causal video VAE, and a 16-billion-parameter multimodal diffusion transformer. The deployment uses aligned autoregressive distribution-matching distillation, long-horizon optimization, bounded KV-state inference, and deployment-oriented scheduling to reduce train–inference mismatch and temporal drift during streaming generation. The repository reports 30 FPS end-to-end throughput at 720 × 1248 in its deployment benchmark. The repository provides deployment code and model checkpoints, a local server with a browser interface, and instructions for CUDA-based inference. It is licensed under Apache 2.0.
TurkishExporter.net is a business directory used to find suppliers and manufacturers in Turkey.
IndiaMART is an online business-to-business marketplace for browsing and contacting suppliers in India.
Thomasnet is an industrial directory for finding manufacturers and suppliers in the United States, Canada, and Mexico, including sources for private-label and custom products.
screenpipe is a local AI-agent memory layer that continuously captures computer history on macOS, Windows, and Linux. It records screen frames with OCR, accessibility data, microphone and system audio with transcripts, and application activity, storing the underlying history locally. Agents can search the history through a local REST API, database, or MCP server and use it for tasks such as meeting summaries, follow-ups, and workflow automations. Users can exclude apps, windows, URLs, or time periods and redact sensitive fields on the device; the page describes the project as source-available.
Gemini Deep Think is an AI reasoning system used in mathematical research. In the cited example, a mathematician used it to prove lemmas after human experimentation led to a stronger problem statement.
OpenClaude is an open-source, terminal-first coding-agent CLI for cloud and local model providers. It connects to OpenAI-compatible APIs, Gemini, GitHub Models, Codex, Ollama, Atomic Chat, and other supported backends, providing prompts, streaming output, Bash and file tools, grep, glob, agents, tasks, MCP, slash commands, web search and fetch, and image inputs for compatible providers. The CLI supports guided provider setup with saved profiles, conversation continuation and forking, detached local background sessions, model-specific agent routing, repository maps based on PageRank-ranked code structure, and a headless bidirectional-streaming gRPC server for integrations, CI/CD pipelines, and custom interfaces. A bundled VS Code extension provides launch integration, in-editor chat, provider-aware controls, and theme support. OpenClaude runs on Node.js 22 or newer, is distributed through npm and an Arch Linux AUR package, and can use local inference or remote APIs. Its repository is licensed MIT for the project's modifications and states that it is an independent community project derived from and substantially modified from the Claude Code codebase, without Anthropic affiliation.
A workstation for cybersecurity investigation and forensics that hosts autonomous incident-response agent harnesses. The video states that five winning agents from the SANS Institute's Find Evil! hackathon were made available on it.
yk is a meta-tracing technology that adds a just-in-time compiler to existing C-based language interpreters. It records interpreter execution, identifies hot loops, generates optimized machine code, and falls back to the interpreter when necessary while preserving the interpreter's language behavior. Its implementation uses ykllvm and serialized intermediate representations, with promotion hints, deoptimization, safepoints, and shadow stacks to manage compiled execution. The project is presented as a way to optimize interpreters such as Lua without maintaining a separate language implementation.
ykllvm is a fork of LLVM used by YK to compile C-based interpreters. It inserts trace-recording functions into the compiled interpreter and serializes a simplified representation of the interpreter's LLVM IR into the executable, allowing YK to construct and compile execution traces at runtime.
Clang is a C language family compiler front end and tooling infrastructure for LLVM, supporting C, C++, Objective-C, OpenCL, and CUDA. It provides GCC-compatible and MSVC-compatible compiler drivers, along with libraries for clients such as refactoring, static analysis, code generation, and IDE integration. In the yk build process described by the video, Clang compiles the C interpreter with yk-specific flags. The project is distributed under the Apache 2 license.
Lua is a programming language with a standard C interpreter used as the baseline for the Lua Mandelbrot benchmark. In the cited demonstration, this interpreter is the target to which yk adds an automatically generated just-in-time compiler.
MicroPython is a lean implementation of Python 3 designed to run on microcontrollers and other constrained environments. It provides a compiler and runtime with an interactive REPL, script execution from a built-in filesystem, a subset of the Python standard library, and hardware-specific modules such as `machine` for low-level device access. The implementation is written in C99, supports multiple architectures, and uses features including a mark-sweep garbage collector, frozen bytecode, a native machine-code emitter, and configurable compile-time options to fit within limited code space and RAM. MicroPython is open-source software released primarily under the MIT license and is developed openly on GitHub; the project also provides the pyboard, an official microcontroller board for running it on bare metal.
YK Lua is a fork of the standard Lua virtual machine adapted to work with YK, a meta-tracing technology that adds a just-in-time compiler to existing C-based interpreters. YK-specific hints such as `yk_promote` expose values, including constants, for optimization; the video also discusses deoptimization, safepoints, shadow stacks, and the risk of divergent traces. A Mandelbrot benchmark demonstrated roughly a four-times speedup compared with standard Lua.
YK MicroPython is an early-stage fork of the MicroPython interpreter enabled with yk meta-tracing technology. The project is presented as adding just-in-time compilation to the existing C-based interpreter; the video reports that its benchmark example runs at about twice the speed of standard MicroPython.
Skill Cabinet is a local catalog for agent skills installed on a machine. It scans user-level skill locations such as .agents, .claude, .codex, .cursor plugins, Hermes profiles, and other ~/.* /skills folders, then lets users filter skills by drawer, metadata, risk, invocation mode, and status; inspect rendered or source bodies, YAML frontmatter, extra files, symlinks, duplicates, origins, and broken links; and review disk usage. It runs with Node 20 or later through npx skill-cabinet, starting a server bound to 127.0.0.1 and opening the catalog in a browser. Users can quarantine skills to ~/.skill-cabinet/quarantine and restore them, or delete individual skills or groups; deletion removes folders, files, or symlinks from the scanned locations, while a symlink's target is retained. The project is distributed under the MIT license.
SkillRadar is open-source discovery, security, ranking, and routing infrastructure for Agent Skills and Codex. It discovers public SKILL.md files, parses their contents, performs conservative static checks for capabilities such as shell commands, dynamic execution, secret access, networking, package installation, filesystem writes, and deployment tooling, then classifies and ranks candidates in a safety-gated registry. D and Blocked candidates are kept audit-only and excluded from automatic routing. Its Codex plugin provides task-to-Top-3 routing, skill search, provenance and safety inspection, and a read-only Skill Budget Doctor. Routing can use a bundled offline registry and returns relevance, SkillRadar score, security grade, provenance, reasons, and match details without executing candidate repositories or depending on live GitHub discovery. The repository includes radar data, a matching system, a router-quality benchmark, a local registry UI, and daily bot-refreshed generated data; it is released under the MIT license.
genart-skill is a Claude Code plugin that provides an AI coding skill for generative art. It teaches deterministic, hash-seeded compositions, including seeding a pseudorandom number generator from a token hash, using named sub-streams, rendering the same composition at different resolutions, and designing traits and rarity tables for editions. It covers Canvas 2D, p5.js, Three.js/WebGL, and SVG, with guidance on ethics, platform-specific workflows, debugging, keyboard shortcuts, PNG and video export, and print and pen-plotter output. The plugin includes scripts for checking same-machine repeatability, distinctness, global-state isolation, and feature stability; rendering a hash-specific PNG or contact sheet; measuring rarity across a census; and exporting batches of images. The scripts can be run in a project and use Playwright when browser rendering is required. The repository explicitly limits its determinism claims, noting that GPU shader compilers, floating-point behavior, rasterizers, MSAA, and JavaScript engines can prevent cross-machine equivalence. It is installed through the Claude Code plugin marketplace and is maintained by Camille Roux. The repository says that CI tests the scripts against a known-good fixture and a deliberately broken variant, while a monthly check verifies that platform-documentation URLs still resolve. It is licensed under the MIT License.
vol-rs is a Rust port of the Volatility 3 memory-forensics framework. It reads the same memory-image formats, runs the same plugins and options, and is designed to produce byte-for-byte equivalent output while running faster than the Python implementation. The project ports Volatility 3's 197 plugins and command-line interface, including plugin listing, CSV rendering, filtering, symbol handling, and configuration options. Its built-in components include symbol decompression, RC4, DES, AES, YARA matching, x86 disassembly, PNG writing, and tar-archive creation, avoiding the Python version's external dependencies for those functions. Symbol packs can be loaded from configured directories, and Windows symbols can be fetched from Microsoft's symbol server and cached locally. The repository reports identical plugin help for all 197 plugins, matching output across tested Windows and Linux captures, and substantially shorter benchmark runtimes on those captures. Testing remains incomplete for real macOS images, 32-bit images, several image formats, and additional kernel versions. vol-rs is distributed under the Volatility Software License 1.0 as an addition to Volatility 3 and is not affiliated with or endorsed by the Volatility Foundation.
A Claude Code skill for converting scripts into MiniMax H3 shot lists and directing character performance. It breaks scenes into shots based on emotional-beat density, using controlled comparisons and documented inferences about MiniMax H3 output; its central rule is to split shots that contain too many facial-expression beats because H3 may otherwise produce a frozen face. The skill also covers dialogue-tag effects on frame allocation, camera placement, observable body actions for silent characters, reference-image use for shapes, edit-based handling of pauses and emotional transitions, PSNR-based verification, dubbing constraints, tail degradation, and a six-step breakdown workflow. It is installed with `npx skills add https://github.com/phileiny/h3-storyboard-skill --skill h3-storyboard` or by copying the skill directory into `.claude/skills/`. The repository distinguishes controlled findings from partly verified and inferred rules, and describes its scope as serialized short-form drama produced with local ComfyUI and MiniMax H3.
Keyword Pro is a standalone, local-first keyword research console developed as an independent open-source project. It uses the user's own DataForSEO credentials rather than a hosted account, signup, or subscription system, and stores research sessions in a local database. Its guided report can make up to 19 keyword-research calls covering search volume, CPC, difficulty, intent, trends, SERP competitors, and related terms. The application supports 94 markets and 46 languages, validates country-language combinations, skips unsupported paid calls, and provides an advanced mode with 332 keyword-oriented DataForSEO endpoint definitions. It displays estimated costs before paid runs, saves completed reports for later reopening without another API call, and exports results as PDF, CSV, and JSON. The application runs locally with Node.js, pnpm, OpenSSL, and Podman or Docker, binding to 127.0.0.1 by default for a single local user. DataForSEO credentials stored by the application are encrypted with AES-256-GCM. The project is distributed under the Apache License 2.0.
FixAnything is an open-source video-refinement tool that repairs rendering artifacts from 3D representations, including 3D Gaussian splats, NeRFs, meshes, and sparse point clouds. It repurposes the pretrained Wan2.1 video diffusion model with minimal modification and fine-tuning, using a FixAnything LoRA to turn a rendered camera-path video into a refined video. The inference pipeline accepts a video file or folder of frames; its documented workflow uses 61-frame renderings resized to 832×480 and writes a generated video, the resized input, and a side-by-side comparison. An optional pipeline reconstructs scenes from a small set of photographs with MapAnything, renders the reconstruction along an interpolated camera path, and then refines that rendering. The repository contains inference code and downloads the underlying Wan2.1 model and FixAnything weights; the code and weights are released under the Apache 2.0 License.
Procedura is an open-source agentic 3D-modeling tool from SpatiaOS that converts text prompts into editable procedural assemblies rather than point clouds or triangle meshes. It generates OpenSCAD source with named parts and typed mates, plans and builds parts incrementally, and can use reference images and optional Blender render feedback during refinement. Optional passes assign per-part PBR materials and plan articulation, exporting motion to OpenUSD and URDF with Isaac-based validation. It runs locally as a Bun/TypeScript pipeline using a configurable OpenAI-compatible, Gemini, or local model endpoint; it does not provide hosted inference or API keys. The pipeline uses a Manifold-capable OpenSCAD build to compile the generated programs and Blender for renders, and includes a web Studio for composing runs and inspecting intermediate artifacts. The repository is MIT-licensed.
CDAF is an open sidecar format and toolkit for video that stores a timestamped plain-text description beside the corresponding video file. Its Python library and CLI can generate, parse, validate, read, and report the status of `.cdaf` files, while an agent skill teaches video agents to check for a matching sidecar before processing footage. Each sidecar contains a minimal versioned header with the video filename, SHA-256 hash, byte size, duration, generator, and creation time, followed by sections such as summary, timestamped segments, transcript, on-screen text, and tags. Conforming tools verify the video's freshness and refuse to use a stale sidecar after the video changes; the format is model-agnostic even though the included generator uses the Gemini Files API, with optional local-model support. The repository includes a normative specification, reproducible sidecar-versus-direct-video benchmarks, an agent skill installable with `npx cdaf-skill`, and CLI commands for generation, validation, reading, and status checks. The core validation functions require only the Python standard library; generation requires Python 3.10 or later and a user-supplied Gemini API key. The project is licensed under MIT.
An Apache-2.0-licensed serving setup for Qwen3.8-27B on a single 24 GB NVIDIA GPU, built around vLLM and exposing an OpenAI-compatible API with optional key authentication. Its preparation pipeline starts from a W4A16 AutoRound model, requantizes the language-model head and embedding matrices to int8, requantizes the MTP components, and applies vLLM patches for the serving stack. The runtime combines int8 tensor-core GEMMs with an fp16 recurrent state, continuous batching, split-KV verification attention, and speculative decoding through either Qwen's MTP path or the optional DFlash2 block drafter; DFlash2 can also draft from the request's cached context for document-reproduction workloads. The repository provides Docker Compose profiles for batch throughput and single-user latency, prebuilt container images, model-download and requantization scripts, benchmark and quality-test tools, and launchers for fast, long-context, and experimental KVarN cache modes. The standard configurations target roughly 64k to 150k tokens of context, while KVarN and alternative int4 KV-cache paths extend the claimed capacity to the 256k-token range with lossy KV quantization. The setup requires a recent NVIDIA driver and a compatible Ampere-or-newer GPU; the repository notes that its benchmark figures were measured on an RTX 3090 subject to a 250 W power limit.
cc-prune is a Python command-line tool for surgical context recovery in Claude Code sessions. It removes already-processed tool output from a Claude transcript while preserving conversation text, reasoning, thinking blocks, transcript structure, and tool-call relationships; it does not summarize or truncate the retained content. The workflow measures actual API usage, inspects transcript byte buckets, optionally audits individual results, and clears selected buckets such as tool results, inputs, attachments, or orphaned command output. Clear operations use a character-length threshold, can protect a tail of recent records, create numbered snapshots, validate invariants, and support restoration. Its validation checks include unchanged record and UUID/parent-UUID chains, byte-identical thinking signatures, and matching tool-result and tool-use identifiers. The tool operates on Claude Code JSONL transcripts under the user's .claude/projects directory and requires the session to be stopped before editing. It supports usage verification after pruning, snapshot listing and restoration, and density calculations based on measured token counts. The project states that the Claude Code transcript format is undocumented and version-sensitive, and that cleared results may need to be regenerated; it is distributed under the MIT license and requires Python 3.9 or later with no dependencies.
An Apache-2.0 C++23/GGML port of SkinTokens and TokenRig for automatic 3D mesh rigging on CPU or Vulkan. It takes a static GLB mesh, uses a Michelangelo point encoder, Qwen3-based TokenRig policy, FSQ expansion, condition encoder, and SkinVAE decoder to predict a skeleton and per-vertex skin weights, then exports a rigged GLB. It can also generate weights for an existing static or animated skeleton, inspect GLB assets, and retarget recognized SOMA30 motion to the 52-joint Mixamo order or to a generated humanoid rig. The project provides a command-line interface, shared library, C11-compatible API, and a Go/WebGL demo; converted GGUF model weights are available separately, with optional Vulkan support.