2,454 tools and products — trending open source, and what gets used in AI and other work.
FluidUse is an open-source macOS library and demo application for local computer use on Apple silicon. It reads forms in running Mac applications or browsers through the macOS Accessibility API, uses an on-device decision model to match visible fields with supplied profile data, and types or selects the resulting answers in the real application. A WebFormDriver supports embedded WKWebView forms. The project converts CUA-S1-FORMS to Core ML and runs it through FluidAudio; it also includes laya-based typed decision models for choice, score, and yes/no-style questions. DocumentEntities can turn PDF or text files containing label-and-value lines into a profile, while a predetermined-answer sheet handles question-style fields outside the model's decisions. The repository states that processing stays on the machine and that the tool does not read résumés, reason about dropdown options, write free text, upload files, or click Submit unless enabled. FluidUse is licensed under Apache 2.0 and requires Accessibility access for its launching terminal.
AI Manager is a local-first desktop application for installing, updating, configuring, and repairing AI coding command-line tools on a user's computer. It provides a single interface for checking installed tools, managing versions, connecting tools to reviewed or custom HTTPS AI-service endpoints, and viewing the effective endpoint and environment-variable overrides without editing shell profiles or PATH files. Its sidebar includes software management, API endpoints, skills, MCP connections, global prompts, sessions, and settings. It installs skills from trusted GitHub catalogues or local ZIP files, supports stdio, HTTP, and SSE MCP servers, and searches local coding-tool sessions that can be resumed in a terminal. Tool behavior is controlled by a capability registry, and the application uses local signals for checks rather than claiming a process is healthy without evidence. The application is built with Tauri 2 and React 18 and is released under the GNU AGPL-3.0-or-later. It has no analytics or telemetry SDK, stores product data locally, and offers signed update verification and platform-specific installers. Provider credentials are currently stored in plaintext in the local database, backups, and SQL configuration exports. AI Manager began as a fork of CC Switch, with inherited portions retaining their original MIT licensing.
JevRouter is a local-first agent capability router for models, subagents, skills, MCP tools, CLIs, and DSH plugins. It presents these capabilities as one candidate set and uses Jev for typed selection while enforcing availability, permissions, risk, and confirmation policies. It is decision-only by default and does not execute selected tools implicitly. The `route` operation answers one capability-selection question; `plan` selects capabilities for multiple steps using serial, batch, or decomposed strategies, with optional beam sequence search and contextual routing. Candidates can be supplied inline or discovered from Skill directories, MCP configurations, CLI help output, and DSH manifests. The router preserves Jev probabilities and provider responses, validates supplied permissions and input schemas, filters unavailable or disallowed candidates, and can select a safe fallback while retaining the original Jev choice. The CLI, Node.js SDK, HTTP server, and MCP adapter expose the routing contract. Decisions and plans are written as append-only receipts with provenance hashes, and a local dashboard reads those files without making provider requests or uploading data. It supports the official Jev API and OpenRouter, has an offline demo provider, requires Node.js 20 or later, and is distributed under the MIT license.
jev-voice-browser is a Node application that controls a headed Chromium browser by voice or typed commands. It streams partial speech transcripts from the browser's Web Speech API to a Node server, which uses TypeSafe's Jev model to classify the intent, select a page element or site, extract candidate text and URL spans, determine whether the command is complete or addressed to the browser, and assess whether an action is destructive. Playwright then executes navigation, searches, typing, clicking, scrolling, history, and tab actions. The application sends the page URL, title, detected site, a compact snapshot of visible elements, and recent actions to Jev for each transcript update. Code—not the model—constructs URLs, copies text verbatim, and applies policy thresholds to decide whether to act, wait, ignore, request confirmation, or show numbered alternatives. Recent action context supports corrections such as reversing an action, rejecting a selected result, or choosing another candidate. It requires Node.js, Chrome or Edge for microphone input, Playwright's Chromium installation, and a TypeSafe API key. The server can launch a persistent-profile Chromium window or attach to an existing browser through CDP; it also supports headless operation and typed commands when no microphone is available.
reflex is an open-source local decision model for classifying tickets, documents, photos, and other state against typed questions with fixed answer choices. It runs the state through an open-weights model in one forward pass, branches each question independently, reads only the answer-label scores rather than generating free text, and returns probabilities for yes/no, choice, or ordered-score questions. It can read each question in multiple option orders and average the results to reduce position bias; optional temperature calibration adjusts confidence estimates for a target workload. The project provides a Python engine, an HTTP endpoint at /v1/systemone, GPU and Apple Silicon execution, an SGLang backend, a WebGPU browser demonstration, and optional LoRA training and distillation tools. The default serving configuration uses a frozen Qwen checkpoint, and the repository is licensed under MIT; model weights have their own licenses.
Laya is an open-source, Jev-compatible System-1 decision model package by Convai Innovations for running typed decision questions from Node.js and TypeScript through ONNX Runtime. It accepts a state such as a ticket, email, or JSON object and returns calibrated answers for choice questions with per-option probabilities, ordered score questions with a probability distribution and expected level, or yes/no statements with a probability of true. Questions in one call are batched into a single model run, and the package uses the same request and response shape as the Python reference implementation without requiring Python or PyTorch at runtime. The package downloads and caches ONNX model weights from Hugging Face by default, or can load a local ONNX bundle; it requires Node.js 20 or newer. The JavaScript package is MIT-licensed, while the Laya model weights are published under Apache 2.0.
DocJev is an open-source Python library, CLI, and local app for classifying and splitting PDF, DOCX, and PPTX files with user-written natural-language rules. LiteParse extracts page text locally, then the hosted Jev decision engine assigns document categories or identifies boundaries between component documents; optional LlamaParse tiers provide cloud OCR for difficult inputs. Classification returns a category, probabilities, and review flags. Splitting returns ordered categories and contiguous page ranges, with optional export of each segment as a PDF and review reasons for uncertain boundaries. DOCX and PPTX processing requires LibreOffice, and inference is not offline because Jev is a hosted service. The project is distributed under the Apache-2.0 license.
Jevmind is a Python decision layer for software agents. It turns agent decisions into typed answers such as yes/no results, choices, and scores with confidence values, then applies a code-defined gate: uncertain answers escalate, trusted positive answers may act, and trusted negative answers are held. Its decisions are recorded in an append-only, hash-chained ledger, graded against later outcomes, and optionally recalibrated with Platt scaling. It includes nine skills for compacting tool output, navigating repositories, reviewing diffs, routing tasks to model tiers, checking completion claims, curating training data, walking linked markdown vaults, guarding commands, and running a structured-state control-loop demo. The default local brain uses readable rules, reflexes, and BM25 retrieval without a language model; an optional Jev HTTP brain can answer the same typed questions, while replay mode uses recorded answers. Jevmind runs as a command-line tool and MCP server, includes a Claude Code hook for shell-command screening, supports Python 3.10+, has no runtime dependencies, and is distributed under the MIT license.
jeff is a self-hosted, MIT-licensed drop-in implementation of TypeSafe's jev System One API, powered by the GLiFormer encoder model. It accepts classification questions through a compatible SDK or the POST /v1/systemone endpoint and returns choice selections, ordered scores, and noul yes/no probabilities. The server supports CUDA, Apple MPS, CPU, and ONNX Runtime deployments, with batching, bearer-key authentication, rate limits, request limits, health and statistics endpoints, and deployment configurations for Modal. Its repository reports lower hosting costs than jev but weaker results on reasoning-heavy tasks.
A personal Codex skill and orchestration workflow for splitting software-development tasks between GPT-6 Astra and DeepSeek V4.1 Flash. Astra handles planning, architecture, high-stakes decisions, task briefs, and final review; Flash handles scoped repository discovery, implementation, testing, debugging, and routine verification, then returns a patch and evidence for Astra's acceptance pass. The workflow supports existing plans and phased implementation bundles, normally uses one writer, and includes verification, checkpointing, dry-run installation, backups, and guarded undo receipts. It is an early release requiring a Codex client with native subagents, Python 3.11 or newer, GPT-6 Astra as the root model, and an already configured Codex Router route for DeepSeek V4.1 Flash. It installs a skill, a named builder agent, and a scoped workflow policy without changing credentials, the root model, or Router configuration; no third-party Python dependencies are required.
Null Motion is a local browser-based motion-design preview and export tool that presents a finished advertisement alongside the black-and-white HyperFrames drafts used to plan it. It includes an earlier launch editor and ships with example references and authored draft scenes. The project uses FFmpeg scene detection to divide reference videos into sections while preserving their cuts. Each section is represented as a 640×360 HTML scene on a paused GSAP timeline, with beats such as text, logos, phones, windows, chats, cards, lists, charts, and clouds. During playback, the active draft is driven by the film’s clock and stays synchronized frame by frame with the reference video. For export, a hidden video plays at half speed; requestVideoFrameCallback captures decoded frames, the drafts are rendered at each frame’s timestamp, WebCodecs encodes the video, and mp4-muxer produces an MP4 with copied reference audio. The tool runs with Node.js 22 or newer and requires Chrome or Edge for WebCodecs export. Reference media remains local, and the repository states that the drafts are authored rather than AI-generated. No project-wide open-source license has been selected.
Mr. Mak Workspace is a local Windows desktop workspace for Codex CLI and Claude Code CLI. It combines real CLI chats with a project workspace for folders, research, images, files, Markdown documents, project reports, and saved project materials. The application provides chat history and tabs, agent activity indicators, file and folder insertion, project cards, previews for media files, and a right rail for project skills, knowledge, workflows, MCP connections, settings, and help. The repository includes fourteen project skills covering areas such as planning, handoffs, image references, media generation, Three.js, Blender, video inspection, and dictation. It can also provide an optional voice coordinator using Codex CLI and an OpenAI API key; external providers and MCP connections are optional and use the user's own accounts. The Windows installer includes the local Node service but does not bundle the supported CLIs or provider accounts. The native application targets Windows x64, while browser report previews can run elsewhere. Application code and original workflow documentation are MIT-licensed, with third-party skill materials retaining their original licenses.
ThinkingOrbs is a SwiftUI package by Haplo LLC that provides dotted 3D loading indicators for AI and agent interfaces. It includes nine hand-tuned designs representing states such as searching, solving, listening, connecting, composing, and waiting; eight use rotated, depth-shaded, z-sorted 3D forms and one uses a morphing outline. The package provides regular and small sizes, status-label views with optional shimmering text, localization and accessibility labels, speed and pause controls, and public frame geometry for custom renderers. Each orb is drawn with grayscale dots in a Canvas inside a TimelineView. Timelines pause when an orb is offscreen or the app is backgrounded, while Reduce Motion displays a representative static frame. The package supports iOS 17+, macOS 14+, tvOS 17+, watchOS 10+, and visionOS 1+, and is distributed as a Swift Package under the MIT license.
ddc is a Rust command-line Android decompiler that converts DEX bytecode from APKs and related containers into readable Java. It supports multi-dex applications, XAPK/APKS/APKM packages, DEX versions 035–041, lambdas, string concatenations, and selected Android and Kotlin-specific output transformations. Its progressive-analysis interface provides subcommands for inspecting application metadata, manifests, resources, strings, cross-references, classes, methods, disassembly, and class hierarchies without requiring a complete decompilation first. The repository reports that its generated Java output passes a javac syntax check across its validation corpus and uses deadline-bounded processing for pathological classes. The CLI can be installed through Homebrew, Scoop, cargo-binstall, or built from Rust source. It provides bilingual English and Simplified Chinese messages, emits reproducible output without timestamps, and is released under the MIT license.
Herdr GPUI is a native Rust/GPUI desktop client for a separately installed Herdr daemon. It displays the daemon's terminal sessions, split panes, workspaces, Git worktrees, and agent activity without running another terminal emulator or wrapping the TUI. The daemon owns terminal processes and session state. Herdr GPUI connects to its local or saved remote daemon through the Herdr client socket, receives binary protocol frames and surface patches, renders the terminal surfaces, and sends semantic input back; closing the GUI leaves the daemon and its sessions running. The project is distributed from the repository as a signed, notarized macOS app through Homebrew, with additional Linux packages, experimental Windows builds, Nix support, and source-build instructions. The daemon must be installed separately. It is licensed under Apache-2.0.
Glyd is a Rust lossless-compression library and command-line tool with a C ABI, streaming interfaces, and Python and Go bindings. It targets object storage, data lakes, logs, telemetry, backups, versioned exports, RPC payloads, and caches as an alternative to LZ4, Snappy, and zstd. Its record mode detects delimited records, SQL dumps, JSON lines, and varying-shape logs, converts fields or template variables into typed streams, and compresses them in parallel. Base mode compresses a new version against an older version by using regions of the old data as history. The format uses independently decodable blocks and parallel paths; higher levels use separated token, offset, length, and literal streams with SIMD-oriented decoding, while long-distance matching can find repeated content up to 128 MB back. The repository also provides shape dictionaries for small objects, packs for many small objects, container and JPEG handling at selected levels, and a cold context-mixing level for infrequently read data. The separate glyd-store crate and CLI select similar stored objects as delta bases using fingerprints, cap delta chains, and support local directories and S3-compatible backends. The codec, CLI, C ABI, and bindings use the BSD 3-Clause License or GPL-2.0; glyd-store uses the Business Source License 1.1, with commercial production use requiring a license and conversion to Apache-2.0 four years after each release.
Veo is an AI video-generation engine that creates short video clips from prompts. The demonstrations describe it as better suited to slower, emotional scenes than to high-action sequences.
Musk Miners is a U.S.-based crypto-mining equipment store and hosting service. It sells ASIC mining machines and supplies, including Zcash and Bitcoin miners, and offers equipment hosting with installation, miner operation, monitoring, maintenance, power, and technical support. Its stated process is to order a machine, sign a mining contract, and have the miner hosted and operated by Musk Miners. The company advertises white-glove service, one year of free maintenance for specified components, 24/7 support, and hosting powered by clean energy at stated rates.
Kerbal Space Program is a spacecraft-flight simulator that receives translated controller commands and simulates the resulting spacecraft operations.
Kerbal RPC is software that reads MIDI events from a DJ controller and translates them into control commands sent to Kerbal Space Program. The creator described it as code made available to supporters through Patreon.
Mido is a Python library for interpreting MIDI messages, including events received from a DJ controller. It can be used to translate MIDI input into commands for other software.
Paperclip is an open-source, self-hosted control plane for organizing and running teams of AI agents. Its Node.js server and React UI let users define company goals, assign agents roles and tasks, and monitor work, budgets, approvals, costs, and audit activity from a shared dashboard. Agents run through scheduled or event-based heartbeats. The system maintains organization charts, goal ancestry, task dependencies, persistent session state, isolated workspaces, scoped secrets, runtime skill injection, and atomic task checkout with budget enforcement. It supports agent adapters and runtimes including Claude Code, Codex, Cursor, Bash, HTTP/webhook agents, and plugins; recurring routines can create tracked issues and wake assigned agents. Paperclip supports multiple isolated companies in one deployment, governance and approval gates, cost hard stops, agent pause/resume/termination, company export and import with secret scrubbing, and optional OpenTelemetry and Sentry integrations. It can be installed through the Paperclip CLI or run from source; local setups use embedded PostgreSQL and file storage, while production deployments can use an external PostgreSQL database. The project is distributed under the MIT license.
OpenBao is an open-source software solution for managing, storing, and distributing sensitive data such as secrets, certificates, and keys. It encrypts arbitrary key/value secrets before writing them to persistent storage, supports storage on disk and PostgreSQL, and can generate system-specific dynamic credentials such as AWS or SQL database secrets on demand. Generated secrets have leases that clients can renew through built-in APIs and that OpenBao can automatically revoke; revocation can also target individual secrets or groups such as all secrets accessed by a user or of a particular type. OpenBao also provides encryption and decryption without storing the data, along with audit and key-rolling capabilities. The project is community-run under open-governance principles and is implemented in Go, with a command-line server and web UI.
Cosign is a professional-reputation product for the startup community that serves as a directory of people and companies. It records public and private endorsements, professional intent, questions, announcements, and connections so users can discover talent and make trusted introductions. Its reputation model emphasizes granular signals such as mentors, investors, colleagues, and people or products a user is willing to endorse.
Better Auth is an authentication library used to add sign-in and protect an application's administrative interface.
Rebuild YouTube with AI is a live, hands-on cohort course from ByteByteGo about building and deploying a YouTube-like video site with AI coding tools. It covers decomposing the product into core surfaces, generating interface mockups, implementing React and Next.js pages, modeling data in Postgres, adding authentication and uploads, generating media, storing video with Cloudflare, and deploying with Vercel. The course also implements search and related videos with multimodal embeddings and pgvector, then discusses recommender-system limitations, watch-time tracking, AI-driven Playwright testing, and production workflows. Sessions are recorded, and enrollment includes course materials and access to the cohort recordings after the live course.
Harbor Freight is a retail store and online shop for low-priced power tools, generators, jacks, tool boxes, hand tools, and related equipment. Its product departments include automotive, welding, power tools, air tools, lawn and garden, safety, electrical, material handling, plumbing, painting, and construction hardware. The company also offers replacement parts, tool-protection plans, commercial accounts, store pickup and location services, and membership discounts through its retail network.
Home Depot is identified in the video as a source of scrap lumber and ordinary construction supplies.
GitHub Actions Runner Images is the source repository used by GitHub to create VM images for GitHub-hosted Actions runners and by Microsoft for Azure Pipelines hosted agents. It defines Linux, macOS, and Windows images for x64 and Arm64 architectures, including their YAML labels and preinstalled software. The repository contains image definitions, build sources, release information, and policies for software installation, image support, version updates, and deprecation. Images are typically updated weekly; the `-latest` labels migrate gradually to newer stable operating-system versions, while GA images support at most the two newest OS versions and beta images are used for feedback before general availability. The repository also documents package-manager-based installation, supported tool-version strategies, and instructions for building a VM image from the source.
Mobile MCP is an open-source Model Context Protocol server from Mobile Next for automating and scraping native iOS and Android applications on simulators, emulators, and connected real devices. It exposes a platform-agnostic set of MCP tools for device management, app installation and control, taps, swipes, text input, screenshots, screen recording, device logs, crash reports, deep links, orientation, location, and clipboard operations. The server primarily drives applications through native accessibility trees and structured UI-element data, falling back to screenshots and coordinate-based actions when required. It can run locally over stdio or as a Streamable HTTP server, and can use Mobile Next Cloud to reserve remote physical devices through the same tool interface. It is distributed as the `@mobilenext/mobile-mcp` npm package and requires platform tooling such as Xcode command-line tools or the Android SDK for device access.
LAPTOP is a meme coin that the video describes as being created by Hunter Biden. Its token supply is partly distributed through an airdrop; tokens are burned when associated predictions come true, while tokens tied to unsuccessful predictions are directed to charity.
OpenRig is an open-source, self-hosted multi-agent harness for coordinating AI coding agents such as Claude Code and Codex as one managed team. It uses a local daemon, CLI, terminal UI, MCP server, SQLite, tmux, and runtime adapters. Users define agent topologies in YAML with pods, seats, communication edges, and continuity policies, then use commands such as `rig up`, `rig send`, `rig broadcast`, `rig chatroom`, `rig down --snapshot`, and `rig discover`/`rig adopt` to launch, coordinate, inspect, snapshot, restore, or take over existing sessions. Its MCP tools let agents manage their own topology, while the TUI exposes topology, seat, project, feed, terminal, and system views. OpenRig requires Node.js and tmux; herdr, cmux, and Docker are optional for terminal workspaces and service-backed rigs. It is distributed through the `@openrig/cli` npm package and licensed under Apache 2.0.
Topcoat is a modular, batteries-included Rust framework for building full-stack web applications with server-side rendering, components, API routes, Tailwind CSS support, asset handling, and client-side interactivity. Its `view!` macro combines HTML-like templates with Rust control flow, while components can be asynchronous and access server-side resources such as databases directly. Topcoat evaluates `$()` expressions on the server for the initial render and translates them to JavaScript for browser-side reactivity without a WebAssembly bundle or separate client build step. Server-dependent updates can be implemented with `#[shard]` components, which re-render on the server when their arguments change and replace the resulting HTML in place. It also provides signals, event handlers, server procedures, request-context and per-request memoization, live updates, module-based route discovery, cookies and sessions, a CLI formatter, vendored editable UI components, and an asset bundler. The repository describes it as early-stage and experimental, with breaking changes expected.
Whiteboard is an open-source desktop app for humans and coding agents to architect software in a shared workspace. It connects to agents such as Claude Code and Codex, giving them an SDK for drawing diagrams and describing their work on an in-app canvas. Visualizations including sequence diagrams, entity-relationship diagrams, and agent-trace quotations can link directly to underlying code, while the code-review interface provides VS Code keybindings and LSP support. Whiteboard includes a semantic, AST-aware diff viewer written in Rust. It summarizes large added functions as pseudocode and can collapse or hide changes such as tests and documentation; a WebAssembly-based plugin system makes these rules customizable. Its decision-log tools let agents query and link their traces so users can inspect requirements, implementation choices, and autonomous decisions. The app runs against local checkouts, is distributed for macOS, Windows, Ubuntu, and Fedora, and is licensed under the MIT License. It is self-hostable, although hosted team functionality is planned. Current limitations include no file editing within Whiteboard and limited support for reviewing multiple repositories together.
NeoHorse is TokenRhythm's family of open-weight causal language models for text-based agent workflows, including tool use, coding, and instruction following. NeoHorse-1 uses routing-guided agentic post-training: a harness assigns tasks to a heterogeneous model pool, records tool interactions and outcomes, estimates capability demand, and feeds capability-level feedback into later training mixtures. Updated models can return to the harness in an evaluation-selection-update loop intended as a prototype for recursive self-improvement. NeoHorse-1 is released in approximately 4B and 9B parameter variants, post-trained from Qwen3.5 models, with text input and text output interfaces and deployment examples for vLLM and SGLang. The related NeoHorse-Jev-4B model uses prefill-only inference to return structured Choice, Noul, and Score decisions and probabilities for routing requests, selecting tools, or controlling workflows. The models are released under the Apache License 2.0.
Jive is an open-source terminal coding agent that plans work as executable graphs. It replaces sequential tool calls with graph calls: the planner produces a directed acyclic graph containing tool calls and Jev calls, allowing the agent to perform multi-step profiling, dataset analysis, and repetitive tasks with fast decisions inserted between harder reasoning steps. The project emphasizes a lightweight, token-efficient agent scaffold with eager execution, flexible graphs, and support for system-one and system-two models. It can be installed with the project's shell installer and is distributed under the MIT license.
cli-faq-shortcuts is an agent skill that analyzes a developer's coding-agent history and turns repeated requests into project-specific slash-command shortcuts for Claude Code, Codex, and Cursor. Its local extractor reads prompts from the relevant Claude Code, Codex, and Cursor history files, excluding subagent turns and scripted runs. The agent groups differently worded prompts by intent, counts recurring asks, checks existing shortcuts and project tools, and presents candidates for the user to select. Each selected shortcut is written as a SKILL.md file under the project’s .claude/skills directory, linked through .agents/skills so the supported agents can load it; the repository also records mappings in CLAUDE.md or AGENTS.md. Shortcuts can point to existing scripts and queries, include project-specific cautions, and use a dry run before actions that write or send data. The extractor can be run independently with Python and reads local files without making network calls. The repository supports installation on macOS, Linux, and Windows, and is licensed under Apache-2.0.
Keel is a local-first native macOS coding workspace that runs coding agents through a Rust/GPUI application. It combines sessions, a composer, transcripts, a terminal, change inspection, settings, Git history, crash recovery, and an agent graph showing active agents, subagents, and files as shared context; agents can be steered or stopped from the graph. It connects Claude Code, Codex, Cursor, Grok, Hermes, and pi through the Agent Client Protocol, and also includes an embedded DeepSeek agent loop with host-enforced tool focus. For unpinned tasks, a local Laya selector running through Core ML or an optional hosted Jev selector chooses among host-prepared routes or abstains. Keel validates each choice before applying it, falls back to the ordinary route when a choice is rejected, stale, or expired, and records candidates, validation, fallback, and observed outcome in bounded decision receipts. Decision receipts can be exported as JSONL and summarized with CLI reports. Keel also provides headless and daemon modes, while each coding provider retains its own authentication, model configuration, and tool loop. The repository supports Apple Silicon and macOS 15 or later for the packaged Laya model. It can be built from source, but the repository does not currently publish downloadable app releases; packaged builds are ad hoc signed development builds. Automatic training, public installer distribution, cloud sync, and native Codex realtime voice are not implemented or included in this build.
An agent skill for Claude that guides logo projects from discovery briefs through concept development, SVG construction, testing, refinement, and delivery of brand assets. Its workflow researches category conventions in a classified library of more than 1,400 reference SVG logos, develops 8–12 concepts, builds selected concepts in SVG, and pauses at a checkpoint for direction selection before producing a full kit. Dependency-free Python tools audit SVG structure and craft rules, generate concept and test sheets, check marks at sizes down to 16 px in one-colour and reversed treatments, compare them in contextual and competitor-shelf tests, render PNGs, create presentation boards, and export favicon, app-icon, web-manifest, lockup, and other variants. It can also produce icon sets and brand-guideline materials. The skill follows the Agent Skills format, supports Claude Code, Claude.ai, Claude Desktop, and compatible agents, requires Python 3.8+ for its tools, and is distributed under the MIT license; the reference logo files are third-party trademarks excluded from that license.
LaunchVideo is a tool that turns a website URL or prompt into a 20–40-second launch video. Opus 5.5 writes a single HTML film, while a serverless OpenComputer agent checks the scene at multiple timestamps and renders each frame in headless Chromium under a controlled virtual clock. The frames are encoded into an MP4 with ffmpeg and uploaded to Vercel Blob. The project includes a Next.js web form, OpenComputer agent configuration, API routes for job creation and status, and a CLI or one-click deployment path. URL inputs are fetched for page metadata, headings, colors, and Google Fonts before generation. Rendering runs in a fresh Amazon Linux arm64 microVM and supports deterministic HTML scenes subject to restrictions such as no video, audio, iframe, CSS transitions, random values, or external images.
YuE2 Studio is a Windows desktop application for local AI song generation. It takes a musical style and lyrics, uses YuE2 to compose a melody and chords as ABC notation, engraves the result as sheet music, and renders the composition as a full song with vocals. The score can be edited and rendered again with different sounds, while saved audio codes support exact replay and variation without recomposing. The studio also supports score-first composition, melody transcription for covers through SheetSage2, lyric-level karaoke timing, six-stem separation, audio-to-MIDI conversion, LoRA loading and training, audio processing with VST3 plugins, and an optional local or OpenRouter writing assistant. It can be controlled by AI agents through a local MCP server. It is distributed as a native Windows installer with auto-update or a portable folder and does not require Python or Node.js at runtime. Generation runs offline through the bundled yue2.cpp engine on NVIDIA, AMD, or Intel GPUs, with Vulkan support for the latter two described as experimental; CPU execution is also available. The studio is MIT-licensed, while the YuE2-3B, YuE2 VAE, and SheetSage2 models are CC BY-NC 4.0 and restrict generated songs to non-commercial use unless other terms are obtained.
Interference Search is a research project and Python library for reasoning over explicit states rather than a single linear language-model transcript. It expands live branches in parallel, lets the environment execute their moves, merges branches that reach the same state, uses a trained judge to discard states unlikely to reach the goal, and advances the surviving frontier one level at a time. The method is classical and does not claim quantum speedup. The repository implements Countdown arithmetic search and program search, including move generation, exact solving, judge training, sandboxed program tests, behavioral merging, benchmarks, raw results, failed experiments, a research log, and the accompanying paper. Its reported experiments compare the approach with linear language-model reasoning and other search strategies; it also includes a one-line baseline and tools for reproducing the Countdown results. The package requires Python 3.10 or newer and runs with PyTorch. Language-model experiments use MLX on Apple silicon, while the core search, judge, and tests can run on other PyTorch-supported systems. It is distributed under the Apache 2.0 license.
BridgeClip is an open-source AI desktop app from BridgeMind that turns local videos or podcast, YouTube, and Twitch VOD links into captioned short-form clips. It downloads the source when needed, transcribes speech, uses OpenRouter to identify promising moments and plan the clips, reframes them to 9:16 or 16:9, and renders the results locally with FFmpeg and word-by-word captions. It supports smart face and screen-share framing, multiple caption styles, clip libraries and job tracking, and optional publishing or scheduling through Zernio. The app runs on the user's computer without a BridgeMind backend; provider keys and transcription or planning data are sent to the selected OpenRouter services, while rendering is local. macOS and Windows releases are provided, Linux development is supported, and the project is MIT licensed.
Omarchy Meeting Recorder is a local meeting-recording application for Omarchy on Hyprland, built with Rust, GTK 4 and libadwaita. It captures microphone audio and the computer's output as separate tracks, then uses whisper-rs and a local Whisper model to produce a timestamped transcript with speaker labels. Recordings can also be imported from files; imported single-track audio uses NVIDIA Nemotron 3 Diarization through ONNX Runtime to distinguish speakers locally. The application provides an editable transcript, waveform playback synchronized with transcript lines, speaker renaming, chapter markers, pause and crash recovery, command-line controls, and optional bar-widget integration. Chapters can be generated by the configured Omarchy coding agent; in that case, the transcript text is sent to that agent, while recording and transcription otherwise remain on the user's computer. It stores meetings as folders containing audio, Markdown transcripts, metadata, and separate track files. The project is distributed under the MIT license and can be installed from the Omarchy package repository or built from source.
jev-use is a macOS voice and text computer-use harness built on Jev. It reads the frontmost app's Accessibility tree rather than taking screenshots, sends the command, app and window names, visible targets, and recent actions to Jev, then selects and performs operations such as clicking, typing, pressing keys, scrolling, opening items, and arranging windows. After each operation it reads the Accessibility tree again and continues until the task is complete, blocked, or waiting; low-confidence or destructive actions stop for confirmation. The app supports hold-to-talk, typed commands, hands-free listening, and an optional “Hey Jev” wake phrase, with speech handled by Apple Speech. It requires macOS 14.2 or later and Xcode, has no external dependencies, and stores the TypeSafe API key in the macOS Keychain.
mu (μ) is a coding agent with a judgment kernel, built on pi and AionUi. A small judge handles bounded decisions around context admission, command risk and approval, task framing, memory, completion, drift, browser actions, and multi-agent coordination, while the main model performs the coding work. Each decision can be active, shadow, or off and can use Jev, the local Laya judge, or another LLM; verdicts, probabilities, and timings are recorded in a local ledger that can be inspected from the command line or desktop app. The system admits tool output chunk by chunk, archives or removes stale context, applies rule-based safety checks before judgment, and uses a shared board to route findings among non-editing sub-agents. It provides an npm-installed command-line interface and a native desktop app, supports model-provider connections and importing Claude Code or Codex conversations, and states that it is in early development with no release yet; names, settings, and formats may change.
Jevry is an open-source, MIT-licensed desktop browser agent created by Michael Swissa. It is an Electron browser that combines a language model for planning and text or visual reasoning with TypeSafe's Jev model for selecting typed actions from controls observed on the current page. Chromium executes the selected operations through guarded native input rather than generated JavaScript, arbitrary selectors, or shell commands. Its loop observes page text and compatible controls, compiles supported operation-and-target choices such as clicking, typing, scrolling, or stopping, asks Jev to choose, validates the choice against the current document and target, executes it, records an input receipt, and checks the resulting evidence before continuing. It supports browser tasks such as forms, page-grounded questions, research and citation, documentation navigation, catalogs, and selected visual or numeric games. The project is an experimental developer preview for macOS and Windows; it requires a TypeSafe Jev connection plus a text-model connection, and some capabilities—including uploads, extensions, password-manager workflows, and certain cross-origin interactions—remain unsupported or unverified.
product-film is a Claude Code skill for creating product films in Remotion, including landing-page loops, launch videos, promos, and demo reels from a product's real components, design tokens, logo, and stated claims. It first inspects the product's design rules, tokens, components, live site, logo, and claims, then interviews the user about the film's duration, placement, music, features, and visual ingredients. It records the resulting brand kit in videos/BRAND.md, creates a beat sheet, analyzes the music's beat grid, builds time-based Remotion scenes, and supports review through style frames, stills, contact sheets, and a draft render. The final workflow renders a 240 fps master with motion blur and can produce a muted loop, a music version, a WebM file, and a poster; verification scripts decode the outputs and check colors, duration, and loop continuity. It runs in a local Claude Code toolchain rather than the Claude chat apps, requires Node.js, Bun, and uv, and is distributed under the MIT license for its own text and code.
Jevgrep is a command-line tool for coding agents that finds relevant files and source context in unfamiliar repositories from natural-language questions. It uses Jev to judge relevance across folders, files, and declarations, explores qualifying branches of the repository hierarchy, selects files through content previews, identifies useful source units and surrounding context, and returns a summary followed by file locations, reading leads, and verbatim excerpts with line references. Python and TypeScript/JavaScript support declaration parsing, while other text uses a fallback. It also includes an agent skill that teaches supported coding agents when to invoke the CLI and how to use its results. The CLI is distributed through npm as @dzhng/jevgrep, requires Node.js 22 or later on macOS or Linux, and requires authentication with a supported AI provider. Searches send eligible source content to Jev through the selected provider; local filtering excludes ignored, hidden, dependency/build, binary, and obvious credential files but is not a guarantee that sensitive data is removed. The project is MIT-licensed.
reladraw is a text language and command-line tool for creating diagrams with relative positioning. It sits between automatic layout systems such as Mermaid, Graphviz, and D2 and manually positioned editors such as draw.io: users declare nodes, relationships, labels, and relative placement without specifying absolute coordinates. The parser and layout engine resolve the declarations into a diagram, while the tool reports issues such as overlaps, crossed edges, and text overflow without requiring a rendered preview. Its TypeScript CLI converts `.reladraw` files to SVG, and the repository also provides a browser version and an optional agent skill. The project is early-stage, its syntax is subject to change, and the code is licensed under Apache-2.0.
CogSend is a self-hosted social media scheduler that runs on the user's Cloudflare account. It lets users write a draft, customize it per platform, add images and alt text, automatically split long drafts into threads, and publish immediately or schedule posts for Mastodon, Bluesky, LinkedIn, Threads, and X. The scheduler reports delivery results per account, retries temporary failures, and supports cancellation, rescheduling, and manual retries; it also provides link previews and publication-failure insights. The application stores data in Cloudflare D1 and media in R2, sends posts through the user's API credentials, encrypts credentials, and protects the administrator account with two-factor authentication. It provides a personal API key and an MCP server for scripts, Shortcuts, Claude Code, Codex, or other MCP clients to draft, schedule, and publish posts. It is built with SvelteKit on Cloudflare Workers, uses Drizzle with D1 SQLite, and is licensed under the MIT License.
disktree is a disk-usage treemap for finding and removing files and directories that occupy space, developed in Rust with GPUI. It scans a home directory, selected directory, mounted volume, or whole disk; draws nested blocks sized by actual disk usage; and supports navigation by size, file count, or age, filtering, hidden-file inclusion, apparent-size measurement, and classification of data such as caches, build output, media, and repositories. Users can mark items, review the complete selection, estimate the resulting free space, and move selections to trash or delete them permanently with confirmation and removal safeguards. The scan accounts for hardlinks, avoids following symlinks by default, and aggregates directory sizes in parallel. It runs on Linux, macOS, and Windows, with platform-specific disk-usage and trash behavior, and is licensed under the MIT License.
INKWAVE is an open-source browser-based 4v4 turf-war shooter built with three.js. Players paint arena surfaces, swim through their team's ink to move quickly and refill their tank, use weapons and special abilities, and compete against bots across three stages; the team with the most painted ground after three minutes wins. It includes seven weapon types, squid movement, map viewing with Super Jumps, character customization, and keyboard, mouse, and gamepad controls. Ink is painted into texture-space regions on a GPU-managed 4K atlas, while a coarse CPU grid tracks turf scores and gameplay queries. Shaders add height, gloss, wetness, drying, and wall-drip effects. Stages are defined as data and mirrored by 180 degrees for symmetrical team layouts; characters use procedurally generated geometry, materials, a 60-bone rig, and code-driven animations. Weapons, actors, match logic, effects, HUD, and audio communicate through typed events, and deterministic freeze-and-step tooling supports reproducible bot simulations and filmstrip generation. The game has no build step for local development and can run from a static file server; its included server starts the game at localhost:8490. Chrome and Edge are the target browsers, with Firefox supported and Safari described as slower. The repository is licensed under MIT and states that INKWAVE is an independent project unaffiliated with Nintendo.
Tidewater is a browser-based island fishing game. Players fish from a pier, beach, or boat, manage line tension while fighting catches, sell fish to a village vendor, and spend the proceeds on equipment such as stronger line, reels, rods, a larger hold, fuel capacity, an engine, a fish finder, and deck lights. The island and surrounding ocean can be explored on foot, by swimming underwater, or by boat, with fishing conditions varying by water depth and time of day. The game runs directly on WebGPU and WGSL using its own rendering engine rather than a framework. Its systems include a four-cascade FFT ocean based on Tessendorf spectra, shallow-water swash simulation, boat wakes, underwater lighting and caustics, a physically based atmosphere, volumetric clouds, dynamic lighting, positional audio, and browser-saved progress. It requires a recent browser with WebGPU and a capable GPU; the first load compiles several hundred shaders. The source code is released under the MIT license and is deployed as a static build through GitHub Pages.
Rooms is a macOS menu-bar window-management app that saves each project as a named room containing selected windows and their layout. Pressing ⌥Space and entering a room name restores the windows on the current screen while hiding or parking unrelated windows; rooms can also be assigned direct keyboard shortcuts. It offers Focus, Columns, Grid, custom grid, and Stack layouts, measures application minimum sizes, and remembers separate arrangements for laptop and external displays. Rooms stores room and parked-window data locally on the Mac and makes no network connections. It requires Accessibility permission, supports macOS 14 or later on Apple Silicon and Intel, and is distributed under the MIT license.
Phantomat is a Hyprland plugin that replaces workspaces with a zoomable, infinite 2D canvas for windows. Windows can be placed freely, while linked monitors show adjacent parts of the same canvas and move together. It supports Wayland and X11 applications, including games running under Wine and Proton. The plugin provides a zoomed-out navigation view with title- and app-name search, camera movement to matching windows, recent-window switching, panning, zooming, window movement and resizing, minimap navigation, grid arrangement, undo and redo, fullscreen and screen-fill modes, and optional pinned windows. It remembers window positions and camera state across restarts and includes a live tuner for lens distortion, blur, grid, parallax, HUD, colors, and related settings. Phantomat is distributed as source code under the BSD 3-Clause license. It requires Hyprland 0.56 or newer, Lua configuration support, GCC 15 or newer, and Hyprland development files; the plugin must be rebuilt after a Hyprland update. It was formerly called Spatial Overview, so its configuration files and commands retain that name. The project began as a fork of hyprland-scroll-overview and is described as less tested outside its Omarchy, Hyprland, NVIDIA, and two-monitor development setup.
omnibin is a Linux tool that exposes binaries from Nixpkgs on the PATH without installing all of their packages in advance. It mounts a FUSE-backed filesystem over the Nix store and uses package metadata to resolve executable names and versioned forms; package files are fetched lazily from cache.nixos.org when they are read, then served locally. The omnibin command can locate the newest or all available versions of a binary, including historical Nixpkgs versions. It can run in a shell, as a NixOS module, or in a container; the container requires /dev/fuse and the SYS_ADMIN capability. The project is distributed under the MIT license.
changelog.earth is an environmental news project that presents selected reporting as planetary patch notes. It collects stories from RSS feeds, GDELT, and Spaceflight News, validates and deduplicates them, then uses Groq to select stories and write game-style patch titles while retaining the original headline, publisher summary, publication details, and source link. Scheduled archive jobs save editions that the homepage, API, and RSS feed serve; the RSS feed provides the latest published updates with stable IDs, and saved editions remain available when collection or upstream services fail. The project is an MIT-licensed open-source application built with React, TypeScript, Tailwind CSS, and Next.js, with deployment support for Vercel and Cloudflare.
Flux is an open-source device-integration tool that connects an Omarchy desktop to an Android phone or Mac over a local network, or through Tailscale when away from home. It transfers files, clipboard text, and links; reads phone notifications; sends SMS; controls media; synchronizes Do Not Disturb; runs desktop commands; and uses the phone as a camera or microphone. It can also mirror the phone screen, support fingerprint-based sudo approval, and expose coding-agent output on the phone. The project includes a command-line interface, background daemon, native Qt desktop window, Omarchy shell plugin, Android app, and macOS app. The desktop initiates network connections, so the default Omarchy firewall does not require an additional inbound rule. Flux for Android requires Android 10 or later, while the macOS app requires macOS 14 or later.
ritridata is a Rust command-line tool for read-only inspection of disk images and synthetic image generation. It reads MBR and GPT partition metadata, checks GPT header and entry-array CRCs, reports bounds and overlapping partitions, reads sectors and byte ranges, produces hex dumps and hashes, and exposes public file-signature definitions through ID lookup. Its generator creates blank, MBR, and GPT test images, including synthetic invalid-CRC and truncated fixtures, with JSON output for structured commands. The tool operates offline and accepts regular .img, .raw, and .dd files; input images are opened read-only, while generated files use an exclusive new-file operation. It does not scan or carve files, parse filesystems, recover deleted data, repair images, or substitute GPT backup entries for an invalid primary. The repository is licensed under AGPL-3.0-only.