APEX is an open, verification-first RTL design for an LLM-inference tile developed with Sigmantic AI. It implements one transformer decoder layer—including attention, softmax, RMSNorm, RoPE, SwiGLU, residual processing, and KV-cache compression—and runs Qwen2.5-0.5B through the verified pipeline on FPGA hardware. The KV-cache codec sits in the datapath: cached keys and values are compressed as they are produced and decompressed when consumed, while an importance unit allocates precision according to which parts of the context matter. Each RTL block is checked against an executable NumPy golden model with bit-exact comparisons and mutation-tested testbenches. The repository covers the inference tile rather than a complete system; DRAM control, PCIe, and NoC are out of scope.
barehands is an open-source, webcam-powered hand-tracked interface for AI assistants. It runs in Chrome with a webcam and overlays notes, images, and 3D models as cards over the camera feed; users manipulate them with gestures such as pinching, throwing, stretching, force-pulling, and two-finger dragging. The interface uses a local Python server and loads Google MediaPipe for hand tracking and three.js for 3D rendering. Markdown folders, including Obsidian vaults, can serve as note sources, while configured media folders provide images, transparent props, and 3D models. An AI assistant can control the interface through local files and shell scripts: state files drive the on-screen ring's status, and the board protocol presents content on the stage. The repository states that it can be used, modified, and built on commercially within a business, with redistributed versions remaining under the same license.
CapRover is an open-source, self-hosted platform as a service and web-server manager for deploying applications, databases, and websites on a user's own server. It provides a web GUI and CLI for managing deployments and supports applications built with Node.js, Python, PHP, ASP.NET, Ruby, and Go, along with databases and services such as MariaDB, MySQL, MongoDB, PostgreSQL, and WordPress. Under the interface, CapRover uses Docker and Docker Swarm for containerization and clustering, Nginx for customizable load balancing, and Let's Encrypt for SSL certificate management. It is designed to handle deployment and server-management tasks without requiring direct Docker or Nginx configuration, and the project states that applications continue working if CapRover is removed.
Career Ops is an open-source AI job-search system that runs locally through compatible AI coding CLIs. It scans Greenhouse, Ashby, Lever, and company career pages; compares listings with a CV and evaluates them in an A–H report with a global 1–5 score. The system uses Playwright to navigate career pages, can process multiple offers with sub-agents, generates ATS-oriented CV PDFs tailored to individual job descriptions, records applications in a tracker, and researches companies and potential contacts. Its separate Block G assesses posting legitimacy, including scam or ghost-job risk, while Block H drafts additional material only for highly scored roles. It is intended to filter opportunities rather than submit applications, which users review and submit themselves.
DesktopFly is a macOS desktop companion that places a procedural 3D fruit fly on a transparent, click-through overlay. The fly walks across window edges, grooms, sleeps, flies, and responds to cursor movement, clicks, window changes, keyboard activity, time of day, and thermal state without requiring permissions or entitlements. A separate interactive brain window displays 23,210 FlyWire neuron soma positions and live spikes. Its behavior is driven by a 1 kHz leaky-integrate-and-fire simulation of a 668-neuron circuit containing approximately 19,000 real synaptic connections. Cursor approach is converted into looming input for LC4 and LPLC2 visual neurons; escape is produced only when activity reaches the DNp01/Giant Fiber escape command neuron through the modeled synapses. Other circuit outputs drive walking speed, steering, grooming, backward movement, wing effort, escape maneuvers, arousal, and spontaneous takeoff. The loop also feeds gait rhythm to ascending proprioceptive neurons and fast cursor motion to sensory wind partners. The body geometry and sensory transduction are procedural or modeled, while the downstream connectivity and synaptic strengths come from FlyWire data. The project targets macOS 13 or later and requires the Xcode Command Line Tools. Menu-bar controls provide pausing, brain-window visibility, escape testing, movement between displays, adding or removing flies, and starting a collective startle response; the brain window can stimulate nearby circuit neurons for 400 ms. The code is MIT-licensed, while derived FlyWire data files are CC BY-NC 4.0.
Endoplexity is a Chrome MV3 side panel and local bridge for agentic control of the browser a person is already using. It lets Claude or Cursor agent CLIs operate existing tabs and logged-in sessions by clicking, typing, filling forms, navigating, switching tabs, uploading files, and reading pages, without sending requests directly to a model or requiring a separate metered API key. The side panel owns the Chrome DevTools Protocol connection and communicates over an origin-pinned WebSocket with a bridge on localhost; the bridge exposes browser tools over token-gated MCP/HTTP and enforces the safety policy, including approval gates and autonomy modes. Pages are provided to the agent as accessibility-tree snapshots rather than raw HTML, and actions return the resulting page state so subsequent turns can resend only changed lines. Its file-reading tool extracts text from PDFs, DOCX, XLSX, PPTX, CSV, JSON, Markdown, and other text files.
fx is an open-source coding-agent harness and command-line interface written in Zig by Vercel Labs. It is designed as a compact Unix-like alternative to a terminal IDE, with interactive and one-shot requests for inspecting and modifying repository code, running shell commands, and saving or resuming sessions. The agent supports skills, MCP tools, plugins, subagents, permission rules, and headless requests, and can be embedded natively or through WebAssembly. It is model-agnostic, supports local and cloud inference, is distributed under the Apache-2.0 license, and is marked experimental by its repository.
icm-architect is a Claude skill that designs processes, ideas, and problems as ICM (Interpretable Context Methodology) workspaces, using folder structure as agent architecture, or restructures an existing folder, repository, or vault. Numbered folders encode sequencing, hierarchy scopes context, and plain Markdown files store state, allowing an agent to orient and act by reading the relevant files. In build mode, it extracts stages, human approval points, and stable versus per-run elements, then scaffolds a workspace using one of six forms: Pipeline, Umbrella, Record library, Knowledge bundle, Context map, or System map. In restructure mode, it audits files as catalog, contract, factory, product, or dead, proposes a migration map for approval, then migrates and validates the result. The walk test checks whether an agent with no prior memory can orient, act, and report status from the workspace alone. It can be installed as a Claude Code skill or uploaded to Claude applications, and is MIT licensed.
IP as Logo Skill is a compact Agent Skill in the open Agent Skills format that guides compatible AI agents in turning a product brief into highly simplified, rounded mascot-character concepts. It gathers product context, proposes three design directions, and after approval generates six independent full-resolution square candidates using one dominant silhouette of roughly four to seven large shapes, two mascot or IP colors, a named solid background color, thick rounded forms, and lower-left or lower-right corner emergence. Its default batch uses two variants per direction with a three-left, three-right composition split rather than a contact sheet. Familiar animals are the default subjects, while objects, machines, fantasy artifacts, and other unusual subjects require a clear product-related reason. The skill provides instructions and prompts rather than an image generator, can be installed with the Agent Skills CLI, and is intended for agents such as Codex, Coze, Doubao, YouMind, Manus, Gemini Apps, and Replit Agent when paired with a supported image model.
J-Space Cognition Suite is a model-agnostic, inference-time control system packaged as a cross-platform AI-agent Skill for deep reasoning, long-horizon work, tool use, verification, and recovery. It leaves model weights and training unchanged, organizing an agent's working representations through one entry point and nine selectively loaded modules supported by four references. Its fast, full, and loop operating modes use selective workspace loading, a shared broadcast hub for constraints and values, dense internal reasoning traces, explicit intermediate steps before conclusions, metacognitive routing of confidence and failure signals, and bounded empirical verification. An optional standard-library controller records durable task state, including goals, next actions, checkpoints, open questions, seams, and recovery state. The suite is intended for low-friction integration with AI hosts that support Skills or equivalent system/developer-instruction mechanisms.
Microsandbox is a local-first microVM runtime and library for running untrusted workloads, including AI-agent tasks, user code, plugins, CI jobs, development environments, scrapers, and automation. It uses hardware-level isolation and runs standard OCI container images from registries such as Docker Hub and GHCR, with Docker-like image, command, shell, and volume workflows. The project provides SDKs for TypeScript, Rust, Python, Go, and Ruby, along with a CLI for booting and controlling sandboxes. Applications can create microVMs as child processes without a setup server or long-running daemon; sandboxes can also run detached for long-lived sessions. Agent Skills and an MCP server support agent-created sandboxes, and the project documents secret handling intended to keep keys out of the VM. It runs on Linux, macOS, and Windows, with platform-specific virtualization requirements, and is described as beta software.
MoneyPrinterTurbo is an AI short-video generation tool that turns a topic or keyword into a high-definition video. Its automated workflow generates or accepts a script, extracts search keywords, matches footage from local materials or supported stock sources, synthesizes narration, creates configurable subtitles, adds background music, and renders the result. It provides WebUI, API, CLI, and AI-agent interfaces; supports batch generation, multiple languages, portrait and landscape formats, adjustable clip durations, and multiple text-to-speech and large-language-model providers. It can also generate visual material through supported text-to-video models and automatically publish completed videos to TikTok, Instagram, and YouTube Shorts.
OmniRoute is an open-source, self-hosted AI gateway developed by diegosouzapw. It exposes model providers through a single OpenAI-compatible endpoint and routes requests according to provider availability, quotas, cost, and latency, with automatic fallback and model chaining. The gateway is designed for local-first operation, keeping provider keys locally and supporting encrypted key storage. Its documented features include quota-aware scheduling, context compression through the RTK and Caveman mechanisms, and integrations for MCP and A2A workflows. The repository describes compatibility with coding clients such as Claude Code, Codex, Cursor, OpenCode, Cline, and Copilot, along with desktop and progressive web app interfaces. It is distributed under the MIT license and supports hundreds of providers and more than a thousand model identifiers.
Remocn is a copy-paste component library for building videos with Remotion. It provides animations, transitions, backgrounds, scenes, fades, wipes, and kinetic titles as Remotion code that developers add to their projects through the shadcn component registry and can modify directly. Its components use Remotion's `useCurrentFrame()`, `interpolate()`, and `spring()` APIs. Component pages provide live previews using `@remotion/player`, allowing frame-by-frame scrubbing. Remotion is required as a prerequisite, and the project also provides an optional agent skill for setting up and working on Remotion video projects.
Semantica is an open-source Python infrastructure layer for building context graphs and knowledge graphs for AI systems. It ingests enterprise and other multi-source data, extracts entities and relationships, flags conflicting facts, merges duplicates, and supports ontology management, knowledge modeling, graph analytics, and causal reasoning. The system records provenance and execution trails for the context supplied to an AI system and the decisions produced from it, making those relationships and decisions queryable. Its repository describes deterministic graph construction, reasoning, and provenance that do not require an LLM; it explains the data and policies outside an LLM rather than exposing the model's internal reasoning. Semantica supports RDF and labeled-property-graph storage, W3C standards, self-hosted deployment, and installation with pip.
TencentDB Agent Memory is a self-hostable, team-level memory hub for AI agents. It extracts conversations and tasks into reusable Chat Memory and Skills, and converts documents and code into an LLM Wiki and CodeGraph. These assets can be reviewed, versioned, governed, shared, routed, and reused across agents, frameworks, and team members; existing documents, codebases, and agent sessions can also be imported to reduce cold-start work. The system provides a shared memory server and a proxy that preserves the agent protocol, so supported clients can use the same memory by pointing their base URL at the proxy without plugins, hooks, or an MCP server. The repository deploys memory-core, memory-hub, and the proxy together, with a local management panel and configuration for separate memory and proxy LLM parameters. Documented integrations include DeepSeek Harness, Claude Code, Codex, CodeBuddy, WorkBuddy, Hermes, and OpenClaw.
TrueForge is an open-source, vendor-neutral agent harness developed by TrueFoundry. It runs the agent execution loop, including model calls, MCP tool use, skills, sandboxed code and file execution, approvals, context management, and session state. Agents can use configured model providers, remote MCP servers, git-backed SKILL.md instruction packs, and on-demand sandboxes; context features include subagents, deferred tool loading, Code Mode, large-result offloading, and compaction. TrueForge exposes agents through a bundled chat UI, an HTTP API with a TypeScript SDK, and an embeddable UI SDK. It supports local mode with SQLite and hosted deployments using Postgres and Redis through Docker Compose or Helm; the repository warns that local mode is intended for personal use on localhost rather than production or internet-facing deployment.
Viscose is a portfolio carousel rendered as a single full-screen WebGL fragment shader rather than a grid of DOM elements. Project cards travel around a mostly off-screen ring that can be turned by scrolling, dragging, swiping, or clicking; momentum carries the ring and it snaps the nearest card to the front. Signed distance fields make neighboring cards melt together as they approach and form thin, sagging threads as they separate, while pointer interaction softens the field, shifts nearby cards, and creates connecting effects. The same shader pass renders the threads, the refracting glass-like edges at the top and bottom of the viewport, and the cursor tag. On touch devices, swiping turns the ring and press-and-hold triggers the hover effects. The project includes a development-only lil-gui panel with roughly 136 live parameters and is built with Next.js, React, Three.js, GSAP, and Tailwind CSS; it requires Node.js 20 or newer to run locally.
World Monitor is an open-source global intelligence dashboard developed in the koala73/worldmonitor project. It aggregates curated global and regional news, geopolitical information, infrastructure signals, and financial data into a unified situational-awareness interface, producing AI-synthesized briefs and tracking military, economic, disaster, and escalation signals. The application provides both a 3D globe and a WebGL flat map using a shared map-layer catalog, a server-authoritative Country Instability Index, and finance views for stock exchanges, commodities, crypto, and composite market signals. It supports local AI through Ollama without API keys, as well as other documented AI providers, and exposes an MCP server for programmatic access by agents and scripts. The repository supports multiple site variants from one codebase, including world, tech, finance, commodity, energy, and happy views. It also distributes a Tauri desktop application for macOS, Windows, and Linux, and documents self-hosting through Vercel, Docker, or static deployment.
Zeron is a local control layer for coding agents, including Claude Code, Codex, Cursor, Grok, Hermes, and Pi. Each device runs a small engine that stores agent sessions locally by default, without requiring an account or network connection. Optional synchronization lets users start an agent on one device and follow or control it from another; an always-on machine such as a VPS can keep agents running after a laptop is closed. The Linux installer starts a daemon that persists across reboots, while macOS users can install the desktop release or build from source. Zeron is licensed under the MIT License. The videos refer to the project as Comet.
Searchable transcript of Trending Open-Source GitHub Projects : MoneyPrinterTurbo, OmniRoute, CapRover & DesktopFly #285 — ManuAGI - AutoGPT Tutorials (17:38). Search for a phrase, then click its timestamp to jump straight to that moment in the video.
Captions sourced from the original video on YouTube, published by ManuAGI - AutoGPT Tutorials. The video, its captions and all related intellectual property remain the property of their respective owners; AINotes claims no ownership. Provided for research, accessibility and search — see the Transcript Notice and Copyright Policy.
00:00 Every week, developers release powerful open-source GitHub projects that make coding, building, and shipping software faster. This is your weekly GitHub project update video covering top trending open-source GitHub projects this week. From AI coding agents to local first tools and self-hosted platforms, you'll discover useful and trending developer tools worth adding to your stack.
00:20 Without wasting time, let's get started. >> Before we jump into today's project updates, here's a quick announcement for everyone. We've launched a brand new YouTube channel called AI Agent Studio dedicated entirely to AI Agent projects, tutorials, and tools. So, if you're interested in staying up to date with the latest AI Agent open source projects, learning how to build your own agents, or exploring cuttingedge agent frameworks, make sure to check it out.
00:49 Subscribe now to get weekly videos, in-depth guides, and realtime project breakdowns. The link is right there in the description. Don't miss it. All right, let's get into today's video. >> Project number one, Money Printer Turbo. Generate HD short videos from a topic. Money printer Turbo is an open-source tool that generates HD short videos from a topic or keyword.
01:11 Give it a subject and its AI workflow writes the narration, matches footage from Pixels, Pixabay or Coverer, adds voice over, subtitles and background music, then renders a portrait or landscape clip. It supports voice engines and LLM providers, batch generation, multiple languages, and one-click publishing to Tik Tok and Instagram. Use it via web UI, API, CLI, or an AI agent.
01:34 Install it and generate your video. Project number two, Career Ops. Turn your AI coding CLI into a job search center. Career Ops is an open-source AI job search system that turns any AI coding CLI into a command center. Paste a job link and it evaluates it against your CV with an A tof rubric into a 1:5 score. Flags scams and ghost jobs, then drafts a tailored CV, cover letter, and application email.
02:05 It never submits. It scans portals, tracks your pipeline, and researches companies and contacts. It runs locally in cloud code, codeex or open code. Install it and run your search. Project number three, Miss Semantica, graph infrastructure for accountable, auditable AI agents. Semantica is an open- source Python infrastructure layer that gives AI agents a knowledge graph with audit trails.
02:29 It sits under your LLM, vector store, and agent framework, ingesting data from databases, data bricks, and snowflake. Then extracting entities and relations into a context graph. Every fact carries W3C Bravo Provenence. Every decision becomes a traceable queryable node and reasoning runs deterministically through retay data log and sparkle without an LLM storage spans RDF and property graph backends.
02:56 Install it and ground your agents. Project number four micro sandbox fast local microVMs for untrusted workloads. Micro sandbox is an open-source local first microVM runtime and library for untrusted workloads. AI agents, user code, plugins, CI jobs, and scrapers in fast hardware isolated microVMs. It runs OCI container images with Docker-Like commands and embeds in your code, spawning a VM as a child process.
03:26 No server or demon. Secrets never enter the VM. SDKs ship for Rust, Python, TypeScript, Go, and Ruby, plus a CLI, and MCP server. It runs on Linux, Mac OS, and Windows. Install it and sandbox your workloads. Project number five, EC Omni Route. One endpoint for 340 AI providers. Auto fallback included. Isa Omni Route is a free open-source AI gateway that puts 340 providers and 1,200 plus models, Kimmy, Claude, GPT, Gemini, DeepSeek, and more behind one Open AI compatible endpoint.
04:05 Point any tool, clawed code, codeex, or cursor at it, and it routes requests across providers with quotaaware auto fallback, sliding to the next model when quota runs out or a provider fails. It compresses context to save tokens, includes an MCP server and A2A, and runs local first with encrypted keys. Install it and never hit a limit. Project number six, Dansent DB agent memory, team memory hub of reusable assets for agents.
04:32 Dancent DB agent memory is an open-source self-hostable memory hub for teams of AI agents. It turns conversations, documents, and code into four reusable assets. Chat memory, version skills, an LLM wiki, and a code graph. A human controlled panel governs ownership, versions, and visibility, so you equip agents with what they need and share safely. You connect an agent by pointing its base URL at the proxy.
04:57 No plug-in needed. New agents import repos, docs, and sessions. Deploy it and give your agents memory. Project number seven, World Monitor. Realtime global intelligence dashboard for situational awareness. World monitor is an open-source intelligence dashboard that pulls news, geopolitics, and infrastructure signals into one view. It aggregates 500 feeds across 15 categories and synthesizes them into AI written briefs, plots events on a 3D globe or a WebGL map.
05:28 It correlates military, economic and disaster signals, scores country instability and tracks exchanges, commodities and crypto. You run it locally with Olama and no API keys or reach it via MCP REST and CLI. Clone it and watch the world. Project number eight, JSpace Cognition Suite inference time cognitive control layer for any model. JSpace Cognition Suite is an open-source model agnostic control layer for reasoning, long horizon tasks, tool use, and verification.
06:00 Packaged as an agent skill, it adds no weights, fine-tuning, or hidden service. Text alone turns a model's accessible working representations into a managed workspace. A gate picks fast, full, or loop mode, then applies protocols like selective loading, a broadcast hub, metacognitive control, and empirical verification. An optional standard library controller externalizes loop state so work resumes across seams.
06:28 Install it and use /jspace. Project number nine. Veasif x a tiny native coding agent for terminals. Vasif is a command line coding agent built by Verso Labs. Written in Zigg as a single native binary. It runs inside a repository, inspects and changes code, executes shell commands, and continues save sessions. It works as an MCP client and supports skills, plugins, sub aents, and headless requests through FX ask.
06:54 It stays local, avoids product telemetry, and works with different inference providers. It suits developers who want a lightweight embeddible coding agent for terminals and automated workflows. Explore FX directly in your terminal. Project number 10, Caprover self-hosted deployment platform without Docker expertise. Caprover is a self-hosted platform as a service that deploys and manages applications and databases on a private server.
07:21 It supports NodeJS, Python, PHP, ASP, Net, Ruby, and databases like Mariab, MySQL, MongoDB, and Postgress. Under the hood, it runs Docker Swarm for containerization, EngineX for load balancing, and let's encrypt for free SSL certificates. All controlled through a simple web interface or a command line tool for scripting. Removing Caprover leaves deployed apps running since there is no lockin.
07:48 It suits developers who want server control without manual Docker or EngineX configuration, cutting hosting costs compared to manage platforms. Set up a server and deploy your first app. Project number 11. Remoten copy paste animation components for remote videos. Remokin is a component library for building videos with Remotion. Following a copypaste model similar to Shad Gumontim, instead of coding fades, wipes, and kinetic titles from scratch, developers run a shadum ad command to pull a ready-made animation into
08:20 their project and own the resulting code directly. Every component uses remotions frame interpolate and spring functions correctly, and each one has a live preview through the remotion player for scrubbing frame by frame. It also connects with AI coding agents like Claude Code through an installable skill that sets up a project and opens a preview studio.
08:41 It suits solo builders and small teams who need a polished product demo video without writing animation code from scratch. Explore the component registry and add one to your project. Project number 12, Comet. Control coding agents from any device. Comet is a tool that lets developers control coding agents like Claude Code and Codeex from any of their devices.
09:01 A small engine runs on each device and keeps sessions synced so a developer can start an agent on one machine, then follow and drive that same session from another. Installing the engine as a Damon on an always machine such as a VPS or spare box keeps agents running after a laptop closes. It includes a command line tool with a terminal interface that attaches to the Damon for status checks, updates, and session control.
09:27 It suits developers who run coding agents on multiple machines and want continuity between them. Install the Damon and connect your devices. Project number 13. Why ICM architect? Folder structures that give AI agents context. ICM architect is a clawed skill that turns any process, idea, or problem into a structured workspace of folders and markdown files.
09:50 Treating folder structure itself as the agents architecture. It follows a method called interpretable context methodology where numbered folders carry sequencing, hierarchy carries context scoping, and plain files carry state. So, one model reading the right files can do work that would otherwise need a multi- aent setup. It offers a build mode that scaffolds a fresh workspace from six proven forms and a restructure mode that audits an existing folder and migrates it.
10:18 It installs into clawed code or clawed apps as a skill. It suits developers and teams organizing AI agent workflows who want a system a human can also read directly. Try it on your own project folder. Project number 14. Bare hands. Control your AI's screen with hand gestures. Bare hands is a webcam powered tool that turns hand movements into a screen interface for an AI assistant.
10:41 Notes, images, and 3D models float over the camera feed as glass cards that a person can pinch, throw, stretch, or pull apart using bare hands with no headset or controllers. It runs from a local Python server using Google Media Pipe for handtracking and 3D JS for 3D rendering, both loaded from public CDNs. Any folder of markdown files, including an Obsidian Vault, can serve as a note source on the board.
11:06 It connects to an AI like Claude through simple file-based and local commands, letting the assistant change an on-screen status ring or add cards to the board. It suits developers experimenting with physical gesture-based interfaces for AI assistants. Clone the repo and wave at your webcam to try it. Project number 15, DHUS IPS logo. A skill for consistent AI made mascot logos.
11:29 Dehus logo is a compact agent skill that guides an AI agent to generate simplified mascot style logos rather than full character illustrations. It enforces a strict visual formula, one dominant silhouette built from around six to 10 basic shapes, one or two colors on a separate solid background, thick rounded forms with no sharp details, a lower corner crop, and only light tonal shading.
11:55 It follows the open agent skills format, so it works with any compatible AI agent rather than one specific product. Developers install it by copying its single instruction file into a project skills folder, then ask their agent for a mascot logo in plain language. The skill also defines rejection rules, so overly complex or overly flat results get flagged instead of quietly accepted.
12:18 It suits designers and developers who want repeatable brand ready mascot logos from AI image generation. Try it on your next brand mascot. Project number 16, Endoplexity browser automation using subscriptions you already pay for. Endexity is a Chrome extension that lets a coding assistant CLI act on the web page a person is actually viewing, clicking, typing, filling forms, and moving between tabs.
12:44 It runs through an existing clawed or cursor subscription instead of a metered API key using a local bridge that connects the browser side panel to the CLI over MCP. Pages reach the model as an accessibility tree snapshot rather than raw HTML which keeps data light and each action returns the resulting page so the agent can act and reread in one turn.
13:05 Irreversible actions like submitting or purchasing paws for human approval and three autonomy modes control how much the agent can do unsupervised. Everything runs locally on the loop back address keeping browser sessions and files on the person's own machine. It suits developers who already pay for claude or cursor and want an agent that can operate their real loggedin browser.
13:25 Set it up and hand it your first browsing task. Project number 17, True Forge, an open vendor neutral runtime for AI agents. True Forge is an open-source agent harness from True Foundry that runs the execution loop behind an AI agent, model calls, MCP tool use, skills, sandboxing, approvals, context management, and session state. It works with OpenAI, Anthropic, Google Gemini, and other model providers, or any OpenAI compatible endpoint, so a team can swap models without rebuilding an agent.
14:01 Developers connect models, MCP servers, and skills through ship cataloges, then build agents through a chat interface, an HTTP API with a TypeScript SDK or an embedded UI. It runs locally with a single process and equalite for testing or in a shared deployment using Docker Compose or Kubernetes with Postgress and Reddus. Skills load as getbacked instruction packs inside a sandbox only when needed and human approval gates cover sensitive actions.
14:28 It suits enterprise teams who want agents that run on infrastructure and models they control rather than a single vendor stack. Explore the quick start and connect your first model and tools. Project number 18, Apex inference chip. Open hardware that compresses an LLM's memory on chip. Apex is an open chip designed for running large language model inference built as a transformer decoder layer in real hardware description language rather than software.
14:55 Instead of storing conversation memory in full precision, it compresses the key value cache directly inside its data path, quantizing keys and values as they're written and decompressing them as they're read, so reading speed stays flat as context grows. An importance tracking unit decides how much precision each cached region deserves, spending more bits on parts of the context that matter most.
15:18 Every hardware block is checked bit forbit against a numpy reference model and the design runs QN2.5-0.5B on real FPGA hardware with larger QN2.5-7B tokens verified through the same software pipeline. It suits chip designers and AI hardware researchers interested in memory efficient verifiable inference hardware rather than a finished consumer product.
15:44 Explore the repository and reproduce the verification results yourself. Project number 19, Desktop Fly. A real fly brain simulation living on your Mac. Desktop Fly is a Mac OS desktop app that places a 3D Fruit Fly on your screen controlled by a live spiking simulation built from the real Flywire Fly Connect. It runs a leaky integrate and fire circuit of real neurons and synapses.
16:07 So behavior like walking, grooming, sleeping, and fleeing the cursor comes from actual neural activity, not scripted animation. It reads permissionfree Mac OS signals like cursor position, window movement, clicks, typing rhythm, and system temperature and feeds them into the circuit as sensory input. A separate window shows the live brain with neurons flashing as they fire, and you can click regions to stimulate them directly.
16:32 Everything runs locally on the machine with no cloud processing involved. This suits developers and neuroscience enthusiasts curious about connect-driven behavior. Clone the repository and build it to watch a real fly brain drive a desktop companion. Project number 20. Visose. A portfolio carousel rendered entirely as one shader. Visose is a portfolio carousel built with Nex.js, React, 3JS, and GSAP.
16:56 Rendered as a single WebGL fragment shader. It solves the problem of showing project cards in a way that feels fluid instead of static using signed distance fields. So neighboring cards melt together as they close up and stretch into thin threads as they pull apart. Scrolling, dragging, or swiping turns the ring, and hovering softens the surface around a card and webs threads between it and its neighbors.
17:20 Everything runs client side in the browser with a dev panel exposing every tunable setting for live adjustment. This suits front-end developers exploring shader-driven interfaces for showcasing creative work. Clone the project and turn the ring yourself. Thanks for watching. See you in the next update.