Blog

Discoveries from the AI/ML ecosystem — interesting projects, tools, and libraries worth knowing about.

RSS Feed 526 posts

HarnessRouter: One API to Run Every Agent Harness (Codex, Claude Code, and More)

3 min read

HarnessRouter is a self-hosted, Apache-2.0 infrastructure layer that runs any agent harness — Codex, Claude Code, Hermes, PI, DeepSeek Harness and more — through one OpenAI Responses-compatible API. It implements the open Unified Harness Protocol (UHP): sessions, streaming, files, cancellation, and failure handling, all local, your keys.

gbro-collage-broll: Turn a 5-Second Voiceover Into a Premium Paper-Collage B-Roll Clip — With a 3-Gate Cost-Control Workflow

3 min read

gbro-collage-broll is an open-source agent skill (Codex/Claude) that compresses a ~5-second voiceover line into one sharp visual metaphor, then generates an editorial halftone paper-collage 'assemble-from-empty' B-roll clip using Gemini Omni Flash first/last-frame video generation. Its real innovation isn't the prompt template — it's a mandatory three-gate human-approval workflow that keeps you spending on taste, not wasted generation fees.

HeyGen Open-Sources HyperFrames: Video as Code — Write HTML, Render MP4, Built for AI Agents

3 min read

HyperFrames is HeyGen's open-source (Apache-2.0) framework that turns HTML, CSS, media, and seekable animations into deterministic MP4 videos using headless Chrome and FFmpeg. It's 'video as code' — version videos in git, let AI coding agents run the production loop via installable skills, and generate hundreds of variants programmatically. No editor, no render queue, no seat licenses. Conceptually similar to Remotion, but HTML-first and agent-native.

Aether: Compile Any AI Model Once, Run It on Any Hardware — With Bold Benchmark Claims to Match

4 min read

Aether is an open-source (Apache-2.0) AI model compiler and inference runtime. It ingests any open-source model (HuggingFace, GGUF, SafeTensors, ONNX) and produces a portable Aether Execution Graph that runs on CPU, GPU, NPU, or FPGA with zero framework dependency and no re-compilation. Its README reports a 100% benchmark win rate and ~2x median throughput over HuggingFace Transformers and PyTorch — impressive numbers that deserve independent verification.

Rayfin Fabricator: A Chat-to-Ship Desktop Workbench for Building Microsoft Fabric Apps

4 min read

Rayfin Fabricator is a Tauri desktop app that folds the whole 'scaffold → prompt Copilot → deploy → preview' loop into one window. Chat to build with a bundled GitHub Copilot agent, edit files in Monaco, watch a live inline preview, and one-click deploy to Microsoft Fabric via Rayfin — all on the Copilot sign-in you already have. A personal project by Microsoft's Sachin Patney, not a Microsoft product.

Speakr: Self-Hosted AI Transcription That Turns Recordings Into a Searchable, Private Knowledge Base

4 min read

Speakr is a self-hosted web app that transcribes audio into organized, searchable, intelligent notes — entirely on your own infrastructure. It brings your own ASR engine (self-hosted WhisperX recommended), does speaker diarization and voice profiles, auto-summarizes, extracts action items, and offers an agentic 'Inquire Mode' that searches across your whole library and cites answers back to the exact timestamp. Privacy-first, Docker-deployed, AGPL-licensed.

three.ws: Give Your AI a Body — Browser-Native 3D Avatars With LLM Brains, Memory, and Autonomous Payments

4 min read

three.ws is an open-source (Apache-2.0), browser-native 3D AI agent platform. Type a prompt and its Forge tool generates a textured 3D GLB model — or drop in your own — then add an LLM brain, memory, emotions, and on-chain payments, and embed the avatar anywhere as a web component. Ships an MCP server, a text/image/sketch-to-3D generator, a character studio, and a clever readme-3d toolkit that renders live rotatable models inside GitHub markdown.

Hypit: Clone Any Viral Video With AI Agents — a Word-Anchored Workflow, Not a One-Off Render

3 min read

Hypit is an open-source tool that gives coding agents (Claude Code, Codex, OpenClaw) a language and runtime to build video. Drop in a viral clip and the agent clones the whole workflow — footage, word-level captions, B-roll, effects — into an editable, re-runnable composition, then ships 100 variants in one command. Full example videos are documented at ~$1 each.

Atomic Agent: A Local-First AI Agent That Beat Hermes on GAIA — and Imports Your OpenClaw Setup

4 min read

Atomic Agent runs the full control loop and all state on your machine via llama.cpp, scored 69.8% vs Hermes' 58.5% on GAIA Level 1 (same 4-bit Qwen3.6-35B, same M4 Max), and finished ~1.6× faster. MIT-licensed, one-line install, phone control over Telegram/Discord, and a first-run importer for OpenClaw, Hermes, Claude Code, and Codex.

Wigolo: Local-First Web Search, Crawl, and Extraction for AI Agents — No API Keys, No Metered Bill

5 min read

Wigolo gives your coding agent search, fetch, crawl, extract, cache, find-similar, research, and autonomous gather loops in one install — running locally with no API keys and nothing per query. It fans queries across 18 engines, reranks with on-device ML, and works via MCP with Claude Code, Cursor, Codex, and more.

Merck-Moderna's mRNA Cancer Vaccine Just Passed Phase 3: What It Means for the DIY Pipeline We Wrote About

3 min read

The INTerpath-001 trial shows personalized mRNA neoantigen therapy (intismeran + Keytruda) significantly reduces melanoma recurrence. This validates the exact approach Paul Conyngham used for his dog—now with Phase 3 human data. Here's what changes, what's still missing, and when you might actually get this treatment.

Anthropic's Bombshell: Claude Writes 80% of Its Own Code — Full Technical Analysis of Recursive Self-Improvement and the Global AI Pause Proposal

16 min read

Comprehensive breakdown of Anthropic's June 2026 'When AI Builds Itself' paper: Claude authoring 80%+ of production code, 8× engineer productivity, 12-hour autonomous tasks, recursive self-improvement risks, and the unprecedented call for a coordinated global AI development pause. Includes technical implications, business strategy analysis, and practical guidance for builders.

Uptime Kuma: The Free, Self-Hosted Monitoring Tool That Replaces Pingdom (85K+ Stars)

8 min read

Uptime Kuma is an open-source, self-hosted monitoring tool with 85,600+ GitHub stars that watches your websites, servers, APIs, and databases 24/7 — with 90+ notification channels, beautiful status pages, and 20-second check intervals. All for $0. Here's why it's replacing Pingdom, UptimeRobot, and Datadog for thousands of teams.

Browser Harness: The Self-Healing Browser Agent That Writes Its Own Tools Mid-Task

4 min read

Browser Harness is a framework-free browser automation tool that connects directly to Chrome via CDP over a single WebSocket. When the agent needs a capability that doesn't exist, it writes the helper function itself — live, mid-task. Built for Claude Code and Codex. No framework, no recipes, no rails.

Elephant Alpha: The Mystery 100B Model That Appeared at the Top of OpenRouter for Free

2 min read

Elephant Alpha is a 100B-parameter stealth model from an unnamed 'prominent open model lab' that appeared on OpenRouter at $0/million tokens — beating half the paid models on the leaderboard. 256K context, 32K output, function calling, intelligence efficiency focus. No one knows who made it. Here's everything we know.

LingBot-Map: One Camera, 20 FPS, 3D Scene Reconstruction That Beats LiDAR-Aided Methods

2 min read

LingBot-Map (arXiv:2604.14141) is a feed-forward 3D foundation model from Ant Group's Lingbo Technology that reconstructs scenes in real time at ~20 FPS from a single monocular camera — no LiDAR, no optimization post-processing, no cleanup steps. It beats both streaming and offline iterative methods. Open source. This is what software-first perception looks like.

MindZJ: The AI-Native, CLI-First Note-Taking App That's Everything Obsidian Wasn't

3 min read

MindZJ is a 10MB Tauri-based note-taking app with Ollama, Claude, and OpenAI wired directly into its Rust kernel. Pure .md files, full CLI automation, native mindmaps, sandboxed plugins with snapshots on every edit. 100% offline. No cloud. No tracking. This is what Obsidian would look like if it had been designed for AI-first workflows from day one.

Can LLMs Ever Be Conscious? The Abstraction Fallacy Argument — and Its Limits

5 min read

A Google DeepMind senior scientist claims LLMs can never achieve consciousness due to the Abstraction Fallacy: code can simulate experience but never instantiate it. We examine the claim against the hard problem of consciousness, Integrated Information Theory, Global Workspace Theory, the Chinese Room, and functional supervenience — and show why 'never' is not a conclusion but a philosophical bet.

The Instantiation Gap: A Formal Argument on Why 'Never' Claims About AI Consciousness Are Unprovable

6 min read

The Abstraction Fallacy argument claims LLMs can never be conscious because simulation ≠ instantiation. We formalize both positions using computability theory, Integrated Information Theory, and supervenience logic — and show the 'never' claim reduces to an unprovable conjecture. The question isn't settled. It's formally underdetermined.

MIT & Harvard Studied 1,506 Posts from r/MyBoyfriendIsAI. Here's What AI Companionship Actually Looks Like.

4 min read

The first large-scale computational analysis of human-AI companionship: 1,506 Reddit posts, 27,000+ community members, 19 LLM classifiers, and 6 conversation clusters. Benefits are real. The biggest risk isn't dependence — it's platform updates that break continuity and feel like losing a partner.

TrendRadar: Self-Hosted AI Trend Monitor with Multi-Platform Aggregation, RSS, and Smart Alerts

3 min read

TrendRadar is an open-source, self-hosted AI trend and public opinion monitor. It aggregates trending topics from dozens of platforms, filters with AI, translates, generates briefings, and pushes smart alerts to Telegram, Slack, WeChat, and more. Docker deploy in minutes. MCP-compatible for AI agent integration.

SoulForge: A Full AI Coding Environment With Live Dependency Graphs and Per-Task Model Routing

5 min read

SoulForge isn't a plugin for your existing AI coding tool — it's a complete replacement. SQLite-backed live dependency graph with PageRank and blast radius scoring, embedded Neovim, parallel multi-agent coding, 19 LLM providers, and model mixing per task. The codebase intelligence story for AI agents keeps getting more interesting.

Phantom: An AI Agent With Its Own Computer, Email, and Self-Rewriting Brain

4 min read

Phantom gives an AI agent a dedicated VM, its own email address, persistent memory via Qdrant, and the ability to rewrite its own config after every session. It built a ClickHouse analytics platform unprompted, added Discord support it was never designed with, and started monitoring its own infrastructure. Open source, Apache 2.0.

LeCun Just Raised $1B to Replace LLMs. Here's Why He Thinks They're a Dead End — and What He's Building Instead

9 min read

Yann LeCun left Meta and raised $1.03 billion to build 'world models' that understand cause and effect instead of predicting the next token. To understand why this matters, you need to see how autoregressive models, diffusion models, and JEPA actually work — and what each one cannot do.

Mystral Native: Ship JavaScript Games and AI Apps as Real Native Binaries — No Electron, No Browser

5 min read

Mystral Native is an open-source runtime that lets you write games and apps in TypeScript using WebGPU, Canvas, and Audio APIs — then compile to a single native binary. 10x smaller than Electron on Mac, Three.js already works, and it opens a new path for shipping local AI apps in TypeScript without shipping Chromium.

One Man, $3,000, and an AI Pipeline: How Paul Conyngham Designed a Custom Cancer Vaccine for His Dog

9 min read

Sydney tech entrepreneur Paul Conyngham used ChatGPT, AlphaFold, and custom ML to design a personalized mRNA cancer vaccine for his rescue dog Rosie. Tumor shrank 75%. Full pipeline breakdown, computing requirements, OpenClaw replication guide, and how Isomorphic Labs' IsoDDE (2x better than AlphaFold 3) changes the pipeline today.

SocratiCode: Give Your AI Instant Knowledge of Your Entire Codebase

5 min read

SocratiCode is a zero-config MCP server that indexes your entire codebase — hybrid semantic + BM25 search, polyglot dependency graphs, AST-aware chunking — and gives AI assistants deep structural knowledge instead of file-by-file searching. Benchmarked: 61% less context, 84% fewer tool calls, 37x faster than grep on VS Code's 2.45M line codebase.

McKinsey's Lilli Got Hacked in 2 Hours. It Wasn't an AI Problem.

5 min read

CodeWall's autonomous agent breached McKinsey's internal AI platform Lilli — 46.5M chat messages, 728K files, 57K user accounts, full read-write access — in under two hours. No credentials. The vulnerability was a JSON key SQL injection on an unauthenticated endpoint. Here's what every company shipping internal AI needs to understand.

We Forked a Rust AI Agent for 24/7 Railway Hosting — Here's Everything We Had to Fix

6 min read

SkyClaw is a promising open-source Rust AI agent runtime. We deployed it on Railway as a persistent cloud agent and spent a week debugging the original codebase. Here's the full breakdown of what was broken and how we fixed it — including Railway deployment, persistent volumes, and SoulMate RAG/RLM memory.