Alias ArchiveArchive in progress
Archive in progress

Posts

Published notes

Work logs, tutorials, and processes that occasionally need more room.

CONTENT
P

All48 posts

Most recently active first
O18

The Photoreal Lip-Sync Wall: Testing and Abandoning wav2lip, MuseTalk, and LatentSync

A complete test record of three open-source lip-sync routes on a 16GB consumer GPU: measurement methods, hidden failures, and the evidence behind abandoning real-time lip sync.

Abstract

An evaluation of local real-time lip sync for a photoreal digital companion. wav2lip256 and MuseTalk v15 were measured on an RTX 5070 Ti with 16GB of VRAM for frame rate, latency, memory use, and mouth motion; LatentSync was tested offline. The record covers seven quiet failures, including memory fallback, reversed coordinate conventions, and duplicated-frame stutter. The conclusion is that photoreal real-time expression and lip sync currently require either far more than consumer-grade VRAM or a qualitative compromise in image quality or latency. The project therefore abandoned live lip sync in favour of prerendered clips and an independent voice layer. All figures apply only to the tested versions and machine.

Read post ↗
O17

Treating the Browser as a Display: A Host-Agnostic Stage and Three Infrastructure Battles

An architecture that lets one AI companion inhabit a browser tab, a standalone window, and a desktop wallpaper: move all logic into the server and reduce the page to a display.

Read post ↗
O16

A Homegrown Brain: Four Memory Layers, Fast and Slow Models, and a Catchphrase War

From borrowing an open-source dialogue engine to a fully independent brain: one relationship timeline, four memory layers, a self-reinforcing style infection, and latency-based model roles.

Read post ↗
O15

Building a Digital Companion: A Journal and Reading Map

Two weeks after starting from an open-source desktop companion, the result was a wholly self-owned system. This is both the chronological account and the index to five technical records.

Read post ↗
O14

A Selfie System That Chooses Its Own Outfit: Controlled Randomness and Image–Caption Consistency

Designing an AI companion that sends selfies in tune with a conversation: a layered outfit engine, explicit image–caption constraints, restrained proactive sharing, and two image-to-image API failures that returned success while ignoring the reference.

Read post ↗
O13

Valves and Redaction: Content Isolation for a Two-Context Companion

How an AI companion isolates two content contexts structurally: a fail-closed valve, segment-level memory masking, constant bridge text, and three isolation failures.

Read post ↗
O12

Contact Is the Boundary: 228 Prompts and Ten Failed Experiments

Across 228 prompts from five published AI-video workflows, not one described physical contact between characters. Ten subsequent generation tests reinforced that boundary from the other side.

Read post ↗
O11

Let the storyboard follow the image: keyframes and animation for a painterly pilot

Building a pipeline for a 90-second dialogue-free animated pilot: why locking composition into the storyboard failed, and where painterly texture was lost during video generation.

Read post ↗
O10

Claude Fable 5 Route: Rebuilding The Vigil with Human Keyframes

A Claude Fable 5-driven Flova restart: the creator owned visual judgment and Midjourney keyframes, while the Agent drove the CLI, per-shot review, and orchestration to produce a 14.2-second atmospheric film.

Read post ↗
O09

Codex 5.6 Sol Route: Why the Five-Shot Sylas Greenfield Film Failed

A Codex 5.6 Sol-driven Flova experiment completed its storyboard, per-shot generation, sound, and export, yet the finished film still scored only 3–4/10.

Read post ↗
O08

Flova End-to-End Test: The Project Ran, the Film Still Failed

After shot-by-shot generation failed, the same action sequence was handed to Flova for director audit, storyboarding, keyframes, four-shot generation, sound, and export. The system delivered a 38.24-second film, but character contamination, style drift, weak contact logic, and false-positive self-review meant the workflow still failed.

Read post ↗
O07

AI Action Direction Failure Log: Shot-by-Shot Generation and the Shift to Flova

A stone-giant combat sequence becomes a test of Niji, Image2, Seedance, Dreamina, BytePlus, and Kling—and of why action causality, spatial continuity, and three-dimensional distance ultimately broke the shot-by-shot workflow.

Read post ↗
O05

Web Combat Specimen: From the Main Game to a Deployable Demo

A record of extracting Watchman's movement, attacks, evasion, poise break, and execution loop into a web demo, including asset boundaries, feel calibration, Web export, and loading interaction.

Read post ↗
O06

Rebuilding a Company Website: From Technical Decisions to Asset Handoff

A structured account of a company-site rebuild, covering requirements, stack selection, localization, deployment, domain and ICP work, and final transfer of control.

Read post ↗
M02

Ableton Cover Project: A Modular Workflow for Filling Missing Parts

Using Make Me Wanna Die as a case study, this article documents how a three-piece band used Ableton Live to supply missing drums, bass, and rhythm guitar while preserving the parts played live.

Read post ↗
R19

Building the Data Before the Model: a Spatial-Regression Pipeline

Most of a spatial study happens before the regression: turning raw points and boundaries into one clean layer with real distances, then the OLS → spatial → local modeling ladder that sits on top.

Read post ↗
O04

From Irregular Orders to Structured Records: Building an AI Intake Tool

A reusable method for parsing orders sent as messages, spreadsheets, images, or scans, then routing uncertainty through confidence checks and human review before writing to a shared table.

Read post ↗
O03

Building Alias Archive: Multi-Agent Collaboration, Astro Migration, and Static Delivery

A complete account of Alias Archive, from its anonymous content boundary and four-field trilingual model to agent handoffs, Astro migration, Impeccable-led interface refinement, and Cloudflare Pages delivery.

Read post ↗
O02

AI Presenter Workflow: Coordinating Codex, HeyGen, and HyperFrames

From one assignment to an English virtual presenter, Chinese subtitles, and a social-ready master—including the revisions and two unresolved defects.

Read post ↗
G09

AI-Assisted Pixel Boss Production: The Stone Giant Pipeline

A production note on moving from incompatible high-resolution concept art to a game-ready pixel boss, including style anchoring, animation constraints, and iteration failures.

Read post ↗
G08

Watchman's Opening Cinematic: Silent Narrative and Production Workflow

A record of Watchman's 59-second silent opening cinematic, from concept keyframes and segmented image-to-video generation to editing, OGV conversion, and Godot integration.

Read post ↗
G07

Asset Pipeline: Repacking, Manifests, and Management Tools

Purchasing assets is only the beginning; the real cost of hundreds of packs lies in atlas repacking, manifest generation, and long-term management.

Read post ↗
G06

Balance Tools: Data-Driven Tuning and Assertion Checks

Move values out of the code and into balance.json, tune them through a web console, and use assertions to preserve the underlying logic.

Read post ↗
G05

Night Sail Workflow: Autonomous AI Development Behind Test Gates

A record of an unattended development workflow built from written task briefs, automated tests, adversarial review, and delivery reports.

Read post ↗
G04

World Structure: Hub, Bonfire Return, and Regional Transit

A hub-and-spoke world built around bonfire saves, death returns, regional lamp transit, and the linear pacing of Chapter One.

Read post ↗
G03

Combat Feel Design: Stance Breaks, Executions, and Pacing

How Watchman uses stance, stance breaks, executions, and parry windows to establish the weight of Soulslike combat.

Read post ↗
G02

Single-Weapon System: Forging, Transformations, and Iron-Sword Builds

A single iron sword carries the full game, while forging, transformations, scabbards, and whetstones preserve build depth.

Read post ↗
G01

Project Kickoff: Solo Soulslike Development with AI Collaboration

An overview of Watchman's project direction and the AI-assisted production model used to organize solo development from a director's position.

Read post ↗
O01

AI Music-Video Production: From Still Generation to Editing and Release

An eight-stage AI image workflow covering song approval, still generation, animation, quality review, colour matching, editing, and release.

Read post ↗
M01

Producing a Character Theme with Suno: Prompt Structure and Arrangement Constraints

Starting from character design, this article documents practical constraints for Style prompts, vocals, harmony, dynamics, and song structure.

Read post ↗
R18

Monopoly: one seller, and the wedge it opens

Foundations note — why marginal revenue falls below price, the two-step quantity-then-price optimum, the Lerner markup, and the deadweight loss against the competitive benchmark.

Read post ↗
R17

Perfect competition

Foundations note — the four assumptions, why the firm sets P = MC, the short-run shutdown rule, and how free entry drives long-run profit to zero at the efficient scale.

Read post ↗
R16

Cost curves: what marginal, average, and the long-run envelope encode

Foundations note — the short-run cost family, why marginal cost cuts the average minima, and how the long-run average cost is the envelope of every plant the firm could build.

Read post ↗
R15

The production function

Foundations note — output from inputs, marginal versus average product, why returns diminish in the short run, and what the isoquant's slope encodes.

Read post ↗
R14

Utility maximization: where preferences meet the budget

Foundations note — the consumer's optimum as a tangency, the equal-marginal 'bang per buck' principle, the Lagrange multiplier as the marginal utility of income, and when the tangency fails.

Read post ↗
R13

The budget constraint

Foundations note — the budget line, why its slope is a pure relative price, and the difference between a parallel shift and a pivot.

Read post ↗
R12

Preferences and utility

Foundations note — the axioms that let a preference ordering be written as a utility function, and what the indifference curve and the MRS actually encode.

Read post ↗
R11

Fixed effects: letting each group be its own control

The within estimator sweeps out every time-invariant group confounder at once — a lot of omitted-variable bias for free — at the price of discarding all the between-group variation.

Read post ↗
R10

Instrumental variables: when the regressor is part of the problem

If an explanatory variable is correlated with the error, OLS is inconsistent. An instrument that moves the regressor without touching the error restores identification — under one assumption you can never test.

Read post ↗
R09

Robust standard errors: fixing the inference, not the estimate

The sandwich estimator repairs OLS standard errors under heteroskedasticity while leaving the coefficients untouched — what it does, the HC0–HC3 variants, and when to reach for weighting instead.

Read post ↗
R08

Weighted least squares fixes variance, not endogeneity

Inverse-variance weighting is the right tool for heteroskedasticity and the wrong tool for a biased model. A note on what changing the weights does and does not repair.

Read post ↗
R07

What a local coefficient means

Geographically weighted regression gives you one coefficient per place instead of one for the whole map. What that surface says — and the three ways it can quietly mislead you.

Read post ↗
R06

The LM decision rule: choosing between spatial models

Moran's I tells you spatial dependence exists but not which kind. The Lagrange-multiplier tests — and especially their robust variants — are how you decide between a spatial error and a spatial lag.

Read post ↗
R05

The two workhorse spatial models — and the one that nests them

Spatial error, spatial lag, and the Durbin model that contains both: where the dependence lives, whether there are spillovers, and why the coefficient stops being the marginal effect.

Read post ↗
R04

When the average hides the story: Moran's I and its local map

A global autocorrelation statistic tells you clustering exists; its local decomposition tells you where, and of what kind — along with the traps that make both easy to misread.

Read post ↗
R03

Choosing a spatial weight matrix

Before any spatial test or model, you have to declare who counts as whose neighbour. The W matrix is that declaration — an assumption, not a fact handed over by the data.

Read post ↗
R02

Spatial dependence and spatial heterogeneity are not the same problem

Two different ways geography enters a regression — a neighbour effect versus a relationship that refuses to stay constant — and why the distinction dictates the whole modelling chain.

Read post ↗
R01

How proximity gets priced: the hedonic idea

Why a house price is really a bundle of implicit prices, and how a log-linear regression reads the market's marginal valuation of things like access to a school or a station.

Read post ↗