Cardona CT Lab AI Tooling / Design & Accessibility
Projects Models Code lab Image lab Mesh lab

Cardona CT Lab / Living Research

Local AI Lab

Self-hosted intelligence, measured in public.

PlantForge, CodeForge, ImageForge, and MeshForge are applied local intelligence tracks. This whitepaper records what each system can do, what it cannot do, and the evidence required before capability expands.

Local stack
InterfaceForge / Open WebUI
RuntimeOllama + ComfyUI + 3D
AuthorityDeterministic systems
Local-first / Human-governed

Research snapshot

What are we trying to prove?

Useful local intelligence can grow without quietly expanding authority. Every improvement must be visible in repeatable evidence.

01

Measure, do not imply

Versions, test results, failures, and promotion decisions stay visible.

02

Separate language from authority

Models explain or propose. Deterministic systems and people approve consequential action.

03

Promote only on evidence

A newer or larger model is a challenger until it passes the defined gate without regression.

Current state

Four systems, four bounded roles

Loading the model registry.

ImageForge visual lab

See the image system mature

Reference direction, local baselines, and trained candidates remain visibly distinct. A good reference is not evidence that the local model produced it.

Publication boundary: only approved, public-safe outputs appear here. Private prompts, evaluator rubrics, rejected generations, and training assets remain outside the site.

MeshForge mesh lab

Build useful geometry without pretending it is CAD

The first Stable Fast 3D runtime completed locally. Its artifact reopened correctly, but topology quality blocked promotion.

MeshForge measured candidate A wireframe object representing the first measured Stable Fast 3D candidate and its withheld promotion. RUNNABLE / PROMOTION WITHHELD

Candidate 01

Runtime proven, topology rejected

Stable Fast 3D produced a reopenable GLB in 70.89 seconds at 6171.84 MiB peak VRAM. The artifact contained 4,136 non-manifold edges, so it remains evidence rather than a promoted result.

01Geometry

Silhouette, completeness, scale, and manifold checks.

02Topology

Non-manifold edges, intersections, density, and repair distance.

03Materials

UV integrity, texture coverage, and visible seams.

04Export

Human-approved GLB or OBJ only. CAD claims require a separate parametric track.

Operating boundary: MeshForge may generate review candidates. It cannot approve topology, overwrite source assets, or publish an export.

Capability register

Model comparison

"Current" means selected for its bounded role. It does not mean autonomous, generally capable, or safe outside that role.

ProjectModelBaseRoleEvaluationSystem evidenceDecision
Loading model evidence.

Measured growth

Capability history

Bars show comparable suite pass rates. Missing measurements remain unmeasured, not zero.

Loading benchmark history.

Public capability lab

See the code mature

Curated demonstrations show how implementation quality changes across releases. They are not private evaluator cases or hidden-test outputs.

Publication boundary: prompts, judges, reference implementations, and raw candidates from sealed evaluations remain private. These specimens are purpose-built for explanation.

Decision record

Promotion log

Every retained, rejected, or testing decision carries a reason and an evidence class.

Loading promotion decisions.

Promotion method

Intelligence grows through gates

Capability is promoted one bounded role at a time. A model never inherits authority from a good demo.

  1. 01
    Define the role

    Name the task, inputs, outputs, and forbidden actions.

  2. 02
    Freeze the evaluation

    Use deterministic tests, hostile probes, and untouched holdouts.

  3. 03
    Run in shadow

    Retain evidence without changing source, hardware, or production state.

  4. 04
    Compare correction distance

    Measure useful completion, missed requirements, regressions, and unnecessary change.

  5. 05
    Promote or reject

    Advance only when the full gate passes and the authority boundary stays intact.

Operating boundary

The model is not the authority.

PlantForge cannot water a plant. CodeForge cannot commit, publish, or deploy. ImageForge cannot select or publish its own output. MeshForge cannot approve or export a generated asset. These are structural constraints, not promises in a prompt.

PlantForgeRecommend

PlantOS validates evidence and owns care-state decisions.

No actuation route
CodeForgePropose

Disposable worktrees and authenticated evidence contain model output.

Human commit required
ImageForgeGenerate

Versioned workflows and frozen briefs produce candidates for review.

Human selection required
MeshForgeConstruct

Local image-to-mesh systems produce bounded geometry candidates.

Human export required

Evaluation history

What changed, and why

Loading the evidence timeline.

Next evidence

What would change the chart?

PlantForge

Physical sensor commissioning, a healthy 24-hour observation window, and a candidate that clears every raw-model gate without regression.

Open PlantForge whitepaper

CodeForge

Untouched holdout success, repeatable useful patches, lower correction distance, and preserved containment under a stronger local coding model.

Review current evidence

ImageForge

A reproducible SDXL baseline, measurable prompt fidelity, visual consistency, and a challenger that improves quality without weakening provenance.

Open the visual lab

MeshForge

A frozen image-to-mesh suite, topology checks, cleanup-distance evidence, and a 12 GB compatible local baseline.

Open the mesh lab