Artifact storage cleanupReport artifact2026-07-03

Hugging Face Specialist Model Archive v1

This archive makes artifact storage explicit: unique posttrainllm model outputs live on Hugging Face, while plain upstream base-model caches are removed locally instead of being mirrored under posttrainllm.

Headline Numbers

Evidence class applies to each value: measured = retained evaluation or runtime result; historical = recorded result not freshly reproduced; derived = calculated or decided from records; observed = current artifact state; not-measured = an explicit evidence gap.

posttrainllm HF repos

6all have first-class case studiesobserved

Local model cache

clearedafter upload/remote-size verificationobserved

Storage policy

HF firstR2 remains optional private cache or legacy mirrorderived

Competitive Context

SystemMetricScoreSize / ClassComparable?Readout
Hugging Face artifact storagepublic model distributionactivemodel reposDirectCurrent target for public weights, adapters, and large specialist artifacts.
Local Mac cachedurable artifact storagerejected~30GB cleanedDirectUseful during training, but not the system of record after a model is uploaded or known re-downloadable.
Cloudflare R2artifact storage roleoptionalprivate cache / legacy mirrorDirectionalNo longer the default public artifact store for model weights.

Direct rows share this artifact's eval setup. Directional rows are useful market context but should not be read as leaderboard claims.

Uploaded posttrainllm artifacts

ArtifactHF repoStatusReadout
pace-intent-router-v8pace-intent-router-v8release-ready / rejected winner57.1% sealed accuracy at 3.8ms mean; source-matched 95.5% did not generalize
mt4b_fusedqwen3-4b-file-ops-distilledrelease-readyFile-ops hard gate 58% -> 100%; breadth regression disclosed
mt4b_rest_fusedqwen3-4b-rest-fusedrouted specialistFresh 12/12 vs 9/12 depth win; 25/45 vs 30/45 breadth reject
mt4b_mb_fusedqwen3-4b-multibackend-distilledarchive / failed attemptNegative-transfer comparison artifact: depth 100%, breadth 31%
vibethinker-3b-mlxvibethinker-3b-mlxarchiveLocal MLX conversion of the VibeThinker 3B reasoning specialist
vibe_distill_fusedvibethinker-3b-agentic-distilledarchiveposttrainllm distilled VibeThinker variant; needs eval promotion before product use

Deleted upstream caches

Local cacheUpstream repoReason
mxbai-embed-large-v1mixedbread-ai/mxbai-embed-large-v1Public upstream cache; no posttrainllm delta
qwen3-embedding-0.6bQwen/Qwen3-Embedding-0.6BPublic upstream cache; no posttrainllm delta
qwen3-vl-2b-instructQwen/Qwen3-VL-2B-InstructPublic upstream cache; no posttrainllm delta

Release Blockers

Archive entries are not ship decisions

Public weights can be useful as evidence without being the selected model for Pace or any product lane.

Unblock: Promote only candidates with a current factory run, eval report, package metadata, and routed-use decision.

VibeThinker distilled eval needs promotion

The weights are preserved, but the public artifact should not imply a measured win until the eval evidence is attached.

Unblock: Run the factory eval gate and publish a before/after report before using it as a specialist package.

Evidence

Next Release Action

Use this as the storage index; use the six dedicated model case studies for quality and decision evidence.