Mage-Flow
Strongest single pickup. MIT-licensed T2I plus instruction editing at 4B — it fits our 96 GB box easily and it is commercially clean, which is exactly the constraint that keeps biting us on CC-BY-NC weights. Worth a ComfyUI lane next to the Ideogram path.
antares-1b
1.8B RL-trained vulnerability-detection terminal agent. Small enough for a 24 GB laptop-GPU node or a Spark to hold permanently. Natural backing model for the cso audit skill, which currently spends frontier tokens on work a specialist could pre-filter.
Nanbeige4.2-3B
Apache-2.0, 4.2B, EN/ZH. The only new release in the window sized for a Jetson-class edge node or an always-resident lane. Worth a same-prompt run against whatever currently holds the small-model slot.
OvisOCR2
Not new today but climbing — 0.8B base, Apache-2.0, tuned for document parsing to markdown with tables and formulas. Directly relevant to the OCR intake alias in the doorbell registry, where the current backing model is much larger than the job needs.
MOSS-Transcribe-Diarize
Already our Scribe backend on our 96 GB box, and still trending at 86 with 111K downloads two months post-release. No action — just confirmation that the single-pass ASR + diarization call was the right one.
Kimi K3
Flag, not an emergency. We run the kimi CLI as a super-team lane. The provenance dispute is about the weights and the vendor, not about our usage — but it is a fact worth knowing before K3 goes anywhere near client-facing work.
Stack v3
4.9T tokens, ODC-By, 15.9TB for the training subset. Only matters if the fleet-model-9B fine-tune line restarts — but it is now the most current code corpus in the open, and the old Stack v2 base is three years stale.
Solar-Open2 / Motif-3
250B and 315B MoE. Both are past what our 96 GB box holds at usable quant, and Motif-3 ships with no license file. Watch for GGUF; do not plan around either.