Core AI model zoo

d1-omni-600M: an encoder decision model with an image and an audio tower, on Core AI

2026-10-08. LiquidAI/d1-omni-600M (revision 414f8d64…, LFM Open License v1.0) answers typed questions (noul, choice, score) about a state (text or JSON, with images or one voice clip) with a probability for every option, in one forward pass and without generating text. The port re-authors its three networks in plain PyTorch, exports them in fp16, removes their debug locations, compiles them ahead of time, and gates every stage against the publisher’s own model in fp32 on the CPU. Card: models/d1-omni-600m/. Scripts: conversion/d1_omni/. Swift host: apps/D1Omni/. Transcripts below are models/d1-omni-600m/gate-d1-omni-600m-*.json; records outside the repo are marked “lane” and live under $ZOO_WORK_ROOT/_d1_omni/. One fact per line, with the round it came from. Mac = Apple M4 Max, macOS 27.0 (26A428), Xcode 27.0 RC, coreai-build 3600.83.1, coreai-torch 0.4.1, coreai-core 1.0.0b2.

What the checkpoint is

The oracle and the bar

The graph’s form

Precision and compute unit

The image and audio host

Strip

The Swift host

The small buckets (L64 / L128)

Speed on the Mac

iPhone