How it works
One package, every chip.
Apple exposes no public API for Neural Engine placement, so a model compiles to one portable package and runs unchanged from an M1 to an M5 or an A-series iPhone. Every build reads the real per-op hardware placement and diffs it against the last build, so a silent CPU fallback cannot ship.