Skip to content

Intern-Decision: int8 Showdown snapshot - #22

Merged
Alex-Wengg merged 1 commit into
mainfrom
feat/intern-decision-int8
Sep 30, 2026
Merged

Alex-Wengg merged 1 commit into
mainfrom
feat/intern-decision-int8

Conversation

@Alex-Wengg

Copy link
Copy Markdown
Member

InternDecisionModelStore.ensure(.showdown) now pins the int8 buckets of FluidInference/intern-decision-0.8b-showdown-coreml at 878e1404.

Measured on the Swift runtime, one bucket in use, 60 calls: 374 MB process footprint + 480 MB mapped weights (about 0.85 GB) vs 529 MB + 955 MB for fp16 (about 1.5 GB); p50 unchanged at 90 ms. In the harness, 15-0 vs poke-env's max-power player against 14-1 for fp16 (15 battles each). Download 1.9 GB instead of 3.5 GB.

A re-pinned snapshot also removes packages and .mlmodelc folders the new asset list no longer names, so an older precision left in the cache folder cannot be preferred by the loader.

🤖 Generated with Claude Code

ensure(.showdown) now pins the int8 buckets of
FluidInference/intern-decision-0.8b-showdown-coreml (revision 878e1404):
same 90 ms per decision as fp16, about 0.85 GB in memory with one bucket
in use instead of about 1.5 GB, 1.9 GB to download instead of 3.5 GB,
15-0 vs 14-1 against poke-env's max-power player over 15 battles.

A re-pinned snapshot also removes packages and compiled models the new
asset list no longer names, so an older precision left in the cache
folder cannot be preferred by the loader.

Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
@Alex-Wengg
Alex-Wengg merged commit c9e967f into main Sep 30, 2026
1 check passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant