coder543 commited on
Commit
7d6b508
·
verified ·
1 Parent(s): 62b5b3b

Use qualified default specialization for Watch Core ML

Browse files
Files changed (3) hide show
  1. README.md +4 -4
  2. coreml-watch/bundle.tar +2 -2
  3. coreml-watch/loading.json +11 -10
README.md CHANGED
@@ -141,7 +141,7 @@ signatures and operation counts. Hugging Face provides file checksums for each r
141
 
142
  ## Experimental Core ML bundle for Apple Watch
143
 
144
- `coreml-watch/bundle.tar` contains compiled watchOS 27 Core ML graphs, runtime configuration, frontend/decoder sidecars, license files and `loading.json` with measured per-file loading baselines. Extract the uncompressed POSIX ustar archive into one model directory. The existing Core AI bundles remain available separately. These Watch graphs request CPU and Neural Engine execution using public Core ML APIs.
145
 
146
  This English TDT configuration uses up-to-15-second chunks, a CPU decoder and six resident encoder stages. Native word timings are available. A 90-second numerical qualification matched source token IDs in three runs.
147
 
@@ -149,8 +149,8 @@ Measured on Apple Watch Ultra 4, watchOS 27.0.1, Release runtime, resident graph
149
 
150
  | Measurement | Apple Watch Ultra 4 |
151
  | --- | ---: |
152
- | Processing throughput, 20-second input | 25.3× real time |
153
- | Observed first preparation | 189.7 s |
154
  | Download size | 668.6 MB |
155
 
156
- Preparation is an observed first load, not a controlled cold-cache benchmark. Processing includes the frontend, model work, cache updates and host decoding where applicable; it excludes preparation, warmup and result writing. A separate hardware trace confirmed ANE execution of encoder stage 0; this does not establish whole-pipeline placement or ALU utilization. First preparation on other devices may differ.
 
141
 
142
  ## Experimental Core ML bundle for Apple Watch
143
 
144
+ `coreml-watch/bundle.tar` contains compiled watchOS 27 Core ML graphs, runtime configuration, frontend/decoder sidecars, license files and `loading.json` with measured per-file loading baselines. Extract the uncompressed POSIX ustar archive into one model directory. The existing Core AI bundles remain available separately. These Watch graphs request CPU and Neural Engine execution using public Core ML APIs, with the default specialization strategy.
145
 
146
  This English TDT configuration uses up-to-15-second chunks, a CPU decoder and six resident encoder stages. Native word timings are available. A 90-second numerical qualification matched source token IDs in three runs.
147
 
 
149
 
150
  | Measurement | Apple Watch Ultra 4 |
151
  | --- | ---: |
152
+ | Processing throughput, 20-second input | 27.2× real time |
153
+ | Observed first preparation | 70.2 s |
154
  | Download size | 668.6 MB |
155
 
156
+ Preparation is an observed first load of the default strategy with existing caches preserved, not a controlled cold-cache benchmark. Processing includes the frontend, model work, cache updates and host decoding where applicable; it excludes preparation, warmup and result writing. A separate hardware trace confirmed ANE execution of encoder stage 0; this does not establish whole-pipeline placement or ALU utilization. First preparation on other devices may differ.
coreml-watch/bundle.tar CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:144a3edee53a3d0adcc5fe300e4f367fa177cec921bb93229e80d87de53e3855
3
- size 668620800
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:26dbd456056ca3d10e07e087857318733dd0dcfd5fbca06edb57bf38f5b483f1
3
+ size 668641280
coreml-watch/loading.json CHANGED
@@ -1,14 +1,15 @@
1
  {
 
2
  "device": "Apple Watch Ultra 4",
3
- "os": "watchOS 27.0.1",
4
- "measurement_scope": "Observed first model preparation; not a controlled cache-purge measurement. Includes any on-device specialization.",
5
  "cold_load_seconds": {
6
- "subsampling": 4.074,
7
- "encoder_0": 31.299,
8
- "encoder_1": 30.97,
9
- "encoder_2": 30.778,
10
- "encoder_3": 30.999,
11
- "encoder_4": 30.549,
12
- "encoder_5": 30.814
13
- }
 
14
  }
 
1
  {
2
+ "specialization_strategy": "default",
3
  "device": "Apple Watch Ultra 4",
4
+ "os": "Version 27.0.1 (Build 24R365)",
 
5
  "cold_load_seconds": {
6
+ "subsampling": 0.6473711666767485,
7
+ "encoder_0": 11.457231166670681,
8
+ "encoder_1": 11.467421708337497,
9
+ "encoder_2": 11.44494304167165,
10
+ "encoder_3": 11.574859916669084,
11
+ "encoder_4": 12.016754333337303,
12
+ "encoder_5": 11.474410083334078
13
+ },
14
+ "measurement": "Observed first strategy-specific load; existing caches preserved. Actual per-file preparation calibrates the ETA."
15
  }