pin v2 payload: engine 2610041643001, cuda 2610040656001, cuda12 2610040730001, sbsa 2610040704001, vulkan 2610041656001, macos 2610041000001, media 2610040945001
Browse files- .gitattributes +4 -0
- components/engine/2610041643001/BUILD_INFO.engine.md +57 -0
- components/engine/2610041643001/opencoti-0.10.5-c7-2610041643001 +3 -0
- components/engine/2610041643001/opencoti-0.10.5-c7-2610041643001.exe +3 -0
- components/vulkan/2610041656001/BUILD_INFO.vulkan.md +33 -0
- components/vulkan/2610041656001/ggml-vulkan-win-x86_64.dll +3 -0
- components/vulkan/2610041656001/ggml-vulkan-x86_64.so +3 -0
.gitattributes
CHANGED
|
@@ -230,3 +230,7 @@ components/engine/2610041105001/opencoti-0.10.5-c7-2610041105001 filter=lfs diff
|
|
| 230 |
components/engine/2610041105001/opencoti-0.10.5-c7-2610041105001.exe filter=lfs diff=lfs merge=lfs -text
|
| 231 |
components/vulkan/2610041121001/ggml-vulkan-x86_64.so filter=lfs diff=lfs merge=lfs -text
|
| 232 |
components/vulkan/2610041121001/ggml-vulkan-win-x86_64.dll filter=lfs diff=lfs merge=lfs -text
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 230 |
components/engine/2610041105001/opencoti-0.10.5-c7-2610041105001.exe filter=lfs diff=lfs merge=lfs -text
|
| 231 |
components/vulkan/2610041121001/ggml-vulkan-x86_64.so filter=lfs diff=lfs merge=lfs -text
|
| 232 |
components/vulkan/2610041121001/ggml-vulkan-win-x86_64.dll filter=lfs diff=lfs merge=lfs -text
|
| 233 |
+
components/engine/2610041643001/opencoti-0.10.5-c7-2610041643001 filter=lfs diff=lfs merge=lfs -text
|
| 234 |
+
components/engine/2610041643001/opencoti-0.10.5-c7-2610041643001.exe filter=lfs diff=lfs merge=lfs -text
|
| 235 |
+
components/vulkan/2610041656001/ggml-vulkan-x86_64.so filter=lfs diff=lfs merge=lfs -text
|
| 236 |
+
components/vulkan/2610041656001/ggml-vulkan-win-x86_64.dll filter=lfs diff=lfs merge=lfs -text
|
components/engine/2610041643001/BUILD_INFO.engine.md
ADDED
|
@@ -0,0 +1,57 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
# engine 2610041643001
|
| 2 |
+
|
| 3 |
+
Component of the opencoti pin format v2 (docs/protocols/PIN_FORMAT.md).
|
| 4 |
+
|
| 5 |
+
| file | platform | kind | sha256 | bytes |
|
| 6 |
+
|---|---|---|---|---|
|
| 7 |
+
| `opencoti-0.10.5-c7-2610041643001` | any | bin | `95672d829b5eb61cca1aa3e23c30ab3d4cd0bd678a2c0504585b7834e08d1a92` | 134677035 |
|
| 8 |
+
| `opencoti-0.10.5-c7-2610041643001.exe` | win-x86_64 | bin | `95672d829b5eb61cca1aa3e23c30ab3d4cd0bd678a2c0504585b7834e08d1a92` | 134677035 |
|
| 9 |
+
|
| 10 |
+
ABI `ggml` = `41ba479c24d7e3a624c7b13aa3cbe357e41fac04a504c54ec2ee170de0bff9a1` (exported) β sha256 of this list:
|
| 11 |
+
|
| 12 |
+
```
|
| 13 |
+
llama.cpp/ggml/include/ggml-alloc.h 94e4cd069b9313b2ceb35dacec901981e0bb478d8bb31035b7126be091998c23
|
| 14 |
+
llama.cpp/ggml/include/ggml-backend.h 764d1640ad62579513c4837539708f89e8603c9dbeae053a13d1cbd84a82c865
|
| 15 |
+
llama.cpp/ggml/include/ggml-cuda.h 674ea064021cf5b9970e78f1209e513dd23350b94a44a705a27eec1423fe4b0a
|
| 16 |
+
llama.cpp/ggml/include/ggml-metal.h 322f36cd30f3e9e7aad7b5b9bc63078012fd0b7706ac4e24f5721b604f3d8980
|
| 17 |
+
llama.cpp/ggml/include/ggml-neo-pipeline.h 190f39e183542d16588a3104d8cb4359bf52fc7a62afe8d8cc4c822f4085086d
|
| 18 |
+
llama.cpp/ggml/include/ggml-oc-kvarn-range.h 76635fdd8978632f670640ddf5125a666441e3d47c680bb8b5c0a099e4861202
|
| 19 |
+
llama.cpp/ggml/include/ggml-oc-ledger.h af6dc423ee7c6c3f67a6a187d6dfbfeb94e52ddc1db27c45ee7440289bc44282
|
| 20 |
+
llama.cpp/ggml/include/ggml-vulkan.h 7eae5dad2cc7bb4d3eca828539f441816d5fd59fdc7b224d49aa229fc7b7248c
|
| 21 |
+
llama.cpp/ggml/include/ggml.h f2bb8b127e3e601a78fe32242ffab7235e6a4db3bab5370b72fb6bacba2b2433
|
| 22 |
+
llama.cpp/ggml/include/opencoti-devtype.h 94dea946173449bb76bffd7f733ab822fe658fe7198992de750ac012d29308ec
|
| 23 |
+
llama.cpp/ggml/src/ggml-backend-impl.h f9fdfe6bae61d1387593b40fc0f90ceee490c6a339ec6d7388c51e4fca318d8f
|
| 24 |
+
llamafile/gpu_backend.h d17c504fbab50ff9376105a1f2c79d99fcde4b431886b17b90718a2a359278de
|
| 25 |
+
```
|
| 26 |
+
|
| 27 |
+
ABI `media` = `b099409b3029a6e380268e31d1caefa6706569961d32e72b3543fbc73903d95c` (exported) β sha256 of this list:
|
| 28 |
+
|
| 29 |
+
```
|
| 30 |
+
llamafile/oc-audiocpp/audiocpp.h 6ad0b55d5f28adc9ad60a45bf9a7caa428c4f5441389a6932d6670a8e2a90d73
|
| 31 |
+
llamafile/oc-codec/oc_codec_sidecar.h f2c9001d1414aa2a55533b867b79c52a52c682558e0234c5bebd30bf97b91005
|
| 32 |
+
```
|
| 33 |
+
|
| 34 |
+
ABI `loader` = `d6ff34022f1d3fb5af56ca0c80e3f4e552bf09f1bebdf0cb5f314a454bf38dbc` (exported) β sha256 of this list:
|
| 35 |
+
|
| 36 |
+
```
|
| 37 |
+
.cosmocc/4.0.2/bin/ape-m1.c 78af79e20abbb3f99a97355959b59bd58949564064adb754f0302ebabb4bf7a3
|
| 38 |
+
```
|
| 39 |
+
|
| 40 |
+
The APE: one file for Linux x86_64 / aarch64, Windows and macOS arm64; the `.exe` row is the same bytes under the name Windows needs.
|
| 41 |
+
|
| 42 |
+
Chain through 0556 shutdown-gpu-policy (bug-3919; docs/features/server_shutdown.md), on top of 0555 graceful-shutdown and 0554 abi-export:
|
| 43 |
+
- `POST /shutdown` β loopback only (403 otherwise), 503 while the model is still loading, behind the API key when one is set. The same path as SIGTERM. Feature `graceful_shutdown_v1`.
|
| 44 |
+
- The stop path frees the media engines, the models and the contexts (`shutdown: models and contexts freed`), then for every Vulkan device waits for it to be idle and destroys it (`ggml_vulkan: shutdown: Vulkan0 idle and destroyed`), then the instance (`β¦ N device(s) destroyed, 0 kept, instance destroyed`), exit 0.
|
| 45 |
+
- `OPENCOTI_SHUTDOWN_GPU=auto|off|free|device|full` selects how far that goes; `auto` = `full` on every platform. `off` = nothing released (the exit as it was before 0555): `shutdown: GPU state left to the process exit (β¦)`. `free` = models only; `device` = devices destroyed, instance kept. Feature `shutdown_gpu_policy_v1`.
|
| 46 |
+
- Deadline `OPENCOTI_SHUTDOWN_DEADLINE_MS` (default 10000, 0 = none): `the teardown did not finish within the deadline, exiting now`, exit 0.
|
| 47 |
+
- Parent watch: `OPENCOTI_PARENT_PID=<pid>` (on Windows a Windows process id) β `parent watch: armed on pid N`; when that process is gone, `shutdown: parent pid N is gone` and the same stop path. `NOT armed` when the pid cannot be watched. Feature `parent_watch_v1`.
|
| 48 |
+
- `--list-devices` destroys the Vulkan instance before it exits; `OPENCOTI_LIST_DEVICE_IDS=1` appends ` id=<PCI address | uuid:<32 hex>>` to each line (the stock line is unchanged without it). Feature `list_device_ids_v1`.
|
| 49 |
+
- The Vulkan teardown needs the `vulkan` component of this index; with an older Vulkan library the engine says the library predates the device teardown and leaves the devices.
|
| 50 |
+
No interface header changed: the `ggml`, `media` and `loader` digests are those of engine 2610040950001, so every other component fits unchanged.
|
| 51 |
+
|
| 52 |
+
AMD Radeon on Windows: AMD Software 26.9.2 (display driver 32.0.32015.2008) or later is the minimum for Vulkan. On 26.8.1 (32.0.31041.1004) a model left idle on an RX 9070 XT times out in the driver (0x141, amdkmdag.sys) with any Vulkan application, this engine and stock llama.cpp alike; the engine does not work around it.
|
| 53 |
+
|
| 54 |
+
Gated on this exact file: Linux x86_64 (bs2), `.opencoti/graceful-shutdown-gate.sh` 15/15 β POST /shutdown and SIGTERM on Vulkan (exit 0 in 1.9-3.1 s, device and instance destroyed), non-loopback 403, stop during a stream, deadline, parent watch armed / not armed, the `off` / `free` / `device` modes, CUDA POST /shutdown, list-devices teardown and ids. NOT run on this file: Windows and macOS, Linux aarch64, the bs2 2M-context swarm.
|
| 55 |
+
Run by xollama on the engine one patch earlier (2610041105001, same stop path with the default `full`), Windows 11 + RX 9070 XT on driver 32.0.32015.2008: POST /shutdown + reopen 10 of 10 clean, 103.3-103.6 tok/s; route, teardown lines and parent watch as specified (mails #768, #778).
|
| 56 |
+
|
| 57 |
+
Image / video features are listed for the platforms they were gated on.
|
components/engine/2610041643001/opencoti-0.10.5-c7-2610041643001
ADDED
|
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:95672d829b5eb61cca1aa3e23c30ab3d4cd0bd678a2c0504585b7834e08d1a92
|
| 3 |
+
size 134677035
|
components/engine/2610041643001/opencoti-0.10.5-c7-2610041643001.exe
ADDED
|
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:95672d829b5eb61cca1aa3e23c30ab3d4cd0bd678a2c0504585b7834e08d1a92
|
| 3 |
+
size 134677035
|
components/vulkan/2610041656001/BUILD_INFO.vulkan.md
ADDED
|
@@ -0,0 +1,33 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
# vulkan 2610041656001
|
| 2 |
+
|
| 3 |
+
Component of the opencoti pin format v2 (docs/protocols/PIN_FORMAT.md).
|
| 4 |
+
|
| 5 |
+
Built from the engine source of build 2610041643001; gated floor (engine-min) 2610041643001; gated with 2610041643001.
|
| 6 |
+
|
| 7 |
+
| file | platform | kind | sha256 | bytes |
|
| 8 |
+
|---|---|---|---|---|
|
| 9 |
+
| `ggml-vulkan-x86_64.so` | x86_64 | vulkan | `8875b020518166b37de8680c4b2c3b0cf17744d8ebea49d7cc2e43879ae650dd` | 56290760 |
|
| 10 |
+
| `ggml-vulkan-win-x86_64.dll` | win-x86_64 | vulkan | `bfe2fbb87331c81618b06dcc5d3da5e20b3cf4032b0a2e2fc07dd8faf77f143b` | 57255849 |
|
| 11 |
+
|
| 12 |
+
ABI `ggml` = `41ba479c24d7e3a624c7b13aa3cbe357e41fac04a504c54ec2ee170de0bff9a1` (exported) β sha256 of this list:
|
| 13 |
+
|
| 14 |
+
```
|
| 15 |
+
llama.cpp/ggml/include/ggml-alloc.h 94e4cd069b9313b2ceb35dacec901981e0bb478d8bb31035b7126be091998c23
|
| 16 |
+
llama.cpp/ggml/include/ggml-backend.h 764d1640ad62579513c4837539708f89e8603c9dbeae053a13d1cbd84a82c865
|
| 17 |
+
llama.cpp/ggml/include/ggml-cuda.h 674ea064021cf5b9970e78f1209e513dd23350b94a44a705a27eec1423fe4b0a
|
| 18 |
+
llama.cpp/ggml/include/ggml-metal.h 322f36cd30f3e9e7aad7b5b9bc63078012fd0b7706ac4e24f5721b604f3d8980
|
| 19 |
+
llama.cpp/ggml/include/ggml-neo-pipeline.h 190f39e183542d16588a3104d8cb4359bf52fc7a62afe8d8cc4c822f4085086d
|
| 20 |
+
llama.cpp/ggml/include/ggml-oc-kvarn-range.h 76635fdd8978632f670640ddf5125a666441e3d47c680bb8b5c0a099e4861202
|
| 21 |
+
llama.cpp/ggml/include/ggml-oc-ledger.h af6dc423ee7c6c3f67a6a187d6dfbfeb94e52ddc1db27c45ee7440289bc44282
|
| 22 |
+
llama.cpp/ggml/include/ggml-vulkan.h 7eae5dad2cc7bb4d3eca828539f441816d5fd59fdc7b224d49aa229fc7b7248c
|
| 23 |
+
llama.cpp/ggml/include/ggml.h f2bb8b127e3e601a78fe32242ffab7235e6a4db3bab5370b72fb6bacba2b2433
|
| 24 |
+
llama.cpp/ggml/include/opencoti-devtype.h 94dea946173449bb76bffd7f733ab822fe658fe7198992de750ac012d29308ec
|
| 25 |
+
llama.cpp/ggml/src/ggml-backend-impl.h f9fdfe6bae61d1387593b40fc0f90ceee490c6a339ec6d7388c51e4fca318d8f
|
| 26 |
+
llamafile/gpu_backend.h d17c504fbab50ff9376105a1f2c79d99fcde4b431886b17b90718a2a359278de
|
| 27 |
+
```
|
| 28 |
+
|
| 29 |
+
Built from the 0556 tree: states the `ggml` interface in its bytes, exports `ggml_backend_vk_oc_shutdown` and `ggml_backend_vk_oc_shutdown_steps` β the teardown the engine calls on a normal exit (wait for each device to be idle, destroy it, destroy the instance; `device` mode keeps the instance). A Vulkan device's id is its PCI address, or `uuid:<deviceUUID>` where the driver has no `VK_EXT_pci_bus_info` (Windows). Carries 0551 vk-adopt-stack as before.
|
| 30 |
+
|
| 31 |
+
AMD Radeon on Windows: AMD Software 26.9.2 (display driver 32.0.32015.2008) or later is the minimum. Older drivers time out with a model left idle (0x141 in amdkmdag.sys), with any Vulkan application.
|
| 32 |
+
|
| 33 |
+
Gated: Linux x86_64 on bs2 (NVIDIA, Vulkan) with engine 2610041643001 β the graceful-shutdown gate 15/15. This Windows DLL (mingw cross lane) has NOT been run. The Windows DLL of the previous component (2610041121001, which differs by the split-step export only) was run by xollama on Windows 11 with an RX 9070 XT, driver 32.0.32015.2008: generate + POST /shutdown + reopen 10 of 10 clean (mail #778).
|
components/vulkan/2610041656001/ggml-vulkan-win-x86_64.dll
ADDED
|
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:bfe2fbb87331c81618b06dcc5d3da5e20b3cf4032b0a2e2fc07dd8faf77f143b
|
| 3 |
+
size 57255849
|
components/vulkan/2610041656001/ggml-vulkan-x86_64.so
ADDED
|
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:8875b020518166b37de8680c4b2c3b0cf17744d8ebea49d7cc2e43879ae650dd
|
| 3 |
+
size 56290760
|