Model card: FluidUse
Browse files
README.md
CHANGED
|
@@ -21,8 +21,7 @@ generated tokens. Weights are unchanged from
|
|
| 21 |
[`convaiinnovations/laya`](https://huggingface.co/convaiinnovations/laya) `multilingual/` at
|
| 22 |
revision `1c5edc17a7acd8701df6fc341c0d179f1c62c982`.
|
| 23 |
|
| 24 |
-
Runs through [
|
| 25 |
-
macOS 14+ / iOS 17+.
|
| 26 |
|
| 27 |
```swift
|
| 28 |
let laya = try await LayaManager.load() // downloads the 128 + 512 buckets and tokenizer.json
|
|
@@ -33,9 +32,10 @@ print(answer.noul!) // P(true)
|
|
| 33 |
```
|
| 34 |
|
| 35 |
```bash
|
| 36 |
-
swift run -c release
|
| 37 |
--instructions "What does the customer want?" --options "refund|order status|technical help"
|
| 38 |
-
swift run -c release
|
|
|
|
| 39 |
```
|
| 40 |
|
| 41 |
## Files
|
|
@@ -49,7 +49,7 @@ swift run -c release fluidaudiocli laya-tetris # headless Tetris played by lay
|
|
| 49 |
| `tokenizer.json` | | mmBERT / Gemma vocabulary (256k), byte fallback |
|
| 50 |
|
| 51 |
Each bucket is a complete FP16 model (614 MB, 393 MB of which is the embedding table) with
|
| 52 |
-
32 option slots. `
|
| 53 |
the state on the right for the largest one, exactly like laya's `max_len`.
|
| 54 |
|
| 55 |
Inputs: `input_ids` int32 `[1, L]`, `attention_mask` int32 `[1, L]`, `marker_map` float32
|
|
|
|
| 21 |
[`convaiinnovations/laya`](https://huggingface.co/convaiinnovations/laya) `multilingual/` at
|
| 22 |
revision `1c5edc17a7acd8701df6fc341c0d179f1c62c982`.
|
| 23 |
|
| 24 |
+
Runs through [FluidUse](https://github.com/FluidInference/FluidUse) (`LayaManager`) on macOS 14+.
|
|
|
|
| 25 |
|
| 26 |
```swift
|
| 27 |
let laya = try await LayaManager.load() // downloads the 128 + 512 buckets and tokenizer.json
|
|
|
|
| 32 |
```
|
| 33 |
|
| 34 |
```bash
|
| 35 |
+
swift run -c release FluidUseLaya answer --state "…" --type choice \
|
| 36 |
--instructions "What does the customer want?" --options "refund|order status|technical help"
|
| 37 |
+
swift run -c release FluidUseLaya tetris # headless Tetris played by laya decisions
|
| 38 |
+
swift run -c release LayaTetrisDemo # SwiftUI demo
|
| 39 |
```
|
| 40 |
|
| 41 |
## Files
|
|
|
|
| 49 |
| `tokenizer.json` | | mmBERT / Gemma vocabulary (256k), byte fallback |
|
| 50 |
|
| 51 |
Each bucket is a complete FP16 model (614 MB, 393 MB of which is the embedding table) with
|
| 52 |
+
32 option slots. `FluidUse` picks the smallest loaded bucket that fits a prompt and truncates
|
| 53 |
the state on the right for the largest one, exactly like laya's `max_len`.
|
| 54 |
|
| 55 |
Inputs: `input_ids` int32 `[1, L]`, `attention_mask` int32 `[1, L]`, `marker_map` float32
|