MLX
mlx-lm
flowcast
loudink
loudink-v1.1
voice-agent
apple-silicon
compact-ir
dictation
computer-use
skills
record-and-replay
Instructions to use nsalerni/loudink-v1.1 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use nsalerni/loudink-v1.1 with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir loudink-v1.1 nsalerni/loudink-v1.1
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
loudink-v1.1 1.1.1: P0 lean/mid prompts + correction micro (writing neural 100%)
Browse files- README.md +2 -2
- SHA256SUMS +10 -10
- STATUS.md +14 -7
- inference_config.json +3 -3
- manifest.json +4 -4
- product_scorecard.json +15 -15
- publish_report.json +1 -1
- skill_store/skills/invoice_chase.json +1 -1
- skill_store/skills/morning_focus.json +1 -1
- skill_store/skills/standup_post.json +1 -1
- writer_unified/adapter_config.json +7 -7
- writer_unified/adapters.safetensors +1 -1
README.md
CHANGED
|
@@ -15,7 +15,7 @@ tags:
|
|
| 15 |
library_name: mlx-lm
|
| 16 |
---
|
| 17 |
|
| 18 |
-
# loudink-v1.1 v1.1.
|
| 19 |
|
| 20 |
On-device dual-stack for **LoudInk / Flowcast** on Apple Silicon, plus **user skills**:
|
| 21 |
|
|
@@ -30,7 +30,7 @@ This package is the **v1.1 canary** (PR-A→E). Prefer dual-package client routi
|
|
| 30 |
|
| 31 |
| Slice | Shipped | Neural | Latency |
|
| 32 |
|-------|---------|--------|---------|
|
| 33 |
-
| Writing daily | **100%** | **
|
| 34 |
| Computer daily | **100%** | — | mostly fast-path |
|
| 35 |
| Multi-step | **100%** | — | mostly fast-path |
|
| 36 |
| **Daily-Mac overall (n=69)** | **100%** | — | — |
|
|
|
|
| 15 |
library_name: mlx-lm
|
| 16 |
---
|
| 17 |
|
| 18 |
+
# loudink-v1.1 v1.1.1 (canary)
|
| 19 |
|
| 20 |
On-device dual-stack for **LoudInk / Flowcast** on Apple Silicon, plus **user skills**:
|
| 21 |
|
|
|
|
| 30 |
|
| 31 |
| Slice | Shipped | Neural | Latency |
|
| 32 |
|-------|---------|--------|---------|
|
| 33 |
+
| Writing daily | **100%** | **100.0%** | p50 **184 ms** |
|
| 34 |
| Computer daily | **100%** | — | mostly fast-path |
|
| 35 |
| Multi-step | **100%** | — | mostly fast-path |
|
| 36 |
| **Daily-Mac overall (n=69)** | **100%** | — | — |
|
SHA256SUMS
CHANGED
|
@@ -1,19 +1,19 @@
|
|
| 1 |
-
|
| 2 |
-
|
| 3 |
887cad558e8fde9d4c357f2cdae4507e4e614b1f0dbad190d1fa9d21714ad92f V1_1_COMPLETE.md
|
| 4 |
3965991f7c62bfb29001f9b19f704deefbc890dc85649e4d72fd57243ab317f0 client_contract.md
|
| 5 |
-
|
| 6 |
4159a5a3cd7ed71c4da7a6712b2ec5049d7c1ec2e84cbe538178136e5e99d4f3 ir/adapter_config.json
|
| 7 |
868d11b4360748cb3c493e26e2ee4cf4eef937eb0c9755988c364a741e5d3a82 ir/adapters.safetensors
|
| 8 |
69abbc73e77287e6646bbc323a3fd6869218124b12512cd262a40d0b034ddda2 ir_honesty_fallback/adapter_config.json
|
| 9 |
02b521c2348477cf03f667c7b500f0be82bfb937277806e3f4e562b9593a1bbe ir_honesty_fallback/adapters.safetensors
|
| 10 |
-
|
| 11 |
-
|
| 12 |
8856981bbea72862e2d44a6d5f4de9e540e1793ed1674fcc0dd7ace2364be5f0 router.json
|
| 13 |
12db4143f37da30858061eb64bf3fb645d15e195bb18b9274f3620640eb63f37 skill_store/index.json
|
| 14 |
-
|
| 15 |
-
|
| 16 |
-
|
| 17 |
ceb709ccf7a40ea608c5cc74c247071bbcb0980017f1811f7c55604b4241ab7c writer/promoted_core_100.safetensors
|
| 18 |
-
|
| 19 |
-
|
|
|
|
| 1 |
+
e79d4a3ff4bb2a05b7dcc13afd816abcea782913f0a7552b89e1d6083388e54b README.md
|
| 2 |
+
9cb14787b48df0d0271729464ea2b032562c941d54bb701d3f6c154201b01157 STATUS.md
|
| 3 |
887cad558e8fde9d4c357f2cdae4507e4e614b1f0dbad190d1fa9d21714ad92f V1_1_COMPLETE.md
|
| 4 |
3965991f7c62bfb29001f9b19f704deefbc890dc85649e4d72fd57243ab317f0 client_contract.md
|
| 5 |
+
7b0558b82dcc6016ba38f17ecee0073e0f6a03c9f399ac133f6ae7925a82cf6c inference_config.json
|
| 6 |
4159a5a3cd7ed71c4da7a6712b2ec5049d7c1ec2e84cbe538178136e5e99d4f3 ir/adapter_config.json
|
| 7 |
868d11b4360748cb3c493e26e2ee4cf4eef937eb0c9755988c364a741e5d3a82 ir/adapters.safetensors
|
| 8 |
69abbc73e77287e6646bbc323a3fd6869218124b12512cd262a40d0b034ddda2 ir_honesty_fallback/adapter_config.json
|
| 9 |
02b521c2348477cf03f667c7b500f0be82bfb937277806e3f4e562b9593a1bbe ir_honesty_fallback/adapters.safetensors
|
| 10 |
+
62aae5e5d5e8e1c1a2914b9d3e11427e8636826981bc69fb67522fdbefb72c4f manifest.json
|
| 11 |
+
314f9d8f56364e2726458768848aab60ba1dda02ad28b4823edb142cc97c4719 product_scorecard.json
|
| 12 |
8856981bbea72862e2d44a6d5f4de9e540e1793ed1674fcc0dd7ace2364be5f0 router.json
|
| 13 |
12db4143f37da30858061eb64bf3fb645d15e195bb18b9274f3620640eb63f37 skill_store/index.json
|
| 14 |
+
a7f482c3bae12b95d8fae950642701280e881796de2e18b8044e92d9f883a7dd skill_store/skills/invoice_chase.json
|
| 15 |
+
bcadd34579d6b57d4d584158d58aba534b072ac13603564c824c36addd4ad8de skill_store/skills/morning_focus.json
|
| 16 |
+
d3313b56bbc1571e2af3ee7e1c3be8398d37b657fa90a351abed811d9f95ee26 skill_store/skills/standup_post.json
|
| 17 |
ceb709ccf7a40ea608c5cc74c247071bbcb0980017f1811f7c55604b4241ab7c writer/promoted_core_100.safetensors
|
| 18 |
+
4a1802e622de9412ca8f675639ecc6f7cc0f91ba3ed7c09662f1d8b7b6ec8c9a writer_unified/adapter_config.json
|
| 19 |
+
fce889b02639a5ad3b99b8184a714a84e84cb469a68084b142f242aa7ce6a8a9 writer_unified/adapters.safetensors
|
STATUS.md
CHANGED
|
@@ -7,7 +7,7 @@ Last updated: 2026-07-08
|
|
| 7 |
| Line | Repo | Role |
|
| 8 |
|------|------|------|
|
| 9 |
| **Stable GA** | [nsalerni/loudink-v1](https://huggingface.co/nsalerni/loudink-v1) | **0.4.0** freeze |
|
| 10 |
-
| **Canary** | [nsalerni/loudink-v1.1](https://huggingface.co/nsalerni/loudink-v1.1) | **1.1.
|
| 11 |
|
| 12 |
Rebuild canary:
|
| 13 |
|
|
@@ -33,19 +33,26 @@ python -m gemmaflow_tune.cli.publish_sota --model loudink-v1.1 \
|
|
| 33 |
|--------|-----|--------|
|
| 34 |
| Skill invoke (n=19) | ≥95% named / ≥85% paraphrase | **100%** shipped |
|
| 35 |
| Daily-Mac overall (n=69) | no drop vs 0.4 | **100%** shipped |
|
| 36 |
-
| Writing neural | hold ≥80% | **
|
| 37 |
-
| Writing p50 | ≤250 ms | **
|
| 38 |
| Computer / multi-step | hold 100% | **100%** |
|
| 39 |
| IR honesty (n=112) | ≥45% neural | **46.4%** (v0.4 lineage) |
|
| 40 |
|
| 41 |
-
### Daily-Mac scorecard (v1.1
|
| 42 |
|
| 43 |
| Slice | Shipped | Neural | p50 |
|
| 44 |
|-------|---------|--------|-----|
|
| 45 |
-
| Writing (n=25) | **100%** | **
|
| 46 |
| Computer (n=44) | **100%** | — | ~0 ms FP |
|
| 47 |
| Multi-step (n=25) | **100%** | — | ~0 ms FP |
|
| 48 |
-
| **Overall (n=69)** | **100%** | — | mean ~
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 49 |
|
| 50 |
Scorecard: `artifacts/loudink_v1/product_scorecard.json`
|
| 51 |
Skill bench: `artifacts/benchmarks/loudink_v1_1/skill_invoke/`
|
|
@@ -53,7 +60,7 @@ Skill bench: `artifacts/benchmarks/loudink_v1_1/skill_invoke/`
|
|
| 53 |
## Production stack (v1.1 canary)
|
| 54 |
|
| 55 |
```
|
| 56 |
-
writer: artifacts/
|
| 57 |
polish: structural (+ daily rules)
|
| 58 |
ir: artifacts/sft_loudink_v1_1_skill_invoke_ir/adapters
|
| 59 |
ir_fallback: artifacts/sft_loudink_v1_ir_honesty/adapters
|
|
|
|
| 7 |
| Line | Repo | Role |
|
| 8 |
|------|------|------|
|
| 9 |
| **Stable GA** | [nsalerni/loudink-v1](https://huggingface.co/nsalerni/loudink-v1) | **0.4.0** freeze |
|
| 10 |
+
| **Canary** | [nsalerni/loudink-v1.1](https://huggingface.co/nsalerni/loudink-v1.1) | **1.1.1** skills + P0 speed/accuracy |
|
| 11 |
|
| 12 |
Rebuild canary:
|
| 13 |
|
|
|
|
| 33 |
|--------|-----|--------|
|
| 34 |
| Skill invoke (n=19) | ≥95% named / ≥85% paraphrase | **100%** shipped |
|
| 35 |
| Daily-Mac overall (n=69) | no drop vs 0.4 | **100%** shipped |
|
| 36 |
+
| Writing neural | hold ≥80% | **100%** (P0) |
|
| 37 |
+
| Writing p50 | ≤250 ms | **184 ms** |
|
| 38 |
| Computer / multi-step | hold 100% | **100%** |
|
| 39 |
| IR honesty (n=112) | ≥45% neural | **46.4%** (v0.4 lineage) |
|
| 40 |
|
| 41 |
+
### Daily-Mac scorecard (v1.1.1 P0)
|
| 42 |
|
| 43 |
| Slice | Shipped | Neural | p50 |
|
| 44 |
|-------|---------|--------|-----|
|
| 45 |
+
| Writing (n=25) | **100%** | **100%** | **184 ms** |
|
| 46 |
| Computer (n=44) | **100%** | — | ~0 ms FP |
|
| 47 |
| Multi-step (n=25) | **100%** | — | ~0 ms FP |
|
| 48 |
+
| **Overall (n=69)** | **100%** | — | mean ~63 ms |
|
| 49 |
+
|
| 50 |
+
### P0 speed/accuracy (1.1.1)
|
| 51 |
+
|
| 52 |
+
- Lean + mid dictation prompts (cut full ~260-token prefill on short/mid cases)
|
| 53 |
+
- `writing_daily` in dictation early-stop suites + **coverage-gated** sanitized stop (no multi-clause truncation)
|
| 54 |
+
- Wait-no / money / time correction micro SFT → residual neural fails closed
|
| 55 |
+
- Writer: `artifacts/sft_loudink_v1_1_correction_micro/adapters`
|
| 56 |
|
| 57 |
Scorecard: `artifacts/loudink_v1/product_scorecard.json`
|
| 58 |
Skill bench: `artifacts/benchmarks/loudink_v1_1/skill_invoke/`
|
|
|
|
| 60 |
## Production stack (v1.1 canary)
|
| 61 |
|
| 62 |
```
|
| 63 |
+
writer: artifacts/sft_loudink_v1_1_correction_micro/adapters
|
| 64 |
polish: structural (+ daily rules)
|
| 65 |
ir: artifacts/sft_loudink_v1_1_skill_invoke_ir/adapters
|
| 66 |
ir_fallback: artifacts/sft_loudink_v1_ir_honesty/adapters
|
inference_config.json
CHANGED
|
@@ -2,7 +2,7 @@
|
|
| 2 |
"model_name": "loudink-v1.1",
|
| 3 |
"model_tag": "loudink-v1.1",
|
| 4 |
"product": "loudink-v1.1",
|
| 5 |
-
"version": "1.1.
|
| 6 |
"runner_kind": "loudink_v1",
|
| 7 |
"hf_repo": "nsalerni/loudink-v1.1",
|
| 8 |
"predecessor": "nsalerni/loudink-v1",
|
|
@@ -38,8 +38,8 @@
|
|
| 38 |
"metrics_v11": {
|
| 39 |
"daily_mac_overall_shipped": 1.0,
|
| 40 |
"daily_mac_writing_shipped": 1.0,
|
| 41 |
-
"daily_mac_writing_neural":
|
| 42 |
-
"daily_mac_writing_p50_ms":
|
| 43 |
"daily_mac_computer_shipped": 1.0,
|
| 44 |
"daily_mac_multi_step_shipped": 1.0,
|
| 45 |
"daily_mac_n": 69,
|
|
|
|
| 2 |
"model_name": "loudink-v1.1",
|
| 3 |
"model_tag": "loudink-v1.1",
|
| 4 |
"product": "loudink-v1.1",
|
| 5 |
+
"version": "1.1.1",
|
| 6 |
"runner_kind": "loudink_v1",
|
| 7 |
"hf_repo": "nsalerni/loudink-v1.1",
|
| 8 |
"predecessor": "nsalerni/loudink-v1",
|
|
|
|
| 38 |
"metrics_v11": {
|
| 39 |
"daily_mac_overall_shipped": 1.0,
|
| 40 |
"daily_mac_writing_shipped": 1.0,
|
| 41 |
+
"daily_mac_writing_neural": 1.0,
|
| 42 |
+
"daily_mac_writing_p50_ms": 184,
|
| 43 |
"daily_mac_computer_shipped": 1.0,
|
| 44 |
"daily_mac_multi_step_shipped": 1.0,
|
| 45 |
"daily_mac_n": 69,
|
manifest.json
CHANGED
|
@@ -1,6 +1,6 @@
|
|
| 1 |
{
|
| 2 |
"product": "loudink-v1.1",
|
| 3 |
-
"version": "1.1.
|
| 4 |
"pipeline_namespace": "loudink_v1_1",
|
| 5 |
"runner_kind": "loudink_v1",
|
| 6 |
"gate_mode": "warn",
|
|
@@ -50,8 +50,8 @@
|
|
| 50 |
"metrics_v11": {
|
| 51 |
"daily_mac_overall_shipped": 1.0,
|
| 52 |
"daily_mac_writing_shipped": 1.0,
|
| 53 |
-
"daily_mac_writing_neural":
|
| 54 |
-
"daily_mac_writing_p50_ms":
|
| 55 |
"daily_mac_computer_shipped": 1.0,
|
| 56 |
"daily_mac_multi_step_shipped": 1.0,
|
| 57 |
"daily_mac_n": 69,
|
|
@@ -59,5 +59,5 @@
|
|
| 59 |
"skill_invoke_n": 19,
|
| 60 |
"ir_honesty_neural_n112": 0.4643
|
| 61 |
},
|
| 62 |
-
"notes": "1.1.
|
| 63 |
}
|
|
|
|
| 1 |
{
|
| 2 |
"product": "loudink-v1.1",
|
| 3 |
+
"version": "1.1.1",
|
| 4 |
"pipeline_namespace": "loudink_v1_1",
|
| 5 |
"runner_kind": "loudink_v1",
|
| 6 |
"gate_mode": "warn",
|
|
|
|
| 50 |
"metrics_v11": {
|
| 51 |
"daily_mac_overall_shipped": 1.0,
|
| 52 |
"daily_mac_writing_shipped": 1.0,
|
| 53 |
+
"daily_mac_writing_neural": 1.0,
|
| 54 |
+
"daily_mac_writing_p50_ms": 184,
|
| 55 |
"daily_mac_computer_shipped": 1.0,
|
| 56 |
"daily_mac_multi_step_shipped": 1.0,
|
| 57 |
"daily_mac_n": 69,
|
|
|
|
| 59 |
"skill_invoke_n": 19,
|
| 60 |
"ir_honesty_neural_n112": 0.4643
|
| 61 |
},
|
| 62 |
+
"notes": "1.1.1 canary P0: lean/mid dictation prompts + writing_daily early-stop coverage fix + wait-no correction micro SFT. Writing neural 100% (was 92%), overall 69/69, p50 writing ~184ms. Skills PR-A\u2013E unchanged."
|
| 63 |
}
|
product_scorecard.json
CHANGED
|
@@ -2,7 +2,7 @@
|
|
| 2 |
"reports": [
|
| 3 |
{
|
| 4 |
"source": "artifacts/benchmarks/loudink_v1/daily_mac/loudink_v1_daily_mac_measurement.json",
|
| 5 |
-
"label": "loudink-v1
|
| 6 |
"writing": {
|
| 7 |
"shipped": {
|
| 8 |
"total": 25,
|
|
@@ -11,19 +11,19 @@
|
|
| 11 |
},
|
| 12 |
"neural": {
|
| 13 |
"total": 25,
|
| 14 |
-
"passed":
|
| 15 |
-
"accuracy":
|
| 16 |
"missing": 0
|
| 17 |
},
|
| 18 |
-
"polish_lift": 0.
|
| 19 |
"latency": {
|
| 20 |
"n": 25,
|
| 21 |
-
"mean":
|
| 22 |
-
"p50":
|
| 23 |
-
"p95":
|
| 24 |
-
"max":
|
| 25 |
-
"model_invoked_p50":
|
| 26 |
-
"model_invoked_p95":
|
| 27 |
"model_invoked_n": 20,
|
| 28 |
"fast_path_rate": 0.19999999999999996
|
| 29 |
}
|
|
@@ -79,12 +79,12 @@
|
|
| 79 |
},
|
| 80 |
"latency": {
|
| 81 |
"n": 69,
|
| 82 |
-
"mean":
|
| 83 |
"p50": 0.05,
|
| 84 |
-
"p95":
|
| 85 |
-
"max":
|
| 86 |
-
"model_invoked_p50":
|
| 87 |
-
"model_invoked_p95":
|
| 88 |
"model_invoked_n": 20,
|
| 89 |
"fast_path_rate": 0.7101449275362319
|
| 90 |
}
|
|
|
|
| 2 |
"reports": [
|
| 3 |
{
|
| 4 |
"source": "artifacts/benchmarks/loudink_v1/daily_mac/loudink_v1_daily_mac_measurement.json",
|
| 5 |
+
"label": "loudink-v1.1 daily-mac product (n=69)",
|
| 6 |
"writing": {
|
| 7 |
"shipped": {
|
| 8 |
"total": 25,
|
|
|
|
| 11 |
},
|
| 12 |
"neural": {
|
| 13 |
"total": 25,
|
| 14 |
+
"passed": 25,
|
| 15 |
+
"accuracy": 1.0,
|
| 16 |
"missing": 0
|
| 17 |
},
|
| 18 |
+
"polish_lift": 0.0,
|
| 19 |
"latency": {
|
| 20 |
"n": 25,
|
| 21 |
+
"mean": 175.00264003708958,
|
| 22 |
+
"p50": 183.71283309534192,
|
| 23 |
+
"p95": 322.333584073931,
|
| 24 |
+
"max": 329.85158381052315,
|
| 25 |
+
"model_invoked_p50": 196.95341680198908,
|
| 26 |
+
"model_invoked_p95": 322.333584073931,
|
| 27 |
"model_invoked_n": 20,
|
| 28 |
"fast_path_rate": 0.19999999999999996
|
| 29 |
}
|
|
|
|
| 79 |
},
|
| 80 |
"latency": {
|
| 81 |
"n": 69,
|
| 82 |
+
"mean": 63.43863769459767,
|
| 83 |
"p50": 0.05,
|
| 84 |
+
"p95": 291.4309168700129,
|
| 85 |
+
"max": 329.85158381052315,
|
| 86 |
+
"model_invoked_p50": 196.95341680198908,
|
| 87 |
+
"model_invoked_p95": 322.333584073931,
|
| 88 |
"model_invoked_n": 20,
|
| 89 |
"fast_path_rate": 0.7101449275362319
|
| 90 |
}
|
publish_report.json
CHANGED
|
@@ -1,7 +1,7 @@
|
|
| 1 |
{
|
| 2 |
"hf_repo": "nsalerni/loudink-v1.1",
|
| 3 |
"output_dir": "artifacts/hf_publish/loudink-v1.1",
|
| 4 |
-
"version": "1.1.
|
| 5 |
"measured_adapter_bundle_mib": 70.14,
|
| 6 |
"files": [
|
| 7 |
"README.md",
|
|
|
|
| 1 |
{
|
| 2 |
"hf_repo": "nsalerni/loudink-v1.1",
|
| 3 |
"output_dir": "artifacts/hf_publish/loudink-v1.1",
|
| 4 |
+
"version": "1.1.1",
|
| 5 |
"measured_adapter_bundle_mib": 70.14,
|
| 6 |
"files": [
|
| 7 |
"README.md",
|
skill_store/skills/invoice_chase.json
CHANGED
|
@@ -6,7 +6,7 @@
|
|
| 6 |
"chase invoices"
|
| 7 |
],
|
| 8 |
"source": "user_recording",
|
| 9 |
-
"created_at": "2026-07-09T05:
|
| 10 |
"skill_version": 1,
|
| 11 |
"schema_version": "1.1.0-pra",
|
| 12 |
"steps": [
|
|
|
|
| 6 |
"chase invoices"
|
| 7 |
],
|
| 8 |
"source": "user_recording",
|
| 9 |
+
"created_at": "2026-07-09T05:57:44+00:00",
|
| 10 |
"skill_version": 1,
|
| 11 |
"schema_version": "1.1.0-pra",
|
| 12 |
"steps": [
|
skill_store/skills/morning_focus.json
CHANGED
|
@@ -6,7 +6,7 @@
|
|
| 6 |
"morning focus workflow"
|
| 7 |
],
|
| 8 |
"source": "user_recording",
|
| 9 |
-
"created_at": "2026-07-09T05:
|
| 10 |
"skill_version": 1,
|
| 11 |
"schema_version": "1.1.0-pra",
|
| 12 |
"steps": [
|
|
|
|
| 6 |
"morning focus workflow"
|
| 7 |
],
|
| 8 |
"source": "user_recording",
|
| 9 |
+
"created_at": "2026-07-09T05:57:44+00:00",
|
| 10 |
"skill_version": 1,
|
| 11 |
"schema_version": "1.1.0-pra",
|
| 12 |
"steps": [
|
skill_store/skills/standup_post.json
CHANGED
|
@@ -6,7 +6,7 @@
|
|
| 6 |
"daily standup slack"
|
| 7 |
],
|
| 8 |
"source": "user_recording",
|
| 9 |
-
"created_at": "2026-07-09T05:
|
| 10 |
"skill_version": 1,
|
| 11 |
"schema_version": "1.1.0-pra",
|
| 12 |
"steps": [
|
|
|
|
| 6 |
"daily standup slack"
|
| 7 |
],
|
| 8 |
"source": "user_recording",
|
| 9 |
+
"created_at": "2026-07-09T05:57:44+00:00",
|
| 10 |
"skill_version": 1,
|
| 11 |
"schema_version": "1.1.0-pra",
|
| 12 |
"steps": [
|
writer_unified/adapter_config.json
CHANGED
|
@@ -1,15 +1,15 @@
|
|
| 1 |
{
|
| 2 |
-
"adapter_path": "artifacts/
|
| 3 |
"batch_size": 1,
|
| 4 |
"clear_cache_threshold": 0,
|
| 5 |
"config": null,
|
| 6 |
-
"data": "data/
|
| 7 |
"fine_tune_type": "lora",
|
| 8 |
"gpu_mode": "shared",
|
| 9 |
"grad_accumulation_steps": 8,
|
| 10 |
"grad_checkpoint": true,
|
| 11 |
-
"iters":
|
| 12 |
-
"learning_rate":
|
| 13 |
"lora_parameters": {
|
| 14 |
"dropout": 0.05,
|
| 15 |
"keys": [
|
|
@@ -45,9 +45,9 @@
|
|
| 45 |
"project_name": null,
|
| 46 |
"report_to": null,
|
| 47 |
"resume_adapter_file": "artifacts/sft_loudink_v1_rft_daily/adapters/adapters.safetensors",
|
| 48 |
-
"save_every":
|
| 49 |
-
"seed":
|
| 50 |
-
"steps_per_eval":
|
| 51 |
"steps_per_report": 10,
|
| 52 |
"test": false,
|
| 53 |
"test_batches": 500,
|
|
|
|
| 1 |
{
|
| 2 |
+
"adapter_path": "artifacts/sft_loudink_v1_1_correction_micro/adapters",
|
| 3 |
"batch_size": 1,
|
| 4 |
"clear_cache_threshold": 0,
|
| 5 |
"config": null,
|
| 6 |
+
"data": "data/sft_loudink_v1_1_correction_micro",
|
| 7 |
"fine_tune_type": "lora",
|
| 8 |
"gpu_mode": "shared",
|
| 9 |
"grad_accumulation_steps": 8,
|
| 10 |
"grad_checkpoint": true,
|
| 11 |
+
"iters": 60,
|
| 12 |
+
"learning_rate": 6e-07,
|
| 13 |
"lora_parameters": {
|
| 14 |
"dropout": 0.05,
|
| 15 |
"keys": [
|
|
|
|
| 45 |
"project_name": null,
|
| 46 |
"report_to": null,
|
| 47 |
"resume_adapter_file": "artifacts/sft_loudink_v1_rft_daily/adapters/adapters.safetensors",
|
| 48 |
+
"save_every": 20,
|
| 49 |
+
"seed": 41,
|
| 50 |
+
"steps_per_eval": 20,
|
| 51 |
"steps_per_report": 10,
|
| 52 |
"test": false,
|
| 53 |
"test_batches": 500,
|
writer_unified/adapters.safetensors
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 22292714
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:fce889b02639a5ad3b99b8184a714a84e84cb469a68084b142f242aa7ce6a8a9
|
| 3 |
size 22292714
|