docs(#000057): correct cost claim — <$0.10/1k-q is hermes-8B only, not qwen
$0.10/1k-q overstated the qwen-27B case. Honest range: ~$0.07-0.16 per 1,000 queries of GPU electricity — hermes-8B $0.07-0.09 (under a dime), qwen-27B $0.12-0.16 (over a dime; claim_lattice dearer than quote from more prefilled context). Fixes the §5.4 'either rig' claim.
This commit is contained in:
parent
53db4ad717
commit
105b890e41
1 changed files with 3 additions and 1 deletions
|
|
@ -192,7 +192,9 @@ contaminated window integral:
|
|||
|
||||
Hermes substrate ≈ **half** qwen's per-query GPU energy — but that is the
|
||||
four-way confound (§2), dominated by 8 B vs 27 B. A grounded answer costs
|
||||
**< $0.10 / 1,000 queries** of GPU electricity on either rig.
|
||||
**~$0.07–0.16 per 1,000 queries** of GPU electricity — hermes-8B
|
||||
$0.07–0.09, qwen-27B $0.12–0.16 (claim_lattice dearer than quote, more
|
||||
context prefilled).
|
||||
|
||||
### 5.5 Quality delta (value side) — does the substrate earn its cost?
|
||||
|
||||
|
|
|
|||
Loading…
Add table
Add a link
Reference in a new issue