Claim

Qwen3.6-35B-A3B Q4_K_M Ryzen AI Max+ 395-ի վրա (llama.cpp Vulkan) — share of output characters spent on reasoning: 86.5 %

active lab_unit_replicated strix.qwen36.interactive2.c1.reasoning-share-4k

86.5%

Չափումը

Beelink GTR9 Pro — AMD Ryzen AI Max+ 395, Radeon 8060S (gfx1151), 128 GB LPDDR5X-8000 unified համակարգի վրա, llama.cpp (server, Vulkan backend), b9049 (server_fingerprint b9049-2496f9c14) runtime-ով և ggml-org/Qwen3.6-35B-A3B-GGUF @ baec3ebee244 (Q4_K_M) մոդելով չափված «share of output characters spent on reasoning» մեծությունը կազմել է 86.5 % (median over valid requests of reasoning characters as a share of all output characters)։ Չափումը կատարվել է interactive-assistant-v2@2026-08-02 սառեցված workload-ի տակ; ապացույցի մակարդակը՝ lab_unit_replicated, 6 վավեր գործարկում 2 ֆիզիկական յունիթի վրա։ Արժեքը վերաստացվում է գործարկումների հում գրառումներից ամեն CI build-ում։

Ձևակերպում

Median share of the generated output that the model spent on the reasoning block rather than on the visible answer, counted in characters, with reasoning enabled at a 4096-token budget, human-task corpus.

Claim-երի ձևակերպումները հրապարակվում են միայն անգլերեն — սա կանոնական, մեջբերելի տարբերակն է։

Ապացույցներ

Համակարգ
Beelink GTR9 Pro — AMD Ryzen AI Max+ 395, Radeon 8060S (gfx1151), 128 GB LPDDR5X-8000 unified
Runtime
llama.cpp (server, Vulkan backend), b9049 (server_fingerprint b9049-2496f9c14)
Մոդելի արտեֆակտ
ggml-org/Qwen3.6-35B-A3B-GGUF @ baec3ebee244 (Q4_K_M)
Շրջանակ
interactive-assistant-v2@2026-08-02
Ագրեգացիա
median over valid requests of reasoning characters as a share of all output characters
Ապացույցի մակարդակ
lab_unit_replicated
Կարգավիճակ
active
Հրապարակված է
02 սեպտեմբերի, 2026 թ.
Սահմանափակումներ
Three repeated runs across two commercially identical units — diagnostic depth, not a cross-unit qualification. Counted in characters, not tokens: reasoning and answer text tokenize differently, so this is a proxy for how the token budget was spent, not the spend itself. Requests whose reasoning consumed the whole budget and left no answer are excluded here and counted by the answerless claim at the same 4096-token budget, which reads the same six runs — read the two together.
Գործարկումներ

Արժեքի ստացումը

Արժեքն ամեն CI build-ում վերաստացվում է գործարկումների հում գրառումներից այս հարցումով — հրապարակված թիվը չի կարող լուռ շեղվել իր ապացույցներից։

catalog/claims/sql/reasoning-share-4k.sql

Որտեղ է մեջբերվում

Կայքի էջերը, որտեղ այս թիվն է ցուցադրվում.

Չափված զույգեր

Ինչպես մեջբերել

AGmind Systems Lab. Qwen3.6-35B-A3B Q4_K_M on Ryzen AI Max+ 395 (llama.cpp Vulkan): share of output characters spent on reasoning — 86.5 % (median over valid requests of reasoning characters as a share of all output characters; 6 runs on 2 units; evidence level lab_unit_replicated; workload interactive-assistant-v2@2026-08-02). Claim strix.qwen36.interactive2.c1.reasoning-share-4k. https://agmind.ai/claims/strix.qwen36.interactive2.c1.reasoning-share-4k/
@misc{agmind_strix_qwen36_interactive2_c1_reasoning_share_4k,
  author       = {{AGmind Systems Lab}},
  title        = {Median share of the generated output that the model spent on the reasoning block rather than on the visible answer, counted in characters, with reasoning enabled at a 4096-token budget, human-task corpus.},
  howpublished = {\url{https://agmind.ai/claims/strix.qwen36.interactive2.c1.reasoning-share-4k/}},
  note         = {Claim strix.qwen36.interactive2.c1.reasoning-share-4k: 86.5 \%; evidence level lab\_unit\_replicated; scope interactive-assistant-v2@2026-08-02},
  year         = {2026}
}

Մեքենայաընթեռնելի. /claims/strix.qwen36.interactive2.c1.reasoning-share-4k.json · BibTeX · CSL-JSON · ամբողջ ռեեստրը · փոփոխությունների հոսք

Ունե՞ք համադրելի սարքաշար։ Այս claim-ը կարելի է վերարտադրել. /reproduce/

Հրապարակված արդյունքների ուղղումները արձանագրվում են հրապարակայնորեն. errata

← Բոլոր claim-երը