Skip to content

Category

small language model

339 papers

#small language model Open access Aug 2026

Integration of Cultural Literacy in the Development of Listening Skills Teaching Materials in BIPA Learning

Abstract This study aims to develop listening skills teaching materials in the Indonesian Language for Foreign Speakers (BIPA) course based on cultural literacy for seventh-semester students of the Indonesian Language and Literature Education Study Program at PGRI Silampari University in the 2025/2026 academic year. This study uses a Research and Development (R&D) approach with the ADDIE development model which includes the stages of analysis, design, development, implementation, and evaluation. Data collection techniques were carried out through interviews, questionnaires, and tests. Data analysis was carried out using the Aiken's V formula to measure the level of validity, student response questionnaire analysis to measure practicality, and the N-gain test to determine the level of effectiveness of the teaching materials through a one-group pretest-posttest design. The results of the study showed that the developed listening skills teaching materials based on cultural literacy met the criteria of validity, practicality, and effectiveness. The results of expert validation showed an average value of 0.87 with a very valid category. The results of the practicality test showed a percentage of 85.83% in the small group test and 91.74% in the large group test with a very practical category. In addition, the results of the effectiveness test showed an increase in the average score of students from 40.64 in the pretest to 88.54 in the posttest with an N-gain value of 0.80 which is included in the high category and a percentage of 80.24% with effective criteria. Thus, the cultural literacy-based listening skills teaching materials are effectively used in learning the BIPA course to improve students' listening skills.

Dian Ramadan Lazuardi, Agung Nugroho · 0 citations
#small language model Open access Aug 2026

Source-Preserving Prompt Augmentation and Category-Adaptive Inference for Qwen3-1.7B: A Prospective Seed-Replication Study

This preprint investigates whether a fixed compound inference intervention can improve the performance of a small language model without updating its parameters. The experimental program compares structured prompt replacement, natural-language rewriting, source-preserving semantic augmentation, and a category-adaptive Qwen inference profile. Evaluation uses 48 IFBench and 48 LiveBench tasks with research-authored semantic-stress variants, programmatic scorers, pinned model and benchmark revisions, and matched stochastic seeds.In the prospectively specified P1.3 seed replication, Qwen3-1.7B improved from a mean objective score of 0.272 under raw non-thinking inference to 0.360 under source-preserving augmentation combined with category-adaptive inference. The paired effect was +0.088 (95% percentile-bootstrap CI: 0.018 to 0.160). Component analysis found a positive adaptive-inference-profile effect of +0.061, while the incremental contribution of semantic augmentation under matched adaptive inference remained uncertain at +0.027 (95% CI: −0.046 to 0.099).The study confirms the complete intervention on the fixed 96-task set under new stochastic draws; it does not establish independent task-level replication, cross-model generalization, or semantic augmentation as the active causal component. The package also increased prompt length, latency, and test-time computation. Version 1.1 provides detailed inference settings, benchmark identifiers, systems-cost accounting, causal boundaries, representative interventions, and reproducibility information.

Thibaud Peverelli · 0 citations
#small language model Open access Aug 2026

Source-Preserving Prompt Augmentation and Category-Adaptive Inference for Qwen3-1.7B: A Prospective Seed-Replication Study

This preprint investigates whether a fixed compound inference intervention can improve the performance of a small language model without updating its parameters. The experimental program compares structured prompt replacement, natural-language rewriting, source-preserving semantic augmentation, and a category-adaptive Qwen inference profile. Evaluation uses 48 IFBench and 48 LiveBench tasks with research-authored semantic-stress variants, programmatic scorers, pinned model and benchmark revisions, and matched stochastic seeds.In the prospectively specified P1.3 seed replication, Qwen3-1.7B improved from a mean objective score of 0.272 under raw non-thinking inference to 0.360 under source-preserving augmentation combined with category-adaptive inference. The paired effect was +0.088 (95% percentile-bootstrap CI: 0.018 to 0.160). Component analysis found a positive adaptive-inference-profile effect of +0.061, while the incremental contribution of semantic augmentation under matched adaptive inference remained uncertain at +0.027 (95% CI: −0.046 to 0.099).The study confirms the complete intervention on the fixed 96-task set under new stochastic draws; it does not establish independent task-level replication, cross-model generalization, or semantic augmentation as the active causal component. The package also increased prompt length, latency, and test-time computation. Version 1.1 provides detailed inference settings, benchmark identifiers, systems-cost accounting, causal boundaries, representative interventions, and reproducibility information.

Thibaud Peverelli · 0 citations
#small language model Open access Aug 2026

"Conserved" Has Two Distinct Roots, and Noether Explains Only One ── Shorten the Pendulum and the Energy Rises by 2.000000 While E/omega Does Not Move ── What Separates Them Is Not Symmetry but Slowness ── [Paper 310]

This corpus has cited Noether’s theorem in 27 papers. This paper asks whether every conserved quantity comes from a symmetry──the answer is no. No new mathematical theorem and no new law is claimed. Scope of this paper (scope note): No new mathematical theorem and no new law is claimed──adiabatic invariants, action variables, Noether’s theorem, Ehrenfest’s adiabatic hypothesis and the adiabatic invariance of the magnetic moment are all standard. We do not build mechanics──all we use is one pendulum and the area of one ellipse. We do not prove adiabatic invariance──we do not enter the proof that E/omega is invariant. We check it numerically and name the separator. We do not prove Noether’s theorem──it is merely cited. We do not treat KAM theory──the survival of invariants in non-integrable systems is beyond our tools. We do not adjudicate interpretations of quantum mechanics──Section 6 points only at the algebraic agreement E/omega=(n+1/2)hbar, and enters neither the proof of the adiabatic theorem nor the measurement problem. We do not conflate this with thermodynamic adiabaticity──“adiabatic” here means slow, not thermally isolated. That is a different subject from Paper 126’s integrating factor. Relation to earlier papers: The corpus has cited Noether’s theorem in 27 papers, 341 times──this paper places beside it a conserved quantity Noether does not explain. Paper 16 showed that the equals sign has distinct roots──this paper shows that “conserved” has them. The same form, applied to a claim rather than a symbol. Paper 300 showed that whether two things share a root is decidable, and listed five criteria──this paper applies those criteria to an actual case. Paper 95 separated convention from fact──the invariance of E/omega is a fact, not a convention, and moreover an approximate fact. Paper 180 showed that “the classical limit” is not one limit──the “slowly” of Section 6 is one more limit. What is added is showing numerically that E rises by 2.000000 while E/omega does not move, checking in three cases that the phase-space ellipse keeps its area, giving the drift as exp(-1/epsilon) rather than a power, at 3.72x10^-44, and applying Paper 300’s criteria to conclude distinct roots. First, we build a case where energy is not conserved. Shorten a pendulum’s string slowly from 1.00 m to 0.25 m and E rises by 2.000000 (Section 2). Second, this is the core of the paper. And still E/omega does not move──the Lagrangian depends explicitly on time, so Noether returns nothing (Sections 2 and 3). Third, what is conserved is an area. The phase-space ellipse changes shape and keeps its area (Section 4). Fourth, the separator is slowness. At epsilon=0.01 the drift is 3.72x10^-44──smaller than any power of epsilon (Section 5). Fifth, the same quantity is the quantum number. E/omega=(n+1/2)hbar, so move omega slowly and n does not change (Section 6). Sixth, the two roots can be adjudicated. Applying Paper 300’s criteria (does the agreement persist under motion) returns distinct roots (Section 7). This corpus has cited Noether’s theorem in 27 papers, 341 times──a continuous symmetry gives a conserved quantity. But the theorem never says that is all of them. Shorten a pendulum’s string slowly from 1.00 m to 0.25 m and E rises by 2.000000──the work of pulling enters, the Lagrangian depends on time, and Noether returns nothing. And still E/omega does not move. What is conserved is not a quantity but an area──the phase-space ellipse runs its semi-axes from 1.414214 to 0.707107 and from 1.414214 to 2.828427, and Area/2pi stays at 1.000000. One thing separates them──how slowly it is moved. For a smooth change the drift is not a power of epsilon but exp(-1/epsilon), so at epsilon=0.01 it is 3.72x10^-44──smaller than any power of epsilon. That is why it looks exact. It is not zero. On the quantum side the same quantity is the quantum number──E/omega=(n+1/2)hbar, and Ehrenfest in 1917 used this in reverse: what may be quantised is what is adiabatically invariant. Applying Paper 300’s criteria returns no four times over──under one word, “conserved,” there are two distinct roots. Noether is exact and narrow; the adiabatic invariant is approximate and wide──they trade strength against reach, and neither sits above the other. *Revision Record Second edition (2026-08-30): The subject of this paper has been replaced. The first edition was titled “There Are Three Ways to Show ‘Not Computable,’ and the Equivalence Is a Theorem While the Thesis Is Not,” but its content duplicated Paper 260, “The Equivalence Is a Theorem and the Thesis Is Not”── the three starting points, the difference in status between theorem and thesis, and even the 19729 digits of Ackermann’s A(4,2) all agreed, and Paper 260 has priority. The second edition removes the computability material entirely and refers to Paper 260 for it. The replacement subject, the adiabatic invariant, was chosen because the ground beside the corpus’s heaviest anchor──Noether’s theorem, in 27 papers and 341 places──was empty. On the making of this work: The ideas and content of this work stem from the author's own considerations. Assistance from an AI (a large language model) was used for structuring, English translation, and checking the algebra. Any remaining errors or misinterpretations are solely the author's. Feedback and corrections are sincerely appreciated. ----- 体系はネーターの定理を 27 編で引いてきた。本稿が問うのは、保存量はすべて対称性から来るのかである──答は、来ないである。新しい数学定理も新しい法則も主張しない。 本稿の射程(射程注記):新しい数学定理も新しい法則も主張しない──断熱不変量、作用変数、ネーターの定理、エーレンフェストの断熱仮説、磁気モーメントの断熱不変性は、いずれも標準的である。力学を作らない──使うのは一つの振り子と、一つの楕円の面積だけである。断熱不変量を証明しない──E/omega が不変であることの証明には立ち入らない。数値で確かめ、何が分離子かだけを言う。ネーターの定理を証明しない──引くだけである。 KAM 理論を扱わない──可積分でない系での不変量の生き残りは、本稿の道具では扱わない。量子力学の解釈を判定しない──第6節は E/omega=(n+1/2)hbar という代数的一致を指すだけであり、断熱定理の証明にも、測定の問題にも立ち入らない。熱力学の断熱と混同しない──本稿の「断熱」は「ゆっくり」の意味であり、熱の出入りのことではない。論文126 の積分因子とは別の話である。既刊との関係:体系はネーターの定理を 27 編・341 箇所で引いてきた──本稿はその隣に、ネーターが説明しない保存量を置く。論文16 は「等号にも別根がある」を示した──本稿は「保存する」に別根があると示す。同じ型を、記号ではなく主張に当てる。論文300 は同根か別根かは判定できると示し、五つの基準を並べた──本稿はその基準を実際に一件に適用する。論文95 は規約と事実を分けた──E/omega の不変性は規約ではなく事実であり、しかも近似的な事実である。論文180 は「古典極限」は一つの極限ではないと示した──第6節の「ゆっくり」ももう一つの極限である。加えたのはE が 2.000000 倍になるのに E/omega が動かないことを数で示したこと、位相空間の楕円で面積が保たれることを三例で確かめたこと、ずれが epsilon の冪ではなく exp(-1/epsilon) であることを 3.72x10^-44 という数で出したこと、論文300 の判定基準を当てて別根と結論したことである。 第一に、エネルギーが保存しない場面を作る。振り子の糸を 1.00 m から 0.25 m へゆっくり縮めると、E は 2.000000 倍になる(第2節)。 第二に、これが本稿の芯である。それでも E/omega は動かない──ラグランジアンが時間に依存するのでネーターは何も与えない(第2節・第3節)。 第三に、保存しているのは面積である。位相空間の楕円は形を変えて面積を変えない(第4節)。 第四に、分離子は速さである。 epsilon=0.01 でずれは 3.72x10^-44──epsilon のどの冪よりも小さい(第5節)。 第五に、同じ量が量子数である。 E/omega=(n+1/2)hbar であり、ゆっくり動かせば n は変わらない(第6節)。 第六に、二つの根は判定できる。論文300 の基準(変数を動かしても一致し続けるか)にかけると別根と出る(第7節)。 体系はネーターの定理を 27 編・341 箇所で引いてきた──連続対称性があれば保存量がある、と。だが保存量がそれで全部だとは、定理は言っていない。振り子の糸を 1.00 m から 0.25 m へゆっくり縮めると、E は 2.000000 倍になる──糸を引いた仕事が入るからで、ラグランジアンが時間に依存し、ネーターは何も返さない。それでも E/omega は動かない。保存しているのは量ではなく面積である──位相空間の楕円は半軸が 1.414214->0.707107 と 1.414214->2.828427 に変わりながら、面積/2pi は 1.000000 のままである。分けるものは一つ──どれだけゆっくり動かすか。なめらかに動かせばずれは epsilon の冪ではなく exp(-1/epsilon) で、epsilon=0.01 では 3.72x10^-44──epsilon のどの冪よりも小さい。だから厳密に見える。しかしゼロではない。同じ量が量子側では量子数である──E/omega=(n+1/2)hbar であり、エーレンフェスト 1917 はこれを逆に使って「量子化してよいのは断熱不変量である」と置いた。論文300 の判定基準を当てると、四つとも「いいえ」が返る──同じ「保存する」の下に、別根が二つある。ネーターは厳密で狭く、断熱不変量は近似的で広い──強さと適用範囲を交換しているだけであり、どちらが上位ということはない。 *改訂記録 第2版(2026-08-30):本稿は主題を入れ替えた。 第1版は「「計算できない」の示し方は三つあり、同値性は定理だが、テーゼは定理ではない」と題していたが、 その内容は論文260「同値性は定理であり、テーゼは定理ではない」と重複していた── 三つの出発点、定理とテーゼの身分の差、アッカーマン関数 A(4,2) の 19729 桁まで一致しており、 先行するのは論文260 である。第2版は計算可能性の主題を全て削除し、 論文260 を参照先とする。入れ替えた主題(断熱不変量)は、 体系の最も重い錨であるネーターの定理(27 編・341 箇所)の隣が空いていたことから選んだ。 作成にあたって:本稿の着想と内容は、著者自身の考察に基づくものです。文章の構成整理や英訳、数式の確認には AI(大規模言語モデル)の助力を得ました。最終的な内容の解釈や誤りがあれば、それらはすべて著者の責に帰します。お気づきの点があれば、ご教示いただければ幸いです。

Yuuki Yamagishi · 0 citations
#small language model Open access Aug 2026

Reproducibility files for: CXRG-SVLM: A Parameter-Efficient Small Vision-Language Model for Multi-View Chest X-ray Report Generation in Resource-Constrained Environments.

This repository contains the official implementation, training and inference scripts, pre-extracted dataset splits (train/val/test), custom clinical entity evaluation metrics, and model weights for the CXRG-SVLM architecture. The pipeline integrates a frozen RAD-DINO medical vision encoder with a 4-bit quantized Qwen2.5-3B-Instruct language model via QLoRA.

Muhammad Fareed, Muhammad Awais Sattar · 0 citations
#small language model Open access Aug 2026

There Is Only One Kind of Constant You Can Ask "Has It Changed?" About ── Oklo Holds It Below 5.0x10^-18 per Year ── Six Decimal Places Beyond the Digit Eddington Argued Over ── [Paper 314]

Paper 10 established that only dimensionless numbers can be fine-tuned. This paper asks whether the same restriction applies to “has it changed”──the answer is it does. And one of them has actually been measured. No new mathematical theorem and no new law is claimed. Scope of this paper (scope note): No new mathematical theorem and no new law is claimed──the Oklo bound on alpha, the definition of the fine-structure constant and Duff’s point that a dimensionful constant’s variation cannot be asked about are all standard. We do not build the theorem──“only dimensionless numbers can be tuned” is Paper 10’s conclusion, received here as a premise. What this paper adds is the measurement side alone. We do not build nuclear physics──we do not enter the calculation from the samarium-149 resonance to the bound on alpha. We quote the resulting number. We do not derive the value of alpha──no attempt is made to obtain 1/alpha from theory. We do not enter the road Paper 101 treated as Eddington’s fall. We do not adjudicate the quasar result──whether Webb et al. are right is not treated. We only count by what factor it disagrees with Oklo. We do not discuss theories of variation──models in which alpha can vary (scalar fields and the like) are not treated. We do not enter anthropic reasoning──that context belongs to Papers 10, 64 and 69; this paper looks only at “has it changed”. Relation to earlier papers: Paper 10 established that only dimensionless numbers can be fine-tuned, showing c, hbar and G to be dimensionful, i. e. choices of unit──this paper takes that theorem as it stands and does not reopen it. It adds one thing only: what measurement actually says. Paper 101 treated as a caution Eddington’s “derivation” of 1/alpha as 136, and his adding one after measurement said 137──this paper places, in those same digits, how far measurement has since reached. Paper 15 computed 1/alpha=137.035999 for itself──this paper uses those digits. Paper 285 showed that “forbidden” can be written in powers of alpha──if alpha moved, those rates would move too. Paper 95 separated convention from fact──this paper’s core is the single point that “has it changed” does not form a sentence on the convention side. What is added is converting the Oklo bound into a per-year rate of 5.0x10^-18, stretching it over the age of the universe to 7.0x10^-8, comparing it with the digit Eddington argued over, and counting the disagreement with the quasar claim as a factor of 114. First, “has c changed” has no truth value. It cannot be told apart from a change of units (Section 2). Second, this is the core of the paper. The Oklo natural reactor holds alpha below 5.0x10^-18 per year (Section 3). Third, stretched over the age of the universe that is 7.0x10^-8 (Section 3). Fourth, Eddington argued over the integer digit; Oklo reaches the sixth decimal (Section 4). Fifth, the quasar claim disagrees with Oklo by a factor of 114 (Section 5). Sixth, the separator is whether a change of units erases it (Section 6). Paper 10 established that only dimensionless numbers can be fine-tuned. The same restriction applies to “has it changed”──“c became 1% smaller” cannot be told apart from making the metre 1% longer, and a sentence asserting one of two indistinguishable things has no truth value. It is not false; it is not a sentence. The question can be put only to dimensionless numbers such as alpha and m_p/m_e──and one of them has actually been measured. The natural fission that ran at Oklo in Gabon two billion years ago bounds the drift of alpha through the samarium-149 resonance at below 5.0x10^-18 per year──stretched over the age of the universe that is only 7.0x10^-8 (assuming a constant rate, and saying nothing about what came before). Eddington argued between 136 and 137 in the integer digit; Oklo reaches the sixth decimal──six orders below. And this paper does not try to derive alpha. It looks only at whether it moved──that self-limitation is its answer to Paper 101’s caution. The quasar claim disagrees with Oklo by a factor of 114──not adjudicated here; they look at different epochs and are compatible if the rate is not constant. One thing separates them──whether a change of units erases it. It is not a matter of accuracy but of the shape of the question, and it is settled before any measurement. Paper 10 settled what can be tuned; this paper says that one of them has been measured and has not moved──beside the theorem, a number. On the making of this work: The ideas and content of this work stem from the author's own considerations. Assistance from an AI (a large language model) was used for structuring, English translation, and checking the algebra. Any remaining errors or misinterpretations are solely the author's. Feedback and corrections are sincerely appreciated. ----- 論文10 は「ファインチューニングできるのは無次元量だけである」と示した。本稿が問うのは、同じ制限が「変化したか」にも掛かるかである──答は、掛かるである。そしてその一つは実際に測られている。新しい数学定理も新しい法則も主張しない。 本稿の射程(射程注記):新しい数学定理も新しい法則も主張しない──オクロ天然原子炉の alpha への制限、微細構造定数の定義、ダフによる「次元を持つ定数の変化は問えない」という指摘は、いずれも標準的である。定理を作らない──「調整できるのは無次元量だけである」は論文10 の結論であり、本稿はそれを前提として受け取る。本稿が加えるのは測定の側だけである。原子核物理を作らない──サマリウム 149 の共鳴から alpha の制限を導く計算には立ち入らない。結果の数を引く。 alpha の値を導かない──1/alpha を理論から出そうとしない。論文101 がエディントンの滑落として扱った道には入らない。クェーサーの結果を判定しない──ウェッブらの主張が正しいかどうかは扱わない。オクロと何倍食い違うかを数えるだけである。時間変化の理論を論じない──alpha が変わりうる模型(スカラー場など)は扱わない。人間原理に立ち入らない──論文10・64・69 が扱った文脈であり、本稿は「変化したか」だけを見る。既刊との関係:論文10 は「ファインチューニングできるのは無次元量だけである」を確立し、c・hbar・G が次元を持つ=単位の選択だと示した──本稿はその定理をそのまま受け取り、蒸し返さない。加えるのは「では実際に測るとどうか」の一点だけである。論文101 はエディントンが 1/alpha を 136 と「導き」、測定が 137 と分かってから足したことを戒めとして扱った──本稿はその同じ桁に、測定がどこまで踏み込んだかを置く。論文15 は 1/alpha=137.035999 を自前で計算した──本稿はその桁を使う。論文285 は「禁じられている」が alpha の冪で書けると示した──alpha が動けばその率も動くという接続がある。論文95 は規約と事実を分けた──本稿の芯は「変わったか」という問いが、規約の側では文にならないという一点である。加えたのはオクロの制限を年あたりの率 5.0x10^-18 に直したこと、それを宇宙年齢に引き伸ばして 7.0x10^-8 と出したこと、エディントンが争った桁と比べたこと、クェーサーの主張との食い違いを 114 倍と数えたことである。 第一に、「c は変わったか」は真偽を持たない。単位の選び方と区別できないからである(第2節)。 第二に、これが本稿の芯である。オクロ天然原子炉が alpha を年あたり 5.0x10^-18 未満に押さえている(第3節)。 第三に、宇宙年齢に引き伸ばしても 7.0x10^-8 である(第3節)。 第四に、エディントンが争ったのは第 3 位、オクロが押さえたのは第 9 位である(第4節)。 第五に、クェーサーの主張はオクロと 114 倍食い違う(第5節)。 第六に、分離子は「単位を変えて消せるか」である(第6節)。 論文10 は「ファインチューニングできるのは無次元量だけである」を確立した。同じ制限が「変わったか」にも掛かる──「c が 1% 小さくなった」はメートルを 1% 長くしたと区別できず、区別できない二つを述べる文は真偽を持たない。偽なのではなく、文になっていない。問える相手は alpha や m_p/m_e のような無次元量だけである──そしてその一つは実際に測られている。ガボンのオクロで 20 億年前に起きた天然の核分裂は、サマリウム 149 の共鳴を通じて alpha の変化を押さえており、年あたり 5.0x10^-18 未満──宇宙年齢まで引き伸ばしても7.0x10^-8にしかならない(率が一定だと仮定した場合であり、それ以前については何も言っていない)。エディントンが 136 か 137 かで争ったのは整数の位で、オクロが押さえているのは小数第 6 位──6 桁下である。そして本稿は alpha の値を導こうとしない。動いたかどうかだけを見る──この自己限定が、論文101 の戒めに対する答である。クェーサーからの主張はオクロと114 倍食い違う──どちらが正しいかは本稿では判定しない。違う時代を見ているので、率が一定でなければ両立しうる。分けるものは一つ──単位を変えて消せるかどうか。測定精度の問題ではなく、問いの形の問題であり、測定の前に決まっている。論文10 が何を調整できるかを確定し、本稿はその一つが実際に測られていて動いていないと書いた──定理の隣に、数がある。 作成にあたって:本稿の着想と内容は、著者自身の考察に基づくものです。文章の構成整理や英訳、数式の確認には AI(大規模言語モデル)の助力を得ました。最終的な内容の解釈や誤りがあれば、それらはすべて著者の責に帰します。お気づきの点があれば、ご教示いただければ幸いです。

Yuuki Yamagishi · 0 citations
#small language model Open access Aug 2026

What Saturates Is Not the Earthquake but the Scale ── Measured by m_b, the 2011 Tohoku Earthquake's Energy Is Estimated at One 5623rd ── A Scale Saturates When the Source Duration Exceeds the Period It Observes ── [Paper 282]

“Magnitude” is not one thing. There are four──M_L, m_b, M_S, M_w──and all but M_w hit a ceiling for great earthquakes. This paper asks whether what hits the ceiling is the earthquake or the scale──the answer is the scale. No new mathematical theorem and no new law is claimed. Scope of this paper (scope note): No new mathematical theorem and no new law is claimed──M_w=frac23log_10M_0-6.07, log_10E=1.5M+4.8, the saturation values of each scale, and the rupture duration of the Tohoku earthquake are all standard. We do not build seismology──all we use is two linear expressions and one division. We do not predict earthquakes──time, place, and size are not treated at all. We do not discuss source processes──neither rupture propagation nor slip distribution is treated. The duration is merely placed as a representative value. We claim no accuracy for the saturation values──M_L ~ 6.8, m_b ~ 6.5, M_S ~ 8.5 are approximate, moving by about +/-0.3 with network and procedure. We do not assert the omega^-2 model──the 1.7501 of Section 5 is an upper bound the model gives, not a value that accounts for the observed 0.60. This paper does not explain the gap between model and observation. We do not say M_w is perfect──M_w carries its own error in estimating M_0; it merely does not saturate. We assert no individual earthquake’s values──the M_w and M_S of the tables are representative values widely used in the literature, differing by about +/-0.1 between agencies. Relation to earlier papers: Paper 117 treated the Gutenberg--Richter power law──that concerns the relation between the number and the size of earthquakes, while this paper concerns the construction of the scale itself. The material is the same earthquakes; the question differs. Paper 190 measured “rare” on a logarithmic scale──this paper likewise treats how a difference of 0.6 on a logarithmic scale becomes a factor of 7.94. Paper 255 separated two things called “accuracy,” only one of which calibration removes──the saturation here is on the side that calibration does not remove, arising from the construction of the scale. Paper 140 separated symmetry fixing ratios from dynamics fixing the scale──this paper treats the case where the scale side breaks. What is added is arranging the four scales by observed period, computing that m_b estimates the Tohoku energy at one 5623rd, confirming that the same procedure is off by only a factor of 1.41 for small earthquakes, and writing the ratio 7.5 of rupture duration to observed period as the separator. First, the four scales observe different periods. M_L at 0.1 s, m_b at 1 s, M_S at 20 s, and M_w choosing no period at all (Section 2). Second, this is the core of the paper. m_b saturates at M ~ 6.5, so measuring the M_w=9.0 Tohoku earthquake with it estimates the energy at one 5623rd (Section 3). Third, even M_S falls short by a factor of 7.94. The difference of 0.6 between M_S=8.4 and M_w=9.0 is a factor of 7.9433 in energy (Section 3). Fourth, it does not happen for small earthquakes. For 1995 Southern Hyogo the difference is 0.10, a factor of only 1.41 in energy (Section 4). Fifth, the cause is the duration of rupture. The Tohoku rupture lasted 150 s, 7.5 times the 20 s that M_S observes (Section 5). Sixth, the separator is whether the source duration exceeds the observed period. M_w alone does not saturate because it chooses no period and measures M_0 directly (Section 6). When magnitude hits a ceiling for great earthquakes, it is not the earthquake that hits the ceiling. Measuring the 2011 Tohoku earthquake with m_b gives 6.5, that is an energy estimated at one 5623rd──of the same event, m_b says moderate and M_w says fourth largest ever recorded. Even M_S falls short by 7.9433──the difference is only 0.6, but it is 0.6 on a logarithmic scale. And it does not happen for small earthquakes──for 1995 Southern Hyogo it stays at 1.4125. The scale is not broken; it is being used outside its range. The cause is the duration of rupture──the Tohoku rupture lasted 150 s, 7.5 times the 20 s that M_S observes. A 20 s wave carries only part of a 150 s event. One thing separates them──whether the source duration exceeds that scale’s observed period. If it does, no amount of calibration removes the saturation. If it does not, the four scales agree well. M_w alone escapes saturation not because it is superior but because it has no period to compare against. On the making of this work: The ideas and content of this work stem from the author's own considerations. Assistance from an AI (a large language model) was used for structuring, English translation, and checking the algebra. Any remaining errors or misinterpretations are solely the author's. Feedback and corrections are sincerely appreciated. ----- 「マグニチュード」は一つではない。 M_L・m_b・M_S・M_w の四つがあり、M_w 以外は大地震で頭打ちになる。本稿が問うのは、頭打ちになっているのは地震か、尺度かである──答は、尺度である。新しい数学定理も新しい法則も主張しない。 本稿の射程(射程注記):新しい数学定理も新しい法則も主張しない──M_w=frac23log_10M_0-6.07、log_10E=1.5M+4.8、各尺度の飽和値、東北地震の破壊継続時間は、いずれも標準的である。地震学を作らない──使うのは二つの一次式と、一つの割り算だけである。地震を予知しない──発生時期も場所も規模も一切扱わない。震源過程を論じない──破壊の伝播も、すべりの分布も扱わない。継続時間を代表値として置くだけである。飽和値の精度を主張しない──M_L ~ 6.8、m_b ~ 6.5、M_S ~ 8.5 はおおよその値であり、観測網と手続きによって +/-0.3 程度動く。 omega^-2 模型を主張しない──第5節の 1.7501 は模型が与える上限であって、実測の 0.60 を説明しきる値ではない。模型と実測の差そのものを、本稿は説明しない。 M_w が完全だと言わない──M_w にも M_0 の推定誤差があり、飽和しないというだけである。個別の地震の値を主張しない──表の M_w・M_S は文献で広く用いられている代表値であり、機関によって +/-0.1 程度異なる。既刊との関係:論文117 はグーテンベルク=リヒターの冪則を扱った──あちらは地震の個数と大きさの関係であり、本稿は尺度そのものの構成である。同じ地震を材料にしているが、問いが違う。論文190 は「稀」を対数の目盛りで測った──本稿も対数目盛りの上で 0.6 という差が 7.94 倍になることを扱う。論文255 は「精度」が二つあり較正で消えるのは一方だけだと分けた──本稿の飽和は較正で消えない側であり、尺度の構成に由来する。論文140 は対称性が比を決め力学が尺度を決めると分けた──本稿は尺度の側が壊れる場合を扱う。加えたのは四つの尺度を観測周期で並べたこと、m_b が東北地震のエネルギーを 5623 分の一に見積もると計算したこと、同じ手続きが小さい地震では 1.41 倍しかずれないと確かめたこと、破壊継続時間と観測周期の比 7.5 を分離子として書いたことである。 第一に、四つの尺度は測る周期が違う。 M_L が 0.1 秒、m_b が 1 秒、M_S が 20 秒、M_w は周期を選ばない(第2節)。 第二に、これが本稿の芯である。 m_b は M ~ 6.5 で頭打ちになるので、M_w=9.0 の東北地震を測るとエネルギーを 5623 分の一に見積もる(第3節)。 第三に、M_S でも 7.94 倍足りない。 M_S=8.4 と M_w=9.0 の差 0.6 は、エネルギーでは 7.9433 倍である(第3節)。 第四に、小さい地震では起きない。1995 年兵庫県南部では差が 0.10、エネルギーで 1.41 倍にとどまる(第4節)。 第五に、原因は破壊の継続時間である。東北の破壊は 150 秒続き、M_S の見る 20 秒の 7.5 倍である(第5節)。 第六に、分離子は「震源時間が観測周期を超えるか」である。 M_w だけが飽和しないのは、周期を選ばず M_0 を直接測るからである(第6節)。 大地震でマグニチュードが頭打ちになるのは、地震が頭打ちになっているのではない。2011 年東北地震を m_b で測ると 6.5、すなわちエネルギーを 5623 分の一に見積もる──同じ地震を、m_b は中規模だと言い、M_w は史上第四位だと言う。 M_S でも 7.9433 倍足りない──差は 0.6 にすぎないが、対数目盛りの上の 0.6 だからである。そして小さい地震では起きない──1995 年兵庫県南部では 1.4125 倍にとどまる。尺度は壊れているのではなく、範囲の外で使われている。原因は破壊の継続時間である──東北の破壊は 150 秒続き、M_S の見る 20 秒の 7.5 倍だった。20 秒の波は、150 秒の出来事の一部しか運ばない。分けるものは一つ──震源の継続時間が、その尺度の観測周期を超えているかどうか。超えていれば、どんなに較正しても飽和は消えない。超えていなければ、四つの尺度はよく一致する。 M_w だけが飽和しないのは優れているからではなく、比べるべき周期を持たないからである。 作成にあたって:本稿の着想と内容は、著者自身の考察に基づくものです。文章の構成整理や英訳、数式の確認には AI(大規模言語モデル)の助力を得ました。最終的な内容の解釈や誤りがあれば、それらはすべて著者の責に帰します。お気づきの点があれば、ご教示いただければ幸いです。

Yuuki Yamagishi · 0 citations
#small language model Open access Aug 2026

One Number Sets the Limit of Forecasting ── Each Extra Day Costs 1.5874 Times the Initial Accuracy ── Observe 10 Times More Precisely and You Gain Only 4.98 Days ── [Paper 289]

That a weather forecast cannot reach beyond a certain horizon is due neither to missing equations nor to slow computers. This paper asks what sets the limit──the answer is one number, the error doubling time tau_d. No new mathematical theorem and no new law is claimed. Scope of this paper (scope note): No new mathematical theorem and no new law is claimed──exponential error growth, the doubling time, the limit of predictability, and the value tau_dapprox 1.5 days are all standard. We do not build meteorology──all we use is one exponential and its inverse. We do not discuss chaos──the Lorenz equations, attractors, and bifurcations are not treated at all. We do not discuss numerical weather prediction──grid resolution, parameterisation, and data assimilation are not treated. We do not say the error grows exactly exponentially──e^lambda t holds only while the error is small, and growth stops near saturation. The computations here are confined to the linear-growth regime. We assert no value for tau_d──1.5 days is a representative value widely used in the literature, and it moves from about 1 to 2.5 days with season, region, and variable. Section 5 shows the size of that dependence itself. We do not say there is a single exponent──the real atmosphere has different growth rates at different scales, with smaller eddies growing faster. A single tau_d is a crude approximation. We do not deny that forecasts improve──forecasts have in fact grown longer. What this paper says is only that the growth is logarithmic, not that improvement is pointless. Relation to earlier papers: Paper 253 showed that time is one-dimensional because prediction demands it, not because a law says so──this paper turns how far that demand can be met into a number. Paper 195 separated “stable” into six words──that paper is a classification of stability; this one is a time scale of predictability, the same hyperbolicity as material with a different question. Paper 190 measured “rare” on a logarithmic scale──the return here is likewise logarithmic. Paper 266 showed that the premise of the sampling theorem is never met──“knowing the initial state exactly” here is likewise a premise never met, the same figure. What is added is writing the price per day as the fixed factor 1.5874, computing the accuracy needed for 14->21->30->60 days as 25.40 / 1625.5 / 1.70x10^9, writing backwards that 10 times the observation gains only 4.98 days, and sweeping tau_d from 1.0 to 2.5 to show the answer moving from 65536 to 84.4. First, the price per day is a fixed factor. With tau_d=1.5 days, each extra day costs 1.5874 times the initial accuracy (Section 2). Second, this is the core of the paper. Going from 14 to 21 days costs 25.40 times; to 30 days, 1625.5 times; to 60 days, 1.70x10^9 times (Section 2). Third, read backwards, the return is logarithmic. Observing 10 times more precisely gains only 4.98 days (Section 3). Fourth, even 10^9 times gains only 44.85 days (Section 3). Fifth, the familiar “about two weeks” comes from here. If the initial error is 10^-3 of saturation, the forecastable span is 14.95 days (Section 4). Sixth, the separator is tau_d itself. At tau_d=1.0 day the same extension costs 65536 times; at 2.5 days only 84.4──everything rides on one number (Section 5). What sets the limit of forecasting is neither the equations nor the computers, but one number, the error doubling time tau_d. At tau_d=1.5 days, each extra day costs 1.5874 times the initial accuracy──the factor is the same wherever the day is added, but the extension adds while the price multiplies, so one week costs 25.4, two weeks 645, six weeks 1.7 billion. Read backwards, observing 10 times more precisely gains only 4.98 days, and even 10^9 times gains 44.85. The familiar “about two weeks” comes from this one line──14.95 days at an initial error of 10^-3 of saturation. One thing separates them──tau_d itself. At 1.0 day the same extension costs 65536; at 2.5 days, 84.4. A factor of 776 arises from a single number. So the work of extending forecasts and the work of measuring tau_d carry the same weight. On the making of this work: The ideas and content of this work stem from the author's own considerations. Assistance from an AI (a large language model) was used for structuring, English translation, and checking the algebra. Any remaining errors or misinterpretations are solely the author's. Feedback and corrections are sincerely appreciated. ----- 天気予報がある日数より先を当てられないのは、方程式が足りないからでも、計算機が遅いからでもない。本稿が問うのは、何が限界を決めているかである──答は、誤差の二重時間 tau_d という一つの数である。新しい数学定理も新しい法則も主張しない。 本稿の射程(射程注記):新しい数学定理も新しい法則も主張しない──誤差の指数増大、二重時間、予測可能性の限界、tau_dapprox 1.5 日という値は、いずれも標準的である。気象学を作らない──使うのは一つの指数関数と、その逆関数だけである。カオスを論じない──ローレンツ方程式も、アトラクタも、分岐も一切扱わない。数値予報を論じない──格子解像度も、パラメタリゼーションも、データ同化も扱わない。誤差が厳密に指数増大すると言わない──e^lambda t が成り立つのは誤差が小さいあいだだけであり、飽和に近づけば増大は止まる。本稿の計算は線形増大の領域に限る。 tau_d の値を主張しない──1.5 日は文献で広く用いられる代表値であり、季節・領域・変数によって 1 日から 2.5 日程度まで動く。第5節はこの依存の大きさそのものを示す。単一の指数だと言わない──実際の大気には尺度ごとに違う成長率があり、小さい渦ほど速く育つ。単一の tau_d は粗い近似である。予報の改善を否定しない──現に予報は延びてきた。本稿が言うのはその延び方が対数的であるということだけであり、改善が無意味だとは言わない。既刊との関係:論文253 は時間が一本なのが法則ではなく「予言できる」という要求だと示した──本稿はその要求が、どこまでなら満たせるかを数にする。論文195 は「安定」が六つの別の言葉だと分けた──あちらは安定性の分類、本稿は予測可能性の時間尺度であり、同じ双曲性を材料にして問いが違う。論文190 は「稀」を対数の目盛りで測った──本稿の見返りも対数である。論文266 は標本化定理の前提が決して満たされないと示した──本稿の「初期値を正確に知る」も決して満たされない前提であり、構図が同じである。加えたのは一日あたりの代償を 1.5874 倍という一定倍率として書いたこと、14->21->30->60 日の必要精度を 25.40/1625.5/1.70x10^9 倍と計算したこと、観測 10 倍が 4.98 日にしかならないと逆から書いたこと、tau_d を 1.0 から 2.5 まで振って答が 65536 倍から 84.4 倍まで動くと示したことである。 第一に、一日ごとの代償は一定倍率である。 tau_d=1.5 日なら、一日延ばすたびに初期値の精度が 1.5874 倍要る(第2節)。 第二に、これが本稿の芯である。14 日を 21 日にするのに 25.40 倍、30 日にするのに 1625.5 倍、60 日にするのに 1.70x10^9 倍(第2節)。 第三に、逆から見ると見返りは対数的である。観測を 10 倍精密にしても、延びるのは 4.98 日だけである(第3節)。 第四に、10 億倍にしても 44.85 日である(第3節)。 第五に、約二週間という数がここから出る。初期誤差が飽和の 10^-3 なら、予報可能な期間は 14.95 日(第4節)。 第六に、分離子は「指数か多項式か」である。 tau_d を 1.0 日にすると同じ延長に 65536 倍要り、2.5 日なら 84.4 倍で済む──すべてが一つの数に乗っている(第5節)。 予報の限界を決めているのは、方程式でも計算機でもなく、誤差の二重時間 tau_d という一つの数である。 tau_d=1.5 日なら、一日延ばすたびに初期値の精度が 1.5874 倍要る──どこで延ばしても倍率は同じだが、延長は足し算で、代償は掛け算なので、一週間で 25.4 倍、二週間で 645 倍、一か月半で 17 億倍になる。逆から見れば、観測を 10 倍精密にしても延びるのは 4.98 日であり、10 億倍にしても 44.85 日である。よく言われる「約二週間」も、この一行から出る──初期誤差が飽和の 10^-3 なら 14.95 日。分けるものは一つ──tau_d そのもの。1.0 日なら同じ延長に 65536 倍要り、2.5 日なら 84.4 倍で済む。776 倍の違いが、たった一つの数から生まれる。だから予報を延ばす仕事と、tau_d を測る仕事は、同じ重さを持っている。 作成にあたって:本稿の着想と内容は、著者自身の考察に基づくものです。文章の構成整理や英訳、数式の確認には AI(大規模言語モデル)の助力を得ました。最終的な内容の解釈や誤りがあれば、それらはすべて著者の責に帰します。お気づきの点があれば、ご教示いただければ幸いです。

Yuuki Yamagishi · 0 citations
#small language model Open access Aug 2026

Executable Memory and World Coupling: Code as Cognitive Interface in a Self-Modifying Simulation

Large language model (LLM) agents have recently explored executable memory—compiling agent memory into code snippets that an external LLM interprets at inference time. We argue that this paradigm remains tied to a single architectural choice: the executor is an external model, the memory is a personal profile, and the code never participates in the agent's own memory economy. We present a cognitive simulation engine in which executable code is stored as unit-level memory entries and executed by a deterministic rule engine inside the simulation itself. A memory entry carrying an EXPR: prefix is a small program—an arithmetic expression over engine parameters and state variables—interpreted each generation; its result feeds directly into the unit's behavioral circuits. Code memory participates in the engine's memory economy: entries decay, are reinforced by hits, are evicted by capacity limits, and pass the same verification gates as any mechanism. Units acquire executable fragments by foraging, coupling energy gain with behavioral information transfer. Experiments show that (i) code memory measurably alters survival dynamics (extinction-count growth reduced by roughly 97% at threat 1.0); (ii) the survival benefit of code is stratified by strategy—decay reinforcement confers +15 generations at threat 1.5, healing reinforcement +10, while aggressive threat clearance confers no gain (clearing danger memories also clears the fear that drives defensive behavior); (iii) beyond a critical threat intensity (3.0) no code strategy confers benefit—a measured capability boundary; (iv) external trigger coupling: a unit's code can read an external trigger state (cognition) and, when the external signal is present, deterministically clear its own threat memories while writing an externally observable trace—with the external signal absent, the same code is inert, demonstrating that perception is a necessary component of the response; (v) cognitive code is acquired, not inherited: newly born units without the code fragment cannot perceive the external state, making cognition an evolvable individual trait; and (vi) when defensive and adversarial code coexist, an arms race emerges from primitive operations alone. We also report an unexpected semantics of negative-valued code, its diagnosis, and its redesign as a candidate inhibitory mechanism. The architecture points toward self-modifying systems in which memory, behavior, perception, and robustness converge on a single executable substrate.

Yizhang Hu · 0 citations
#small language model Open access Aug 2026

Can AI Replace Humans? — How Human Knowledge Continues to Grow in the Era of Large Language Models

Recommendation algorithms determine what people see, and may also influence the perspectives through which people understand the world. As Large Language Models (LLMs) enter the domains of knowledge acquisition and information comprehension, this influence may extend further into human understanding, judgment, and modes of thinking. If people rely long-term on a small number of general-purpose LLMs, a new homogenization of knowledge sources and modes of interpretation may emerge. A possible direction is the joint development of general-purpose LLMs and vertical small models: general models provide breadth, while vertical models leverage industry-specific and professional data to provide depth and novelty. LLMs as knowledge tools do not imply that humans will be replaced. LLMs provide the knowledge foundation and computational power; humans provide direction. In this collaborative process, individuals internalize knowledge and continuously push toward the unknown through judgment, reasoning, verification, reorganization, and Global Self-Consistency. When a problem reaches the point where existing knowledge can no longer explain it, new thinking emerges naturally. As more and more individuals explore in different directions, innovation points accumulate and connect, ultimately forming a new knowledge foundation. When there are enough points, they connect into surfaces; when there are enough surfaces, they form a new knowledge foundation.

ling liu · 0 citations
#small language model Open access Aug 2026

Drude Was Right Because Two Errors Cancelled ── The Heat Capacity Too Large by 82.68 and the Squared Speed Too Small by 180.70 ── Their Product Is 2.1856, Against the Actual Discrepancy of 2.1932 ── [Paper 274]

Drude (1900) obtained the ratio of thermal to electrical conductivity in a metal from classical theory alone, and got a value close to measurement. This paper asks why he was right──the answer is not that the premises were correct. Two premises erred largely in opposite directions and very nearly cancelled on taking a ratio. No new mathematical theorem and no new law is claimed. Scope of this paper (scope note): No new mathematical theorem and no new law is claimed──the Wiedemann-Franz law, the Drude model, Sommerfeld's Lorenz number, and the expressions for the Fermi speed and electronic heat capacity are all standard. No solid-state physics is built──what is used is two ratios and one product. The Drude model is not re-derived──L=frac32(k/e)^2 is cited only, and the relaxation-time approximation is not entered. The Sommerfeld expansion is not treated──the origin of the pi^2/3 is cited. Failures of the Wiedemann-Franz law are not treated──that it fails at low temperature and under inelastic scattering is not treated. This paper looks at one point near room temperature. No precision is claimed for the measured values──the values for copper and silver are representative and vary with temperature and purity. The 82.68 and 180.70 are not claimed as precise──they are numbers for seeing orders of magnitude, obtained by inserting representative values of the Fermi temperature and speed. That the agreement fell within 0.35% is partly because those representative values happened to sit well. Sommerfeld is not confused with another work──the 4 occurrences in earlier papers are Bohr-Sommerfeld quantisation (Paper 180), signal velocity (Paper 210) and a biographical mention (Paper 125), not the free-electron model. Relation to earlier papers: Paper 265 showed that what separated the two low-temperature models is the density of states──this paper also treats an electronic heat capacity, but asks after the cancellation of errors. Paper 255 showed that accuracy is two things and calibration removes only one──this paper treats a case where two errors survive as a product. Paper 190 measured rare on a logarithmic scale──this paper likewise writes the error as a factor. Paper 271 counted how far mean field errs by dimension──this paper counts two errors that cancel. Paper 232 counted the standings of one equals sign──this paper asks after the standing of “he was right”. What is added is putting the excess of the heat capacity at 82.68 and the shortfall of the squared speed at 180.70, confirming that their product 2.1856 agrees with the actual discrepancy 2.1932 to 0.35%, and writing that without taking a ratio classical theory is out by orders of magnitude. First, set the two answers out. Drude gives L=frac32(k/e)^2=1.11388x10^-8 and Sommerfeld L=(pi^2)/(3)(k/e)^2=2.44300x10^-8 (Section 2). Second, measurement lies near the latter. Copper 2.23x10^-8 and silver 2.31x10^-8──Drude is about half (Section 2). Third, this is the core of the paper. For the heat capacity, classical theory is 82.68 times too large──because only the fraction kT/E_F of the electrons contributes, and it took all of them to (Section 3). Fourth, the other errs the other way. For the squared speed, classical theory is 180.70 times too small──because what carries the heat is not the thermal speed but the Fermi speed (Section 3). Fifth, the product remains.180.70/82.68=2.1856──agreeing with the actual ratio of the two Lorenz numbers, 2.1932, to 0.35%(Section 4). Sixth, the separator is taking a ratio. Look at kappa alone or at sigma alone and classical theory is out by orders of magnitude──only on taking the ratio do the errors cancel down to a factor of about 2 (Section 5). Drude was right not because the premises were correct. Two premises erred by two orders of magnitude each, in opposite directions, and cancelled on taking a ratio──the heat capacity too large by 82.68, the squared speed too small by 180.70, and their product 2.1856. That agrees with the actual ratio of the two Lorenz numbers, 2.1932, to 0.35%. And both errors came from one and the same oversight──that electrons obey Fermi statistics. One oversight moved two quantities in opposite directions, and they met inside the ratio and vanished. The separator is taking a ratio──cease to take it and, with the same theory, the error appears as a factor of 82.68. One thing separates them──writing down which quantity one means when one says he was right. Write it down, and being right about the ratio separates from being out by two orders about the heat capacity. Do not write it down, and one reads a fortunate cancellation as evidence that the premises were sound. On the making of this work: The ideas and content of this work stem from the author's own considerations. Assistance from an AI (a large language model) was used for structuring, English translation, and checking the algebra. Any remaining errors or misinterpretations are solely the author's. Feedback and corrections are sincerely appreciated. ----- ドルーデ(1900)は古典論だけで金属の熱伝導と電気伝導の比を出し、実測に近い値を得た。本稿が問うのは、なぜ当たったのかである──答は、前提が正しかったからではない。二つの前提が逆向きに大きく外れ、比を取ったときにほとんど打ち消し合ったからである。新しい数学定理も新しい法則も主張しない。 本稿の射程(射程注記):新しい数学定理も新しい法則も主張しない──ヴィーデマン=フランツ則、ドルーデ模型、ゾンマーフェルトのローレンツ数、フェルミ速度と電子比熱の表式は、いずれも標準的である。固体物理を作らない──使うのは二つの比と、一つの積だけである。ドルーデ模型を再導出しない──L=frac32(k/e)^2 を引くだけであり、緩和時間近似の中身に立ち入らない。ゾンマーフェルト展開を扱わない──pi^2/3 の由来は引用である。ヴィーデマン=フランツ則の破れを扱わない──低温や非弾性散乱で破れることは扱わない。本稿は室温付近の一点だけを見る。実測値の精度を主張しない──銅と銀の値は代表値であり、温度や純度によって幅がある。82.68 と 180.70 を精密な値として主張しない──フェルミ温度と速度に代表値を入れて出した桁を見るための数である。一致が 0.35% に収まったのは、代表値がよく揃っていたためでもある。ゾンマーフェルトを別人と混同しない──既刊の「ゾンマーフェルト」4 件はボーア=ゾンマーフェルト量子化(論文180)と信号速度(論文210)と伝記(論文125)であり、自由電子模型ではない。既刊との関係:論文265 は低温比熱で二つの模型を分けたのが状態密度だと示した──本稿も電子比熱を扱うが、問うのは誤りの打ち消しである。論文255 は「精度」が二つあり較正で消えるのは一方だけだと示した──本稿は二つの誤差が積で残る場合を扱う。論文190 は「稀」を対数の目盛りで測った──本稿も外れ方を倍率で書く。論文271 は平均場の外れ方を次元ごとに数えた──本稿は外れ方が二つあって打ち消し合う場合を数える。論文232 は同じ等号の身分を数えた──本稿は「当たった」の身分を問う。加えたのは比熱の過大を 82.68 倍、速度の二乗の過小を 180.70 倍と数で出したこと、その積 2.1856 が実際のずれ 2.1932 と 0.35% で一致することを確かめたこと、比を取らなければ古典論が桁で外れると書いたことである。 第一に、二つの答を並べる。ドルーデは L=frac32(k/e)^2=1.11388x10^-8、ゾンマーフェルトは L=(pi^2)/(3)(k/e)^2=2.44300x10^-8(第2節)。 第二に、実測は後者に近い。銅 2.23x10^-8、銀 2.31x10^-8──ドルーデは半分である(第2節)。 第三に、これが本稿の芯である。比熱について、古典論は 82.68 倍過大である──kT/E_F の割合の電子しか効かないのに、全部が効くとしたからである(第3節)。 第四に、もう一つは逆向きに外れる。速度の二乗について、古典論は 180.70 倍過小である──実際に運ぶのは熱速度ではなくフェルミ速度だからである(第3節)。 第五に、積が残る。180.70/82.68=2.1856──二つのローレンツ数の実際の比 2.1932 と、0.35% で一致する(第4節)。 第六に、分離子は「比を取ること」である。 kappa だけ、sigma だけを見れば古典論は桁で外れる──比を取ったときにだけ、誤りが約分されて 2 倍程度に縮む(第5節)。 ドルーデが当たったのは、前提が正しかったからではなかった。二つの前提が逆向きに二桁ずつ外れ、比を取ったときに打ち消し合ったからである──比熱は 82.68 倍過大、速度の二乗は 180.70 倍過小で、積は 2.1856。二つのローレンツ数の実際の比 2.1932 と、0.35% で一致する。そして二つの誤りは同じ一つの見落としから出ていた──電子がフェルミ統計に従うことである。一つの見落としが、二つの量を逆向きに動かし、比の中で出会って消えた。分離子は比を取ることである──比を取るのをやめれば、同じ理論のまま誤りが 82.68 倍として現れる。分けるものは一つ──どの量について「当たった」と言っているのかを書き出すこと。書き出せば、比については当たり、比熱については二桁外していることが分かる。書き出さなければ、打ち消し合った偶然を、前提の正しさの証拠だと読んでしまう。 作成にあたって:本稿の着想と内容は、著者自身の考察に基づくものです。文章の構成整理や英訳、数式の確認には AI(大規模言語モデル)の助力を得ました。最終的な内容の解釈や誤りがあれば、それらはすべて著者の責に帰します。お気づきの点があれば、ご教示いただければ幸いです。

Yuuki Yamagishi · 0 citations
#small language model Open access Aug 2026

"Forbidden" Means "Slower by a Factor of 7.62x10^7" ── Hydrogen's 2s->1s Does Not Fail to Happen; It Takes 0.1216 Seconds ── The 21 cm Line Takes 11 Million Years, Yet 10^60 Atoms Emit 2.9x10^45 Photons a Second ── [Paper 285]

A selection rule is written as “this transition is forbidden.”Yet forbidden transitions are actually observed. This paper asks what “forbidden” means as a number──the answer is a matter of rate. No new mathematical theorem and no new law is claimed. Scope of this paper (scope note): No new mathematical theorem and no new law is claimed──the E1 selection rules, two-photon emission, hyperfine structure, the lifetime of the 21 cm line, and the alpha^2 estimate are all standard. We do not build quantum electrodynamics──all we use is three lifetimes and two divisions. We do not derive the selection rules──we do not enter the Wigner--Eckart theorem. That is the side of Paper 140. We do not compute transition probabilities──1.596x10^-9, 0.1216, and 3.47x10^14 are merely quoted from the literature, and this paper only takes ratios. We do not say alpha^2 is an exact ratio──the actual E1-to-M1 ratio moves by orders with the matrix elements. alpha^2=5.325x10^-5 is an estimate, not a prediction. We do not say selection rules are meaningless──a factor of 7.6x10^7 is indistinguishable from prohibition in practice in many settings. What this paper says is only that “forbidden” and “slow” have different standings. We do not discuss interstellar physics──the number 10^60 is a representative value for showing orders, and cloud mass, temperature, and excitation mechanism are not treated at all. Relation to earlier papers: Paper 140 treated selection rules from the group-theoretic side, showing that symmetry fixes ratios──this paper is on the rate side, writing the same selection rules as lifetimes. It treats the place where a theorem’s zero set is not in fact zero. Paper 190 measured “rare” on a logarithmic scale──the 23.34 decades here must likewise be read logarithmically. Paper 272 showed that one and the same “doubling” opens by a factor of 6.70 on the stimulus side──this paper is the case where one and the same electromagnetic transition opens by 2.17x10^23 on the rate side. Paper 269 showed that what a crystal forbids is not the symmetry but the periodicity──this paper likewise asks what the word “forbid” contains. There the prohibition is exact; here it is a matter of rate, and the two stand in contrast. What is added is lining up three transition lifetimes in one table and measuring a spread of 2.17x10^23, writing the 7.6190x10^7 of 2s->1s as the numerical content of “prohibition”, filling in why the 21 cm line is visible as “10^60 atoms give 2.88x10^45 per second”, and placing as the separator that what goes to zero is a matrix element, not a rate. First, we compare two lifetimes in the hydrogen atom.2p->1s (E1 allowed) is 1.596x10^-9 s; 2s->1s (E1 forbidden) is 0.1216 s (Section 2). Second, this is the core of the paper. The ratio is 7.6190x10^7──it is not forbidden. It is merely slow (Section 2). Third, the estimate can be written with the fine-structure constant. The E1-to-M1 ratio is alpha^2=5.325x10^-5, and two-photon is alpha^4=2.836x10^-9 (Section 3). Fourth, the 21 cm line is slower still. Lifetime 3.47x10^14 s =11 million years, 2.17x10^23 times that of 2p (Section 4). Fifth, it is nevertheless visible because of numbers. With 10^60 hydrogen atoms, 2.88x10^45 emit each second (Section 5). Sixth, the separator is “exactly zero, or merely small”. What a selection rule sets to zero is a matrix element within one approximation, not the total rate (Section 6). “Forbidden,” as a number, means a matter of rate. Hydrogen’s 2s->1s is E1 forbidden, yet it does not fail to happen; it takes 0.1216 s──7.6190x10^7 times, that is 7.88 decades, slower than 2p->1s. The 21 cm line is slower still, once in 11 million years, 2.17x10^23 times that of 2p──and it is nevertheless visible not because the rate rose but because the count is large: 10^60 atoms emit 2.88x10^45 photons a second. One thing separates them──that what is exactly zero is not the rate but a matrix element within one approximation. The selection rule speaks about that matrix element; the experiment measures the total rate. The zero is a property of the inside of the E1 approximation, not a property of nature. So when a “forbidden” line is seen in the sky, no contradiction has occurred. On the making of this work: The ideas and content of this work stem from the author's own considerations. Assistance from an AI (a large language model) was used for structuring, English translation, and checking the algebra. Any remaining errors or misinterpretations are solely the author's. Feedback and corrections are sincerely appreciated. ----- 選択則は「この遷移は禁じられている」と書かれる。しかし禁じられた遷移は現に観測されている。本稿が問うのは、「禁じられている」が数として何を意味するかである──答は、率の大小である。新しい数学定理も新しい法則も主張しない。 本稿の射程(射程注記):新しい数学定理も新しい法則も主張しない──E1 選択則、二光子放出、超微細構造、21 cm 線の寿命、alpha^2 という目安は、いずれも標準的である。量子電磁力学を作らない──使うのは三つの寿命と、二つの割り算だけである。選択則を導出しない──ウィグナー=エッカートの定理には立ち入らない。それは論文140 の側である。遷移確率を計算しない──1.596x10^-9 も 0.1216 も 3.47x10^14 も文献値を引くだけであり、本稿は比を取るのみである。 alpha^2 を厳密な比だと言わない──E1 と M1 の実際の比は行列要素によって桁で動く。 alpha^2=5.325x10^-5 は目安であって、予言値ではない。選択則が無意味だと言わない──7.6x10^7 倍という差は実務上は禁止と変わらない場面が多い。本稿が言うのは「禁止」と「遅い」は身分が違うということだけである。星間物質の物理を論じない──10^60 という数は桁を示すための代表値であり、雲の質量も温度も励起機構も一切扱わない。既刊との関係:論文140 は選択則を群論の側から扱い、対称性が比を決めることを示した──本稿は率の側であり、同じ選択則を寿命の数で書く。定理の零集合が、実際には零でないところを扱う。論文190 は「稀」を対数の目盛りで測った──本稿の 23.34 桁も対数で読まねばならない。論文272 は同じ「二倍」が刺激の側で 6.70 倍ひらくことを示した──本稿は同じ一つの電磁遷移が、率の側で 2.17x10^23 倍ひらく場合である。論文269 は結晶が禁じているのが対称性でなく周期性だと示した──本稿も「禁じる」という語の中身を問う。あちらは厳密な禁止、こちらは率の大小であり、対照になっている。加えたのは三つの遷移の寿命を一表に並べて 2.17x10^23 倍の開きを測ったこと、2s->1s の 7.6190x10^7 倍を「禁止」の数値的な中身として書いたこと、21 cm 線が見える理由を「10^60 個で毎秒 2.88x10^45 個」と個数で埋めたこと、零になるのが率でなく行列要素だと分離子に据えたことである。 第一に、水素原子で二つの寿命を比べる。2p->1s(E1 許容)は 1.596x10^-9 秒、2s->1s(E1 禁制)は 0.1216 秒(第2節)。 第二に、これが本稿の芯である。比は 7.6190x10^7 倍──禁じられてはいない。遅いだけである(第2節)。 第三に、目安は微細構造定数で書ける。 E1 と M1 の比は alpha^2=5.325x10^-5、二光子は alpha^4=2.836x10^-9(第3節)。 第四に、21 cm 線はさらに遅い。寿命 3.47x10^14 秒 =1100 万年、2p の 2.17x10^23 倍(第4節)。 第五に、それでも見えるのは個数のためである。水素原子が 10^60 個あれば、毎秒 2.88x10^45 個が光る(第5節)。 第六に、分離子は「厳密に零か、小さいだけか」である。選択則が零にするのは一つの近似における行列要素であって、全体の率ではない(第6節)。 「禁じられている」は、数として率の大小を意味する。水素の 2s->1s は E1 禁制だが、起きないのではなく 0.1216 秒かかる──2p->1s の 7.6190x10^7 倍、すなわち 7.88 桁遅いだけである。21 cm 線はさらに遅く、1100 万年に一度、2p の 2.17x10^23 倍である──それでも見えるのは率が上がったからではなく、個数が大きいからであり、10^60 個あれば毎秒 2.88x10^45 個が光る。分けるものは一つ──厳密に零になっているのは、率ではなく、一つの近似における行列要素だということ。選択則はその行列要素について語り、実験は全体の率を測る。零は E1 という近似の内側の性質であって、自然の性質ではない。だから「禁じられた」線が空に見えても、矛盾は起きていない。 作成にあたって:本稿の着想と内容は、著者自身の考察に基づくものです。文章の構成整理や英訳、数式の確認には AI(大規模言語モデル)の助力を得ました。最終的な内容の解釈や誤りがあれば、それらはすべて著者の責に帰します。お気づきの点があれば、ご教示いただければ幸いです。

Yuuki Yamagishi · 0 citations

From tech blogs

See all →
Microsoft Research Blog Aug 31, 2026

GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models

What if pathology foundation models could do more with less? GigaPath-Flash and GigaTIME-Flash cut computational demands while maintaining strong performance, opening the door to larger studies and broader exploration. The post GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models appeared first on Microsoft Research.