<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Awareness Software Group research</title>
    <link>https://awarenesssoftwaregroup.com/research/</link>
    <description>Open research on how AI models actually learn. One finding per entry, in plain English first and in full technical detail underneath, including the approaches that turned out not to work.</description>
    <language>en</language>
    <atom:link href="https://awarenesssoftwaregroup.com/feed.xml" rel="self" type="application/rss+xml" />
    <lastBuildDate>Mon, 28 Sep 2026 02:57:23 +0000</lastBuildDate>
    <item>
      <title>A Clock, Not a Warning</title>
      <link>https://awarenesssoftwaregroup.com/research/a-clock-not-a-warning/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-clock-not-a-warning/</guid>
      <description>Moving the moment a model learns, its approach to the edge of chaos moved too, but always a fixed fraction of the way there: it tracks training progress,</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Boost Then Decay</title>
      <link>https://awarenesssoftwaregroup.com/research/boost-then-decay/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/boost-then-decay/</guid>
      <description>A learning-rate schedule that speeds up early and slows down late made small models learn as fast as our boost and end as accurate as the standard</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Early Towards the Edge</title>
      <link>https://awarenesssoftwaregroup.com/research/early-towards-the-edge/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/early-towards-the-edge/</guid>
      <description>A small model&#x27;s recurrence moves from very stable towards the edge of chaos early in training, half-way at the same step in every run, long before it</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Fast or Accurate</title>
      <link>https://awarenesssoftwaregroup.com/research/fast-or-accurate/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/fast-or-accurate/</guid>
      <description>Our short early learning-rate boost made small models learn fastest; the standard warm-up-and-decay schedule made them end the most accurate.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Grow the Model, Save a Quarter</title>
      <link>https://awarenesssoftwaregroup.com/research/grow-the-model-save-a-quarter/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/grow-the-model-save-a-quarter/</guid>
      <description>Starting a model small and doubling it twice reached the target with 26 percent less compute than training at full size, and when to grow barely mattered.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Growth Did Not Travel</title>
      <link>https://awarenesssoftwaregroup.com/research/growth-did-not-travel/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/growth-did-not-travel/</guid>
      <description>The model-growth recipe that saved a quarter of the compute on one task rarely reached the target on a second, slower task, because its timings were set</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Is Learning a Lucky Accident?</title>
      <link>https://awarenesssoftwaregroup.com/research/is-learning-a-lucky-accident/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/is-learning-a-lucky-accident/</guid>
      <description>Testing whether sudden learning is randomness knocking a model out of a rut.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>It Tries The Wrong Rule First</title>
      <link>https://awarenesssoftwaregroup.com/research/it-tries-the-wrong-rule-first/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/it-tries-the-wrong-rule-first/</guid>
      <description>Instead of asking how often the model is right, we asked what it is doing. Early on it behaves like a simpler rule that does not work.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Late, Not Delayed</title>
      <link>https://awarenesssoftwaregroup.com/research/late-not-delayed/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/late-not-delayed/</guid>
      <description>Run three times longer with little data, a small model learned the rule late but always while fitting its examples, never after. No grokking.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Learning It Again Is Not Like Learning It</title>
      <link>https://awarenesssoftwaregroup.com/research/learning-it-again-is-not-like-learning-it/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/learning-it-again-is-not-like-learning-it/</guid>
      <description>We changed the task underneath a trained model and watched it learn again.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Memorised, Never Generalised</title>
      <link>https://awarenesssoftwaregroup.com/research/memorised-never-generalised/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/memorised-never-generalised/</guid>
      <description>Trained repeatedly on a small fixed set of addition problems, a small model memorised them in a hundred steps and never solved a new one in ten thousand.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>No Delay at the Edge</title>
      <link>https://awarenesssoftwaregroup.com/research/no-delay-at-the-edge/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/no-delay-at-the-edge/</guid>
      <description>With enough training data to learn the rule at all, a small model learned it while fitting its data, never long after.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Only as Good as Its Forecast</title>
      <link>https://awarenesssoftwaregroup.com/research/only-as-good-as-its-forecast/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/only-as-good-as-its-forecast/</guid>
      <description>A free formula timed a training intervention nearly as well as watching the run, for model sizes inside its range, and no better than chance at the edges.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Research questions</title>
      <link>https://awarenesssoftwaregroup.com/research/questions/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/questions/</guid>
      <description>The 10 questions our public AI training research is organised around, each with its current answer and the published results behind it.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Frame Decides the Sign</title>
      <link>https://awarenesssoftwaregroup.com/research/the-frame-decides-the-sign/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-frame-decides-the-sign/</guid>
      <description>Whether bigger models reorganise less while learning depends on which of two standard measures you use: one rises with size, the other falls.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Habit It Did Not Need</title>
      <link>https://awarenesssoftwaregroup.com/research/the-habit-it-did-not-need/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-habit-it-did-not-need/</guid>
      <description>Early on the model falls into a simple wrong habit. We took the habit away.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Scoring Made The Steps</title>
      <link>https://awarenesssoftwaregroup.com/research/the-scoring-made-the-steps/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-scoring-made-the-steps/</guid>
      <description>We thought our way of keeping score was blurring distinct stages of learning together. It turned out to be creating them.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Training Curve Cannot Tell You</title>
      <link>https://awarenesssoftwaregroup.com/research/the-training-curve-cannot-tell-you/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-training-curve-cannot-tell-you/</guid>
      <description>Two models with near-identical training curves, one building machinery and one polishing machinery it was handed. Only looking inside separates them.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>We Gave It The Answer Anyway</title>
      <link>https://awarenesssoftwaregroup.com/research/we-gave-it-the-answer-anyway/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/we-gave-it-the-answer-anyway/</guid>
      <description>We wired a working shortcut straight into the model.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Critical Amount Per Step</title>
      <link>https://awarenesssoftwaregroup.com/research/a-critical-amount-per-step/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-critical-amount-per-step/</guid>
      <description>Learning time follows one curve in material per training step, and the standard critical batch size formula fits it within 2.2 percent across an 86-fold</description>
      <pubDate>Sun, 27 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Rise, Not a Peak</title>
      <link>https://awarenesssoftwaregroup.com/research/a-rise-not-a-peak/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-rise-not-a-peak/</guid>
      <description>Five independently trained copies of a public language model all reorganise their internal state as they learn to copy, then level off.</description>
      <pubDate>Sun, 27 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Speed-Accuracy Frontier</title>
      <link>https://awarenesssoftwaregroup.com/research/a-speed-accuracy-frontier/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-speed-accuracy-frontier/</guid>
      <description>Faster learning rates made small models learn sooner but end less accurate.</description>
      <pubDate>Sun, 27 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Task Battery With Designed Headroom</title>
      <link>https://awarenesssoftwaregroup.com/research/a-task-battery-with-designed-headroom/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-task-battery-with-designed-headroom/</guid>
      <description>Our quality checks all passed because the benchmark scored 99% and nothing could fail.</description>
      <pubDate>Sun, 27 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Third Was Model Size</title>
      <link>https://awarenesssoftwaregroup.com/research/a-third-was-model-size/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-third-was-model-size/</guid>
      <description>With every model built the same size, a bigger vocabulary still slows learning, but about a third less than an earlier estimate that let model size change</description>
      <pubDate>Sun, 27 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Batch and Length Are Substitutes</title>
      <link>https://awarenesssoftwaregroup.com/research/batch-and-length-are-substitutes/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/batch-and-length-are-substitutes/</guid>
      <description>Bigger batches and longer sequences both speed learning, but each helps less when the other is large. A law that treats them as independent misses that.</description>
      <pubDate>Sun, 27 Sep 2026 00:00:00 +0000</pubDate>
    </item>
  </channel>
</rss>
