<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Awareness Software Group research</title>
    <link>https://awarenesssoftwaregroup.com/research/</link>
    <description>Open research on how AI models actually learn. One finding per entry, in plain English first and in full technical detail underneath, including the approaches that turned out not to work.</description>
    <language>en</language>
    <atom:link href="https://awarenesssoftwaregroup.com/feed.xml" rel="self" type="application/rss+xml" />
    <lastBuildDate>Tue, 29 Sep 2026 07:17:04 +0000</lastBuildDate>
    <item>
      <title>A Delay by Most Measures</title>
      <link>https://awarenesssoftwaregroup.com/research/a-delay-by-most-measures/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-delay-by-most-measures/</guid>
      <description>The easy-sum warm-up that looked like a trap mostly delayed learning: given longer, most runs learned the hard sum, several thousand steps later than</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Higher Ceiling, Not a Head Start</title>
      <link>https://awarenesssoftwaregroup.com/research/a-higher-ceiling-not-a-head-start/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-higher-ceiling-not-a-head-start/</guid>
      <description>Given more training, models that started on a random mix of easier tasks reached 94 percent on the hard task while models trained on it alone stalled at</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Learning-Rate Warm-Up Does the Job</title>
      <link>https://awarenesssoftwaregroup.com/research/a-learning-rate-warm-up-does-the-job/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-learning-rate-warm-up-does-the-job/</guid>
      <description>Starting training with small steps and ramping up did as well as our best easy-data warm-up, and better than the easy-task mix.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Sweet Spot, Not the Nearest</title>
      <link>https://awarenesssoftwaregroup.com/research/a-sweet-spot-not-the-nearest/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-sweet-spot-not-the-nearest/</guid>
      <description>Before a hard memory task, the best single warm-up was a version easier by a clear margin.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Task Battery With Designed Headroom</title>
      <link>https://awarenesssoftwaregroup.com/research/a-task-battery-with-designed-headroom/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-task-battery-with-designed-headroom/</guid>
      <description>Our quality checks all passed because the benchmark scored 99% and nothing could fail.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Against the Best Plain Run, Not Established</title>
      <link>https://awarenesssoftwaregroup.com/research/against-the-best-plain-run-not-established/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/against-the-best-plain-run-not-established/</guid>
      <description>On the second task, starting with easier sums solved the hard sum on 11 of 16 runs against 6 for the best ordinary recipe, but the accuracy difference is</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>At Width 96 the Gap Holds Steady</title>
      <link>https://awarenesssoftwaregroup.com/research/at-width-96-the-gap-holds-steady/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/at-width-96-the-gap-holds-steady/</guid>
      <description>In a model twice as wide, trained 2.5 times longer, the easy-task warm-up stayed about five points ahead but did not pull further away, unlike in the</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Easier Data Adds Nothing to a Rate Warm-Up</title>
      <link>https://awarenesssoftwaregroup.com/research/easier-data-adds-nothing-to-a-rate-warm-up/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/easier-data-adds-nothing-to-a-rate-warm-up/</guid>
      <description>Combined with a standard learning-rate warm-up, our best easy-data warm-up added nothing. The best result of the whole study needed no special data at all.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Forty Percent Fewer Steps</title>
      <link>https://awarenesssoftwaregroup.com/research/forty-percent-fewer-steps/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/forty-percent-fewer-steps/</guid>
      <description>On 32 fresh runs, starting with a short mix of easier sums reached the solution to a hard sum in about 40 percent fewer training steps than the best</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>It Must Be an Early Phase</title>
      <link>https://awarenesssoftwaregroup.com/research/it-must-be-an-early-phase/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/it-must-be-an-early-phase/</guid>
      <description>Easy sums helped a small model learn a hard sum as a short phase at the start. Mixed in for the whole run, the same sums left every run unable to learn it.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>It Replicates on a Second Target</title>
      <link>https://awarenesssoftwaregroup.com/research/it-replicates-on-a-second-target/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/it-replicates-on-a-second-target/</guid>
      <description>On a different hard sum, the easy-sum mix again beat the best ordinary recipe on fresh runs, solving it on 25 of 32 runs against 2.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>No Other Claim Ran Hot</title>
      <link>https://awarenesssoftwaregroup.com/research/no-other-claim-ran-hot/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/no-other-claim-ran-hot/</guid>
      <description>We checked every standing speed-up result in the archive for the mistake that undid our warm-up results. Only the warm-up results themselves had made it.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>On Fresh Seeds, the Mixture Wins</title>
      <link>https://awarenesssoftwaregroup.com/research/on-fresh-seeds-the-mixture-wins/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/on-fresh-seeds-the-mixture-wins/</guid>
      <description>On the adding task, starting with a mix of easier sums beat the best ordinary recipe, learning-rate warm-up included, on 32 new runs: 91 percent against</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>One Nearby Version Does Most of It</title>
      <link>https://awarenesssoftwaregroup.com/research/one-nearby-version-does-most-of-it/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/one-nearby-version-does-most-of-it/</guid>
      <description>Warming up on a single easier version close to the hard task gave most of the benefit, on every run. Mixing several easier versions added at most a little.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Order Can Make the Warm-Up Harmful</title>
      <link>https://awarenesssoftwaregroup.com/research/order-can-make-the-warm-up-harmful/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/order-can-make-the-warm-up-harmful/</guid>
      <description>The same three easier sums helped a small model when shuffled and stopped it learning the hard sum at all when given from hardest to easiest.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Order Did Not Help, Mixing Might</title>
      <link>https://awarenesssoftwaregroup.com/research/order-did-not-help-mixing-might/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/order-did-not-help-mixing-might/</guid>
      <description>Working up from easy to hard did not speed learning at matched compute; a random mix of easier tasks early ended clearly higher, which we are now testing</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Plain Gets There Eventually</title>
      <link>https://awarenesssoftwaregroup.com/research/plain-gets-there-eventually/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/plain-gets-there-eventually/</guid>
      <description>Given three times as long, ordinary training learned the harder sum on 10 of 16 runs. The easy-sum mix got there in about a third of the steps.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Research questions</title>
      <link>https://awarenesssoftwaregroup.com/research/questions/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/questions/</guid>
      <description>The 10 questions our public AI training research is organised around, each with its current answer and the published results behind it.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Slower Than Tuned, Not Hotter</title>
      <link>https://awarenesssoftwaregroup.com/research/slower-than-tuned-not-hotter/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/slower-than-tuned-not-hotter/</guid>
      <description>A check on one earlier result: its task learns fastest at twice the learning rate it was run at, so it did not have the too-large-rate problem.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Formula Could Have Failed</title>
      <link>https://awarenesssoftwaregroup.com/research/the-formula-could-have-failed/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-formula-could-have-failed/</guid>
      <description>Our formula for how long to train a model you borrow from had never faced a test it could fail. We built one.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Mixture Gets There Sooner</title>
      <link>https://awarenesssoftwaregroup.com/research/the-mixture-gets-there-sooner/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-mixture-gets-there-sooner/</guid>
      <description>Given twice as long, ordinary training on the adding task mostly caught up. What the easy-sum mix really did was solve it in about a third of the steps.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Tuned on Its Own Target, Plain Never Solves It</title>
      <link>https://awarenesssoftwaregroup.com/research/tuned-on-its-own-target-plain-never-solves-it/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/tuned-on-its-own-target-plain-never-solves-it/</guid>
      <description>Given four ordinary settings tuned on the second hard sum itself, plain training solved it on none of 16 runs; the easy-sum mix solved it on 12.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Two Easy Sums Are Traps</title>
      <link>https://awarenesssoftwaregroup.com/research/two-easy-sums-are-traps/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/two-easy-sums-are-traps/</guid>
      <description>Warming up on either of the two easiest sums alone stopped a small model learning the hard sum at all.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Clock, Not a Warning</title>
      <link>https://awarenesssoftwaregroup.com/research/a-clock-not-a-warning/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-clock-not-a-warning/</guid>
      <description>Moving the moment a model learns, its approach to the edge of chaos moved too, but always a fixed fraction of the way there: it tracks training progress,</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Short Warm-Up Is Enough</title>
      <link>https://awarenesssoftwaregroup.com/research/a-short-warm-up-is-enough/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-short-warm-up-is-enough/</guid>
      <description>Spending only the first 6 percent of training on a random mix of easier tasks lifted a small model from 58 to 83 percent on a hard one.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
  </channel>
</rss>
