<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Awareness Software Group research</title>
    <link>https://awarenesssoftwaregroup.com/research/</link>
    <description>Open research on how AI models actually learn. One finding per entry, in plain English first and in full technical detail underneath, including the approaches that turned out not to work.</description>
    <language>en</language>
    <atom:link href="https://awarenesssoftwaregroup.com/feed.xml" rel="self" type="application/rss+xml" />
    <lastBuildDate>Tue, 29 Sep 2026 03:20:12 +0000</lastBuildDate>
    <item>
      <title>A Delay by Most Measures</title>
      <link>https://awarenesssoftwaregroup.com/research/a-delay-by-most-measures/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-delay-by-most-measures/</guid>
      <description>The easy-sum warm-up that looked like a trap mostly delayed learning: given longer, most runs learned the hard sum, several thousand steps later than</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Learning-Rate Warm-Up Does the Job</title>
      <link>https://awarenesssoftwaregroup.com/research/a-learning-rate-warm-up-does-the-job/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-learning-rate-warm-up-does-the-job/</guid>
      <description>Starting training with small steps and ramping up did as well as our best easy-data warm-up, and better than the easy-task mix.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Sweet Spot, Not the Nearest</title>
      <link>https://awarenesssoftwaregroup.com/research/a-sweet-spot-not-the-nearest/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-sweet-spot-not-the-nearest/</guid>
      <description>Before a hard memory task, the best single warm-up was a version easier by a clear margin.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Task Battery With Designed Headroom</title>
      <link>https://awarenesssoftwaregroup.com/research/a-task-battery-with-designed-headroom/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-task-battery-with-designed-headroom/</guid>
      <description>Our quality checks all passed because the benchmark scored 99% and nothing could fail.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Against the Best Plain Run, Not Established</title>
      <link>https://awarenesssoftwaregroup.com/research/against-the-best-plain-run-not-established/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/against-the-best-plain-run-not-established/</guid>
      <description>On the second task, starting with easier sums solved the hard sum on 11 of 16 runs against 6 for the best ordinary recipe, but the accuracy difference is</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Easier Data Adds Nothing to a Rate Warm-Up</title>
      <link>https://awarenesssoftwaregroup.com/research/easier-data-adds-nothing-to-a-rate-warm-up/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/easier-data-adds-nothing-to-a-rate-warm-up/</guid>
      <description>Combined with a standard learning-rate warm-up, our best easy-data warm-up added nothing. The best result of the whole study needed no special data at all.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>It Replicates on a Second Target</title>
      <link>https://awarenesssoftwaregroup.com/research/it-replicates-on-a-second-target/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/it-replicates-on-a-second-target/</guid>
      <description>On a different hard sum, the easy-sum mix again beat the best ordinary recipe on fresh runs, solving it on 25 of 32 runs against 2.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>No Other Claim Ran Hot</title>
      <link>https://awarenesssoftwaregroup.com/research/no-other-claim-ran-hot/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/no-other-claim-ran-hot/</guid>
      <description>We checked every standing speed-up result in the archive for the mistake that undid our warm-up results. Only the warm-up results themselves had made it.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>On Fresh Seeds, the Mixture Wins</title>
      <link>https://awarenesssoftwaregroup.com/research/on-fresh-seeds-the-mixture-wins/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/on-fresh-seeds-the-mixture-wins/</guid>
      <description>On the adding task, starting with a mix of easier sums beat the best ordinary recipe, learning-rate warm-up included, on 32 new runs: 91 percent against</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>One Nearby Version Does Most of It</title>
      <link>https://awarenesssoftwaregroup.com/research/one-nearby-version-does-most-of-it/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/one-nearby-version-does-most-of-it/</guid>
      <description>Warming up on a single easier version close to the hard task gave most of the benefit, on every run. Mixing several easier versions added at most a little.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Order Can Make the Warm-Up Harmful</title>
      <link>https://awarenesssoftwaregroup.com/research/order-can-make-the-warm-up-harmful/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/order-can-make-the-warm-up-harmful/</guid>
      <description>The same three easier sums helped a small model when shuffled and stopped it learning the hard sum at all when given from hardest to easiest.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Order Did Not Help, Mixing Might</title>
      <link>https://awarenesssoftwaregroup.com/research/order-did-not-help-mixing-might/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/order-did-not-help-mixing-might/</guid>
      <description>Working up from easy to hard did not speed learning at matched compute; a random mix of easier tasks early ended clearly higher, which we are now testing</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Research questions</title>
      <link>https://awarenesssoftwaregroup.com/research/questions/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/questions/</guid>
      <description>The 10 questions our public AI training research is organised around, each with its current answer and the published results behind it.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Slower Than Tuned, Not Hotter</title>
      <link>https://awarenesssoftwaregroup.com/research/slower-than-tuned-not-hotter/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/slower-than-tuned-not-hotter/</guid>
      <description>A check on one earlier result: its task learns fastest at twice the learning rate it was run at, so it did not have the too-large-rate problem.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Formula Could Have Failed</title>
      <link>https://awarenesssoftwaregroup.com/research/the-formula-could-have-failed/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-formula-could-have-failed/</guid>
      <description>Our formula for how long to train a model you borrow from had never faced a test it could fail. We built one.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Mixture Gets There Sooner</title>
      <link>https://awarenesssoftwaregroup.com/research/the-mixture-gets-there-sooner/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-mixture-gets-there-sooner/</guid>
      <description>Given twice as long, ordinary training on the adding task mostly caught up. What the easy-sum mix really did was solve it in about a third of the steps.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Tuned on Its Own Target, Plain Never Solves It</title>
      <link>https://awarenesssoftwaregroup.com/research/tuned-on-its-own-target-plain-never-solves-it/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/tuned-on-its-own-target-plain-never-solves-it/</guid>
      <description>Given four ordinary settings tuned on the second hard sum itself, plain training solved it on none of 16 runs; the easy-sum mix solved it on 12.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Two Easy Sums Are Traps</title>
      <link>https://awarenesssoftwaregroup.com/research/two-easy-sums-are-traps/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/two-easy-sums-are-traps/</guid>
      <description>Warming up on either of the two easiest sums alone stopped a small model learning the hard sum at all.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Clock, Not a Warning</title>
      <link>https://awarenesssoftwaregroup.com/research/a-clock-not-a-warning/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-clock-not-a-warning/</guid>
      <description>Moving the moment a model learns, its approach to the edge of chaos moved too, but always a fixed fraction of the way there: it tracks training progress,</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Higher Ceiling, Not a Head Start</title>
      <link>https://awarenesssoftwaregroup.com/research/a-higher-ceiling-not-a-head-start/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-higher-ceiling-not-a-head-start/</guid>
      <description>Given more training, models that started on a random mix of easier tasks reached 94 percent on the hard task while models trained on it alone stalled at</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Short Warm-Up Is Enough</title>
      <link>https://awarenesssoftwaregroup.com/research/a-short-warm-up-is-enough/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-short-warm-up-is-enough/</guid>
      <description>Spending only the first 6 percent of training on a random mix of easier tasks lifted a small model from 58 to 83 percent on a hard one.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Speed-Accuracy Frontier</title>
      <link>https://awarenesssoftwaregroup.com/research/a-speed-accuracy-frontier/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-speed-accuracy-frontier/</guid>
      <description>Faster learning rates made small models learn sooner but end less accurate.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>At Width 96 the Gap Holds Steady</title>
      <link>https://awarenesssoftwaregroup.com/research/at-width-96-the-gap-holds-steady/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/at-width-96-the-gap-holds-steady/</guid>
      <description>In a model twice as wide, trained 2.5 times longer, the easy-task warm-up stayed about five points ahead but did not pull further away, unlike in the</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Being Choosy About Data Costs More Than It Saves</title>
      <link>https://awarenesssoftwaregroup.com/research/being-choosy-about-data-costs-more-than-it-saves/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/being-choosy-about-data-costs-more-than-it-saves/</guid>
      <description>Inspect four batches, train only on the most informative. It sounds efficient.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Bigger Batch, Higher Limit</title>
      <link>https://awarenesssoftwaregroup.com/research/bigger-batch-higher-limit/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/bigger-batch-higher-limit/</guid>
      <description>The fastest learning rate a small model can tolerate roughly triples from batch 16 to batch 256, consistent with the common square-root rule.</description>
      <pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate>
    </item>
  </channel>
</rss>
