<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Awareness Software Group research</title>
    <link>https://awarenesssoftwaregroup.com/research/</link>
    <description>Open research on how AI models actually learn. One finding per entry, in plain English first and in full technical detail underneath, including the approaches that turned out not to work.</description>
    <language>en</language>
    <atom:link href="https://awarenesssoftwaregroup.com/feed.xml" rel="self" type="application/rss+xml" />
    <lastBuildDate>Wed, 30 Sep 2026 01:32:38 +0000</lastBuildDate>
    <item>
      <title>A Harder Maximum, and It Helps Again</title>
      <link>https://awarenesssoftwaregroup.com/research/a-harder-maximum-and-it-helps-again/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-harder-maximum-and-it-helps-again/</guid>
      <description>On a harder version of the maximum task, starting on easier versions saved about a quarter of the steps, correcting our earlier reading that the speed-up</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Sweet Spot, Not the Nearest</title>
      <link>https://awarenesssoftwaregroup.com/research/a-sweet-spot-not-the-nearest/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-sweet-spot-not-the-nearest/</guid>
      <description>Before a hard memory task, the best single warm-up was a version easier by a clear margin.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Differences Help Too</title>
      <link>https://awarenesssoftwaregroup.com/research/differences-help-too/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/differences-help-too/</guid>
      <description>Warming up on easy subtraction problems sped up learning a hard addition problem almost as much as easy addition did.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>For the Maximum, It Slows Learning</title>
      <link>https://awarenesssoftwaregroup.com/research/for-the-maximum-it-slows-learning/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/for-the-maximum-it-slows-learning/</guid>
      <description>When the hard task was to output the larger of two remembered symbols, starting on easier versions slowed learning by about 1,300 steps instead of</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Kept In, They Get in the Way</title>
      <link>https://awarenesssoftwaregroup.com/research/kept-in-they-get-in-the-way/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/kept-in-they-get-in-the-way/</guid>
      <description>Even with the hard sum three quarters of the training data, keeping easy sums in the mix stopped every run learning it, worse than no easy sums at all.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The LSTM&#x27;s Best Rate, and the Speed-Up Holds</title>
      <link>https://awarenesssoftwaregroup.com/research/lstm-best-rate-found/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/lstm-best-rate-found/</guid>
      <description>Tuned to its best learning rate, ordinary training of a small LSTM still took about 3,800 more steps than starting on easier sums, and solved the task</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>On a Transformer, No Speed-Up</title>
      <link>https://awarenesssoftwaregroup.com/research/on-a-transformer-no-speed-up/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/on-a-transformer-no-speed-up/</guid>
      <description>Given enough runs and time, a small transformer learned the hard sum more often and sooner without the easy-sum warm-up.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>On a Transformer, Not Established</title>
      <link>https://awarenesssoftwaregroup.com/research/on-a-transformer-not-established/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/on-a-transformer-not-established/</guid>
      <description>On a small transformer, the easy-sum start got more runs to the target, but most runs of both recipes had not got there by the end, so the result is not</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>On an LSTM It Holds</title>
      <link>https://awarenesssoftwaregroup.com/research/on-an-lstm-it-holds/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/on-an-lstm-it-holds/</guid>
      <description>On a second kind of recurrent network, an LSTM, the easy-sum start solved the hard sum on 15 of 16 runs; ordinary training on 5.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Predicting The Cost From The Task</title>
      <link>https://awarenesssoftwaregroup.com/research/predicting-the-cost-from-the-task/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/predicting-the-cost-from-the-task/</guid>
      <description>What it costs to take something away from a model turns out to be set by the structure of the exercise, not by the model. Change the exercise and it moves.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Research questions</title>
      <link>https://awarenesssoftwaregroup.com/research/questions/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/questions/</guid>
      <description>The 10 questions our public AI training research is organised around, each with its current answer and the published results behind it.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Subtraction Too</title>
      <link>https://awarenesssoftwaregroup.com/research/subtraction-too/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/subtraction-too/</guid>
      <description>With subtraction instead of addition as the hard task, starting on easier subtractions still roughly halved the training steps needed.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The First Piece Is The Load-Bearing One</title>
      <link>https://awarenesssoftwaregroup.com/research/the-first-piece-is-the-load-bearing-one/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-first-piece-is-the-load-bearing-one/</guid>
      <description>We expected the hardest part of a task to be the one you cannot take away from a model. It is the easiest part, the one it learns first.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Operation and the Component</title>
      <link>https://awarenesssoftwaregroup.com/research/the-operation-and-the-component/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/the-operation-and-the-component/</guid>
      <description>Easy sums sharing nothing with the hard sum still sped up learning, and easy sums sharing a part of it sped it up about as much again.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Tuned on Its Own Target, Plain Never Solves It</title>
      <link>https://awarenesssoftwaregroup.com/research/tuned-on-its-own-target-plain-never-solves-it/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/tuned-on-its-own-target-plain-never-solves-it/</guid>
      <description>Given four ordinary settings tuned on the second hard sum itself, plain training solved it on none of 16 runs; the easy-sum mix solved it on 12.</description>
      <pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Delay by Most Measures</title>
      <link>https://awarenesssoftwaregroup.com/research/a-delay-by-most-measures/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-delay-by-most-measures/</guid>
      <description>The easy-sum warm-up that looked like a trap mostly delayed learning: given longer, most runs learned the hard sum, several thousand steps later than</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Higher Ceiling, Not a Head Start</title>
      <link>https://awarenesssoftwaregroup.com/research/a-higher-ceiling-not-a-head-start/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-higher-ceiling-not-a-head-start/</guid>
      <description>Given more training, models that started on a random mix of easier tasks reached 94 percent on the hard task while models trained on it alone stalled at</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Learning-Rate Warm-Up Does the Job</title>
      <link>https://awarenesssoftwaregroup.com/research/a-learning-rate-warm-up-does-the-job/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-learning-rate-warm-up-does-the-job/</guid>
      <description>Starting training with small steps and ramping up did as well as our best easy-data warm-up, and better than the easy-task mix.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Task Battery With Designed Headroom</title>
      <link>https://awarenesssoftwaregroup.com/research/a-task-battery-with-designed-headroom/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/a-task-battery-with-designed-headroom/</guid>
      <description>Our quality checks all passed because the benchmark scored 99% and nothing could fail.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Against the Best Plain Run, Not Established</title>
      <link>https://awarenesssoftwaregroup.com/research/against-the-best-plain-run-not-established/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/against-the-best-plain-run-not-established/</guid>
      <description>On the second task, starting with easier sums solved the hard sum on 11 of 16 runs against 6 for the best ordinary recipe, but the accuracy difference is</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>At Width 24, Undecided</title>
      <link>https://awarenesssoftwaregroup.com/research/at-width-24-undecided/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/at-width-24-undecided/</guid>
      <description>In a model half the size, the easy-sum start leaned the same way, needing about three quarters of the steps, but the difference was not clear of the noise.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>At Width 96 It Holds</title>
      <link>https://awarenesssoftwaregroup.com/research/at-width-96-it-holds/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/at-width-96-it-holds/</guid>
      <description>In a model twice as wide, on a hard sum matched to it, the easy-sum start reached the solution in about 46 percent fewer steps, and every run of both</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>At Width 96 the Gap Holds Steady</title>
      <link>https://awarenesssoftwaregroup.com/research/at-width-96-the-gap-holds-steady/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/at-width-96-the-gap-holds-steady/</guid>
      <description>In a model twice as wide, trained 2.5 times longer, the easy-task warm-up stayed about five points ahead but did not pull further away, unlike in the</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Easier Data Adds Nothing to a Rate Warm-Up</title>
      <link>https://awarenesssoftwaregroup.com/research/easier-data-adds-nothing-to-a-rate-warm-up/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/easier-data-adds-nothing-to-a-rate-warm-up/</guid>
      <description>Combined with a standard learning-rate warm-up, our best easy-data warm-up added nothing. The best result of the whole study needed no special data at all.</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Forty Percent Fewer Steps</title>
      <link>https://awarenesssoftwaregroup.com/research/forty-percent-fewer-steps/</link>
      <guid isPermaLink="true">https://awarenesssoftwaregroup.com/research/forty-percent-fewer-steps/</guid>
      <description>On 32 fresh runs, starting with a short mix of easier sums reached the solution to a hard sum in about 40 percent fewer training steps than the best</description>
      <pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate>
    </item>
  </channel>
</rss>
