<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Haga Blog</title>
    <link>https://haga.mushoodhanif.com/blog</link>
    <description>Independent physical-AI verification, world-model evaluation, and sim-to-real benchmarking methodology.</description>
    <language>en-us</language>
    <lastBuildDate>Tue, 28 Jul 2026 00:00:00 GMT</lastBuildDate>
    <atom:link href="https://haga.mushoodhanif.com/rss.xml" rel="self" type="application/rss+xml"/>
    <item>
      <title>Sim-to-real benchmark design: what to measure when hardware is scarce</title>
      <link>https://haga.mushoodhanif.com/blog/sim-to-real-benchmark</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/sim-to-real-benchmark</guid>
      <pubDate>Tue, 28 Jul 2026 00:00:00 GMT</pubDate>
      <description>A credible sim-to-real benchmark needs protocol discipline, claim boundaries, and an explicit failure taxonomy — not just a leaderboard of successes.</description>
      <category>Evaluation</category>
    </item>
    <item>
      <title>Embodied AI trust gap: stop mixing demos, benchmarks, and deployment claims</title>
      <link>https://haga.mushoodhanif.com/blog/embodied-ai-trust-gap</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/embodied-ai-trust-gap</guid>
      <pubDate>Mon, 27 Jul 2026 00:00:00 GMT</pubDate>
      <description>Embodied-AI marketing outpaces measurement. Separate policy robustness, world-model consistency, and hardware evidence before you call a system deployable.</description>
      <category>Physical AI</category>
    </item>
    <item>
      <title>Mobile robot navigation evaluation: success rate is not enough</title>
      <link>https://haga.mushoodhanif.com/blog/mobile-robot-navigation-evaluation</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/mobile-robot-navigation-evaluation</guid>
      <pubDate>Sun, 26 Jul 2026 00:00:00 GMT</pubDate>
      <description>Navigation can succeed in a clean map and still hide weak localization and brittle obstacle handling. Probe those failure modes before you claim robustness.</description>
      <category>Navigation</category>
    </item>
    <item>
      <title>Robot simulator evaluation: choose for verification, not only for cinema</title>
      <link>https://haga.mushoodhanif.com/blog/robot-simulator-evaluation</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/robot-simulator-evaluation</guid>
      <pubDate>Sat, 25 Jul 2026 00:00:00 GMT</pubDate>
      <description>Simulators differ in physics fidelity, sensor modeling, and repeatability. The right choice for evaluation is not always the most cinematic one.</description>
      <category>Simulation</category>
    </item>
    <item>
      <title>Pick-and-place verification: beyond the one-video demo</title>
      <link>https://haga.mushoodhanif.com/blog/pick-and-place-verification</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/pick-and-place-verification</guid>
      <pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate>
      <description>Automated pick and place needs reproducible policy evaluation under physics stress — not a single successful rollout video.</description>
      <category>Manipulation</category>
    </item>
    <item>
      <title>Why robot hands are still hard — and how independent eval makes progress measurable</title>
      <link>https://haga.mushoodhanif.com/blog/why-robot-hands-are-hard</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/why-robot-hands-are-hard</guid>
      <pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate>
      <description>Dexterous robot hands still lack a shared measurement standard outside the demo reel. Independent grasp evaluation turns highlight clips into comparable evidence.</description>
      <category>Physical AI</category>
    </item>
    <item>
      <title>World model evaluation: from plausible clips to reproducible claims</title>
      <link>https://haga.mushoodhanif.com/blog/world-model-evaluation</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/world-model-evaluation</guid>
      <pubDate>Wed, 22 Jul 2026 00:00:00 GMT</pubDate>
      <description>Realistic video is not proof of physical correctness. Independent world-model evaluation measures consistency, failure modes, and claim boundaries.</description>
      <category>World-model verification</category>
    </item>
    <item>
      <title>Sim-to-real trust gap: why physics trust is the real bottleneck</title>
      <link>https://haga.mushoodhanif.com/blog/sim-to-real-trust-gap</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/sim-to-real-trust-gap</guid>
      <pubDate>Tue, 21 Jul 2026 00:00:00 GMT</pubDate>
      <description>The sim-to-real gap is a trust problem, not only a hardware problem. Independent physics evaluation changes what you can claim before deployment.</description>
      <category>Physical AI</category>
    </item>
    <item>
      <title>Physics-IQ held-out cohort: real 0%, CogVideoX 100% via static_hover</title>
      <link>https://haga.mushoodhanif.com/blog/physics-iq-held-out-cohort</link>
      <guid isPermaLink="true">https://haga.mushoodhanif.com/blog/physics-iq-held-out-cohort</guid>
      <pubDate>Mon, 20 Jul 2026 00:00:00 GMT</pubDate>
      <description>Protocol v1 held-out results on Physics-IQ: real quiet controls stay quiet, CogVideoX I2V fails uniformly via static_hover — with claim boundaries.</description>
      <category>World-model verification</category>
    </item>
  </channel>
</rss>