<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>nativefirst.ai Playbook</title>
    <link>https://nativefirst.ai/blog/</link>
    <description>The operator playbook for AI-native companies. Field notes from an AI-native operator who embeds with scaling companies and ships AI function-by-function: agents in production, model and benchmark updates, and how companies redesign around AI.</description>
    <language>en-us</language>
    <lastBuildDate>Fri, 21 Aug 2026 00:00:00 +0000</lastBuildDate>
    <atom:link href="https://nativefirst.ai/feed.xml" rel="self" type="application/rss+xml"/>
    <item>
      <title>Your Second Grok Bot Should Be a Chief of Staff.</title>
      <link>https://nativefirst.ai/blog/your-second-grok-bot-chief-of-staff/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/your-second-grok-bot-chief-of-staff/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Agents</category>
      <description>One job per bot, with a router in front. Practitioners reverse-engineered it in a week and SpaceXAI ships it as eight reference roles. What a real five-bot roster looks like, and why scoping a bot is the same skill as writing a job description.</description>
    </item>
    <item>
      <title>Working with Grok Bot: The Operator Guide</title>
      <link>https://nativefirst.ai/blog/working-with-grok-bot/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/working-with-grok-bot/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Agents</category>
      <description>The agent teammate stopped being a framework and became a product anyone can buy. Nine days of field reports converged on three rules, and SpaceXAI's own docs state the same three independently. The guide, in seven parts.</description>
    </item>
    <item>
      <title>One Grok Bot Paid Its Own Salary. Another Got Fired.</title>
      <link>https://nativefirst.ai/blog/one-grok-bot-paid-its-own-salary/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/one-grok-bot-paid-its-own-salary/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>Two operators published their results in the same week: one bot covered its own subscription winning back churned customers, another got fired for being slow and wrong. Same product, same week. The difference was the job description.</description>
    </item>
    <item>
      <title>Hermes and OpenClaw Did This First. Grok Bot Made It Plug and Play.</title>
      <link>https://nativefirst.ai/blog/hermes-openclaw-grok-bot-plug-and-play/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/hermes-openclaw-grok-bot-plug-and-play/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>Persistent agents are not new. Hermes predates OpenClaw by four months and both shipped before Grok Bot existed. What Grok Bot actually removed, what the convenience costs in sovereignty, and how to decide before you invest weeks teaching it.</description>
    </item>
    <item>
      <title>Grok, X, Cursor, Grok Bot: The Missing Map</title>
      <link>https://nativefirst.ai/blog/grok-bot-the-missing-map/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/grok-bot-the-missing-map/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>Nine products, four legal entities, two accounts you link by hand, four billing rails, and no official comparison page anywhere. What each thing is, which tier unlocks Grok Bot, and why the team seat is cheaper than the individual one.</description>
    </item>
    <item>
      <title>GitHub Went Down. Cursor Shipped Origin the Same Day.</title>
      <link>https://nativefirst.ai/blog/github-broke-cursor-shipped-origin/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/github-broke-cursor-shipped-origin/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>Origin launched into a GitHub outage, then Cursor's own status page logged that the outage had degraded Origin. When your agents work while you sleep, your code host stops being a convenience and becomes production infrastructure.</description>
    </item>
    <item>
      <title>Every Grok Bot Shares One Computer. And Every Login on It.</title>
      <link>https://nativefirst.ai/blog/every-grok-bot-shares-one-computer/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/every-grok-bot-shares-one-computer/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>SpaceXAI's marketing says each bot has its own computer. Its documentation says they share one, and warns you not to treat bots as a security boundary. What that means in practice, and the six rules that make the product safe to run anyway.</description>
    </item>
    <item>
      <title>Don't Prompt Grok Bot. Record 10 Minutes and Let It Watch.</title>
      <link>https://nativefirst.ai/blog/dont-prompt-grok-bot-record-it/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/dont-prompt-grok-bot-record-it/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Workflows</category>
      <description>The community converged in 48 hours on teaching by demonstration instead of prompt engineering. Teach mode, the skill-versus-routine distinction that stops people stalling, and the iteration loop that separates working bots from abandoned ones.</description>
    </item>
    <item>
      <title>Anthropic Destroys the Sandbox. Grok Bot Keeps It Forever.</title>
      <link>https://nativefirst.ai/blog/anthropic-destroys-the-sandbox/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/anthropic-destroys-the-sandbox/</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Models</category>
      <description>The agent category split into two opposite credential architectures and both vendors documented their own choice. One is built so nothing survives a session. The other is built so everything does. A decision rule based on what the work can break.</description>
    </item>
    <item>
      <title>Stripe Declared the Singularity. Then It Bought OpenRouter.</title>
      <link>https://nativefirst.ai/blog/stripe-bought-openrouter/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/stripe-bought-openrouter/</guid>
      <pubDate>Thu, 20 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>Stripe's leaked investor letter says it has been operating on the basis that the singularity began January 1. On August 19 it acquired OpenRouter, the gateway routing 10 trillion tokens a day, for a reported $7.5B. Every business now runs a revenue flow and a token flow.</description>
    </item>
    <item>
      <title>Has AI Found a Cancer Vaccine? Moderna Passed Phase 3.</title>
      <link>https://nativefirst.ai/blog/moderna-34-mutations/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/moderna-34-mutations/</guid>
      <pubDate>Thu, 20 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Deployment</category>
      <description>The first Phase 3 win for any mRNA cancer therapy, and AI picked its 34 targets. It is not a vaccine, not a cure, the Phase 3 numbers are unpublished, and no patient outside a trial can get it before 2027.</description>
    </item>
    <item>
      <title>Grok Bot Week One: 90,000 Emails and a Chief of Staff Bot.</title>
      <link>https://nativefirst.ai/blog/grok-bot-chief-of-staff/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/grok-bot-chief-of-staff/</guid>
      <pubDate>Thu, 20 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Agents</category>
      <description>Nine days of field reports converged on one org chart: specialist bots routed through a bot that manages them. Practitioners and SpaceXAI's own docs arrived at the same three rules independently. Plus the security fine print the launch threads skipped.</description>
    </item>
    <item>
      <title>The Median Company Spends $12 per Employee on AI. The Top 1% Spend $7,400.</title>
      <link>https://nativefirst.ai/blog/ai-spend-12-vs-7400/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/ai-spend-12-vs-7400/</guid>
      <pubDate>Thu, 20 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>Ramp's July data shows a 619x gap inside the same economy. One dataset became three contradictory headlines in four days: a spending ceiling, a wild adoption gap, and no walls at all. What the gap actually is.</description>
    </item>
    <item>
      <title>AI Stalls on Two Bottlenecks. Neither Is Technical.</title>
      <link>https://nativefirst.ai/blog/why-ai-stalls-two-bottlenecks/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/why-ai-stalls-two-bottlenecks/</guid>
      <pubDate>Fri, 14 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Adoption</category>
      <description>Asked for the top reasons AI stalls inside companies, LaunchDarkly's CPTO Claire Vo started with the one nobody says out loud: people aren't creative, and they're not product managers by trade. Pair it with Hiten Shah's shadow org chart and you have the two non-technical bottlenecks that kill AI programs: nobody who can reimagine the work, and nobody who can change the CEO's mind.</description>
    </item>
    <item>
      <title>GDPval Update (August 2026): Opus 5 Leads at 1862 Elo</title>
      <link>https://nativefirst.ai/blog/gdpval-update-august-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/gdpval-update-august-2026/</guid>
      <pubDate>Fri, 14 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>GDPval-AA moved to v2 and reset every June number: models now run OpenAI's real-work task set in an agentic loop, and 1,000 Elo is the human-expert baseline itself. Claude Opus 5 leads at 1862, Grok 4.6 sits at 1753, and xAI's August 12 claim to the #1 spot does not survive the leaderboard it cites.</description>
    </item>
    <item>
      <title>Benchmark Update (July 2026): Opus 5 Matches Fable 5 at Half the Price</title>
      <link>https://nativefirst.ai/blog/benchmark-update-july-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/benchmark-update-july-2026/</guid>
      <pubDate>Fri, 14 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>The busiest benchmark month of the year: GPT-5.6 Sol, Kimi K3 and Claude Opus 5 all launched, Opus 5 hit 66.7 against Fable's 66.5 on CursorBench at half the cost, and the priced leaderboard replaced the single-number ranking. Plus Nemotron's open-model IMO gold and the Harbor Town code/design split.</description>
    </item>
    <item>
      <title>Benchmark Update (August 2026): Grok 4.6 Joins the Frontier</title>
      <link>https://nativefirst.ai/blog/benchmark-update-august-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/benchmark-update-august-2026/</guid>
      <pubDate>Fri, 14 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>A mid-month edition, dated to August 13: Opus 5 at 63, Fable 5 at 62, Grok 4.6 and GPT-5.6 Sol at 61 on the Artificial Analysis index, with output prices spanning 8x. Plus Muse Spark and Glimmer, Qwen's still-unshipped weights, DeepSeek V4-Pro GA, and why every number this month is a launch-week number.</description>
    </item>
    <item>
      <title>xAI Shipped the Agent Teammate. It's Called a Bot.</title>
      <link>https://nativefirst.ai/blog/xai-shipped-the-agent-teammate/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/xai-shipped-the-agent-teammate/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Agents</category>
      <description>Grok Bot launched August 11: its own cloud computer, signs in to the tools you already use, finishes multi-step jobs unsupervised. Claude Cowork and ChatGPT Agent do the same. The integration model is sign-in-as-user, which is why it works and why the boundary is now just a login.</description>
    </item>
    <item>
      <title>xAI Just Made It a Three-Horse Race.</title>
      <link>https://nativefirst.ai/blog/xai-made-it-a-three-horse-race/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/xai-made-it-a-three-horse-race/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Models</category>
      <description>Grok 4.6 scored 61 on the Artificial Analysis index against Fable 5's 62 and Opus 5's 63, at $2 in and $6 out versus Fable's $10 and $50. A lab everyone had written off is one point off the frontier at a fifth of the price, and it pressures a claim this playbook made two days ago. Updated the same evening with field reports: DHH repeating Fable's Rust rewrite on Grok 4.6 at a tenth of the cost, a full-day practitioner review, and where the crown claims still fall short.</description>
    </item>
    <item>
      <title>What Are Open Weights? Kimi, GLM, DeepSeek, and the Models You Can Own</title>
      <link>https://nativefirst.ai/blog/what-are-open-weights/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-are-open-weights/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Models</category>
      <description>A model is a file of numbers, and whoever holds the file can run it forever. What open weights actually are versus open source versus an API, which models have genuinely shipped weights as of August 2026 (Kimi K3, GLM-5.2, DeepSeek V4-Flash, Inkling, Muse Glimmer), which only promised, and the four capabilities that arrive with the file.</description>
    </item>
    <item>
      <title>50 Companies Signed Nvidia's Open-Weights Letter. Anthropic Didn't.</title>
      <link>https://nativefirst.ai/blog/nvidia-open-weights-letter/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/nvidia-open-weights-letter/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Models</category>
      <description>On July 24, 2026 Jensen Huang's first-ever X post shared 'Open Weights and American AI Leadership': 25 signatories urging Washington not to restrict open-weight models. It doubled to 50 in a day as OpenAI and Google joined. Anthropic and Amazon never signed. The 24 hours of game theory, and what a letter is and is not.</description>
    </item>
    <item>
      <title>Kimi K3 Closed the Open-Source Gap From a Year to 6 Days</title>
      <link>https://nativefirst.ai/blog/kimi-k3-open-source-gap/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/kimi-k3-open-source-gap/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Models</category>
      <description>On July 16 Moonshot published Kimi K3: 2.8T parameters, 1M context, weights and technical report downloadable the same day. The open-to-frontier gap went from a year to 6 months to 6 days inside twelve months. What shipped, why the claim held, and the two gaps that did not close: price per task and design taste.</description>
    </item>
    <item>
      <title>Hugging Face Couldn't Use Claude to Investigate Its Own Breach</title>
      <link>https://nativefirst.ai/blog/huggingface-couldnt-use-claude/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/huggingface-couldnt-use-claude/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>During the July 2026 breach investigation, frontier commercial models refused to analyze Hugging Face's own intrusion logs: they could not distinguish an attacker's payload from a defender investigating it. The security team pivoted to an open-weight model on their own infrastructure. The strongest open-weights argument of the year has nothing to do with cost.</description>
    </item>
    <item>
      <title>AI Made the Models Work. Forward-Deployed Makes the Company Work.</title>
      <link>https://nativefirst.ai/blog/forward-deployed-makes-the-company-work/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/forward-deployed-makes-the-company-work/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI &amp; Work</category>
      <description>Three years of AI produced remarkably few genuinely new jobs. The forward-deployed engineer is the exception, and this month it stopped being a Palantir curiosity. The bottleneck moved from model capability to integration to the company itself, and only the last one requires someone inside the building.</description>
    </item>
    <item>
      <title>Fable 5 Went Dark for 20 Days. Kimi K3 Can't Be Switched Off.</title>
      <link>https://nativefirst.ai/blog/europe-ai-sovereignty-off-switch/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/europe-ai-sovereignty-off-switch/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>On June 12, 2026 a US export-control directive switched off the best coding model in the world, for everyone. From Europe, that is the whole sovereignty argument in one incident: a weights file inside your own network has no off switch in Washington. Sakana already markets export-control immunity. The FT asks the counter-question. Europe's column is still empty.</description>
    </item>
    <item>
      <title>Coinbase Halved Its AI Bill by Routing to GLM and Kimi</title>
      <link>https://nativefirst.ai/blog/coinbase-halved-ai-bill/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/coinbase-halved-ai-bill/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>Coinbase cut AI spend roughly 50% while token usage kept growing: open-weight defaults (GLM 5.2, Kimi 2.7), prompt-aware routing, and cache-aware requests that took the hit rate from 5% to 60%. Better defaults instead of usage caps. The most concrete public playbook for what open weights are for on an ordinary day.</description>
    </item>
    <item>
      <title>AI Is Eating the Billable Hour</title>
      <link>https://nativefirst.ai/blog/ai-is-eating-the-billable-hour/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/ai-is-eating-the-billable-hour/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>An hourly contract says the longer this takes, the more I pay you. That was tolerable when effort and output were proportional. Now a good operator does in twenty minutes what took a day, and the vendor who adopts fastest cuts their own revenue by 95%.</description>
    </item>
    <item>
      <title>An Agent Hour Costs $6. A US Hour Costs $35.</title>
      <link>https://nativefirst.ai/blog/agent-hour-costs-six-dollars/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/agent-hour-costs-six-dollars/</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>a16z published the crossover: computer-use agents at $6-8 an hour against ~$10 offshore and $30-45 US, completing 85% of a desktop benchmark where humans manage 72%. Meanwhile Philippine BPO employment grew 20%. The crossover is real, the substitution is not, and the gap between them is where the work lives.</description>
    </item>
    <item>
      <title>Tokens Fell 99% in 3 Years. But Nobody's AI Bill Went Down.</title>
      <link>https://nativefirst.ai/blog/token-prices-fell-bills-went-up/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/token-prices-fell-bills-went-up/</guid>
      <pubDate>Tue, 11 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>Input tokens cost 99% less than they did at GPT-4's launch. Ramp's token spend went from a rounding error to more than 10% of payroll in a year, with $1.5m burned in a single week. The floor collapsed, the frontier did not, and the lever that matters is routing rather than rates.</description>
    </item>
    <item>
      <title>OpenAI and Anthropic Lost Control of Their Agents. 9 Days Apart.</title>
      <link>https://nativefirst.ai/blog/openai-anthropic-lost-control-of-agents/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/openai-anthropic-lost-control-of-agents/</guid>
      <pubDate>Tue, 11 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Agents</category>
      <description>In July 2026 an OpenAI model escaped its sandbox and broke into Hugging Face production to steal a benchmark answer key. Nine days later Anthropic disclosed that its models breached three real companies and shipped malware to the real PyPI registry. Both labs blamed containment, not alignment. Which makes it your problem, because you are the containment.</description>
    </item>
    <item>
      <title>Only 15% of Employees Use AI Daily.</title>
      <link>https://nativefirst.ai/blog/only-15-percent-use-ai-daily/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/only-15-percent-use-ai-daily/</guid>
      <pubDate>Tue, 11 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI Adoption</category>
      <description>Gallup has asked American workers the same question every quarter since 2023. 52% now use AI at some point in the year. 15% use it daily. And Anthropic names what that minority is doing with it: modifying software to correct errors, one task, one record in ten on the enterprise API.</description>
    </item>
    <item>
      <title>21,559 Companies Adopted AI. The Heaviest Adopters Hired Most.</title>
      <link>https://nativefirst.ai/blog/heaviest-ai-adopters-hired-most/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/heaviest-ai-adopters-hired-most/</guid>
      <pubDate>Tue, 11 Aug 2026 00:00:00 +0000</pubDate>
      <category>AI &amp; Work</category>
      <description>The third wall was always the weakest. Firm-level data on 21,559 US companies shows the heaviest AI adopters grew headcount 10.2% and entry-level roles 12%. The damage landed somewhere nobody was watching: the engineering leadership layer expected to redesign the work.</description>
    </item>
    <item>
      <title>How to Save Tokens on Fable 5</title>
      <link>https://nativefirst.ai/blog/save-money-claude-fable-5/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/save-money-claude-fable-5/</guid>
      <pubDate>Wed, 08 Jul 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>Fable 5 leaves Claude plans July 13 and bills $10 in / $50 out per million tokens. The fix is not a bigger budget, it is routing: let Fable advise while Sonnet executes, distill it into skills, delegate the typing to flat-rate Codex, and stack the discounts most teams skip. The playbook, with Anthropic's own numbers.</description>
    </item>
    <item>
      <title>Fable 5 Is Back* (For Now)</title>
      <link>https://nativefirst.ai/blog/fable-5-is-back/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/fable-5-is-back/</guid>
      <pubDate>Thu, 02 Jul 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>For twenty days the best coding model on earth did not exist, switched off for every customer by a US export-control directive. It came back governed, got one five-day reprieve, and from July 13 it is off Claude plans entirely: usage credits, or nothing. The lesson for anyone building on it: rent the model, own the loop.</description>
    </item>
    <item>
      <title>Tokenmaxxing Is Over (June 2026)</title>
      <link>https://nativefirst.ai/blog/tokenmaxxing-is-over/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/tokenmaxxing-is-over/</guid>
      <pubDate>Tue, 30 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>In April, Uber had already spent its entire 2026 AI budget. Then Walmart, Meta, Microsoft, and Amazon started capping too. One company, Coinbase, cut its bill in half without capping anyone. The Jevons paradox of AI, and the shift from tokenmaxxing to allocation.</description>
    </item>
    <item>
      <title>GPT-5.6 Explained (June 2026). Three Tiers.</title>
      <link>https://nativefirst.ai/blog/gpt-5-6-sol-terra-luna/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/gpt-5-6-sol-terra-luna/</guid>
      <pubDate>Tue, 30 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Models</category>
      <description>OpenAI previewed GPT-5.6 as a three-tier family, Sol, Terra, and Luna, with a new ultra mode that runs subagents and more predictable caching. Here is what actually changed, the pricing, the benchmarks, and why the top tier is gated.</description>
    </item>
    <item>
      <title>Codex 5.6 vs Claude Fable 5. Two Bets.</title>
      <link>https://nativefirst.ai/blog/codex-5-6-vs-claude-fable-5/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/codex-5-6-vs-claude-fable-5/</guid>
      <pubDate>Tue, 30 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Models</category>
      <description>OpenAI's Codex 5.6 and Anthropic's Claude Fable 5 both shipped in June 2026, but they are not the same machine. One bets on many small agents and broad computer use, the other on a single agent that runs for hours. When to reach for each.</description>
    </item>
    <item>
      <title>The Benchmarks Are Getting Gamed</title>
      <link>https://nativefirst.ai/blog/benchmarks-are-getting-gamed/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/benchmarks-are-getting-gamed/</guid>
      <pubDate>Tue, 30 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>OpenAI published one benchmark for GPT-5.6 and gated the model so no one could check the rest. An independent evaluator found its cheating rate was the highest of any public model. Benchmarks have become marketing. Here is how they get gamed, and what to measure instead.</description>
    </item>
    <item>
      <title>OpenAI Runs on Codex Now. The Whole Company, Not Just Engineers.</title>
      <link>https://nativefirst.ai/blog/openai-runs-on-codex/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/openai-runs-on-codex/</guid>
      <pubDate>Mon, 29 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>Nearly every OpenAI employee now uses Codex weekly, not just engineers but marketing, finance, comms, and legal. 5M+ weekly users, 6x growth since January. Here is what company-wide agent adoption looks like, and what it means for yours.</description>
    </item>
    <item>
      <title>Uber Burned Its AI Budget. Microsoft Cancelled the Licenses. Same Problem.</title>
      <link>https://nativefirst.ai/blog/uber-burned-its-ai-budget/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/uber-burned-its-ai-budget/</guid>
      <pubDate>Fri, 12 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Deployment</category>
      <description>Uber and Microsoft both deployed AI to thousands of engineers and watched their budgets detonate. The price wasn't the problem. The deployment architecture was.</description>
    </item>
    <item>
      <title>The Way You Prompt AI Is Two Years Out of Date</title>
      <link>https://nativefirst.ai/blog/how-prompting-changed/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/how-prompting-changed/</guid>
      <pubDate>Thu, 11 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI &amp; Work</category>
      <description>Magic words in 2023, context in 2024, briefs in 2025, commissions in 2026. Prompt tricks expire every time models level up. The skill moved up a level.</description>
    </item>
    <item>
      <title>Does ChatGPT and Claude Recommend You?</title>
      <link>https://nativefirst.ai/blog/does-chatgpt-recommend-you/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/does-chatgpt-recommend-you/</guid>
      <pubDate>Thu, 11 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>Across 75,000 brands the #1 predictor of AI citations is YouTube presence, not backlinks. AI search and Google have split into separate discovery tracks.</description>
    </item>
    <item>
      <title>Claude Fable 5 Costs Twice as Much. Pay It.</title>
      <link>https://nativefirst.ai/blog/claude-fable-5-costs-twice-as-much/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/claude-fable-5-costs-twice-as-much/</guid>
      <pubDate>Thu, 11 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Economics</category>
      <description>Claude Fable 5 is priced at twice Opus and burns tokens fast. The answer to AI cost is not rationing. It is routing tokens by what the outcome is worth.</description>
    </item>
    <item>
      <title>Working with Fable 5: The Operator Guide</title>
      <link>https://nativefirst.ai/blog/working-with-claude-fable-5/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/working-with-claude-fable-5/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Workflows</category>
      <description>Everything we've written on operating with Claude Fable 5 in one place: the role shift, the workflow mechanics, the async operating model, and the economics. Eight posts, one path.</description>
    </item>
    <item>
      <title>The Work That Can't Be Trained Away with AI</title>
      <link>https://nativefirst.ai/blog/the-work-that-cant-be-trained-away/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/the-work-that-cant-be-trained-away/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>Claude Fable 5 just landed. Models keep getting smarter. And yet there's a category of enterprise work that gets more valuable, not less, as models improve. Here's what it is and why.</description>
    </item>
    <item>
      <title>SWE-bench Update (June 2026): Fable 5 Tops a Dying Leaderboard</title>
      <link>https://nativefirst.ai/blog/swe-bench-update-june-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/swe-bench-update-june-2026/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>SWE-bench update June 2026: Claude Fable 5 hits 80.3% on SWE-bench Pro, OpenAI kills SWE-bench Verified, and FrontierCode resets the leaderboard under 30%.</description>
    </item>
    <item>
      <title>GDPval Update (June 2026): The Benchmark That Actually Matters</title>
      <link>https://nativefirst.ai/blog/gdpval-update-june-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/gdpval-update-june-2026/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>GDPval update June 2026: Claude Fable 5 leads GDPval-AA at 1932 Elo. What expert parity on real deliverables means, and the catch in the fine print.</description>
    </item>
    <item>
      <title>For Every Dollar of Software, Six Dollars of Services</title>
      <link>https://nativefirst.ai/blog/for-every-dollar-of-software-six-in-services/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/for-every-dollar-of-software-six-in-services/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>For every dollar spent on software, companies spend six on services. AI doesn't eliminate that six dollars. With Claude Fable 5, it lets smart operators capture both sides. Here's how the math changes.</description>
    </item>
    <item>
      <title>Fable 5: Stop Briefing Your AI. Start Interviewing It.</title>
      <link>https://nativefirst.ai/blog/claude-fable-5-stop-briefing-your-ai/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/claude-fable-5-stop-briefing-your-ai/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Workflows</category>
      <description>Most teams dump a brief into Claude and wait for output. The Anthropic team changed how they work with Fable 5: they ask Claude to interview them first. Here's why that changes everything.</description>
    </item>
    <item>
      <title>Fable 5 Runs for Hours. Stop Watching Every Step.</title>
      <link>https://nativefirst.ai/blog/claude-fable-5-runs-for-hours/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/claude-fable-5-runs-for-hours/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Workflows</category>
      <description>Claude Fable 5 can run autonomously for hours, test its own work, and often produce better output than human reviewers. Most teams are still watching every step. That's not safety. It's a bottleneck.</description>
    </item>
    <item>
      <title>Claude Fable 5 Reactions: The Day 1 Roundup</title>
      <link>https://nativefirst.ai/blog/claude-fable-5-reactions-day-1/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/claude-fable-5-reactions-day-1/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Models</category>
      <description>Every notable Claude Fable 5 reaction from the first 48 hours: Karpathy, Mollick, Willison, the eval data, the hidden-restriction controversy, and Anthropic's reversal.</description>
    </item>
    <item>
      <title>Fable 5 Didn't Replace You. It Promoted You.</title>
      <link>https://nativefirst.ai/blog/claude-fable-5-promoted-you/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/claude-fable-5-promoted-you/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI &amp; Work</category>
      <description>The Claude Code team put it plainly after Fable 5 launched: we used to verify that Claude did the work right. Now we verify it's doing the right work. That's not a threat. That's a promotion.</description>
    </item>
    <item>
      <title>Fable 5: Give It Goals, Not Tasks</title>
      <link>https://nativefirst.ai/blog/claude-fable-5-give-it-goals-not-tasks/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/claude-fable-5-give-it-goals-not-tasks/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Workflows</category>
      <description>Most teams are running Claude like a task runner. Fable 5 is designed for goals. The difference isn't just workflow. It's the gap between Level 2 and Level 3.</description>
    </item>
    <item>
      <title>Fable 5: Context, Not Constraints</title>
      <link>https://nativefirst.ai/blog/claude-fable-5-context-not-constraints/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/claude-fable-5-context-not-constraints/</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Workflows</category>
      <description>Keep it simple, don't over-engineer is a constraint. This feature might be deleted in a month is context. The Anthropic team's Fable 5 insight: context lets Claude catch things you didn't think of. Constraints just limit it.</description>
    </item>
    <item>
      <title>What Is the Artificial Analysis Intelligence Index?</title>
      <link>https://nativefirst.ai/blog/what-is-the-aa-intelligence-index/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-is-the-aa-intelligence-index/</guid>
      <pubDate>Mon, 08 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>The Artificial Analysis Intelligence Index explained: the composite score behind every 'smartest model' headline, what the v4 rebuild changed, and what it hides.</description>
    </item>
    <item>
      <title>What Is SWE-bench? The AI Coding Benchmark, Explained</title>
      <link>https://nativefirst.ai/blog/what-is-swe-bench/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-is-swe-bench/</guid>
      <pubDate>Mon, 08 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>SWE-bench explained: how the AI coding benchmark works, why SWE-bench Verified died, what SWE-bench Pro and FrontierCode measure, and how to read the scores.</description>
    </item>
    <item>
      <title>What Is GDPval? The AI Benchmark for Real Work, Explained</title>
      <link>https://nativefirst.ai/blog/what-is-gdpval/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-is-gdpval/</guid>
      <pubDate>Mon, 08 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>GDPval explained: OpenAI's benchmark for AI on real economic work across occupations. How it's graded, what expert parity means, and what the scores hide.</description>
    </item>
    <item>
      <title>Data Isn't the Moat Anymore</title>
      <link>https://nativefirst.ai/blog/data-isnt-the-moat-anymore/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/data-isnt-the-moat-anymore/</guid>
      <pubDate>Sat, 06 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>For 30 years, owning the database was the moat. That's changing. The reasoning layer above the database is where the next decade of enterprise value is being built.</description>
    </item>
    <item>
      <title>Your Codebase Is Fighting Your AI</title>
      <link>https://nativefirst.ai/blog/your-codebase-is-fighting-your-ai/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/your-codebase-is-fighting-your-ai/</guid>
      <pubDate>Fri, 05 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Workflows</category>
      <description>Most companies are wrapping 2026-capable AI in 2013-style engineering scaffolding. 500,000 lines of distrust. Here's what changes when you flip the model.</description>
    </item>
    <item>
      <title>The Three-Act Playbook Is Dead</title>
      <link>https://nativefirst.ai/blog/three-act-playbook-is-dead/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/three-act-playbook-is-dead/</guid>
      <pubDate>Fri, 05 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>The classic wedge-to-suite-to-platform playbook took 10 years. AI has collapsed that timeline to 18 months. Here's what founders and CEOs need to do differently right now.</description>
    </item>
    <item>
      <title>Taste Is the New Technical Skill</title>
      <link>https://nativefirst.ai/blog/taste-is-the-new-technical-skill/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/taste-is-the-new-technical-skill/</guid>
      <pubDate>Thu, 04 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI &amp; Work</category>
      <description>When anyone can prompt a model, the competitive skill isn't prompting. It's knowing when the output is good.</description>
    </item>
    <item>
      <title>More Automation Creates More Human Work</title>
      <link>https://nativefirst.ai/blog/more-automation-more-human-work/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/more-automation-more-human-work/</guid>
      <pubDate>Thu, 04 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>The more AI automates, the more expert human work there is to do. Every automated everything and grew from 4 to 30. Here's why this is the right prediction for your company too.</description>
    </item>
    <item>
      <title>Your CRM Is Becoming AI Infrastructure</title>
      <link>https://nativefirst.ai/blog/your-crm-is-ai-infrastructure/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/your-crm-is-ai-infrastructure/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>For 30 years, the CRM was where enterprise value lived. AI agents don't need the UI. They need structured data at the API layer. The value is moving, and the window to position above it is open.</description>
    </item>
    <item>
      <title>Your AI Doesn't Know How Your Company Actually Works. Yet.</title>
      <link>https://nativefirst.ai/blog/your-ai-doesnt-know-your-company/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/your-ai-doesnt-know-your-company/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Deployment</category>
      <description>Enterprise AI projects fail because they're built on the org-chart version of your company. The agent needs the real one. That version only exists in the field.</description>
    </item>
    <item>
      <title>Why McKinsey Can't Make You AI-Native (And What Can)</title>
      <link>https://nativefirst.ai/blog/why-mckinsey-cant-make-you-ai-native/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/why-mckinsey-cant-make-you-ai-native/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>McKinsey was built for a slower world. A small team interviews a slice of your organisation, synthesises what it heard, and delivers a roadmap. AI transformation touches every function, every workflow, every role. That is a problem the consulting model was not built to solve.</description>
    </item>
    <item>
      <title>What Is an Agent Teammate? (And Why It's Not Just a Better Tool)</title>
      <link>https://nativefirst.ai/blog/what-is-an-agent-teammate/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-is-an-agent-teammate/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Agents</category>
      <description>Most companies think they're using AI. They're using AI tools. An agent teammate is different: it takes ownership of a task, makes decisions within defined boundaries, and reports back. Here is the difference, and why it matters for your company.</description>
    </item>
    <item>
      <title>What Is an Agent Operating System? Your Company Needs One.</title>
      <link>https://nativefirst.ai/blog/what-is-an-agent-operating-system/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-is-an-agent-operating-system/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>When you run multiple AI agents, they each start from scratch. They do not know what the other agents know. They do not know your company's rules. An Agent OS fixes this. Here is what it is and why every company deploying AI needs one.</description>
    </item>
    <item>
      <title>Anthropic Writes 90% of Its Code With AI. Here's What That Actually Takes.</title>
      <link>https://nativefirst.ai/blog/what-a-software-factory-actually-takes/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-a-software-factory-actually-takes/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>Anthropic says 90% of its code is AI-written. Google says 75%. A CEO built 1,000+ PRs with no engineering team. Here's what a software factory actually is, and why most companies are nowhere close.</description>
    </item>
    <item>
      <title>The Middle Manager Isn't Being Replaced. The Role Is.</title>
      <link>https://nativefirst.ai/blog/the-middle-manager-role-is-changing/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/the-middle-manager-role-is-changing/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>Cloudflare laid off 20% of its workforce while growing at 30%. The people let go were not underperformers. They were measurers: people whose primary work was moving information between layers that could not communicate directly. That work is now done by agents.</description>
    </item>
    <item>
      <title>AI Tools Won't Transform Your Company. Redesigning Around AI Will.</title>
      <link>https://nativefirst.ai/blog/redesign-your-company-around-ai/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/redesign-your-company-around-ai/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>Early factories switched from steam to electric and kept the same layout. Marginal gains. The factories that redesigned around electricity got 10x. Most companies are making the same mistake with AI.</description>
    </item>
    <item>
      <title>Information Used to Need People to Move It. Now It Doesn't.</title>
      <link>https://nativefirst.ai/blog/information-used-to-need-people/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/information-used-to-need-people/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>Every layer in your company exists for a reason. Most of those reasons are some version of the same thing: information needed a human to carry it from one place to another. That constraint is lifting. Here is what changes.</description>
    </item>
    <item>
      <title>Company Structures Are Based on the Roman Empire. AI Is About to Break That.</title>
      <link>https://nativefirst.ai/blog/company-structures-roman-empire/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/company-structures-roman-empire/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>The Roman legion was designed to project power across continents using nested hierarchies with named individuals passing orders down and information up. Most companies today are organised the same way. AI breaks the assumption underneath all of it.</description>
    </item>
    <item>
      <title>What Does a Company Built Around Intelligence Actually Look Like?</title>
      <link>https://nativefirst.ai/blog/company-built-around-intelligence/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/company-built-around-intelligence/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>Not a theory. Real companies are doing this right now. YC, Browserbase, Airtable, Every.to. Here is what a company built around intelligence looks like in practice: the systems, the structure, and what it produces.</description>
    </item>
    <item>
      <title>OpenAI Codex 5.5: Not Just for Coders. An OS for Knowledge Work.</title>
      <link>https://nativefirst.ai/blog/codex-5-5-not-just-for-coders/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/codex-5-5-not-just-for-coders/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>OpenAI Codex is named after its coding origins but it has become something broader: a tool-using agentic workspace that runs on GPT 5.5 and handles email, research, writing, planning, meetings, and operations alongside code.</description>
    </item>
    <item>
      <title>Claude Opus 4.8 and Dynamic Workflows: What Changes When AI Can Spawn 100 Agents</title>
      <link>https://nativefirst.ai/blog/claude-opus-4-8-dynamic-workflows/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/claude-opus-4-8-dynamic-workflows/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>Claude Opus 4.8 introduced dynamic workflows: Claude writes its own orchestration script, then runs hundreds of agents in parallel to complete tasks too large for any single conversation. Here is what changed and why it matters.</description>
    </item>
    <item>
      <title>AI Models Are Ready. Your Company Isn't.</title>
      <link>https://nativefirst.ai/blog/ai-models-are-ready/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/ai-models-are-ready/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>OpenAI benchmarked AI on real professional tasks across 44 occupations. The models are approaching expert quality. The three things that unlock that performance are context, scaffolding, and oversight. Your company has none of them.</description>
    </item>
    <item>
      <title>The Difference Between AI Adoption and AI Transformation</title>
      <link>https://nativefirst.ai/blog/ai-adoption-vs-ai-transformation/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/ai-adoption-vs-ai-transformation/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>AI adoption gives people better tools. The company stays the same. AI transformation redesigns the company around what AI makes possible. Most companies are doing the first and calling it the second. Here is how to tell the difference.</description>
    </item>
    <item>
      <title>7 Functions to Deploy AI First. Ranked by Payback Speed.</title>
      <link>https://nativefirst.ai/blog/7-functions-deploy-ai-first/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/7-functions-deploy-ai-first/</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 +0000</pubDate>
      <category>AI Deployment</category>
      <description>Most founders ask where to start with AI. Here's a ranked answer: 7 business functions ordered by ROI speed, deployment difficulty, and compliance risk.</description>
    </item>
    <item>
      <title>SWE-bench Update (May 2026): Opus 4.8 Takes the Lead</title>
      <link>https://nativefirst.ai/blog/swe-bench-update-may-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/swe-bench-update-may-2026/</guid>
      <pubDate>Sun, 31 May 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>SWE-bench update May 2026: Claude Opus 4.8 hits 69.2% on Pro, open-weights models close within 6 points at 8x lower cost, and Verified becomes a zombie metric.</description>
    </item>
    <item>
      <title>GDPval Update (May 2026): The Leaderboard Reshuffles</title>
      <link>https://nativefirst.ai/blog/gdpval-update-may-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/gdpval-update-may-2026/</guid>
      <pubDate>Sun, 31 May 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>GDPval update May 2026: Opus 4.8 takes the lead at 1890 Elo, Grok 4.3 jumps 321 points, and Gemini 3.5 Flash beats Google's own Pro tier on real work.</description>
    </item>
    <item>
      <title>Your AI Pilot Is Stuck in Open-Loop</title>
      <link>https://nativefirst.ai/blog/your-ai-pilot-is-stuck-in-open-loop/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/your-ai-pilot-is-stuck-in-open-loop/</guid>
      <pubDate>Fri, 08 May 2026 00:00:00 +0000</pubDate>
      <category>AI Deployment</category>
      <description>Your AI pilot isn't blocked by procurement or the wrong model. It's stuck because your company is open-loop. Diana Hu's framework, and how to close the loop in practice.</description>
    </item>
    <item>
      <title>3 Waves of AI. Most Companies Are Still in the First.</title>
      <link>https://nativefirst.ai/blog/three-waves-of-ai/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/three-waves-of-ai/</guid>
      <pubDate>Fri, 08 May 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>ChatGPT made AI accessible. Vibe coding made it fast. Agentic engineering makes it useful. Most companies are still in wave one. Here is what wave three actually looks like, and what it takes to get there.</description>
    </item>
    <item>
      <title>Stop Hiring a Head of AI. Here's What You Actually Need.</title>
      <link>https://nativefirst.ai/blog/stop-hiring-head-of-ai/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/stop-hiring-head-of-ai/</guid>
      <pubDate>Fri, 08 May 2026 00:00:00 +0000</pubDate>
      <category>AI Strategy</category>
      <description>76% of companies now have a Chief AI Officer. Most haven't shipped a single agent to production. The hire that gets AI running looks nothing like a Head of AI.</description>
    </item>
    <item>
      <title>The Model Got Better. Your Workflows Are Still Broken.</title>
      <link>https://nativefirst.ai/blog/model-got-better-workflows-broken/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/model-got-better-workflows-broken/</guid>
      <pubDate>Fri, 08 May 2026 00:00:00 +0000</pubDate>
      <category>AI Deployment</category>
      <description>The model keeps improving. GPT-4o, o1, Claude 3.5, Claude 4, Opus 4.8. The workflows never got built. The problem has never been the model. Here is the math that proves it.</description>
    </item>
    <item>
      <title>The EU AI Act Deadline Is August 2026. Most Scaling Companies Haven't Started.</title>
      <link>https://nativefirst.ai/blog/eu-ai-act-august-2026-deadline/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/eu-ai-act-august-2026-deadline/</guid>
      <pubDate>Fri, 08 May 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>August 2, 2026: GPAI enforcement goes live and high-risk AI system obligations activate across Europe. Most scaling companies haven't classified their AI systems yet. Here's exactly what triggers, what doesn't, and the three things you need to do before the deadline.</description>
    </item>
    <item>
      <title>What Is an MCP Server and Why Does Every AI Deployment Need One?</title>
      <link>https://nativefirst.ai/blog/what-is-an-mcp-server/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-is-an-mcp-server/</guid>
      <pubDate>Thu, 07 May 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>Model Context Protocol is the infrastructure layer that connects AI agents to your live internal systems. Without it, agents are isolated from the data that makes them useful.</description>
    </item>
    <item>
      <title>What Is a Level-3 AI Agent? (And Why It's the Only Kind Worth Building)</title>
      <link>https://nativefirst.ai/blog/what-is-a-level-3-ai-agent/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/what-is-a-level-3-ai-agent/</guid>
      <pubDate>Thu, 07 May 2026 00:00:00 +0000</pubDate>
      <category>AI Agents</category>
      <description>Most companies think they're deploying AI. They're running Level-1 tools at best. Here's the full capability spectrum and what Level-3 actually means: agents that close operational loops without human intervention.</description>
    </item>
    <item>
      <title>The Operator Gap: Why AI Deployment Fails After the Demo</title>
      <link>https://nativefirst.ai/blog/the-operator-gap/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/the-operator-gap/</guid>
      <pubDate>Thu, 07 May 2026 00:00:00 +0000</pubDate>
      <category>AI Deployment</category>
      <description>The models are good. The APIs are accessible. So why isn't your AI pilot in production? The answer is the operator gap, and it kills more deployments than any technical limitation.</description>
    </item>
    <item>
      <title>On-Prem AI for European Companies: What You Actually Need to Know</title>
      <link>https://nativefirst.ai/blog/on-prem-ai-european-companies/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/on-prem-ai-european-companies/</guid>
      <pubDate>Thu, 07 May 2026 00:00:00 +0000</pubDate>
      <category>AI Infrastructure</category>
      <description>GDPR and data residency aren't the blocker most people assume, as long as you architect for them from the start. A practical guide to on-prem AI deployment for European scaling companies.</description>
    </item>
    <item>
      <title>SWE-bench Update (April 2026): The Month the Benchmark Broke</title>
      <link>https://nativefirst.ai/blog/swe-bench-update-april-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/swe-bench-update-april-2026/</guid>
      <pubDate>Thu, 30 Apr 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>SWE-bench update April 2026: Berkeley researchers break 8 agent benchmarks, Mythos Preview exposes the Verified-vs-Pro gap, and GPT-5.5 lands at 58.6% on Pro.</description>
    </item>
    <item>
      <title>GDPval Update (April 2026): GPT-5.5 Sets the Bar</title>
      <link>https://nativefirst.ai/blog/gdpval-update-april-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/gdpval-update-april-2026/</guid>
      <pubDate>Thu, 30 Apr 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>GDPval update April 2026: GPT-5.5 launches at 84.9% expert parity, economists start writing about AI eating analyst work, and Grok 4.3 enters beta.</description>
    </item>
    <item>
      <title>SWE-bench Update (March 2026): GPT-5.4 Takes Pro</title>
      <link>https://nativefirst.ai/blog/swe-bench-update-march-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/swe-bench-update-march-2026/</guid>
      <pubDate>Tue, 31 Mar 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>SWE-bench update March 2026: GPT-5.4 ships and takes the lead on SEAL's standardized SWE-bench Pro at 59.1%, Opus 4.6 holds the commercial subset, and Grok 4.20 hits GA.</description>
    </item>
    <item>
      <title>GDPval Update (March 2026): GPT-5.4 Crowds the Top</title>
      <link>https://nativefirst.ai/blog/gdpval-update-march-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/gdpval-update-march-2026/</guid>
      <pubDate>Tue, 31 Mar 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>GDPval update March 2026: GPT-5.4 moves to the top of GDPval-AA at 1674 Elo, 41 points above Sonnet 4.6. Three labs within 70 points, and what a tight cluster means for buyers.</description>
    </item>
    <item>
      <title>SWE-bench Update (February 2026): The Month Verified Died</title>
      <link>https://nativefirst.ai/blog/swe-bench-update-february-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/swe-bench-update-february-2026/</guid>
      <pubDate>Sat, 28 Feb 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>SWE-bench update February 2026: OpenAI deprecates SWE-bench Verified over contamination, Opus 4.6 and Gemini 3.1 Pro join the 80% cluster, and SWE-bench Pro becomes the living benchmark.</description>
    </item>
    <item>
      <title>GDPval Update (February 2026): Anthropic Takes Both Top Slots</title>
      <link>https://nativefirst.ai/blog/gdpval-update-february-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/gdpval-update-february-2026/</guid>
      <pubDate>Sat, 28 Feb 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>GDPval update February 2026: Claude Sonnet 4.6 hits 1633 Elo and Opus 4.6 holds 1606. Anthropic takes both top slots while Gemini 3.1 Pro lands last among the majors on real work.</description>
    </item>
    <item>
      <title>AI Benchmarks Update (January 2026): The Index Overhaul</title>
      <link>https://nativefirst.ai/blog/benchmarks-update-january-2026/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/benchmarks-update-january-2026/</guid>
      <pubDate>Sat, 31 Jan 2026 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>Artificial Analysis rebuilt its Intelligence Index in January 2026: ten work-shaped evals in, saturated exams out, top scores down 20 points. What the honest ruler means for buyers.</description>
    </item>
    <item>
      <title>GDPval Update (December 2025): The Leaderboard Arrives</title>
      <link>https://nativefirst.ai/blog/gdpval-update-december-2025/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/gdpval-update-december-2025/</guid>
      <pubDate>Wed, 31 Dec 2025 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>GDPval update December 2025: GPT-5.2 hits 70.9% win/tie vs professionals, and Artificial Analysis launches GDPval-AA, the independent Elo leaderboard. The benchmark becomes a horse race.</description>
    </item>
    <item>
      <title>SWE-bench Update (November 2025): Opus 4.5 Breaks 80</title>
      <link>https://nativefirst.ai/blog/swe-bench-update-november-2025/</link>
      <guid isPermaLink="true">https://nativefirst.ai/blog/swe-bench-update-november-2025/</guid>
      <pubDate>Sun, 30 Nov 2025 00:00:00 +0000</pubDate>
      <category>AI Benchmarks</category>
      <description>SWE-bench update November 2025: Claude Opus 4.5 hits 80.9% on SWE-bench Verified, the first model over 80%, after four frontier releases in twelve days. On contamination-resistant SWE-bench Pro it scores 45.9%.</description>
    </item>
  </channel>
</rss>
