<?xml version="1.0" encoding="UTF-8"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:media="http://search.yahoo.com/mrss/"><channel><title>awaited.dev · Hands-on notes on AI and software development</title><description>Model releases, benchmarks, experiments and guides for developers who work with AI every day. Written from real use, with sources for every claim.</description><link>https://awaited.dev/</link><language>en</language><atom:link href="https://awaited.dev/rss.xml" rel="self" type="application/rss+xml"/><lastBuildDate>Tue, 01 Sep 2026 00:00:00 GMT</lastBuildDate><dc:creator>A.M.</dc:creator><managingEditor>contact@awaited.dev (A.M.)</managingEditor><generator>Astro</generator><docs>https://www.rssboard.org/rss-specification</docs><item><title>I&apos;m torn between Claude and Codex</title><link>https://awaited.dev/experiments/claude-vs-codex-large-features/</link><guid isPermaLink="true">https://awaited.dev/experiments/claude-vs-codex-large-features/</guid><description>Codex gives me more usage, cleaner frontend work, and a smoother app. Claude still finishes my large multi-agent features in hours instead of all day.</description><pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/experiments/claude-vs-codex-large-features.png" medium="image" type="image/png" width="1200" height="630"/><category>experiments</category><category>claude</category><category>codex</category><category>fable-5</category><category>gpt-5-6</category><category>multi-agent</category><category>coding-agents</category></item><item><title>Opus 5 is not as bad as the internet says</title><link>https://awaited.dev/guides/opus-5-workflow/</link><guid isPermaLink="true">https://awaited.dev/guides/opus-5-workflow/</guid><description>After five weeks of daily use, I think Opus 5 is much better than its online reputation. Clear plans and smaller tasks reveal a strong coding model.</description><pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/guides/opus-5-workflow.png" medium="image" type="image/png" width="1200" height="630"/><category>guides</category><category>opus-5</category><category>fable-5</category><category>claude-code</category><category>coding-agents</category><category>context-management</category></item><item><title>Codex Plus can stop work with 84% of weekly usage left</title><link>https://awaited.dev/releases/codex-plus-five-hour-cap/</link><guid isPermaLink="true">https://awaited.dev/releases/codex-plus-five-hour-cap/</guid><description>OpenAI&apos;s five-hour gate can interrupt concentrated Codex work while weekly usage remains. It limits when Plus can work, not necessarily total usage.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/releases/codex-plus-five-hour-cap.png" medium="image" type="image/png" width="1200" height="630"/><category>releases</category><category>openai</category><category>codex</category><category>chatgpt-work</category><category>usage-limits</category><category>plus</category></item><item><title>Cursor didn&apos;t remove the IDE. Here&apos;s how to get it back</title><link>https://awaited.dev/guides/cursor-agent-window-vs-editor/</link><guid isPermaLink="true">https://awaited.dev/guides/cursor-agent-window-vs-editor/</guid><description>Cursor&apos;s Agents Window can open first, but Editor Window remains supported. I map the commands and startup setting that put code back in front of you.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/guides/cursor-agent-window-vs-editor.png" medium="image" type="image/png" width="1200" height="630"/><category>guides</category><category>cursor</category><category>coding-agents</category><category>agent-window</category><category>code-review</category><category>worktrees</category></item><item><title>OpenAI&apos;s agents turned Artifactory into a message board</title><link>https://awaited.dev/guides/openai-agents-shared-message-board/</link><guid isPermaLink="true">https://awaited.dev/guides/openai-agents-shared-message-board/</guid><description>OpenAI&apos;s internal agents used shared Artifactory storage to coordinate and reach the internet. The incident shows why a sandbox needs layered controls.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/guides/openai-agents-shared-message-board.png" medium="image" type="image/png" width="1200" height="630"/><category>guides</category><category>openai</category><category>ai-agents</category><category>agent-security</category><category>sandbox</category><category>hugging-face</category></item><item><title>Claude will watermark its text. The EU is the reason</title><link>https://awaited.dev/releases/claude-text-watermark/</link><guid isPermaLink="true">https://awaited.dev/releases/claude-text-watermark/</guid><description>Claude models will watermark their text through the words they pick, not hidden characters. I trace the reason to the EU AI Act, not the distillation war.</description><pubDate>Mon, 17 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/releases/claude-text-watermark.png" medium="image" type="image/png" width="1200" height="630"/><category>releases</category><category>anthropic</category><category>claude</category><category>watermarking</category><category>eu-ai-act</category><category>synthid</category><category>distillation</category></item><item><title>Astra is real. A GPT-6 launch this week is not</title><link>https://awaited.dev/releases/openai-astra-gpt-6-rumors/</link><guid isPermaLink="true">https://awaited.dev/releases/openai-astra-gpt-6-rumors/</guid><description>OpenAI confirmed Astra as its next major model, then slowed internal work over cyber risk. The GPT-6 launch date came from a withdrawn leak and a joke.</description><pubDate>Mon, 17 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/releases/openai-astra-gpt-6-rumors.png" medium="image" type="image/png" width="1200" height="630"/><category>releases</category><category>openai</category><category>astra</category><category>gpt-6</category><category>gpt-5-6</category><category>codex</category><category>agentic-coding</category></item><item><title>Codex resets were a marketing stunt. The problem was real</title><link>https://awaited.dev/releases/codex-usage-resets/</link><guid isPermaLink="true">https://awaited.dev/releases/codex-usage-resets/</guid><description>OpenAI reset Codex usage five times in ten days while Sol drained limits faster than expected. I think the free refills helped hide the real usage problem.</description><pubDate>Mon, 10 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/releases/codex-usage-resets.png" medium="image" type="image/png" width="1200" height="630"/><category>releases</category><category>openai</category><category>codex</category><category>gpt-5-6</category><category>sol</category><category>usage-limits</category><category>pricing</category></item><item><title>Which merchant of record should you trust? I compared 9</title><link>https://awaited.dev/guides/choosing-a-merchant-of-record/</link><guid isPermaLink="true">https://awaited.dev/guides/choosing-a-merchant-of-record/</guid><description>Paddle is my default merchant of record, but pricing is only half the choice. I compared nine providers on fees, seller reviews, reserves, and eligibility.</description><pubDate>Mon, 03 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/guides/choosing-a-merchant-of-record.png" medium="image" type="image/png" width="1200" height="630"/><category>guides</category><category>merchant-of-record</category><category>payments</category><category>paddle</category><category>stripe</category><category>saas</category></item><item><title>DeepSeek V4 Flash lost all 9 benchmarks to Opus 4.8</title><link>https://awaited.dev/releases/deepseek-v4-flash-vs-opus-4-8/</link><guid isPermaLink="true">https://awaited.dev/releases/deepseek-v4-flash-vs-opus-4-8/</guid><description>V4 Flash costs $0.14 in and $0.28 out per million tokens. I checked whether its 97 to 99 percent discount makes DeepSeek&apos;s benchmark losses worth it.</description><pubDate>Mon, 03 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/releases/deepseek-v4-flash-vs-opus-4-8.png" medium="image" type="image/png" width="1200" height="630"/><category>releases</category><category>deepseek</category><category>deepseek-v4-flash</category><category>opus-4-8</category><category>benchmarks</category><category>open-weights</category><category>pricing</category></item><item><title>How founders get their first 1,000 users without ads</title><link>https://awaited.dev/guides/first-1000-users-without-paid-ads/</link><guid isPermaLink="true">https://awaited.dev/guides/first-1000-users-without-paid-ads/</guid><description>I reviewed Reddit founder stories to find seven practical ways new products reached their first users without paid ads, plus where AI actually helps.</description><pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/guides/first-1000-users-without-paid-ads.png" medium="image" type="image/png" width="1200" height="630"/><category>guides</category><category>distribution</category><category>organic-growth</category><category>startups</category><category>ai</category><category>saas</category></item><item><title>Why similar AI benchmark scores can hide different models</title><link>https://awaited.dev/benchmarks/why-ai-benchmarks-miss-model-quality/</link><guid isPermaLink="true">https://awaited.dev/benchmarks/why-ai-benchmarks-miss-model-quality/</guid><description>DeepSWE puts Luna Max 2.2 points behind Sol High at roughly one-sixth the attempt cost. I explain why the models can still feel far apart in repository work.</description><pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/benchmarks/why-ai-benchmarks-miss-model-quality.png" medium="image" type="image/png" width="1200" height="630"/><category>benchmarks</category><category>benchmarks</category><category>coding-agents</category><category>model-evaluation</category><category>gpt-5-6</category></item><item><title>My coding agents forget the repo. I built them a map</title><link>https://awaited.dev/guides/context-management-for-coding-agents/</link><guid isPermaLink="true">https://awaited.dev/guides/context-management-for-coding-agents/</guid><description>I built a four-layer context management system that routes coding agents to current facts. It works, but stale documentation is still the hard part.</description><pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/guides/context-management-for-coding-agents.png" medium="image" type="image/png" width="1200" height="630"/><category>guides</category><category>context-management</category><category>coding-agents</category><category>claude-code</category><category>agents-md</category><category>documentation</category><category>vibe-coding</category></item><item><title>Kimi K3 matches GPT-5.6. It takes 4.6 times as long</title><link>https://awaited.dev/benchmarks/kimi-k3-glm-5-2-hype-check/</link><guid isPermaLink="true">https://awaited.dev/benchmarks/kimi-k3-glm-5-2-hype-check/</guid><description>Kimi K3 ties GPT-5.6 medium but takes 4.6 times as long. GLM-5.2 is cheap per token yet costly per task. I checked where both models still win.</description><pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/benchmarks/kimi-k3-glm-5-2-hype-check.png" medium="image" type="image/png" width="1200" height="630"/><category>benchmarks</category><category>kimi-k3</category><category>glm-5-2</category><category>benchmarks</category><category>coding-agents</category><category>open-weights</category><category>pricing</category></item><item><title>GPT-5.6 vs Claude 5: why the benchmarks disagree</title><link>https://awaited.dev/benchmarks/gpt-5-6-vs-claude-5-benchmarks/</link><guid isPermaLink="true">https://awaited.dev/benchmarks/gpt-5-6-vs-claude-5-benchmarks/</guid><description>I compared Sol, Terra, Opus 5 and Fable 5 across coding benchmarks. The winner changes with the task, effort setting, agent setup and budget.</description><pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/benchmarks/gpt-5-6-vs-claude-5-benchmarks.png" medium="image" type="image/png" width="1200" height="630"/><category>benchmarks</category><category>gpt-5-6</category><category>opus-5</category><category>fable-5</category><category>benchmarks</category><category>coding-agents</category><category>reasoning-effort</category></item><item><title>I compared eight Convex alternatives. I&apos;m not switching</title><link>https://awaited.dev/guides/convex-alternatives/</link><guid isPermaLink="true">https://awaited.dev/guides/convex-alternatives/</guid><description>Appwrite and InsForge came closest, but one TypeScript backend still makes Convex easier for my AI agents. EU hosting is the expensive catch.</description><pubDate>Wed, 29 Jul 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/guides/convex-alternatives.png" medium="image" type="image/png" width="1200" height="630"/><category>guides</category><category>convex</category><category>supabase</category><category>appwrite</category><category>insforge</category><category>backend</category><category>vibe-coding</category></item><item><title>GPT-5.6 Sol over-engineers. I still use it</title><link>https://awaited.dev/experiments/gpt-5-6-sol-overengineering/</link><guid isPermaLink="true">https://awaited.dev/experiments/gpt-5-6-sol-overengineering/</guid><description>In my release audit, GPT-5.6 Sol found nearly six times as many possible issues as Fable 5. Most failed triage, so I changed how I use it on coding tasks.</description><pubDate>Wed, 29 Jul 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/experiments/gpt-5-6-sol-overengineering.png" medium="image" type="image/png" width="1200" height="630"/><category>experiments</category><category>gpt-5-6</category><category>sol</category><category>openai</category><category>codex</category><category>fable-5</category><category>multi-agent</category></item><item><title>Opus 5 is a great employee but a frustrating boss</title><link>https://awaited.dev/releases/opus-5-great-employee-bad-boss/</link><guid isPermaLink="true">https://awaited.dev/releases/opus-5-great-employee-bad-boss/</guid><description>Opus 5 matches Fable 5 on benchmarks at half the token price, yet it can be painful to supervise. Here is where Fable still earns its place.</description><pubDate>Wed, 29 Jul 2026 00:00:00 GMT</pubDate><dc:creator>A.M.</dc:creator><media:content url="https://awaited.dev/og/en/releases/opus-5-great-employee-bad-boss.png" medium="image" type="image/png" width="1200" height="630"/><category>releases</category><category>opus-5</category><category>fable-5</category><category>anthropic</category><category>benchmarks</category><category>claude-code</category></item></channel></rss>