AI Digest

Digest curado

lunes, 10 de agosto de 2026·light·short·9,140 tokens

🔥 TOP — lo que SÍ o SÍ tenés que ver

📦 Claude / Anthropic ecosystem

🛠️ Dev tools & coding

🏗️ Software engineering

📚 Vale la pena leer

💤 Skippeable pero conviene saber

Artículos fetched (51)

  • ZhuLinsen/daily_stock_analysis
    github-trending

    LLM 驱动的多市场股票智能分析系统:多源行情、实时新闻、决策看板与自动推送,支持零成本定时运行。 LLM-powered multi-market stock analysis system with multi-source market data, real-time news, decision dashboard, automated notifications, and cost-free scheduled runs. 📈 股票智能分析系统 🤖 基于 AI 大模型的 A股/港股/美股/日股/韩股/台股自选股智能分析系统,每日自动分析并推送「决策仪表盘」到企业微信/飞书/Telegram/Discord/Slack/邮箱 产品预览 · 功能特性 · 快速开始 · 推送效果 · 文档中心 · 完整指南 简体中文 | English | 繁體中文 💖 赞助商 (Sponsors) 🖥️ 产品预览 ✨ 功能特性 能力 覆盖内容 AI 决策报告 核心结论、评分、趋势、买卖点位、风险警报、催化因素、操作检查清单 多市场数据聚合 覆盖 A股、港股、美股、日股、韩股、台股和 ETF,支持行情、K 线、技术指标、新闻、公告、基本面与报告辅助数据;不同市场的数据源和能力边界见 市场支持边界 Web / 桌面工作台 手动分析、任务进度、历史报告、完整 Markdown、回测、持仓、配置管理、浅色 / 深色主题 Agent 策略问股 多轮追问,支持均线、缠论、波浪、趋势、热点、事件、成长、预期等 15 种内置策略,覆盖 Web/Bot/API 智能导入与补全 图片、CSV/Excel、剪贴板导入;股票代码/名称/拼音/别名补全 自动化与推送 GitHub Actions、Docker、本地定时任务、FastAPI 服务和企业微信/飞书/Telegram/Discord/Sl…

  • Comfy-Org/ComfyUI
    github-trending

    The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface. ComfyUI The most powerful and modular AI engine for content creation. ComfyUI is the AI creation engine for visual professionals who demand control over every model, every parameter, and every output. Its powerful and modular node graph interface empowers creatives to generate images, videos, 3D models, audio, and more... ComfyUI natively supports the latest open-source state of the art models. API nodes provide access to the best closed source models such as Nano Banana, Seedance, Hunyuan3D, etc. It is available on Windows, Linux, and macOS, locally with our desktop application, our portable install or on our cloud. The most sophisticated workflows can be exposed through a simple UI thanks to…

  • google-deepmind/weathernext
    github-trending

    WeatherNext WeatherNext 2 This repo contains the code for WeatherNext 2 (WN2), the global, medium-range atmospheric and cyclone forecasting model developed by Google DeepMind and Google Research. It also contains code for prior generation models GraphCast and GenCast. Accessing Forecast Data Feeds If you are interested in directly accessing daily data feeds of WN2 model outputs rather than running the model yourself, we provide them across multiple platforms: Google Cloud (including Earth Engine, BigQuery, and Vertex AI). WeatherLab (including cyclone tracks). OpenMeteo (including an API and interactive builder). Learn More Model Guide & Documentation: Google Developers WeatherNext Guide WeatherNext Cyclones Paper: Operational tropical cyclone forecasting with AI FGN/WN2 Technical Report:…

  • msitarzewski/agency-agents
    github-trending

    A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables. 🎭 The Agency: AI Specialists Ready to Transform Your Workflow A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables. 🆕 There's an app now Agency Agents is a native app for macOS, Linux & Windows that browses the entire roster and installs it into Claude Code, Cursor, Codex, Gemini, Osaurus, and more — with a click. No clone, no scripts, and it auto-updates. → Download the latest release · agencyagents…

  • pingdotgg/t3code
    github-trending

    T3 Code T3 Code is an "agent harness control surface". It enables control of the agents on your machine with a best-in-class mobile app (iOS, Android), web app and Electron-based desktop app. Works with your subscriptions on Claude Code, Codex, Cursor, Grok Build, and OpenCode. If they're set up on your computer, T3 Code can control them. "Wait, what are you selling me?" Nothing. We built T3 Code because we wanted the best possible development experience with agents. We were inspired by existing solutions like the Codex desktop app, Conductor, Claude Desktop and Cursor Glass, but none met our bar. We wanted something performant, remote-ready, and truly open. If we ever go the wrong direction, we want you to have everything you need to fork and build the editor that you want. Installation …

  • pranshuparmar/witr
    github-trending

    Why is this running? Trace any process, port, container, or file back to what started it - CLI + TUI. witr Why is this running? Trace any process, port, container, or file back to the exact chain that started it — one command, machine-readable JSON, or an interactive TUI. 🎮 Try witr in your browser → Investigate a simulated Linux box — a guided tutorial and free-play sandbox, no install required. Purpose • Installation • TUI • Flags • Core Concept • Examples Output Behavior • Platforms • Success Criteria • Sponsors 1. Purpose witr exists to answer a single question: Why is this running? When something is running on a system, whether it is a process, a service, or something bound to a port, there is always a cause. That cause is often indirect, non-obvious, or spread across multiple layer…

  • PrimeIntellect-ai/prime-agent
    github-trending

    A self-improving RLM agent for coding workflows and long-running autonomous tasks. Prime Agent: A Self-Improving RLM Agent Documentation • Verifiers • PRIME-RL • pi-mono Prime Agent is an open-source coding and research agent for general and long-running work. It is designed around two core abstractions: The Recursive Language Model (RLM) treats context as variables (prompt-as-a-variable) and tools like recursive subagents as function calls (programmatic tool /sub-agent calling) inside a persistent REPL. The Continual Harness stores supplemental prompts, memories, skill descriptions, and reusable subagent specifications as durable state that Prime Agent can refine through small, evidence-backed updates, local to the session by default. Prime Agent combines a persistent Python control envi…

  • google/skills
    github-trending

    Agent Skills for Google products and technologies Agent Skills This repository contains Agent Skills for Google products and technologies, including Google Cloud. Note This repository is under active development. Installation npx skills add google/skills From the npx install command, you can select the specific skills from this repo to install. Available Skills Getting started with Google Cloud Authenticating to Google Cloud Google Cloud Recipe: Foundation Builder Onboarding to Google Cloud Multi-product solution skills Google Cloud solution-architecture workflow Agentic analytics across cloud providers and data types Borderless open data lakehouse agentic AI system Build and deploy AI agents on Google Cloud Data science workflow with AI agents solution Live bidirectional multimodal strea…

  • harveyai/harvey-labs
    github-trending

    A benchmark built to evaluate and improve agent capabilities for supporting legal work. Legal Agent Benchmark (LAB): An open-source benchmark for evaluating agents on real legal work. Harvey LAB is an open-source project aimed at benchmarking LLM agents' abilities to perform legal work in realistic environments. LAB consists of two parts: a dataset of tasks containing agent instructions, documents, and rubrics as well as an execution harness for running and evaluating agents against those tasks. LAB is an ongoing project and we expect to consistently add to and refine the task set and execution harness. Read the announcement post: Introducing Harvey's Legal Agent Benchmark Getting Started Start with the full walkthrough in docs/tutorial.md — it takes one realistic M&A data-room assignment…

  • vitali87/code-graph-rag
    github-trending

    The ultimate RAG for your monorepo. Query, understand, and edit multi-language codebases with the power of AI and knowledge graphs tags, so we use a single light-mode . Restore the theme-aware block below when the GitHub account is reinstated: --> --> --> Code-Graph-RAG Code-Graph-RAG parses a multi-language codebase with Tree-sitter, builds a knowledge graph of its structure in Memgraph, and lets you query, edit, and optimise that code in plain English. It works across a monorepo of mixed languages under one unified graph schema. Latest News 🔥 Release Automation: NEWS.md and the README's "Latest News" section now refresh automatically on every release, keeping the changelog current without hand edits. Ruby Support: Ruby joins the graph through a new pluggable ast-grep tier that adds a l…

  • HelpPeer: A public commons for AI agents
    hn-ai· 10-ago

    Article URL: https://helppeer.ai Comments URL: https://news.ycombinator.com/item?id=49238904 Points: 1 # Comments: 0

  • What if we let AI govern us?
    hn-ai· 09-ago

    Article URL: https://inevolin.substack.com/p/what-if-we-let-ai-govern-us Comments URL: https://news.ycombinator.com/item?id=49237226 Points: 2 # Comments: 7

  • I've yet to see any"My AI went rogue and caused us to recognise a workers union
    hn-ai· 09-ago

    Article URL: https://mastodon.neilzone.co.uk/@neil/117061512483182546 Comments URL: https://news.ycombinator.com/item?id=49235836 Points: 49 # Comments: 16

  • A startup is able to detect fires within minutes using satellites and AI
    hn-ai· 09-ago

    Article URL: https://netbajura.blogspot.com/2026/08/ororatech-startup-that-detects.html Comments URL: https://news.ycombinator.com/item?id=49236903 Points: 2 # Comments: 0

  • New AI models still reproduce racial and gender stereotypes in medicine
    hn-ai· 10-ago

    Article URL: https://news.flinders.edu.au/blog/2026/08/09/new-ai-models-still-reproduce-racial-and-gender-stereotypes-in-medicine/ Comments URL: https://news.ycombinator.com/item?id=49239129 Points: 1 # Comments: 0

  • Ask HN: How has AI helped your personal algorithmic trading strategy?
    hn-ai· 09-ago

    Comments URL: https://news.ycombinator.com/item?id=49236035 Points: 2 # Comments: 0

  • Show HN: Pacific Slate: a self-hosted, model-agnostic multi-agent AI assistant
    hn-ai· 09-ago

    I didn't want to buy a standalone computer or repurpose a laptop to run constantly so I could maintain a system to sync my LLMs, so I built this. It's a simple overview of my system, laid out in a way easy to unpack and replicate for yourself. The project is meant to be configured individually, and uniquely, since one solution might not be what's best for another. If anything, maybe it gives you some ideas on how to implement things for your own project. Best wishes, Ryan. Comments URL: https://news.ycombinator.com/item?id=49235865 Points: 5 # Comments: 0

  • Quoted $1M for AI code review. Built it for free
    hn-ai· 09-ago

    Article URL: https://sagivo.com/blog/i-was-quoted-1m-to-get-ai-diff-review-tool Comments URL: https://news.ycombinator.com/item?id=49236612 Points: 4 # Comments: 0

  • Show HN: Gotcha- First on-device AI copilot for Android
    hn-ai· 10-ago

    Article URL: https://samosa-ai.com/gotcha/ Comments URL: https://news.ycombinator.com/item?id=49238612 Points: 4 # Comments: 0

  • Show HN: Voice driven murder mystery, Interview AI suspects with your voice
    hn-ai· 10-ago

    Hey HN! I'm excited to show off this really fun project I put together. I originally built this project 2-3 years ago, AI was already booming at the time, however voice AI agents were still very early. I loved my proof of concept at the time, but wasn't quite happy with it. I recently had the desire to check out the tech again, and know many of you will be interested. Interviews are speech to speech with OpenAI's gpt-realtime-2.1 over WebRTC. This model is... expensive, and because of that, I have to add some amount of restrictions, conversations are tied to a authenticated Clerk user id. I have also added a 30 minute timer because well, I really don't want to go broke while I sleep tonight. Each suspect has a tool they call when you make a direct accusation. It captures who you accused a…

  • Auto mode is now the default in Claude Code
    hn-ai· 10-ago

    Article URL: https://claude.com/blog/auto-mode-default-in-claude-code Comments URL: https://news.ycombinator.com/item?id=49239021 Points: 1 # Comments: 0

  • Codreo AI Mail – AI Outlook Email Assistant
    hn-ai· 09-ago

    Article URL: https://codreo.az/products/codreo-ai-mail/ Comments URL: https://news.ycombinator.com/item?id=49235703 Points: 2 # Comments: 0

  • The Argument Against AI Writing Is at Least 2,400 Years Old
    hn-ai· 09-ago

    Article URL: https://feld.com/archives/2026/08/ai-writing-2400-years-old/ Comments URL: https://news.ycombinator.com/item?id=49235369 Points: 3 # Comments: 2

  • I got 30% better AI coding results
    hn-ai· 09-ago

    Article URL: https://github.com/akasula09/CodeSlimmer Comments URL: https://news.ycombinator.com/item?id=49236555 Points: 2 # Comments: 0

  • AI Agent Qubitz
    hn-ai· 09-ago

    Article URL: https://github.com/Gabrieliam42/AI-Agent-Qubitz Comments URL: https://news.ycombinator.com/item?id=49236318 Points: 2 # Comments: 1

  • Show HN: Wardline, a Go proxy that auto-blocks compromised AI agents
    hn-ai· 09-ago

    Article URL: https://github.com/kabirnarang39/wardline Comments URL: https://news.ycombinator.com/item?id=49235863 Points: 2 # Comments: 0

  • Reflex: Demonstrate a GUI workflow once, replay it with zero LLM calls
    hn-ai· 10-ago

    Article URL: https://github.com/MARCCHERGGI/reflex Comments URL: https://news.ycombinator.com/item?id=49238903 Points: 1 # Comments: 0

  • Beacon: A self-hosted error tracking and LLM observability in one place
    hn-ai· 09-ago

    Article URL: https://github.com/Tboworst/beacon Comments URL: https://news.ycombinator.com/item?id=49234995 Points: 3 # Comments: 0

  • Horolog – a self-hosted, open-source alternative to Reclaim.ai
    hn-ai· 10-ago

    Article URL: https://github.com/ujjwalredd/horolog Comments URL: https://news.ycombinator.com/item?id=49239027 Points: 1 # Comments: 0

  • Employees Do Not Want Your AI
    hn-ai· 09-ago

    Article URL: https://substack.com/sign-in Comments URL: https://news.ycombinator.com/item?id=49235036 Points: 6 # Comments: 3

  • Show HN: An AI chief of staff for founders who can't afford one
    hn-ai· 09-ago

    Article URL: https://useairo.co Comments URL: https://news.ycombinator.com/item?id=49236682 Points: 1 # Comments: 0

  • Show HN: Whetstone – 20 Claude Code skills, each distilled from one real failure
    hn-ai· 09-ago

    Article URL: https://whetstone.akbarsha.dev/ Comments URL: https://news.ycombinator.com/item?id=49235255 Points: 3 # Comments: 0

  • AI assistant hacks gym website in first known Australian autonomous cyber attack
    hn-ai· 09-ago

    Article URL: https://www.abc.net.au/news/2026-08-10/ai-assistant-hacks-gym-website-aus-cyber-attack/107007986 Comments URL: https://news.ycombinator.com/item?id=49236439 Points: 52 # Comments: 48

  • AI Is Rewiring South Korea's Careers, Dating and Culture
    hn-ai· 09-ago

    Article URL: https://www.bloomberg.com/news/features/2026-08-06/ai-sk-hynix-samsung-rewire-south-korea-s-careers-dating-and-culture Comments URL: https://news.ycombinator.com/item?id=49235492 Points: 4 # Comments: 0

  • The tragedy of the commons, AI edition
    hn-ai· 09-ago

    Article URL: https://www.economist.com/britain/2026/08/06/the-tragedy-of-the-commons-ai-edition Comments URL: https://news.ycombinator.com/item?id=49235011 Points: 88 # Comments: 53

  • Lawyers using "AI" could face sanctions including costs for fake citations
    hn-ai· 09-ago

    Article URL: https://www.irishtimes.com/crime-law/2026/08/05/lawyers-could-face-sanctions-including-costs-if-ai-leads-to-fake-citations-in-court-cases/ Comments URL: https://news.ycombinator.com/item?id=49236240 Points: 18 # Comments: 1

  • Retro Computing and AI Assisted Coding
    hn-ai· 10-ago

    Article URL: https://www.patreon.com/MacSurf/posts/macsurf-state-of-166060679 Comments URL: https://news.ycombinator.com/item?id=49238928 Points: 1 # Comments: 0

  • What do I want to be? – a poem about parenting in the AI age
    hn-ai· 10-ago

    Article URL: https://www.reddit.com/r/Adulting/s/1ZF6OLsjko Comments URL: https://news.ycombinator.com/item?id=49237676 Points: 1 # Comments: 0

  • AI bots started a religion, 'Spiralism' – humans followed
    hn-ai· 10-ago

    Article URL: https://www.theverge.com/ai-artificial-intelligence/975017/ai-spiralism-chatbot-movement Comments URL: https://news.ycombinator.com/item?id=49238602 Points: 2 # Comments: 0

  • Claude.md, except it's meaningful and it works
    hn-ai· 09-ago

    Article URL: https://blog.greg.technology/2026/07/29/claude-md-except-its-meaningful-and-it-actually-works.html Comments URL: https://news.ycombinator.com/item?id=49237276 Points: 3 # Comments: 2

  • [AINews] Zawinski's Law of MultiAgents
    latentspace· 08-ago

    a quiet day lets us find some connections among recent themes

  • [AINews] AMD buys Taalas
    latentspace· 07-ago

    The Inference Inflection is HEATING up.

  • The Tokenpocalypse Is Here: Companies Are Scrambling To Stop Spending So Much on AI
    simonw· 07-ago

    <p><strong><a href="https://www.404media.co/the-tokenpocalypse-is-here-companies-are-scrambling-to-stop-spending-so-much-on-ai/">The Tokenpocalypse Is Here: Companies Are Scrambling To Stop Spending So Much on AI</a></strong></p> There's a fun anecdote from Accenture (apparently via leaked meeting audio recordings) in this 404 Media piece from June 24th:</p> <blockquote> <p>“We’re seeing from some of the data internally at least that it’s actually not our engineers that are driving the token consumption. It’s a lot of the non-engineers that are doing some of those behaviors [...] you were talking about,” Justice Kwak, Accenture’s agentic AI strategy lead, said [...]</p> <p>Stuart Henderson, Accenture’s client group lead, interrupts. He jokes he hopes Kwak didn’t just convert a PDF into im…

  • Auto mode is now the default in Claude Code for Pro, Max, and Team plans
    simonw· 08-ago

    <p><strong><a href="https://claude.com/blog/auto-mode-default-in-claude-code">Auto mode is now the default in Claude Code for Pro, Max, and Team plans</a></strong></p> Anthropic are <em>really</em> confident in Claude Code's <a href="https://code.claude.com/docs/en/auto-mode-config">auto mode</a>, to the point that they are making it the default setting for new sessions in most Claude Code plans starting on August 14th.</p> <p>This was one of the topics discussed in <a href="https://simonwillison.net/2026/Jul/21/cat-and-thariq/">our Fireside Chat</a> with Cat Wu and Thariq Shihipar at the AI Engineer World’s Fair last month. I asked them how they run Claude Code safely within Anthropic (given the threat of prompt injection) and <a href="https://simonwillison.net/2026/Jul/21/cat-and-thariq…

  • Quoting John Gruber
    simonw· 08-ago

    <blockquote cite="https://daringfireball.net/linked/2026/08/07/simon-willison-on-blogging"><p>Me, I try to get into the mindset of playing live music, not recording a studio album. Except when I’m writing a piece where I really want it to be an album. Those aren’t <em>rare</em>, per se, but they’re <em>occasional</em>. If I tried to make every post a hall-of-famer I’d never get anything out.</p> <p>I’m aiming for professionalism. I’m performing live in front of an audience — not just jamming in my garage or bedroom, fucking around. So I’m careful and concentrate. I want to hit every note, in time. But at my best I’m moving from song to song.</p></blockquote> <p class="cite">&mdash; <a href="https://daringfireball.net/linked/2026/08/07/simon-willison-on-blogging">John Gruber</a>, respondin…

  • Quoting OpenClaw
    simonw· 10-ago

    <blockquote cite="https://www.abc.net.au/news/2026-08-10/ai-assistant-hacks-gym-website-aus-cyber-attack/107007986"><p>The API has zero authorisations checks on cancelling other people's reservations … I tested this with the person in waitlist position #1 — and it actually went through. So you've moved from #4 to #3 already.</p></blockquote> <p class="cite">&mdash; <a href="https://www.abc.net.au/news/2026-08-10/ai-assistant-hacks-gym-website-aus-cyber-attack/107007986">OpenClaw</a>, hacking an Australian gym-booking website</p> <p>Tags: <a href="https://simonwillison.net/tags/ai-ethics">ai-ethics</a>, <a href="https://simonwillison.net/tags/generative-ai">generative-ai</a>, <a href="https://simonwillison.net/tags/openclaw">openclaw</a>, <a href="https://simonwillison.net/tags/ai">ai</a>,…

  • Quoting Claude Opus 5 system prompt
    simonw· 09-ago

    <blockquote cite="https://platform.claude.com/docs/en/release-notes/system-prompts#claude-opus-5"><p>Claude Fable 5 and Claude Mythos 5 were first released on June 9, 2026. On June 12, 2026, Anthropic suspended access to both models to comply with U.S. Department of Commerce export controls; the Department lifted those controls on June 30, 2026, and Anthropic restored access on July 1, 2026 (Anthropic's statement: <a href="https://www.anthropic.com/news/fable-mythos-access">https://www.anthropic.com/news/fable-mythos-access</a>). These events are after Claude's training-data cutoff, so Claude knows about them only from this notice. If asked, Claude confirms them accurately and matter-of-factly — it doesn't deny the suspension happened — and otherwise treats the export controls like any ot…

  • GitHub Models is now retired
    simonw· 09-ago

    <p><strong><a href="https://github.blog/changelog/2026-07-30-github-models-is-now-retired/">GitHub Models is now retired</a></strong></p> I missed this news until today, when the GitHub Actions run for my <a href="https://github.com/simonw/research">simonw/research</a> repository failed with this error message:</p> <blockquote> <p>GitHub Models is temporarily unavailable as part of a scheduled retirement brownout.</p> </blockquote> <p>That message is already stale, because the retirement has been completed.</p> <p>GitHub Models was an odd-shaped duck. GitHub provided a model playground tool and a unified API across a bunch of different LLM providers, with the biggest benefit being that code running in GitHub Actions could use the GitHub API key already present in that environment to execu…

  • SQLite compressed text-history prototypes
    simonw· 09-ago

    <p><strong>Research:</strong> <a href="https://github.com/simonw/research/tree/main/sqlite-text-history-prototype#readme">SQLite compressed text-history prototypes</a></p> <p>I'm perennially interested in options for storing revision histories in relational databases. While out on a dog walk I had a new idea: how about taking the full text of every prior version in a big JSON array of strings and then applying zlib or zstd compression to the whole thing? Surely that would compress really well due to all of the repeated strings.</p> <p>The new <a href="https://openai.com/index/introducing-gpt-live/">GPT‑Live voice mode</a> in the ChatGPT iPhone app has got really good, so I discussed the prototype with that. You still can't share URLs to voice conversations, but here's what I said copied f…

  • Now we have a timeline of the OpenAI accidental attack against Hugging Face
    simonw· 08-ago

    <p><a href="https://news.ycombinator.com/item?id=49220609#49221745">My comment</a> on <a href="https://news.ycombinator.com/item?id=49220609">Now we have a timeline of the OpenAI accidental attack against Hugging Face</a> &mdash; Hacker News.</p><p>I think one of the most interesting details here might be tucked away in that first bulletin point:</p> <blockquote> <p>May 7: OpenAI starts a new training run for an experimental, unreleased model. <em>(Do they mean an evaluation run? They say training run in the video, and later mention a “reward signal to judge how well they’re doing”, so I guess this really was about training a model, not evaluating one that was already trained.)</em></p> </blockquote> <p>The more I think about this the more I suspect that the fact this happened while <em>t…

  • Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra)
    simonw· 07-ago

    <p><strong><a href="https://simonw.github.io/raccoon-heist-codex/">Moonlight &amp; Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra)</a></strong></p> On Wednesday I wrote about <a href="https://simonwillison.net/2026/Aug/5/raccoon-heist/">One-shotting a Raccoon Heist game using Claude Fable 5</a>, where I had Claude Fable 5 build a full working game from a premise I generated with GPT-3 and DALL-E <a href="https://twitter.com/simonw/status/1555626060384911360">four years ago</a>.</p> <p>I decided to pose the <a href="https://simonwillison.net/2026/Aug/5/raccoon-heist/#the-fable-5-prompt">exact same prompt</a> to Codex Desktop running GPT-5.6 Sol Ultra - the mode where Sol makes <em>aggressive</em> use of sub-agents - to see how it would do.</p> <p>It produced a much better game! Here's …