Dagens Vibes — 2. august 2026

Dagens hovedvibe: Astra gør AI-matematik maskincheckbar, mens resten af feedet bygger arbejdsformen rundt om modellerne — målbare constraints, mindre specs, specialiserede subagents og bedre kurskorrektioner.

Fra X-feedet

OpenAIs interne Astra-model har leveret ti nye resultater på gamle problemer i matematik og teoretisk datalogi. Manuskripterne er menneskeforberedte, men argumenterne modelgenererede og formaliseret i Lean; uafhængig faglig granskning mangler stadig, men det her er dagens tunge signal.

Ten advances in mathematics and theoretical computer scienceOpenAI · argumenter, manuskripter og Lean-certifikaterhttps://openai.com/index/ten-advances-in-mathematics/
Chubby♨️
Chubby♨️@kimmonismus

HOLY: OpenAI says its *unreleased* Astra model (GPT6?) produced ten advances on long-standing open problems across mathematics, quantum complexity and theoretical computer science. Among them: – The first explicit non-sofic group – Connes’s rigidity conjecture disproved – Quantum parallel repetition proved for general two-player entangled games – Ehrhart’s volume conjecture proved – The first improved general sphere-packing exponent since 1978 OpenAI says the core arguments were generated by Astra. The model then formalized the proofs in Lean, producing machine-checkable certificates alongside a 249-page manuscript. The successful solution runs would cost only roughly $2,000 in tokens at Sol API rates. Scientific reasoning is becoming a genuine model capability much much faster than most people expected. I am so freaking hyped. Breakthroughs every day. The day before yesterday, an 80% price cut for Terra and Luna; yesterday, the DeepSeek 4 flash release with insane evaluations and prices. Today, more breakthroughs with an unreleased model. I love it! OpenAI is on such a great run!

OpenAIs Astra-resultater
♥ 3.623↻ 272💬 144🔖 803
https://x.com/kimmonismus/status/2083484340512604323

“LLM’er er en ny slags compiler” er mindre bombastisk, end det lyder: giv agenten målbare constraints, lad den prøve sig frem, og verificér output automatisk. Pixelperfekt DOCX-rendering via Word Online er et godt eksempel på, at harnesset er halvdelen af intelligensen.

kache
kache@yacineMTB

LLMs are a new kind of compiler https://t.co/wrQfyb9rTp

Eksempel på autoverificerbar agentudvikling
Citeret tweet
Baudouin@netapy

This technique is a much bigger deal than it seems For a few $ you can fully replicate any product you have access to just by bruteforcing agents against it We did this internally at Ordalie with microsoft word online to get pixel perfect docx rendering and it worked insanely well It's about expressing the constraints and levers in the harness. agents excel at tasks they can autoverify -- expect much more of this!

♥ 427↻ 10💬 6🔖 289
♥ 308↻ 9💬 19🔖 143
https://x.com/yacineMTB/status/2083521294860001737

Birgitta Böckeler skiller spec-first, spec-anchored og spec-as-source ad — og finder mest markdown-bureaukrati, når workflowet bliver større end problemet. Små iterative specs overlever hypen bedst.

SDD
Understanding Spec-Driven DevelopmentMartin Fowler · Kiro, spec-kit og Tessl i praksishttps://martinfowler.com/articles/exploring-gen-ai/sdd-3-tools.html
Matt Pocock
Matt Pocock@mattpocockuk

Great guest article from Birgitta on Martin Fowler's site: https://t.co/i3KmU0IdLM

♥ 30↻ 0💬 1🔖 61
https://x.com/mattpocockuk/status/2083563358180012470

Dagens bedste smagsmaskine: saml UI’er, lad en agent dissekere dem ugentligt, og gem kun de stærke mønstre som DESIGN.md-skabeloner. Designsmag som vedligeholdt kontekst, ikke som endnu en magisk prompt.

Machina
Machina@EXM7777

over the past year i've slowly acquired real taste for frontend design... all i did was collect inspirations on one side and working UI components on the other here's the loop i'm using for that: > every design you like on X becomes a bookmark > X MCP plugged into my Hermes agent walks the bookmarks weekly > Hermes spawns a Gemini 3.1 Pro subagent to tear each design down > the keepers become DESIGN․md templates on top of that, a cron job searches Reddit and GitHub for "UI component" and "frontend skill" hundreds of indie devs are shipping tools for hyper specific features, and the good ones get folded into my frontend design plugin for Claude Code

♥ 168↻ 9💬 23🔖 235
https://x.com/EXM7777/status/2083551514648518710

Sol Advisor formaliserer den orchestration, Batty allerede læner sig imod: Sol holder arkitektur og accept, Luna tager rutinen, Terra den brede implementering, og en frisk Sol læser diffen. Mere interessant end “multi-agent” som mærkat er den eksplicitte routing og obligatoriske review.

SA
sol-advisorGitHub · capability-routed Codex-plugin med frisk reviewhttps://github.com/DannyMac180/sol-advisor
Dan McAteer
Dan McAteer@daniel_mac8

Codex users: here's an amazing way to take advantage of the efficiency and capability of GPT-5.6 Luna Max. It's called 'sol-advisor'. 1. GPT-5.6 Sol High as orchestrator 2. GPT-5.6 Luna Max as implementer for routine tasks 3. GPT-5.6 Terra Max as implementer for complex tasks 4. Fresh GPT-5.6 Sol instance as reviewer A free and open-source plugin for Codex. Installation instructions below 👇 I dare you to install it, and tell me if you hit your weekly usage limits using 'sol-advisor'. I doubt you can.

Sol Advisor i Codex
♥ 1.943↻ 148💬 54🔖 3.505
https://x.com/daniel_mac8/status/2083607027813662810

Og når et langt agentjob driver: stop, skriv én ordentlig kurskorrektion med hjælp fra en sidechat, og fortsæt derfra. Kontinuerlig mikrostyring er bare promptversionen af at rive i rattet.

eric provencher
eric provencher@pvncher

Pro tip with codex. If you’re working on very long ambitious tasks, you’re better off interrupting and formulating a detailed course correction prompt, than continuously steering. Side chats help a lot with making those prompts!

♥ 432↻ 17💬 19🔖 212
https://x.com/pvncher/status/2083602795391782927

Pi-påmindelsen: /session viser tokenforbrug og pris, mens /compact kan få en besked med, der styrer selve komprimeringen. Context management er stadig håndværk; vinduet rydder ikke sig selv af høflighed.

Pi
Pi@pidotdev

Pi helps you use tokens as efficiently as possible. Here’s a cheatsheet of commands to try: - /session: shows token use, cost, messages - /compact: summarizes context. Send a message with it to steer the summary - tune compaction in your settings json file for more control https://t.co/3V2boXjMQ1

Pi-cheatsheet til session og compaction
♥ 503↻ 21💬 7🔖 217
https://x.com/pidotdev/status/2083515588551577926

Nyhedsbonus

Coinkite udvidede i nat sin Coldcard-advarsel til flere modeller: en fejl reducerede entropien i device-genererede seeds, og eksisterende seeds bliver ikke sikre af en firmwareopdatering. Firmaet har frigivet fixes, men anbefaler nye seeds og migrering af midlerne; open source reddede ikke de gamle wallets af sig selv.

RNG
Coldcard Security AdvisoryCoinkite · opdateret 1. august 2026https://blog.coinkite.com/coldcard-mk3-seed-generation-warning/