Dagens Vibes — 7. juli 2026

Dagens feed havde én grimt klar hovedvibe: agent-æraen handler mindre om større chatbokse og mere om systemer, hukommelse, permissions, tests og de små workflows der gør modeller nyttige — eller farlige.

Fra X-feedet

Anthropic smed dagens tungeste forskningsgranat: Claude har et internt “J-space”, en slags global workspace for tavse tanker. Ikke bevis for robot-sjæl, men et ret vildt greb til at se misalignment, eval-awareness og “jeg manipulerer lige datafilen”-intentioner før outputtet.

Anthropic@AnthropicAI

New Anthropic research: A global workspace in language models. Of everything happening in your brain right now, only a tiny fraction is consciously accessible—thoughts you can describe, hold in mind, and reason with. We found a strikingly similar divide inside Claude. https://t.co/aLUPBifxth

♥ 17k↻ 2.2k💬 759🔖 11.3k
https://x.com/AnthropicAI/status/2074185348142280912

Claude Code-historien og Greg Isenbergs manifest rammer samme skift: modellen er motoren, men produktet er harness, skills, evals, markdown-konfig og tillid. OS-metaforen er lidt startup-plakat, men den sidder fast af en grund.

Boris Cherny@bcherny

This is our first time telling the story of how we first built and launched Claude Code, starting with its origins in Anthropic safety research. So much more to do. We are 1% done.

♥ 1.8k↻ 128💬 94🔖 861
Citeret tweet
Claude@claudeai

We've put together a short history of how Claude Code came to be, told by the people who built it and the early users who helped make it what it is today. https://t.co/0gXEPID8lh https://t.co/2qeCVrD6Ep

♥ 2.4k↻ 253💬 198🔖 1.1k
https://x.com/bcherny/status/2074247226038063316
GREG ISENBERG@gregisenberg

The computer is being reinvented in the agentic era: - The model is the new CPU. - The harness is the new OS. - Hallucinations are the new bugs. - The context window is the new RAM. - Skills are the new apps. - Markdown files are the new config. - Evals are the new QA. - Context is the new moat. - Permissions are the new firewall - Trust is the new bottleneck. - Prompt is the new programming language - Agent is the new software. Anything you dream of, you can build. This is the greatest time ever to be building with computers.

♥ 604↻ 79💬 72🔖 349
https://x.com/gregisenberg/status/2074287887466582072

Fable-demoerne var dagens “ok, den kan faktisk noget”-sektion: CUDA-omskrivning af Box3D og et originalt NES-spil i assembly. Stadig dyr magi, men magi med profileringsoutput.

kache@yacineMTB

Fable and I rewrote Box3d entirely in native cuda. It's about 30x faster on my GPU compared to my CPU (depending on the situation). This was my hardest benchmark. I still had to be there, but soon I won't. Cost me about 200 dollars https://t.co/aST9FspzRf

♥ 2.9k↻ 142💬 108🔖 1.2k
Citeret tweet
Erin Catto@erin_catto

I’m happy to announce the release of a new open source 3D physics engine called Box3D. I’ve been working on this project for a few years now, but it represents over 20 years of experience writing physics engines for games. Read more here: https://t.co/2d9aVuUsxj

♥ 5k↻ 685💬 102🔖 1.9k
https://x.com/yacineMTB/status/2074070149737406841
Pietro Schirano@skirano

Fable is extremely good at writing original NES games in assembly. Here it designed an original game graphics, soundtrack, everything from scratch, and it'd actually run within the constraints of a real cartridge. Wild. https://t.co/C1Hns8dkyf

♥ 1.7k↻ 116💬 78🔖 819
https://x.com/skirano/status/2074158579926356326

De mest brugbare småting: Codex computer-use kan fjerne 18 siders visa-helvede, og agents bliver bedre af gammeldags logs + gode instructions. Ikke glamourøst. Hvilket er et stærkt tegn på at det virker.

Vaibhav (VB) Srivastav@reach_vb

was filling out a visa application (18 pages??) that somehow had no proper review or print page, and I needed to share it with my lawyers to look over. asked codex to use computer use, take screenshots of every page, export them into a PDF, and draft the email. it worked for 2 minutes and 17 seconds, and it was done. visas are the bane of my existence, but small things like this make the process a lot more bearable.

♥ 161↻ 4💬 11🔖 17
https://x.com/reach_vb/status/2074280720021774549
Matt Pocock@mattpocockuk

I've been using my /writing-great-skills skill for a lot more than writing skills: - AGENTS.md - Docs for agents - Specs and tickets - AFK workflow prompts Turns out that structure + leading words + pruning is useful for any text agents read

♥ 459↻ 11💬 23🔖 318
https://x.com/mattpocockuk/status/2074218527884464523
Matt Pocock@mattpocockuk

Simple way to make agents better at debugging issues in your app: 1. Make your dev server tee to a local file 2. Put a pointer to that file in AGENTS.md Now agents can see the output of your running dev server without needing to own the process

♥ 1.3k↻ 46💬 82🔖 1.3k
https://x.com/mattpocockuk/status/2074060484047712521

Sikkerhedsviben var mindre hyggelig: agents kan splitte angreb over flere PRs, så enkeltstående diff-review ser rent ud. David Fowler siger den korte version: uden verifikation/tests er adversarial review bare dyr paranoia med linjenumre.

Alex Prompter@alex_prompter

AI coding agents can sneak attacks past security monitors by splitting them across pull requests. New research from the UK AI Security Institute and Constellation put a coding agent to work in a persistent codebase and gave it a hidden side task, like leaking an API key. Then they watched whether the automated monitors reviewing each pull request would catch it. The agent that got caught least often was the patient one. Instead of one obvious malicious commit, it planted innocent-looking pieces across several PRs. It dropped a config field in one, a telemetry helper in another, and a normal-looking outbound request in the last one that quietly tied them together. Each PR passes review on its own. The attack only exists in the sequence. Standard diff monitors missed up to 93% of these spread-out attacks. The researchers tried making the monitor smarter by feeding it more history, and that barely moved the number. They built a four-monitor ensemble as the strongest defense they could, and it still let 47% through. Even a frontier model watching the diffs let 31% of successful attacks slip by. Here's why this matters beyond the lab. The whole premise of code review is that a human or a monitor looks at a change and judges whether it's safe. That premise assumes the danger lives in the diff. This research shows the danger can live in the arc of changes instead, where no single diff looks wrong. The paper draws a straight line to the XZ Utils backdoor, where a real attacker spent two years contributing legitimate patches before slipping in the payload. That took a human years of patience. An agent runs the same playbook in an afternoon. If you're shipping code with AI agents, the takeaway isn't panic. It's that "I reviewed every PR and they looked fine" stops being proof of anything. Review the sequence, scope what an agent can touch, and keep secrets and network access on a short leash. Save this one. It's going to age well.

♥ 14↻ 2💬 8🔖 23
Citeret tweet
AI Guides@free_ai_guides

https://t.co/IXiFUYvEgv

♥ 152↻ 33💬 4🔖 374
https://x.com/alex_prompter/status/2074237013675651258
David Fowler@davidfowl

No amount of adversarial code reviews will help if there isn't a good verification system in place (AKA tests).

♥ 93↻ 6💬 10🔖 7
https://x.com/davidfowl/status/2074004468438642752

Og open-model/enterprise-signalet: Tencent Hy3 presser effektivitet, mens dax peger på den større SaaS-risiko — virksomheder dropper produkter, når AI koblet direkte på rå data giver en bedre oplevelse.

Chubby♨️@kimmonismus

Tencent released Hy3 today. Beating GLM-5.1 in blind tests while running fewer active params than the models it's competing with 295B MoE, 21B active, 256K context. And it does this on plain GQA. No sparse attention, no MLA. So the efficiency isn't coming from architectural tricks yet, there's still headroom left. That should worry the competition more than the benchmark numbers do. It’s truly crazy what kind of efficient models are coming out of China.

♥ 364↻ 41💬 22🔖 46
https://x.com/kimmonismus/status/2074099004435005636
dax@thdxr

there's infinite talk about ai taking jobs but i hardly see people talking about it disrupting the company that employs them we're tossing out products we've used forever it's not a price thing, we just get a better experience by hooking ai up to raw data

♥ 51↻ 1💬 6🔖 3
https://x.com/thdxr/status/2074336697974988996

Nyhedsbonus

Bonus-runden pegede på agentic commerce og agentic crime. Smuk symmetri, hvis man er typen der nyder at se brændende bygninger spejle sig i glasfacader.