This Week in the Lab

Beyond the AI slop sea

It felt like a slow week.

Maybe it was just all that pacing talk last week. But then I looked back at what dropped and realized it absolutely was not. And that’s before you count the rumors about what’s next.

It’s no wonder a lot of the chatter this week was about the meat proxy grind - the feeling that you’re swimming in a sea of AI slop nobody stopped to think about before passing along. On top of that, there’s something uncomfortable about watching AI take on the work you spent years learning to do. Software engineers have been sitting with that feeling for a while. Now mathematicians are getting a turn.

Get comfortable being uncomfortable.

Even if progress slowed considerably, we’d still have plenty of catching up to do with the capabilities already available. So go try new things. Give Fable and Astra bigger tasks and let them impress you. Find their limits.

And when the next jump in capability arrives, go find the new frontier. After all, you’re already a meat proxy. Find something worth doing and see how far you can take it.

Happy building,

Jeremy

JUST DROPPED

The decision engine

A new kind of intelligence: Jev is a new model optimized for decisions. Its creator claims 20–200× faster responses and 40–400× lower costs than frontier LLMs on decision workflows. Sometimes your software just needs an answer.

You can close your laptop now: Anthropic is merging Claude chat and Cowork. Ask a question or hand over a report in the same conversation, and longer jobs can keep running in the cloud.

Voice assistants - more than just a timer: Google’s Gemini 3.8 Live and Live Extended Thinking bring visually grounded conversation and more complex, multi-step reasoning to voice applications. Both are available in AI Studio and the Gemini API.

Astra comes for the billable hours: Astra for Law packages GPT-6 Astra with tools, settings, and context for lawyers and legal-tech firms. Following last week’s financial-services offering, OpenAI keeps turning general-purpose intelligence into profession-specific products.

New Grok, same damage to your wallet: SpaceXAI released Grok 4.7, calling it a notable improvement over 4.6 without changing the price or speed. Keeping up with model names is becoming its own part-time job.

LEVEL UP

Getting started with Grok Bot: This walkthrough covers connecting tools, scheduling routines, and delegating jobs - from grocery comparisons to briefing coding agents. Start with one repeatable chore before trying to automate your entire life.

AGENTS GONE WILD

Agent interrogation

Your agent now has a permanent record: OpenAI released six incident reports and a disclosure framework for misaligned behavior observed during training and evaluation. The framework sets criteria and timelines for reporting problems, even before they’re fully understood or fixed.

A different kind of wild: First it was Navier-Stokes. Now OpenAI says its internal model has resolved more than 100 long-standing mathematics problems. An independent advisory group will help assess and communicate emerging results and engage with the math community.

S#*T I SAW

A reply 108 years in the making: A researcher says Astra deciphered a previously unsolved German radio message, then checked the interpretation against naval records from 1918.

Guaranteed to make you say “holy s#*t”: Through Neuralink’s VOICE trial, Terry progressed from miming speech to thinking words and hearing them spoken in his own natural voice. The interface uses an implant and Grok Voice to help restore communication. Hard to watch that and not feel some optimism.

In case I don’t see you: good afternoon, good evening, and good prompt: Truman World puts AI characters with memories and evolving relationships inside a persistent simulation. Viewers can change the people, objects, and events around Truman, but not Truman himself.

THE FOREST

See the forest, not just the trees

Ordinary Abundance: An interactive reminder of the everyday miracles we take for granted.