← Tech / AI / IT Monitor Index Tech / AI Generated 2026-05-28 19:30 UTC

Tech / AI / IT Monitor

May 28, 2026 · Based on tweets from the last 24 hours · 162 tweets analyzed · model: ollama-cloud/glm-5.1:cloud

Executive Summary

Claude Opus 4.8 launched to widespread developer attention, with early reports highlighting significantly reduced laziness, 4x fewer coding faults, and a restructured /fast mode pricing (2x cost instead of 6x) that makes it competitive with GPT 5.5. Hermes Agent rapidly integrated Opus 4.8 alongside a new MCP Catalog and Krea 2 image generation support. On the hardware front, NVIDIA's DGX Spark is gaining traction as a serious local AI compute device, while tinygrad began tearing down RTX PRO 6000 Blackwell GPUs for its tinybox pro. Open model usage in production reportedly 3x'd over the past month relative to closed models, signaling a shift in enterprise AI deployment preferences.

Key Events

Analysis

Patterns: The dominant theme is the rapid maturation of AI coding agents and the models powering them. Opus 4.8's launch dominated discourse, with developer attention focused on its /effort toggle, dynamic workflows, and dramatically restructured pricing — suggesting Anthropic is directly responding to GPT 5.5's cost-performance ratio. Meanwhile, the "harness vs. model" debate (per @kunchenguid) gained traction: developers increasingly recognize the model is the primary value driver, and the agent harness just needs to not hold it back.

Escalation trend: Open models are gaining serious ground in production. The 3x increase in open model usage at Factory and continued excitement around Qwen, Nemotron, and MiniMax suggest the closed-model dominance narrative may be softening. Pricing sensitivity is escalating — Anthropic's 5x cost premium (per @jezell) is creating real churn pressure.

Hardware divergence: DGX Spark is being positioned as the "serious compute" choice vs. Mac's "desk aesthetic," while tinygrad's RTX PRO 6000 teardown signals continued demand for custom CUDA infrastructure. The hardware startup space is being warned against tight model-hardware coupling, as models evolve far faster than PCBs.

What to watch next: Opus 4.8 real-world benchmark results vs. GPT 5.5 over the coming week; whether Hermes Agent's MCP Catalog drives broader MCP adoption; NVFP4 quantization developments from @Ex0byt; and whether the open-model production adoption trend continues to accelerate.

Tweet Feed

Claude Opus 4.8 Release & Capabilities

@nummanali · 2026-05-28T16:55

Opus 4.8 → tweet link

@kunchenguid · 2026-05-28T17:03

opus 4.8 is here!

biggest surprise that's easy to miss is that /fast mode with 4.8 is now only 2x cost (instead of 6x before) at 2.5x speed

it's no longer prohibitively expensive, and makes it finally able to compete with gpt 5.5 fast mode for my interactive work → tweet link

@nummanali · 2026-05-28T16:59

I think I might go full Claude with the new Opus 4.8 release

  • stays on track for long running tasks
  • behaves like an experienced engineer
  • no need for constant my check ins

Sounds a lot like codex → tweet link

@nummanali · 2026-05-28T17:01

Opus 4.8 is 4x less likely to create faulty code!

That's a good sign !!

I simply could let trust Claude in Enterprise repos

Will go full ham on it over the next week → tweet link

@nummanali · 2026-05-28T17:03

Don't go below XHigh reasoning on Opus 4.8 for consistent performance

Thanks @danshipper ! → tweet link

@Teknium · 2026-05-28T18:54

Opus 4.8 the least lazy model ever? → tweet link

@victormustar · 2026-05-28T17:44

opus 4.8 new ultracode /effort don't know if it's good but this animation is wild 😅 → tweet link

@TheAhmadOsman · 2026-05-28T17:48

I don't care about Opus 4.8 tbh, if I pay for it today it'll be nerfed in a week anyway lol → tweet link

@nummanali · 2026-05-28T18:33

When did the /effort toggle in CC become so hot!? → tweet link

@nummanali · 2026-05-28T18:04

Dynamic Workflows is underrated, really impressed by my first usage of it

Enable it by default with /effort ultracode → tweet link

@nummanali · 2026-05-28T18:09

Love the fullscreen Claude Code TUI

See the new dynamic /workflows UI in the video, really impressed!

Two-Stage pipeline: 1. Orient on scope 2. Fan our 12 parallel agents

All managed programmatically by Opus 4.8 → tweet link

Hermes Agent & Nous Research

@Teknium · 2026-05-28T17:43

Opus 4.8 is now supported in Hermes Agent ^_^ → tweet link

@Teknium · 2026-05-28T19:05

MCP Catalog in Hermes Agent now to have preconfigured already available MCP's from trusted sources! → tweet link

@Teknium · 2026-05-27T20:52

Check out Krea's latest image generation model in Hermes Agent, we've added direct support for their API!

Access today with hermes update or wait for our next major release version coming soon → tweet link

@Teknium · 2026-05-28T08:21

Added Krea 2 support to the Fal image gen tool provider options

Coming soon to Nous Portal subscription access! → tweet link

@Teknium · 2026-05-27T22:20

Okay if you are on the latest Hermes Agent update and use OpenAI OAuth, how is it doing now? Reliably working yet? → tweet link

AI Coding Agents & Developer Tools

@nummanali · 2026-05-28T08:54

I absolutely love the sidebar in the Codex App

It is the first UX of it's kind that gives a complete 360 of what your agent is doing - Task Status - Sub Agents - Commands - Changes - Tools used (Linear etc)

The apps performance has greatly improved too, really enjoying it! → tweet link

@kunchenguid · 2026-05-27T19:56

"claude code/codex is amazing look at what it did for me"

for most of such tweets, you can swap the name to any other oss agent and it would be equally amazing

what's amazing is the model

the harness just need to give the model access to the world and not hold it back → tweet link

@jezell · 2026-05-28T17:11

Coding Agent 101. You need to have a test driven mindset period. If you don't, you might as well not even use AI because you are a worse, more expensive, version of your former self. → tweet link

@thdxr · 2026-05-28T03:06

one thing i've been enjoying doing is any medium to large size task, i'll ask OpenCode to split it into groups of work we can tackle one at a time

then we go through the list one by one

often times i can confirm tests and commit after each group so it's like a save point → tweet link

@TheAhmadOsman · 2026-05-28T07:12

Agentic harnesses are so bloated right now btw

This is the slowest and worst they'll ever be → tweet link

@steipete · 2026-05-28T13:52

Every claw release spins up hundreds of CI machines to QA test and eventually creates a ledger. → tweet link

@steipete · 2026-05-28T13:26

Hit GitHub's rate limit one too many times, so I built octopool: a Cloudflare Worker that pools your team's PATs + GitHub App installations behind a shared read cache.

Self-host on Cloudflare. Drop-in gh shim. → tweet link

@steipete · 2026-05-28T18:51

build the thing that builds the thing. → tweet link

@thdxr · 2026-05-28T14:03

does anyone i know have a connection with someone who works on LetsEncrypt? have a concept i want to run by them → tweet link

@hnasr · 2026-05-28T14:02

Why It took Postgres years to do async → tweet link

@nummanali · 2026-05-28T17:28

The future of software development is with Linear

I have no doubt that it is a giant and has become a titan

Checkout Diffs, I've used for a few weeks and it's a banger → tweet link

@jezell · 2026-05-28T03:58

PDFs in your the terminal. Which other agents are doing this? Shouldn't they all? → tweet link

Open Source & Local AI Models

@TheAhmadOsman · 2026-05-28T09:02

Models I am excited to run locally these next few months

@sudoingX · 2026-05-28T07:48

the best model on a single 3090 is still qwen 27B dense. nothing comes close. → tweet link

@TheAhmadOsman · 2026-05-28T06:35

Just realized this article has almost half a million views

Opensource AI Will Win → tweet link

@badlogicgames · 2026-05-27T19:20

jesus, qwen3-tts is FANTASTIC. going for full local stt/tts/llm with parakeet, qwen3-tts, and gemma 4 via llama.cpp for my little robot. excite, excite! → tweet link

@badlogicgames · 2026-05-28T00:01

let's see if i get a workable Rust/MLX qwen3-tts engine by the morning. took [reference] and added kv cache, now adding 6-bit quant and MLX-format support. → tweet link

@badlogicgames · 2026-05-27T20:16

gpt 5.5 designing a binary protocol. oh no... → tweet link

Hardware & Compute

@sudoingX · 2026-05-28T16:41

people keep asking me dgx spark vs mac mini/studio. let me clear it.

if you want serious local AI work (fine-tunes, agentic loops, multi-modal, nvidia ecosystem like nemo / NIM / cuda kernels), dgx spark wins. nothing comparable at this tier for cuda + bandwidth + nvidia native stack.

mac mini/studio is great if you want a beautiful desk and a general dev machine. but the moment your workload is "actually run modern AI agents + fine-tune locally," macs hit walls cuda doesn't.

different lanes. one is for the desk aesthetic. the other is for the compute. → tweet link

@sudoingX · 2026-05-28T18:46

people keep asking me dgx spark vs halo strix and the truth is i don't have a halo strix in hand.

until i can test it the same way i tested the 1080 and the dgx i'm not putting a take out there. → tweet link

@sudoingX · 2026-05-28T15:40

my dgx spark is pulling an all nighter on biomechanical 3D visualizations via hermes agent /goal mode running qwen 3.6 35B-A3B at Q8.

this is what owning your cognition looks like. your thoughts stay on your hardware. your work runs on your schedule.

dgx spark is the most underrated machine in ai right now and it's not even close. → tweet link

@tinygrad · 2026-05-28T02:20

1 of 8 NVIDIA RTX PRO 6000 Blackwell being torn down for tinybox pro install. Don't worry, it's only $10,000 if you shear one of the ribbon cables. → tweet link

@tinygrad · 2026-05-28T18:24

tinygrad will write that C for you. Our new driver compiles all interaction with the GPU to C, so once it's running the CPU does next to nothing. → tweet link

@Ex0byt · 2026-05-27T22:08

Wow.. looks like you folks really like this one… we'll do more on nvfp4 → tweet link

@TheAhmadOsman · 2026-05-28T03:11

Please stop pitching me hardware startups that are tightly coupled with models

No, printing model architectures on hardware isn't smart, it's a waste of PCBs and memory

GTX 1080s from 10 years ago could run today's models, but a model on a PCB today won't be used in 10 years → tweet link

OpenAI & GPT 5.5

@gdb · 2026-05-27T22:42

Underappreciated how capable GPT-5.5 is at cybersecurity: → tweet link

@gdb · 2026-05-28T02:57

please report any ChatGPT bugs in the thread below — team (and codex) working super hard to resolve them: → tweet link

@gdb · 2026-05-27T20:37

Codex for parallel browser-using subagents: → tweet link

@gdb · 2026-05-27T20:27

bring-your-own MCP servers: → tweet link

@gdb · 2026-05-27T23:49

OpenAI for self-improving tax agents: → tweet link

@gdb · 2026-05-28T17:10

How @CGRTeams is working with @OpenAI to improve motorsports performance: → tweet link

Industry & Enterprise AI

@LinusEkenstam · 2026-05-28T18:38

Pre, Seed, A,B,C,D,E,F,G…. series H 🫠

$65.000.000.000 in fresh capital

at a whopping

$965.000.000.000 valuation → tweet link

@jezell · 2026-05-28T06:47

Anthropic costs 5x, so 35% more revenue = getting 73% fewer tasks done than OpenAI. Good luck with the churn Dario. → tweet link

Benchmarks & Research

@victormustar · 2026-05-28T15:22

RT @adithya_s_k: Introducing Repo2RLEnv

Turn any repository into runnable, verifiable coding environments built from real PRs and commits… → tweet link

@LinusEkenstam · 2026-05-28T16:03

I'm extremely interested too see how this will work in real world scenarios.

If this version of self-improvement can deliver on all the metrics provided in the paper. We are truly in for some interesting times ahead.

What I personally like with this approach is that it's open → tweet link

@victormustar · 2026-05-28T11:14

RT @skalskip92: RF-DETR is now available in @huggingface transformers

state of the art in both detection and segmentation, outperforming Y… → tweet link

@TheAhmadOsman · 2026-05-28T07:21

RT @hardmaru: For over a decade, we've accepted that end-to-end backprop is the only way to train deep networks. But holding the entire net… → tweet link

@victormustar · 2026-05-28T12:54

Hugging Face just shipped one of the most requested features on /models 🕺

a "Base only" toggle that hides all the finetunes, quants, adapters & merges! → tweet link

Reinforcement Learning & PufferLib

@jsuarez · 2026-05-28T13:40

PufferLib trains small task-specific reinforcement learning models at up to 20M steps/second in 5,000 lines of CUDA C on a single GPU. Stop using awful DSLs just because they are pretending to be Python! → tweet link

@jsuarez · 2026-05-28T00:23

The problem with RL algo research on PufferLib is that the baseline is too damn good. Multi-million step sparsity? Cool, that will be a hard benc--oh solved out of the box. This has happened 3 times in the last 3 days. → tweet link

@jsuarez · 2026-05-28T03:10

PufferLib has a mandate from heaven to make reinforcement learning agents trainable on your laptop. Source: the Pope. → tweet link

xAI & Grok

@nummanali · 2026-05-27T21:36

Grok Build was advertised as a massive first block on my For You page

What I can discern: - The new @xai team is in 6th gear and moving fast - X will be the conversion point for its 500M+ users - @grok will be special after its Colossus 2 run

A new lead player soon? → tweet link

Misc Dev Tools & Platforms

@mraleph · 2026-05-28T10:36

RT @FlutterDev: Flutter in your pocket, on your web browser, and now... in your car 🛠️

See how Flutter is helping Toyota deliver an intuit… → tweet link

@Teknium · 2026-05-28T05:49

Is this why they are overloaded? → tweet link