Executive Summary
The last 24 hours saw significant activity around open-weight AI models and local inference, with Qwen 3.6-35B emerging as a community-favorite model and MiniMax teasing M3. OpenAI's Codex platform experienced multiple infrastructure issues (OAuth, subscription endpoints) that were patched throughout the day, while the open-source Hermes Agent framework rapidly grew to 7,400+ community members. A notable debate flared on AI coding agent viability, with multiple developers arguing that Codex and Claude Code suit "vibe coding" while Cursor remains better for production software. Sam Altman announced a $250M OpenAI Foundation commitment, and YC revealed it has built 350+ internal agent tools.
Key Events
- Qwen 3.6-35B gains traction as top open-weight model, with community polls and benchmarks favoring it; MiniMax M3 teased as next release → link
- OpenAI Codex suffers OAuth and subscription endpoint issues, patched via
hermes updateand v1.15.11 of OpenCode → link - Sam Altman announces $250M OpenAI Foundation commitment for measurement, transition support, and broadly shared prosperity → link
- GPT-5.5 and GPT-5.3-Codex sunset announced for June 2nd in Codex fleet management → link
- vLLM merges Rust frontend, addressing CPU-side frontend bottleneck as GPU speeds improve → link
- Ex0byt releases Cortex-Conv, a bio-plausible neural model 43x smaller than predecessors at 96.8% MNIST accuracy, running in-browser via WebGPU → link
- Marlin-2B open video VLM released (Apache 2.0) for timestamped video understanding → link
- Y Combinator reveals 350+ internal agent tools and self-improving skill infrastructure → link
- GitHub impersonation attack flagged: @LLMMart issued unsigned commits using a developer's email; Vigilant Mode recommended → link
- Micron and SK Hynix both join $1 trillion market cap club within a 24-hour span → link
- OpenCode x MiMo V2.5 launched with 1M context, reasoning, text, and image — free for limited time → link
Analysis
Local AI is hitting an inflection. Multiple posts document runnable agentic workloads on decade-old consumer GPUs (GTX 1080 8GB). The gap between perceived and actual local capability is now primarily a knowledge gap (right quant, engine, and flags), not a hardware gap. This threatens cloud API dependency and reinforces the open-weight community's momentum.
Coding agent positioning is stratifying. A clear three-tier narrative is forming: Codex/Claude Code for quick "vibe coding," Cursor for production-grade work, and local/open agents for ownership of the cognition stack. The SWE-Bench benchmark controversy (GPT-5.5 vs. Claude Sonnet discrepancy) underscores that benchmarks poorly capture real-world coding quality — mirroring @thdxr's analogy that disabling syntax highlighting makes an editor faster on paper but worse in practice.
Open-source ecosystem acceleration. The Hermes Agent community's organic growth to 7,400+ builders, combined with vLLM's Rust merge, LanceDB's world model framework call, Holocron as a Mintlify alternative, and PrismML's 1-bit image models, signals open-source AI tooling is compounding faster than proprietary platforms can lock in.
What to watch: Qwen 3.7 open-weight commitment (or lack thereof), Codex fleet transition after GPT-5.2/5.3 sunset on June 2, MiniMax M3 capabilities, and whether local agent frameworks can sustain the current velocity on consumer hardware.
Tweet Feed
AI Model Releases & Updates
@alexocheema · 2026-05-27T17:34
Qwen3.6-35B winning. MiniMax M2.7 also popular in the comments (and M3 coming soon). → tweet
@Teknium · 2026-05-26T21:25
Got that latest Qwen for ya in Hermes Agent now! → tweet
@SkylerMiao7 · 2026-05-27T11:55
MSA tech blog coming soon. And M3 :) → tweet
@gdb · 2026-05-26T21:39
GPT-5.5 is a uniquely good coding model → tweet
@goospaceport · 2026-05-26T19:24
That comedown feelings called Post Banger Open Weights and there is no cure aside from more Banger Open Weights dropping unfortunately. Qwen 3.6 is REALLY GOOD still and a true model for the middle so you can get by until next banger weight drop, focus on outputs. → tweet
@sudoingX · 2026-05-26T19:48
will qwen 3.7 ship open weights or run away like openai did? cast your vote. and if they should, drop why in the replies. → tweet
@victormustar · 2026-05-26T19:56
cool new release: a tiny open video VLM that understands what happens in videos and when 👀 Marlin-2B (Apache 2.0!) can caption clips into timestamped events, or find a natural-language moment inside the video → tweet
@Ex0byt · 2026-05-27T00:59
Meet Cortex-Conv: - 34,106 → 43x smaller! - 96.8% on MNIST/Fashion MNIST - 32+8 iterations per sample → ~30% faster - loads in a browser tab in ~3s (WebGPU) - 720 KB weights ship with the page → tweet
@Ex0byt · 2026-05-27T14:06
Boom!.. want to see native EAGLE3 specdec with all open models and engines for faster local ai inference! → tweet
Codex & Agent Infrastructure
@thdxr · 2026-05-27T18:21
codex subscription endpoints have had some issues over the past week - you may have noticed some stalling in OpenCode. we put in some patches in v1.15.11 that should fix it → tweet
@Teknium · 2026-05-27T02:50
If you have been experiencing issues with OpenAI Codex oAuth, it is now fixed. OpenAI has been fixing it in the background and then seemed to change their whole spec today to solve it, you must
hermes updateto fix. → tweet
@nummanali · 2026-05-27T12:25
I'll raise you one Vaibhav. Permanent Codex Server Daemon for any device: codex remote-control start. FYI you will need to switch to managed Codex install for auto updating: curl -fsSL https://t.co/thfVANUgOM | sh → tweet
@jezell · 2026-05-27T06:21
Time to revisit the codex app server protocol with the semi-official SDK. Last time I implemented a codex app server client was right after it was released, and it was full of deadlocks. Must be better now right? → tweet
@gdb · 2026-05-27T15:47
Codex for transcribing and answering questions about a meeting in real time: → tweet
@gdb · 2026-05-27T06:16
codex is great for any kind of work done with a computer: → tweet
@badlogicgames · 2026-05-27T16:28
RT @thsottiaux: To simplify our Codex compute fleet management, we will be sunsetting GPT-5.2 and GPT-5.3-Codex in Codex on June 2nd when l… → tweet
@Teknium · 2026-05-26T20:54
Supergrok plan limits have been reset for Hermes Agent users! Go wild! → tweet
@Teknium · 2026-05-27T17:20
RT @NousResearch: Hermes Agent now has a built-in MCP Catalog → tweet
@Teknium · 2026-05-27T18:19
384 people in the Hermes Agent Jam Session in our discord, happening now! → tweet
Coding Agents Debate & Developer Tools
@sudoingX · 2026-05-27T07:17
saying this out loud again. codex and claude code are for vibe coding. cursor is for actual software development. i've used all three across real workloads. right now i'm spending most of my cursor credits REWORKING things codex and claude code shipped that didn't hold up. → tweet
@thdxr · 2026-05-27T12:10
imagine a benchmark of two editors — one opens a file 10x faster! wow it must be better. but oh wait they just disabled syntax highlighting. real products have to do things that make it worse on benchmarks but better in practice → tweet
@steipete · 2026-05-27T01:57
autoreview is the most impactful skill I've added to my stack. It automatically reviews your code before landing a PR. Finds so many edge cases. Sometimes it runs for hours. → tweet
@steipete · 2026-05-26T23:52
All the deps around opus are old or terrible, so vibed my own and replaced octoscript and opus-native. Performance of modern wasm on node/V8 is ~equivalent to native. → tweet
@steipete · 2026-05-26T23:56
Also extracted our image-logic into a separate library. Rastermill - Portable image processing for Node agents. Uses Wasm+Rust to be fast. → tweet
@sudoingX · 2026-05-27T11:31
if mythos is so good, why does claude code feel this broken? maybe the model is fine and the bloated harness around it is the actual problem. → tweet
@jsuarez · 2026-05-26T21:26
I don't care what your model benchmarks say. Codex xhigh just tried to add bubble sort to my high-perf C experience buffer → tweet
@thdxr · 2026-05-27T17:47
a lot of programmers work in the tech, a lot more of them work outside of tech. tech makes a lot of revenue per employee - if they had to pay 2x per programmer it probably wouldn't matter. very different for everyone else. worth thinking about in terms of coding agent spend → tweet
@nummanali · 2026-05-27T11:33
Lazy mans coach — Pin a chat in ChatGPT / Claude. "Be my personal coach, I want to lose weight and be my optimal self." I sent a voice note saying what I'd eaten for Eid and it logged it, then gave follow up advice → tweet
@nummanali · 2026-05-26T23:59
Imagine if all your digital interactions were through an intermediate layer of Generative UI powered by an ultrafast swarm of agents with a main orchestrator. If I had spare time, I'd build a device specially to interact with the heuristics of LLMs at an OS level → tweet
Open Source & Infrastructure
@badlogicgames · 2026-05-27T01:11
RT @vllm_project: 🦀 The Rust frontend is officially merged into vLLM! As GPUs get faster, the frontend has become a real share of CPU time… → tweet
@nummanali · 2026-05-27T18:21
🚨@LLMMart has tried to impersonate me on GitHub 🚨 They did unsigned commits using my email address. All OSS devs should enable Vigilant Mode on GitHub. It forces an unverified badge on unsigned commits, keeps you safe! → tweet
@jezell · 2026-05-27T14:37
RT @psviderski: Fly is starting to do public releases for Corrosion and shipped v1.0.0! 🚀 It's a gossip-based eventually consistent distri… → tweet
@jezell · 2026-05-27T05:43
RT @__morse: introducing holocron — it's an open source alternative to Mintlify, as a self hostable Vite plugin. it supports the same exact… → tweet
@ollama · 2026-05-27T00:13
RT @agupta: v0.13.0 of Exo just dropped, and it's one I've been looking forward to for a while. tl;dr swapping out anthropic for @ollama cl… → tweet
@TheAhmadOsman · 2026-05-27T19:47
Let me make Local AI easy for you. Give Codex Cli the article below & tell it: Infer the right Inference Engine from your hardware + article, use uv+venv, pick right kernels, tune flags, batching, KVCache, etc → tweet
@steipete · 2026-05-26T23:49
What do people use for SSO/SCIM/Endpoint Security in 2026. As we're hiring people for the OpenClaw Foundation, I gotta level up. → tweet
AI Industry & Strategy
@sama · 2026-05-27T16:44
AI should dramatically increase quality of life and individual freedoms for people around the world. The OpenAI Foundation is making an initial $250M commitment to measurement, transition support, and new approaches to broadly shared prosperity. → tweet
@jezell · 2026-05-27T17:48
RT @a16z: OpenAI and Anthropic are effectively telling the market they can't solve every problem with a generic AI coworker. You don't pou… → tweet
@jezell · 2026-05-27T17:37
I know people like to make fun of gstack, but honestly @garrytan and the @ycombinator team seem to get the reality of where agents need to go better than most people I've heard talking about the state of agents and where we are headed. → tweet
@jezell · 2026-05-27T16:42
RT @ycombinator: Over the past year, we've been building our own internal agent infrastructure at YC: over 350 tools, self-improving skill… → tweet
@TheAhmadOsman · 2026-05-27T04:26
The company that fumbled AI the most so far has been Microsoft. I remember how the entire internet was using them as an example of great cunning lol → tweet
@TheAhmadOsman · 2026-05-27T03:21
The amount of software that will need to be written for AI in high performance and specialized languages far exceeds your imagination and it WILL NOT be vibecoded → tweet
@TheAhmadOsman · 2026-05-27T01:54
Opensource AI MUST WIN — Ahmad, The OpenSource Man → tweet
@TheAhmadOsman · 2026-05-26T21:55
People keep asking me why do I focus on fundamentals instead of agents or shiny products. Shortcuts don't compound. Models are still improving, Agents come and go, Frameworks churn, Products age fast. Fundamentals stick. I'm not optimizing for the next launch, I'm optimizing for the next decade. → tweet
@LinusEkenstam · 2026-05-27T11:49
I feel that 2026/27 will be the year that e-commerce will be re-written from core principles. Personal shopping agents, hyper personalized assets, universal carts and virtual try-on are some of the things that will make the entire e-commerce experience 10x better. → tweet
Local AI & Consumer Hardware
@sudoingX · 2026-05-26T19:06
most of you have a gpu in a drawer or a secondary pc collecting dust right now that can run more intelligence locally than you think. i've been running current open weight agents on a 10 year old gtx 1080 8gb pascal card this week. 656k context on gemma 4 e4b. 248k context on qwen 3.5 9b. agents firing real tool calls at 18-20 tok/s. → tweet
@sudoingX · 2026-05-26T19:58
the chatgpt moment for small gpus is coming. and most of you are sleeping on it. a 5 month old open weight model running on a 10 year old gpu just executed tool calls cleanly, autonomously, end to end. no cloud. zero api. → tweet
@sudoingX · 2026-05-27T07:07
let me put it like this: the chatgpt moment for small gpus is coming. the day a single 24gb gpu runs a model that makes people forget they ever needed an api. we're closer than you think. → tweet
@alexocheema · 2026-05-26T20:47
What model do you like best on 128GB local inference devices e.g. MacBook Pro / DGX Spark / Strix Halo? → tweet
@sudoingX · 2026-05-27T16:55
RT: dgx spark with hermes agent /goal is the most underemployed combo in local ai right now. → tweet
Research & Benchmarks
@jezell · 2026-05-27T18:58
RT @lancedb: World model research is fragmented: every paper reimplements its own data pipeline, baselines, and eval harness. Comparing… → tweet
@jsuarez · 2026-05-26T22:20
Another massive fail. Cites PPO-v3 + DreamerV3 on percentile scaling for robust advantage scaling. Pretty nifty right? Except I'm the last author on PPO-v3 and the paper states that DreamerV3's scaling tricks generally do not work at all. → tweet
@jsuarez · 2026-05-27T18:50
Get all your papers accepted with this one closely guarded insider tip! Just fucking lie! → tweet
@badlogicgames · 2026-05-27T03:20
a model on its own is not concious/sentient under popular theories/frameworks of conciousness. not because its a big matmul machine. but because it lacks things like continuity/"state", self-maintenance, believe consolidation, embodiement, grounding, feedback loops, etc. → tweet
@steipete · 2026-05-27T12:04
RT @koltregaskes: Many developers have suspected for months that GPT-5.5 outperforms Claude Sonnet for coding. But SWE-Bench reported near-… → tweet
@badlogicgames · 2026-05-27T00:39
this is going to be super duper interesting! i wonder what sparse attention methods, if any, the closed big labs use. from the outside it looks like the open weights labs are innovating hard here. → tweet
Community & Events
@sudoingX · 2026-05-27T11:23
7,400+ actual builders in one place. not lurkers. the hermes agent community grew faster than any agent community on x by orders of magnitude. all organic. zero paid promotion. → tweet
@levelsio · 2026-05-27T16:20
🏆 Round 1 of judging the Vibe Jam of 2026 sponsored @cursor_ai + @boltdotnew + @heyglif + @tripoai is finished now. We've gone through almost 1000 submissions. → tweet
@levelsio · 2026-05-27T17:21
Just saw a guy at Munich airport code completely without AI like some kind of maniac 🤯 → tweet
@sudoingX · 2026-05-27T16:55
anthropic's marketing operation is genuinely impressive. amazing salesmen using every tactic in the book to convince you that their calculator went through grief, sadness, fear, humility, and even blackmail before producing 2+2=4. → tweet
Hardware & Semiconductors
@TrungTPhan · 2026-05-27T16:10
Micron and SK Hynix shareholders after both joined the $1T club within a 24-hour span → tweet
@TrungTPhan · 2026-05-27T14:14
J.R. Simplot — Born in 1909, made billions with Idaho potatoes then helped build the now $1T Micron Technology in 1980. Lived almost a century going from potato chips to semiconductor chips. → tweet