Executive Summary
This reporting period saw a significant AI development cycle with OpenAI's release of GPT-Realtime-2 to the API, featuring reasoning capabilities and voice-to-voice translation. The local AI ecosystem continues maturing with Nous Research's Hermes Agent framework gaining substantial community adoption and new features including cronjob automation and browser integration. Developer tooling advances include Node.js 26.1.0 with native FFI support and benchmark comparisons between Qwen 3.6 27B and fine-tuned variants demonstrating competitive performance on consumer hardware. The industry continues debating compute as the primary competitive moat, with training automation potentially displacing traditional "learn to code" pathways.
Key Events
-
OpenAI releases GPT-Realtime-2 to API with reasoning capabilities — Sam Altman announces voice interaction momentum; model includes reasoning and translation features → tweet
-
Nous Research releases Hermes Agent v0.11.0+ with multiple updates — New features include profile creation with
--no-skillsflag for blank slate agents, cronjob support for programmatic tasks without agent overhead, and Lightpanda browser backend integration → tweet -
Node.js 26.1.0 ships with native FFI module — Major addition enabling foreign function interface calls; Platformatic team sponsored development → tweet
-
Qwen 3.6 27B benchmarks on consumer hardware show competitive agentic performance — Testing on RTX 5090 24GB demonstrates vanilla model completing tasks with fewer tool calls than fine-tuned variants while generating more reasoning content per message → tweet
-
pi package manager completes namespace migration — Version 0.74.0 moves from @mariozechner to @earendil-works; users must update imports → tweet
-
SubQ model introduced as sub-quadratic sparse LLM breakthrough — First model built on fully sub-quadratic sparse architecture → tweet
Analysis
AI Development Patterns: Voice interaction with AI is accelerating, with OpenAI noting users increasingly prefer voice for context-heavy tasks. The Hermes Agent ecosystem shows healthy open-source momentum with third-party skills (ComfyUI video grid, cron automation) and growing adoption. The trend toward agent customization—blank slate profiles, skill registries—suggests developers are moving beyond monolithic agent solutions.
Benchmarking Maturation: Consumer hardware benchmarks are becoming more rigorous. The shift from single-run happy-path tests to multi-run variance capture, error injection, and adversarial scenarios indicates the field is demanding higher standards for agent evaluation. The observation that vanilla models can outperform fine-tuned variants on tool-use tasks challenges assumptions about specialized training.
Developer Tooling Renaissance: Node.js FFI addition and pi package namespace changes signal ecosystem maintenance and evolution. The continued emphasis on constraint-based development (ICQ vs. Teams efficiency debate) suggests developers are increasingly aware of bloat concerns.
What to Watch: xAI's compute capacity utilization by competitors (Anthropic reportedly renting capacity) raises questions about competitive moats. Training automation predictions indicate significant industry disruption. Hermes Agent's new minimal TUI mode and proper tracing implementation may set new standards for coding agent UX.
Tweet Feed
AI Model Releases & Research
@sama · 2026-05-07T18:55
people are really starting to use voice to interact with AI, especially when they have a lot of context to dump. GPT-Realtime-2 comes to the API today; it is a pretty big step forward. (we are working on improvements to voice in chat.) → tweet
@gdb · 2026-05-07T18:01
You can now just build amazing voice agents, with the GPT-Realtime-2 reasoning model in our API: → tweet
@louszbd · 2026-05-07T15:54
RT @vincentweisser: We are releasing Lab RL just works across almost any verifiable domain We want to enable everyone to train their own a… → tweet
@louszbd · 2026-05-07T06:54
RT @alex_whedon: Introducing SubQ - a major breakthrough in LLM intelligence. It is the first model built on a fully sub-quadratic sparse-… → tweet
Hermes Agent & Local AI Development
@sudoingX · 2026-05-07T13:50
here are the results after testing carnice v2 vs qwen 3.6 27b dense base on my new agentic tool use task, same setup top to bottom. ... both succeeded: closed plan-execute-verify loop wrote self_report.md then read it back before claiming done zero hallucinated tools both caught the prompt-spec drift on send_final_answer → tweet
@sudoingX · 2026-05-07T12:15
so yesterday i ran carnice-v2 27b on 5090 and a self-report card task and called it the rare-bird loop closer, 19 tool calls clean. today vanilla qwen 3.6 ran the exact same task on the same hardware and i saw fewer tool calls. more reasoning per call. closed the same verify loop. → tweet
@sudoingX · 2026-05-07T16:09
watch this video of qwen 3.6 27b dense vanilla running my new autonomous tool use benchmark via @NousResearch hermes agent on my 5090 24gb laptop gpu at 16 tok/sec average. ... the part that surprised me lives in the hermes agent reasoning field, not visible to most viewers but pulled from the session trace afterward. mid-run the model literally wrote: "Wait, I need to use the send_final_answer tool... but looking at my available tools, I don't see send_final_answer listed." → tweet
@sudoingX · 2026-05-07T16:24
it's so easy to get started in local ai actually. the only real wall is vram math. practical heuristic for a single gpu:
24gb = 27B Q4_K_M at 262k context (qwen 3.6, carnice-v2) 16gb = 13B Q5_K_M at 32k or 9B Q8_0 at 64k 12gb = 8B Q5_K_M at 16k 8gb = 4B Q4_K_M at 8k → tweet
@sudoingX · 2026-05-07T10:17
anon, if you're new to local ai or agentic workflows, learn these three tools before anything else.
tmux - persistent sessions that survive disconnects. your agents keep running whether you're watching or not. termius - ssh from your phone. full terminal access from anywhere. tailscale - mesh your machines. access any device from any device. → tweet
@sudoingX · 2026-05-07T06:27
patrick just asked the question that exposes how underserved the local ai vision + design + browser-use lane actually is right now. cms migration with theme replication, web design context preserved, all chained on consumer hardware. → tweet
@sudoingX · 2026-05-07T14:47
any designers out there? i'm working on something and need a logo. not a paid ask but if i pick yours for the launch i'll RT + name you publicly + amplify. → tweet
@sudoingX · 2026-05-07T08:54
people asking me about the framework desktop with amd 395+ ai max 128gb a lot lately, honest answer is i haven't tested it so i can't give you the receipts you'd actually use. current data point on the workstation tier from what i've actually run: dgx spark is the beast right now, 128gb unified, GB10 silicon, sustained inference under load. → tweet
@sudoingX · 2026-05-07T14:58
every ai lab in the world is compute constrained right now. xai is the only one with so much compute they can sell capacity to anthropic. xai bros, what's happening? is this a good sign you've won the compute war, or a bad sign your moat is getting rented to competitors? → tweet
@Teknium · 2026-05-07T11:35
You can now use
hermes profile create <name> --no-skillsto create a new agent with no built in skills whatsoever, start with a blank slate, fresh canvas! → tweet
@Teknium · 2026-05-07T02:52
Traditional cron jobs are great for silent tasks on a machine, and Hermes Agent cronjobs are great for extending that to your agent, but why not utilize the gateway and hermes' cron to access things that don't need to cost an agent's time across any messenger service you have connected? → tweet
@Teknium · 2026-05-07T02:20
RT @UncleHODL: @Liftaris1 @NousResearch @Teknium @imbabybrooklyn It even works on a tiny screen AND through Termux on mobile. The wannabe h… → tweet
@Teknium · 2026-05-07T14:20
RT @lightpanda_io: Lightpanda is now a browser backend in Hermes by @NousResearch. Open source autonomous agent. Open source browser bui… → tweet
@Teknium · 2026-05-07T03:15
RT @solana_devs: Appreciate everyone that came out for the Hermes Agent workshop with @spacemandev → tweet
Developer Tools & Frameworks
@badlogicgames · 2026-05-07T17:11
RT @nodejs: Node.js 26.1.0 is out, with a new
node:ffimodule,crypto.randomUUIDv7(), and many more features and bug fixes. → tweet
@thdxr · 2026-05-07T17:01
we sponsored the ffi work with @matteocollina and the team from @platformatic exciting to have the resources to push on these things - we did it for OpenTUI but will usher in an era of high performance libraries for ecosystem → tweet
@thdxr · 2026-05-07T22:54
preview of our minimal mode that doesn't run as a fullscreen TUI we're designing this carefully, it never rewrites your scrollback which means there's some tradeoffs but it'll be the only coding agent that doesn't have flickering or weird layout shifts → tweet
@badlogicgames · 2026-05-07T15:31
People of https://t.co/TgG5bkXUdV. If you've installed pi 0.73.1, you will now be notified to do a
pi updatefor 0.74.0. Starting today, all pi packages on NPM are in the earendil-works namespace, instead of mariozechner. → tweet
@badlogicgames · 2026-05-07T14:50
People of https://t.co/TgG5bkXUdV. Update to the latest pi now (0.73.1), which is the last version coming from the
@mariozechnernamespace on NPM. Next release in 10 minutes (0.74.0) will be in the@earendil-worksnamespace on NPM. → tweet
@thdxr · 2026-05-07T00:12
we're taking most of opencode's logic and breaking it down into internal plugins to force our plugin api to get better in that process i added proper tracing - so now i can what plugins are hooking and track down any perf issues → tweet
@gdb · 2026-05-06T22:41
codex is for everyone → tweet
@steipete · 2026-05-06T21:52
closed source, open source, nothing can stop codex. → tweet
@jezell · 2026-05-06T21:35
Looks like the Responses API returns zeros for usage data when context_management + threshold are sent @stevendcoffey. Seems like a bug? → tweet
@MatejKnopp · 2026-05-07T17:31
I've been writing objective c for 20 something years - since Tiger. TIL about binary form of ternary operator. NSScreen *screen = window.screen ?: [NSScreen mainScreen]; What. → tweet
Software Engineering Practices & Observations
@hnasr · 2026-05-07T14:03
I believe one reason for the bloat in modern applications is that we build them on powerful computers, which hide coding inefficiencies. ... I remember running ICQ on an Intel 90 MHz single-core machine with 64 MB of RAM on Windows 95. It launched instantly and the chat functionality worked seamlessly. In contrast, Teams can take several seconds, sometimes even minutes, to start up on my 64 GB, Intel 3.0 GHz, 16-core machine. → tweet
@hnasr · 2026-05-07T13:57
We tend to take what others built for gospel and Follow it blindly. Some even defend it and form tribes. We rarely question it because we don't understand the reason behind it and the problem it is solving. → tweet
@badlogicgames · 2026-05-07T18:57
it is absolutely crazy to me that our entire industry has succumbed to hyper waterfall. because that's what ya'll are doing with your massive plans and beads and dark factories. have you learned nothing? → tweet
@badlogicgames · 2026-05-07T18:54
the absolute worst part about the state of our industry is that excellent engineers are now entirely incapable of formulizing useful bug reports. we are at the level of "my computer doesn't turn on". → tweet
@badlogicgames · 2026-05-07T17:17
recommended reading, as always with entire stuff. lots of low hanging fruit for agentic search if people started to remember information retrieval from ca. 2004. → tweet
@jezell · 2026-05-07T17:17
I used to say an agent is just a for loop. That may still be true for demos, but it's kind of crazy how much work a good agent harness is compared to a for loop. Once you start adding steering, nice status messages, support for shell, computer use, a TUI, a GUI, compaction, lossless persistence, a protocol for surfacing all this to clients in many languages, etc. → tweet
Industry Trends & Commentary
@TheAhmadOsman · 2026-05-07T15:24
Training models is gonna become the new learn to code and will be automated to a high degree Novel research is another thing and has already been automated to some extent Synthetic data might be the only thing yet to crack, matter of time Compute is the ONLY TRUE moat → tweet
@TheAhmadOsman · 2026-05-06T21:03
What most don't get is that Compute is the only real moat → tweet
@TheAhmadOsman · 2026-05-06T21:52
I should add that I wouldn't wanna be compute poor comes 2029 Bookmark this → tweet
@thdxr · 2026-05-07T05:14
i pay attention to: 99%: using our product in dumb/simple ways and will never change behavior 0.01%: aliens who are showing us the distant future ignore the reminder: "pro" users who think they invented some clever workflow every week but get less done than the 99% group → tweet
@thdxr · 2026-05-07T15:28
here's how the whole "taste" trend is gonna go everyone is gonna talk about taste and then apply it by being opinionated about their products "i craft the perfect experience" and it's going to narrow the products audience to the point of not being sustainable → tweet
@thdxr · 2026-05-07T14:44
the way everyone does branding/marketing is trying to imagine how they want people to see them but the task is figuring out who you actually are because that's the only thing you can actually pull off → tweet
@thdxr · 2026-05-07T03:38
i never make plans i hate looking at markdown i don't wanna read markdown files i just plan by having it make changes to the code then i look at the code to see what sucks then i prompt again → tweet
@kunchenguid · 2026-05-07T15:16
what are the most important human skills now? i think at least one of them is - "learning how to use tokens to buy time" people who have 2400 hours of time per day will outperform people who have 8 → tweet
@victormustar · 2026-05-07T08:46
RT @nathanhabib1011: The SWE-bench Verified leaderboard on @huggingface now compares almost 50 models... Community benchmarking > closed b… → tweet
Flutter & Mobile Development
@ASalvadorini · 2026-05-07T17:46
addPostFrameCallback is totally legit to do something after your page has built. Now, if your build() only builds, you'll build first the initial state s0 (empty page pattern), and only after you tell the notifier to fetch the data. That's how you handle errors too
flutter
@ASalvadorini · 2026-05-07T12:24
At @ElisaOyj Viihde we process 177 EPG channels, 24 hours programs, shorter last minutes, over 3 weeks. Data is lazy loaded on demand. We just introduced paginated navigation with sections, and it feels faster than ever 🔥
flutter #flutterdev #flutterweb
@ASalvadorini · 2026-05-07T06:40
RT @AbdallahSh07: We know the community has been asking for Flutter/Dart skills! So excited they are out now! → tweet
@ASalvadorini · 2026-05-07T05:02
Despite what people say, I'm also using Antigravity + Gemini in my free time and weekends, and so far I'm quite happy with it ☺️
SI #Gemini #antigravity #Flutter #Flutterdev
Hardware & Systems
@levelsio · 2026-05-07T18:43
SQLite supports databases up to 281 terabytes in size SQLite is a highly optimized piece of software and can easily write 500,000 rows per second with proper batching → tweet
@levelsio · 2026-05-07T18:30
SQLite is free → tweet
@alexinexxx · 2026-05-07T03:16
visited @FrameworkPuter office today with @0xSero i'm assembling a team → tweet
@FrameworkPuter · 2026-05-07T01:29
RT @dcapitella: The new Framework Laptop 13 Pro is a thing of beauty! Thank you @FrameworkPuter for sending over a pre-release engineering… → tweet
@thdxr · 2026-05-07T15:33
we are adding 3000 new subscribers per day → tweet
AI Integration & Design
@MengTo · 2026-05-07T04:46
I did a 50-min podcast with @gregisenberg on how I use Google's DESIGN.md + custom skills to give AI a real design system → tweet
@levelsio · 2026-05-06T23:15
My two favorite AI companies working together! → tweet
Developer Community
@badlogicgames · 2026-05-07T16:29
RT @matteocollina: NodeConf EU is back in Bologna, Italy and it's going to be special 🇮🇹 Great talks, better people, and the kind of hallw… → tweet
@sudoingX · 2026-05-07T12:37
flagging this for the hermes agent team. i hit the same issue on mobile termius after a few prompts on small screens, ui disappears or gets stuck. people who run agents on the go would get real value from a fix here. → tweet
@sudoingX · 2026-05-07T06:39
when you tmux ls and you not see 5 sessions running agentic work you're ngmi → tweet
@sudoingX · 2026-05-07T04:24
just woke up and as always i must decide. if i build what i'm building for all of us, i don't get to post and x won't pay. if i post and make content, i don't get to build. every day this trade. fuck. → tweet
Report generated from Twitter/X monitoring. All times UTC. Includes 170 tweets filtered to tech/AI/software/IT content.