← Tech / AI / IT Monitor Index Tech / AI Generated 2026-09-23 19:13 UTC

Tech / AI / IT Monitor

September 23, 2026 · Based on tweets from the last 24 hours · 194 tweets analyzed · model: ollama-cloud/glm-5.2:cloud

Executive Summary

The AI ecosystem saw significant model releases over the last 24 hours, headlined by Anthropic's Opus 5.5 and OpenAI's GPT-6 Sol and Luna, both demonstrating massive leaps in coding, reasoning, and cost efficiency for agent workloads. Open-source and local AI continued to gain momentum, highlighted by the release of FLUX 3 Action (a 7B world-action model), Unsloth surpassing 500M downloads on Hugging Face, and new highly-optimized quantizations enabling complex agentic workflows on consumer 12GB GPUs. Developer tooling also evolved rapidly, with coding agents like Amp, Buzz, and local stacks (e.g., Hermes, llama.cpp) proving capable of building enterprise-grade intranets and fixing decade-old bugs autonomously.

Key Events

Analysis

The current trend highlights a rapid bifurcation in AI usage: ultra-fast, high-reasoning models for interactive coding (Opus 5.5) and highly efficient, "slow" models for autonomous background work (GPT-6 Luna). Furthermore, open-source local models (like Bonsai 2 and Unsloth's quants) are proving capable of complex multi-file engineering on standard consumer hardware (e.g., 12GB RTX 3060). Another major pattern is the focus on inference economics—Anthropic's aggressive cache-read pricing and Apple's token-saving document retrieval finetune show that the battlefront is shifting from raw intelligence to efficient context and agent management.

Watch next for further integration of "Jev" style decision models into coding harnesses, as well as hardware-level optimizations for Apple Silicon (M5 Ultra) and clustered consumer setups to run increasingly dense models.

Tweet Feed

AI Model Releases & Benchmarks

@TheAhmadOsman · 2026-09-22T19:30

Pacing the frontier but we got a few more releases for you guys please meet Opus 5.5 → tweet

@jezell · 2026-09-22T19:24

RT @thsottiaux: GPT-6 Sol and Luna are out. Not only are they a very significant improvement across the board, but also in writing and gene… → tweet

@victormustar · 2026-09-23T17:52

RT @bfl_ai: Introducing FLUX 3 Action. An open weights 7B World Action Model that achieves first place on the RoboLab benchmark. It outpe… → tweet

@victormustar · 2026-09-23T18:15

Alert: Apple just dropped a new model on Hugging Face. It's a Qwen3.5-9B finetune that turns long documents into small page images to save tokens, then pulls up the full text of only the pages relevant to your question 💡 https://t.co/TAfCmVlXJn → tweet

@victormustar · 2026-09-22T20:11

RT @AdinaYakup: Ant Group @TheInclusionAI just released a new image model 🔥 Ming Image-0.1 Design - 6B + MIT licensed - Focused on text… → tweet

@victormustar · 2026-09-23T10:31

Opus 5.5 made this galloping horse (entirely in code every pixel drawn procedurally). One self-contained HTML file. Vanilla JS + Canvas 2D. No images or libraries. 128×96 pixels, articulated legs driven by inverse kinematics, 12-pose gallop... Something is happening... → tweet

@kunchenguid · 2026-09-22T23:15

ok i've used opus 5.5 enough now to have an informed opinion and i just want to say - hallelujah!!! opus 5.5 solved SO MANY problems. let me list them below in increasing importance 1. verbosity is now used more appropriately ... → tweet

@kunchenguid · 2026-09-23T15:51

just took gpt 6 luna for a spin, and… it’s a very weird model 1. it’s insanely cheap, even with the point below considered 2. it’s extremely slow, not in terms of time to first token or toks/sec, but how many turns it takes to get something done... → tweet

@swyx · 2026-09-23T06:43

can confirm. ran @latentspacepod AINews side by side with 6 Sol and the difference was night and day... 5.5 Opus is the new default model for AINews going forward. so much more concise and tasteful reporting, with much less slopese than even 5 Opus. → tweet

Open Source & Local AI

@victormustar · 2026-09-23T16:33

RT @UnslothAI: Unsloth has surpassed 500M model downloads on Hugging Face! 🦥🤗 Qwen3.8-27B GGUF is already Unsloth’s #1 most-downloaded mod… → tweet

@sudoingX · 2026-09-23T18:35

this is what 12gb of vram builds in 2026, absolute magic

rtx 3060 12gb, #1 gpu on steam bonsai 2 27b + mtp, 5.95 gb of weights hermes agent, 5 hours, 328k tokens written... this entire game was built by bonsai2, a qwen 3.8 27b dense compressed to ternary... → tweet

@sudoingX · 2026-09-23T00:01

my local ai stack is 5 pieces and every one of them is non negotiable. if you are starting on your own gpu, this is the shortest path i know.

llama.cpp for one card, vllm for more than one... hermes agent as the harness... browser-use with playwright when you want it on the web... tmux for everything... tailscale and termius for reach. → tweet

@TheAhmadOsman · 2026-09-23T04:26

If your grandma cannot run Local AI on her laptop from 2013 we'll have failed our mission That's the bar we're setting for ODS We're gonna make Local AI the default... → tweet

@victormustar · 2026-09-23T15:07

RT @NVIDIAAI: When several people talk at once, a transcript can get messy fast. Our new Nemotron 3 Diarization model tracks who spoke whe… → tweet

@Prince_Canuma · 2026-09-23T16:46

That was blazing fast!🔥 Nemotron 3 Diarization is now on MLX-Audio → tweet

Developer Tools & Coding Agents

@sqs · 2026-09-22T23:38

I know people who spent $12k in tokens and built nothing. @Westpac spent $12k in tokens on @AmpCode and built a new intranet for their 35,000 employees (vs. ~$3–5M pre-AI cost). "Amp is unlocking a huge amount of value in engineering" from https://t.co/udHaDS16UB → tweet

@thdxr · 2026-09-23T01:14

button on flight check-in page wasn't working so i told opencode to download the source and find the bug it figured out there was an invisible form field that was required and it asked me for the values and submitted AI finally fixed all the shitty forms on the internet → tweet

@jack · 2026-09-23T17:40

RT @blocks: An open source team moved its day-to-day development into Buzz. Across matched workweeks, MeshLLM merged 56% more PRs and cut m… → tweet

@jack · 2026-09-22T19:03

RT @John_Ely_21m: I’ve been building software for 25 years and my Buzz agents are my favorite dev team I’ve ever had the pleasure of workin… → tweet

@steipete · 2026-09-23T08:13

CodexBar learned a few new tricks in the last few weeks! TypeSafe, Nous Portal, Muse Code, CodeRabbit, Replicate, HuggingFace, Pi, Helmcode, v0, Charm Hyper, GitKraken AI, Bifrost My ambitions for this app REALLY outgrew the app name. → tweet

@sqs · 2026-09-23T18:53

Some other popular agents cost ~25-60% more than Amp, based on est prices w/their different compaction thresholds. Interesting and obv incomplete comparison (thus names withheld). Amp's compaction exploits persistent thread storage... → tweet

@MatejKnopp · 2026-09-23T16:20

Gemini feedback on a PR: "If DestroyWindow(hwnd) fails because the window has been created on another thread this function will get stuck in a infinte loop. You should copy the handles write the loop like this so it doesn't happen." My dear clanker, if we somehow end up with HWNDs created on different thread we have way bigger problem... → tweet

@jezell · 2026-09-23T16:05

Honestly, what @rakyll's team at Google is working on is the most significant contribution Google is making to the agent space. Gemini still lags, A2A is useless, but OpenAI and Anthropic haven't pushed a single line of useful infra code to github and AX is pretty neat. → tweet

@steipete · 2026-09-22T20:53

Had ChatGPT sometimes crashing on me after updating to macOS 27 and... Astra found a ~14 year old bug in libuv. → tweet

Hardware & Performance

@TheAhmadOsman · 2026-09-22T22:02

Getting these 8x RTX PRO 6000 bad boys ready for MiMo-V2.6-Pro-RL → tweet

@NaderLikeLadder · 2026-09-23T16:09

RT @NVIDIARTXSpark: We got @UnslothAI a DGX Station! @DanielHanChen and @NaderLikeLadder checked out Unsloth’s new @Dell Pro Max with GB30… → tweet

@TheAhmadOsman · 2026-09-23T02:18

You run Kernels not models The model is just a graph The Inference Engine is scheduler / optimizer / executor But the actual work? That happens in the Kernels... This is why Inference Engines and the Kernels implemented within them matter → tweet

@ivanfioravanti · 2026-09-23T08:54

400W under heavy load on M5 Ultra is not trivial. It's doubled compared to M3 Ultra. This is the review to watch 👇 → tweet

@FrameworkPuter · 2026-09-23T18:41

This may be the strongest model currently for the 192GB Framework Desktop, but can also run today with either SSD streaming on a single 128GB or clustering together two 128GB machines. → tweet

AI Research & Tooling

@LinusEkenstam · 2026-09-23T07:46

I love seeing how fast the frontier labs are jumping in on creating forks. Just today we’ve seen omni-jev and now djev from Google Gemma team. It will be clear in a few weeks, just how powerful JEV is going to be inside harnesses and applications. → tweet

@steipete · 2026-09-23T17:40

RT @CAIS: We are releasing HLE-Diamond, a refined subset of Humanity’s Last Exam (HLE), following a year-long process of cleaning and refin… → tweet

@victormustar · 2026-09-23T10:09

RT @MhYin76491: Introducing GAE — Geometry-Native Autoencoder. The key choice is where generation happens. GAE lets video models generate d… → tweet

@victormustar · 2026-09-23T12:17

RT @mishig25: New JS package: @huggingface/lerobot 🤖 Read LeRobot datasets on the Hub straight from the browser. No download. → tweet

@victormustar · 2026-09-23T13:00

RT @vanstriendaniel: Trained a Jev-style classifier on @huggingface Jobs for ~$1.50. It's a 194M GLiNER2 model that suggests task tags for… → tweet

@kunchenguid · 2026-09-23T04:27

since it’s been a good day for anthropic with a strong opus 5.5 release, i’m going to highlight one more thing that may not be obvious across xai, openai and anthropic - 1. anthropic is the only provider that does not charge 2x for long context requests 2. anthropic’s latest models have ridiculously low pricing for cached read... → tweet

@uwteam · 2026-09-23T05:14

Wczoraj w nocy ogłoszono nowego buga w silniku Wordpressa. Ponownie jest to RCE, czyli atakujący może wykonać dowolny kod na serwerze, a co za tym idzie przejąć całą stronę i wszystkie dane. Tradycyjnie: Zaktualizujcie Wordpressy 😉 → tweet