← Tech / AI / IT Monitor Index Tech / AI Generated 2026-09-10 19:13 UTC

Tech / AI / IT Monitor

September 10, 2026 · Based on tweets from the last 24 hours · 217 tweets analyzed · model: ollama-cloud/glm-5.2:cloud

Executive Summary

The last 24 hours were dominated by the release of DeepSeek V4.1 Flash, an open-weight 552B MoE model utilizing a novel split-brain architecture that disrupts the current price-to-performance frontier, matching or beating proprietary models on agentic benchmarks. OpenAI simultaneously launched GPT-Live-1 for natural voice agents in the API and expanded ChatGPT Voice usage limits, while the developer community continued to test "Astra" (GPT-6) for complex 3D generation and coding workflows. On the hardware and local execution front, developers are aggressively optimizing consumer GPUs with speculative decoding and quantized KV caches, while early details emerge about NVIDIA's "RTX Spark" hardware and its built-in "OpenShell" agentic OS runtime. Finally, OpenAI welcomed AI safety pioneer Paul Christiano to its team amidst escalating friction between the open-source community and proprietary labs over open-weight regulations.

Key Events

Analysis

Tweet Feed

AI Model Releases & Performance

@Ex0byt · 2026-09-10T13:03

DeepSeek V4.1-Flash is out and overclocks V4-Pro on the agentic/coding/cyber stack. Most importantly, speed and cost thanks to RL and a new split-brain design architecture with native vision. → tweet link

@sudoingX · 2026-09-10T07:07

deepseek keeps punching above its weight man, this is insane. the new v4.1 flash is a 552B MoE with a new asymmetric trick, 8B active parameters on input, 16B on output, and look what that buys in the table. it beats their own v4 PRO on almost every agentic bench... → tweet link

@thdxr · 2026-09-10T06:34

it's crazy that a minor update to a flash model is surpassing the previous pro model. deepseek should probably double down on their flash models. it fills a crazy spot in the price/performance spectrum → tweet link

@TheAhmadOsman · 2026-09-10T13:07

DeepSeek V4.1 Flash being this good makes Anthropic’s $2T IPO sound crazy stupid. That’s why they hate Opensource AI and wanna ban it → tweet link

@victormustar · 2026-09-10T16:04

I can confirm it now, DeepSeek V4.1 Flash is amazing 🥹 → tweet link

@alexinexxx · 2026-09-10T14:55

RT @eliebakouch: Deepseek V4.1 Flash 552B total, 8/16B active with a new arch trained on 45T tokens, there are different active parameters… → tweet link

@RayFernando1337 · 2026-09-10T16:16

This Cognition drop is bigger than most people realize due to the attention to detail they spent post-training Kimi K3. Only a handful of labs are able to curate enough engineers with taste to double down on the details that matter for this work. → tweet link

@ivanfioravanti · 2026-09-10T06:28

Unbelievable intelligence compressed in a Flash model! Probably going big big big is not the right answer. Smaller, less expensive and more capable models is the real way. → tweet link

Developer Tools & Local Inference

@sudoingX · 2026-09-10T01:00

psa for every llama.cpp user: update your build. if your llama.cpp is older than b10354, you are missing the MTP speculative decode path for qwen 3.8 27b dense. that is somewhere between 33 and 145% decode speed measured across 68 community rigs... → tweet link

@sudoingX · 2026-09-10T00:01

most of the people telling you never quantize kv cache have never owned a 24gb card, and it shows. on an h100 cluster, sure, run f16 kv forever. but local ai does not live there. local ai lives on a 3090 with 24gb... → tweet link

@sqs · 2026-09-10T08:10

Amp portals just got 70% faster (on p95 subresource load times) At this rate, it's almost like soon you could be using them for more than just dev... → tweet link

@sqs · 2026-09-10T04:45

Context analysis, now on Amp thread usage/cost pages. If you have 377 MCP tools and AGENTS.md files junking up your context, now you can ask Amp and it'll tell you that (it can see this info too). → tweet link

@Teknium · 2026-09-10T08:01

You can now install plugins in Hermes Agent from private repos! It'll use your stored GitHub credentials now. → tweet link

@Teknium · 2026-09-10T15:38

Big upgrade for subagent observability and management in Hermes Agent :) → tweet link

@thdxr · 2026-09-10T01:06

RT @jlongster: we're making worktrees better in opencode. plugins can now register different strategies for managing them, so you can use t… → tweet link

@jsuarez · 2026-09-10T15:40

RT @VincentMoens: Just released torchrl 0.14. We had a lot of fun integrating Microduck in it! Other than that, expect DreamerV3 faster th… → tweet link

@ivanfioravanti · 2026-09-10T14:31

Qwen 3.8 Flash Next on DwarfStar reached 224K of context on an M4 Max 64GB! Let's go 🚀 Gonna test and merge now PR #10 from @p0ly that gives a +30% boost to prefill on Metal 4 kernel (M5+ chips) → tweet link

OpenAI / GPT-Live-1 / Astra

@victormustar · 2026-09-10T18:57

RT @OpenAIDevs: GPT-Live-1 is now available in the API. Bring ChatGPT’s natural back-and-forth to your app, with voice agents that listen… → tweet link

@jxnlco · 2026-09-10T17:14

RT @ChatGPT: Now everyone can put data to work. We’re introducing a new Data agent in ChatGPT Work so you can turn your company’s data int… → tweet link

@jezell · 2026-09-09T20:22

RT @juberti: Just increased the ChatGPT Voice usage limits: Go: 3h Live mini, Plus: 3h Live, Pro 100: 15h Live, Pro 200: Unlimited → tweet link

@kunchenguid · 2026-09-09T20:01

i had to really push astra on this, but it did end up coding this nice transition from sunset -> nightfall -> moonrise -> milky way. there's not a single image in here. it's all just JS code. i'm mildly impressed → tweet link

@LinusEkenstam · 2026-09-09T19:47

RT @LinusEkenstam: Astra was able to recreate my Studio in Blender from 5 photos + 3 panos, I shared the Blender files 3 days ago, it too… → tweet link

@badlogicgames · 2026-09-10T16:26

astra writes absolutely terrible C++ now. → tweet link

@jezell · 2026-09-10T14:43

Codex sparkles now and also queues follow up questions in the latest build (kind of like steering messages in reverse, for the human). Interesting. → tweet link

Hardware & Agentic OS

@sudoingX · 2026-09-10T18:45

so after posting this i went and read nvidia's own launch materials, and it turns out they already built the agentic OS layer. it is called OpenShell, a runtime for secure agent execution shipping with rtx spark, built with microsoft into windows itself. → tweet link

@sudoingX · 2026-09-10T18:20

dear nvidia, you are about to launch rtx spark, the most exciting consumer hardware in years. a grace cpu and a blackwell gpu on one package sharing up to 128GB of unified memory, a petaflop in a laptop... ship first class linux support, bless distros like omarchy the way dell just did on xps, and rtx spark becomes the default machine of the agentic os era → tweet link

@ivanfioravanti · 2026-09-10T08:24

M5 Ultra 512GB is a must have for me now 😎 → tweet link

@TheAhmadOsman · 2026-09-10T08:08

This new DeepSeek V4.1 Flash model will make GPUs prices go even higher … way higher. $40k RTX PRO 6000 incoming. Bookmark this tweet → tweet link

AI Strategy, Safety & Industry Dynamics

@sama · 2026-09-09T19:57

Welcome, Paul. Grateful you are doing this, and all you have done for AI safety. Excited to work together again. → tweet link

@gdb · 2026-09-09T20:50

we used our models to find and fix critical vulnerabilities in our own systems, as part of a 250+ person effort within openai. playbooks/learnings/architecture which we hope can be helpful for others to make the most of the defenders' window: → tweet link

@sudoingX · 2026-09-10T07:32

BREAKING: Anthropic CEO Dario Amodei is concerned after the new DeepSeek V4.1 Flash beat Opus 5 on agentic benchmarks while being free for anyone to download, and is said to be calling an emergency meeting with the Pentagon to discuss whether open weights this capable should be considered as national security risk. → tweet link

@TheAhmadOsman · 2026-09-10T00:31

Any calls for pausing AI for safety reasons is a PSYOP. OpenAI and Anthropic can withstand not putting out models while all the smaller / opensource labs lose their momentum and shut down. Then they will lobby to resume. This is only happening because Opensource AI is catching up → tweet link

@kunchenguid · 2026-09-10T08:50

you know why apple’s so behind on AI? at a time when everybody and their mom could build an agent, how could apple basically still have nothing to show? well.. that is because they locked themselves into this privacy narrative... meanwhile, meta has invested in research, bought compute, acquired talent, created a reasonable model, and shipped AI products that are looking more and more promising every day → tweet link

@levelsio · 2026-09-10T17:07

The AI chat apps will slowly eat up most services and provide them to users directly, many times without even an app or interface, just do whatever the user wants. Dario Anthropic said it himself "in the end there will be just Anthropic and world governments"... → tweet link