Executive Summary
Andrej Karpathy's announcement that he has joined Anthropic as an individual contributor dominated the day's tech discourse, seen as a major talent win for the company amid intensifying frontier LLM competition. Google I/O took center stage with the unveiling of Gemini Omni (multimodal generation from any input), Gemini 3.5 Flash powering dynamic UI generation in Google Search, and the Antigravity CLI/SDK launch. In the open-source and local AI sphere, llama.cpp's new Multi-Token Prediction (MTP) support for Qwen 3.6 delivered up to 78% throughput gains, reinforcing the viability of local-first model deployment. Meanwhile, Apple's $599 MacBook Neo—built on repurposed slightly-defective A18 Pro chips—illustrated a quietly powerful silicon strategy, and developer-tooling conversations intensifed around agentic orchestration, Cursor's Composer 2.5, and the growing AI bot spam problem on social platforms.
Key Events
-
Andrej Karpathy joins Anthropic as an IC to return to R&D, calling the next few years at the LLM frontier "especially formative" and noting he will resume education work in time. → link
-
Google I/O: Gemini Omni announced, a new model that can create anything from any input, starting with video generation. → link
-
Google Search now generates UIs on the fly using Gemini Flash 3.5 based on user queries. → link
-
Antigravity CLI and SDK launched globally, with early testers calling Gemini 3.5 Flash their favorite harness. → link
-
llama.cpp adds MTP support for Qwen 3.6 family, boosting Qwen3.6-27B dense generation from 25→45 tok/s (+78%) on A10G. → link
-
Apple's $599 MacBook Neo revealed to run on slightly defective A18 Pro chips (5 GPU cores vs 6), a strategy Apple has used for years to maximize silicon yield. → link
-
Cursor Composer 2.5 released, built on Moonshot's Kimi K2.5 open-source base. → link
-
Levelsio raises alarm over AI bot reply spam on X, noting bots are now quote-tweeting as a new attack vector. → link
-
Flutter 3.44.0 release notes published. → link
-
tinygrad opens Exabox preorders, positioning it as the cheapest deployable compute at launch. → link
Analysis
Talent consolidation at frontier labs. Karpathy's move to Anthropic—especially as an IC rather than in a leadership/education role—signals that top AI talent sees the next 2-3 years of foundation model development as decisive. Anthropic's recruiting run (satirized with the Wemby joke) is now a visible pattern, and the community reaction split between excitement and concern about concentration of expertise.
Local AI reaching pragmatic parity. Between Qwen 3.6 27B's MTP-enabled performance jump, the RTX 3090 as the recommended entry point, and dgx spark's underserved ecosystem, the local-first movement is coalescing around specific hardware-software stacks. The repeated "Opus 4.5 at home" framing suggests open-source models are closing the perceived quality gap with frontier proprietary models for coding tasks.
Agentic orchestration over raw intelligence. Multiple threads (Teknium/Hermes Agent, sudoingX's async agent bus, Codex "ask mode" requests) converge on the idea that the bottleneck is workflow architecture, not model capability. Agents committing state to git/folders asynchronously, rather than broadcasting via tmux, emerged as a concrete best practice.
What to watch: Google I/O follow-through on Gemini Omni capabilities and Antigravity adoption; Anthropic's first Karpathy-influenced research outputs; whether llama.cpp MTP support extends to other model families; Apple's defective-chip supply constraint on MacBook Neo; and the escalating AI bot spam arms race on X.
Tweet Feed
AI Industry Moves & Talent
@karpathy · 2026-05-19T15:05
Personal update: I've joined Anthropic. I think the next few years at the frontier of LLMs will be especially formative. I am very excited to join the team here and get back to R&D. I remain deeply passionate about education and plan to resume my work on it in time. → tweet link
@levelsio · 2026-05-19T17:35
My favorite person in AI joined my favorite AI company (well, shared favorite with xAI) → tweet link
@alexinexxx · 2026-05-19T16:47
my new litmus test is people's reaction to karpathy joining anthropic → tweet link
@TheAhmadOsman · 2026-05-19T17:23
Karpathy noooo \n\nYou weren't supposed to join the evil camp bro → tweet link
@TrungTPhan · 2026-05-19T15:36
Anthropic's new valuation after Andrej Karpathy joined the AI startup as an IC → tweet link
@badlogicgames · 2026-05-19T15:17
i was quite excited about everything i heard about his plans for education. kinda sad this won't happen for a while. but in return we'll get the most excellent flicker :D → tweet link
@TrungTPhan · 2026-05-19T17:41
"Ok, Google IO is about to start. Have Karpathy announce he's joining Anthropic." → tweet link
Google I/O — Gemini & Antigravity
@levelsio · 2026-05-19T18:49
RT @OfficialLoganK: Introducing Gemini Omni 🔮........ Omni is our new model that can create anything from any input — starting with video (… → tweet link
@ASalvadorini · 2026-05-19T18:09
RT @rakyll: Google Search is now generating UIs with Gemini Flash 3.5 on the fly based on your query, completely custom generated on the fl… → tweet link
@ASalvadorini · 2026-05-19T18:09
RT @rakyll: Antigravity CLI and SDK are now both available globally! With Gemini 3.5 Flash, Antigravity is my favorite harness. Fast and in… → tweet link
@ASalvadorini · 2026-05-19T18:07
We've to give it to @Google : they cooked 🔥\n\n#GoogleIO → tweet link
@ASalvadorini · 2026-05-19T17:27
Gemini Omni\n\n#GoogleIO → tweet link
@ASalvadorini · 2026-05-19T18:37
@ulusoyapps should we give 3.5 Flash a spin? 🤔\n\n#GoogleIO @antigravity → tweet link
@ASalvadorini · 2026-05-19T18:55
Testing #Flutter app with @antigravity 2.0 \n\n#GoogleIO → tweet link
@ASalvadorini · 2026-05-19T18:18
This is confusing at least 🤔\n\n#GoogleIO @antigravity → tweet link
@ASalvadorini · 2026-05-19T17:34
Now that you built a new os from scratch with 1k$ please finish Fuchsia next month\n\n#GoogleIO #Fuchsia #os @antigravity @Google → tweet link
@ASalvadorini · 2026-05-19T17:04
10 years since they're AI first company is a huge lie \n\n#GoogleIO → tweet link
@ASalvadorini · 2026-05-19T15:10
Only a couple of hours to go 🔥\n\n#GoogleIO #flutter → tweet link
Local AI & Open Source Models
@TheAhmadOsman · 2026-05-19T00:41
Gentle reminder that all you need to start with Local AI is:\n\n- 2x RTX 3090s (pick up for $700-$900 on r/hardwareswap)\n- Qwen 3.6 27B / Gemma 4 31B\n- Your favorite agent (Claude Code / OpenCode / etc)\n- Self-hosted SearXNG for web access\n\nAnd you got yourself Opus 4.5 at home → tweet link
@victormustar · 2026-05-18T19:27
llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀\n\nQwen3.6-27B dense generation (on A10G):\nFrom 25 tok/s → 45 tok/s (+78%).\n\nTwo flags on llama-server:\n--spec-type draft-mtp --spec-draft-n-max 2 → tweet link
@sudoingX · 2026-05-19T04:53
130 bookmarks. that means 130 of you are planning to try this.\n\nstop planning. go download qwen 3.6 27B dense Q4. compile llama.cpp from source. load it on your 3090.\n\nyou'll have the best local coding model running in under 30 minutes. i just told you the answer. now go use it. → tweet link
@TheAhmadOsman · 2026-05-19T04:42
Examples from 2 repos where I have been recognized recently\n\nThis stuff makes me so happy. It means I have impact and I am using that impact right → tweet link
@TheAhmadOsman · 2026-05-18T22:04
I have ran LLMs locally that had 2048 / 4096 / 8192 context windows\n\nThat alone keeps me from ever getting one-shot by an AI into a psychosis\n\nBeing able to tinker with these things is legit good for you CogSec, just saying → tweet link
@TheAhmadOsman · 2026-05-18T20:11
I keep getting this question\n\n- Where do I start with Local AI and selfhosting LLMs / models?\n\nHere is a thread for what I consider my most important educational content\n\nOpensource AI will win → tweet link
@TheAhmadOsman · 2026-05-18T19:05
Life after an RTX 3090 > Life before an RTX 3090 btw → tweet link
@Ex0byt · 2026-05-18T21:02
the 2.7, 2.8, 3.6, 3.7, 4.8, 5.2, 5.6 classes are coming -- get ready. → tweet link
@sudoingX · 2026-05-19T02:37
dgx spark has the most underserved ecosystem for what the hardware is capable of. complete white space. → tweet link
Agentic Orchestration & Developer Tools
@Teknium · 2026-05-19T18:18
RT @boxmining: This hits different @Teknium. Been saying the real power isn't in the model, it's in the workflow. Orchestration > raw intel… → tweet link
@sudoingX · 2026-05-18T19:36
yeah, tmux broadcasts will always be spotty, that is synchronous messaging, miss the moment and the message is gone.\n\nwhat worked for me was to stop having agents talk to each other at all. they go through the repo instead. an agent commits its state, the next one pulls and reads it, async, and git never loses a message because it is just files with history.\n\ni pushed it further, each agent has its own folder, pending, ongoing, done. a task is a file you drop in an agent's pending folder, it picks it up and moves it through.\n\nthat is the bus, durable by default. → tweet link
@sudoingX · 2026-05-19T03:38
running 6 agents at once is sweaty work, and nobody posts that part.\n\ni have two dedicated merge agents handling the merges and i am still sweating. the agents do the typing.\n\nkeeping all 6 pointed at the right thing, unblocked, not drifting, that is still all you.\n\nit does not get easier, it gets faster. faster is its own kind of load. → tweet link
@sudoingX · 2026-05-19T10:39
quick honest one. if you are using ai as a chatbot in a browser tab, you genuinely do not need any of this.\n\nit kicks in once your agents start doing real work, running for hours, spread across more than one machine, holding state you cannot lose. → tweet link
@RydMike · 2026-05-19T08:31
Today's Codex hot take:\nCodex needs ask mode like Cursor. It is such a nice simple safety net when you only want to analyze and explore a code base, no need for "Don't do anything only answer the questions" extra part in the prompt that does not always work. → tweet link
@thdxr · 2026-05-19T02:50
the whole sdk category that stainless was in never made much sense to me\n\nit turned what should be a simple run anywhere process and put it behind a cloud service that forced all these awkward workflows\n\nwe've been building on and sponsoring heyapi for the past year → tweet link
@crystalsssup · 2026-05-19T04:14
RT @cursor_ai: Composer 2.5 is built on the same open-source base as Composer 2, Moonshot's Kimi K2.5. → tweet link
@badlogicgames · 2026-05-19T16:54
wait, prompts are code, files are state?\n\nYC catching up to ca. Q2 2025 → tweet link
@sudoingX · 2026-05-19T02:13
a test for anything in your stack: remove it. if your day breaks, it is foundation. if nothing changes, it is decoration.\n\nmost of what people agonize over is decoration. the framework, the agent IDE, the prompt library, the model of the week, swap any of it and your work does not wobble.\n\nthe five in this post are the other kind. pull one and your flow stops. → tweet link
Research & Learning
@jsuarez · 2026-05-19T17:03
Reinforcement learning research with Joseph Suarez → tweet link
@jsuarez · 2026-05-19T13:53
SOTA over Puffer 4 on day 1. Next stream starts in a few hours! → tweet link
@badlogicgames · 2026-05-19T13:14
recommended reading! such a nice exploration of applying the "old" and new in ML to a tangible problem. having nostalgic feels when reading LDA. → tweet link
@badlogicgames · 2026-05-19T11:22
recommended reading, on a variation of @_can1357 's hash line read/edit tools for agents. → tweet link
Hardware & Silicon
@TrungTPhan · 2026-05-19T16:38
Macbook Neo is best example of a side benefit for Apple's custom silicon: next level expertise at re-using slightly defective chips.\n\nThe $599 hit laptop built on A18 Pro chips, which was first used for iPhone 16 (2024).\n\nApple needs chips for 200m+ iPhones a year and manufacturing process obviously imperfect (a lot of them have slight defects).\n\nSome defective chips are fine to power cheaper devices. → tweet link
@TrungTPhan · 2026-05-19T17:56
RT @bearlyai: Other than the A18 Pro chip (hard to board swap), Apple's $599 Neo is likely the most repairable Macbook ever: \n\n> modular pi… → tweet link
@TheAhmadOsman · 2026-05-19T02:22
Which NVIDIA GPU for Local AI you ask?\n\n(From a new article I am working on) → tweet link
@tinygrad · 2026-05-18T23:50
To everyone who wants to invest in tiny, preorder an exabox. At launch, it will be the cheapest compute you can buy. Deploy it and run it and make returns!\n\nWe don't want VCs who invest other people's pensions with only upside potential for them. We want people in the trenches. → tweet link
Platform Issues — AI Bot Spam
@levelsio · 2026-05-19T18:01
The AI reply problem is so big right now that I have no choice other than to restrict my replies to this\n\nYou can still QT me though and I can see it\n\nI started seeing AI bots now also discover they can QT btw (new attack vector) → tweet link
Security & Practical AI Usage
@levelsio · 2026-05-19T11:54
A nice way to stay safe is to ask Claude Code to audit your devices\n\nI do same on my VPS servers, so today I tried it on my MacBook Pro and it's pretty good at it too\n\nIt found lots of stuff that was not secured, I actually forgot to enable FileVault when I got this new MBP in 2025\n\nJust ask it "can you security audit my computer" → tweet link
Developer Ecosystem & Tools
@jezell · 2026-05-19T07:33
Flutter 3.44.0 release notes are up → tweet link
@badlogicgames · 2026-05-19T22:19
it's 2026 and publishing to maven central is still an absolute horror show. incredible. → tweet link
@badlogicgames · 2026-05-19T07:57
what good is an "industry standard" saying the frontmatter must be valid YAML, when flicker company then turns around and just accepts any invalid YAML, so other harnesses have to also implement it that way?\n\nthis is very upsetting. → tweet link
@steipete · 2026-05-19T10:24
RT @boxmining: OpenClaw 5.18 (@openclaw) feels like one of those updates that quietly makes agents way less annoying to run. → tweet link
@steipete · 2026-05-19T10:06
RT @sorenbs: The Bun Rust port has fixed at least one critical memory leak 🤘 → tweet link
@badlogicgames · 2026-05-19T00:44
Due to popular demand, you can now browse all my recommended reading/viewing suggestions here: → tweet link
@kunchenguid · 2026-05-19T02:45
i've been in the front row seat of tech companies' AI adoption and layoffs\n\nshould i make a post / video to explain what's happening? → tweet link
@ASalvadorini · 2026-05-19T03:07
RT @chimon1984: There's an imbalance between code generation speed and the experience required to validate it. → tweet link
@nummanali · 2026-05-19T14:03
TLDR - Everyone becomes a Product Engineer → tweet link