# This Week in All Things AI - Week 22-2026

*Sunday 24th May 2026 to Saturday 30th May 2026*

By [This Week in All Things AI](https://paragraph.com/@twiata) · 2026-05-31

---

This Week in All Things AI covers key developments in models, agents, tools, infrastructure, and policy curated from discussions in the [**All Things AI Telegram group.**](https://t.me/+PnCnwhgH8V4yMTFl)  
  
If you follow AI for work, research, investing, or just to understand where the technology is heading, this [**weekly brief**](https://paragraph.com/@twiata) is a concise way to scan the most important launches, risks, and resources in a few focused minutes.

Major developments this week included xAI completing training on Grok V9-Medium (1.5T parameters) with a public release expected in 2–3 weeks, alongside plans to open-source its current v8-small model by year-end. OpenRouter raised $113M at a $1.3B valuation as Xiaomi's MiMo slashed API pricing by up to 99% and MiniMax teased its M3 sparse-attention architecture. Anthropic released Opus 4.8, Stanford launched the open-source OpenJarvis agent framework, and Cursor announced its inaugural Compile conference. China's Shanghai Futures Exchange began researching AI token futures as Tomasz Tunguz argued that competitive advantage now lies in the orchestration layer rather than foundation models. The week closed with Kirkland & Ellis committing $500M to build a proprietary in-house AI platform and ElevenLabs launching Dubbing v2, a model that preserves speaker emotion, tone, and pacing across 90+ languages.

The sections that follow walk through these items day by day, with short context and links so you can dive deeper into the pieces most relevant to your work or interests.

* * *

Sunday 24th May 2026
--------------------

Monday 25th May 2026
--------------------

Elon Musk announced that training of Grok foundation model V9-Medium with 1.5 trillion parameters is complete, with positive evaluation results. 

Supplementary training included substantial Cursor data, with fine-tuning now underway and reinforcement learning scheduled to start in a few days. 

Public release is expected in 2-3 weeks, marking a major upgrade over the current 0.5T v8-small model, particularly for challenging coding tasks.

He also announces that xAI will open-source its current 0.5T Grok model (v8-small) by the end of 2026, stating it should remain quite useful even after newer models launch.

[![User Avatar](https://storage.googleapis.com/papyrus_images/5848bccca58719b811c7fb02b5f45961d7b1ebf8c07f3b15d4288a08097885c5.jpg)](https://twitter.com/elonmusk)

[Elon Musk](https://twitter.com/elonmusk)

[@elonmusk](https://twitter.com/elonmusk)

[](https://twitter.com/elonmusk/status/2058787384364265734)

Grok foundation model V9-Medium (1.5T) has finished training. Evals look good. A lot of Cursor data was added in supplementary training and there is more to come.  
  
Fine-tuning is underway and reinforcement learning begins in a few days. 2 to 3 weeks to public release.  
  
This will

[70.2K](https://twitter.com/elonmusk/status/2058787384364265734)[

5:48 AM • May 25, 2026

](https://twitter.com/elonmusk/status/2058787384364265734)

[![User Avatar](https://storage.googleapis.com/papyrus_images/5848bccca58719b811c7fb02b5f45961d7b1ebf8c07f3b15d4288a08097885c5.jpg)](https://twitter.com/elonmusk)

[Elon Musk](https://twitter.com/elonmusk)

[@elonmusk](https://twitter.com/elonmusk)

[](https://twitter.com/elonmusk/status/2058796067592736866)

We will open source the 0.5T model towards the end of this year. It should still be quite useful.

[2,932](https://twitter.com/elonmusk/status/2058796067592736866)[

6:23 AM • May 25, 2026

](https://twitter.com/elonmusk/status/2058796067592736866)

Tuesday 26th May 2026
---------------------

Wednesday 27th May 2026
-----------------------

OpenRouter raised $113M led by CapitalG, a source says at a $1.3B valuation, and now processes 25T tokens across 400+ models weekly, up from 5T six months ago

[![User Avatar](https://storage.googleapis.com/papyrus_images/d4e597dab8b7cc67732a61849ae3e345cf14e720929622cccccefc98f223556f.jpg)](https://twitter.com/alexatallah)

[Alex Atallah](https://twitter.com/alexatallah)

[@alexatallah](https://twitter.com/alexatallah)

[](https://twitter.com/alexatallah/status/2059287308151538174)

We're seeing a Cambrian explosion of AI models, and it's happening on OpenRouter.  
  
The future of AI is neurodiversity:  
\- Agents choosing the most cost-effective model/provider/tool for the task  
\- Agents orchestrating multiple models for the smartest result  
\- Advanced security and

[![User Avatar](https://storage.googleapis.com/papyrus_images/df50010f9252281a515979ea904651dc32f918d2e7e0026f60c64bdb64a859c1.jpg)](https://twitter.com/OpenRouter)

[OpenRouter](https://twitter.com/OpenRouter)

[@OpenRouter](https://twitter.com/OpenRouter)

[](https://twitter.com/OpenRouter/status/2059277623629664758)

Today we’re announcing our $113M Series B led by @CapitalGVC.  
  
Over the last 6 months, weekly volume on OpenRouter grew from 5T to 25T tokens as AI rapidly shifts from experimentation into production.  
  
We’re excited for what comes next.

![](https://storage.googleapis.com/papyrus_images/9f7b1e11299ed572f781dd490dfea9677e71e16eec7f696b2dbc0cc913f84362.png)

[219](https://twitter.com/alexatallah/status/2059287308151538174)[

2:55 PM • May 26, 2026

](https://twitter.com/alexatallah/status/2059287308151538174)

[![User Avatar](https://storage.googleapis.com/papyrus_images/d6718c8993e1c26bda0cc4dbafe736fd6d9e09903337fe733f206537039cbeb7.jpg)](https://twitter.com/deedydas)

[Deedy](https://twitter.com/deedydas)

[@deedydas](https://twitter.com/deedydas)

[](https://twitter.com/deedydas/status/2059298453872947623)

OpenRouter is now serving 1.5 quadrillion tokens/yr!  
  
That token run rate is:  
— 15-30% of Google APIs  
— 20-40% of OpenAI  
— >50% of Microsoft Azure Foundry  
  
That's 15x larger than when we invested a year ago.  
  
Revenue has already doubled since this $1.3B round was done in Feb!

![](https://storage.googleapis.com/papyrus_images/7bdf8af8fd117f18952a2ef1edac6319589df27f2e4aaa04d35f763c650c3109.jpg)

[287](https://twitter.com/deedydas/status/2059298453872947623)[

3:39 PM • May 26, 2026

](https://twitter.com/deedydas/status/2059298453872947623)

[https://www.nytimes.com/2026/05/26/business/dealbook/openrouter-ai-models-fundraising.html](https://www.nytimes.com/2026/05/26/business/dealbook/openrouter-ai-models-fundraising.html)

Xiaomi MiMo announced permanent API price cuts of up to 99% for MiMo-V2.5-Pro and MiMo-V2.5 models, with unified pricing across context lengths and new low rates like $0.0036 per 1M input cache hit tokens for Pro. 

Token plans received major upgrades delivering 5–8× more credits at the same price points, with all existing subscriber credits fully reset as a thank-you, while MiMo-V2.5-TTS stays free temporarily.

[![User Avatar](https://storage.googleapis.com/papyrus_images/29ec46febc6b0f0b74ad155ac4427abc394c06da62c703ab55a17af48c52b6c9.jpg)](https://twitter.com/XiaomiMiMo)

[Xiaomi MiMo](https://twitter.com/XiaomiMiMo)

[@XiaomiMiMo](https://twitter.com/XiaomiMiMo)

[](https://twitter.com/XiaomiMiMo/status/2059314052892099070)

![🚀](https://abs-0.twimg.com/emoji/v2/72x72/1f680.png) Better inference efficiency, lower costs, broader access.  
  
MiMo-V2.5 Series API pricing is now permanently reduced — by up to 99% compared to previous pricing.  
![✨](https://abs-0.twimg.com/emoji/v2/72x72/2728.png) Unified pricing across all context lengths.  
MiMo Token Plans have also been upgraded:  
• 5–8× more usable tokens

![](https://storage.googleapis.com/papyrus_images/eb75340e02d28725f72f7ab88e889bffbe806fb1f2db2959a0986c79e1aeddf3.jpg)

[4,185](https://twitter.com/XiaomiMiMo/status/2059314052892099070)[

4:41 PM • May 26, 2026

](https://twitter.com/XiaomiMiMo/status/2059314052892099070)

[![User Avatar](https://storage.googleapis.com/papyrus_images/0c8630743c0f7dee330a7190946ce21ca1623d789f6c901b98b390b513d3cbff.jpg)](https://twitter.com/_LuoFuli)

[Fuli Luo](https://twitter.com/_LuoFuli)

[@\_LuoFuli](https://twitter.com/_LuoFuli)

[](https://twitter.com/_LuoFuli/status/2059618247553745204)

Behind the MiMo API Price Reduction:  
The deepest price cut, up to 99%, is for Input (Cache Hit). The core reason is our inference framework now supports hierarchical KV cache optimization for SWA. Production inference engine tests show this optimization increases cached token

[1,736](https://twitter.com/_LuoFuli/status/2059618247553745204)[

12:50 PM • May 27, 2026

](https://twitter.com/_LuoFuli/status/2059618247553745204)

[

Xiaomi MiMo Api Open Platform - Token Plan Global Launch
--------------------------------------------------------

One-time purchase unlocks both MiMo-V2.5 flagship models, plus TTS model free across all tiers for a limited time. Unleash powerful productivity with Xiaomi MiMo

https://platform.xiaomimimo.com

![Xiaomi MiMo Api Open Platform - Token Plan Global Launch](https://storage.googleapis.com/papyrus_images/91be98414ef86e708fc4a9476ad5f75ff54a5e7b1643e8bd3bc5115fec813c32.jpg)

](https://platform.xiaomimimo.com/docs/en-US/news/v2.5-price-update)

Skyler Miao, Head of Engineering at MiniMax\_AI, posted a teaser for the upcoming M3 model showing their new sparse attention architecture.  The diagram details a two-stage GQA-based system: an Index Branch quickly scans and selects top-k relevant token blocks via block max pooling, then a Sparse Branch performs full attention only on those blocks. 

This design delivers 9.7x faster prefill and 15.6x faster decoding at 1M token context versus M2, enabling efficient long-context AI without prohibitive compute costs.

[![User Avatar](https://storage.googleapis.com/papyrus_images/2930c6605c16327f97ba57cad844a8460623a1ffab2506ca8da904c280fb6ef6.jpg)](https://twitter.com/SkylerMiao7)

[Skyler Miao](https://twitter.com/SkylerMiao7)

[@SkylerMiao7](https://twitter.com/SkylerMiao7)

[](https://twitter.com/SkylerMiao7/status/2059285750458544561)

Something BIG is coming

![](https://storage.googleapis.com/papyrus_images/c402531f837c491e8af2bc4de468af856c923311eb7c46b000c22baae8e7d7f9.jpg)

[3,341](https://twitter.com/SkylerMiao7/status/2059285750458544561)[

2:49 PM • May 26, 2026

](https://twitter.com/SkylerMiao7/status/2059285750458544561)

[![User Avatar](https://storage.googleapis.com/papyrus_images/87e3fff1c6cbe57cb9eafb066ac3e82ef761170b401b43d5695cc622c01e2330.jpg)](https://twitter.com/eliebakouch)

[elie](https://twitter.com/eliebakouch)

[@eliebakouch](https://twitter.com/eliebakouch)

[](https://twitter.com/eliebakouch/status/2059321928205156568)

new minimax sparse attention compared to deepseek v3.2 (DSA) and v4 (CSA)  
  
main changes:  
\- based on GQA not MLA  
\- block level selection like in CSA but attention is done on the real KV, not in the compressed dimension

![](https://storage.googleapis.com/papyrus_images/97a162aa6dd77ce7decda5ee4272692e2ec48ec81b1bb052edfcc391cd343db7.jpg)

[![User Avatar](https://storage.googleapis.com/papyrus_images/2930c6605c16327f97ba57cad844a8460623a1ffab2506ca8da904c280fb6ef6.jpg)](https://twitter.com/SkylerMiao7)

[Skyler Miao](https://twitter.com/SkylerMiao7)

[@SkylerMiao7](https://twitter.com/SkylerMiao7)

[](https://twitter.com/SkylerMiao7/status/2059285750458544561)

Something BIG is coming

![](https://storage.googleapis.com/papyrus_images/c402531f837c491e8af2bc4de468af856c923311eb7c46b000c22baae8e7d7f9.jpg)

[697](https://twitter.com/eliebakouch/status/2059321928205156568)[

5:13 PM • May 26, 2026

](https://twitter.com/eliebakouch/status/2059321928205156568)

[![User Avatar](https://storage.googleapis.com/papyrus_images/0d7ef457f13ae0587c4a85261dece652c2acafb85aedeecb2d52a399b4631708.jpg)](https://twitter.com/rameswar08)

[Rameswar](https://twitter.com/rameswar08)

[@rameswar08](https://twitter.com/rameswar08)

[](https://twitter.com/rameswar08/status/2059305973236596879)

For non ai people:  
  
most ai models work like this,  
  
every word looks at every other word to understand context  
  
great for accuracy, terrible for speed at massive scale  
  
MiniMax's sparse attention changes that  
  
instead of processing an entire 1m token context deeply, the model

[221](https://twitter.com/rameswar08/status/2059305973236596879)[

4:09 PM • May 26, 2026

](https://twitter.com/rameswar08/status/2059305973236596879)

Thursday 28th May 2026
----------------------

Cursor Announces Invite-Only Compile Conference in San Francisco

The one-day Compile event happens June 16 at Fort Mason, bringing together engineers, researchers, and builders for discussions on AI-native development. Speakers like Cursor's Michael Truell and Ryo Lu, plus guests from Every, Shopify, and indie makers, will work through ideas live on a chalkboard stage. It's waitlist-only now—sign up at cursor.com/compile and invitations go out via email—with a call for papers still open on rethinking systems and simplifying complex ideas.

[![User Avatar](https://storage.googleapis.com/papyrus_images/9036a3c92d415f41257c8e035d81b161c8c2bb45207ddeeba39395e5059e8b2c.jpg)](https://twitter.com/cursor_ai)

[Cursor](https://twitter.com/cursor_ai)

[@cursor\_ai](https://twitter.com/cursor_ai)

[](https://twitter.com/cursor_ai/status/2059673762728116442)

We're hosting an event on June 16th in San Francisco.  
  
Compile is a one-day event that brings together engineers, researchers, designers, and builders of all kinds to discuss the future of software.  
[

![](https://storage.googleapis.com/papyrus_images/7499b8fae246161f705b524dc21d40cd56610248e743cbb0e773aaa5e1b6001c.jpg)

cursor.com

Cursor · Compile
----------------

Compile is Cursor's inaugural conference — bringing together developers, researchers, and teams shaping the future of AI-native development.





](https://t.co/8YERlPFooL)

[1,224](https://twitter.com/cursor_ai/status/2059673762728116442)[

4:31 PM • May 27, 2026

](https://twitter.com/cursor_ai/status/2059673762728116442)

Tomasz Tunguz of Theory Ventures with a post where in he articulates that in the AI era, the core differentiator in software is not the model itself but the “harness” layer that tames a general LLM into a reliable, domain-specific agent by combining seven capabilities (context, tools, orchestration, state, sandboxing, observability, and cost optimization).

The article argues that when everyone can access similar foundation models, competitive advantage shifts to whoever builds the best harness around the model. This harness “domesticates” a powerful but wild LLM into a dependable system that can safely execute real workflows in specific industries.

[

Software After AI
-----------------

Software is no longer about UX & data. It is about the harness, the layer that turns an LLM into a reliable agent. Seven components define the new stack.

https://tomtunguz.com

![Software After AI](https://storage.googleapis.com/papyrus_images/b939cd5fc8c4ab8b428ed11791356b9be27f65d3441948e5cf292bcff5db5e65.png)

](https://tomtunguz.com/harnessing-ai/)

In response [Brent Maxwell](https://t.me/brentmaxwell) writes

In my engineering teams, nobody can tell the difference between gpt-5.5, claude-4.7, gemini-3.5-flash or composer-2.5.

There is no winner right now in the model wars.

The harnesses are totally the right thing to focus on - they are making or breaking development processes for us. When we figure out how to manate the harness better, we ship faster with less time stuck in PR. When one of our devs just uses the default agent config in the IDE, it really doesn't give them the same amount of leverage.

Shanghai Futures Exchange Explores AI Token Contracts

China’s Shanghai Futures Exchange is in early-stage research on futures contracts tied to AI tokens, the smallest unit of information processed and billed for by AI models. These contracts would let companies hedge against volatile AI compute costs along the AI supply chain, similar in spirit to commodity or energy futures.

While U.S. exchanges like CME and ICE are moving toward futures tied to GPU rental/compute capacity, China’s concept would be directly linked to AI token consumption used for pricing AI services. This represents a different abstraction layer: the U.S. focuses on hardware capacity, whereas China targets the usage-based “digital fuel” that powers AI models.

China’s daily AI token usage has exploded roughly 1,000x since early 2024, reaching more than 140 trillion tokens a day by March 2026, underscoring surging demand and cost exposure. Token-based derivatives are being framed as a potential new asset class, with figures like BlackRock’s CEO noting that futures on compute could become a distinct financial market.

[

Exclusive: China works on AI token futures market, sources say, in race with US
-------------------------------------------------------------------------------

China is designing a futures market for AI tokens, sources familiar with ​the matter said, as the country potentially takes a different tack to U.S. exchanges developing compute power futures to tap ‌the rapidly growing appetite to hedge AI costs.

https://www.reuters.com

![Exclusive: China works on AI token futures market, sources say, in race with US](https://storage.googleapis.com/papyrus_images/a0514061bfed20e9453e5b89e884d96d47739c0c68cd6cca5ab5ab7b1a8c85d4.jpg)

](https://www.reuters.com/world/china/china-works-ai-token-futures-market-sources-say-race-with-us-2026-05-28/)

Friday 29th May 2026
--------------------

Some links about Anthropic's Opus 4.8 launch  including a tweet thread from Anthropic with some guidance on devs as they migrate to Opus 4.8 and some commentary from Dan Shipper of Every who had access to Opus 4.8 for the past two weeks

[![User Avatar](https://storage.googleapis.com/papyrus_images/e6f736ef39a919cd51cb27919ba706916ba8718627b6635dde46a15532a2c415.png)](https://twitter.com/ClaudeDevs)

[ClaudeDevs](https://twitter.com/ClaudeDevs)

[@ClaudeDevs](https://twitter.com/ClaudeDevs)

[](https://twitter.com/ClaudeDevs/status/2060043208277811437)

Opus 4.8 is live in Claude Code today.  
  
A few things worth knowing: ![🧵](https://abs-0.twimg.com/emoji/v2/72x72/1f9f5.png)

[![User Avatar](https://storage.googleapis.com/papyrus_images/dc69ec28be72de3fb258b612f16a89ba4ffa6bf64d6b63683c69051b6c38c662.jpg)](https://twitter.com/claudeai)

[Claude](https://twitter.com/claudeai)

[@claudeai](https://twitter.com/claudeai)

[](https://twitter.com/claudeai/status/2060042702150930686)

Introducing Claude Opus 4.8: it builds on Opus 4.7 with sharper judgment, more honesty about its own progress, and the ability to work independently for longer than its predecessors.  
  
Available today at the same price.

![](https://storage.googleapis.com/papyrus_images/3c2216a3788af875df288a3d45ef0c92260a5e559f01e486713f7c64f9612558.png)

[9,585](https://twitter.com/ClaudeDevs/status/2060043208277811437)[

4:59 PM • May 28, 2026

](https://twitter.com/ClaudeDevs/status/2060043208277811437)

[![User Avatar](https://storage.googleapis.com/papyrus_images/9a24e89bcea6545e47519412d53c6a01304c5edd527986785aa7374097b8e18a.jpg)](https://twitter.com/danshipper)

[Dan Shipper 📧](https://twitter.com/danshipper)

[@danshipper](https://twitter.com/danshipper)

[](https://twitter.com/danshipper/status/2060043738752422304)

BREAKING:  
  
Anthropic just dropped Opus 4.8—and it is a MONSTER  
  
We've been testing for about a week [@every](https://twitter.com/every) and our verdict is they could've just called it Opus 5, it's that good.  
  
Here's our vibe check:  
  
\- Beats GPT-5.5 on Senior Engineer bench. On our toughest benchmark Opus

![](https://pbs.twimg.com/amplify_video_thumb/2060043248962347013/img/AQmmXi73h30gCM5S.jpg)

[1,625](https://twitter.com/danshipper/status/2060043738752422304)[

5:01 PM • May 28, 2026

](https://twitter.com/danshipper/status/2060043738752422304)

[

Vibe Check: Opus 4.8-Anthropic Should've Rounded Up to 5
--------------------------------------------------------

Opus 4.8 tops both our Senior Engineer benchmark and our writing tests. It's the most complete model we've tested. We just wish it had an app to match.

https://every.to

![Vibe Check: Opus 4.8-Anthropic Should've Rounded Up to 5](https://storage.googleapis.com/papyrus_images/5c42f7aa840419aab244f090fcb0320e33371de69331289467f439db213f712b.png)

](https://every.to/vibe-check/opus-4-8-vibecheck)

Stanford Unveils OpenJarvis: Efficient AI Agents for Your Devices

Stanford researchers from Hazy Research and Scaling Intelligence Lab released OpenJarvis v1.0, an open-source framework that builds personal AI agents to run locally on devices. It emphasizes 'Intelligence per Watt' with swappable components like local models, engines, agents, tools for apps like Slack and email, and self-improvement features. Users get eight ready-to-run agents for tasks such as morning briefings and code review, installable via a simple one-line script across CLI, web, desktop, or messaging apps. Benchmarks show it handles most queries at interactive speeds with far lower costs and latency than cloud systems, positioning it as a privacy-focused alternative to massive data centers.

[![User Avatar](https://storage.googleapis.com/papyrus_images/360192767b4c6a90de38a33b6b2bcf21453710aeb0b586794d46ae7f2e62a348.jpg)](https://twitter.com/JonSaadFalcon)

[Jon Saad-Falcon](https://twitter.com/JonSaadFalcon)

[@JonSaadFalcon](https://twitter.com/JonSaadFalcon)

[](https://twitter.com/JonSaadFalcon/status/2060054559142326468)

The dominant story in AI has been the growing cloud: bigger clusters, larger models, more gigawatts.  
  
We believe the future is in the opposite direction: on-device inference, smaller models, watts instead of gigawatts.  
  
Today we're releasing [@OpenJarvisAI](https://twitter.com/OpenJarvisAI) v1.0: a personal AI

![](https://pbs.twimg.com/amplify_video_thumb/2060052644547493888/img/-mK4VSutVh3zOZdq.jpg)

[527](https://twitter.com/JonSaadFalcon/status/2060054559142326468)[

5:44 PM • May 28, 2026

](https://twitter.com/JonSaadFalcon/status/2060054559142326468)

[![User Avatar](https://storage.googleapis.com/papyrus_images/6554c2c190bb0f3d57e27c016668ea0d295784032c70633fd725426489c1f047.png)](https://twitter.com/LambdaAPI)

[Lambda](https://twitter.com/LambdaAPI)

[@LambdaAPI](https://twitter.com/LambdaAPI)

[](https://twitter.com/LambdaAPI/status/2060073096741298522)

Most agent frameworks are built around one cloud model. Swap in a local model, performance drops.  
  
[@OpenJarvisAI](https://twitter.com/OpenJarvisAI) fixes the harness, not the model. Result: 77% of the accuracy gap recovered, 800x lower cost per query, 4x lower latency.  
  
Built on Lambda. Open-sourced from Stanford.

![](https://storage.googleapis.com/papyrus_images/3ecfda3ed114b8d804bde777cb092b3c905a9f757aed90b6a21832708c7ac148.jpg)

[14](https://twitter.com/LambdaAPI/status/2060073096741298522)[

6:57 PM • May 28, 2026

](https://twitter.com/LambdaAPI/status/2060073096741298522)

[![User Avatar](https://storage.googleapis.com/papyrus_images/f73727fc676c1130922ca8515403ca893b686bec8f2fc949d9eac2632c13336e.jpg)](https://twitter.com/SnorkelAI)

[Snorkel AI](https://twitter.com/SnorkelAI)

[@SnorkelAI](https://twitter.com/SnorkelAI)

[](https://twitter.com/SnorkelAI/status/2060060964800844080)

Huge congrats to @jonsaadfalcon, [@Avanika15](https://twitter.com/Avanika15), [@Azaliamirh](https://twitter.com/Azaliamirh) and the [@HazyResearch](https://twitter.com/HazyResearch) team on [@OpenJarvisAI](https://twitter.com/OpenJarvisAI) — out today.  
  
For two years, they've been making the case that AI inference belongs on hardware people already own, not just in megawatt data centers. Excited to support the

![](https://storage.googleapis.com/papyrus_images/fe4b4b2c0910e460a4c4bfa678a16972b44bfb9d48d9bad826f8c665be36f9bc.jpg)

[![User Avatar](https://storage.googleapis.com/papyrus_images/360192767b4c6a90de38a33b6b2bcf21453710aeb0b586794d46ae7f2e62a348.jpg)](https://twitter.com/JonSaadFalcon)

[Jon Saad-Falcon](https://twitter.com/JonSaadFalcon)

[@JonSaadFalcon](https://twitter.com/JonSaadFalcon)

[](https://twitter.com/JonSaadFalcon/status/2060054559142326468)

The dominant story in AI has been the growing cloud: bigger clusters, larger models, more gigawatts.  
  
We believe the future is in the opposite direction: on-device inference, smaller models, watts instead of gigawatts.  
  
Today we're releasing [@OpenJarvisAI](https://twitter.com/OpenJarvisAI) v1.0: a personal AI

![](https://pbs.twimg.com/amplify_video_thumb/2060052644547493888/img/-mK4VSutVh3zOZdq.jpg)

[18](https://twitter.com/SnorkelAI/status/2060060964800844080)[

6:09 PM • May 28, 2026

](https://twitter.com/SnorkelAI/status/2060060964800844080)

via [shrwn](https://t.me/shrwn)

Free Google x Kaggle vibe coding course

\* Mid June  
\* 1-2hrs/day for 5 days   
\* Free

[

5-Day AI Agents: Intensive Vibe Coding Course With Google
---------------------------------------------------------

June 15 - 19, 2026

https://www.kaggle.com

![5-Day AI Agents: Intensive Vibe Coding Course With Google](https://storage.googleapis.com/papyrus_images/7283a732ac78ac9af8ed90ff179a25fe0fee8b8b77bd438b35bdd273f2e4402a.png)

](https://www.kaggle.com/competitions/5-day-ai-agents-intensive-vibecoding-course-with-google)

Saturday 30th May 2026
----------------------

Kirkland & Ellis LLP is one of the world’s most successful and profitable “Big Law” firms with approximately 4,000 attorneys across 23 offices. It consistently ranks at or near the top of revenue and profitability metrics.

Kirkland & Ellis is committing $500 million over the next 3–4 years to build its own proprietary AI platform and custom tools. This is one of the largest and most ambitious technology investments ever announced by a law firm.

Key details include:

*   Funding & Timeline: Funded entirely from the firm’s own revenue (aligning with Chair Jon Ballis’s philosophy of investing roughly 1% of revenue in new initiatives — which would be ~$100M+ in a $10B+ year). More than $100 million is expected in 2026 alone, with the balance spread over the following years.
    
*   Goals: Create a broad, firm-wide AI platform that captures and deploys the firm’s “collective intelligence” to support lawyers across practices “start to finish” on client work. The aim is to move beyond reliance on multiple third-party tools toward more integrated, customized capabilities.
    
*   Development Approach: Informed by input from 250 lawyers (including 100 partners) on real workflows and needs. A team of 180 tech professionals is involved, working with undisclosed external partners/companies to help build the technology. Crucially, the resulting tools and IP will not be commercialized or sold to other law firms — a deliberate contrast to some other firm-vendor partnerships (e.g., Freshfields’ work with Anthropic).
    
*   Technical Elements: Involves on-premise GPU environments and Microsoft Azure-based AI infrastructure, including facilities for training and inference. The firm is actively hiring for AI-related roles (dozens of positions), including high-compensation roles like AI Infrastructure Directors.
    
*   Context Within the Firm: This builds on existing efforts. Kirkland already deploys third-party legal AI tools such as Harvey across its attorneys. It has a history of building proprietary technology in-house, including SideTrack (a tool for investment fund work, particularly around MFN issues) and earlier databases like CTRAN for M&A competitive intelligence. It also maintains internal innovation teams, AI Innovation Advisors embedded in practice groups, and responsible AI governance structures.
    

This move sits at the center of a key debate in legal technology: buy vs. build for AI. Many top firms rely heavily on specialized platforms like Harvey, CoCounsel (Thomson Reuters), or Lexis+ AI. Kirkland’s approach prioritizes:

*   Greater control over data, customization, and roadmap.
    
*   Using the firm’s own vast proprietary knowledge and deal experience as a competitive differentiator (rather than everyone having access to similar generic or vendor tools).
    
*   Enhanced security, confidentiality, and alignment with internal workflows.
    
*   Long-term ownership rather than ongoing licensing dependency.  
    

[

Kirkland & Ellis has form for building its own technology. The $500m AI play is its biggest yet. - Legal IT Insider
-------------------------------------------------------------------------------------------------------------------

Kirkland & Ellis announced that it would spend $500m over the next three to four years developing its own custom AI tools and services.

https://legaltechnology.com

![Kirkland & Ellis has form for building its own technology. The $500m AI play is its biggest yet. - Legal IT Insider](https://storage.googleapis.com/papyrus_images/c7da21b0c6cd10bed62137c92bdbb08e9ac9abfb29b899e64a66f7e7818cc0a9.png)

](https://legaltechnology.com/kirkland-ellis-has-form-for-building-its-own-technology-the-500m-ai-play-is-its-biggest-yet/)

ElevenLabs launches Dubbing v2, which it says preserves the original speaker's emotion, tone, and pacing across 90+ languages while staying synced to content

[![User Avatar](https://storage.googleapis.com/papyrus_images/f3a777f97fb00291b89ed14e346d3ed4027dc7b009c58933107578c2f2fa138d.jpg)](https://twitter.com/0xaneri)

[aneri](https://twitter.com/0xaneri)

[@0xaneri](https://twitter.com/0xaneri)

[](https://twitter.com/0xaneri/status/2060038836877705516)

Introducing Dubbing v2.  
  
For the first time, AI dubbing preserves how something was said, not just what was said. Dubbing v2 reads the original audio directly rather than just the transcript, so your emotion, tone, and delivery carries across 100+ languages.  
  
Every system before

![](https://pbs.twimg.com/amplify_video_thumb/2060033890493054979/img/M6XA_inLKPBXkj1-.jpg)

[49](https://twitter.com/0xaneri/status/2060038836877705516)[

4:41 PM • May 28, 2026

](https://twitter.com/0xaneri/status/2060038836877705516)

[

Introducing Dubbing v2: our revolutionary new dubbing model
-----------------------------------------------------------

Introducing Dubbing v2: our revolutionary new dubbing model, which carries the emotion and performance of the original speaker across every language.

https://elevenlabs.io

![Introducing Dubbing v2: our revolutionary new dubbing model](https://storage.googleapis.com/papyrus_images/53ad1bd9d51342150da444c542a8df00289be37c69868d3c416a3d20674983f9.webp)

](https://elevenlabs.io/blog/introducing-dubbing-v2)

* * *

Below is my personal website which aggregates links to many of my socials as well as the various content and community that I curate. Feel free to share this link to others who you think may find this content/community useful to them

[**https://linktr.ee/goolamabbas**](https://linktr.ee/goolamabbas)

The cover image of this newsletter via generated via the **Krea 2 Large** model within the [**Krea**](https://www.krea.ai/refer/EJQQP9FJ) tool via the following prompt

![](https://paragraph.com/editor/callout/information-icon.png)

Automobile inspired modern curved building , where the curved building's façade meticulously crafted from gears, cogs, and mechanical parts. The building's façade design embodies the fusion of nature and machinery, as evidenced by the intertwining vines and mechanical components

---

*Originally published on [This Week in All Things AI](https://paragraph.com/@twiata/this-week-in-all-things-ai-week-22-2026)*
