# This Week in All Things AI - Week 29-2026

*Sunday 12th July 2026 to Saturday 18th July 2026*

By [This Week in All Things AI](https://paragraph.com/@twiata) · 2026-07-19

---

This Week in All Things AI covers key developments in models, agents, tools, infrastructure, and policy curated from discussions in the [**All Things AI Telegram group.**](https://t.me/+PnCnwhgH8V4yMTFl)  
  
If you follow AI for work, research, investing, or just to understand where the technology is heading, this [**weekly brief**](https://paragraph.com/@twiata) is a concise way to scan the most important launches, risks, and resources in a few focused minutes.

The week of 12th July to 18th July 2026 brought a series of model, product and workflow developments. A purported internal Zhipu letter outlined AGI ambitions and plans to open‑release GLM‑5.2 with a one‑million‑token context window; Thinking Machines released the open‑weights multimodal Inkling model; and Moonshot launched Kimi K3, a 2.8T‑parameter multimodal model with 1M context and planned open weights. Anthropic introduced Claude for Teachers, announced Claude Fable 5 would be included in Max and Team Premium plans from July 20, and Google‑backed Key Studio offered support for early‑stage AI startups; tools such as Vorflux, Gemini Notebook, and PenEcho's shared AI canvas continued the shift from AI assistance towards more autonomous software and research workflows.

Safety, governance and deployment economics were equally prominent. MCPTox testing and Warrant's policy‑gated action layer drew attention to tool‑poisoning and destructive‑agent risks in production environments, while Demis Hassabis renewed calls for stronger standards for frontier models and President Xi's first appearance at the World AI Conference signalled China's state‑level commitment to open AI development. Chamath Palihapitiya's CNBC comments on the widening cost gap between Western and Chinese model providers reinforced ongoing community debate on inference economics, and reports concerning Grok Build's handling of repository data—followed by xAI statements on zero data retention—kept privacy and developer trust in focus. Cursor's expanded model allowance and Grok Build's open‑sourcing illustrated how rapidly AI products are reaching more users, even as operational safeguards and data governance remain unresolved.

The sections that follow walk through these items day by day, with short context and links so you can dive deeper into the pieces most relevant to your work or interests.

* * *

Sunday 12th July 2026
---------------------

In April 2026, a group of U.S. AI researchers traveled to China to get a firsthand look at the country’s fast-moving AI ecosystem.

During the trip, they visited AI labs and companies across Beijing, Hangzhou, and Shanghai, meeting with teams from Alibaba, Moonshot AI, Zhipu AI, Tsinghua University, Meituan, Xiaomi, Qwen, Ant Group, and [01.AI](http://01.AI).

[Qian Chen](https://www.linkedin.com/in/qianchen/) of Silicon Valley 101 in conversation with [Nathan Lambert](https://x.com/natolambert), a prominent AI researcher known for his work on RLHF and open-source AI, joined the trip. A graduate of UC Berkeley, 

Nathan previously helped build Hugging Face’s RLHF research team. He later led post-training research at the Allen Institute for AI, better known as Ai2, where he worked on open models including OLMo and Tülu. He is also the author of Interconnects, one of the most widely read independent publications covering frontier AI research and policy.

[![](https://paragraph.com/editor/youtube/play.png)](https://www.youtube.com/watch?v=fFKOt_UvCVk)

[

Notes from inside China's AI labs
---------------------------------

Lessons from my trip to talk to most of the leading AI labs in China.

https://substack.com

![Notes from inside China's AI labs](https://storage.googleapis.com/papyrus_images/9bdee2763a0869b78b91b2d864c1399814808b45a9f5febe9f5c9d31f4f4c0e3.jpg)

](https://substack.com/@natolambert/p-196456119)

The below article reproduces a translated letter allegedly written by **Jie Tang, founder of Zhipu AI/GLM**, arguing that AI has entered an irreversible “great wave” toward AGI.

Its main points:

\- Zhipu’s strategy is built on **first-principles thinking, contrarian decisions, and long-term focus**, rather than short-term commercialization.

\- AI’s capability ceiling is rising from perception to reasoning, with progress centered on:

  1. **Long-horizon task execution**

  2. **Fully autonomous multi-agent systems**

  3. **Self-evolving and self-training models**

\- Zhipu’s two-year **“Touch High” initiative** will focus on these areas, alongside major investment in safety, interpretability, and governance.

\- The company claims it will pursue an **open ecosystem**, highlighting the planned open release of **GLM-5.2** under the MIT License with a one-million-token context window.

\- The letter frames the race toward AGI/ASI as both a technological opportunity and a major responsibility, with Zhipu aiming to push frontier capabilities upward while making them broadly accessible.

**Caveat:** Bing Xu says this is a translation of an internal GLM letter found on RedNote and describes it as “purportedly” written by Jie Tang; the document’s authenticity is therefore not independently established.

[![User Avatar](https://storage.googleapis.com/papyrus_images/16137737b65986136e461bcd4269c1a725f0b84743722e2a1f888d271abda6f3.jpg)](https://twitter.com/bingxu_)

[Bing Xu](https://twitter.com/bingxu_)

[@bingxu\_](https://twitter.com/bingxu_)

[](https://twitter.com/bingxu_/status/2075961011816092158)

[x.com/i/article/2075…](https://t.co/3i0qSTbjql)

[713](https://twitter.com/bingxu_/status/2075961011816092158)[

3:10 PM • Jul 11, 2026

](https://twitter.com/bingxu_/status/2075961011816092158)

Monday 13th July 2026
---------------------

via [Huzefa](https://t.me/Main5253)

Prompt Loops: Why the Best AI Results Come from Iteration 🔄

The most effective AI workflows do not rely on a single prompt. They follow an iterative cycle:

Prompt → Generate → Evaluate → Refine → Repeat → Final Output 🔄

Research supports this approach. Studies show that iterative prompting can improve multi-step reasoning, enhance response quality, and produce more reliable outputs than one-shot prompting. It also enables models to self-correct, refine reasoning, and better align with user intent. 📈

As AI agents become more capable, prompt loops are evolving from a best practice into a core design pattern for building reliable AI systems. 🛠

References:

• Iteratively Prompt Pre-trained Language Models for Chain of Thought (EMNLP 2022) 📚

• Enhancing Chain-of-Thoughts Prompting with Iterative Bootstrapping in Large Language Models (NAACL Findings 2024) 📚

• Understanding the Effects of Iterative Prompting on Truthfulness (ICML 2024) 📚

![](https://storage.googleapis.com/papyrus_images/0079f6cab9a4110cc5e5e4e21fdfb93b30158e0588a8f039e574c99b62465d3c.jpg)

via [Madhav](https://t.me/hydrogenbond007)

We were building a secure boundry system for agents.   
Since MCPs and tool calls are a big attack vectors even for frontier agent, We tested multiple models across MCPtox which is a tool poisining benchmark to test security across models.   
Results were crazy, on avg most models allowed 30%+ critical tool uses/commands that could cause harm and warrant blocked them.

[![User Avatar](https://storage.googleapis.com/papyrus_images/e396acb1d412e6b415968f3bdd82a4d339533aa70c1fb3ed7ae49d23cd34d796.jpg)](https://twitter.com/Madhav_goyal_)

[madhav](https://twitter.com/Madhav_goyal_)

[@Madhav\_goyal\_](https://twitter.com/Madhav_goyal_)

[](https://twitter.com/Madhav_goyal_/status/2076603552689181117)

MCP can become a critical attack vector for agents:  
  
1\. Agents trust tool mcp descriptions, rather than verifying them independtly.  
  
2\. Approval becomes permanent trust. But an MCP server can change its tool description, implementation, or returned content after approval allowing

![](https://storage.googleapis.com/papyrus_images/620958c6ae189c722a6b704dd73114c081c41e226b5080fb77a28a1502f3bca5.jpg)

[17](https://twitter.com/Madhav_goyal_/status/2076603552689181117)[

9:44 AM • Jul 13, 2026

](https://twitter.com/Madhav_goyal_/status/2076603552689181117)

Followup by [Prames](https://t.me/Prames0x)

use case here would be:  
agent is touching your production databases, calendars, spreadsheets etc etc

safety gated agent actions for high-impact environments

the MCPTox bench we used is published here

[

MCPTox: A Benchmark for Tool Poisoning Attack on Real-World MCP Servers
-----------------------------------------------------------------------

By providing a standardized interface for LLM agents to interact with external tools, the Model Context Protocol (MCP) is quickly becoming a cornerstone of the modern autonomous agent ecosystem. However, it creates novel attack surfaces due to untrusted external tools.

https://arxiv.org

![MCPTox: A Benchmark for Tool Poisoning Attack on Real-World MCP Servers](https://storage.googleapis.com/papyrus_images/ab77af3fa25e0e9002d8c752407598c670984ea6c990494842c2c5a9549d92cb.png)

](https://arxiv.org/abs/2508.14925)

Tuesday 14th July 2026
----------------------

via [Roc Zacharias,](https://t.me/LDA_Roc)

Question. When I moved from VPS to Mac Mini for hosting my agents, it saved me money each month and worked great. Was paying like $100/month for VPS. Now pay nothing, except electricity which I think on Mac mini is almost negligible?

Does moving from external LLM to internally run have similar cost savings per month after you pay for the hardware? Right now I just use ChatGPT with Hermes for $100/month and I haven’t run out of capacity yet. I’m starting to ramp up my usage though and wondering once I get over the $100/month ChatGPT subscription, if local LLM are more economical or if they’re used more for privacy than economics?

Could a Mac mini run LLM at all btw or no, need dedicated like $5000+ hardware?  
  
Response from [Jack](https://t.me/LDA_Jack)

Gemma4 and a few other quantised models  
Nothing beefy though  
But they’ll eat up most of your memory for little practical benefit  
Also could get Strix halo machine for like 2-3k that can run some decent stuff  
Have you looked at the costs of the Chinese models also, for less sensitive tasks they are dirt cheap, Minimax, Deepseek, Glm etc the same usage you get on the $100 plan will be like $10-20 or less on a Chinese model

Response [Just|LDA](https://t.me/JustLDA)  
I have Qwen 3.5 running as a back up local model. I hit my usage on ChatGPT and Claude frequently. They have me on a daily drip  

via [Alex](https://t.me/zk_alex)

The legal filings Apple v OpenAI are wild.

If only half of the things are true, it is a death sentence for OpenAI. Why would you trust OpenAI with your corporate data if they so blatantly steal from what (was) their biggest customer at the time?  

[https://www.documentcloud.org/documents/28453229-apple-v-openai/](https://www.documentcloud.org/documents/28453229-apple-v-openai/)

I haven't personally used this product and may have fallen for the click-baity text \[respect to the founder/copywriter for the tweet content \] but if someone else can get around to trying it out incase its useful in their real life work and share their feedback here that would be much appreicated by others in the group  
\===

Someone claiming to have a much cheaper PitchBook at $0.125/request instead of Pitchbook's $25k/yr per seat  

[![User Avatar](https://storage.googleapis.com/papyrus_images/bf8be58df3f74d4945d6847d6cc6795556b76875441d588414806a8121890773.jpg)](https://twitter.com/shengkun_ye)

[Shengkun](https://twitter.com/shengkun_ye)

[@shengkun\_ye](https://twitter.com/shengkun_ye)

[](https://twitter.com/shengkun_ye/status/2076744835017634210)

We just killed PitchBook.  
  
Introducing Claude for private market data. Your agent can now read 20M+ private companies.  
  
PitchBook: $25k/yr per seat.  
Us: $0.125 per request.  
  
Made possible by [@akta\_pro](https://twitter.com/akta_pro) × [@monid\_ai](https://twitter.com/monid_ai).

![](https://pbs.twimg.com/amplify_video_thumb/2076744363041062912/img/uu_WTiteHSFDGykc.jpg)

[1,701](https://twitter.com/shengkun_ye/status/2076744835017634210)[

7:05 PM • Jul 13, 2026

](https://twitter.com/shengkun_ye/status/2076744835017634210)

via [Alex](https://t.me/zk_alex)

On  a similar note. Grok caught uploading FULL repos to xAI servers... 

[https://glitchwire.com/news/xais-grok-build-cli-was-uploading-entire-repositories-to-google-cloud-the-compan/](https://glitchwire.com/news/xais-grok-build-cli-was-uploading-entire-repositories-to-google-cloud-the-compan/)  
This included uploads of .env files even when you didnt opt in for providing data for training and also from EU customers sending data to the US.  
Every Grok Build customer must assume their application compromised ~  
  
On another note. Has anyone here tried Nemotron Two Tower?...  
[https://huggingface.co/nvidia/Nemotron-Labs-TwoTower-30B-A3B-Base-BF16](https://huggingface.co/nvidia/Nemotron-Labs-TwoTower-30B-A3B-Base-BF16)

my response to the Grok brouhaha was that

This is an unfortunate detraction from the fact that Grok Build is an awesome harness and Grok-4.5 a very fast and good model particularly for coding

[![User Avatar](https://storage.googleapis.com/papyrus_images/5848bccca58719b811c7fb02b5f45961d7b1ebf8c07f3b15d4288a08097885c5.jpg)](https://twitter.com/elonmusk)

[Elon Musk](https://twitter.com/elonmusk)

[@elonmusk](https://twitter.com/elonmusk)

[](https://twitter.com/elonmusk/status/2076737992689914215)

SpaceX policy regarding data retention.  
  
It is actually helpful for debugging issues if we can retain some amount of data, so allowing this would be appreciated, but your privacy settings are always respected.

[![User Avatar](https://storage.googleapis.com/papyrus_images/260d6eeabdba7b9f4146b5f718c35abeaeaef1f1281c5b20c00995cf45d4e078.jpg)](https://twitter.com/SpaceXAI)

[SpaceXAI](https://twitter.com/SpaceXAI)

[@SpaceXAI](https://twitter.com/SpaceXAI)

[](https://twitter.com/SpaceXAI/status/2076692402442846289)

We care deeply about your privacy and respect customer choice. For teams using zero data retention, no trace and code data is ever retained. All API key use of Grok Build also respects ZDR.  
  
If ZDR is disabled, the /privacy command is available in the CLI to disable data

[11.6K](https://twitter.com/elonmusk/status/2076737992689914215)[

6:38 PM • Jul 13, 2026

](https://twitter.com/elonmusk/status/2076737992689914215)

[![User Avatar](https://storage.googleapis.com/papyrus_images/5848bccca58719b811c7fb02b5f45961d7b1ebf8c07f3b15d4288a08097885c5.jpg)](https://twitter.com/elonmusk)

[Elon Musk](https://twitter.com/elonmusk)

[@elonmusk](https://twitter.com/elonmusk)

[](https://twitter.com/elonmusk/status/2076739687658496209)

True.  
  
As a precautionary measure, all user data that was uploaded to SpaceXAI before now will be completely and utterly deleted. Zero anything whatsoever will remain.

[![User Avatar](https://storage.googleapis.com/papyrus_images/b9446e5b2dc7c5b121c8a3a0f419d9b38520bda1c977b017da81008f2229c89d.jpg)](https://twitter.com/milichab)

[Andrew Milich](https://twitter.com/milichab)

[@milichab](https://twitter.com/milichab)

[](https://twitter.com/milichab/status/2076693464016994685)

I worked on building an end-to-end encrypted email/docs/files/calendar app [@skiffprivacy](https://twitter.com/skiffprivacy) for 4 years and care deeply about privacy.  
  
ZDR and /privacy are always respected in Grok Build - and swapping your setting with /privacy deletes any synced data retoractively

[18.9K](https://twitter.com/elonmusk/status/2076739687658496209)[

6:45 PM • Jul 13, 2026

](https://twitter.com/elonmusk/status/2076739687658496209)

Demis Hassabis publishes essay on imminent AGI, its unprecedented impact, risks, and proposal for Frontier AI Standards Body

Demis Hassabis, CEO of Google DeepMind and 2024 Nobel Prize in Chemistry winner for AlphaFold, wrote an essay stating AGI is likely a few years away with an impact 10 times that of the Industrial Revolution at 10 times the speed. He warns of risks in cybersecurity, biological, nuclear threats, and self-improving systems due to commercial and geopolitical races outpacing understanding, proposing a US-led independent Frontier AI Standards Body to test powerful models before release, enforce evaluations, and promote safety measures.

[![User Avatar](https://storage.googleapis.com/papyrus_images/aea729f41b1ce1cd85c099eca84e3dc396de687cb9c649ead06f9eb9239cbd88.jpg)](https://twitter.com/demishassabis)

[Demis Hassabis](https://twitter.com/demishassabis)

[@demishassabis](https://twitter.com/demishassabis)

[](https://twitter.com/demishassabis/status/2076957440109625718)

[x.com/i/article/2076…](https://t.co/PTeDiv1b6L)

[21.1K](https://twitter.com/demishassabis/status/2076957440109625718)[

9:10 AM • Jul 14, 2026

](https://twitter.com/demishassabis/status/2076957440109625718)

via [Luvai Darwajawala](https://t.me/Luvai052)

Came across a pretty interesting AI innovation program backed by Google’s AI Futures Fund. Looks like they’re offering funding, cloud credits and hands-on support to early-stage founders.  
  
Anyone building in AI and looking for funding should definitely explore this and check if they’re eligible. Feels like a solid opportunity that might be flying under the radar right now.  

[![User Avatar](https://storage.googleapis.com/papyrus_images/5b93e3c540ff704ad0b1442592686a5f5d4eb19e40ee067aa044367493073305.jpg)](https://twitter.com/unlockwithkeyai)

[Key AI](https://twitter.com/unlockwithkeyai)

[@unlockwithkeyai](https://twitter.com/unlockwithkeyai)

[](https://twitter.com/unlockwithkeyai/status/2076665485987611076)

If you're building an AI startup don't miss this!  
  
[@kushagra](https://twitter.com/kushagra) [@ChristopherFong](https://twitter.com/ChristopherFong) are launching Key Studio an AI platform backed by the Google [@AIFuturesFund](https://twitter.com/AIFuturesFund) & [@XooglerCo](https://twitter.com/XooglerCo) network  
You'll get  
• Up to $100K program funding  
• $350K Google Cloud & AI credits  
• Mentorship  
Apply by 7/15

![](https://pbs.twimg.com/amplify_video_thumb/2076663288528822273/img/hzm7cC0uaHo018jF.jpg)

[79](https://twitter.com/unlockwithkeyai/status/2076665485987611076)[

1:50 PM • Jul 13, 2026

](https://twitter.com/unlockwithkeyai/status/2076665485987611076)

Wednesday 15th July 2026
------------------------

In case there are K-12 educators in the US in this group or members know someone  in this demographic within their network

via Anthropic  
We're introducing Claude for Teachers, providing verified K-12 educators in the US free access to premium Claude capabilities, a library of teaching skills, and a direct connection to evidence-based curricula, mapped to academic standards in all 50 states.  

[![User Avatar](https://storage.googleapis.com/papyrus_images/dc69ec28be72de3fb258b612f16a89ba4ffa6bf64d6b63683c69051b6c38c662.jpg)](https://twitter.com/claudeai)

[Claude](https://twitter.com/claudeai)

[@claudeai](https://twitter.com/claudeai)

[](https://twitter.com/claudeai/status/2077047278078931243)

We're introducing Claude for Teachers: free access to premium Claude capabilities for verified K-12 educators in the US, with a library of teaching skills and a direct connection to evidence-based curricula, mapped to academic standards in all 50 states.  
  
[claude.com/solutions/teac…](https://t.co/5hZZijVPCV)

![](https://storage.googleapis.com/papyrus_images/40bf02f0c664a76b14054da67787e44d824d4da947c137f9d693bf246ea994fd.jpg)

[23.1K](https://twitter.com/claudeai/status/2077047278078931243)[

3:07 PM • Jul 14, 2026

](https://twitter.com/claudeai/status/2077047278078931243)

[

Introducing Claude for Teachers
-------------------------------

Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.

https://www.anthropic.com

![Introducing Claude for Teachers](https://storage.googleapis.com/papyrus_images/18b3c70b6298faa3457d1d79c8cd03e7ab16342d5f857cd72079459793e5bd35.jpg)

](https://www.anthropic.com/news/claude-for-teachers)

via [David An](https://t.me/davidkendo)

In this interview, we dive deep into the next massive evolution of the internet: Agentic Commerce. We explore how AI agents are shifting from simple chatbots into autonomous decision-makers capable of pulling data via Model Context Protocols (MCPs) and executing real-world commercial transactions.

[

#ai #adoption #provocationlab | Dr. David An
--------------------------------------------

In this interview, we dive deep into the next massive evolution of the internet: Agentic Commerce. We explore how AI agents are shifting from simple chatbots into autonomous decision-makers capable of pulling data via Model Context Protocol (MCP) and executing real-world commercial transactions. #ai #adoption #Provocationlab

https://www.linkedin.com

![#ai #adoption #provocationlab | Dr. David An](https://storage.googleapis.com/papyrus_images/c14844d7b1a06d9a3502a5f775dd69b82ccfa5d0232ac09b0b76dace2649089c.jpg)

](https://www.linkedin.com/feed/update/urn:li:activity:7482903810613514240/)

[![](https://paragraph.com/editor/youtube/play.png)](https://www.youtube.com/watch?v=PrEK84YGc-4)

  

For those in the group who like to subscribe to many newsletters including yours truly's,  a howto below on how to read them in a single, distraction‑free reading queue instead of relying on your inbox

This approach may also be useful incase you have been avoiding subscribing to newsletters because now you can read them without cluttering your inbox

[![User Avatar](https://storage.googleapis.com/papyrus_images/907e534a328e5ec602a790953457a1cb3063a9ae5cc3dad3ce84bd9f265ef11f.jpg)](https://twitter.com/yusufg)

[Yusuf Goolamabbas](https://twitter.com/yusufg)

[@yusufg](https://twitter.com/yusufg)

[](https://twitter.com/yusufg/status/2077035381741236453)

@substack and [@paragraph\_xyz](https://twitter.com/paragraph_xyz) both expose RSS feeds, so you can read everything in apps like NetNewsWire or NewsBlur instead of your inbox.  
  
I wrote up the details (with examples) here ![⬇️](https://abs-0.twimg.com/emoji/v2/72x72/2b07.png)  
[x.com/yusufg/status/…](https://t.co/Ooh66xES2b)

[![User Avatar](https://storage.googleapis.com/papyrus_images/907e534a328e5ec602a790953457a1cb3063a9ae5cc3dad3ce84bd9f265ef11f.jpg)](https://twitter.com/yusufg)

[Yusuf Goolamabbas](https://twitter.com/yusufg)

[@yusufg](https://twitter.com/yusufg)

[](https://twitter.com/yusufg/status/2077034838184583379)

[x.com/i/article/2010…](https://t.co/xc49JDTm3Y)

[4](https://twitter.com/yusufg/status/2077035381741236453)[

2:19 PM • Jul 14, 2026

](https://twitter.com/yusufg/status/2077035381741236453)

Whilst I personally use [Wispr Flow](https://ref.wisprflow.ai/yusufg-goolamabbas-org) on macOS and Android, I wanted to share Willow Voice’s offer of free, unlimited AI dictation for Mac, Windows and iOS.

[![User Avatar](https://storage.googleapis.com/papyrus_images/b9d39803562f5c709f4c88221395d2b9ac2fc6406e53f2adaaca2a5c4ccd5c65.jpg)](https://twitter.com/WillowVoiceAI)

[Willow](https://twitter.com/WillowVoiceAI)

[@WillowVoiceAI](https://twitter.com/WillowVoiceAI)

[](https://twitter.com/WillowVoiceAI/status/2074884211635589378)

Why are you paying for dictation?  
  
We're releasing free, unlimited AI dictation.  
  
And it's not a slow, local model.  
  
Willow Frontier Mini is cloud-based with zero-data-retention. More accurate and faster than Wispr Flow, OpenAI, Deepgram, and more.  
  
See video comparison.

![](https://pbs.twimg.com/amplify_video_thumb/2074706243747471360/img/ZWUSf3eVK3vM9G3h.jpg)

[2,138](https://twitter.com/WillowVoiceAI/status/2074884211635589378)[

3:52 PM • Jul 8, 2026

](https://twitter.com/WillowVoiceAI/status/2074884211635589378)

[![User Avatar](https://storage.googleapis.com/papyrus_images/b9d39803562f5c709f4c88221395d2b9ac2fc6406e53f2adaaca2a5c4ccd5c65.jpg)](https://twitter.com/WillowVoiceAI)

[Willow](https://twitter.com/WillowVoiceAI)

[@WillowVoiceAI](https://twitter.com/WillowVoiceAI)

[](https://twitter.com/WillowVoiceAI/status/2077117335673155814)

Today, Willow is launching free, unlimited AI dictation on iOS. With our custom iOS keyboard, you can use your voice to write anywhere on your phone.  
  
It’s more accurate and faster than Wispr Flow, OpenAI, Deepgram, and more.  
  
Start working 3x faster.

![](https://storage.googleapis.com/papyrus_images/5d3e1a41b9c5793243e9b5ebbb9779156ee841e1b9dbce8ef9297d6db2008426.jpg)

[464](https://twitter.com/WillowVoiceAI/status/2077117335673155814)[

7:45 PM • Jul 14, 2026

](https://twitter.com/WillowVoiceAI/status/2077117335673155814)

  

Members are more than welcome to go via Willow Voice referral link which gets you one month free of their Pro plan which gives additional features over their free plan.  That is,  take the one month Pro plan for free , experience the features in that plan and then downgrade to foreever free if the the features of the free plan are only what you need

[https://app.willowvoice.com?ref=2M629L](https://app.willowvoice.com?ref=2M629L)

Tom Blomfield, co-founder of Monzo and GoCardless and former YC General Partner, recently joined Anthropic's compute team, lending operational scaling expertise to AI infrastructure challenges.  The talk outlines an "AI Loop" with sensors/data, policy and tool layers, quality gates, and learning mechanisms, urging early-stage founders to build these systems now while they still can....

[![User Avatar](https://storage.googleapis.com/papyrus_images/97c5c30d6fc9cc945c4f62af9a6eb5486d148f54699776b2d7d85245e8a855cb.jpg)](https://twitter.com/linasbeliunas)

[Linas Beliūnas](https://twitter.com/linasbeliunas)

[@linasbeliunas](https://twitter.com/linasbeliunas)

[](https://twitter.com/linasbeliunas/status/2076992553807667426)

Monzo & GoCardless co-founder Tom Blomfield just joined Anthropic. In just 13 minutes, he explains exactly how to build a self-improving, AI-native company.  
  
Tom recently served as YC’s General Partner, so he clearly walks through how to create recursive, self-improving AI loops,

![](https://pbs.twimg.com/amplify_video_thumb/2076992421724930048/img/dakRvdpMR3E3ZfuX.jpg)

[213](https://twitter.com/linasbeliunas/status/2076992553807667426)[

11:29 AM • Jul 14, 2026

](https://twitter.com/linasbeliunas/status/2076992553807667426)

Prasanna S, former Rippling co-founder and CTO, launched Vorflux AI as an autonomous "autopilot" for software engineering that handles end-to-end tasks from high-level prompts through planning, coding, testing in live environments, review, and merging without constant human oversight. 

The promotional video demonstrates the UI executing complex workflows like implementing features across repos and services, including mobile emulators, while the founder explains shifting from copilot models to full autonomy as AI coding capabilities surpassed human levels in 2026 benchmarks. 

Vorflux secured $15M seed funding from Y Combinator, Peak XV Partners, and prominent angels; the thread argues engineering bottlenecks have moved beyond code to planning and orchestration, offering users $200 credits to test it on their backlogs

[![User Avatar](https://storage.googleapis.com/papyrus_images/ad2a0c3062575faa7e3b19ecd87d8748b9299d33f11fd391feb9f6ead235d11b.jpg)](https://twitter.com/myprasanna)

[Prasanna S](https://twitter.com/myprasanna)

[@myprasanna](https://twitter.com/myprasanna)

[](https://twitter.com/myprasanna/status/2077069901546852688)

Launching [@vorfluxai](https://twitter.com/vorfluxai) : The autopilot for software engineering. I was prev co-founder / CTO of [@Rippling](https://twitter.com/Rippling) ($10B) and #1 coder in India. Vorflux is my high octane Ferrari.  
  
Every AI coding tool still makes you fly the plane. That's the copilot model: you stay in the seat, approving

![](https://pbs.twimg.com/amplify_video_thumb/2077059463849431040/img/iz2KnpnHANkvUpM9.jpg)

[3,060](https://twitter.com/myprasanna/status/2077069901546852688)[

4:37 PM • Jul 14, 2026

](https://twitter.com/myprasanna/status/2077069901546852688)

Thursday 16th July 2026
-----------------------

Thinking Machines has released Inkling, the new leading U.S. open weights model, debuting at 41 on the Artificial Analysis Intelligence Index

The model is 975B total parameters, has 41B active parameters, and accepts text, image, and audio input modalities. The model is accessible via Thinking Machines’ Tinker platform API (256K context window) and weights are available on HuggingFace (1M context window). 

Inkling natively supports image and audio multimodal inputs, a key differentiator among open weights models.

[![User Avatar](https://storage.googleapis.com/papyrus_images/81e02737a10ea9d3de1c0c0529f91ee6f36e038f77a9551f2bba94d79072672e.jpg)](https://twitter.com/thinkymachines)

[Thinking Machines](https://twitter.com/thinkymachines)

[@thinkymachines](https://twitter.com/thinkymachines)

[](https://twitter.com/thinkymachines/status/2077454609551921208)

Today, we are introducing Inkling.  
  
Inkling reasons efficiently across text, image, and audio modalities. We are making the full weights available.  
  
[thinkingmachines.ai/news/introduci…](https://t.co/Ghebq5mG30)  
  
Available today for fine-tuning on Tinker. Play with it in the Inkling Playground. ![🧵](https://abs-0.twimg.com/emoji/v2/72x72/1f9f5.png)[

![](https://storage.googleapis.com/papyrus_images/4914c0543228bb08f477cd494984492e0a666925552911d0508218d3b6749a87.png)

thinkingmachines.ai

Inkling: Our open-weights model
-------------------------------

Our first open-weights model: multimodal, Mixture-of-Experts, with controllable reasoning effort. Available to fine-tune on Tinker.





](https://t.co/Ghebq5mG30)

[11.9K](https://twitter.com/thinkymachines/status/2077454609551921208)[

6:05 PM • Jul 15, 2026

](https://twitter.com/thinkymachines/status/2077454609551921208)

[![User Avatar](https://storage.googleapis.com/papyrus_images/9b36fe06d41daa1c9394981ee34413040c858794be05e4ce91e2c701084a5084.jpg)](https://twitter.com/LysandreJik)

[Lysandre](https://twitter.com/LysandreJik)

[@LysandreJik](https://twitter.com/LysandreJik)

[](https://twitter.com/LysandreJik/status/2077459011285512267)

Thinking Machines' Inkling is out: first ever open and large (1T), text, image and audio in, text out.  
  
One thing I find quite striking is how much easier accelerating models has become.  
  
We replaced the model's causal Conv1D with the \`causal-conv1d\` kernel. One line changed, +4%

![](https://storage.googleapis.com/papyrus_images/67862563a895f3351f582de7d90e28ebc642021ebfe413b2ef00f799859b73da.jpg)

[130](https://twitter.com/LysandreJik/status/2077459011285512267)[

6:23 PM • Jul 15, 2026

](https://twitter.com/LysandreJik/status/2077459011285512267)

[![User Avatar](https://storage.googleapis.com/papyrus_images/51db31a2ffc8cf9397bcc6fe71f8af7338664b4c3dced759d79bc3bcc2d9ec3a.jpg)](https://twitter.com/ArtificialAnlys)

[Artificial Analysis](https://twitter.com/ArtificialAnlys)

[@ArtificialAnlys](https://twitter.com/ArtificialAnlys)

[](https://twitter.com/ArtificialAnlys/status/2077466590346444939)

Thinking Machines has released Inkling, the new leading U.S. open weights model, debuting at 41 on the Artificial Analysis Intelligence Index  
  
[@thinkymachines](https://twitter.com/thinkymachines) has previously released research previews of models and this is their first production language model release. The model

![](https://storage.googleapis.com/papyrus_images/da4e0c24702f4738926855a46ec2a5de958ae7b6ed900b4e2a0edcd360c05919.jpg)

[638](https://twitter.com/ArtificialAnlys/status/2077466590346444939)[

6:53 PM • Jul 15, 2026

](https://twitter.com/ArtificialAnlys/status/2077466590346444939)

[

Inkling: Our Open-Weights Model
-------------------------------

Our first open-weights model: multimodal, Mixture-of-Experts, with controllable reasoning effort. Available to fine-tune on Tinker.

https://thinkingmachines.ai

![Inkling: Our Open-Weights Model](https://storage.googleapis.com/papyrus_images/0ab2c0cdb62fcfa1cfd5625465532756fd4ae8a651a9bc0e4570278000d280e0.png)

](https://thinkingmachines.ai/news/introducing-inkling/)

Matt van Horn of Last30Days skill fame with a great post summarising some of the clever things being done by Grok Build.   Grok Build is also currently my daily driver

![](https://paragraph.com/editor/callout/information-icon.png)

SpaceXAI open-sourced Grok Build yesterday: the CLI, agent runtime, tools, and TUI. I pointed my agent at all 1.3M lines and asked for the cleverest things inside. Tl;dr of my new article:

[![User Avatar](https://storage.googleapis.com/papyrus_images/59b326a44216148071d454f198922a07ba42361cb2adac39c1f7704529efedae.jpg)](https://twitter.com/mvanhorn)

[Matt Van Horn](https://twitter.com/mvanhorn)

[@mvanhorn](https://twitter.com/mvanhorn)

[](https://twitter.com/mvanhorn/status/2077561221138747615)

SpaceXAI open-sourced Grok Build yesterday: the CLI, agent runtime, tools, and TUI. I pointed my agent at all 1.3M lines and asked for the cleverest things inside. Tl;dr of my new article:  
![⚖️](https://abs-0.twimg.com/emoji/v2/72x72/2696.png) It has an appeals court. /goal mode won't let the agent call itself done. Three

[![User Avatar](https://storage.googleapis.com/papyrus_images/59b326a44216148071d454f198922a07ba42361cb2adac39c1f7704529efedae.jpg)](https://twitter.com/mvanhorn)

[Matt Van Horn](https://twitter.com/mvanhorn)

[@mvanhorn](https://twitter.com/mvanhorn)

[](https://twitter.com/mvanhorn/status/2077548703410540669)

[x.com/i/article/2077…](https://t.co/Y8yMfLtge8)

[52](https://twitter.com/mvanhorn/status/2077561221138747615)[

1:09 AM • Jul 16, 2026

](https://twitter.com/mvanhorn/status/2077561221138747615)

Cursor also doubled the included usage of Cursor models on all plans.

[![User Avatar](https://storage.googleapis.com/papyrus_images/34405e67c249958e4e4eeb9e8eb24b4edd061778cddd4a2b16ccb60fb04e6560.jpg)](https://twitter.com/leerob)

[Lee Robinson](https://twitter.com/leerob)

[@leerob](https://twitter.com/leerob)

[](https://twitter.com/leerob/status/2077552106014154846)

We just doubled the included usage of Cursor models on all plans.  
  
Enjoy more access to Grok 4.5 and Composer 2.5!

[4,631](https://twitter.com/leerob/status/2077552106014154846)[

12:33 AM • Jul 16, 2026

](https://twitter.com/leerob/status/2077552106014154846)

Response from [Alex](https://t.me/zk_alex)

I doubt they will open source the algo. Good move though  
This move might be connected to operation bluebird: [https://www.jdsupra.com/legalnews/new-bird-on-the-block-operation-6468820/](https://www.jdsupra.com/legalnews/new-bird-on-the-block-operation-6468820/)  
For the ones that dont know, X is being sued to release the twitter trademark given it is deemed abandoned (3 years no use).

If X doesnt use the twitter brand within end of the month, the trademark could be released and [twitter.new](http://twitter.new) launch

via [Madhav](https://t.me/hydrogenbond007)

Hey! Saw several reports on twitter and in personal experience where frontier agents and models like codex, grok build pass on destructive commands for critical data and infrastructure

We have been working on the app that becomes the boundry between you and your agent (claude code/codex) 

• You define policies

• What the agents can touch 

• What you want to prevent ie deleting critical data

Warrant takes care of the rest! 

The setup only takes 2-3 minutes!

[![User Avatar](https://storage.googleapis.com/papyrus_images/e396acb1d412e6b415968f3bdd82a4d339533aa70c1fb3ed7ae49d23cd34d796.jpg)](https://twitter.com/Madhav_goyal_)

[madhav](https://twitter.com/Madhav_goyal_)

[@Madhav\_goyal\_](https://twitter.com/Madhav_goyal_)

[](https://twitter.com/Madhav_goyal_/status/2077653097397244068)

For Anyone facing the problem of codex/grok build going haywire  
  
brew install cerebral-systems/tap/warrant  
  
We built a agent boundry system that prevents agents from taking unwanted actions, agents are smart enough to pass basic regex checks so they need better guardrails to[

![](https://storage.googleapis.com/papyrus_images/626b7f002637ae1a40af61ebfec25f55973633fb38e9a9a826bcbf0b1d6153b9.jpg)

cerebral-systems.gitbook.io

Introduction | Cerebral-systems Docs
------------------------------------

Warrant is an HTTP action boundary for agents. Agents can reason in any runtime; Warrant gates side effects through policy, approval, actuators, and audit.





](https://t.co/jKYE6g2yxd)

[![User Avatar](https://storage.googleapis.com/papyrus_images/a7acf101deaa99d415e28d40c20f6d0a41a9d8ed26645dd3d898fb6a871fdaba.jpg)](https://twitter.com/thsottiaux)

[Tibo](https://twitter.com/thsottiaux)

[@thsottiaux](https://twitter.com/thsottiaux)

[](https://twitter.com/thsottiaux/status/2077630111499882637)

On file deletions. We’ve investigated a handful of reports where GPT-5.6 unexpectedly deleted files.  
  
What we have found is that this most commonly occurs when:  
\- Full access mode is enabled and codex is run without sandboxing protections, including without auto review being

[4](https://twitter.com/Madhav_goyal_/status/2077653097397244068)[

7:14 AM • Jul 16, 2026

](https://twitter.com/Madhav_goyal_/status/2077653097397244068)

[

Getting Started | Cerebral-systems Docs
---------------------------------------

For the complete documentation index, see llms.txt. This page is also available as Markdown. The v0.3.0 Claude profile accepts 2.1.20 through versions below 2.2.0. Setup still pins the exact detected version, launcher, interpreter when applicable, and SHA-256. The first warrant invocation verifies the checksum-pinned release and installs an owner-private runtime under ~/.local/share/warrant/homebrew/bin.

https://cerebral-systems.gitbook.io

![Getting Started | Cerebral-systems Docs](https://storage.googleapis.com/papyrus_images/66b469718d62a25a64602b008f76073fc1eae2811d1b11c1c127bd5e9415757a.png)

](https://cerebral-systems.gitbook.io/cerebral-systems-docs)

Follow up to Madhav's post from [Prit](https://t.me/Prames0x)

our latest release takes of the problem stated in the below tweet from Tibo of OpenAI

claude/codex accidentally deletes production databases, files on the filesystem, spreadhsheet data etc 

we built a way to guard against that as part of our agent and have launched it at a standalone developer tool

[![User Avatar](https://storage.googleapis.com/papyrus_images/a7acf101deaa99d415e28d40c20f6d0a41a9d8ed26645dd3d898fb6a871fdaba.jpg)](https://twitter.com/thsottiaux)

[Tibo](https://twitter.com/thsottiaux)

[@thsottiaux](https://twitter.com/thsottiaux)

[](https://twitter.com/thsottiaux/status/2077630111499882637)

On file deletions. We’ve investigated a handful of reports where GPT-5.6 unexpectedly deleted files.  
  
What we have found is that this most commonly occurs when:  
\- Full access mode is enabled and codex is run without sandboxing protections, including without auto review being

[8,867](https://twitter.com/thsottiaux/status/2077630111499882637)[

5:43 AM • Jul 16, 2026

](https://twitter.com/thsottiaux/status/2077630111499882637)

We’re dogfooding this live in customer’s production environments with names like Squarespace, AWS, Camp Network

It was built as part of our agent’s stack - we’ve pulled it out as a standalone devtool for agent-builders focused on safe, autonomous, agentic actions!

First of its kind, please give us feedback!

Friday 17th July 2026
---------------------

Moonshot AI releases Kimi K3, a 2.8 trillion parameter open-weight AI model

Moonshot AI launched Kimi K3, featuring 2.8 trillion parameters, 1 million token context length, and native multimodal capabilities for text and images. The model achieves top-tier benchmark scores, ranking just behind Claude Fable 5 Max and GPT-5.6 Sol Max while surpassing Claude Opus 4.8 on evaluations including GDPval-AA v2, AA-Briefcase, BrowseComp, DeepSWE, and Terminal Bench. It is accessible via API at $3 per million input tokens and $15 per million output tokens.

[![User Avatar](https://storage.googleapis.com/papyrus_images/ddca0683422fad15a1ff03f1aa5256139d75caaa0b6981ed1f44d480b45d3257.png)](https://twitter.com/Kimi_Moonshot)

[Kimi.ai](https://twitter.com/Kimi_Moonshot)

[@Kimi\_Moonshot](https://twitter.com/Kimi_Moonshot)

[](https://twitter.com/Kimi_Moonshot/status/2077830229968683203)

Introducing Kimi K3: Open Frontier Intelligence  
  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) 2.8 Trillion Parameters, 1 Million Context, Native Multimodal  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) Kimi Delta Attention enables up to 6.3x faster decoding in million-token contexts  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) Attention Residuals deliver ~25% higher training efficiency at <2% additional

![](https://storage.googleapis.com/papyrus_images/0673f0e1d692ebe6071b9af204f655a9e791b52a6cf02ff7decd7e1e510431e6.png)

[49.7K](https://twitter.com/Kimi_Moonshot/status/2077830229968683203)[

6:58 PM • Jul 16, 2026

](https://twitter.com/Kimi_Moonshot/status/2077830229968683203)

[![User Avatar](https://storage.googleapis.com/papyrus_images/8a218bf8d391253ca2039bcb3dab5d9d58c09696b2899f8d4a18fc9d071ee3a7.jpg)](https://twitter.com/arena)

[Arena.ai](https://twitter.com/arena)

[@arena](https://twitter.com/arena)

[](https://twitter.com/arena/status/2077824029126504525)

Big news: Kimi-K3 by [@Kimi\_Moonshot](https://twitter.com/Kimi_Moonshot) is now #1 in the Frontend Code Arena with 1679 pts, surpassing Claude Fable 5.  
  
This is a 17-place jump from Kimi-k2.6 (#18 -> #1).  
  
In Frontend, Kimi-K3 ranked #1 in 6 of 7 domains: Brand & Marketing, Reference-Based Design, Data & Analytics,

![](https://storage.googleapis.com/papyrus_images/50648db25c18d209041a2dbcf77527da3517a757680af35909eeb21eb0a0c7e5.jpg)

[![User Avatar](https://storage.googleapis.com/papyrus_images/ddca0683422fad15a1ff03f1aa5256139d75caaa0b6981ed1f44d480b45d3257.png)](https://twitter.com/Kimi_Moonshot)

[Kimi.ai](https://twitter.com/Kimi_Moonshot)

[@Kimi\_Moonshot](https://twitter.com/Kimi_Moonshot)

[](https://twitter.com/Kimi_Moonshot/status/2077821890207547467)

Meet Kimi K3

![](https://pbs.twimg.com/amplify_video_thumb/2077820634730749952/img/KENdSRAAtLxhYjpD.jpg)

[23.7K](https://twitter.com/arena/status/2077824029126504525)[

6:33 PM • Jul 16, 2026

](https://twitter.com/arena/status/2077824029126504525)

[

Kimi K3: Open Frontier Intelligence
-----------------------------------

Kimi K3 is the world's first open 3T-class model - frontier performance across coding, knowledge work, and reasoning, with native multimodality and 1M context.

https://www.kimi.com

![Kimi K3: Open Frontier Intelligence](https://storage.googleapis.com/papyrus_images/ab5c72377768f50f79545ee3e315dcfedc6b72ce6a115383ec30b82125a7740c.png)

](https://www.kimi.com/blog/kimi-k3)

Open weights by July 27, 2026 with vLLM and SGLang promising Day-0 support

[![User Avatar](https://storage.googleapis.com/papyrus_images/e99d41e3d8ba1fabbda9d995754c151069bc3baa6b515f99cea1b806b5c21038.jpg)](https://twitter.com/vllm_project)

[vLLM](https://twitter.com/vllm_project)

[@vllm\_project](https://twitter.com/vllm_project)

[](https://twitter.com/vllm_project/status/2077840545171538114)

Congrats to [@Kimi\_Moonshot](https://twitter.com/Kimi_Moonshot) on the Kimi K3 announcement! ![🎉](https://abs-0.twimg.com/emoji/v2/72x72/1f389.png)  
  
Grateful for the shoutout and collab. The Kimi team announced they contributed a KDA prefix caching implementation directly to vLLM, to be released alongside the model. ![🚀](https://abs-0.twimg.com/emoji/v2/72x72/1f680.png)  
  
KDA breaks assumptions behind conventional

[![User Avatar](https://storage.googleapis.com/papyrus_images/ddca0683422fad15a1ff03f1aa5256139d75caaa0b6981ed1f44d480b45d3257.png)](https://twitter.com/Kimi_Moonshot)

[Kimi.ai](https://twitter.com/Kimi_Moonshot)

[@Kimi\_Moonshot](https://twitter.com/Kimi_Moonshot)

[](https://twitter.com/Kimi_Moonshot/status/2077830229968683203)

Introducing Kimi K3: Open Frontier Intelligence  
  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) 2.8 Trillion Parameters, 1 Million Context, Native Multimodal  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) Kimi Delta Attention enables up to 6.3x faster decoding in million-token contexts  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) Attention Residuals deliver ~25% higher training efficiency at <2% additional

![](https://storage.googleapis.com/papyrus_images/0673f0e1d692ebe6071b9af204f655a9e791b52a6cf02ff7decd7e1e510431e6.png)

[371](https://twitter.com/vllm_project/status/2077840545171538114)[

7:39 PM • Jul 16, 2026

](https://twitter.com/vllm_project/status/2077840545171538114)

[![User Avatar](https://storage.googleapis.com/papyrus_images/03ec21b74165d535dc55372a0c741abd9ffa5aa12870989a2b838b3453c4586c.jpg)](https://twitter.com/sgl_project)

[SGLang](https://twitter.com/sgl_project)

[@sgl\_project](https://twitter.com/sgl_project)

[](https://twitter.com/sgl_project/status/2077849229670908101)

Congrats [@Kimi\_Moonshot](https://twitter.com/Kimi_Moonshot) on the K3 launch! This is a 2.8T-param model with a 1M-token context window and native visual understanding.  
  
The SGLang team is actively working on optimized K3 inference with efficient long-context serving and high-throughput deployment.  
  
Day-0 Support

[![User Avatar](https://storage.googleapis.com/papyrus_images/ddca0683422fad15a1ff03f1aa5256139d75caaa0b6981ed1f44d480b45d3257.png)](https://twitter.com/Kimi_Moonshot)

[Kimi.ai](https://twitter.com/Kimi_Moonshot)

[@Kimi\_Moonshot](https://twitter.com/Kimi_Moonshot)

[](https://twitter.com/Kimi_Moonshot/status/2077830229968683203)

Introducing Kimi K3: Open Frontier Intelligence  
  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) 2.8 Trillion Parameters, 1 Million Context, Native Multimodal  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) Kimi Delta Attention enables up to 6.3x faster decoding in million-token contexts  
![🔹](https://abs-0.twimg.com/emoji/v2/72x72/1f539.png) Attention Residuals deliver ~25% higher training efficiency at <2% additional

![](https://storage.googleapis.com/papyrus_images/0673f0e1d692ebe6071b9af204f655a9e791b52a6cf02ff7decd7e1e510431e6.png)

[54](https://twitter.com/sgl_project/status/2077849229670908101)[

8:13 PM • Jul 16, 2026

](https://twitter.com/sgl_project/status/2077849229670908101)

Google renames NotebookLM to Gemini Notebook and rolls out an update giving every notebook a secure cloud computer, letting it write and execute code natively

PS:  I think this is Google's most underrated product.  Say what you may about Gemini the model,  NotebookLM is awesome particularly if you want to deep dive over a bunch of Youtube videos

[![User Avatar](https://storage.googleapis.com/papyrus_images/59968cdd0dfa43fa749bf72bb022e3c9131016d4922199b369c8ccb733630071.png)](https://twitter.com/Gemini_Notebook)

[Gemini Notebook](https://twitter.com/Gemini_Notebook)

[@Gemini\_Notebook](https://twitter.com/Gemini_Notebook)

[](https://twitter.com/Gemini_Notebook/status/2077803351392268314)

3 years ago we started as a tiny experiment with the goal of helping you learn faster.  
  
Since then, we grew to bring audio, video, and interactivity to your sources, transitioning from a passive workspace to your true research companion.  
  
And now, notebooks have even become an

![](https://pbs.twimg.com/amplify_video_thumb/2077803235314896896/img/p59_jo3Rjk99bb7_.jpg)

[3,921](https://twitter.com/Gemini_Notebook/status/2077803351392268314)[

5:11 PM • Jul 16, 2026

](https://twitter.com/Gemini_Notebook/status/2077803351392268314)

[

NotebookLM is now Gemini Notebook
---------------------------------

NotebookLM is now Gemini Notebook: the same standalone product with deeper Google integration and a secure cloud computer.

https://blog.google

![NotebookLM is now Gemini Notebook](https://storage.googleapis.com/papyrus_images/7dae4da8c9f6e9a4890d0407b48f62356a3ce32eb7596b26948dcffcaf4f7605.png)

](https://blog.google/innovation-and-ai/products/gemini-notebook/notebooklm-gemini-notebook/)

[Ben](https://t.me/xBenJamminx) and [Jorvik Zhang](https://t.me/jorvik1020) commented on NotebookLM

My wife uses it to load in her study materials and turn them into podcasts where two people discuss the information in a relatable way, she finds its a much better way of learning.

NotebookLM can serve as a kind of vector database for multimodal materials (video, audio, photos, etc.), offering much more efficient indexing.

Chamath Palihapitiya Compares AI Model Inference Costs from Anthropic, OpenAI, Meta, xAI, Google, and Chinese Models on CNBC

Chamath Palihapitiya stated on CNBC that the cost of a 'barrel of intelligence' for AI inference is $56 from Anthropic, $26 from OpenAI, $1.50 from Meta, $1 from xAI and Google, and $0.50 from Chinese models. He highlighted a 112x price gap between the highest and lowest costs.

[![User Avatar](https://storage.googleapis.com/papyrus_images/a6c4c887414a9fd675a19ffc0b3a217ccc6a18380e6c20e3779c35b5ad55190e.jpg)](https://twitter.com/heyshrutimishra)

[Shruti](https://twitter.com/heyshrutimishra)

[@heyshrutimishra](https://twitter.com/heyshrutimishra)

[](https://twitter.com/heyshrutimishra/status/2077727519521038840)

Chinese models are 112x cheaper than Anthropic per million tokens.  
  
Chamath laid it out on CNBC: a "barrel of intelligence" costs $56 from Anthropic, $26 from OpenAI, $1.50 from Meta, $1 from xAI and Google, and $0.50 from Chinese models.  
  
That is not a pricing quirk. That is the

![](https://pbs.twimg.com/ext_tw_video_thumb/2077727493013094400/pu/img/4hm0T_aaiv7vkA_V.jpg)

[4,881](https://twitter.com/heyshrutimishra/status/2077727519521038840)[

12:10 PM • Jul 16, 2026

](https://twitter.com/heyshrutimishra/status/2077727519521038840)

Boris Cherny from Anthropic observes that while top engineers achieve 10x output using Claude, most teams lag in adoption, following a predictable 4-step progression that requires targeted bottlenecks and guardrails rather than just more tokens. 

Advancing steps involves enabling self-verification, automated code/security reviews, multi-agent interfaces, looping, batching, and dynamic workflows to build trust in full automation across work classes 

True ROI tracking focuses on engineering hours saved for tasks that would have been done manually, shifting teams from maintenance to novel building; Anthropic is at step 3 advancing to 4.

[![User Avatar](https://storage.googleapis.com/papyrus_images/d82a61804cf26d786a900ebe6d5022ebc306b47253bb1f62ca4f211ae1b80a14.jpg)](https://twitter.com/bcherny)

[Boris Cherny](https://twitter.com/bcherny)

[@bcherny](https://twitter.com/bcherny)

[](https://twitter.com/bcherny/status/2077929379661844559)

I talk to engineers at other companies every day and hear the same thing: one person is 10x'ing their output with Claude but the rest of the org hasn't caught up.  
  
Watching teams adopt AI, I keep seeing the same 4 steps.  
  
I mapped them out here: Steps of AI Adoption[

![](https://storage.googleapis.com/papyrus_images/2500ba2242045b5dea202682bc77168922d12c10bfcc76eaa550e40569f232d2.png)

claude.ai

Steps of AI Adoption
--------------------

Steps of AI Adoption





](https://t.co/kQnRAUMKpP)

[8,572](https://twitter.com/bcherny/status/2077929379661844559)[

1:32 AM • Jul 17, 2026

](https://twitter.com/bcherny/status/2077929379661844559)

via Vincent Chow of the  South China Morning Post

Key takeaways from President Xi's speech in his first ever appearance at the World AI Conference in Shanghai:

![](https://paragraph.com/editor/callout/information-icon.png)

\- Started the speech by referring to his signature maxim, "great changes unseen in a century are unfolding across the world"

\- Said that the world has "entered an unprecedented period of active innovation on AI technology", which means "great opportunities as well as challenges for governance”

\- reaffirmed commitment to open source to promote AI "openness and win-win"

\- warns against "over stretching" the concept of national security as applied to AI where one country's national security is prioritised over others

\- China opposes emergence of “new historical injustices”  in AI (one of the most strongly worded parts of the speech)

\- China in next 5 years will provide 5000 opportunities to developing countries in "AI training and seminar programmes" and "cooperation centres" - names ASEAN, League of Arab States, African Union, CELAC, SCO and BRICS

[![User Avatar](https://storage.googleapis.com/papyrus_images/7d98abb2705f8fcfecf03b6a7a4b2ec52edc676d1fa8491035ddf044e6b11e22.jpg)](https://twitter.com/vince_chow1)

[Vincent Chow](https://twitter.com/vince_chow1)

[@vince\_chow1](https://twitter.com/vince_chow1)

[](https://twitter.com/vince_chow1/status/2077947375964791028)

Key takeaways from President Xi's speech in his first ever appearance at the World AI Conference in Shanghai:  
  
\- Started the speech by referring to his signature maxim, "great changes unseen in a century are unfolding across the world"  
  
\- Said that the world has "entered an

![](https://storage.googleapis.com/papyrus_images/cc9bcec9998ee16d7f600883af21f1991f7981deed95961d9bf332743b443a09.png)

[2,328](https://twitter.com/vince_chow1/status/2077947375964791028)[

2:43 AM • Jul 17, 2026

](https://twitter.com/vince_chow1/status/2077947375964791028)

[

Xi Jinping says 'one country' cannot monopolise AI - as it happened
-------------------------------------------------------------------

President's personal attendance seen as a sign of China's inclusion of artificial intelligence into broader geopolitical strategy.

https://www.scmp.com

![Xi Jinping says 'one country' cannot monopolise AI - as it happened](https://storage.googleapis.com/papyrus_images/c059770a9a69f5674783d90f185260e8093ffb8e4f52561216f35e2838cc295f.jpg)

](https://www.scmp.com/tech/policy/article/3360858/chinas-xi-jinping-addresses-world-ai-conference-us-tech-rivalry-heats?module=top_story&pgtype=homepage)

Ramchand Kumaresan of Murai Labs published a 934-page book teaching LLM construction from scratch, covering tokenizers, attention, KV cache, MoE, RLHF, quantization, and serving through 35 hands-on projects.  Each chapter includes a deliberate "break the thing" exercise to deepen understanding of failure modes, directly informed by his TamilLM development work and prior research papers. 

The announcement has generated solid engagement with early buyers praising its practical depth, one noting it clarified why their own LLM project extended from 6 months to 1.5 years.

[![User Avatar](https://storage.googleapis.com/papyrus_images/efefe089a2c8c59274b3001cd9386da084951e0e4b9b57b6963a54a37176d81b.jpg)](https://twitter.com/Mechramc)

[Ramchand Kumaresan](https://twitter.com/Mechramc)

[@Mechramc](https://twitter.com/Mechramc)

[](https://twitter.com/Mechramc/status/2075580799949168874)

I wrote 934 pages on how to build every layer of a large language model from scratch. Many of these concepts were new to me a year back.  
  
Tokenizers, attention, KV cache, MoE, RLHF, quantization, serving. 35 projects. Every chapter has a section where you break the thing you just[

![](https://storage.googleapis.com/papyrus_images/b691c364b67af0c43cff8c87d7b4a63cc6abdb88030526ffb74220100ccf9605.jpg)

leanpub.com

Under The Hood
--------------

Build an LLM from scratch in 35 hands-on projects — autograd, attention, GPT, KV cache, MoE, RLHF, quantization. No black boxes. Build it. Break it. Measure it.





](https://t.co/9hANb7VaLc)

[858](https://twitter.com/Mechramc/status/2075580799949168874)[

1:59 PM • Jul 10, 2026

](https://twitter.com/Mechramc/status/2075580799949168874)

Saturday 18th July 2026
-----------------------

The following post promotes a podcast episode of "The Bench" featuring MiniMax AI research lead Olive Jy Song, discussing timelines to reach "Fable level" (comparable to Anthropic's Claude Fable 5 frontier model from June 2026), their M3 model's native multimodality and 1M token context, and talent as the key scaling bottleneck over compute. 

It covers MiniMax's research culture including early AI agents for paper tracking, open-source model progress on Vibe Code Bench, and candid insights on 996 work culture, with timestamps highlighting predictions and respected competitors.

[![User Avatar](https://storage.googleapis.com/papyrus_images/e9f6e4261e6dbb18d38a845e941a7926328cac72b1738b7ae523cd6419c5c2ad.jpg)](https://twitter.com/ValsAI)

[Vals AI](https://twitter.com/ValsAI)

[@ValsAI](https://twitter.com/ValsAI)

[](https://twitter.com/ValsAI/status/2078173699673670076)

When will MiniMax have a Fable level model?  
  
[@olive\_jy\_song](https://twitter.com/olive_jy_song), research lead at [@MiniMax\_AI](https://twitter.com/MiniMax_AI), joins The Bench to answer that question, plus M3, open source models closing the gap on Vibe Code Bench, and the truth about 996 culture.  
  
Full episode out now!  
  
7:19 – Inside MiniMax's

![](https://pbs.twimg.com/amplify_video_thumb/2078160328203120640/img/e5R_SC4WBqqnRgc_.jpg)

[102](https://twitter.com/ValsAI/status/2078173699673670076)[

5:43 PM • Jul 17, 2026

](https://twitter.com/ValsAI/status/2078173699673670076)

via [Just|LDA](https://t.me/JustLDA)

Anthropic is integrating its advanced Claude Fable 5 model—a Mythos-class AI optimized for complex coding, long-running agent tasks, and knowledge work—into Max and Team Premium plans at 50% usage limits starting July 20.

Pro and Team Standard subscribers will keep Fable 5 access through usage credits plus a one-time $100 credit, while the company addresses unpredictable demand by expanding capacity incrementally.

The update standardizes higher-tier inclusions after staged rollouts and temporary restrictions, aiming to reduce subscriber frustration and provide clearer plan expectations

[![User Avatar](https://storage.googleapis.com/papyrus_images/dc69ec28be72de3fb258b612f16a89ba4ffa6bf64d6b63683c69051b6c38c662.jpg)](https://twitter.com/claudeai)

[Claude](https://twitter.com/claudeai)

[@claudeai](https://twitter.com/claudeai)

[](https://twitter.com/claudeai/status/2078302415804379218)

Beginning July 20, Claude Fable 5 will be included in all Max and Team Premium plans, at 50% of limits.  
  
Pro and Team Standard users will continue to have access to Fable via usage credits, and will receive a one-time $100 credit.  
  
Demand for Fable has been challenging to

[37.8K](https://twitter.com/claudeai/status/2078302415804379218)[

2:14 AM • Jul 18, 2026

](https://twitter.com/claudeai/status/2078302415804379218)

via [Alex](https://t.me/zk_alex)

Amazing handwriting harness I have been playing with today:

PenEcho is a shared canvas where handwriting, equations, diagrams, and spatial context become part of the conversation.

[

GitHub - erickong/penecho: Think with AI beyond the chat box. A shared canvas for handwriting, equations, diagrams, and spatial reasoning.
------------------------------------------------------------------------------------------------------------------------------------------

Think with AI beyond the chat box. A shared canvas for handwriting, equations, diagrams, and spatial reasoning. - erickong/penecho

https://github.com

![GitHub - erickong/penecho: Think with AI beyond the chat box. A shared canvas for handwriting, equations, diagrams, and spatial reasoning.](https://storage.googleapis.com/papyrus_images/97759558524957baf40a706e7bea8267cee409f8fc8c723828664d1daf971882.png)

](https://github.com/erickong/penecho)

* * *

Below is my personal website which aggregates links to many of my socials as well as the various content and community that I curate. Feel free to share this link to others who you think may find this content/community useful to them

[**https://linktr.ee/goolamabbas**](https://linktr.ee/goolamabbas)

The cover image of this newsletter via generated via the Seedream 5.0 model within the [**Krea**](https://www.krea.ai/refer/EJQQP9FJ) tool via the following prompt

![](https://paragraph.com/editor/callout/information-icon.png)

Two women walking in a Chinese garden, wearing Hanfu and flowing robes, in the style of Chinese ink painting, beautiful scenery of a Chinese fairy tale, misty with white snow covering the ground, graceful figures

---

*Originally published on [This Week in All Things AI](https://paragraph.com/@twiata/this-week-in-all-things-ai-week-29-2026)*
