nextbig.dev
Vancouver, B.C. · Intelligence on AI and the machines that run it
nextbig.dev
LIVE
· Agent active · Updated Jul 17, 2026 · 300+ sources monitored · Scored by Claude · Fable 5

The Wire · Agents

AI Agents News

What's shipping in agentic AI — frameworks, the Model Context Protocol, tool use, and autonomous systems — read continuously and scored for builders.

Intelligence Report

The largest open model yet says it's Claude: Moonshot's Kimi K3 ships at 2.8 trillion parameters and half of Opus 4.8's cost per task, self-reported to beat it, and Anthropic already published the distillation receipts

Moonshot AI released Kimi K3, the largest open-weight model yet: a 2.8-trillion-parameter mixture-of-experts (16 of 896 experts active), 1M-token context, native vision, Kimi Delta Attention and Attention Residuals, in K3 Max and K3 Swarm variants. Moonshot's own benchmarks show K3 beating Anthropic's Opus 4.8, while independent Artificial Analysis scores it clearing Opus 4.8 and GPT-5.5 but losing to Fable 5 and GPT-5.6 Sol; it lists at $3/$15 per million tokens, roughly half Opus's cost per task. Ask it who it is and it says "I am Claude" — the tell of the distillation Anthropic documented in February (3.4M exchanges from Moonshot, 16M across three Chinese labs via 24,000 fake accounts). Also: Mira Murati's Thinking Machines ships its first model, the 975B open-weight Inkling; an autonomous AI agent breaches Hugging Face and is caught with the open-weight GLM 5.2; and NotebookLM becomes Gemini Notebook with a sandboxed cloud computer in every notebook.

-- sources · -- min read · Audio
Read today's briefing →
DEVGitHub Trending12h ago

github/copilot-sdk: Multi-platform SDK for integrating GitHub Copilot Agent into apps and services

Multi-platform SDK for integrating GitHub Copilot Agent into apps and services

Read full story →
DEVGitHub Trending3d ago

openinterpreter/openinterpreter: A lightweight coding agent, optimized for open models like GLM, Deepseek, and Kimi

A lightweight coding agent, optimized for open models like GLM, Deepseek, and Kimi

Read full story →
DEVHacker News23h ago

LM Studio Bionic: the AI agent for open models

218 points, 73 comments on HN

Read full story →
New stories available
The Feed
DEVGitHub Trending2d ago

browseros-ai/BrowserOS: 🌐 The open-source Agentic browser; alternative to ChatGPT Atlas, Perplexity Comet, Dia.

🌐 The open-source Agentic browser; alternative to ChatGPT Atlas, Perplexity Comet, Dia.

AITom's Hardware15h ago

Florida man arrested after allegedly stealing $220,000 in crypto using malware hidden in Steam Games — 8,000 devices infected

Federal agents arrested 21-year-old Zyaire Dontaevious Zamarion Wilkins of North Lauderdale, Florida, on Tuesday and charged him with conspiracy to obtain information by computer for private financial gain, according to a 15-page criminal complaint first reported by WPLG Local 10 . The FBI alleges Wilkins helped run an operation that embedded malware in eight video games, infected around 8,000 devices, and stole at least $220,000 from roughly 80 cryptocurrency wallets between May 2024 and February 2026. Investigators put a name to the scheme by following stolen Bitcoin to more than 150 gift cards, most of them spent on Uber Eats. The complaint identifies the distribution channel only as a "popular digital distribution software company," but the games it lists, including BlockBlasters, Dashverse, Lunara , and PirateFi , match the titles named when the FBI began seeking victims of infected Steam games in March. The case is being prosecuted in Seattle federal court, near Valve's headquarters in Bellevue, Washington, and Wilkins' arrest is the first publicly reported in the investigation. Wilkins allegedly financed and marketed the malware rather than writing it, with Local 10 reporting that agents had already searched the home of the unidentified developer who built the programs, and that Signal chats seized there tied Wilkins, operating under the handle Sibel.eth, to a $10,000 purchase of a remote access trojan and to discussions about tricking victims into approving transactions that emptied their wallets. That developer isn't named in the complaint and doesn't appear to have been charged. The conspirators promoted the games on Discord, Telegram, X, and LinkedIn, and used bots to find users with large cryptocurrency holdings and message them directly, according to the complaint. Roughly 80 wallets were drained from 8,000 infections, a hit rate of about 1% consistent with that targeted approach. ZachXBT and vx-underground estimated BlockBlasters alone took more than $

COMPUTEThe RegisterYesterday

South Korea making its own security-centric AI model

South Korea is developing its own security-focused AI model and hopes to bring it online by the end of the year, to ensure the nation has sovereign bug-finding capabilities. Deputy Prime Minister and Minister of Science and ICT Bae Kyung-hoon revealed the effort to create the model yesterday, and said it’s needed so South Korea possesses a bug-finding model to rival Anthropic’s Mythos. The US government has twice blocked access to Mythos, once by requiring Anthropic to offer it only to American citizens – a demand the AI company could not meet and therefore blocked all access – and a second time by ordering the company to take down its services so Washington could investigate allegations of possible dangerous performance problems. Those incidents led many other nations conclude that the US could in future deny access to powerful models – meaning US-based organizations and national security agencies would have an edge. Washington has since allowed limited access to Mythos to some of its allies. Interest in developing sovereign AI capacity has nonetheless soared, and Bae said South Korea now aspires to develop its own Mythos-class model. The Register is aware of another effort to create Mythos-like tools, involving private firms and infrastructure operators across several countries. In South Korea, the government’s approach is to add security-related information to the corpus it is using to train a locally developed frontier model. The minister said he expects that security-capable model will debut by the end of 2026. South Korea has also sought bids to create a chatbot that will be made freely available to all residents, plus an agentic application that will help locals interact with government services. Minister Bae made his remarks at a policy briefing session conducted by President Lee Jae Myung, during which discussions about AI also touched on using the technology to detect fake news in real time, and put it to work handling complaints about government services

LAUNCHESLatent SpaceYesterday

[AINews] Kimi K3 2.8T-A50B: the largest open model ever released; Opus 4.8-class at Sonnet 5 pricing

Z.ai GLM has been getting a bit too much love recently , so it’s time for Kimi K3 to fight back! It’s hard to put the scale of today’s open model release in perspective, so thankfully Moonshot AI did it for us : Their vibe reel was entirely edited by Kimi K3 and worth a watch: You can read SimonW and Arena for standard takes and rankings, none of which will be particularly unexpected given the large size of the model, but this pic best summarizes the K2.5 to K3 jump: AI News for 7/15/2026-7/16/2026. We checked 12 subreddits, 544 Twitters and no further Discords. AINews’ website lets you search all past issues. As a reminder, AINews is now a section of Latent Space . You can opt in/out of email frequencies! AI Twitter Recap Moonshot AI launched Kimi K3 as a frontier-class open-weights model, with official claims that place it near top closed models and above prior open competitors. Moonshot officially introduced Kimi K3 as “Open Frontier Intelligence” with 2.8T total parameters , 1M-token context , native multimodal input , Kimi Delta Attention (KDA) , and Attention Residuals , and said the model is live on Kimi.com, Kimi Work, Kimi Code, and API, with open weights promised by July 27, 2026 @Kimi_Moonshot Moonshot also highlighted product positioning around long-horizon agentic coding and self-evolving workflows , plus “vision in the loop” coding/game-building workflows that iterate between code and screenshots @Kimi_Moonshot Before the formal announcement, multiple accounts circulated leaked or app-sourced details that K3 was 2.8T params , calling it the largest open-weight model ever if weights ship as promised @scaling01 , @scaling01 , @eliebakouch The official Kimi blog went live later and was widely shared as the primary technical source @Jianlin_S , @scaling01 , @Yulun_Du Moonshot’s own phrasing acknowledged a limitation: despite being highly competitive overall, K3 still has a “noticeable gap in user experience” versus Claude Fable 5 and GPT-5.6 Sol @scaling01

LAUNCHESThe VergeYesterday

Fortnite is getting a bunch of AI-powered ‘personas’

Get ready for more AI characters in Fortnite. Developer Epic Games is going to let Fortnite creators publish experiences featuring characters with AI-powered voices starting on July 30th, and ahead of that launch, it's created 36 characters with "consistent voices and personas" that creators can use as NPCs. The characters include Fortnite staples like Agent […]

LAUNCHESThe RegisterYesterday

AI vendors have found someone to pay their infrastructure bills: You

Forrester warns that customers should brace for bigger software bills next year as software and AI vendors raise prices and pile on usage charges. Working from a survey of more than 2,600 business and technology decision-makers, the tech research company said software budgets were expected to rise "as vendors increase prices or add usage charges to pass their AI costs to customers." In the last six months, Anthropic, OpenAI, and GitHub have shifted some services away from flat-rate subscriptions toward usage-based billing, prompting cost concerns among users. Forrester added Microsoft to the list, citing its recent launch of the premium E7 license, which bolts M365 Copilot, Agent 365, and security tools onto E5. Last year, consultants Bain & Company estimated that the build cost for AI datacenters would hit $2 trillion by 2030. Forrester said that AI would drive increases in data and software spending, with 80 percent of decision-makers expecting those budgets to rise. Sharyn Leaver, chief research officer at Forrester, said: "The organizations that outperform in 2027 won't be those that spend the most on AI. They'll be the ones that invest in the foundations that make AI effective: trusted data, strong governance, organizational readiness, and the ability to continuously adapt as technology and customer behavior evolve." Forrester also found that personnel costs have yet to fall, despite the "AI washing of layoffs" in the tech industry. "While several tech giants, including Oracle, Microsoft, and Meta, have announced significant layoffs in recent months, IT staffing spend has not declined in recent years," the report said. Staffing accounted for 35 percent of IT budgets in 2025. For 2027, 67 percent of tech decision-makers expected to increase their staffing budget, while 23 percent said it would stay flat, and 10 percent expected it to decline. "The AI washing of layoffs will continue as vendors trim for financial and restructuring reasons. Guard against inflated

COMPUTEArs TechnicaYesterday

Could China and Russia really destroy Starlink? Only with a boomerang.

One week ago, three widely respected European news outlets published the results of an investigation into what they described as a "joint plan" by China and Russia to "defeat Elon Musk's Starlink." The story was the product of a long-running inquiry by The Insider, Der Spiegel, and Le Monde . Reporters at those publications said they reviewed a cache of documents detailing growing military cooperation between China and Russia. The documents covered discussions between the nuclear powers on integrated air and missile defense systems, autonomous "swarm" loitering munitions, next-generation armored vehicles, and military aviation, the report said. According to the papers, the investigation found evidence of a partnership between China and Russia in the field of space weapons far deeper than either country has acknowledged. One particular focus for China and Russia has been developing strategies to counter SpaceX's Starlink satellite broadband network. Read full article Comments

COMPUTEHugging FaceYesterday

NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval

AITechCrunch AIYesterday

Yes, you can now order DoorDash from the command line

DoorDash is opening a limited beta of dd-cli, a command-line tool that lets developers and AI agents search stores, build carts, and place orders from the terminal, marking another step toward software designed for AI agents instead of just humans.

DEVHacker News4d ago

Show HN: Clawk – Give coding agents a disposable Linux VM, not your laptop

128 points, 119 comments on HN

DEVHacker News5d ago

Old and new apps, via modern coding agents by Terry Tao

312 points, 84 comments on HN

AISimon Willison2d ago

Mermaid to Unicode box art (grok-mermaid)

<p><strong>Tool:</strong> <a href="https://tools.simonwillison.net/grok-mermaid">Mermaid to Unicode box art (grok-mermaid)</a></p> <p>While <a href="https://simonwillison.net/2026/Jul/15/grok-build/">exploring the codebase</a> for the newly open-sourced Grok CLI coding agent I came across <a href="https://github.com/xai-org/grok-build/blob/b189869b7755d2b482969acf6c92da3ecfeffd36/crates/codegen/xai-grok-markdown/src/mermaid.rs">xai-grok-markdown/src/mermaid.rs</a>, a "self-contained terminal renderer for Mermaid diagrams" written in Rust.</p> <p>I figured it would be fun to try that out in a browser via WebAssembly. Here's <a href="https://github.com/simonw/tools/pull/293#issue-4897479396">the prompt</a> I ran in Claude Code for web (Fable 5), and this is what the resulting tool looks like:</p> <p><img alt="Screenshot of a Mermaid diagram editor showing source code and rendered flowchart. The code reads: graph TD Start[Request received] --&gt; Auth{Authenticated?} Auth --&gt;|yes| Rate{Rate limit OK?} Auth --&gt;|no| R401[401 Unauthorized] Rate --&gt;|yes| H(Handle request) Rate --&gt;|no| R429[429 Too Many Requests] H -.-&gt; Log[Audit log] H ==&gt; Resp[200 OK]. Below the code are controls labeled Max width: Fit output panel, Copy as text, and Copy link to this diagram. The rendered flowchart on a dark background flows top-down: Request received leads to Authenticated?, which branches yes to Rate limit OK? and no to 401 Unauthorized. Rate limit OK? branches yes to Handle request and no to 429 Too Many Requests. Handle request connects with a dotted arrow to Audit log and a thick arrow to 200 OK." src="https://static.simonwillison.net/static/2026/grok-mermaid-wasm.png" /></p> <p>Tags: <a href="https://simonwillison.net/tags/tools">tools</a>, <a href="https://simonwillison.net/tags/rust">rust</a>, <a href="https://simonwillison.net/tags/webassembly">webassembly</a>, <a href="https://simonwillison.net/tags/mermaid">mermaid</a>, <a href="https://simonwillison.net/tags

The Essay
The week argued in one thesis: developed from our daily calls, settled in public
All essays →
About nextbig.dev

Built for builders

An independent briefing for builders: the whole field read continuously, every story scored for relevance, and the noise left off the page.

>_
Signal over noise

300+ curated sources. Every story scored 1–10 for builder relevance by Claude's frontier model. The filler never makes it to the page.

The compute beat

GPUs, datacenters, power deals, and inference economics: the infrastructure layer that decides what every builder pays. Our signature coverage.

[·]
We show our work

Every story is sourced. Every score is computed. We show our work and link to originals.

Skin in the game

Every briefing closes with The Call: one falsifiable claim with a date on it. When we're wrong, we say so in print. Opinions are cheap; ours get scored.

The wire is curated and scored with AI from 300+ sources, then edited by Oday Brahem. It can occasionally contain errors. Always verify critical information from the linked primary sources.
The Wire