> ## Content Index
> Fetch the complete content index at: https://www.thenewway.ai/llms.txt
> Use this file to discover other available public pages before exploring further.

# Mistral ships a trillion-parameter model, Large 4, in API preview, with weights due at month's end
- URL: https://www.thenewway.ai/mistral-ships-a-trillion-parameter-model/
- Published: 2026-10-07T16:31:49.000Z
- Updated: 2026-10-07T16:32:01.000Z
- Description: A trillion-parameter model in preview, scoring 61.7% on the DeepSWE coding benchmark, with weights due this month. Plus Codex users vote for a usage reset.
- Author: The New Way

Mistral put a public preview of Large 4 on Mistral Studio Tuesday: a trillion-parameter model whose weights it says drop at the end of the month, scoring 61.7% on the DeepSWE coding benchmark. Separately, OpenAI's Codex team reset usage limits after 76% of 74,565 poll voters asked for it, on day two of a 28-day pledge to ship an improvement or a reset every day. GitHub shipped two changes to how teams build: a fix for a metrics bug that was hiding Copilot's agent activity, and general availability for stacked pull requests. And Claude Code's newest patch closes a real permission-prompt gap, buried in 92 bullets of plumbing.

📌

****In this issue**  
• [Mistral ships a trillion-parameter model, Large 4, in API preview, with weights due at month's end](https://www.thenewway.ai/mistral-ships-a-trillion-parameter-model/#mistral-ships-a-trillion-parameter-model-large-4-in-api-preview-with-weights-due-at-months-end)  
• [Codex users vote 76% for a usage reset, and OpenAI's Codex team grants it](https://www.thenewway.ai/mistral-ships-a-trillion-parameter-model/#codex-users-vote-76-for-a-usage-reset-and-openais-codex-team-grants-it)  
• [GitHub says a Copilot SDK migration undercounted agent activity in usage metrics](https://www.thenewway.ai/mistral-ships-a-trillion-parameter-model/#github-says-a-copilot-sdk-migration-undercounted-agent-activity-in-usage-metrics)  
• [GitHub's stacked pull requests reach general availability — approvals now survive a rebase](https://www.thenewway.ai/mistral-ships-a-trillion-parameter-model/#githubs-stacked-pull-requests-reach-general-availability-%E2%80%94-approvals-now-survive-a-rebase)  
• [Claude Code 2.1.292 fixes a permission-prompt bypass for network file reads](https://www.thenewway.ai/mistral-ships-a-trillion-parameter-model/#claude-code-21292-fixes-a-permission-prompt-bypass-for-network-file-reads)

---

## What changed this week

### Mistral ships a trillion-parameter model, Large 4, in API preview, with weights due at month's end

Mistral put a public preview of Large 4, "le Chonk," on Mistral Studio Tuesday: a 1-trillion-parameter, natively multimodal model with 49 billion active parameters. On the coding benchmarks Mistral reports, it scores 61.7% on DeepSWE v1.1, 59.4% on SWE-Atlas-QnA, and 28.3% on Terminal-Bench 4\. In a blind human evaluation it ranked second of five, behind only Claude Opus 5\. The louder claim is security. On one test in the independent Artificial Analysis Cyber Index, reproducing and patching a real vulnerability, it scores 82%, and Mistral says Opus 5.5 and GPT-6 Astra "score near zero" because they refuse. The same afternoon, [Anthropic widened access to Opus 5.5](https://x.com/AnthropicAI/status/2107546569654636883?ref=thenewway.ai) for verified security professionals doing defensive work. Weights drop at month's end.

[Introducing Mistral Large 4 | MistralThe most powerful AI platform for enterprises. Customize, fine-tune, and deploy AI assistants, autonomous agents, and multimodal AI with open models.![](https://storage.ghost.io/c/9c/1f/9c1ff71c-2859-44c2-a962-1337982a6484/content/images/icon/favicon-d4c09d52-4a1c-4559-9cf1-5ca67747d24d.ico)Mistral![](https://storage.ghost.io/c/9c/1f/9c1ff71c-2859-44c2-a962-1337982a6484/content/images/thumbnail/cover-mistral-sovereign-20copie-8e82f54b-06c8-44a0-9705-f0c48a753106.jpg)](https://mistral.ai/news/mistral-large-4/?ref=thenewway.ai)

---

### Codex users vote 76% for a usage reset, and OpenAI's Codex team grants it

Tibo Sottiaux of OpenAI's Codex team pledged Sunday that for 28 days, each day brings an improvement or "a full reset," meaning a reset of usage limits. Day two brought both. He listed four updates, among them the [Decisions API](https://developers.openai.com/api/docs/guides/decisions?ref=thenewway.ai), which returns a probability, a category or a score "about 10x faster than the Responses API" at $0.10 per million input tokens. He polled users, and [76% of 74,565 voted](https://x.com/thsottiaux/status/2107576143285219799?ref=thenewway.ai) "needs a reset." His reply: "the reset has been processed."

> We shipped four things that were deemed good to great and some math proofs, but the vote is clear and the community demands a reset. I did calibrate it and it \*seems\* that the game is rigged in reset's favor, but such are the rules at the moment.  
>  
> Therefore ... the reset has been… [https://t.co/mHSI0Pcu4M](https://t.co/mHSI0Pcu4M?ref=thenewway.ai)
> 
> — Tibo (@thsottiaux) [October 7, 2026](https://x.com/thsottiaux/status/2107676072871600470?ref%5Fsrc=twsrc%5Etfw&ref=thenewway.ai)

---

### GitHub says a Copilot SDK migration undercounted agent activity in usage metrics

GitHub explained why some teams watched Copilot's agent activity fall in their usage dashboards while overall usage kept growing. "Several IDEs recently moved Copilot agent sessions to the Copilot SDK," and those sessions "didn't identify which IDE they came from," so the metrics dropped most of that activity or miscounted it as CLI use. VS Code 1.139 has the fix now; Visual Studio, JetBrains, Eclipse and Xcode follow through November. The fix doesn't back-fill history. GitHub's own warning is blunt: "their activity can't be recovered later."

[Update your IDE to restore agent activity in Copilot usage metrics - GitHub ChangelogIf your Copilot usage metrics have shown agent activity or agent lines of code falling while Copilot usage kept growing, we’ve found the cause, and a fix is rolling out…![](https://storage.ghost.io/c/9c/1f/9c1ff71c-2859-44c2-a962-1337982a6484/content/images/icon/cropped-github-favicon-512-4bc060a4-576e-49b4-ad4c-16899061dcf4.png)The GitHub BlogAllison![](https://storage.ghost.io/c/9c/1f/9c1ff71c-2859-44c2-a962-1337982a6484/content/images/thumbnail/664302370-ef253fd4-491c-476f-95bb-ab6fd03ed1c1-ea984001-3a1c-4a85-bb75-12d152918a27.jpg)](https://github.blog/changelog/2026-10-06-update-your-ide-to-restore-agent-activity-in-copilot-usage-metrics?ref=thenewway.ai)

---

### GitHub's stacked pull requests reach general availability — approvals now survive a rebase

GitHub's stacked pull requests, the feature for splitting a large change into smaller reviewable ones, are generally available after a preview in which "repositories using stacks have seen a 9% increase in merged code compared to peers" and the top 1% of repos saw "a 5% improvement in time-to-merge." The GA release fixes a real workflow gap: rebasing a stack used to risk dismissing reviewers' approvals, and now keeps them in place even in repos that dismiss stale approvals on every other push. That's the real fix.

[Stacked pull requests generally available - GitHub ChangelogGitHub stacked pull requests are now generally available. Break large changes into smaller, focused pull requests that you can review independently and merge together. Since the feature went into public…![](https://storage.ghost.io/c/9c/1f/9c1ff71c-2859-44c2-a962-1337982a6484/content/images/icon/cropped-github-favicon-512-25aa3149-ba04-4963-aee7-d7affd1b2d29.png)The GitHub BlogAllison![](https://storage.ghost.io/c/9c/1f/9c1ff71c-2859-44c2-a962-1337982a6484/content/images/thumbnail/StackedPullRequests_NewRelease_Unfurl_Mobile_261006_1321-687d21b5-a052-4d80-a48d-13993a651f8e.png)](https://github.blog/changelog/2026-10-06-stacked-pull-requests-generally-available?ref=thenewway.ai)

---

### Claude Code 2.1.292 fixes a permission-prompt bypass for network file reads

Claude Code 2.1.292 closes a real gap: a security fix stops "PreToolUse hook approvals and auto mode" from "bypassing the permission prompt for file reads from network (UNC) paths." The same release fixes a tampered on-disk settings cache that could have disabled Claude Code's built-in policy plugin, and a sandboxed command that could read staged /ultrareview upload files it shouldn't see. The other 89 bullets are mod and plugin-hook plumbing. Not security fixes.

[Claude Code changelog - Claude Code DocsRelease notes for Claude Code, including new features, improvements, and bug fixes by version.![](https://storage.ghost.io/c/9c/1f/9c1ff71c-2859-44c2-a962-1337982a6484/content/images/icon/android-chrome-192x192-35ade1fe-9ed7-48f9-bd77-769cd0533e21.png)Claude Code Docs![](https://storage.ghost.io/c/9c/1f/9c1ff71c-2859-44c2-a962-1337982a6484/content/images/thumbnail/image-4e36afdf-c1cc-4f2b-8f64-ec9f05a2f76b.png)](https://code.claude.com/docs/en/changelog?ref=thenewway.ai)

---

Know someone who'd want this in their inbox? Forward it — that's how this grows. And if we got something wrong, or you think we buried the real story today, hit reply. A person reads every one.

---

*The New Way is written with AI. It gathers the day's stories, checks them against their sources and drafts every summary. A person decides what runs and reviews every issue before we hit send.*