The one thing
01TypeSafe prices Jev at $0.042 per million input tokens with free output, then gates it behind a waitlist.
Diogo Almeida, who worked on the instruction-following research behind ChatGPT, came out of two years of stealth yesterday with what TypeSafe AI calls a System One model. Jev takes unstructured program state in and returns typed probabilistic decisions with calibrated confidence; it generates no strings, so it cannot invent a field that was never declared. Published price is $0.042 per million input tokens with output free, against the $0.20 to $10 per million TypeSafe lists for frontier models.1
The speed claim is 70ms to 500ms end to end, which TypeSafe puts at 40x to 200x faster on the query shapes it targets — from evaluations it designed, ran on its own laptops and scored against two frontier models it chose as the reference.1 It took 1,085 points on Hacker News overnight.10 Access is a waitlist, and a price you cannot buy at is a press release.
Shipped
02GitHub has disabled SHA-1 in HTTPS
On 15 September GitHub completed a sunset it had scheduled in advance: SHA-1 in HTTPS is now off for github.com and partner CDNs, including Enterprise Cloud and Enterprise Cloud with Data Residency. Enterprise Server is unaffected.2
Gemini 3.8 Live and 3.8 Live Extended Thinking, in the API today
Both are available now through the Gemini API and Google AI Studio at $0.005 per minute of audio input and $0.018 per minute of audio output.3 For a voice agent the useful part is asynchronous function calling — tool calls run in the background while audio keeps streaming — alongside visual grounding, mid-conversation switching across 97 languages, and alphanumeric parsing for confirmation codes.34
Promised, not shipped
TypeSafe Jev — early access off a waitlist, no general availability date.[^1] · Gemini Enterprise for Customer Experience — "coming soon", no date.[^4]
The conversation
01Three vendors published a leaderboard position yesterday. A preprint filed the same afternoon argues the leading coding-agent leaderboard can no longer order its own top entries.6
- The authors of arXiv 2609.17394Preprint read directly; no peer review
Audited 254 SWE-bench submissions across four splits without running a single model. On Verified the top two entries each resolve 396 of 500 instances, and the top ten share 285 successes and 51 failures — leaving 164 instances that separate them at all. Exact paired McNemar tests separate none of the 29 adjacent Verified top-thirty pairs at alpha 0.05, and within-model scaffold ranges reach 29.8 percentage points against the 8.8-point spread across that whole top thirty.6
On SWE-bench Verified the harness you wrap a model in moves the resolution rate by up to 29.8 points, more than three times the spread across the entire top thirty. A rank won by swapping scaffolds is a fact about your plumbing — so pick your agent on your own repository.
The X and Reddit layers did not cover this window: the newest digest for either is 2026-09-15T05:30Z, roughly 24.5 hours old, closing before ours opens, and the accounts roster produced only markdown. Today's community reading is Hacker News and arXiv alone.
In production
01Claude Mythos 5.1 and Fable 5.1 returned intermittent error spikes on 15 September, an incident Anthropic rated major impact and closed after one hour and twenty-four minutes.5 OpenAI ran a separate major-impact incident on gpt-image-2.5-flare for 0.7 hours.9Source · Anthropic and OpenAI status pages, read 16 September 2026
Skip this
04TypeSafe's workflow evaluations The company designed the evaluation, chose the reference models and ran it from its own laptops. We quoted its price and refused its benchmark in the same breath.[^1]
"#1 on Artificial Analysis Speech-to-Speech" A vendor citing a leaderboard on the morning it ships, while its own coverage puts the sibling model second elsewhere.[^8]
Strix's admin access to Baseten's GitHub 258 points on Hacker News yesterday, but dated 1 September and the token was rotated in July. Out of window, already fixed.[^7]
Twelve releases across the watchlist Claude Code 2.1.273, openai-python 3.14.1, gemini-cli 0.60.0, adk-python 2.9.1, anthropic-sdk-python 1.6.0 and seven more, every one with an empty release note.
Everything we saw
141141 candidates scanned · 3 used in this issue — the rest, with the reason each one was left out
| Item | Source | Signal | Call |
|---|---|---|---|
| TypeSafe AI: System One models and Jev | Vendor blog / HN 1,085p | Priced, waitlisted, vendor-run evals | Led the issue as TOO_EARLY |
| SHA-1 in HTTPS on GitHub sunset | GitHub changelog | Completed, breaks old TLS clients | Shipped, BREAKING |
| Gemini 3.8 Live and 3.8 Live Extended Thinking | Google developer blog | In the API today, per-minute pricing | Shipped, USE_IT |
| Coding Agents Have Converged (arXiv 2609.17394) | arXiv cs.SE | 254 submissions audited, no ranking survives | Framed the conversation |
| Anthropic: error spikes on Mythos 5.1 and Fable 5.1 | Anthropic status | Major impact, 1.4 hours | Printed as the production figure |
| OpenAI: elevated errors on gpt-image-2.5-flare | OpenAI status | Major impact, 0.7 hours | Printed alongside |
| Strix: admin access to Baseten's GitHub | Vendor blog / HN 258p | Real finding, remediated in July | Skipped, out of window |
| Enforce GitHub Advanced Security configurations | GitHub changelog | Enterprise admin scope only | No decision for a developer |
| Copilot suggests custom property definitions | GitHub changelog | Public preview, org metadata | Too narrow to print |
| Vercel: Is Agentic tailors its audit by site type | Vercel changelog | Report categorisation change | Watching |
| Hugging Face / IBM: will your agent do it again? | Hugging Face blog | Consistency tooling, no release | Watching |
| Java 27 and Swift 6.4 | HN 325p / 122p | Large releases, off this beat | Off beat |
| AgentGuard: guardrails from anomalous agent traces | arXiv cs.SE | No implementation to run | Watching |
| GitHub trending sweep | Discovery layer | 50 on-beat repositories, 2 admitted after judgement | Watchlist now 35 |
| X accounts roster and keyword digests | Community layer | Newest collection 2026-09-15T05:30Z | Layer did not cover the window |
| Reddit practitioner digest | Community layer | Newest collection 2026-09-15T05:30Z | Layer did not cover the window |
End of feed. That is everything from the window worth your time.
Next issue tomorrow, 06:00 UTC — and if nothing ships, it will say so in two hundred words.