Signing you in...

Please wait while we verify your authentication

Article · Wednesday, September 9, 2026

AI developer tools · What shipped

For a senior engineer who already reads HN. Real changes in AI developer tools today: releases with version numbers, papers with benchmarks, repos that crossed a threshold worth knowing. Skip hype threads, pre-announcement leaks, and recycled summaries. Always link primary sources.

By Marius BongartsTech66 editions
← See today's latest
Editions
14 / 66
Generated by AI overnight from public sources, refreshed daily.
AI developer tools · What shipped
Wednesday, September 9, 2026
AI developer tools · What shipped

Fable 5.1 stays, Tencent Hy4 debuts, Apple Silicon speedup ships

1 min read

Claude Fable 5.1

Continuing previous issue from yesterday, the pricing edge widens further.

Anthropic confirmed Fable 5.1 is now live on all platforms—AWS, Google Cloud, Azure, and the Claude API—with input at $10 per million tokens and cache reads locked at $0.25 [Source: Anthropic]. Mythos 5.1, the unrestricted sibling, routes through trusted-access programs for verified researchers in cybersecurity and life sciences, using identical internals but relaxed guardrails. Terminal-Bench 4.0 holds at 55.8% on agentic coding; Humanity's Last Exam hits 65.0% with tools.

Expect enterprise agents to default to Fable 5.1 now.

Tencent Hy4 open-source

Tencent's 770B mixture-of-experts model enters preview for coding and research.

Hy4 activates only 49 billion parameters per request despite its full size, designed for software engineering, financial analysis, and research workflows [Source: The Daily Star]. Integration into Tencent's CodeBuddy and WorkBuddy tools is underway; the early version trades complexity latency and answer verification costs for scale. Open-source release follows, expanding the field beyond closed-API incumbents.

Watch for latency benchmarks against Fable and Muse on long-context tasks.

Rapid-MLX 0.13.4 on Apple Silicon

Local inference on M-series Macs got measurably faster this week.

Rapid-MLX 0.13.4 ships verified speedups across Qwen3.8-27B: decode throughput jumps 1.43× to 2.34× depending on prompt length, with time-to-first-token improvements from 0.54s to 0.51s on short prompts and 25.04s to 24.66s on 8K contexts [Source: GitHub]. New multi-token prediction qualification lands alongside maintained tier-1 compatibility across Claude Code, Aider, DeepSeek Harness, and four other production agents tested end-to-end. The model catalog expanded to 253 aliases covering text, image, video, and audio, all running as an OpenAI-compatible backend with no external API calls.

Local agent infrastructure just shed another layer of friction.

Sources
Introducing Claude Fable 5.1 and Claude Mythos 5.1 - Anthropic
Introducing Claude Fable 5.1 and Claude Mythos 5.1 - Anthropic
2 hours ago ... ... AI experts, testing whether the models could match human specialists' performance. Mythos 5.1's capabilities are greater than those of Mythos 5. However ...
anthropic.com
AI Summary

Anthropic released Claude Fable 5.1 and Claude Mythos 5.1, new flagship models for coding and knowledge work. Fable 5.1 costs approximately 25% less than its predecessor for typical workloads (up to 45% savings for agentic tasks) due to 75% reduced cache read pricing at $0.25 per million tokens. The model achieves state-of-the-art performance across multiple benchmarks including Terminal-Bench 4.0 (55.8% on agentic coding), CursorBench 3.2.0 (73.4%), and Humanity's Last Exam (65.0% with tools), demonstrating improvements over Fable 5 and Opus 5 across coding, scientific research, and knowledge work tasks. Mythos 5.1 is an identical model with relaxed safeguards available through trusted access programs for verified cybersecurity professionals and life sciences researchers. Fable 5.1 is generally available today on all platforms including AWS, Google Cloud, and Azure with input pricing at $10 per million tokens and output at $50 per million tokens.

Visit source
PUBG parent Tencent releases open-source AI model | The Daily Star
PUBG parent Tencent releases open-source AI model | The Daily Star
9 hours ago ... Chinese technology group Tencent, maker of hit mobile game PUBG, has released a preview of a new open-source artificial intelligence model aimed at software ...
thedailystar.net
AI Summary

Tencent released a preview of Hy4, an open-source AI model with 770 billion parameters (49 billion activated per request) designed for software engineering, research, and financial analysis. The model uses a mixture-of-experts architecture and will be integrated into Tencent's CodeBuddy and WorkBuddy developer tools, though Tencent noted the early version can be slower on complex tasks and may repeatedly verify answers.

Visit source
raullenchai/Rapid-MLX: The fastest local AI engine for Apple Silicon ...
raullenchai/Rapid-MLX: The fastest local AI engine for Apple Silicon ...
5 hours ago ... Twitter / X: Follow @rapidmlx for releases, benchmarks, and project updates. Questions & builds: Ask or share in GitHub Discussions. Feedback & ideas: Report a ...
github.com
AI Summary

Rapid-MLX 0.13.4 ships with verified performance improvements across large models on Apple Silicon. Qwen3.8-27B's decode throughput increased 1.43× to 2.34× across different prompt lengths (128 tokens to 32K), with TTFT improvements from 0.54s down to 0.51s on 128-token prompts and 25.04s to 24.66s on 8K prompts. The release includes new MTP (Multi-Token Prediction) path qualification for Qwen3.8-27B and maintains tier-1 compatibility verification across five production agents (Claude Code, Codex CLI, Hermes, Aider, DeepSeek Harness) tested end-to-end before release. Rapid-MLX also expanded its open model catalog to 253 total aliases covering text, image, video, and audio generation with reproducible benchmarks — all running as a local OpenAI-compatible backend on M-series Macs with no external API calls.

Visit source
Compiled overnight by MorningMail.aiDelivered at 05:10 AM