Edition 12

AI technology intelligence, edited into one clear read.

Latest Tech News

What changed across the AI industry.

All articles

74 articles · Recently published

Models & Platforms

Anthropic has released Claude for Financial Advisors, a plugin that connects Claude to financial-data systems for meeting preparation, follow-up drafting and other administrative work while leaving advice and execution with human professionals.

Model Stack

OpenAI's GPT-6 Astra can finish some agent tasks in fewer steps while raising the stakes of unsupervised execution. Efficiency is part of the case for Astra; it is not a reason to give the model more authority.

Model Stack

Google kept the introductory token rate identical to its predecessor, but one independent test found that higher reasoning effort and extra agent loops pushed the cost of finished work up by roughly 40 percent.

Model Stack

Alibaba’s Qwen3.8 puts a 2.4-trillion-parameter model in public repositories. That is a real expansion of freedom—but mainly for organizations already equipped to operate industrial AI infrastructure.

Model Buzz

What people are discussing about the models they use.

All discussions
Selected discussions5 threads · Latest first
LINUX DO分区:前沿快讯Pricing discussion

LINUX DO Thread Notes All-Day DeepSeek Off-Peak API Pricing on Holidays and Weekends

On September 19, 2026, LINUX DO user SharkyMew reported that DeepSeek API usage is billed at off-peak rates all day during statutory holidays, citing official documentation. DeepSeek's pricing page defines peak hours as Monday through Friday, 01:00–04:00 and 06:00–10:00 UTC, leaving remaining hours off-peak at half the peak rate. The page identifies deepseek-flash as DeepSeek-V4.1-Flash.

In the repliesCommunity members discussed scheduling workloads over weekends and holidays. One user shared a personal example of processing 700 million tokens for 37 yuan. The thread contains no account logs, token mix details, cache-hit breakdowns or billing receipt to verify that figure, so it is presented as an individual anecdote rather than a price guarantee.

DeepSeek released DeepSeek-V4.1-Flash on September 10, 2026 under the API model name deepseek-flash. Actual costs vary with prompt and completion tokens, cache-hit rates and the provider's active pricing rules. The official page explicitly defines weekday peak windows and all other hours as off-peak; the holiday interpretation is the forum author's attribution.

LINUX DO thread (https://linux.do/t/topic/2922294); DeepSeek API Docs Models & Pricing (https://api-docs.deepseek.com/quick_start/pricing/); DeepSeek API Docs Change Log (https://api-docs.deepseek.com/updates/).

LINUX DO分区:前沿快讯Product feedback

LINUX DO Thread Discusses GLM-5.3-FlashX Release, Pricing, and Account Access

LINUX DO user cooor shared a screenshot-led post stating that GLM-5.3-FlashX had been released, prompting a forum discussion on model availability, costs, and plan eligibility.

In the repliesCommenters noted the model feels faster but drains allowances more rapidly, while one user reported being unable to access it via their Coding Plan and cited official documentation calling it unsupported. Commenter yyf123 linked to official Z.AI pricing, which lists GLM-5.3-FlashX at $0.37 per 1M input tokens, $0.075 per 1M cached input tokens, and $1.25 per 1M output tokens—roughly 2.5 times the rates for GLM-5.3-Flash ($0.15 input, $0.03 cached input, $0.50 output), with cached storage listed as limited-time free for both.

Official documentation corroborates the per-token pricing structure, but community observations regarding speed, allowance burn, and Coding Plan restrictions remain anecdotal. The thread lacks formal benchmark scores, architecture specifications, or a comprehensive plan-eligibility matrix.

LINUX DO thread (https://linux.do/t/topic/2917803) and official Z.AI pricing documentation (https://docs.z.ai/guides/overview/pricing).

Redditr/DeepSeekHands-on

Developer Reports Fast Coding Handoff From Codex to DeepSeek 4.1

Reddit user AllenLeftTheBLDNG said they switched from Codex to DeepSeek 4.1 after exhausting Codex usage during a large coding project. The author described DeepSeek 4.1 as cheap, faster and better at following instructions, but supplied no task specification, token counts, billing evidence, controlled baseline or exact variant beyond “4.1.”

In the repliesCommenter Wrong_Visual5077 described having the model write a detailed Markdown specification before building, rather than relying only on planning mode. The commenter said they access DeepSeek and GLM through Claude Code Desktop and Cline Desktop, and personally found Cline’s output choppy at times while offering higher usage caps with slower output. Another user asked about the harness; no controlled replication was provided.

DeepSeek 4.1 coding handoff; reply mentions DS4.1f, GLM 5.3f, Claude Code Desktop and Cline Desktop; task, settings and billing details unspecified.

Personal reports from one Reddit thread, crawled September 17, 2026. Relative timestamps and missing task logs/configuration limit verification; not a benchmark or pricing comparison.

Redditr/LocalLLMHands-on

Developer Reports a 30-Day Local Qwen 3.8 Coding Run With Tool-Loop and Cache Issues

Reddit user Efficient-Part5344 reported running Unsloth Qwen3.8-27B-UD-Q4_K_XL locally for 30 days, using it for daily small coding tasks, overnight runs of more than nine hours and some production services. On an RTX 5070 Ti plus RTX 4070 Super, the author reported 73.8 mean/95.1 peak generation tokens per second, 845.1 mean/1,729.5 peak prompt processing, and 0.48108 MTP acceptance (674/1,401). They also reported slow, token-heavy reasoning, looping tool calls and cache behavior that could trigger full prompt reprocessing.

In the repliesAfter a hardware question, the author identified a Ryzen 7 5700X3D and 32GB RAM, and said subagents can save the main context even without parallel serving. Other commenters welcomed the 30-day run or reported similar speeds, but did not reproduce the measurements.

Local Unsloth Qwen3.8-27B quantization with a custom coding harness; dual-GPU setup and cache/reasoning behavior are setup-specific.

Single-user Reddit report crawled September 17, 2026; relative timestamps are not stable. Measurements are self-reported, not standardized or independently verified.

Redditr/codexHands-on

Community Reports on GPT-6 Astra Quota Limits and Browser Tasks

Reddit user ImpressiveTable3654 reported that running GPT-6 Astra in ChatGPT Work against a "decently large codebase" on a $20 plan routinely hit a session limit after roughly 15 minutes within a single prompt. The author said lighter, non-coding work appeared to consume less file context and allow longer use, but supplied no token counts, telemetry, exact settings or reproducible benchmark.

In the repliesIn visible replies, ActionOrganic4617 claimed Astra completed nine mandatory corporate-training modules in parallel with a 100% score after opening the portal in the Codex browser, muting audio and accelerating playback where possible. Other commenters discussed local models, quotas and HR automation without providing controlled validation. These are individual experiences, not verified platform limits or performance standards.

GPT-6 Astra in ChatGPT Work · Large-codebase coding versus browser-assisted training · $20 plan mentioned by the author; exact model settings, token counts and account telemetry unspecified.

Discovery indexing dates the original post to September 12, 2026. The live Reddit page shows relative timestamps without stable absolute metadata. Speculative HR job-displacement claims were excluded.

Selected and reviewed Sep 19, 2026. Summaries reflect individual posts and replies, not a community-wide verdict.

Skill Guides

Skills for everyday work.

19 capabilities for coding, design, office work, and media production. Updated September 15, 2026.

Coding
Use case

Review existing assertions and identify missing regression test scenarios.

Best for

Prioritizing tests for a feature or pull request

Works with

GitHub Copilot with Agent Skills support

Skill guideRead the update
Source repositoryGitHub

Hot Projects

Useful projects, with the details that matter.

See all Hot Projects

Week 36 · 2026 · Snapshot September 2, 2026 · GitHub Trending weekly figures

Top Pick

Archify

Generate technical diagrams from system descriptions or source-code repositories.

View repository
Why it’s rising
+25,469 stars in the weekly snapshot
Who it’s for
Software architects, engineering leads, and developers
Total stars
42,548 total stars
Latest activity
v2.16.0 · Aug 30
License
MIT
Consideration
Young project; stable release trails main-branch development.

OpenMAIC

Interactive multi-agent classrooms

+8,014weekly snapshot
MIT

OpenSEO

Agent-accessible SEO workflows

+2,625weekly snapshot
MIT

OpenCut

Programmable video editing in a ground-up rewrite

+2,630weekly snapshot
MIT

Apache Maka

Local-first, recoverable agent workspaces

+1,285weekly snapshot
Apache-2.0

Behind the coverage

Meet the editors.

About TechReadly