[ SECTION / POSTS ]
Posts
27 POSTS01
POSTS / 2026.08.03 / 7 MIN
Qwen3.8-Max Released: 2.4T-Parameter Coding Model, Open Weights Next Week
Qwen3.8-Max is out: 2.4T total params (95B active), the first Max-class model to open its weights (HF + ModelScope next week). Autonomous coding case: 265 commits in 16 …
↗
02
POSTS / 2026.08.02 / 8 MIN
Who Killed RSS? The Year Google Closed Reader, the Internet Broke a Little
Google Reader, FeedBurner, the Chrome RSS button, Google Alerts RSS — Google didn't hate RSS. It loved the users RSS brought in, then closed every door behind them. A …
↗
03
POSTS / 2026.08.02 / 8 MIN
AI Code's Bad Patterns Self-Replicate: 3 Instances Become 30, Silently
AI code decay isn't about 'low quality' — it's a structurally different kind of technical debt. Models treat your codebase as a style guide and faithfully replicate …
↗
04
POSTS / 2026.08.02 / 7 MIN
Cursor's Bill Disappeared. The Official Answer: Deliberate Design.
On July 31st, Cursor quietly removed dollar amounts from its usage page, zeroed out API cost fields, and an official employee confirmed it was 'deliberate design.' …
↗
05
POSTS / 2026.08.01 / 8 MIN
AI Agents Go to Work: qm, YC's Open-Source Multiplayer Agent Harness
YC's team open-sourced qm: a multiplayer agent harness where every employee gets an isolated workspace while agents also collaborate in Slack channels, groups, and …
↗
06
POSTS / 2026.08.01 / 9 MIN
2.78 Trillion Parameters, 29GB RAM, Half a Token Per Second: Kimi K3 Actually Runs on a Laptop
WASTE, a pure-C inference engine, squeezes the 2.78-trillion-parameter Kimi K3 into a 64GB laptop: a 982GB container streams expert weights from disk on demand, measured …
↗
07
POSTS / 2026.07.30 / 9 MIN
Six Days, Thousands of Dollars: AI Completed All the Code, But Zero Research
Frontier AI agents worked for six days, spent thousands of dollars, and completed all engineering without human help — but the original paper authors rejected both …
↗
08
POSTS / 2026.07.30 / 6 MIN
AI Office Agents: Fast and Cheap, But Not Ready to Ship
Baidu's OmegaUse-OfficeVal benchmark tested 6 frontier AI agents on 100 real office tasks. The best model scored only 64% of the human baseline, and 38% of tasks scored …
↗
09
POSTS / 2026.07.28 / 7 MIN
Screenshots or Source Code? StateAct and JarvisHub Reveal Two Futures for GUI Agents
July 28's hottest HuggingFace papers decoded: Salesforce's StateAct replaces pixels with program state for long-horizon GUI tasks, while JarvisHub builds an open test …
↗
10
POSTS / 2026.07.28 / 7 MIN
Claude Opus 5 Underrated, Kimi K3 1.56TB Open-Weight, GLM 5.2 Becomes Security Default: July 2026 Model Landscape Shift
July's triple hit: Anthropic's Opus 5 tops coding but benchmarks underrate it, Moonshot's 2.8T-param Kimi K3 goes open-weight with a stricter license, and Zhipu's GLM 5.2 …
↗
POSTS / 2026.08.03 / 7 MIN
Qwen3.8-Max Released: 2.4T-Parameter Coding Model, Open Weights Next Week
Qwen3.8-Max is out: 2.4T total params (95B active), the first Max-class model to open its weights (HF + ModelScope next week). Autonomous coding case: 265 commits in 16 …
↗
02
POSTS / 2026.08.02 / 8 MIN
Who Killed RSS? The Year Google Closed Reader, the Internet Broke a Little
Google Reader, FeedBurner, the Chrome RSS button, Google Alerts RSS — Google didn't hate RSS. It loved the users RSS brought in, then closed every door behind them. A …
↗
03
POSTS / 2026.08.02 / 8 MIN
AI Code's Bad Patterns Self-Replicate: 3 Instances Become 30, Silently
AI code decay isn't about 'low quality' — it's a structurally different kind of technical debt. Models treat your codebase as a style guide and faithfully replicate …
↗
04
POSTS / 2026.08.02 / 7 MIN
Cursor's Bill Disappeared. The Official Answer: Deliberate Design.
On July 31st, Cursor quietly removed dollar amounts from its usage page, zeroed out API cost fields, and an official employee confirmed it was 'deliberate design.' …
↗
05
POSTS / 2026.08.01 / 8 MIN
AI Agents Go to Work: qm, YC's Open-Source Multiplayer Agent Harness
YC's team open-sourced qm: a multiplayer agent harness where every employee gets an isolated workspace while agents also collaborate in Slack channels, groups, and …
↗
06
POSTS / 2026.08.01 / 9 MIN
2.78 Trillion Parameters, 29GB RAM, Half a Token Per Second: Kimi K3 Actually Runs on a Laptop
WASTE, a pure-C inference engine, squeezes the 2.78-trillion-parameter Kimi K3 into a 64GB laptop: a 982GB container streams expert weights from disk on demand, measured …
↗
07
POSTS / 2026.07.30 / 9 MIN
Six Days, Thousands of Dollars: AI Completed All the Code, But Zero Research
Frontier AI agents worked for six days, spent thousands of dollars, and completed all engineering without human help — but the original paper authors rejected both …
↗
08
POSTS / 2026.07.30 / 6 MIN
AI Office Agents: Fast and Cheap, But Not Ready to Ship
Baidu's OmegaUse-OfficeVal benchmark tested 6 frontier AI agents on 100 real office tasks. The best model scored only 64% of the human baseline, and 38% of tasks scored …
↗
09
POSTS / 2026.07.28 / 7 MIN
Screenshots or Source Code? StateAct and JarvisHub Reveal Two Futures for GUI Agents
July 28's hottest HuggingFace papers decoded: Salesforce's StateAct replaces pixels with program state for long-horizon GUI tasks, while JarvisHub builds an open test …
↗
10
POSTS / 2026.07.28 / 7 MIN
Claude Opus 5 Underrated, Kimi K3 1.56TB Open-Weight, GLM 5.2 Becomes Security Default: July 2026 Model Landscape Shift
July's triple hit: Anthropic's Opus 5 tops coding but benchmarks underrate it, Moonshot's 2.8T-param Kimi K3 goes open-weight with a stricter license, and Zhipu's GLM 5.2 …
↗