Qwen3.8-Max网页开发评测升至第二,DeepSeek-V4-Flash刷新智能体成本性能前沿。多款开源工具与端侧部署方案同步涌现。

模型动态

Qwen3.8-Max升至Image-to-WebDev Arena第二

Qwen3.8-Max1631分,仅落后Claude Opus 5 Max39分,图生网页能力突出。[1][2][2][1]

DeepSeek-V4-Flash刷新智能体成本性能前沿

DeepSeek-V4-Flash单任务中位成本仅0.024美元,成为低价位中净提升最明显的模型。[3][3]

Maple-Preview探索三值权重稀疏推理

Maple-Preview总参数20B、激活约1B参数,采用三值权重降低开放模型部署成本。[10]

产品工具

Qwen-Image-3.0-Pro上线Qwen Cloud API

Qwen-Image-3.0-Pro现已上线Qwen Cloud,开发者可通过API调用。https://t.co/wG5vtLAYVp[4][4]

LLM 0.32加入推理轨迹与服务端工具

LLM 0.32新增推理轨迹、OpenAI Responses及内容寻址日志。https://simonwillison.net/2026/Aug/4/new-release-of-llm/#atom-everything[5][6][7]

llm-anthropic 0.26扩展Claude与MCP能力

新版支持Claude 5系列WebSearch、CodeExecution和Anthropic MCP。https://simonwillison.net/2026/Aug/4/llm-anthropic/#atom-everything[8]

Firecrawl推出Rust办公文档转Markdown库

该库可将14种办公格式转为干净的Markdown,便于直接输入LLM。https://t.co/zdc2Z4hJmQ[9]

硬件动态

VibeVoice 1.5B在iPhone实现本地实时生成

社区演示VibeVoice 1.5B在iPhone占用约2.2GB,速度最高达【1.28倍实时】。[11]