模型发布 32 个事件
本分类 RSS ↗今天
-
有限来源模型发布
AA is the reason for Qwen3.8 27B shipped with xhigh
I know why Qwen3.8 27B shipped with xhigh reasoning as default, it's to do its best in benchmarks. Models from top labs often get benchmarked at multiple reasoning levels, but that same treatment doe…
Reddit r/LocalLLaMA -
有限来源模型发布Agent · 编程
Optimizing Qwen3.6 / Qwen3.8-27B on 16GB VRAM: Complete Benchmark Results and Setup Guide (~30-50tps at 32k to 72k context)
This post was made with AI. I tried to remove as much slop as possible and keep it straight to the point to save your time as I know how annoying AI slop posts can be, but I still wanted to retain al…
Reddit r/LocalLLaMA 昨天
-
有限来源模型发布研究 · 评测
在 4 张 RTX 3090 上实测 Qwen3.8-27B译
A while back I made a post about my 4x3090 rig in a Silverstone RV-02 . Check it out if you're a conoissuer of OG PC cases. With the incredible Qwen 3.8 27B release I ran benchmarks. So in case you a…
Reddit r/LocalLLaMA -
有限来源模型发布
Qwen3.8-27B Uncensored Aggressive 发布:K_P 量化 + HauhauCS FastMTP(生成提速最高 3.02 倍)!译
The dense Qwen release is back! Qwen3.8-27B Uncensored Aggressive is out with the complete K_P quant range, Vision, native NextN, and HauhauCS FastMTP. Aggressive here means no refusals, no personali…
Reddit r/LocalLLaMA -
有限来源模型发布商业动态
GPT-5.6 Sol 价格直降 50%译
openai / gpt-5.6-sol GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and mu…
Hacker News Frontpage -
有限来源模型发布
奇怪,没人讨论 Qwen3.8 27B 的 Temperature 设置译
Mind you, it is 1.0 by default, yet everyone is focused on how much the new model thinks, restricting the reasoning budget and/or dropping the reasoning level. Set the temperature to 0.7 and the mode…
Reddit r/LocalLLaMA -
有限来源模型发布
我在 RTX 3090 上把 Qwen3.8-27B 推到单请求 99 tps、批量 1150 tps译
I'm back. Yesterday I released the first version of hyper-optimized Qwen3.8-27B inference engine for a RTX 3090, reaching 82 tps on single request and 672 peak. Over the last 24 hours I've been explo…
Reddit r/LocalLLaMA -
有限来源模型发布
Qwen3.8 27B 在 Artificial Analysis 的成绩离谱!译
submitted by /u/FormOne2615 [link] [comments]
Reddit r/LocalLLaMA -
有限来源模型发布Agent · 编程
Artificial Analysis Agentic 指数:Qwen3.8 27B 超过 Opus 5 Medium译
https://preview.redd.it/xh1rloaf4zjh1.png?width=1628&format=png&auto=webp&s=a536fae1b50b327f2bc55d4f4f874f94ae66e867 Thanks Qwen team! submitted by /u/secopsml [link] [comments]
Reddit r/LocalLLaMA -
有限来源模型发布
Qwen3.8 27B = 压缩进 27B 的 GPT-5.6 Luna译
How crazy is that? submitted by /u/kevinlch [link] [comments]
Reddit r/LocalLLaMA -
有限来源模型发布研究 · 评测
Artificial Analysis 评测:Qwen3.8-27B 与 DeepSeek V4、GPT-5.6 Luna Max 并驾齐驱译
submitted by /u/anderspitman [link] [comments]
Reddit r/LocalLLaMA -
有限来源模型发布
Qwen3.8 27B 在 Artificial Analysis 得 52 分译
Article URL: https://artificialanalysis.ai/models/qwen3-8-27b Comments URL: https://news.ycombinator.com/item?id=49334544 Points: 101 # Comments: 37
Hacker News Frontpage -
有限来源模型发布Agent · 编程
Cursor 发布 GitHub 替代品 Origin译
Aug 17, 2026 · Changelog Cursor can now host your code. Origin begins rolling out today in early beta on all paid plans. We're starting with the essentials, designed for agent scale: repos, pull requ…
Hacker News Frontpage -
有限来源模型发布
Qwen3.8-27B 在 24GB RTX PRO 4000 SFF 上跑 256K 上下文(432 GB/s):开 MTP 达 50 tok/s译
AI INFRASTRUCTURE I gave Qwen3.8's MTP drafter another 69.2 MiB of precision. Throughput fell from 50.44 to 37.02 tokens per second. That result sums up the whole experiment: the best local inference…
Hacker News Frontpage -
有限来源模型发布
Qwen3.8-27B Q8_0 在 Strix Halo 上表现相当惊艳译
Sorry for the slop, but I was impressed by this model as I have been testing Qwen3.8-27B Q8_0 locally on my ROG Flow Z13 (Ryzen AI Max+ 395, 128 GB unified memory) and this model was the only one who…
Reddit r/LocalLLaMA -
有限来源模型发布硬件 · 算力
你的 Qwen3.8-27B 输出速度是多少 token/秒?译
Trying to get a feel for where I stand. If you can list your relevant hardware and model used, that would be awesome. Here's mine: Model: Qwen3.8-27B-heretic-ara, Q5_K_M GGUF T/s : ~30-32 t/sec (I th…
Reddit r/LocalLLaMA 8月13日 · 周四
-
有限来源模型发布
llm-gemini 0.33
Release: llm-gemini 0.33 It's been a while since the last llm-gemini release. This version of the plugin adds support for today's Gemini 3.7 Flash release, plus gemini-3.6-flash , gemini-3.5-flash-li…
Simon Willison's Weblog -
已确认模型发布
Ultrafast 模式预览:GPT-5.6 Sol 提速最高 14 倍译
Lucas Ropek If you’ve ever found yourself wishing that ChatGPT was a little bit quicker on the uptake, OpenAI seems to be answering your prayers. The AI lab has rolled out a new mode called Ultrafast…
OpenAI News · TechCrunch AI -
已确认模型发布开源
Gemini 3.7 Flash 正式发布译
Our most intelligent workhorse model yet for coding and agents. Tulsee Doshi Senior Director, Product Management, on behalf of the Gemini team Today, we’re building on the progress of our widely used…
Ars Technica AI · Google DeepMind Blog 8月12日 · 周三
-
已确认模型发布
OlmoEarth 嵌入发布:从 OlmoEarth Studio 导出自定义嵌入用于下游分析译
📄 Tech Report: https://allenai.org/papers/olmoearth | 📊 Documentation: https://docs.olmoearth.allenai.org/embeddings | 💻 Learn more about OlmoEarth: https://allenai.org/olmoearth OlmoEarth Studio…
Hugging Face Blog 8月11日 · 周二
-
已确认模型发布
Daybreak 系列模型登陆 AWS译
OpenAI and AWS are making Daybreak cybersecurity capabilities available through Amazon Bedrock to support enterprise security workflows.
OpenAI News 8月10日 · 周一
-
有限来源模型发布
Muse Glimmer 正式发布译
Introducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to…
Simon Willison's Weblog -
已确认模型发布
Model ML 用 GPT-5.6 Sol 更高效地完成金融工作译
Model ML uses GPT-5.6 Sol to carry finance work from research and analysis through editable, traceable PowerPoint decks and Excel workbooks.
OpenAI News 8月9日 · 周日
-
有限来源模型发布
引述 Claude Opus 5 系统提示词译
Claude Fable 5 and Claude Mythos 5 were first released on June 9, 2026. On June 12, 2026, Anthropic suspended access to both models to comply with U.S. Department of Commerce export controls; the Dep…
Simon Willison's Weblog 8月8日 · 周六
-
有限来源模型发布Agent · 编程
Claude Code 的 Auto 模式成为 Pro / Max / Team 计划的默认选项译
Auto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode , to the point that they are making it the default setting for new s…
Simon Willison's Weblog 8月7日 · 周五
-
有限来源模型发布
Moonlight & Mayhem(Codex + GPT-5.6 Sol Ultra 打造的浣熊大劫案)译
Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra) On Wednesday I wrote about One-shotting a Raccoon Heist game using Claude Fable 5 , where I had Claude Fable 5 build a full working gam…
Simon Willison's Weblog 8月6日 · 周四
-
已确认模型发布
改进 ChatGPT 中的 GPT‑5.6 Sol,并向免费用户开放 GPT-5.6 Luna译
ChatGPT introduces improved GPT-5.6 Sol with better accuracy and consistency, plus expanded access for free users and unlimited everyday chats with GPT-5.6 Luna.
OpenAI News 7月30日 · 周四
-
已确认模型发布
GPT-5.6 推进性价比前沿译
Explore lower GPT‑5.6 pricing for Luna and Terra—and how OpenAI’s more efficient models help enterprises deploy AI workflows at scale.
OpenAI News 7月29日 · 周三
-
已确认模型发布
GPT-5.6 如何兼得前沿智能与前沿效率译
GPT-5.6 improves AI efficiency across models, inference, and agentic workflows, helping deliver more useful intelligence per dollar.
OpenAI News 7月22日 · 周三
-
已确认模型发布
OpenAI Presence 正式发布译
Introducing OpenAI Presence, a proven enterprise AI agent platform that helps organizations deploy trusted voice and chat agents for customer and internal workflows.
OpenAI News 7月21日 · 周二
-
已确认模型发布
ChatGPT 小微企业计划发布译
OpenAI launches the ChatGPT for Small Businesses program, helping entrepreneurs build AI skills, automate work, and grow with ChatGPT Work.
OpenAI News -
已确认模型发布
Gemini 3.6 Flash、3.5 Flash-Lite 与 3.5 Flash Cyber 发布译
Our newest Gemini models deliver the efficiency, latency, and reliability to build AI agents at scale. Tulsee Doshi Senior Director, Product Management, on behalf of the Gemini team Developers and cu…
Google DeepMind Blog