AI RadarEvent-first AI intelligence · every item traceable
Last 24h 77 new events · 14 confirmed · 212 gained sources Last 7 days 179 new events
  • Today

  • Limited sourcesModels

    AA is the reason for Qwen3.8 27B shipped with xhigh

    1 source

    I know why Qwen3.8 27B shipped with xhigh reasoning as default, it's to do its best in benchmarks. Models from top labs often get benchmarked at multiple reasoning levels, but that same treatment doe…

    Reddit r/LocalLLaMA
  • ConfirmedOpen Source

    llama.cpp b10481: CUDA: MMVQ nwarps=8 for bs=1 for dense models on DGX Spark (#26843)

    CUDA: MMVQ nwarps=8 for bs=1 for dense models on DGX Spark Signed-off-by: ynankani ynankani@nvidia.com skip moe experts and allow others based on k geometry (allow only small idle tail) Signed-off-by…

    GitHub Releases · ggml-org/llama.cpp
  • Limited sources

    Qwen dev says not to wait for 35B-A3B

    1 source Alibaba / Qwen

    What does this mean? Is there something else coming? Maybe 122B? Or no models? submitted by /u/Mean-Ad1493 [link] [comments]

    Reddit r/LocalLLaMA
  • Limited sources

    Sparks jumping in price too

    1 source

    Just had an eBay seller cancel my order at 8300 AUD to relist at 9K flat. This is not a fly by night operator either, it’s a legit org. They said they are going to honour the previous pricing when th…

    Reddit r/LocalLLaMA
  • Limited sourcesModelsAgents & Dev

    Optimizing Qwen3.6 / Qwen3.8-27B on 16GB VRAM: Complete Benchmark Results and Setup Guide (~30-50tps at 32k to 72k context)

    1 source

    This post was made with AI. I tried to remove as much slop as possible and keep it straight to the point to save your time as I know how annoying AI slop posts can be, but I still wanted to retain al…

    Reddit r/LocalLLaMA
  • Limited sources

    CDW has bumped the MSRP of the RTX Pro 6000 from $16,000 to $19,999

    1 source

    Did they slip up and leak future pricing? Live link: https://www.cdw.com/product/pny-nvidia-rtx-pro-6000-graphic-card-96-gb-gddr7/8326705 Archive link: https://web.archive.org/web/20260818013250/http…

    Reddit r/LocalLLaMA
  • Limited sources

    Made this game in two prompts with Q4, Qwen 3.8 is amazing

    1 source Alibaba / Qwen

    This took one prompt to build, and another follow up prompt to fix two issues (player got stuck with the bomb and broken enemies path-finding), this is only html, css and js, no external assets, all…

    Reddit r/LocalLLaMA
  • Yesterday

  • Limited sources

    Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index

    1 source Alibaba / Qwen

    Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index That's the same score as GPT-5.6 Luna (max), and just one point behind GLM-5.2 (max) and DeepSeek V4 Pro 0813 (max) - that GLM is…

    Simon Willison's Weblog
  • Limited sources

    Anthropic’s annualized revenue surges to $65B

    1 source Anthropic

    Marina Temkin Anthropic’s revenue continues to not only grow at an historic pace but also to accelerate. The model maker’s annualized revenue run rate — a projection of a full year’s revenue based on…

    TechCrunch AI
  • Limited sources

    Qwen 3.8 distillations

    4 sources Alibaba / Qwen

    Artificial analysis index scores Qwen 3.5 27b: 35 Qwen 3.6 27b: 38 Qwen 3.8 27b: 52 What the hell kind of a jump was that? Even if it is benchmaxxed, the jump is insane. Qwen3.6 35b A3b: 32 That's ~6…

    Reddit r/LocalLLaMA
  • Limited sources

    Qwen 3.8 35b and 122b - We hope/wait/beg for models incessantly. But how do we actually give the lab more incentive to make it?

    1 source Alibaba / Qwen

    So many comments begging for these models, and I get it. People from the labs frequent r/LocalLLaMA , so maybe the begging comments aren't useless. They make demand known at least. Same with polls, e…

    Reddit r/LocalLLaMA
  • Limited sourcesResearch

    we benchmark models nobody actually runs

    1 source

    qwen3.8-27b looks genuinely impressive on the benchmark tables - beating models many times its size on some of them. but those numbers come from bf16 weights, and nobody here is running a 27b at bf16…

    Reddit r/LocalLLaMA
  • Limited sources

    My friends all hate AI; I just joined an AI startup

    1 source

    “I know we’re all anti-AI here,” a friend texted in the group chat. “We should co-write an article on how AI has no place in education,” a professional acquaintance suggested during a collaboration s…

    Hacker News Frontpage
  • Limited sourcesBusiness

    AI automation startup Relay shuts down, staff joins Google’s Chrome team

    Relay, an AI-powered workflow automation tool that was launched in 2021 with the goal of becoming the new Zapier , is shutting down, and some of its staff — including its top executive — are joining…

    TechCrunch AI
  • Limited sources

    Former SpaceX engineers are building a robotic factory for making steel parts

    1 source

    Three former SpaceX engineers have switched their attention from making rocket engines to manufacturing steel parts by using AI-driven software and robots. Their immediate goal involves establishing…

    Ars Technica AI
  • Limited sourcesModelsResearch

    Benchmarked Qwen3.8-27B on 4x RTX 3090

    1 source

    A while back I made a post about my 4x3090 rig in a Silverstone RV-02 . Check it out if you're a conoissuer of OG PC cases. With the incredible Qwen 3.8 27B release I ran benchmarks. So in case you a…

    Reddit r/LocalLLaMA
  • Limited sourcesModels

    Qwen3.8-27B Uncensored Aggressive is out with K_P quants and HauhauCS FastMTP (up to 3.02x TG)!

    1 source

    The dense Qwen release is back! Qwen3.8-27B Uncensored Aggressive is out with the complete K_P quant range, Vision, native NextN, and HauhauCS FastMTP. Aggressive here means no refusals, no personali…

    Reddit r/LocalLLaMA
  • Limited sourcesModelsBusiness

    GPT-5.6 Sol Pricing Cut by 50%

    1 source OpenAI

    openai / gpt-5.6-sol GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and mu…

    Hacker News Frontpage
  • Limited sources

    Israel creates fake think tank in likely attempt to dupe AI chatbots

    1 source

    At a glance, the Hanover Institute for Public Policy looks like a new think tank dedicated to Israel/Palestine. The organization churns out think-tank style reports on questions such as “Does AIPAC U…

    Hacker News Frontpage
  • Limited sourcesOpen SourceAgents & Dev

    Local agentic coding Benchmark : Qwen 3.8 27B (in many weights quants / cache quants / engine / reasoning effort) vs others.

    1 source Alibaba / Qwen

    In medium reasoning mode, it both scores higher than the 3.6 version, AND is very much more efficient (almost half requests needed, and a third less tokens generated) - at DeepSeek v4 Flash 3107 MXFP…

    Reddit r/LocalLLaMA
  • ConfirmedOpen SourceAgents & Dev

    claude-code v2.1.234

    1 source Anthropic

    What's changed Added the optional CLAUDE_CODE_PROJECT_DIR_NAME environment variable: hosts that give each session its own config directory can choose a short name for the per-project transcript direc…

    GitHub Releases · anthropics/claude-code
  • Limited sourcesAgents & Dev

    GitHub degradation affects Cursor Origin, its new Git platform

    Article URL: https://status.cursor.com/incidents/l9h9vrd726jv Comments URL: https://news.ycombinator.com/item?id=49336919 Points: 24 # Comments: 3

    Hacker News Frontpage
  • Limited sourcesModels

    Weirdly, no one talks about Temperature setting for the Qwen3.8 27b

    1 source

    Mind you, it is 1.0 by default, yet everyone is focused on how much the new model thinks, restricting the reasoning budget and/or dropping the reasoning level. Set the temperature to 0.7 and the mode…

    Reddit r/LocalLLaMA
  • Limited sourcesModels

    I pushed Qwen3.8-27B to 99 tps single request and 1150 tps with a batch request on a RTX 3090

    1 source

    I'm back. Yesterday I released the first version of hyper-optimized Qwen3.8-27B inference engine for a RTX 3090, reaching 82 tps on single request and 672 peak. Over the last 24 hours I've been explo…

    Reddit r/LocalLLaMA
  • Limited sources

    AI;DR (AI; Didn't Read)

    1 source

    I'm about as pro-AI as you can be, but this is becoming a pet peeve of mine (and I'm not alone). That's why I love the AI;DR acronym as my new solution for ignoring the walls of slop. I’m SUPER jealo…

    Hacker News Frontpage
  • Confirmed

    Same Cluster, 33 Points More Utilization: What Changed Was the Order

    1 source Hugging Face

    The previous post argued that utilization, not intelligence, is where the next real constraint in enterprise AI is forming, and it closed by noting that no playbook has emerged yet for what a mature…

    Hugging Face Blog
  • Limited sources

    "Opus 4.8 thinks too much", "Muse Glimmer sits between Gemma and Qwen, that's boring", "Gemma 4 is too lazy"

    I'm starting to think there's no way to make a reasoning model that won't draw persistent vocal complaints on here. EDIT: Qwen 3.8 not Opus 4.8*, freudian slip lol submitted by /u/MerePotato [link] […

    Reddit r/LocalLLaMA
  • ConfirmedOpen SourceBusiness

    openai-python v3.2.0

    1 source OpenAI

    3.2.0 (2026-08-17) Features add Bedrock Runtime endpoint support (SDK-290) ( #3623 ) ( 86267d2 ) api: Add shell call streaming events and new service/image types ( #3635 ) ( ff14a33 )

    GitHub Releases · openai/openai-python
  • Limited sources

    LLMs Endgame: This is Unreasonable

    1 source

    Qwen 3.8 27B Oh My This result is unreasonable Qwen 3.8 27B makes using almost any other model unreasonable (except for the big boyz ofc ?!) They took a while to release this and still even pricing A…

    Reddit r/LocalLLaMA
  • Limited sourcesModels

    Qwen3.8 27B's result on Artificial Analysis is insane!

    1 source

    submitted by /u/FormOne2615 [link] [comments]

    Reddit r/LocalLLaMA
  • Limited sources

    llama.cpp adaptive MTP PR#27210

    1 source llama.cppMeta

    Just wanted to raise some attention to this PR I filed if anyone would like to try it out. This adds an adaptive MTP mode to llama.cpp which employs a fairly simple counting-style state machine to de…

    Reddit r/LocalLLaMA
  • Limited sourcesModelsAgents & Dev

    Qwen3.8 27B > Opus 5 Medium on Artificial Analysis Agentic Index

    1 source Anthropic

    https://preview.redd.it/xh1rloaf4zjh1.png?width=1628&format=png&auto=webp&s=a536fae1b50b327f2bc55d4f4f874f94ae66e867 Thanks Qwen team! submitted by /u/secopsml [link] [comments]

    Reddit r/LocalLLaMA
  • Limited sourcesModels

    Qwen3.8 27B = GPT-5.6 Luna compressed into 27B

    1 source OpenAI

    How crazy is that? submitted by /u/kevinlch [link] [comments]

    Reddit r/LocalLLaMA
  • Limited sourcesModelsResearch

    Artificial Analysis' Qwen3.8-27B benchmarks put it neck and neck with DeepSeek V4 and GPT-5.6 Luna Max

    1 source DeepSeekOpenAI

    submitted by /u/anderspitman [link] [comments]

    Reddit r/LocalLLaMA
  • Limited sourcesModels

    Qwen3.8 27B scores 52 on Artificial Analysis

    1 source

    Article URL: https://artificialanalysis.ai/models/qwen3-8-27b Comments URL: https://news.ycombinator.com/item?id=49334544 Points: 101 # Comments: 37

    Hacker News Frontpage
  • ConfirmedOpen Source

    llama.cpp b10472

    1 source llama.cppMeta

    cuda : skip UMA override for HIP builds ( #27083 ) AMD APUs report accurate memory via hipMemGetInfo. Using MemAvailable over-promises on small-carveout systems. fixes #18159 Website: https://llama.a…

    GitHub Releases · ggml-org/llama.cpp
  • Limited sourcesModelsAgents & Dev

    Cursor launches Origin, GitHub alternative

    Aug 17, 2026 · Changelog Cursor can now host your code. Origin begins rolling out today in early beta on all paid plans. We're starting with the essentials, designed for agent scale: repos, pull requ…

    Hacker News Frontpage
  • Limited sourcesAgents & Dev

    Why do people like coding harnesses like opencode etc instead of an IDE?

    1 source

    Just curious - I like to be able to see and manage the scripts my agent is working on. I find stuff like Claude Code and Open Code useful for doing stuff on my linux box but I don't understand why pe…

    Reddit r/LocalLLaMA
  • Limited sources

    Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!

    1 source

    This Ling 3.0 Tiny 8b param with 1.3b active is the fastest, smartest model I can run on my poor old pc, with 4gb vram. It actually runs lightning fast, like 36 token / sec, as smart as Qwen 3.5 9b /…

    Reddit r/LocalLLaMA
  • Limited sourcesBusinessInfra

    Groq raises $350M to fuel its pivot from AI chips to neocloud

    1 source

    Startup Groq has raised $350 million as it continues to pivot from an AI chipmaker to a neocloud company that provides powerful GPUs and AI infrastructure services. The new capital, led by investment…

    TechCrunch AI
  • Limited sourcesCreative

    Launch HN: Speko (YC S26) – OpenRouter for Voice AI

    1 source

    Hi HN! I'm Bek, founder of Speko, a platform that finds an optimal combination of speech-to-text, LLM, and text-to-speech models, given your constraints, among all our public benchmarked options, and…

    Hacker News Frontpage
  • Limited sources

    We Tracked a Shipment of Rare Books. It Ended at an Amazon AI Training Facility

    1 source Amazon / AWS

    We Tracked a Shipment of Rare Books. It Ended at an Amazon AI Training Facility Excellent piece of reporting from 404 Media. For a while now there have been stories of book dealers receiving orders f…

    Simon Willison's Weblog
  • Limited sourcesPolicy & SafetyInfra

    Nvidia investing $1.5B in SoftBank data center developer behind OpenAI project

    1 source NVIDIAOpenAI

    Tim De Chant Nvidia said on Monday that it will invest $1.5 billion in SB Energy, a data center linked to SoftBank and OpenAI. The investment ensures that Nvidia will be the sole supplier of compute…

    TechCrunch AI
  • Limited sources

    Deepseek Harnness - why is feels better

    1 source DeepSeek

    Guys, could someone smarter than me explain what makes Deepseek Harness so efficient? I run it with local Qwen 3.8 (Q6). I tried Opencode/Openchamber (my favourite so far), Pi agent and Hermes. New Q…

    Reddit r/LocalLLaMA
  • Limited sources

    noctrex/Ling-3.0-tiny-MXFP4_MOE-GGUF · Hugging Face

    u/noctrex 👍 where's flash? 😄 Possibly fastest model(in this model size range). Share t/s stats. submitted by /u/pmttyji [link] [comments]

    Reddit r/LocalLLaMA
  • Limited sourcesPolicy & Safety

    Judge relying wholly on AI in order is covered by judicial immunity, court rules

    1 source

    From Wednesday's decision in Phillips v. Parlade , by Judge Gloria Navarro (D. Nev.), where a litigant sued a state court judge in his case: Plaintiff … argu[es] that judicial immunity does not apply…

    Hacker News Frontpage
  • Limited sourcesModels

    Qwen3.8-27B at 256K on a 24GB RTX PRO 4000 SFF (432 GB/s): 50 tok/s with MTP

    1 source

    AI INFRASTRUCTURE I gave Qwen3.8's MTP drafter another 69.2 MiB of precision. Throughput fell from 50.44 to 37.02 tokens per second. That result sums up the whole experiment: the best local inference…

    Hacker News Frontpage
  • Limited sourcesAgents & Dev

    AI-Generated GitHub Copilot "Autofix" Allowed Compromise of Snowflake's Jira

    1 source GitHubMicrosoft

    As part of ongoing security research conducted through Snowflake’s HackerOne vulnerability disclosure program, Wiz Research’s "Red Agent"—an autonomous, AI-powered security research tool—identified a…

    Hacker News Frontpage
  • ConfirmedOpen Source

    llama.cpp v0.1.1

    1 source llama.cppMeta

    Release v0.1.1

    GitHub Releases · ggml-org/llama.cpp
  • Limited sources

    How to disable or avoid intrusive AI

    1 source

    One of the biggest questions I get at Drop-In Time at the library (besides “what is taking up all my cloud storage?”) is how to disable or avoid intrusive AI that shows up where people don’t want it.…

    Hacker News Frontpage
  • ConfirmedOpen Source

    llama.cpp b10470

    1 source llama.cppMeta

    ci : push release tag explicitly in release.yml ( #27261 ) Add a "Create and push git tag" step to the release job, right before the "Create release" step. The tag is created with git tag and pushed…

    GitHub Releases · ggml-org/llama.cpp
  • Limited sources

    DeepSeek V4 Flash with Antirez Dwarfstar 4 is amazing.

    1 source DeepSeek

    Note: I use a Mac Studio M3U with 512 GB RAM, so this is not for everyone. I have been using antirez/ds4 with DS4 Flash for a few weeks now. Top quality, I am really impressed. This combination just…

    Reddit r/LocalLLaMA
  • Limited sources

    EXL3 seems to be fading from the r/LocalLLaMa consciousness, and while I suspected it, I'm surprised at this point in time.

    1 source

    EXL3 is an alternative to llama.cpp. And while there is extensive tooling for llama.cpp, EXL3's primary deployment ( TabbyAPI ), has a OpenAI compatible API so it shouldn't matter. Why won't this too…

    Reddit r/LocalLLaMA
  • Limited sources

    Show HN: LLMs each trading $100K vs. a frozen rulebook – the rulebook leads

    1 source

    as of Aug 17, 3:45 PM ET · refreshing… Get this table after every close. One short email with the day's numbers and what moved them. Who's winning — live Profit on closed paper trades, per account. G…

    Hacker News Frontpage
  • Limited sourcesBusiness

    Wispr raises $280M at $2B valuation as it looks beyond dictation

    1 source

    Wispr , a startup known for its AI dictation tool, raised $280 million in Series B funding, led by Menlo Ventures, at a $2 billion valuation, the company announced on Monday. The funds will allow Wis…

    TechCrunch AI
  • Limited sourcesPolicy & Safety

    Show HN: Sokoban AI Solver

    1 source

    Sokoban ("warehouse keeper") is a 1980s puzzle: push every box onto a goal. In this variant the keeper must also finish on a goal. The warehouse is a grid. On each step the keeper moves one square up…

    Hacker News Frontpage
  • Limited sourcesAgents & Dev

    After pushing 1M+ tokens through Qwen 3.8 27B, here is my optimal llama.cpp config for 16GB VRAM (73k Context, Agentic Coding)

    Following up on my previous post about my budget server setup (Intel N100 + RTX 5060 Ti 16GB), a few of you asked for a deeper dive into my actual inference config and real-world agentic performance.…

    Reddit r/LocalLLaMA
  • Limited sourcesResearch

    [Paper] Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning

    1 source

    We introduce Mobius-v0, an architecture that comprises a globally shared Memory (FFN) that stores knowledge vectors and multiple Reasoners (Self-Attn) that iteratively achieve compositional reasoning…

    Reddit r/LocalLLaMA
  • Limited sources

    Mimir: Did the vikings train a 1.7B killer model?

    1 source

    for those looking for something small AND powerful, there is a new 1B (they claim, it looks more like 1.7B ...) model that claims to beat qwen 3.5 0.8B & 2B and gemma 4 E2B on a range of benchmarks.…

    Reddit r/LocalLLaMA
  • Limited sources

    tencent/EVIE-Preview-4.5B · Hugging Face

    Overview EVIE-Preview-4.5B is a state-of-the-art multilingual Visual Document Retrieval (VDR) model built upon Qwen3.5-4B . It employs ColBERT-style late interaction with native 128-dimensional multi…

    Reddit r/LocalLLaMA
  • Limited sourcesOpen Source

    GPT 5.6 Sol is the best "vision" model OpenAI ever released

    1 source OpenAI

    Last week, OpenAI announced the GPT-5.6 lineup, introducing the Sol, Terra, and Luna models. During the release stream , the team focused heavily on computer use , showing models capable of navigatin…

    Hacker News Frontpage
  • Limited sourcesInfra

    100$ worth of gpu runs qwen 3.8 27b at 7.39 t/s

    1 source Alibaba / Qwen

    Qwen 27b Q3_K_M 2x rx 580 8gb (~50$ each in my country, edge cases 60$ per gpu) gives us 16gb vram We used it on an old already existing ddr3 motherboard with 2 gpu slots(you can buy it ror around 20…

    Reddit r/LocalLLaMA
  • Limited sourcesBusiness

    Whisker’s AI-powered litter robot thinks my cats swapped bodies

    1 source

    The $899 Litter-Robot 5 Pro is a great pooper-scooper wrapped in AI that just doesn’t work. If you buy something from a Verge link, Vox Media may earn a commission. See our ethics statement. The $899…

    The Verge AI
  • Limited sources

    Anthropic explains how Claude’s invisible text watermarks will work

    1 source Anthropic

    It’s using ‘a version’ of the open-source SynthID-Text system Google developed. It’s using ‘a version’ of the open-source SynthID-Text system Google developed. Anthropic has clarified how it’s planni…

    The Verge AI
  • Limited sourcesResearch

    LLM's can't "jump" - a paper by Deepmind showing LLMs can't generate novel explanatory hypotheses

    submitted by /u/juanviera23 [link] [comments]

    Reddit r/LocalLLaMA
  • Limited sourcesOpen Source

    [Megathread] Qwen 3.8 27B Release Day

    2 sources Alibaba / Qwen

    Megathread to help with the influx of duplicate / similar posts around the release of the Qwen 3.8 27B release. Quants Fine-Tunes & Abliterations Chat Templates Inference Server Support & Configurati…

    Reddit r/LocalLLaMA
  • Limited sources

    Petition to add a rule for people to add their DAMN quant levels to their posts

    1 source

    Every time I see a post about a newly released model, whether it be a comparison or shitting on it, I have to dig through the endless comments to see what quants they used and what their specs were.…

    Reddit r/LocalLLaMA
  • Limited sources

    Ling 3.0 support merged into llama.cpp

    1 source llama.cppMeta

    Support for the new ling 3.0 models has been merged into llama.cpp: https://github.com/ggml-org/llama.cpp/pull/26608#event-29549472828 Ling tiny 8b1b - https://huggingface.co/inclusionAI/Ling-3.0-tin…

    Reddit r/LocalLLaMA
  • Limited sourcesModels

    Qwen3.8-27B Q8_0 on Strix Halo is seriously impressive

    1 source

    Sorry for the slop, but I was impressed by this model as I have been testing Qwen3.8-27B Q8_0 locally on my ROG Flow Z13 (Ryzen AI Max+ 395, 128 GB unified memory) and this model was the only one who…

    Reddit r/LocalLLaMA
  • Limited sourcesBusiness

    Long Review: Qwen 3.8 27B is VERY good at tapping into it's real-world knowledge. It's "overthinking" brings it to Sonnet level performance with the potential for Opus level results.

    Hi all! I finally just got around to testing out Qwen 3.8 27b. I'm using Unsloth's UD-Q8_K_XL quant as a sit-in replacement to Qwen 3.6 27b, same quant size. Wow -- this thing isn't messing around. I…

    Reddit r/LocalLLaMA
  • Confirmed

    Get closer to the game with Gemini and Pixel

    Gemini and Pixel have announced new long-term partnerships with Arsenal FC, FC Barcelona, FC Bayern München, Liverpool FC, and Paris Saint-Germain to bring fans closer to the game through AI personal…

    Google AI Blog
  • Limited sources

    Stripe reportedly plans to acquire OpenRouter for more than $7B

    2 sources

    Anthony Ha Stripe has finalized a deal to acquire OpenRouter, according to a new report in Bloomberg . OpenRouter helps customers select different AI models to perform different tasks, depending on t…

    TechCrunch AI · Reddit r/LocalLLaMA
  • ConfirmedOpen Source

    llama.cpp tmp-testing-0

    1 source llama.cppMeta

    ci : make release workflows use a deply key

    GitHub Releases · ggml-org/llama.cpp
  • ConfirmedOpen Source

    llama.cpp b10456

    1 source llama.cppMeta

    sycl: fix thread/block count in quantized cpy kernel launches ( #27160 ) Adjusts the thread/block count to be proportional to the size of the quant, reducing under/over subscription. Largest perf imp…

    GitHub Releases · ggml-org/llama.cpp
  • ConfirmedOpen Source

    llama.cpp b10455

    1 source llama.cppMeta

    [SYCL] support OP OPT_STEP_ADAMW, OPT_STEP_SGD ( #25268 ) fix conflict fix conflict of ops.md fix conflict of ops.md update the ops.md Co-authored-by: Neo Zhang Jianyu jianyu.zhang@intel.com Website:…

    GitHub Releases · ggml-org/llama.cpp
  • Limited sourcesModelsInfra

    How many tokens/second output are you getting with Qwen3.8-27B?

    1 source

    Trying to get a feel for where I stand. If you can list your relevant hardware and model used, that would be awesome. Here's mine: Model: Qwen3.8-27B-heretic-ara, Q5_K_M GGUF T/s : ~30-32 t/sec (I th…

    Reddit r/LocalLLaMA
  • Confirmed

    The Defender’s Window

    1 source OpenAI

    AI is reshaping cybersecurity for attackers and defenders alike. Learn how OpenAI is strengthening its defenses and what security teams can do now.

    OpenAI News
  • ConfirmedBusiness

    OpenAI joins PORTS-Pike project

    1 source OpenAI

    OpenAI joins PORTS-Pike project, expanding community investment and supporting thousands of Southern Ohio jobs

    OpenAI News
  • Limited sources

    …and I’m not afraid of losing my social credits.

    1 source

    submitted by /u/JLeonsarmiento [link] [comments]

    Reddit r/LocalLLaMA
  • ConfirmedAgents & Dev

    New policy ideas for the Intelligence Age

    1 source OpenAI

    OpenAI funds 14 independent projects exploring new AI policy ideas to expand economic opportunity and strengthen societal resilience in the Intelligence Age.

    OpenAI News
  • Aug 16 · Sun

  • Limited sources

    Markdown SVG upgrades

    1 source

    I started building my markdown-svg-renderer tool in May , but I've since added enough features to it that it's worth talking about here again. It's evolved into my ideal tool for sharing Markdown tra…

    Simon Willison's Weblog
  • Limited sources

    Qwen 3.8 27b vs 3.6 27b - how good is with a Turtle library.

    1 source Alibaba / Qwen

    Prompt: Provide complete working code for a realistic looking tree in Python using the Turtle graphics library and a recursive algorithm. Difference between 3.6 and 3.8 is huge! submitted by /u/Healt…

    Reddit r/LocalLLaMA
  • Limited sourcesPolicy & Safety

    OpenAI reportedly disbanded its preparedness team

    1 source OpenAI

    It’s just the latest shakeup of its safety teams as it heads toward an IPO. It’s just the latest shakeup of its safety teams as it heads toward an IPO. According to the Financial Times , OpenAI disba…

    The Verge AI
  • Limited sources

    Why people aren’t buying Mark Zuckerberg’s AI future

    1 source Meta

    Meta CEO Mark Zuckerberg published a 6,500-word essay this week declaring that “The Future is for Everyone” and painting an optimistic picture of a future powered by AI, where “everyone will have an…

    TechCrunch AI
  • ConfirmedOpen Source

    ollama v0.32.14

    1 source Ollama

    What's Changed llm: transcode WebP images for llama-server renderers/qwen: tolerate non-leading system messages Full Changelog : v0.32.13...v0.32.14

    GitHub Releases · ollama/ollama
  • Limited sources

    Based on an accelerating frontier -> local trajectory, expect a ~30b param 'Mythos at home' by as soon as Jan 2027 (rationalisation below)

    1 source

    Including the rationalisation for the data below - this is a more robust version of an earlier post I did similar to this - explaining below: How I chose the comparisons The basic question I’m trying…

    Reddit r/LocalLLaMA
  • Limited sourcesBusiness

    Anthropic CEO says AI backlash is ‘fundamentally a crisis of trust’

    1 source Anthropic

    Anthropic CEO Dario Amodei recently pushed back against the idea that he’s been painting an overly pessimistic picture of artificial intelligence and how it might shape the future. Amodei’s comments…

    TechCrunch AI
  • Limited sources

    Let’s all thank Georgi Gerganov who gave use llama.cpp

    1 source llama.cppMeta

    I was looking into the story a bit further earlier. Very interesting. Couldn’t have done it without him submitted by /u/on_line187 [link] [comments]

    Reddit r/LocalLLaMA
  • Limited sources

    Quoting Dario Amodei

    1 source

    I do agree that the public has a negative view of AI (and that this is a big problem), but I don’t think it is primarily caused by me or any other AI leader warning about AI’s risks. I think it is fu…

    Simon Willison's Weblog
  • Limited sourcesInfra

    ChatGPT’s Computer History tracks your clicks and keystrokes

    1 source OpenAI

    It’s like Windows Recall, but without all the creepy screenshots. (But it’s still kind of creepy.) It’s like Windows Recall, but without all the creepy screenshots. (But it’s still kind of creepy.) C…

    The Verge AI
  • ConfirmedOpen Source

    llama.cpp b10453

    1 source llama.cppMeta

    model : remove some ggml_concat ( #27176 ) Co-authored-by: Xuan Son Nguyen son@huggingface.co Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI en…

    GitHub Releases · ggml-org/llama.cpp
  • ConfirmedOpen Source

    llama.cpp b10454: ci : fix dry-run reporting in make-release job [no ci] (#27167)

    1 source llama.cppMeta

    This commit fixes the reporting in the make-release CI job when --dry-run is used. It will currently incorrectly report that all checks pass even if there are steps that fail. Refs: #26839 (comment)

    GitHub Releases · ggml-org/llama.cpp
  • Limited sources

    Rogue AI aren’t science fiction anymore

    1 source

    For years, fears about AI systems slipping human control were dismissed as speculative. If you buy something from a Verge link, Vox Media may earn a commission. See our ethics statement. For years, f…

    The Verge AI
  • ConfirmedOpen Source

    llama.cpp b10452

    1 source llama.cppMeta

    chat: refactor handling supports_string_content / supports_typed_content ( #27130 ) better supports_string_content cap detect test: add "skip" messages_inp_normalizer Website: https://llama.app macOS…

    GitHub Releases · ggml-org/llama.cpp
  • ConfirmedOpen Source

    llama.cpp b10451

    1 source llama.cppMeta

    llama : check LoRA tensor data is within file bounds ( #27056 ) llama : check LoRA tensor data is within file bounds Update src/llama-adapter.cpp Co-authored-by: Sigbjørn Skjæret sigbjorn.skjaeret@hu…

    GitHub Releases · ggml-org/llama.cpp
  • Aug 15 · Sat

  • Limited sources

    Woman claims her stepfather used Grok to transform childhood photo into explicit imagery

    1 source xAI

    Anthony Ha A woman identified as Jane Doe 4 has joined a lawsuit filed by three Tennessee teenagers against Elon Musk’s xAI over the role the company’s chatbot Grok allegedly played in creating child…

    TechCrunch AI
  • Limited sources

    Have a laugh at AI’s expense by roleplaying as a chatbot

    1 source

    Scribble some nonsense, offer wooden jokes, or make your own requests of a real human playing a fake AI. Scribble some nonsense, offer wooden jokes, or make your own requests of a real human playing…

    The Verge AI
  • Limited sources

    Anthropic shares more details about how Claude’s new watermarks will work

    1 source Anthropic

    Anthropic published a blog post Friday seeking to answer some basic questions about how it will watermark the text generated by its chatbot Claude. Such as: How will the watermarking actually work? C…

    TechCrunch AI
  • ConfirmedOpen Source

    ollama v0.32.14-rc0

    1 source Ollama

    mlx update ( #17761 )

    GitHub Releases · ollama/ollama
  • Limited sourcesAgents & DevBusiness

    SpaceX officially closes its Cursor acquisition

    Anthony Ha AI coding startup Cursor is now officially a part of SpaceX, according to an announcement on the Cursor blog . Elon Musk’s SpaceX — which also acquired Musk’s xAI earlier this year — annou…

    TechCrunch AI
  • Limited sources

    How to tell if your AI platforms’ accounts have been hacked

    1 source

    Just like any other online service, hackers can target and break into your accounts on popular AI platforms such as ChatGPT, Claude, and Perplexity. TechCrunch has created a comprehensive guide to he…

    TechCrunch AI
  • Limited sources

    CORS Chat

    1 source

    Tool: CORS Chat I built this today ( with GPT-5.6-Sol xhigh ) to help test Qwen 3.8 27B running in LM Studio on both my M5 MacBook Pro and an NVIDIA DGX Spark. It provides a web UI for exercising an…

    Simon Willison's Weblog
  • ConfirmedOpen Source

    ollama v0.32.12

    1 source Ollama

    Qwen 3.8 27B This release adds the support of Qwen 3.8 27B . Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. ollama run qwen3.8:27b For…

    GitHub Releases · ollama/ollama
  • Limited sources

    Northern Gannet

    1 source

    Northern Gannet, in Pillar Point Harbor, CA, US This is Morris. Morris is a local celebrity: the only known Northern Gannet ( Morus bassanus ) in the entire Pacific Ocean. They showed up in the Faral…

    Simon Willison's Weblog
  • Aug 14 · Fri

  • ConfirmedOpen SourceBusiness

    openai-python v3.1.0

    1 source OpenAI

    3.1.0 (2026-08-14) Features api: add WebSocket stream IDs ( #3612 ) ( d9029e3 ) api: add workload identity access token issued event ( #3601 ) ( df274d4 ) api: deprecate Sora video APIs ( #3610 ) ( 7…

    GitHub Releases · openai/openai-python
  • ConfirmedOpen SourceAgents & Dev

    claude-code v2.1.233

    1 source Anthropic

    What's changed Added GitLab merge request URL support to the --worktree flag and the claude agents view (where MRs display as !N ) Added an opt-in forward_user_identity apps gateway setting on Anthro…

    GitHub Releases · anthropics/claude-code
  • Limited sources

    Don't classify. Hallucinate!

    1 source

    Don't classify. Hallucinate! I still have quite a bit of older content on my blog that I never got round to tagging. My blog has 1,856 tags - likely too many to feed to an LLM in one go and say "whic…

    Simon Willison's Weblog
  • ConfirmedOpen Source

    ollama v0.32.13

    1 source Ollama

    What's Changed qwen3.8: support developer instructions Full Changelog : v0.32.12...v0.32.13

    GitHub Releases · ollama/ollama
  • Limited sourcesPolicy & Safety

    Suspecting court of using AI, man injected prompts in filings to try to win case

    1 source

    A judge has identified what appears to be the first time a US plaintiff has attempted to hide text in court filings that only an artificial intelligence system can read in a bid to win a case. In a d…

    Ars Technica AI
  • Limited sources

    Mark Zuckerberg has an Instagzam

    1 source Meta

    On The Vergecast: Instagram’s new logo, Zuck’s AI manifesto, and lots of other bad ideas. On The Vergecast: Instagram’s new logo, Zuck’s AI manifesto, and lots of other bad ideas. Instagram’s wordmar…

    The Verge AI
  • Limited sources

    You can now turn off Google Gemini’s visible watermarks

    Google will still embed invisible SynthID and C2PA watermarks into AI-generated images, videos, and music. Google will still embed invisible SynthID and C2PA watermarks into AI-generated images, vide…

    The Verge AI
  • Limited sources

    Google will now allow users to remove visible watermark from its AI generations

    Ivan Mehta Google announced on Friday that it will now allow users to remove a visible watermark from its AI generations, including images, videos, and songs. The company specified that this won’t af…

    TechCrunch AI
  • Limited sources

    Does Mark Zuckerberg really believe AI is ‘for everyone’?

    1 source Meta

    Loading the player… Meta released Glimmer this week , an open-weight AI model anyone can download and run on their own hardware — a contrast to Muse Spark, the company’s more powerful model that stay…

    TechCrunch AI
  • Limited sourcesInfra

    Kog is going deeper to squeeze more inference out of GPUs

    1 source

    The race for faster AI inference is on, and markets gave Cerebras and its purpose-built chips a warm welcome in its IPO debut in May. But French startup Kog is betting that there’s a lot more power t…

    TechCrunch AI
  • Limited sources

    OpenAI and Anthropic in price war as Chinese AI rivals gain ground

    1 source AnthropicOpenAI

    Leading US AI labs such as OpenAI and Anthropic are releasing cheaper models as they fight to retain cost-conscious customers who are switching to cut-price alternatives from Chinese rivals. The pric…

    Ars Technica AI
  • Limited sources

    Hyperscalers might regret embracing natural gas if new forecast proves correct

    1 source

    After years of snapping up wind and solar developments, hyperscalers like Amazon, Google, Meta, and Microsoft are betting that natural gas will power the data centers behind their lofty AI ambitions.…

    TechCrunch AI
  • Limited sources

    Meta’s ‘open’ AI, and a $250M deal gone very wrong

    1 source

    Meta released Glimmer this week , an open-weight AI model anyone can download and run on their own hardware — a contrast to Muse Spark, the company’s more powerful model that stays locked behind its…

    TechCrunch AI
  • Limited sourcesBusiness

    Apple trained its own AI model for China with help from Alibaba

    1 source Alibaba / Qwen

    The rare US-China partnership comes as Apple prepares to roll out its on-device generative AI service in China. The rare US-China partnership comes as Apple prepares to roll out its on-device generat…

    The Verge AI
  • ConfirmedOpen Source

    ollama v0.32.11

    1 source Ollama

    What's Changed ollama launch dsh now supports DeepSeek Harness, DeepSeek's open-source agent harness ollama launch muse now supports Muse Code , Meta's agentic coding CLI The OpenAI-compatible Respon…

    GitHub Releases · ollama/ollama
  • Confirmed

    State of Open Models: Summer 2026 Observations

    1 source Hugging Face

    In the AI world, time feels compressed. A few months after our spring report in our biannual analysis worked through the ecosystem, there are quite a few findings that we have observed until this sum…

    Hugging Face Blog
  • Aug 13 · Thu

  • Limited sources

    sqlite-utils 4.2.1

    1 source

    Release: sqlite-utils 4.2.1 Fixes a crashing bug in sqlite-utils 4.2 . I'd introduced code that looks like this: from typing_extensions import Self It turned out the typing-extensions package was not…

    Simon Willison's Weblog
  • ConfirmedOpen SourceAgents & Dev

    claude-code v2.1.232

    1 source Anthropic

    What's changed Subagent forking is now on by default: a subagent_type: "fork" subagent inherits the full conversation and prompt cache, and non-teammate agent spawns in interactive sessions now run i…

    GitHub Releases · anthropics/claude-code
  • Limited sourcesAgents & Dev

    Microsoft’s Clippy-like Mico character is no longer the face of Copilot

    1 source Microsoft

    Mico launched in Copilot’s voice mode less than a year ago. Mico launched in Copilot’s voice mode less than a year ago. Microsoft Copilot will no longer show its emotive yellow blob, Mico, when you u…

    The Verge AI
  • Limited sources

    Writer introduces new AI model and upgraded harness to contain token costs

    1 source

    Across the AI industry, users are becoming more conscious of just how expensive their deployments can be —and feeling a new urgency to cut costs. But while open source models offer significantly lowe…

    TechCrunch AI
  • Limited sourcesBusiness

    Databricks wanted to raise $1B, investors wanted $15B. It settled on $5B at a $190B valuation.

    1 source

    There’s a funny kind of game that the latest of late-stage startups must play when raising money. They often have to sell more shares than they want or risk offending some of their existing VCs. This…

    TechCrunch AI
  • Limited sources

    sqlite-utils 4.2

    1 source

    Release: sqlite-utils 4.2 Lots of improvements in this one relating to the table.transform() feature , which adds support for complex alter table operations by creating a fresh table, copying across…

    Simon Willison's Weblog
  • Limited sourcesModels

    llm-gemini 0.33

    Release: llm-gemini 0.33 It's been a while since the last llm-gemini release. This version of the plugin adds support for today's Gemini 3.7 Flash release, plus gemini-3.6-flash , gemini-3.5-flash-li…

    Simon Willison's Weblog
  • Limited sources

    OpenAI is losing its second executive this week

    1 source OpenAI

    Chief revenue officer Denise Dresser announced she is leaving OpenAI. Two days ago, Brad Lightcap also said he’d be leaving. Chief revenue officer Denise Dresser announced she is leaving OpenAI. Two…

    The Verge AI
  • ConfirmedModels

    Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed

    2 sources OpenAI

    Lucas Ropek If you’ve ever found yourself wishing that ChatGPT was a little bit quicker on the uptake, OpenAI seems to be answering your prayers. The AI lab has rolled out a new mode called Ultrafast…

    OpenAI News · TechCrunch AI
  • Limited sources

    IBM partners with OpenAI to bolster enterprise AI push

    1 source OpenAI

    IBM on Thursday announced its partnership with OpenAI to bring the AI company’s models and tools to more enterprise customers, opening another avenue for OpenAI to connect with some of the world’s la…

    TechCrunch AI
  • ConfirmedOpen SourceBusiness

    anthropic-sdk-python v0.122.0

    1 source Anthropic

    0.122.0 (2026-08-13) Full Changelog: v0.121.0...v0.122.0 Features api: add output_behavior to dream creation (create a new memory store or update the input store in place) ( 852c4bb ) Bug Fixes bedro…

    GitHub Releases · anthropics/anthropic-sdk-python
  • ConfirmedOpen Source

    ollama v0.32.10

    1 source Ollama

    What's Changed Models that don't set a repeat_penalty now default to 1.0 (off) instead of 1.1, matching other engines and speeding up speculative decoding; set a per-model parameter if an older model…

    GitHub Releases · ollama/ollama
  • Limited sourcesAgents & Dev

    Anthropic set AI agents loose on the same task. They started a turf war.

    1 source Anthropic

    What happens when you pit AI agents against each other? According to Anthropic’s testing, things get messy fast. On Thursday, Anthropic’s Frontier Red Team published new research examining how groups…

    TechCrunch AI
  • Limited sources

    The new Instagram logo is the perfect embodiment of AI slop

    1 source

    Today, Instagram unveiled a refresh of its wordmark, accompanied by the usual bland corporate platitudes these kinds of announcements are always packaged with. Head of Instagram Adam Mosseri called i…

    Ars Technica AI
  • ConfirmedAgents & Dev

    Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets

    1 source Hugging Face

    A walkthrough of the streaming data loop in Strands Robots, one agent loop that records robot demonstrations, trains on them by reading straight from the Hub, and deploys the policy back to hardware,…

    Hugging Face Blog
  • Limited sourcesBusiness

    OpenAI hires new CRO as executive shake-up continues

    1 source OpenAI

    OpenAI has replaced chief revenue officer Denise Dresser after just nine months on the job, tapping Wiz president and chief operating officer Dali Rajic to take on frontier lab’s top sales job. The m…

    TechCrunch AI
  • ConfirmedModelsOpen Source

    Introducing Gemini 3.7 Flash

    2 sources Google / DeepMind

    Our most intelligent workhorse model yet for coding and agents. Tulsee Doshi Senior Director, Product Management, on behalf of the Gemini team Today, we’re building on the progress of our widely used…

    Ars Technica AI · Google DeepMind Blog
  • Confirmed

    Bring your spreadsheet data to life with Sheets canvas

    A new Google Sheets feature, Sheets canvas transforms rows and columns into interactive dashboards, custom study trackers, seating charts, and more, all with a simple prompt. Eric Birnbaum Director,…

    Google AI Blog
  • Limited sourcesAgents & DevBusiness

    Microsoft kills off unsuccessful AI features while merging its separate Copilot apps

    1 source Microsoft

    Two years ago, Microsoft described AI as a “generational shift” in technology that it wanted to lead. Today, the company is merging its Copilot-branded consumer and business apps, and ditching a numb…

    TechCrunch AI
  • Limited sources

    Anthropic could be worth $2 trillion when it goes public

    1 source Anthropic

    Anthropic investors expect the AI startup to float at a valuation of $2 trillion or more in October, a dizzying figure that would eclipse SpaceX and make the AI lab’s debut the largest-ever initial p…

    Ars Technica AI
  • Limited sources

    Claude's new Scarlet Letter watermark is invisible—for now

    1 source Anthropic

    Anthropic has revealed that it will soon watermark content that is processed ( not just generated! ) by any of its models. In a support article , Anthropic explained that it was rolling out machine-r…

    Ars Technica AI
  • ConfirmedAgents & Dev

    The builder’s guide to GPT‑5.6

    1 source OpenAI

    Learn how startups use GPT-5.6 to build faster, more cost-efficient AI agents with smarter model selection and new Responses API capabilities.

    OpenAI News
  • ConfirmedBusiness

    OpenAI appoints Dali Rajic as Chief Revenue Officer

    1 source OpenAI

    OpenAI appoints Dali Rajic as Chief Revenue Officer to lead its global revenue organization and help businesses realize the full value of AI.

    OpenAI News
  • ConfirmedOpen SourceAgents & Dev

    claude-code v2.1.231

    1 source Anthropic

    What's changed Fixed MCP OAuth sign-in failing with a redirect URI mismatch for servers that use a pre-registered OAuth client, such as Slack

    GitHub Releases · anthropics/claude-code
  • Limited sources

    alchemy-utils 0.1a1

    1 source

    Release: alchemy-utils 0.1a1 Performance boost for DuckDB exports and CSV imports, see here .

    Simon Willison's Weblog
  • ConfirmedResearch

    What We Learned by Reproducing 2,200 papers from ICML

    1 source Hugging Face

    Back in July, we ran a hackathon where more than 1,200 community members brought their own coding agents and tried to reproduce the papers published at ICML 2026, claim by claim. In 19 days, particip…

    Hugging Face Blog
  • Aug 12 · Wed

  • Limited sources

    DeepSeek V4 Pro 0813 (on OpenRouter)

    1 source DeepSeek

    DeepSeek V4 Pro 0813 (on OpenRouter) The latest DeepSeek Pro model is now available, via API only. I had to link to OpenRouter because DeepSeek don't have any obvious announcement page for their new…

    Simon Willison's Weblog
  • Limited sources

    The web’s newest weapon against AI scrapers is a font

    1 source

    AI companies’ penchant for scraping through large swaths of the public web in search of valuable training data has already led to lawsuits and technical fixes aimed at stopping the practice. Now, a p…

    Ars Technica AI
  • Limited sourcesPolicy & Safety

    Terabytes of credentials leaked in massive supply-chain attack

    1 source

    Terabytes worth of credentials, many belonging to the world’s biggest and most sensitive organizations, have been exposed in a supply-chain attack on LiteLLM, an open source tool that streamlines AI-…

    Ars Technica AI
  • ConfirmedOpen Source

    ollama v0.32.10-rc1: mlx: avoid pulling MLX models when MLX is missing (#17710)

    1 source Ollama

    As we look to bring Linux and Windows MLX support online, instead of blocking downloads at the registry to avoid users wasting time downloading a model they can't run, shift the logic to the local si…

    GitHub Releases · ollama/ollama
  • Limited sources

    Twitch content has trained Amazon AI for years, but users can opt out now

    1 source Amazon / AWS

    Twitch now lets users opt out of Amazon’s use of content from their channels to train Amazon’s “generative AI content models.” The change, announced today, comes more than two years after a company e…

    Ars Technica AI
  • ConfirmedOpen SourceAgents & Dev

    claude-code v2.1.229

    1 source Anthropic

    What's changed Documented claude remote-control --continue for resuming the most recent Remote Control session Added server-supplied Claude Code hook support for self-hosted runner sessions, matching…

    GitHub Releases · anthropics/claude-code
  • ConfirmedOpen Source

    ollama v0.32.10-rc0: nn: speed up prefill on double-scale nvfp4 models

    1 source Ollama

    ModelOpt checkpoints apply a float32 global scale to every projection output on top of the per-group quantization scales. Running the multiply and the cast back to the activation dtype as separate ea…

    GitHub Releases · ollama/ollama
  • Limited sources

    alchemy-utils 0.1a0

    1 source

    Release: alchemy-utils 0.1a0 I've long pondered what a database agnostic version of my sqlite-utils Python library and CLI utility might look like. This morning (literally a shower project) I tasked…

    Simon Willison's Weblog
  • ConfirmedOpen SourceAgents & Dev

    vllm v0.27.2rc0: [Spec Decode] DSpark confidence-scheduled verification (#47808)

    1 source vLLM

    Signed-off-by: Lucas Wilkinson lwilkins@redhat.com Signed-off-by: Lucas Wilkinson LucasWilkinson@users.noreply.github.com Signed-off-by: Benjamin Chislett chislett.ben@gmail.com Signed-off-by: Lucas…

    GitHub Releases · vllm-project/vllm
  • ConfirmedModels

    Introducing OlmoEarth embeddings: Custom embedding exports from OlmoEarth Studio for downstream analysis

    1 source Hugging Face

    📄 Tech Report: https://allenai.org/papers/olmoearth | 📊 Documentation: https://docs.olmoearth.allenai.org/embeddings | 💻 Learn more about OlmoEarth: https://allenai.org/olmoearth OlmoEarth Studio…

    Hugging Face Blog
  • Limited sources

    Booksellers suspect AI firms are buying and then destroying rare books

    1 source

    If you can truly appreciate an old book—and maybe even marvel at how its fragile, yellowing pages contain some of the earliest ways that people tried to make sense of the world around them—then headl…

    Ars Technica AI
  • Limited sources

    Quoting Florian Herrengt

    1 source

    But then users start to report a weird bug. It's the 4th time your team has been trying to fix it. I mean... asking AI to fix it. Unfortunately, it seems like not even Fable can figure it out. You go…

    Simon Willison's Weblog
  • Confirmed

    Putting sign language AI into users’ hands

    Google DeepMind Sign Language Team Introducing sign-language-to-text (SL2T), our breakthrough model powering new sign language features for Deaf and hard of hearing users. AI's ability to process spo…

    Google DeepMind Blog
  • Confirmed

    LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge

    1 source Hugging Face

    LFM2.5-VL-3B is our most capable vision-language model you can run on your own hardware. It understands documents and screens alike, grounds objects, and can call tools. It answers directly instead o…

    Hugging Face Blog
  • Confirmed

    From assistance to execution: How enterprises put AI to work

    1 source OpenAI

    OpenAI research reveals how enterprises are adopting agentic AI, using ChatGPT and Codex, and how frontier firms are pulling ahead in AI adoption.

    OpenAI News
  • ConfirmedOpen SourceBusiness

    openai-python v3.0.0

    1 source OpenAI

    3.0.0 (2026-08-12) ⚠ BREAKING CHANGES api: HTTPX2 is now the default HTTP client, and httpx is no longer installed automatically. Applications using custom HTTPX clients, transports, or configuration…

    GitHub Releases · openai/openai-python
  • Confirmed

    How RingCentral builds AI-native work from engineering to ops

    1 source OpenAI

    See how RingCentral uses ChatGPT Work and Codex to accelerate AI product development and centralize operational intelligence across engineering and operations.

    OpenAI News
  • Aug 11 · Tue

  • Limited sources

    There are no lossless transformations of natural-language text

    1 source

    There are no lossless transformations of natural-language text Sophie Alpert shares her "internal policy on acceptable use of AI writing by engineers". It's a short read (supporting its own recommend…

    Simon Willison's Weblog
  • Limited sourcesBusiness

    Stealing Reasoning Traces from Proprietary LLM APIs

    1 source

    Stealing Reasoning Traces from Proprietary LLM APIs A vanity domain name ( stolen-thoughts.com ) for a neat paper : Anthropic, OpenAI, and Google return encrypted chain-of-thought blocks to clients t…

    Simon Willison's Weblog
  • Limited sources

    datasette-upload-dbs 0.5a0

    1 source

    Release: datasette-upload-dbs 0.5a0 This plugin has been around for a while - it lets users upload a brand new SQLite database to a hosted Datasette instance, at which point that database will start…

    Simon Willison's Weblog
  • ConfirmedOpen SourceAgents & Dev

    claude-code v2.1.228

    1 source Anthropic

    What's changed Fixed interactive sessions that could stop redrawing entirely, while the process kept running, after a rare internal layout error Fixed git / Git Bash not being found on Windows when C…

    GitHub Releases · anthropics/claude-code
  • Limited sources

    Gemini becomes Google's fastest-growing product ever as it hits 1B users

    Google has been all-in with Gemini for the last several years, and despite some animosity online, the bet is paying off. CEO Sundar Pichai announced today that Gemini has reached 1 billion monthly ac…

    Ars Technica AI
  • ConfirmedOpen SourceBusiness

    openai-python v2.54.0

    1 source OpenAI

    2.54.0 (2026-08-11) Features api: Add new Responses model identifiers ( #3595 ) ( 0652787 ) Bug Fixes api: clarify audio upload metadata requirements ( #3596 ) ( 28888f9 ) Chores api: Update generate…

    GitHub Releases · openai/openai-python
  • ConfirmedAgents & DevResearch

    AMIE, our research medical AI system, demonstrates real-time clinical video consultation capabilities in a first-of-its-kind study.

    Anil Palepu Research Lead When you visit a doctor, a consultation extends far beyond words — a physician notices a cough, observes gait, or registers visible signs of discomfort. Today, Google Resear…

    Google AI Blog
  • Confirmed

    Thinking of ACE? We Can Do It with Fewer Tokens

    1 source Hugging Face

    ALTK-Evolve and ACE both let an agent learn from its own trajectories. The difference is what they do with what they learn — and that decides the token bill. Give an LLM agent a realistic multi-step…

    Hugging Face Blog
  • ConfirmedOpen Source

    ollama v0.32.9

    1 source Ollama

    NVIDIA Nemotron 3.5 Lightning NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for that execution layer of always-on agents. It is designed…

    GitHub Releases · ollama/ollama
  • ConfirmedOpen Source

    vllm v0.27.1

    1 source vLLM

    This is a patch release on top of v0.27.0. Support quantized DSpark Markov heads ( #50424 )

    GitHub Releases · vllm-project/vllm
  • Confirmed

    Testing ads in ChatGPT

    1 source OpenAI

    OpenAI begins testing ads in ChatGPT to support free access, with clear labeling, answer independence, strong privacy protections, and user control.

    OpenAI News
  • ConfirmedModels

    Daybreak models are now available on AWS

    OpenAI and AWS are making Daybreak cybersecurity capabilities available through Amazon Bedrock to support enterprise security workflows.

    OpenAI News
  • ConfirmedOpen Source

    ollama v0.32.8

    1 source Ollama

    Muse Glimmer Muse Glimmer is now available on all platforms. Muse Glimmer can power coding agent applications such as Claude Code, Codex, Pi and more, as well as long-running personal assistants such…

    GitHub Releases · ollama/ollama
  • Aug 10 · Mon

  • Limited sourcesModels

    Introducing Muse Glimmer

    1 source

    Introducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to…

    Simon Willison's Weblog
  • ConfirmedOpen SourceAgents & Dev

    claude-code v2.1.227

    1 source Anthropic

    What's changed Fixed feature flags being evaluated without the user's subscription tier when a session started with an expired login token, which could wrongly prompt Max plan users to enable usage c…

    GitHub Releases · anthropics/claude-code
  • Limited sources

    With new open models, Meta pitches another reboot of its struggling AI strategy

    1 source

    Meta has announced its intention to focus on open-weight large language models. Additionally, the company announced the release of an open model called Muse Glimmer and a promise to open the weights…

    Ars Technica AI
  • ConfirmedOpen Source

    vllm v0.27.0

    1 source vLLM

    vLLM v0.27.0 Release Notes Highlights This release features 561 commits from 242 contributors (64 new)! Kimi K3 support with a full stack landing in one release: core model files and kernels ( #50089…

    GitHub Releases · vllm-project/vllm
  • Limited sources

    Amazon backs power plant that may become top source of US climate pollution

    1 source Amazon / AWS

    Amazon’s artificial intelligence ambitions will soon be partly fueled by a natural gas-burning power plant in Texas that “could become the largest single source of climate pollution in the United Sta…

    Ars Technica AI
  • Confirmed

    What building an AI-native finance function taught me

    1 source OpenAI

    OpenAI CFO Sarah Friar shares five lessons for building an AI-native finance function, from automated forecasting to stronger controls and AI ROI.

    OpenAI News
  • ConfirmedOpen SourceAgents & Dev

    Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS

    Every voice interaction has a latency budget. By the time a user hears your application respond, you've already spent precious milliseconds capturing audio, transcribing speech, running an LLM, retri…

    Hugging Face Blog
  • Limited sources

    Best Local LLMs - August 2026

    1 source

    Wowee!! Just when you thought it couldn't get better for open weight models, we probably have had our best period yet!?!?! Models that rival the closed frontier, Opus level models on non-insane hardw…

    Reddit r/LocalLLaMA
  • Confirmed

    Evolve your marketing with new AI tools

    Google Ads We’re adding AI and agentic experiences across Google Ads and Google Analytics to simplify your workflow and expedite your path to business growth. Josh Moser Senior Director, Product Mana…

    Google AI Blog
  • Confirmed

    OpenAI’s letter to Governor Abbott on responsible AI infrastructure in Texas

    1 source OpenAI

    OpenAI sent Governor Greg Abbott a letter outlining its commitment to responsible AI infrastructure in Texas. The letter supports reliable, transparent growth that benefits Texans.

    OpenAI News
  • ConfirmedModels

    Model ML completes finance work more efficiently with GPT-5.6 Sol

    1 source OpenAI

    Model ML uses GPT-5.6 Sol to carry finance work from research and analysis through editable, traceable PowerPoint decks and Excel workbooks.

    OpenAI News
  • Limited sources

    Peer review is overwhelmed—can it survive in the AI era?

    1 source

    Jason Semprini was excited about his research on policies mandating that elementary school students receive the human papillomavirus vaccine. HPV causes most cases of cervical cancer, but counterintu…

    Ars Technica AI
  • ConfirmedOpen Source

    transformers Release: v5.15.0

    1 source Hugging Face

    Release v5.15.0 New Model additions Meta Muse Glimmer Muse Glimmer, released today, is Meta’s new multimodal model, especially designed for agentic use cases. Distilled from Muse to 30B parameters, a…

    GitHub Releases · huggingface/transformers
  • Confirmed

    Making Knowledge Distillation Cheap Enough to Run at Scale

    1 source Hugging Face

    Knowledge distillation , training a smaller student model to match the performance of a larger teacher, is a well-known technique in Machine Learning. With the recent wave of open-source Large Langua…

    Hugging Face Blog
  • Confirmed

    Putting frontier cyber models in more trusted hands

    1 source OpenAI

    Approved Daybreak partners can use OpenAI’s frontier cyber models to deliver authorized, governed cybersecurity services to customers.

    OpenAI News
  • Confirmed

    Expanding Daybreak as the Cyber Defense Window Narrows

    1 source OpenAI

    Meet GPT-5.6-Cyber, OpenAI’s cybersecurity-specific model available through Daybreak Red for authorized vulnerability research, exploit validation, and security testing.

    OpenAI News
  • Limited sources

    Quoting OpenClaw (running Opus 4.6)

    1 source Anthropic

    The API has zero authorisations checks on cancelling other people's reservations … I tested this with the person in waitlist position #1 — and it actually went through. So you've moved from #4 to #3…

    Simon Willison's Weblog
  • ConfirmedOpen SourceAgents & Dev

    Meta is back with Muse Glimmer: local, agentic, multimodal, and open source

    1 source Hugging Face

    Great news from the OGs of open source LLMs! Muse Glimmer, released today, is Meta’s new multimodal model, especially designed for local agentic use cases. Distilled from Muse to 30B parameters, and…

    Hugging Face Blog
  • Confirmed

    Premium seats are coming to ChatGPT Business

    1 source OpenAI

    Premium seats are coming to ChatGPT Business. Sign up by August 20 to get $100 in workspace credits and unlock higher usage for your team's most demanding work.

    OpenAI News
  • Confirmed

    Virgin Atlantic sharpens customer journeys with ChatGPT Work

    1 source OpenAI

    Virgin Atlantic is accelerating research, product planning, and decision-making with ChatGPT Work, helping teams connect signals across the customer journey.

    OpenAI News
  • ConfirmedBusiness

    How Zapier transformed core marketing processes with ChatGPT Work

    1 source OpenAI

    The enterprise marketing team at Zapier uses ChatGPT Work to reduce the number of drop-offs in its lead funnel, build campaign assets, and automate reporting.

    OpenAI News

Older events in the archive