2026-08 240 events
Today
-
Limited sourcesModels
AA is the reason for Qwen3.8 27B shipped with xhigh
I know why Qwen3.8 27B shipped with xhigh reasoning as default, it's to do its best in benchmarks. Models from top labs often get benchmarked at multiple reasoning levels, but that same treatment doe…
Reddit r/LocalLLaMA -
ConfirmedOpen Source
llama.cpp b10481: CUDA: MMVQ nwarps=8 for bs=1 for dense models on DGX Spark (#26843)
CUDA: MMVQ nwarps=8 for bs=1 for dense models on DGX Spark Signed-off-by: ynankani ynankani@nvidia.com skip moe experts and allow others based on k geometry (allow only small idle tail) Signed-off-by…
GitHub Releases · ggml-org/llama.cpp -
Limited sources
Qwen dev says not to wait for 35B-A3B
What does this mean? Is there something else coming? Maybe 122B? Or no models? submitted by /u/Mean-Ad1493 [link] [comments]
Reddit r/LocalLLaMA -
Limited sources
Sparks jumping in price too
Just had an eBay seller cancel my order at 8300 AUD to relist at 9K flat. This is not a fly by night operator either, it’s a legit org. They said they are going to honour the previous pricing when th…
Reddit r/LocalLLaMA -
Limited sourcesModelsAgents & Dev
Optimizing Qwen3.6 / Qwen3.8-27B on 16GB VRAM: Complete Benchmark Results and Setup Guide (~30-50tps at 32k to 72k context)
This post was made with AI. I tried to remove as much slop as possible and keep it straight to the point to save your time as I know how annoying AI slop posts can be, but I still wanted to retain al…
Reddit r/LocalLLaMA -
Limited sources
CDW has bumped the MSRP of the RTX Pro 6000 from $16,000 to $19,999
Did they slip up and leak future pricing? Live link: https://www.cdw.com/product/pny-nvidia-rtx-pro-6000-graphic-card-96-gb-gddr7/8326705 Archive link: https://web.archive.org/web/20260818013250/http…
Reddit r/LocalLLaMA -
Limited sources
Made this game in two prompts with Q4, Qwen 3.8 is amazing
This took one prompt to build, and another follow up prompt to fix two issues (player got stuck with the bomb and broken enemies path-finding), this is only html, css and js, no external assets, all…
Reddit r/LocalLLaMA Yesterday
-
Limited sources
Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index
Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index That's the same score as GPT-5.6 Luna (max), and just one point behind GLM-5.2 (max) and DeepSeek V4 Pro 0813 (max) - that GLM is…
Simon Willison's Weblog -
Limited sources
Anthropic’s annualized revenue surges to $65B
Marina Temkin Anthropic’s revenue continues to not only grow at an historic pace but also to accelerate. The model maker’s annualized revenue run rate — a projection of a full year’s revenue based on…
TechCrunch AI -
Limited sources
Qwen 3.8 distillations
Artificial analysis index scores Qwen 3.5 27b: 35 Qwen 3.6 27b: 38 Qwen 3.8 27b: 52 What the hell kind of a jump was that? Even if it is benchmaxxed, the jump is insane. Qwen3.6 35b A3b: 32 That's ~6…
Reddit r/LocalLLaMA -
Limited sources
Qwen 3.8 35b and 122b - We hope/wait/beg for models incessantly. But how do we actually give the lab more incentive to make it?
So many comments begging for these models, and I get it. People from the labs frequent r/LocalLLaMA , so maybe the begging comments aren't useless. They make demand known at least. Same with polls, e…
Reddit r/LocalLLaMA -
Limited sourcesResearch
we benchmark models nobody actually runs
qwen3.8-27b looks genuinely impressive on the benchmark tables - beating models many times its size on some of them. but those numbers come from bf16 weights, and nobody here is running a 27b at bf16…
Reddit r/LocalLLaMA -
Limited sources
My friends all hate AI; I just joined an AI startup
“I know we’re all anti-AI here,” a friend texted in the group chat. “We should co-write an article on how AI has no place in education,” a professional acquaintance suggested during a collaboration s…
Hacker News Frontpage -
Limited sourcesBusiness
AI automation startup Relay shuts down, staff joins Google’s Chrome team
Relay, an AI-powered workflow automation tool that was launched in 2021 with the goal of becoming the new Zapier , is shutting down, and some of its staff — including its top executive — are joining…
TechCrunch AI -
Limited sources
Former SpaceX engineers are building a robotic factory for making steel parts
Three former SpaceX engineers have switched their attention from making rocket engines to manufacturing steel parts by using AI-driven software and robots. Their immediate goal involves establishing…
Ars Technica AI -
Limited sourcesModelsResearch
Benchmarked Qwen3.8-27B on 4x RTX 3090
A while back I made a post about my 4x3090 rig in a Silverstone RV-02 . Check it out if you're a conoissuer of OG PC cases. With the incredible Qwen 3.8 27B release I ran benchmarks. So in case you a…
Reddit r/LocalLLaMA -
Limited sourcesModels
Qwen3.8-27B Uncensored Aggressive is out with K_P quants and HauhauCS FastMTP (up to 3.02x TG)!
The dense Qwen release is back! Qwen3.8-27B Uncensored Aggressive is out with the complete K_P quant range, Vision, native NextN, and HauhauCS FastMTP. Aggressive here means no refusals, no personali…
Reddit r/LocalLLaMA -
Limited sourcesModelsBusiness
GPT-5.6 Sol Pricing Cut by 50%
openai / gpt-5.6-sol GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and mu…
Hacker News Frontpage -
Limited sources
Israel creates fake think tank in likely attempt to dupe AI chatbots
At a glance, the Hanover Institute for Public Policy looks like a new think tank dedicated to Israel/Palestine. The organization churns out think-tank style reports on questions such as “Does AIPAC U…
Hacker News Frontpage -
Limited sourcesOpen SourceAgents & Dev
Local agentic coding Benchmark : Qwen 3.8 27B (in many weights quants / cache quants / engine / reasoning effort) vs others.
In medium reasoning mode, it both scores higher than the 3.6 version, AND is very much more efficient (almost half requests needed, and a third less tokens generated) - at DeepSeek v4 Flash 3107 MXFP…
Reddit r/LocalLLaMA -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.234
What's changed Added the optional CLAUDE_CODE_PROJECT_DIR_NAME environment variable: hosts that give each session its own config directory can choose a short name for the per-project transcript direc…
GitHub Releases · anthropics/claude-code -
Limited sourcesAgents & Dev
GitHub degradation affects Cursor Origin, its new Git platform
Article URL: https://status.cursor.com/incidents/l9h9vrd726jv Comments URL: https://news.ycombinator.com/item?id=49336919 Points: 24 # Comments: 3
Hacker News Frontpage -
Limited sourcesModels
Weirdly, no one talks about Temperature setting for the Qwen3.8 27b
Mind you, it is 1.0 by default, yet everyone is focused on how much the new model thinks, restricting the reasoning budget and/or dropping the reasoning level. Set the temperature to 0.7 and the mode…
Reddit r/LocalLLaMA -
Limited sourcesModels
I pushed Qwen3.8-27B to 99 tps single request and 1150 tps with a batch request on a RTX 3090
I'm back. Yesterday I released the first version of hyper-optimized Qwen3.8-27B inference engine for a RTX 3090, reaching 82 tps on single request and 672 peak. Over the last 24 hours I've been explo…
Reddit r/LocalLLaMA -
Limited sources
AI;DR (AI; Didn't Read)
I'm about as pro-AI as you can be, but this is becoming a pet peeve of mine (and I'm not alone). That's why I love the AI;DR acronym as my new solution for ignoring the walls of slop. I’m SUPER jealo…
Hacker News Frontpage -
Confirmed
Same Cluster, 33 Points More Utilization: What Changed Was the Order
The previous post argued that utilization, not intelligence, is where the next real constraint in enterprise AI is forming, and it closed by noting that no playbook has emerged yet for what a mature…
Hugging Face Blog -
Limited sources
"Opus 4.8 thinks too much", "Muse Glimmer sits between Gemma and Qwen, that's boring", "Gemma 4 is too lazy"
I'm starting to think there's no way to make a reasoning model that won't draw persistent vocal complaints on here. EDIT: Qwen 3.8 not Opus 4.8*, freudian slip lol submitted by /u/MerePotato [link] […
Reddit r/LocalLLaMA -
ConfirmedOpen SourceBusiness
openai-python v3.2.0
3.2.0 (2026-08-17) Features add Bedrock Runtime endpoint support (SDK-290) ( #3623 ) ( 86267d2 ) api: Add shell call streaming events and new service/image types ( #3635 ) ( ff14a33 )
GitHub Releases · openai/openai-python -
Corroborated
Reports say Amazon scanned and destroyed rare books for AI training
For the past year or so, booksellers have suspected that AI firms are buying up huge lots of rare books, then destroying them after scanning them to train AI. But this was hard to prove until now, as…
Ars Technica AI · Hacker News Frontpage · TechCrunch AI -
Limited sources
LLMs Endgame: This is Unreasonable
Qwen 3.8 27B Oh My This result is unreasonable Qwen 3.8 27B makes using almost any other model unreasonable (except for the big boyz ofc ?!) They took a while to release this and still even pricing A…
Reddit r/LocalLLaMA -
Limited sourcesModels
Qwen3.8 27B's result on Artificial Analysis is insane!
submitted by /u/FormOne2615 [link] [comments]
Reddit r/LocalLLaMA -
Limited sources
llama.cpp adaptive MTP PR#27210
Just wanted to raise some attention to this PR I filed if anyone would like to try it out. This adds an adaptive MTP mode to llama.cpp which employs a fairly simple counting-style state machine to de…
Reddit r/LocalLLaMA -
Limited sourcesModelsAgents & Dev
Qwen3.8 27B > Opus 5 Medium on Artificial Analysis Agentic Index
https://preview.redd.it/xh1rloaf4zjh1.png?width=1628&format=png&auto=webp&s=a536fae1b50b327f2bc55d4f4f874f94ae66e867 Thanks Qwen team! submitted by /u/secopsml [link] [comments]
Reddit r/LocalLLaMA -
ConfirmedOpen Source
llama.cpp reaches the v0.1.0 release milestone
llama.cpp is apparently moving to semantic versioning instead of just sequential build numbers (like b10456). The first semantic version tag was created today: https://github.com/ggml-org/llama.cpp/r…
GitHub Releases · ggml-org/llama.cpp · Reddit r/LocalLLaMA · Hacker News Frontpage -
Limited sourcesModels
Qwen3.8 27B = GPT-5.6 Luna compressed into 27B
How crazy is that? submitted by /u/kevinlch [link] [comments]
Reddit r/LocalLLaMA -
Limited sourcesModelsResearch
Artificial Analysis' Qwen3.8-27B benchmarks put it neck and neck with DeepSeek V4 and GPT-5.6 Luna Max
submitted by /u/anderspitman [link] [comments]
Reddit r/LocalLLaMA -
Limited sourcesModels
Qwen3.8 27B scores 52 on Artificial Analysis
Article URL: https://artificialanalysis.ai/models/qwen3-8-27b Comments URL: https://news.ycombinator.com/item?id=49334544 Points: 101 # Comments: 37
Hacker News Frontpage -
ConfirmedOpen Source
llama.cpp b10472
cuda : skip UMA override for HIP builds ( #27083 ) AMD APUs report accurate memory via hipMemGetInfo. Using MemAvailable over-promises on small-carveout systems. fixes #18159 Website: https://llama.a…
GitHub Releases · ggml-org/llama.cpp -
Limited sourcesModelsAgents & Dev
Cursor launches Origin, GitHub alternative
Aug 17, 2026 · Changelog Cursor can now host your code. Origin begins rolling out today in early beta on all paid plans. We're starting with the essentials, designed for agent scale: repos, pull requ…
Hacker News Frontpage -
Limited sourcesAgents & Dev
Why do people like coding harnesses like opencode etc instead of an IDE?
Just curious - I like to be able to see and manage the scripts my agent is working on. I find stuff like Claude Code and Open Code useful for doing stuff on my linux box but I don't understand why pe…
Reddit r/LocalLLaMA -
Limited sources
Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!
This Ling 3.0 Tiny 8b param with 1.3b active is the fastest, smartest model I can run on my poor old pc, with 4gb vram. It actually runs lightning fast, like 36 token / sec, as smart as Qwen 3.5 9b /…
Reddit r/LocalLLaMA -
Limited sourcesBusinessInfra
Groq raises $350M to fuel its pivot from AI chips to neocloud
Startup Groq has raised $350 million as it continues to pivot from an AI chipmaker to a neocloud company that provides powerful GPUs and AI infrastructure services. The new capital, led by investment…
TechCrunch AI -
Limited sources
Qwen 3.8 27B tests show strong results but possible overthinking
Yes, it sucks to waste time waiting on 16K+ reasoning tokens alone. But here's the thing, this is only a 27B model trying to perform on par with 1T+ parameter models. Something has to be sacrificed,…
Simon Willison's Weblog · Reddit r/LocalLLaMA -
Limited sourcesCreative
Launch HN: Speko (YC S26) – OpenRouter for Voice AI
Hi HN! I'm Bek, founder of Speko, a platform that finds an optimal combination of speech-to-text, LLM, and text-to-speech models, given your constraints, among all our public benchmarked options, and…
Hacker News Frontpage -
Limited sources
We Tracked a Shipment of Rare Books. It Ended at an Amazon AI Training Facility
We Tracked a Shipment of Rare Books. It Ended at an Amazon AI Training Facility Excellent piece of reporting from 404 Media. For a while now there have been stories of book dealers receiving orders f…
Simon Willison's Weblog -
Limited sourcesPolicy & SafetyInfra
Nvidia investing $1.5B in SoftBank data center developer behind OpenAI project
Tim De Chant Nvidia said on Monday that it will invest $1.5 billion in SB Energy, a data center linked to SoftBank and OpenAI. The investment ensures that Nvidia will be the sole supplier of compute…
TechCrunch AI -
Limited sources
Deepseek Harnness - why is feels better
Guys, could someone smarter than me explain what makes Deepseek Harness so efficient? I run it with local Qwen 3.8 (Q6). I tried Opencode/Openchamber (my favourite so far), Pi agent and Hermes. New Q…
Reddit r/LocalLLaMA -
Limited sources
noctrex/Ling-3.0-tiny-MXFP4_MOE-GGUF · Hugging Face
u/noctrex 👍 where's flash? 😄 Possibly fastest model(in this model size range). Share t/s stats. submitted by /u/pmttyji [link] [comments]
Reddit r/LocalLLaMA -
Limited sourcesPolicy & Safety
Judge relying wholly on AI in order is covered by judicial immunity, court rules
From Wednesday's decision in Phillips v. Parlade , by Judge Gloria Navarro (D. Nev.), where a litigant sued a state court judge in his case: Plaintiff … argu[es] that judicial immunity does not apply…
Hacker News Frontpage -
Limited sourcesModels
Qwen3.8-27B at 256K on a 24GB RTX PRO 4000 SFF (432 GB/s): 50 tok/s with MTP
AI INFRASTRUCTURE I gave Qwen3.8's MTP drafter another 69.2 MiB of precision. Throughput fell from 50.44 to 37.02 tokens per second. That result sums up the whole experiment: the best local inference…
Hacker News Frontpage -
Limited sourcesAgents & Dev
AI-Generated GitHub Copilot "Autofix" Allowed Compromise of Snowflake's Jira
As part of ongoing security research conducted through Snowflake’s HackerOne vulnerability disclosure program, Wiz Research’s "Red Agent"—an autonomous, AI-powered security research tool—identified a…
Hacker News Frontpage -
ConfirmedOpen Source
llama.cpp v0.1.1
Release v0.1.1
GitHub Releases · ggml-org/llama.cpp -
Limited sources
How to disable or avoid intrusive AI
One of the biggest questions I get at Drop-In Time at the library (besides “what is taking up all my cloud storage?”) is how to disable or avoid intrusive AI that shows up where people don’t want it.…
Hacker News Frontpage -
ConfirmedOpen Source
llama.cpp b10470
ci : push release tag explicitly in release.yml ( #27261 ) Add a "Create and push git tag" step to the release job, right before the "Create release" step. The tag is created with git tag and pushed…
GitHub Releases · ggml-org/llama.cpp -
Limited sources
DeepSeek V4 Flash with Antirez Dwarfstar 4 is amazing.
Note: I use a Mac Studio M3U with 512 GB RAM, so this is not for everyone. I have been using antirez/ds4 with DS4 Flash for a few weeks now. Top quality, I am really impressed. This combination just…
Reddit r/LocalLLaMA -
Limited sources
EXL3 seems to be fading from the r/LocalLLaMa consciousness, and while I suspected it, I'm surprised at this point in time.
EXL3 is an alternative to llama.cpp. And while there is extensive tooling for llama.cpp, EXL3's primary deployment ( TabbyAPI ), has a OpenAI compatible API so it shouldn't matter. Why won't this too…
Reddit r/LocalLLaMA -
Limited sources
Show HN: LLMs each trading $100K vs. a frozen rulebook – the rulebook leads
as of Aug 17, 3:45 PM ET · refreshing… Get this table after every close. One short email with the day's numbers and what moved them. Who's winning — live Profit on closed paper trades, per account. G…
Hacker News Frontpage -
Limited sourcesBusiness
Wispr raises $280M at $2B valuation as it looks beyond dictation
Wispr , a startup known for its AI dictation tool, raised $280 million in Series B funding, led by Menlo Ventures, at a $2 billion valuation, the company announced on Monday. The funds will allow Wis…
TechCrunch AI -
Limited sourcesPolicy & Safety
Show HN: Sokoban AI Solver
Sokoban ("warehouse keeper") is a 1980s puzzle: push every box onto a goal. In this variant the keeper must also finish on a goal. The warehouse is a grid. On each step the keeper moves one square up…
Hacker News Frontpage -
Limited sourcesAgents & Dev
After pushing 1M+ tokens through Qwen 3.8 27B, here is my optimal llama.cpp config for 16GB VRAM (73k Context, Agentic Coding)
Following up on my previous post about my budget server setup (Intel N100 + RTX 5060 Ti 16GB), a few of you asked for a deeper dive into my actual inference config and real-world agentic performance.…
Reddit r/LocalLLaMA -
Limited sourcesResearch
[Paper] Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning
We introduce Mobius-v0, an architecture that comprises a globally shared Memory (FFN) that stores knowledge vectors and multiple Reasoners (Self-Attn) that iteratively achieve compositional reasoning…
Reddit r/LocalLLaMA -
Limited sources
Mimir: Did the vikings train a 1.7B killer model?
for those looking for something small AND powerful, there is a new 1B (they claim, it looks more like 1.7B ...) model that claims to beat qwen 3.5 0.8B & 2B and gemma 4 E2B on a range of benchmarks.…
Reddit r/LocalLLaMA -
Limited sources
tencent/EVIE-Preview-4.5B · Hugging Face
Overview EVIE-Preview-4.5B is a state-of-the-art multilingual Visual Document Retrieval (VDR) model built upon Qwen3.5-4B . It employs ColBERT-style late interaction with native 128-dimensional multi…
Reddit r/LocalLLaMA -
Limited sourcesOpen Source
GPT 5.6 Sol is the best "vision" model OpenAI ever released
Last week, OpenAI announced the GPT-5.6 lineup, introducing the Sol, Terra, and Luna models. During the release stream , the team focused heavily on computer use , showing models capable of navigatin…
Hacker News Frontpage -
Limited sourcesInfra
100$ worth of gpu runs qwen 3.8 27b at 7.39 t/s
Qwen 27b Q3_K_M 2x rx 580 8gb (~50$ each in my country, edge cases 60$ per gpu) gives us 16gb vram We used it on an old already existing ddr3 motherboard with 2 gpu slots(you can buy it ror around 20…
Reddit r/LocalLLaMA -
Limited sourcesBusiness
Whisker’s AI-powered litter robot thinks my cats swapped bodies
The $899 Litter-Robot 5 Pro is a great pooper-scooper wrapped in AI that just doesn’t work. If you buy something from a Verge link, Vox Media may earn a commission. See our ethics statement. The $899…
The Verge AI -
Limited sources
Anthropic explains how Claude’s invisible text watermarks will work
It’s using ‘a version’ of the open-source SynthID-Text system Google developed. It’s using ‘a version’ of the open-source SynthID-Text system Google developed. Anthropic has clarified how it’s planni…
The Verge AI -
Limited sourcesResearch
LLM's can't "jump" - a paper by Deepmind showing LLMs can't generate novel explanatory hypotheses
submitted by /u/juanviera23 [link] [comments]
Reddit r/LocalLLaMA -
Limited sourcesOpen Source
[Megathread] Qwen 3.8 27B Release Day
Megathread to help with the influx of duplicate / similar posts around the release of the Qwen 3.8 27B release. Quants Fine-Tunes & Abliterations Chat Templates Inference Server Support & Configurati…
Reddit r/LocalLLaMA -
Limited sources
Petition to add a rule for people to add their DAMN quant levels to their posts
Every time I see a post about a newly released model, whether it be a comparison or shitting on it, I have to dig through the endless comments to see what quants they used and what their specs were.…
Reddit r/LocalLLaMA -
Limited sources
Ling 3.0 support merged into llama.cpp
Support for the new ling 3.0 models has been merged into llama.cpp: https://github.com/ggml-org/llama.cpp/pull/26608#event-29549472828 Ling tiny 8b1b - https://huggingface.co/inclusionAI/Ling-3.0-tin…
Reddit r/LocalLLaMA -
Limited sourcesModels
Qwen3.8-27B Q8_0 on Strix Halo is seriously impressive
Sorry for the slop, but I was impressed by this model as I have been testing Qwen3.8-27B Q8_0 locally on my ROG Flow Z13 (Ryzen AI Max+ 395, 128 GB unified memory) and this model was the only one who…
Reddit r/LocalLLaMA -
Limited sourcesBusiness
Long Review: Qwen 3.8 27B is VERY good at tapping into it's real-world knowledge. It's "overthinking" brings it to Sonnet level performance with the potential for Opus level results.
Hi all! I finally just got around to testing out Qwen 3.8 27b. I'm using Unsloth's UD-Q8_K_XL quant as a sit-in replacement to Qwen 3.6 27b, same quant size. Wow -- this thing isn't messing around. I…
Reddit r/LocalLLaMA -
Confirmed
Get closer to the game with Gemini and Pixel
Gemini and Pixel have announced new long-term partnerships with Arsenal FC, FC Barcelona, FC Bayern München, Liverpool FC, and Paris Saint-Germain to bring fans closer to the game through AI personal…
Google AI Blog -
Limited sources
Stripe reportedly plans to acquire OpenRouter for more than $7B
Anthony Ha Stripe has finalized a deal to acquire OpenRouter, according to a new report in Bloomberg . OpenRouter helps customers select different AI models to perform different tasks, depending on t…
TechCrunch AI · Reddit r/LocalLLaMA -
ConfirmedOpen Source
llama.cpp tmp-testing-0
ci : make release workflows use a deply key
GitHub Releases · ggml-org/llama.cpp -
ConfirmedOpen Source
llama.cpp b10456
sycl: fix thread/block count in quantized cpy kernel launches ( #27160 ) Adjusts the thread/block count to be proportional to the size of the quant, reducing under/over subscription. Largest perf imp…
GitHub Releases · ggml-org/llama.cpp -
ConfirmedOpen Source
llama.cpp b10455
[SYCL] support OP OPT_STEP_ADAMW, OPT_STEP_SGD ( #25268 ) fix conflict fix conflict of ops.md fix conflict of ops.md update the ops.md Co-authored-by: Neo Zhang Jianyu jianyu.zhang@intel.com Website:…
GitHub Releases · ggml-org/llama.cpp -
Limited sourcesModelsInfra
How many tokens/second output are you getting with Qwen3.8-27B?
Trying to get a feel for where I stand. If you can list your relevant hardware and model used, that would be awesome. Here's mine: Model: Qwen3.8-27B-heretic-ara, Q5_K_M GGUF T/s : ~30-32 t/sec (I th…
Reddit r/LocalLLaMA -
Confirmed
The Defender’s Window
AI is reshaping cybersecurity for attackers and defenders alike. Learn how OpenAI is strengthening its defenses and what security teams can do now.
OpenAI News -
ConfirmedBusiness
OpenAI joins PORTS-Pike project
OpenAI joins PORTS-Pike project, expanding community investment and supporting thousands of Southern Ohio jobs
OpenAI News -
Limited sources
…and I’m not afraid of losing my social credits.
submitted by /u/JLeonsarmiento [link] [comments]
Reddit r/LocalLLaMA -
ConfirmedAgents & Dev
New policy ideas for the Intelligence Age
OpenAI funds 14 independent projects exploring new AI policy ideas to expand economic opportunity and strengthen societal resilience in the Intelligence Age.
OpenAI News Aug 16 · Sun
-
Limited sources
Markdown SVG upgrades
I started building my markdown-svg-renderer tool in May , but I've since added enough features to it that it's worth talking about here again. It's evolved into my ideal tool for sharing Markdown tra…
Simon Willison's Weblog -
Limited sources
Qwen 3.8 27b vs 3.6 27b - how good is with a Turtle library.
Prompt: Provide complete working code for a realistic looking tree in Python using the Turtle graphics library and a recursive algorithm. Difference between 3.6 and 3.8 is huge! submitted by /u/Healt…
Reddit r/LocalLLaMA -
Limited sourcesPolicy & Safety
OpenAI reportedly disbanded its preparedness team
It’s just the latest shakeup of its safety teams as it heads toward an IPO. It’s just the latest shakeup of its safety teams as it heads toward an IPO. According to the Financial Times , OpenAI disba…
The Verge AI -
Limited sources
Why people aren’t buying Mark Zuckerberg’s AI future
Meta CEO Mark Zuckerberg published a 6,500-word essay this week declaring that “The Future is for Everyone” and painting an optimistic picture of a future powered by AI, where “everyone will have an…
TechCrunch AI -
ConfirmedOpen Source
ollama v0.32.14
What's Changed llm: transcode WebP images for llama-server renderers/qwen: tolerate non-leading system messages Full Changelog : v0.32.13...v0.32.14
GitHub Releases · ollama/ollama -
Limited sources
Based on an accelerating frontier -> local trajectory, expect a ~30b param 'Mythos at home' by as soon as Jan 2027 (rationalisation below)
Including the rationalisation for the data below - this is a more robust version of an earlier post I did similar to this - explaining below: How I chose the comparisons The basic question I’m trying…
Reddit r/LocalLLaMA -
Limited sourcesBusiness
Anthropic CEO says AI backlash is ‘fundamentally a crisis of trust’
Anthropic CEO Dario Amodei recently pushed back against the idea that he’s been painting an overly pessimistic picture of artificial intelligence and how it might shape the future. Amodei’s comments…
TechCrunch AI -
Limited sources
Let’s all thank Georgi Gerganov who gave use llama.cpp
I was looking into the story a bit further earlier. Very interesting. Couldn’t have done it without him submitted by /u/on_line187 [link] [comments]
Reddit r/LocalLLaMA -
Limited sources
Quoting Dario Amodei
I do agree that the public has a negative view of AI (and that this is a big problem), but I don’t think it is primarily caused by me or any other AI leader warning about AI’s risks. I think it is fu…
Simon Willison's Weblog -
Limited sourcesInfra
ChatGPT’s Computer History tracks your clicks and keystrokes
It’s like Windows Recall, but without all the creepy screenshots. (But it’s still kind of creepy.) It’s like Windows Recall, but without all the creepy screenshots. (But it’s still kind of creepy.) C…
The Verge AI -
ConfirmedOpen Source
llama.cpp b10453
model : remove some ggml_concat ( #27176 ) Co-authored-by: Xuan Son Nguyen son@huggingface.co Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI en…
GitHub Releases · ggml-org/llama.cpp -
ConfirmedOpen Source
llama.cpp b10454: ci : fix dry-run reporting in make-release job [no ci] (#27167)
This commit fixes the reporting in the make-release CI job when --dry-run is used. It will currently incorrectly report that all checks pass even if there are steps that fail. Refs: #26839 (comment)
GitHub Releases · ggml-org/llama.cpp -
Limited sources
Rogue AI aren’t science fiction anymore
For years, fears about AI systems slipping human control were dismissed as speculative. If you buy something from a Verge link, Vox Media may earn a commission. See our ethics statement. For years, f…
The Verge AI -
ConfirmedOpen Source
llama.cpp b10452
chat: refactor handling supports_string_content / supports_typed_content ( #27130 ) better supports_string_content cap detect test: add "skip" messages_inp_normalizer Website: https://llama.app macOS…
GitHub Releases · ggml-org/llama.cpp -
ConfirmedOpen Source
llama.cpp b10451
llama : check LoRA tensor data is within file bounds ( #27056 ) llama : check LoRA tensor data is within file bounds Update src/llama-adapter.cpp Co-authored-by: Sigbjørn Skjæret sigbjorn.skjaeret@hu…
GitHub Releases · ggml-org/llama.cpp Aug 15 · Sat
-
Limited sources
Woman claims her stepfather used Grok to transform childhood photo into explicit imagery
Anthony Ha A woman identified as Jane Doe 4 has joined a lawsuit filed by three Tennessee teenagers against Elon Musk’s xAI over the role the company’s chatbot Grok allegedly played in creating child…
TechCrunch AI -
Limited sources
Have a laugh at AI’s expense by roleplaying as a chatbot
Scribble some nonsense, offer wooden jokes, or make your own requests of a real human playing a fake AI. Scribble some nonsense, offer wooden jokes, or make your own requests of a real human playing…
The Verge AI -
Limited sources
Anthropic shares more details about how Claude’s new watermarks will work
Anthropic published a blog post Friday seeking to answer some basic questions about how it will watermark the text generated by its chatbot Claude. Such as: How will the watermarking actually work? C…
TechCrunch AI -
ConfirmedOpen Source
ollama v0.32.14-rc0
mlx update ( #17761 )
GitHub Releases · ollama/ollama -
Limited sourcesAgents & DevBusiness
SpaceX officially closes its Cursor acquisition
Anthony Ha AI coding startup Cursor is now officially a part of SpaceX, according to an announcement on the Cursor blog . Elon Musk’s SpaceX — which also acquired Musk’s xAI earlier this year — annou…
TechCrunch AI -
Limited sources
How to tell if your AI platforms’ accounts have been hacked
Just like any other online service, hackers can target and break into your accounts on popular AI platforms such as ChatGPT, Claude, and Perplexity. TechCrunch has created a comprehensive guide to he…
TechCrunch AI -
Limited sources
CORS Chat
Tool: CORS Chat I built this today ( with GPT-5.6-Sol xhigh ) to help test Qwen 3.8 27B running in LM Studio on both my M5 MacBook Pro and an NVIDIA DGX Spark. It provides a web UI for exercising an…
Simon Willison's Weblog -
ConfirmedOpen Source
ollama v0.32.12
Qwen 3.8 27B This release adds the support of Qwen 3.8 27B . Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. ollama run qwen3.8:27b For…
GitHub Releases · ollama/ollama -
Limited sources
Northern Gannet
Northern Gannet, in Pillar Point Harbor, CA, US This is Morris. Morris is a local celebrity: the only known Northern Gannet ( Morus bassanus ) in the entire Pacific Ocean. They showed up in the Faral…
Simon Willison's Weblog Aug 14 · Fri
-
ConfirmedOpen SourceBusiness
openai-python v3.1.0
3.1.0 (2026-08-14) Features api: add WebSocket stream IDs ( #3612 ) ( d9029e3 ) api: add workload identity access token issued event ( #3601 ) ( df274d4 ) api: deprecate Sora video APIs ( #3610 ) ( 7…
GitHub Releases · openai/openai-python -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.233
What's changed Added GitLab merge request URL support to the --worktree flag and the claude agents view (where MRs display as !N ) Added an opt-in forward_user_identity apps gateway setting on Anthro…
GitHub Releases · anthropics/claude-code -
Limited sources
Don't classify. Hallucinate!
Don't classify. Hallucinate! I still have quite a bit of older content on my blog that I never got round to tagging. My blog has 1,856 tags - likely too many to feed to an LLM in one go and say "whic…
Simon Willison's Weblog -
ConfirmedOpen Source
ollama v0.32.13
What's Changed qwen3.8: support developer instructions Full Changelog : v0.32.12...v0.32.13
GitHub Releases · ollama/ollama -
Limited sourcesPolicy & Safety
Suspecting court of using AI, man injected prompts in filings to try to win case
A judge has identified what appears to be the first time a US plaintiff has attempted to hide text in court filings that only an artificial intelligence system can read in a bid to win a case. In a d…
Ars Technica AI -
Limited sources
Mark Zuckerberg has an Instagzam
On The Vergecast: Instagram’s new logo, Zuck’s AI manifesto, and lots of other bad ideas. On The Vergecast: Instagram’s new logo, Zuck’s AI manifesto, and lots of other bad ideas. Instagram’s wordmar…
The Verge AI -
Limited sources
You can now turn off Google Gemini’s visible watermarks
Google will still embed invisible SynthID and C2PA watermarks into AI-generated images, videos, and music. Google will still embed invisible SynthID and C2PA watermarks into AI-generated images, vide…
The Verge AI -
Limited sources
Google will now allow users to remove visible watermark from its AI generations
Ivan Mehta Google announced on Friday that it will now allow users to remove a visible watermark from its AI generations, including images, videos, and songs. The company specified that this won’t af…
TechCrunch AI -
Limited sources
Does Mark Zuckerberg really believe AI is ‘for everyone’?
Loading the player… Meta released Glimmer this week , an open-weight AI model anyone can download and run on their own hardware — a contrast to Muse Spark, the company’s more powerful model that stay…
TechCrunch AI -
Limited sourcesInfra
Kog is going deeper to squeeze more inference out of GPUs
The race for faster AI inference is on, and markets gave Cerebras and its purpose-built chips a warm welcome in its IPO debut in May. But French startup Kog is betting that there’s a lot more power t…
TechCrunch AI -
Limited sources
OpenAI and Anthropic in price war as Chinese AI rivals gain ground
Leading US AI labs such as OpenAI and Anthropic are releasing cheaper models as they fight to retain cost-conscious customers who are switching to cut-price alternatives from Chinese rivals. The pric…
Ars Technica AI -
Limited sources
Hyperscalers might regret embracing natural gas if new forecast proves correct
After years of snapping up wind and solar developments, hyperscalers like Amazon, Google, Meta, and Microsoft are betting that natural gas will power the data centers behind their lofty AI ambitions.…
TechCrunch AI -
Limited sources
Meta’s ‘open’ AI, and a $250M deal gone very wrong
Meta released Glimmer this week , an open-weight AI model anyone can download and run on their own hardware — a contrast to Muse Spark, the company’s more powerful model that stays locked behind its…
TechCrunch AI -
Limited sourcesBusiness
Apple trained its own AI model for China with help from Alibaba
The rare US-China partnership comes as Apple prepares to roll out its on-device generative AI service in China. The rare US-China partnership comes as Apple prepares to roll out its on-device generat…
The Verge AI -
ConfirmedOpen Source
ollama v0.32.11
What's Changed ollama launch dsh now supports DeepSeek Harness, DeepSeek's open-source agent harness ollama launch muse now supports Muse Code , Meta's agentic coding CLI The OpenAI-compatible Respon…
GitHub Releases · ollama/ollama -
Confirmed
State of Open Models: Summer 2026 Observations
In the AI world, time feels compressed. A few months after our spring report in our biannual analysis worked through the ecosystem, there are quite a few findings that we have observed until this sum…
Hugging Face Blog Aug 13 · Thu
-
Limited sources
sqlite-utils 4.2.1
Release: sqlite-utils 4.2.1 Fixes a crashing bug in sqlite-utils 4.2 . I'd introduced code that looks like this: from typing_extensions import Self It turned out the typing-extensions package was not…
Simon Willison's Weblog -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.232
What's changed Subagent forking is now on by default: a subagent_type: "fork" subagent inherits the full conversation and prompt cache, and non-teammate agent spawns in interactive sessions now run i…
GitHub Releases · anthropics/claude-code -
Limited sourcesAgents & Dev
Microsoft’s Clippy-like Mico character is no longer the face of Copilot
Mico launched in Copilot’s voice mode less than a year ago. Mico launched in Copilot’s voice mode less than a year ago. Microsoft Copilot will no longer show its emotive yellow blob, Mico, when you u…
The Verge AI -
Limited sources
Writer introduces new AI model and upgraded harness to contain token costs
Across the AI industry, users are becoming more conscious of just how expensive their deployments can be —and feeling a new urgency to cut costs. But while open source models offer significantly lowe…
TechCrunch AI -
Limited sourcesBusiness
Databricks wanted to raise $1B, investors wanted $15B. It settled on $5B at a $190B valuation.
There’s a funny kind of game that the latest of late-stage startups must play when raising money. They often have to sell more shares than they want or risk offending some of their existing VCs. This…
TechCrunch AI -
Limited sources
sqlite-utils 4.2
Release: sqlite-utils 4.2 Lots of improvements in this one relating to the table.transform() feature , which adds support for complex alter table operations by creating a fresh table, copying across…
Simon Willison's Weblog -
Limited sourcesModels
llm-gemini 0.33
Release: llm-gemini 0.33 It's been a while since the last llm-gemini release. This version of the plugin adds support for today's Gemini 3.7 Flash release, plus gemini-3.6-flash , gemini-3.5-flash-li…
Simon Willison's Weblog -
Limited sources
OpenAI is losing its second executive this week
Chief revenue officer Denise Dresser announced she is leaving OpenAI. Two days ago, Brad Lightcap also said he’d be leaving. Chief revenue officer Denise Dresser announced she is leaving OpenAI. Two…
The Verge AI -
ConfirmedModels
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed
Lucas Ropek If you’ve ever found yourself wishing that ChatGPT was a little bit quicker on the uptake, OpenAI seems to be answering your prayers. The AI lab has rolled out a new mode called Ultrafast…
OpenAI News · TechCrunch AI -
Limited sources
IBM partners with OpenAI to bolster enterprise AI push
IBM on Thursday announced its partnership with OpenAI to bring the AI company’s models and tools to more enterprise customers, opening another avenue for OpenAI to connect with some of the world’s la…
TechCrunch AI -
ConfirmedOpen SourceBusiness
anthropic-sdk-python v0.122.0
0.122.0 (2026-08-13) Full Changelog: v0.121.0...v0.122.0 Features api: add output_behavior to dream creation (create a new memory store or update the input store in place) ( 852c4bb ) Bug Fixes bedro…
GitHub Releases · anthropics/anthropic-sdk-python -
ConfirmedOpen Source
ollama v0.32.10
What's Changed Models that don't set a repeat_penalty now default to 1.0 (off) instead of 1.1, matching other engines and speeding up speculative decoding; set a per-model parameter if an older model…
GitHub Releases · ollama/ollama -
Limited sourcesAgents & Dev
Anthropic set AI agents loose on the same task. They started a turf war.
What happens when you pit AI agents against each other? According to Anthropic’s testing, things get messy fast. On Thursday, Anthropic’s Frontier Red Team published new research examining how groups…
TechCrunch AI -
Limited sources
The new Instagram logo is the perfect embodiment of AI slop
Today, Instagram unveiled a refresh of its wordmark, accompanied by the usual bland corporate platitudes these kinds of announcements are always packaged with. Head of Instagram Adam Mosseri called i…
Ars Technica AI -
ConfirmedAgents & Dev
Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets
A walkthrough of the streaming data loop in Strands Robots, one agent loop that records robot demonstrations, trains on them by reading straight from the Hub, and deploys the policy back to hardware,…
Hugging Face Blog -
Limited sourcesBusiness
OpenAI hires new CRO as executive shake-up continues
OpenAI has replaced chief revenue officer Denise Dresser after just nine months on the job, tapping Wiz president and chief operating officer Dali Rajic to take on frontier lab’s top sales job. The m…
TechCrunch AI -
ConfirmedModelsOpen Source
Introducing Gemini 3.7 Flash
Our most intelligent workhorse model yet for coding and agents. Tulsee Doshi Senior Director, Product Management, on behalf of the Gemini team Today, we’re building on the progress of our widely used…
Ars Technica AI · Google DeepMind Blog -
Confirmed
Bring your spreadsheet data to life with Sheets canvas
A new Google Sheets feature, Sheets canvas transforms rows and columns into interactive dashboards, custom study trackers, seating charts, and more, all with a simple prompt. Eric Birnbaum Director,…
Google AI Blog -
Limited sourcesAgents & DevBusiness
Microsoft kills off unsuccessful AI features while merging its separate Copilot apps
Two years ago, Microsoft described AI as a “generational shift” in technology that it wanted to lead. Today, the company is merging its Copilot-branded consumer and business apps, and ditching a numb…
TechCrunch AI -
Limited sources
Anthropic could be worth $2 trillion when it goes public
Anthropic investors expect the AI startup to float at a valuation of $2 trillion or more in October, a dizzying figure that would eclipse SpaceX and make the AI lab’s debut the largest-ever initial p…
Ars Technica AI -
Limited sources
Claude's new Scarlet Letter watermark is invisible—for now
Anthropic has revealed that it will soon watermark content that is processed ( not just generated! ) by any of its models. In a support article , Anthropic explained that it was rolling out machine-r…
Ars Technica AI -
ConfirmedAgents & Dev
The builder’s guide to GPT‑5.6
Learn how startups use GPT-5.6 to build faster, more cost-efficient AI agents with smarter model selection and new Responses API capabilities.
OpenAI News -
ConfirmedBusiness
OpenAI appoints Dali Rajic as Chief Revenue Officer
OpenAI appoints Dali Rajic as Chief Revenue Officer to lead its global revenue organization and help businesses realize the full value of AI.
OpenAI News -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.231
What's changed Fixed MCP OAuth sign-in failing with a redirect URI mismatch for servers that use a pre-registered OAuth client, such as Slack
GitHub Releases · anthropics/claude-code -
Limited sources
alchemy-utils 0.1a1
Release: alchemy-utils 0.1a1 Performance boost for DuckDB exports and CSV imports, see here .
Simon Willison's Weblog -
ConfirmedResearch
What We Learned by Reproducing 2,200 papers from ICML
Back in July, we ran a hackathon where more than 1,200 community members brought their own coding agents and tried to reproduce the papers published at ICML 2026, claim by claim. In 19 days, particip…
Hugging Face Blog Aug 12 · Wed
-
Limited sources
DeepSeek V4 Pro 0813 (on OpenRouter)
DeepSeek V4 Pro 0813 (on OpenRouter) The latest DeepSeek Pro model is now available, via API only. I had to link to OpenRouter because DeepSeek don't have any obvious announcement page for their new…
Simon Willison's Weblog -
Limited sources
The web’s newest weapon against AI scrapers is a font
AI companies’ penchant for scraping through large swaths of the public web in search of valuable training data has already led to lawsuits and technical fixes aimed at stopping the practice. Now, a p…
Ars Technica AI -
Limited sourcesPolicy & Safety
Terabytes of credentials leaked in massive supply-chain attack
Terabytes worth of credentials, many belonging to the world’s biggest and most sensitive organizations, have been exposed in a supply-chain attack on LiteLLM, an open source tool that streamlines AI-…
Ars Technica AI -
ConfirmedOpen Source
ollama v0.32.10-rc1: mlx: avoid pulling MLX models when MLX is missing (#17710)
As we look to bring Linux and Windows MLX support online, instead of blocking downloads at the registry to avoid users wasting time downloading a model they can't run, shift the logic to the local si…
GitHub Releases · ollama/ollama -
Limited sources
Twitch content has trained Amazon AI for years, but users can opt out now
Twitch now lets users opt out of Amazon’s use of content from their channels to train Amazon’s “generative AI content models.” The change, announced today, comes more than two years after a company e…
Ars Technica AI -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.229
What's changed Documented claude remote-control --continue for resuming the most recent Remote Control session Added server-supplied Claude Code hook support for self-hosted runner sessions, matching…
GitHub Releases · anthropics/claude-code -
ConfirmedOpen Source
ollama v0.32.10-rc0: nn: speed up prefill on double-scale nvfp4 models
ModelOpt checkpoints apply a float32 global scale to every projection output on top of the per-group quantization scales. Running the multiply and the cast back to the activation dtype as separate ea…
GitHub Releases · ollama/ollama -
Limited sources
alchemy-utils 0.1a0
Release: alchemy-utils 0.1a0 I've long pondered what a database agnostic version of my sqlite-utils Python library and CLI utility might look like. This morning (literally a shower project) I tasked…
Simon Willison's Weblog -
ConfirmedOpen SourceAgents & Dev
vllm v0.27.2rc0: [Spec Decode] DSpark confidence-scheduled verification (#47808)
Signed-off-by: Lucas Wilkinson lwilkins@redhat.com Signed-off-by: Lucas Wilkinson LucasWilkinson@users.noreply.github.com Signed-off-by: Benjamin Chislett chislett.ben@gmail.com Signed-off-by: Lucas…
GitHub Releases · vllm-project/vllm -
ConfirmedModels
Introducing OlmoEarth embeddings: Custom embedding exports from OlmoEarth Studio for downstream analysis
📄 Tech Report: https://allenai.org/papers/olmoearth | 📊 Documentation: https://docs.olmoearth.allenai.org/embeddings | 💻 Learn more about OlmoEarth: https://allenai.org/olmoearth OlmoEarth Studio…
Hugging Face Blog -
Limited sources
Booksellers suspect AI firms are buying and then destroying rare books
If you can truly appreciate an old book—and maybe even marvel at how its fragile, yellowing pages contain some of the earliest ways that people tried to make sense of the world around them—then headl…
Ars Technica AI -
Limited sources
Quoting Florian Herrengt
But then users start to report a weird bug. It's the 4th time your team has been trying to fix it. I mean... asking AI to fix it. Unfortunately, it seems like not even Fable can figure it out. You go…
Simon Willison's Weblog -
Confirmed
Putting sign language AI into users’ hands
Google DeepMind Sign Language Team Introducing sign-language-to-text (SL2T), our breakthrough model powering new sign language features for Deaf and hard of hearing users. AI's ability to process spo…
Google DeepMind Blog -
Confirmed
LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge
LFM2.5-VL-3B is our most capable vision-language model you can run on your own hardware. It understands documents and screens alike, grounds objects, and can call tools. It answers directly instead o…
Hugging Face Blog -
Confirmed
From assistance to execution: How enterprises put AI to work
OpenAI research reveals how enterprises are adopting agentic AI, using ChatGPT and Codex, and how frontier firms are pulling ahead in AI adoption.
OpenAI News -
ConfirmedOpen SourceBusiness
openai-python v3.0.0
3.0.0 (2026-08-12) ⚠ BREAKING CHANGES api: HTTPX2 is now the default HTTP client, and httpx is no longer installed automatically. Applications using custom HTTPX clients, transports, or configuration…
GitHub Releases · openai/openai-python -
Confirmed
How RingCentral builds AI-native work from engineering to ops
See how RingCentral uses ChatGPT Work and Codex to accelerate AI product development and centralize operational intelligence across engineering and operations.
OpenAI News Aug 11 · Tue
-
Limited sources
There are no lossless transformations of natural-language text
There are no lossless transformations of natural-language text Sophie Alpert shares her "internal policy on acceptable use of AI writing by engineers". It's a short read (supporting its own recommend…
Simon Willison's Weblog -
Limited sourcesBusiness
Stealing Reasoning Traces from Proprietary LLM APIs
Stealing Reasoning Traces from Proprietary LLM APIs A vanity domain name ( stolen-thoughts.com ) for a neat paper : Anthropic, OpenAI, and Google return encrypted chain-of-thought blocks to clients t…
Simon Willison's Weblog -
Limited sources
datasette-upload-dbs 0.5a0
Release: datasette-upload-dbs 0.5a0 This plugin has been around for a while - it lets users upload a brand new SQLite database to a hosted Datasette instance, at which point that database will start…
Simon Willison's Weblog -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.228
What's changed Fixed interactive sessions that could stop redrawing entirely, while the process kept running, after a rare internal layout error Fixed git / Git Bash not being found on Windows when C…
GitHub Releases · anthropics/claude-code -
Limited sources
Gemini becomes Google's fastest-growing product ever as it hits 1B users
Google has been all-in with Gemini for the last several years, and despite some animosity online, the bet is paying off. CEO Sundar Pichai announced today that Gemini has reached 1 billion monthly ac…
Ars Technica AI -
ConfirmedOpen SourceBusiness
openai-python v2.54.0
2.54.0 (2026-08-11) Features api: Add new Responses model identifiers ( #3595 ) ( 0652787 ) Bug Fixes api: clarify audio upload metadata requirements ( #3596 ) ( 28888f9 ) Chores api: Update generate…
GitHub Releases · openai/openai-python -
ConfirmedAgents & DevResearch
AMIE, our research medical AI system, demonstrates real-time clinical video consultation capabilities in a first-of-its-kind study.
Anil Palepu Research Lead When you visit a doctor, a consultation extends far beyond words — a physician notices a cough, observes gait, or registers visible signs of discomfort. Today, Google Resear…
Google AI Blog -
Confirmed
Thinking of ACE? We Can Do It with Fewer Tokens
ALTK-Evolve and ACE both let an agent learn from its own trajectories. The difference is what they do with what they learn — and that decides the token bill. Give an LLM agent a realistic multi-step…
Hugging Face Blog -
ConfirmedOpen Source
ollama v0.32.9
NVIDIA Nemotron 3.5 Lightning NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for that execution layer of always-on agents. It is designed…
GitHub Releases · ollama/ollama -
ConfirmedOpen Source
vllm v0.27.1
This is a patch release on top of v0.27.0. Support quantized DSpark Markov heads ( #50424 )
GitHub Releases · vllm-project/vllm -
ConfirmedModels
Daybreak models are now available on AWS
OpenAI and AWS are making Daybreak cybersecurity capabilities available through Amazon Bedrock to support enterprise security workflows.
OpenAI News -
Confirmed
Testing ads in ChatGPT
OpenAI begins testing ads in ChatGPT to support free access, with clear labeling, answer independence, strong privacy protections, and user control.
OpenAI News -
ConfirmedOpen Source
ollama v0.32.8
Muse Glimmer Muse Glimmer is now available on all platforms. Muse Glimmer can power coding agent applications such as Claude Code, Codex, Pi and more, as well as long-running personal assistants such…
GitHub Releases · ollama/ollama Aug 10 · Mon
-
Limited sourcesModels
Introducing Muse Glimmer
Introducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to…
Simon Willison's Weblog -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.227
What's changed Fixed feature flags being evaluated without the user's subscription tier when a session started with an expired login token, which could wrongly prompt Max plan users to enable usage c…
GitHub Releases · anthropics/claude-code -
Limited sources
With new open models, Meta pitches another reboot of its struggling AI strategy
Meta has announced its intention to focus on open-weight large language models. Additionally, the company announced the release of an open model called Muse Glimmer and a promise to open the weights…
Ars Technica AI -
ConfirmedOpen Source
vllm v0.27.0
vLLM v0.27.0 Release Notes Highlights This release features 561 commits from 242 contributors (64 new)! Kimi K3 support with a full stack landing in one release: core model files and kernels ( #50089…
GitHub Releases · vllm-project/vllm -
Limited sources
Amazon backs power plant that may become top source of US climate pollution
Amazon’s artificial intelligence ambitions will soon be partly fueled by a natural gas-burning power plant in Texas that “could become the largest single source of climate pollution in the United Sta…
Ars Technica AI -
Confirmed
What building an AI-native finance function taught me
OpenAI CFO Sarah Friar shares five lessons for building an AI-native finance function, from automated forecasting to stronger controls and AI ROI.
OpenAI News -
ConfirmedOpen SourceAgents & Dev
Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS
Every voice interaction has a latency budget. By the time a user hears your application respond, you've already spent precious milliseconds capturing audio, transcribing speech, running an LLM, retri…
Hugging Face Blog -
Limited sources
Best Local LLMs - August 2026
Wowee!! Just when you thought it couldn't get better for open weight models, we probably have had our best period yet!?!?! Models that rival the closed frontier, Opus level models on non-insane hardw…
Reddit r/LocalLLaMA -
Confirmed
Evolve your marketing with new AI tools
Google Ads We’re adding AI and agentic experiences across Google Ads and Google Analytics to simplify your workflow and expedite your path to business growth. Josh Moser Senior Director, Product Mana…
Google AI Blog -
Confirmed
OpenAI’s letter to Governor Abbott on responsible AI infrastructure in Texas
OpenAI sent Governor Greg Abbott a letter outlining its commitment to responsible AI infrastructure in Texas. The letter supports reliable, transparent growth that benefits Texans.
OpenAI News -
ConfirmedModels
Model ML completes finance work more efficiently with GPT-5.6 Sol
Model ML uses GPT-5.6 Sol to carry finance work from research and analysis through editable, traceable PowerPoint decks and Excel workbooks.
OpenAI News -
Limited sources
Peer review is overwhelmed—can it survive in the AI era?
Jason Semprini was excited about his research on policies mandating that elementary school students receive the human papillomavirus vaccine. HPV causes most cases of cervical cancer, but counterintu…
Ars Technica AI -
ConfirmedOpen Source
transformers Release: v5.15.0
Release v5.15.0 New Model additions Meta Muse Glimmer Muse Glimmer, released today, is Meta’s new multimodal model, especially designed for agentic use cases. Distilled from Muse to 30B parameters, a…
GitHub Releases · huggingface/transformers -
Confirmed
Making Knowledge Distillation Cheap Enough to Run at Scale
Knowledge distillation , training a smaller student model to match the performance of a larger teacher, is a well-known technique in Machine Learning. With the recent wave of open-source Large Langua…
Hugging Face Blog -
Confirmed
Expanding Daybreak as the Cyber Defense Window Narrows
Meet GPT-5.6-Cyber, OpenAI’s cybersecurity-specific model available through Daybreak Red for authorized vulnerability research, exploit validation, and security testing.
OpenAI News -
Confirmed
Putting frontier cyber models in more trusted hands
Approved Daybreak partners can use OpenAI’s frontier cyber models to deliver authorized, governed cybersecurity services to customers.
OpenAI News -
Limited sources
Quoting OpenClaw (running Opus 4.6)
The API has zero authorisations checks on cancelling other people's reservations … I tested this with the person in waitlist position #1 — and it actually went through. So you've moved from #4 to #3…
Simon Willison's Weblog -
ConfirmedBusiness
How Zapier transformed core marketing processes with ChatGPT Work
The enterprise marketing team at Zapier uses ChatGPT Work to reduce the number of drop-offs in its lead funnel, build campaign assets, and automate reporting.
OpenAI News -
Confirmed
Virgin Atlantic sharpens customer journeys with ChatGPT Work
Virgin Atlantic is accelerating research, product planning, and decision-making with ChatGPT Work, helping teams connect signals across the customer journey.
OpenAI News -
Confirmed
Premium seats are coming to ChatGPT Business
Premium seats are coming to ChatGPT Business. Sign up by August 20 to get $100 in workspace credits and unlock higher usage for your team's most demanding work.
OpenAI News -
ConfirmedOpen SourceAgents & Dev
Meta is back with Muse Glimmer: local, agentic, multimodal, and open source
Great news from the OGs of open source LLMs! Muse Glimmer, released today, is Meta’s new multimodal model, especially designed for local agentic use cases. Distilled from Muse to 30B parameters, and…
Hugging Face Blog Aug 9 · Sun
-
Limited sourcesModels
Quoting Claude Opus 5 system prompt
Claude Fable 5 and Claude Mythos 5 were first released on June 9, 2026. On June 12, 2026, Anthropic suspended access to both models to comply with U.S. Department of Commerce export controls; the Dep…
Simon Willison's Weblog -
Limited sources
GitHub Models is now retired
GitHub Models is now retired I missed this news until today, when the GitHub Actions run for my simonw/research repository failed with this error message: GitHub Models is temporarily unavailable as…
Simon Willison's Weblog -
Limited sources
SQLite compressed text-history prototypes
Research: SQLite compressed text-history prototypes I'm perennially interested in options for storing revision histories in relational databases. While out on a dog walk I had a new idea: how about t…
Simon Willison's Weblog -
ConfirmedOpen Source
vllm v0.27.0rc2
v0.27.0rc2
GitHub Releases · vllm-project/vllm Aug 8 · Sat
-
Limited sourcesModelsAgents & Dev
Auto mode is now the default in Claude Code for Pro, Max, and Team plans
Auto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode , to the point that they are making it the default setting for new s…
Simon Willison's Weblog -
Limited sourcesAgents & Dev
Now we have a timeline of the OpenAI accidental attack against Hugging Face
OpenAI gave a last-minute presentation at the Black Hat security on Wednesday about "the Hugging Face Incident" ( previously on this blog). The video was published yesterday. It's short and informati…
Simon Willison's Weblog -
Limited sources
DeepMind’s hurricane breakthrough has surprised weather scientists
In October 2025, a storm brewed over the Caribbean Sea. Weather models differed on its trajectory. Would it remain weak and end up in Haiti, or would it intensify and head to Jamaica? Artificial inte…
Ars Technica AI -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.226
What's changed Bug fixes and reliability improvements
GitHub Releases · anthropics/claude-code -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.225
What's changed Added gateway spend-limit support to Claude Code's usage warning; the limit-reached message now names the cap, its reset time, and the operator's message (requires the gateway on 2.1.2…
GitHub Releases · anthropics/claude-code -
Limited sources
Quoting John Gruber
Me, I try to get into the mindset of playing live music, not recording a studio album. Except when I’m writing a piece where I really want it to be an album. Those aren’t rare , per se, but they’re o…
Simon Willison's Weblog Aug 7 · Fri
-
Limited sourcesModels
Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra)
Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra) On Wednesday I wrote about One-shotting a Raccoon Heist game using Claude Fable 5 , where I had Claude Fable 5 build a full working gam…
Simon Willison's Weblog -
Limited sources
OpenAI’s expensive smart speaker will use moving parts to seem “more alive”
OpenAI’s upcoming smart speaker will probably cost over $300, “people familiar with the matter” have told Bloomberg’s Mark Gurman . The generative AI company has “discussed” charging up to $400, Bloo…
Ars Technica AI -
ConfirmedOpen SourceBusiness
anthropic-sdk-python v0.121.0
0.121.0 (2026-08-07) Full Changelog: v0.120.2...v0.121.0 Features api: add mid-conversation-tool-changes-2026-07-01 beta ( c7d1531 ) api: add support for session budgets, advisor tool, pinned inferen…
GitHub Releases · anthropics/anthropic-sdk-python -
Limited sources
The Tokenpocalypse Is Here: Companies Are Scrambling To Stop Spending So Much on AI
The Tokenpocalypse Is Here: Companies Are Scrambling To Stop Spending So Much on AI There's a fun anecdote from Accenture (apparently via leaked meeting audio recordings) in this 404 Media piece from…
Simon Willison's Weblog -
Confirmed
Responding to the next frontier of critical cyber capabilities
OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.
OpenAI News -
Limited sources
AI chatbots have failed people in crisis. Can that be fixed?
This year alone, there have been numerous known instances—often via lawsuits—of AI chatbots (most often, OpenAI’s ChatGPT) that have gone horrifically wrong. A January lawsuit described the story of…
Ars Technica AI -
Limited sources
ByteDance trains massive AI model in bid to rival Anthropic
ByteDance is training an AI model that could approach the size of Anthropic’s most cutting-edge Mythos system, as Chinese companies continue to narrow the gap with the top US labs. The Chinese tech g…
Ars Technica AI -
Confirmed
How HSP GRUPPE builds AI capabilities for tax advisory
Discover how HSP GRUPPE uses ChatGPT Enterprise to boost productivity, improve work quality, and create more capacity for tax advisory and client service.
OpenAI News -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.224
What's changed Added self-hosted environments: claude self-hosted-runner turns your own machines or containers into a place Claude Code web, mobile, and desktop sessions can run, on Team and Enterpri…
GitHub Releases · anthropics/claude-code -
ConfirmedOpen Source
vllm v0.27.0rc1
v0.27.0rc1
GitHub Releases · vllm-project/vllm Aug 6 · Thu
-
Limited sources
datasette-auth-tokens 0.4a13
Release: datasette-auth-tokens 0.4a13 Upgraded for compatibility with `sqlite-utils 4. Tags: datasette
Simon Willison's Weblog -
Limited sources
Suno hopes to go legit with watermarks for AI-generated music
The Internet is awash in AI content, and it’s not always easy to tell it apart from genuine human creations. While images and videos are perhaps the most obvious type of AI slop, streaming music serv…
Ars Technica AI -
Limited sources
Anthropic will design its own hardware to power Claude
Anthropic is hiring a “custom silicon team” to design chips on which to run its models, the company has revealed. Yesterday, Business Insider noticed a job listing for a senior engineer with experien…
Ars Technica AI -
Limited sources
datasette 1.0a38
Release: datasette 1.0a38 This release fixes a SQL injection security issue that affects Datasette instances that serve a mixture of public and private tables in the same database, with access config…
Simon Willison's Weblog -
Confirmed
WeatherNext: AI model achieves breakthrough in forecasting cyclones
WeatherNext team WeatherNext enables accurate cyclone forecasts that can give an extra day of warning. Now we are open sourcing the model. Predicting how dangerous cyclones develop is a longstanding…
Google DeepMind Blog -
ConfirmedModels
Improving GPT‑5.6 Sol in ChatGPT—and expanding access to GPT-5.6 Luna for free users
ChatGPT introduces improved GPT-5.6 Sol with better accuracy and consistency, plus expanded access for free users and unlimited everyday chats with GPT-5.6 Luna.
OpenAI News -
Confirmed
Working with the American Psychological Association on youth mental health and AI
OpenAI and the American Psychological Association advance evidence-based guidance, resources, and safeguards for responsible AI use and youth mental health.
OpenAI News -
ConfirmedOpen SourceAgents & Dev
claude-code v2.1.223
What's changed Added owner wildcard entries ( "owner/*" ) to the strictKnownMarketplaces and blockedMarketplaces managed settings for allowing or blocking all marketplace repos under a GitHub org Add…
GitHub Releases · anthropics/claude-code -
ConfirmedAgents & Dev
Baseten on Hugging Face Inference Providers 🔥
We're thrilled to share that Baseten is now a supported Inference Provider on the Hugging Face Hub! Baseten joins our growing ecosystem, enhancing the breadth and capabilities of serverless inference…
Hugging Face Blog -
Confirmed
From asking to doing: How the world is putting ChatGPT to work
New OpenAI Signals data shows how people use ChatGPT worldwide, with country-level insights on adoption, usage trends, and evolving behavior.
OpenAI News Aug 4 · Tue
-
ConfirmedResearchBusiness
Third-party cyber evaluations involving OpenAI models
OpenAI explains recent third-party cybersecurity evaluation incidents and outlines new safeguards to strengthen AI model testing and evaluation.
OpenAI News -
Confirmed
The latest AI news we announced in July 2026
Here’s a recap of some of our biggest AI updates from July, including three new Gemini models for building AI agents at scale and Gemini Robotics ER 2 to help robots reason, collaborate, and solve re…
Google AI Blog -
Confirmed
New ways to learn and teach with ChatGPT Work and Codex
Explore new education plugins for ChatGPT Work and Codex that help K–12 teachers, college educators, and students learn, teach, research, and build.
OpenAI News Aug 3 · Mon
-
ConfirmedBusiness
Apple is getting this wrong
OpenAI addresses Apple’s baseless lawsuit, corrects claims about its employees, and shares messages documenting what happened.
OpenAI News -
ConfirmedOpen SourceBusiness
openai-python v2.53.0
2.53.0 (2026-08-03) Features api: Add gpt-5.5 and tool name/namespace to Responses types ( #3569 ) ( dd1202d ) Bug Fixes ci: avoid NumPy source builds and duplicate HTTPX coverage ( #3573 ) ( b58332f…
GitHub Releases · openai/openai-python -
ConfirmedOpen SourceBusiness
openai-python v2.52.1
2.52.1 (2026-07-31) Full Changelog: v2.52.0...v2.52.1 Chores ci: pin setup-uv v5 to its underlying commit ( #3560 ) ( cbdc98b )
GitHub Releases · openai/openai-python -
ConfirmedAgents & Dev
Inside our 353,000-person vibe coding course
Google and Kaggle’s “AI Agents: Intensive Vibe Coding” course had expert-led sessions, technical whitepapers and notebooks, and capstone projects. Anant Nawalgaria Group AI Product Manager, Founder o…
Google AI Blog -
ConfirmedCreative
How we built a realtime system for responsive voice AI in six months
GPT-Live enables continuous voice interaction with AI, using a turnless speech model and low-latency architecture for faster, more natural conversations.
OpenAI News -
Confirmed
Circles powers telco personalization with OpenAI technology
Circles uses the OpenAI API and Codex to power AI-native telco experiences, increasing ARPU by 22%, reducing churn by 9%, and improving development efficiency.
OpenAI News Aug 1 · Sat
-
ConfirmedInfra
Ten advances in mathematics and theoretical computer science
OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.
OpenAI News