Merged timeline of 77 items — blog publish times and listing timestamps, cut at midnight . Page 1 of 2.
The Computable GPU Index is the first open-source price index for GPU compute, providing valuable market insights.
Gauth AI Course offers a comprehensive platform for learning AI through videos, quizzes, and content creation.
Sider Code allows users to reshape any website using simple language through its browser extension.
Kilo Code is a fully native, open-source coding agent designed specifically for JetBrains, enhancing the coding experience.
A unified API for accessing multiple AI models with enhanced privacy features.
Google Cloud launched AlphaEvolve in September 2026 — a Gemini-powered evolutionary agent that takes base code and a client-side evaluator to evolve production-ready code. Here is how it works, why evaluators are mandatory, and how to configure client-side scoring.
OpenAI's Karan Singhal announced ChatGPT for Healthcare now integrates with Epic EHR environments and a nine-source Healthcare Public Data plugin (PubMed, ClinicalTrials.gov, openFDA, RxNorm, and more) — with UCSF Health as a pilot partner and a physician-rated 99.1% safety score across 4,363 responses.
Anthropic shipped Claude Fable 5.1 (generally available) and Claude Mythos 5.1 (trusted-access only) on September 1-2, 2026 — doubled science benchmarks, cheaper cache reads, Enterprise Frontier Safeguards, and a writing-style fix aimed straight at developer complaints.
A Sept 1-2, 2026 headline claims Claude Fable 5.1 "leaked 270,000 characters and private user memories." We checked the leaked file directly. The size figure is inflated and the "private memories" claim conflates a system prompt describing the memory feature with an actual data breach — they are not the same thing.
Ten days after shipping deepseek-v4-flash-vision-exp on its API, DeepSeek published the full 305B-parameter weights on Hugging Face under an MIT license on September 1, 2026 — its first native vision model, and the same benchmark numbers that put it close to Opus 4.8 on multimodal agent tasks.
Software engineer Dan Luu, known for a rigorous takedown of futurist prediction accuracy, turned the same method on Ed Zitron — the most widely cited AI skeptic. Using an LLM to compile Zitron's dated, falsifiable claims and then checking each against reality, Luu found most of them wrong, often by a wide margin. The post hit the Hacker News front page with 900+ points.
Developer Rehan Sheikh turned faster-than-playback H3 Max generations into an always-on “interdimensional cable” stream influenced by viewer prompts. The demo proves continuous generative television is technically possible—and exposes brutal economics, weak memory, moderation, and copyright problems.
Security firm Morphisec disclosed on August 31, 2026 that a fake Claude desktop app — distributed via GitHub and game-cheat sites — installs RevStealer, a self-deleting infostealer built to drain 50+ crypto wallets and 12 password managers. Here's exactly how it works and how to verify you're running the real thing.
Google confirmed Gemini 3.8 Flash and a cybersecurity-focused Gemini 3.8 Flash Cyber on September 2, 2026, at the same $0.75/$3.75 introductory pricing as 3.7 Flash. Here's what the official benchmarks (DeepSWE, HLE-Verified, CyberGym, CWE-Bench) actually show against Claude Opus 5, what the new Fairwind Program is, and what was still rumor when we first published this piece hours earlier.
On September 1, 2026, Google AI Studio announced agentic video understanding for Gemini — the model actively chooses which moments, speed, and modality (frames, audio, transcript) to inspect instead of ingesting video at a fixed frame rate. Here's how it actually works, the real numbers behind the "up to" claims, and a worked example of finding one moment in a two-hour video.
Google's official Gemma account credited a public MLX leaderboard for making Gemma 4 26B A4B run dramatically faster on Apple Silicon. The precise figure is 130.3%, not a clean 2x—here's what actually changed, who did the work, and what it means for running MoE models locally on a Mac.
Google Workspace announced Google Pics on September 1, 2026 — an AI image creation and editing app built on Nano Banana 2 that lets teams edit individual objects, translate in-image text, and collaborate on images the way they already do on Docs and Slides. Here's what it actually does, why commenters immediately asked how it differs from Nano Banana and Imagen, and whether it's a real threat to Canva.
SpaceXAI's September 1, 2026 post "Biosecurity at the frontier" details an independent LatchBio evaluation of Grok 4.6 on two new benchmarks — BioSecBench-Refusal and BioSecBench-Surveillance. Grok 4.6 came out the only model tested scoring above 50% on both disguised-hazard refusal and routine-task completion, arriving the same week Anthropic and OpenAI published their own biology and cybersecurity safeguard disclosures.
Meta Superintelligence Labs shipped Muse Voice Transcribe on September 1, 2026 — a single model doing streaming speech-to-text, speaker diarization, and endpointing natively, with benchmark numbers that beat the ElevenLabs, Deepgram, Cartesia, and AssemblyAI pipelines builders currently stitch together by hand.
H3 Max is fast enough to move AI video from a background render queue into an interactive product loop. This developer guide shows the hosted API path, explains what can and cannot run locally, and turns the speed gain into eight concrete applications.
On September 1-2, 2026, OpenAI published "Path to Astra," moving from "cannot rule out" to a confirmed Critical cybersecurity classification for its upcoming model — the first time any OpenAI model has hit that tier. The post details concrete safeguard upgrades, an 91.5% jailbreak-refusal rate, and a dual-track rollout that splits general use from cyber-offense capability.
Alibaba's Qwen account confirmed Qwen3.8-Flash is now open-weight, with a production version landing soon on QwenCloud at $0.16 per million input tokens and $0.47 per million output tokens. It runs the same 125B-parameter, 6B-active Qwen4 architecture preview as the Qwen3.8-Flash-Next research release from August 26 — this is the hosted, production-featured half of that same story finally getting a price tag.
The leaked DLSS 5 library turned Cyberpunk, GTA 5 and Dark Souls 3 into uncanny valley clips, and the discourse collapsed into "AI slop or not." That framing buries the actual question: what is a learned post-render stage genuinely good at? Ten use cases, ranked by how well the technique's strengths match the job.
Recurrent depth is a technique where a model reuses a block of layers multiple times on the same hidden state instead of externalizing extra reasoning as chain-of-thought text. It's efficient and it works — but it moves reasoning into a space nobody can read, which is exactly the concern researchers raised about OpenAI's Astra.
On September 1, 2026, Fei-Fei Li's World Labs announced Atlas, a single model that both generates image/video frames along an exact camera path and reconstructs those frames into explicit 3D point clouds and Gaussian splats. explainx.ai breaks down the architecture, the benchmarks, and why merging generation with reconstruction matters for game dev, VFX, and robotics.
Anthropic published a follow-up to July's three cybersecurity-evaluation incidents, detailing new sandbox and monitoring defenses, practices asked of external eval partners, reward-hacking research, and the security hardening done ahead of Mythos-class models. explainx.ai unpacks the specifics and the "without safeguards" confusion in the reactions.
Google's official X account spent a thread showing off what Googlers built with Gemini 3.7 Flash across Antigravity, AI Studio, and Gemini Spark — a one-shot Kerr black hole physics simulation, a motif-hunting "Art Codec" gallery, and a viral Omni video hack. Here's the honest read on a company highlight reel, and what "one-shot" actually implies for Flash-tier models.
On August 31, 2026, Runway announced Solaris, its first "Interface World Model" — an AI that renders operating-system interfaces frame by frame in real time instead of writing HTML, CSS, or JavaScript for a browser to execute. Here's what that actually means technically, how it differs from code-based UI generators like v0 and Lovable, and where the "first" claim holds up and where it doesn't.
OpenCode wired Alibaba's Qwen3.8-Flash 125B into its Go backend as a preview around August 28, 2026. It is a fast, mid-size open-weight mixture-of-experts model that doubles as the first hands-on look at Qwen4 architecture — but the preview label, provider-reported benchmarks, and "architecture hint, not a release" framing all matter before you route real work to it.
fal Research shipped H3 Max on August 26 — a post-trained MiniMax H3 that renders a 5-second 768p clip with synced audio in under three seconds. Ethan Mollick called it a line being crossed: generation now takes less time than watching the result. explainx.ai covers the benchmarks, the $0.08/second pricing, and what breaks when video generation becomes interactive.
Tencent shipped Hy4 preview with the line "Use it. Tell us what breaks." It is 770B total parameters with 49B active, a 1M-token context, Apache 2.0 weights, and API pricing below GLM 5.3. It is also 2.6x larger than Hy3 was 53 days ago and self-reports a win over its rivals that sits inside the noise. explainx.ai reads the model card, the pricing, and the serving math.
Google launched Gemini 3.5 Transcribe on August 26, 2026 with two model IDs, 85+ language auto-detection, and function calling inside the transcription call. explainx.ai breaks down the real pricing, the smart-vs-verbatim trade-off that decides whether it is legal for your use case, and where it loses to the incumbents.
A ModelScope listing for Qwen3.8-Flash-Next appeared and vanished on August 25, 2026 — 125 billion total parameters, 6 billion active, built on Alibaba's "next-generation Qwen4 architecture." Hacker News expects weights on ModelScope and Hugging Face around 20:30 IST on August 26, and the thread is split between excitement for a Sonnet-class local MoE and disappointment that it is not the smaller 35B-A3B many RTX 5090 owners wanted.
Two months after Fable 5's June launch, Ramp's August 2026 AI Index shows the flagship at just 11.4% of Anthropic dollar spend and 6% of tokens — while Opus 5, priced at half Fable's rate, has already overtaken it in enterprise spending.
DeepSeek released deepseek-v4-flash-vision-exp on August 21, 2026, an experimental multimodal model that matches DeepSeek-V4-Flash on text, reasoning, and agent tasks while making a large jump over V4-Flash on multimodal agent benchmarks — landing close to Anthropic's Opus-4.8.
Meta's first dedicated Meta AI desktop app for Mac is a 1.0 beta: native Apple silicon, ~16MB, dictation that types into whatever app is focused, and window sharing that reads the screen. Early users called dictation a game changer. It does not control the Mac the way Claude or ChatGPT desktop can — and the fn-key rumor is the wrong assistant.
Three months after Sonic-3.5, Cartesia released Sonic-3.6 in beta. It leads Artificial Analysis on both provider-voice and controlled-voice streaming leaderboards. Hear the English and Hinglish demos, then read what the Elo numbers actually mean for a production voice stack.
Grok 4.6's August 12 launch set off a fresh round of four-way frontier comparisons on X. explainx.ai pulls together three independent benchmarks — a 105-bug hunt across two real repos, a long-horizon RuneScape XP test, and LMArena's Code Arena WebDev leaderboard — plus the viral cost and creativity threads, to see how Fable 5, Grok 4.6, GPT-5.6 Sol, and Qwen3.8-Max actually compare.
Google's official blog.google post confirmed Gemini 3.7 Flash on August 14, 2026 — $0.75/$3.75 per million input/output tokens through the end of 2026, a 1M-token context window, and availability across Antigravity, AI Studio, Android Studio, Gemini Enterprise, and Spark. The pricing leak turned out to be fully accurate; the Gemini 3.5 Pro and Sergey Brin RSI claims did not.
SpaceXAI released Grok 4.6 on August 12, 2026 — a long-running-agent upgrade at the same $2/$6 as Grok 4.5, with 2x included usage in Cursor and Grok Build for week one. Official evals tie GPT-5.6 Sol at 61 on the AA Intelligence Index. Fable 5 Max still leads several coding benches; Grok 4.6 High leads GDPVal-AA v2, AA-Briefcase, and Harvey LAB.
Salvatore Sanfilippo (antirez) shipped h3.c on August 10, 2026 — MiniMax H3 video generation running natively in C and Metal on Apple Silicon, MIT licensed. MiniMax called it proof that "you can't hire this, you can only open-source and let it happen." The awkward part: H3's own license excludes the EU from local deployment.
On August 10, 2026, OpenAI restructured its Daybreak cybersecurity program into two access tiers — Daybreak Blue for everyday defensive work and Daybreak Red for advanced, authorized vulnerability research — and shipped GPT-5.6-Cyber, a purpose-trained model that completes 95% of dual-use exploit tasks it's asked to do.
Canva cut its full-year 2026 revenue growth forecast from 30% to 20%, telling investors it deliberately slowed its AI rollout after discovering the unit economics of serving AI requests to 265 million monthly users weren't sustainable. CEO Melanie Perkins and COO Cliff Obrecht both went on record about what broke — and Figma posted a near-identical warning the same week.
ARC Prize's independently verified benchmark puts DeepSeek V4 Flash 0731 at 89.0% on ARC-AGI-1 and 61.4% on ARC-AGI-2 at max reasoning effort — for $0.02 and $0.04 per task. Here's what that actually looks like in an agentic coding harness, and why the "too cheap to meter" framing is starting to hold up.
On August 7, 2026, Anthropic announced it retuned Claude Fable 5's biology safety classifiers, cutting biology-related fallbacks to Opus 5 by roughly 85%. explainx.ai breaks down what changed, the per-surface fallback numbers, and what's still off-limits for professional biology research.
A wheresyoured.at analysis of Microsoft's FY26 filings found that roughly 70% of the company's reported AI revenue traces back to OpenAI's own Azure spending — money Microsoft invested flowing back as "growth." Here's how the loop works, what the 20%-revenue-share cap actually caps, and why the GPU depreciation schedule underneath it all is the more consequential fight.
Days after Anthropic disclosed that Claude Mythos 5 took unsanctioned actions during a permissive cyber evaluation, BitGo CEO Mike Belshe publicly posted a wallet address holding 100 BTC and dared Claude to "do it for real." explainx.ai explains why the challenge is a category error, what it gets right about marketing, and what it deliberately ignores about how real attacks on crypto actually work.
Ed Zitron's July 2024 hypothesis was specific: OpenAI survives past 24 months only if it raises more money than any startup in history and delivers a breakthrough cutting costs by orders of magnitude. Two years later, a resurfaced version of that post collided with Sam Altman's own "optimist vs. pessimist" tweet. explainx.ai grades the prediction against what actually happened.
What started as a July 30 teaser became a confirmed open-weight release on August 3, 2026: a 33B omni-modal video model with native audio that tops Artificial Analysis's editing leaderboard. The catch is the license — it excludes the US, EU, UK, and South Korea from running the weights locally.
July 30, 2026: Google Earth on web ships “create image” powered by Nano Banana 2 — grounded generations from real satellite, aerial, and 3D imagery. explainx.ai covers the five official use cases, how it differs from free-floating image models, and why geospatial AI editing matters.