Sunday, 30 August, 2026

Weekly AI Review β€” August 23, 2026

πŸš€ Model & Product Releases DeepSeek opens a multimodal API with an experimental vision model – DeepSeek released DeepSeek-V4-Flash-Vision-Exp on August 21 and opened its multimodal API to developers. The company describes a sparse mixture-of-experts model with 13 billion active parameters out of 284 billion total, adding image understanding while matching the text-only V4-Flash on […]

Weekly AI Review β€” August 16, 2026

πŸš€ Model & Product Releases Google prices Gemini 3.7 Flash at half its predecessor – Google released the coding and agent model on August 13 with a one-million-token input window, a 64,000-token output limit and multimodal input. Introductory API pricing is $0.75 per million input tokens and $3.75 per million output through the end of […]

Weekly AI Review β€” August 9, 2026

πŸš€ Model & Product Releases Alibaba ships Qwen3.8-Max and promises open weights – Alibaba released the 2.4-trillion-parameter mixture-of-experts model on August 3, with 95 billion parameters active per token and a one-million-token context window. Listed API pricing is $2 per million input tokens and $6 per million output, against $5 and $30 for GPT-5.6 Sol. […]

Weekly AI Review β€” August 2, 2026

πŸš€ Model & Product Releases Moonshot releases Kimi K3 weights, the largest open model yet – Moonshot AI published full checkpoints for Kimi K3 on July 27, a 2.8-trillion-parameter mixture-of-experts model with 104 billion active parameters and a one-million-token context window. The download runs to roughly 1.56 TB across 96 shards on Hugging Face, released […]

AI News of the Week β€” 18/07/2026 (REST API test)

Models and releases OpenAI ships the GPT-5.6 family – OpenAI launched its new flagship GPT-5.6 models in three sizes β€” Luna, Terra and Sol β€” priced from $1 to $5 per million input tokens, all with a 1M-token context window. The release adds programmatic tool calling, multi-agent orchestration and prompt cache breakpoints to the API. […]

RAG vs Obsidian-Based Memory: A Comparative Analysis

Two paradigms for giving language models knowledge they didn’t have at training time β€” one statistical, one curated. When does each one win? Contents Introduction Foundations of RAG Foundations of an Obsidian-Based Memory System Comparison When to Use Which Conclusion References 1. Introduction Large language models have a memory problem. Their parametric memory β€” the […]

First large-scale Blackwell B200 AI cluster goes live in Canada

Alpha Compute announced on 8 May 2026 the handover of its first large-scale NVIDIA Blackwell deployment: a 504-chip B200 server cluster located in Canada, now in final testing and ready for AI compute customers. It is one of the first commercial-scale Blackwell clusters to come online outside the hyperscalers. Blackwell-architecture GPUs pack 208 billion transistors […]

US AI Safety Institute wins pre-release access to Google, Microsoft, xAI models

The US Center for AI Standards and Innovation (CAISI), the successor body to the US AI Safety Institute, announced on 5 May 2026 that it had signed agreements with Google DeepMind, Microsoft, and Elon Musk’s xAI to evaluate their frontier AI models before public release. The deals expand a framework that already covered OpenAI and […]

Google previews ‘Gemini Intelligence’ for Android ahead of I/O

On 13 May 2026, days before its I/O developer conference, Google previewed Gemini Intelligence: a system-level layer that pushes Gemini into the core of Android. The rollout begins this summer on the Samsung Galaxy and Google Pixel lines, then expands across compatible Android devices β€” phones, watches, cars, glasses, and laptops β€” through the rest […]

Google releases Gemma 4, betting on intelligence-per-parameter

Google quietly shipped the Gemma 4 family on 4 May 2026, the next generation of its open-weight models. The release continues Google’s twin-track strategy: keep Gemini at the frontier of closed models while using Gemma to set the bar for what open-weight systems can do. Gemma 4 is explicitly engineered for advanced reasoning and agentic […]