AI
Weekly AI Review β August 23, 2026
π Model & Product Releases DeepSeek opens a multimodal API with an experimental vision model β DeepSeek released DeepSeek-V4-Flash-Vision-Exp on August 21 and opened its multimodal API to developers. The company describes a sparse mixture-of-experts model with 13 billion active parameters out of 284 billion total, adding image understanding while matching the text-only V4-Flash on […]
Weekly AI Review β August 16, 2026
π Model & Product Releases Google prices Gemini 3.7 Flash at half its predecessor β Google released the coding and agent model on August 13 with a one-million-token input window, a 64,000-token output limit and multimodal input. Introductory API pricing is $0.75 per million input tokens and $3.75 per million output through the end of […]
Weekly AI Review β August 9, 2026
π Model & Product Releases Alibaba ships Qwen3.8-Max and promises open weights β Alibaba released the 2.4-trillion-parameter mixture-of-experts model on August 3, with 95 billion parameters active per token and a one-million-token context window. Listed API pricing is $2 per million input tokens and $6 per million output, against $5 and $30 for GPT-5.6 Sol. […]
Weekly AI Review β August 2, 2026
π Model & Product Releases Moonshot releases Kimi K3 weights, the largest open model yet β Moonshot AI published full checkpoints for Kimi K3 on July 27, a 2.8-trillion-parameter mixture-of-experts model with 104 billion active parameters and a one-million-token context window. The download runs to roughly 1.56 TB across 96 shards on Hugging Face, released […]
AI News of the Week β 18/07/2026 (REST API test)
Models and releases OpenAI ships the GPT-5.6 family β OpenAI launched its new flagship GPT-5.6 models in three sizes β Luna, Terra and Sol β priced from $1 to $5 per million input tokens, all with a 1M-token context window. The release adds programmatic tool calling, multi-agent orchestration and prompt cache breakpoints to the API. […]
RAG vs Obsidian-Based Memory: A Comparative Analysis
Two paradigms for giving language models knowledge they didn’t have at training time β one statistical, one curated. When does each one win? Contents Introduction Foundations of RAG Foundations of an Obsidian-Based Memory System Comparison When to Use Which Conclusion References 1. Introduction Large language models have a memory problem. Their parametric memory β the […]
First large-scale Blackwell B200 AI cluster goes live in Canada
Alpha Compute announced on 8 May 2026 the handover of its first large-scale NVIDIA Blackwell deployment: a 504-chip B200 server cluster located in Canada, now in final testing and ready for AI compute customers. It is one of the first commercial-scale Blackwell clusters to come online outside the hyperscalers. Blackwell-architecture GPUs pack 208 billion transistors […]
US AI Safety Institute wins pre-release access to Google, Microsoft, xAI models
The US Center for AI Standards and Innovation (CAISI), the successor body to the US AI Safety Institute, announced on 5 May 2026 that it had signed agreements with Google DeepMind, Microsoft, and Elon Musk’s xAI to evaluate their frontier AI models before public release. The deals expand a framework that already covered OpenAI and […]
Google previews ‘Gemini Intelligence’ for Android ahead of I/O
On 13 May 2026, days before its I/O developer conference, Google previewed Gemini Intelligence: a system-level layer that pushes Gemini into the core of Android. The rollout begins this summer on the Samsung Galaxy and Google Pixel lines, then expands across compatible Android devices β phones, watches, cars, glasses, and laptops β through the rest […]
Google releases Gemma 4, betting on intelligence-per-parameter
Google quietly shipped the Gemma 4 family on 4 May 2026, the next generation of its open-weight models. The release continues Google’s twin-track strategy: keep Gemini at the frontier of closed models while using Gemma to set the bar for what open-weight systems can do. Gemma 4 is explicitly engineered for advanced reasoning and agentic […]
