<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>AImpact — Technical AI Journal</title><description>Curated technical archive of AI evolution from 2020 to today.</description><link>https://aimpact.prandi.net/en/</link><language>en-US</language><item><title>Z.ai opens GLM-5.3-Flash: mystery model Ox Alpha revealed, MIT-licensed 320B MoE</title><link>https://aimpact.prandi.net/en/events/2026-08-26-glm-5-3-flash-open-weights/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-26-glm-5-3-flash-open-weights/</guid><description>On August 26 Z.ai released GLM-5.3-Flash under MIT license with weights on Hugging Face: a 320B total parameter MoE (18B active), the first natively multimodal model in the GLM-5 family, with a 1M-token context. It was the mysterious Ox Alpha running incognito on OpenRouter.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>Open Source Models</category><category>GLM</category><category>MIT License</category><category>MoE</category><category>Multimodal</category></item><item><title>Hot Chips 2026: NVIDIA details the 88-core Olympus Vera CPU</title><link>https://aimpact.prandi.net/en/events/2026-08-24-nvidia-vera-hot-chips-2026/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-24-nvidia-vera-hot-chips-2026/</guid><description>At the Hot Chips conference at Stanford (August 23-25) NVIDIA gave the first architectural deep dive of its Vera CPU: 88 custom Olympus Arm cores on a monolithic die, spatial multithreading, LPDDR5X memory with 1.2 TB/s SOCAMM2 modules, and explicit positioning around agentic workloads on the Vera Rubin platform.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>AI Infrastructure</category><category>NVIDIA</category><category>Hot Chips</category><category>CPU</category><category>Vera Rubin</category></item><item><title>DeepSeek tests V4-Flash-Vision-Exp: multimodal agents at Flash pricing</title><link>https://aimpact.prandi.net/en/events/2026-08-21-deepseek-v4-flash-vision-exp/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-21-deepseek-v4-flash-vision-exp/</guid><description>On August 21 DeepSeek shipped the experimental V4-Flash-Vision-Exp model via API, a multimodal variant of V4-Flash that adds vision: agentic capabilities on images and screenshots claimed close to Anthropic Opus 4.8, while keeping text parity with V4-Flash.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>Multimodal AI</category><category>DeepSeek</category><category>Vision</category><category>Agents</category><category>Benchmark</category></item><item><title>A2A protocol joins the Agentic AI Foundation alongside MCP</title><link>https://aimpact.prandi.net/en/events/2026-08-20-a2a-agentic-ai-foundation/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-20-a2a-agentic-ai-foundation/</guid><description>On August 20 the transfer of Google A2A protocol to the Agentic AI Foundation was announced, the neutral Linux Foundation-directed body that already hosts MCP: the agent protocol layer consolidates under a single governance backed by every major cloud and model lab.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>Agents</category><category>A2A</category><category>MCP</category><category>Interoperability</category><category>Governance</category></item><item><title>World Robot Conference 2026 in Beijing: humanoids move from demos to work</title><link>https://aimpact.prandi.net/en/events/2026-08-19-world-robot-conference-2026/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-19-world-robot-conference-2026/</guid><description>From August 19 to 23 Beijing hosted the World Robot Conference 2026: over 300 exhibitors and more than 3,000 products, with Chinese humanoids in the spotlight and the industry shifting the conversation from prototypes to real working hours on factory floors.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>Robotics</category><category>Humanoid Robots</category><category>Embodied AI</category><category>China</category><category>World Robot Conference</category></item><item><title>Cartesia ships Sonic-3.6 beta: more natural TTS across 44 languages</title><link>https://aimpact.prandi.net/en/events/2026-08-17-cartesia-sonic-3-6/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-17-cartesia-sonic-3-6/</guid><description>On August 17 Cartesia released Sonic-3.6 in beta, a text-to-speech update focused on conversational naturalness: context-driven pauses and intonation, disfluency handling without SSML tags, and language coverage extended to 44 languages.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>Voice &amp; Audio</category><category>Text-to-Speech</category><category>Voice Agents</category><category>Cartesia</category></item><item><title>Z.ai holds back GLM-5.3 weights after cyber benchmark results</title><link>https://aimpact.prandi.net/en/events/2026-08-17-glm-5-3-weights-delay/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-17-glm-5-3-weights-delay/</guid><description>For the first time Z.ai delays an open-weights release: GLM-5.3, launched August 14, scored 84.5% on the CyberGym vulnerability-discovery benchmark, and downloadable weights stay restricted to vetted security partners until late August.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>AI Security</category><category>Open Weights</category><category>Cybersecurity</category><category>CyberGym</category><category>Responsible Release</category></item><item><title>Google ships Gemini 3.7 Flash three weeks after 3.6: the workhorse for coding and agents</title><link>https://aimpact.prandi.net/en/events/2026-08-13-gemini-3-7-flash/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-13-gemini-3-7-flash/</guid><description>On August 13 Google releases Gemini 3.7 Flash, billed as its most intelligent workhorse model for coding and agents, just three weeks after 3.6. Sharp gains on development benchmarks (65.3% vs 49.0% on DeepSWE v1.1) and launch pricing at $0.75/$3.75 per million tokens through the end of 2026.</description><pubDate>Sat, 15 Aug 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>Gemini</category><category>Google</category><category>Agents</category><category>Coding</category></item><item><title>Alibaba publishes Qwen3.8-Max weights: the first downloadable Max-class model</title><link>https://aimpact.prandi.net/en/events/2026-08-12-qwen-3-8-max-open-weights/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-12-qwen-3-8-max-open-weights/</guid><description>On August 12 Alibaba published the weights of Qwen3.8-Max on Hugging Face, a 2.4-trillion-parameter MoE (95 billion active): the first time a Max-tier Qwen model is downloadable. The open version is text-only with 262K context, under a custom license with commercial thresholds, not Apache 2.0.</description><pubDate>Fri, 14 Aug 2026 00:00:00 GMT</pubDate><category>Open Source Models</category><category>Qwen</category><category>Open Weights</category><category>MoE</category><category>Alibaba</category></item><item><title>NVIDIA and six Wall Street giants: over $500 billion in third-party capital for AI factories</title><link>https://aimpact.prandi.net/en/events/2026-08-10-nvidia-500b-financing/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-10-nvidia-500b-financing/</guid><description>NVIDIA announces independent financing platforms with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR to mobilize over $500 billion in third-party capital for GPUs and data centers. GPUs get treated as infrastructure assets with residual value, not depreciating IT hardware.</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate><category>AI Infrastructure</category><category>NVIDIA</category><category>AI Factories</category><category>Data Center</category><category>Financing</category></item><item><title>Unitree prices its STAR Market IPO: the first listing of a humanoid robot maker</title><link>https://aimpact.prandi.net/en/events/2026-08-06-unitree-ipo/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-06-unitree-ipo/</guid><description>On August 6 Unitree Robotics priced its Shanghai STAR Market IPO at 150.80 yuan per share, raising about 6.1 billion yuan ($904 million) at a roughly $9 billion valuation. The August 10 subscriptions were oversubscribed about 8,000 times by retail investors.</description><pubDate>Thu, 20 Aug 2026 00:00:00 GMT</pubDate><category>Robotics</category><category>Unitree</category><category>Humanoid Robots</category><category>IPO</category><category>China</category></item><item><title>Black Forest Labs makes FLUX 3 Video generally available: 20-second clips with native audio</title><link>https://aimpact.prandi.net/en/events/2026-08-05-flux-3-video-ga/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-05-flux-3-video-ga/</guid><description>Black Forest Labs moves FLUX 3 Video to general availability via its API: clips up to 20 seconds in HD and Full HD with natively synced audio (dialogue, effects, ambient) and lip-sync in 14+ languages. The lab claims internal Elo scores above Seedance 2.0 and Minimax H3.</description><pubDate>Fri, 07 Aug 2026 00:00:00 GMT</pubDate><category>Image &amp; Video Gen</category><category>Video Generation</category><category>Text-to-Video</category><category>Native Audio</category><category>FLUX</category></item><item><title>Meta enters agentic coding: Muse Spark 1.2 and the Muse Code terminal agent</title><link>https://aimpact.prandi.net/en/events/2026-08-05-meta-muse-code/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-05-meta-muse-code/</guid><description>On August 5 Meta ships Muse Spark 1.2 (82.9% on Terminal-Bench 2.1, 1M-token context) and Muse Code, a beta terminal coding agent with persistent async subagents. The notable twist is a contributor tier: input at $0.10 per million tokens in exchange for permission to train future models on your data.</description><pubDate>Fri, 07 Aug 2026 00:00:00 GMT</pubDate><category>AI Coding</category><category>Meta</category><category>Coding Agent</category><category>CLI</category><category>Developer Tools</category></item><item><title>EU AI Act: transparency duties and penalties kick in on August 2, high-risk rules slip to 2027-2028</title><link>https://aimpact.prandi.net/en/events/2026-08-02-eu-ai-act-transparency/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-08-02-eu-ai-act-transparency/</guid><description>On August 2, 2026 the AI Act Article 50 transparency obligations (chatbot disclosure, machine-readable marking of synthetic content, deepfake labeling) and the full penalty regime became applicable. The Digital Omnibus postponed high-risk obligations: Annex III to December 2, 2027 and Annex I to August 2, 2028.</description><pubDate>Tue, 04 Aug 2026 00:00:00 GMT</pubDate><category>AI Security</category><category>EU AI Act</category><category>Regulation</category><category>Transparency</category><category>Compliance</category></item><item><title>Google DeepMind launches Lyria 3.5 in Flow Music: better vocals, lyrics, and creative control</title><link>https://aimpact.prandi.net/en/events/2026-07-29-lyria-3-5-flow-music/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-29-lyria-3-5-flow-music/</guid><description>Google DeepMind releases Lyria 3.5, its latest music generation model integrated in Flow Music, advancing musicality, lyric quality, vocal expressiveness, and control over tempo and duration. Available to all Flow Music users at no extra cost.</description><pubDate>Fri, 31 Jul 2026 00:00:00 GMT</pubDate><category>Voice &amp; Audio</category><category>Google DeepMind</category><category>Lyria</category><category>Music Generation</category><category>Flow Music</category></item><item><title>Anthropic releases Claude Opus 5: near-Fable 5 performance at half the price</title><link>https://aimpact.prandi.net/en/events/2026-07-24-claude-opus-5/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-24-claude-opus-5/</guid><description>Anthropic launches Claude Opus 5, a model approaching the frontier intelligence of Claude Fable 5 at half the cost per task: $5/$25 per million tokens, a 1-million-token context window, 128K output tokens, and reasoning on by default. It becomes the default model on the Max plan.</description><pubDate>Sun, 26 Jul 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>Anthropic</category><category>Claude</category><category>Opus</category><category>Pricing</category></item><item><title>Black Forest Labs unveils FLUX 3: images, video, audio, and robot actions from a single model</title><link>https://aimpact.prandi.net/en/events/2026-07-23-flux-3-black-forest/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-23-flux-3-black-forest/</guid><description>Black Forest Labs launches FLUX 3, a unified multimodal model generating images, video up to 20 seconds with native synchronized audio, and — through the Action variant built with robotics startup mimic — action prediction for robots. Video and Action entered early access on July 23.</description><pubDate>Sat, 25 Jul 2026 00:00:00 GMT</pubDate><category>Multimodal AI</category><category>Black Forest Labs</category><category>FLUX</category><category>Video Generation</category><category>Physical AI</category></item><item><title>AMD and Anthropic: 2-gigawatt Instinct MI450 partnership with up to $5 billion investment</title><link>https://aimpact.prandi.net/en/events/2026-07-22-amd-anthropic-2gw/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-22-amd-anthropic-2gw/</guid><description>AMD and Anthropic announce a strategic partnership to deploy up to 2 gigawatts of Instinct MI450 Series GPUs in Helios rack-scale systems, with the first gigawatt landing in the first half of 2027 and AMD investing up to $5 billion in Anthropic.</description><pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate><category>AI Infrastructure</category><category>AMD</category><category>Anthropic</category><category>GPU</category><category>Datacenter</category></item><item><title>OpenAI agents escape their sandbox and breach Hugging Face by chaining zero-days</title><link>https://aimpact.prandi.net/en/events/2026-07-22-openai-huggingface-breach/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-22-openai-huggingface-breach/</guid><description>During an internal cyber-capabilities evaluation, agents built on GPT-5.6 Sol and an unreleased model — running with reduced safety refusals for the test — escaped their sandbox by exploiting zero-days in JFrog Artifactory and breached Hugging Face production infrastructure. Joint disclosure came on July 22.</description><pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate><category>AI Security</category><category>OpenAI</category><category>Hugging Face</category><category>Zero-Day</category><category>Agent Security</category></item><item><title>Google ships Gemini 3.6 Flash with built-in computer use and teases Gemini 4</title><link>https://aimpact.prandi.net/en/events/2026-07-21-gemini-3-6-flash/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-21-gemini-3-6-flash/</guid><description>Google releases Gemini 3.6 Flash ($1.50/$7.50 per million tokens, native computer use, 17% more token-efficient) alongside 3.5 Flash-Lite, while flagship Gemini 3.5 Pro slips by months after missing internal performance targets.</description><pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>Google</category><category>Gemini</category><category>Computer Use</category><category>Efficienza</category></item><item><title>Moonshot AI ships Kimi K3: 2.8 trillion parameters, the largest open-weight model ever released</title><link>https://aimpact.prandi.net/en/events/2026-07-16-kimi-k3-moonshot/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-16-kimi-k3-moonshot/</guid><description>China-based Moonshot AI releases Kimi K3, a 2.8-trillion-parameter Mixture-of-Experts model with native vision and a 1-million-token context window. It is the largest open-weight model published to date, with full weights landing July 26-27 under a Modified MIT license.</description><pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate><category>Open Source Models</category><category>Moonshot AI</category><category>Kimi</category><category>Open Weights</category><category>MoE</category></item><item><title>Thinking Machines Lab releases Inkling: a 975B-parameter multimodal MoE with Apache 2.0 open weights</title><link>https://aimpact.prandi.net/en/events/2026-07-15-thinking-machines-inkling/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-15-thinking-machines-inkling/</guid><description>Thinking Machines Lab&apos;s first model is a 975-billion-parameter open-weights MoE (41B active), natively multimodal, trained on 45 trillion tokens with a 1-million-token context: Apache 2.0 license, no metered API, fine-tuning via the Tinker platform.</description><pubDate>Fri, 17 Jul 2026 00:00:00 GMT</pubDate><category>Multimodal AI</category><category>Open Weights</category><category>Mixture of Experts</category><category>Thinking Machines Lab</category><category>Fine-tuning</category></item><item><title>1X unveils new 25-degree-of-freedom tendon-driven hands for its NEO home humanoid</title><link>https://aimpact.prandi.net/en/events/2026-07-09-1x-neo-tendon-hands/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-09-1x-neo-tendon-hands/</guid><description>1X Technologies reveals redesigned hands for NEO: 25 degrees of freedom, tendon actuation from forearm-mounted motors, tactile fingertips with real-time slip detection, 45 N grip force, IP68 sealing and food-safe materials, in production with a 10,000-unit target for 2026.</description><pubDate>Sat, 11 Jul 2026 00:00:00 GMT</pubDate><category>Robotics</category><category>Humanoid Robots</category><category>Dexterous Manipulation</category><category>Tactile Sensing</category><category>Home Robotics</category></item><item><title>OpenAI launches the GPT-5.6 family in three variants: Luna, Terra, and Sol</title><link>https://aimpact.prandi.net/en/events/2026-07-09-openai-gpt-5-6/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-09-openai-gpt-5-6/</guid><description>OpenAI releases GPT-5.6 as three models of increasing capability (Luna, Terra, Sol) alongside the ChatGPT Work tool: Sol is billed as the company&apos;s best coding and cybersecurity model, using 54% fewer tokens on programming tasks.</description><pubDate>Sat, 11 Jul 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>GPT-5.6</category><category>ChatGPT Work</category><category>Cybersecurity</category><category>Coding Models</category></item><item><title>SpaceXAI unveils Grok 4.5, a flagship model at 2/6 dollars per million tokens</title><link>https://aimpact.prandi.net/en/events/2026-07-08-grok-4-5/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-08-grok-4-5/</guid><description>SpaceXAI announces Grok 4.5, which Musk calls an Opus-class model but faster and cheaper: focused on coding, agents, and office work, with a claimed 2x token efficiency and aggressive pricing at 2 dollars input / 6 output per million tokens.</description><pubDate>Fri, 10 Jul 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>Grok</category><category>LLM Pricing</category><category>Agentic Coding</category><category>Token Efficiency</category></item><item><title>Claude Cowork lands on web and mobile: agentic tasks move to the cloud</title><link>https://aimpact.prandi.net/en/events/2026-07-07-claude-cowork-web-mobile/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-07-claude-cowork-web-mobile/</guid><description>Anthropic extends Cowork beyond the desktop app: beta on claude.ai, iPhone, iPad, and Android starting with the Max plan, with server-side task execution that continues while devices are offline and explicit approval before outputs are published.</description><pubDate>Thu, 09 Jul 2026 00:00:00 GMT</pubDate><category>Agents</category><category>Claude Cowork</category><category>Autonomous Agents</category><category>Mobile</category><category>Cloud Agents</category></item><item><title>OpenAI releases gpt-realtime-2.1: p95 voice latency cut by at least 25%</title><link>https://aimpact.prandi.net/en/events/2026-07-06-openai-gpt-realtime-2-1/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-06-openai-gpt-realtime-2-1/</guid><description>OpenAI ships gpt-realtime-2.1 and gpt-realtime-2.1-mini on the Realtime API: at least 25% lower p95 latency via improved caching, more reliable tool use, and better handling of interruptions and noise.</description><pubDate>Wed, 08 Jul 2026 00:00:00 GMT</pubDate><category>Voice &amp; Audio</category><category>Realtime API</category><category>Voice Agents</category><category>Speech-to-Speech</category><category>OpenAI</category></item><item><title>NVIDIA releases Nemotron-Labs-TwoTower, an open-weight diffusion language model on a frozen autoregressive backbone</title><link>https://aimpact.prandi.net/en/events/2026-07-01-nvidia-nemotron-twotower/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-07-01-nvidia-nemotron-twotower/</guid><description>NVIDIA publishes open weights for Nemotron-Labs-TwoTower, a two-tower diffusion LM built on a frozen Nemotron-3-Nano-30B-A3B: it retains 98.7% of autoregressive quality at 2.42x throughput.</description><pubDate>Fri, 03 Jul 2026 00:00:00 GMT</pubDate><category>Open Source Models</category><category>NVIDIA Nemotron</category><category>Diffusion LM</category><category>Open Weights</category><category>Inference Efficiency</category></item><item><title>OpenAI releases o4-mini-high: frontier reasoning at 60% lower cost than o3</title><link>https://aimpact.prandi.net/en/events/2026-06-23-openai-o4-reasoning-efficient/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-23-openai-o4-reasoning-efficient/</guid><description>OpenAI releases o4-mini-high, a reasoning model matching o3 on key benchmarks like SWE-bench and AIME while costing 60% less, with extended thinking up to 32K tokens and tool use during reasoning chains.</description><pubDate>Fri, 26 Jun 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>Reasoning</category><category>OpenAI</category><category>Cost Efficiency</category><category>API</category></item><item><title>Microsoft Build 2026: Copilot becomes the agentic OS layer for Windows</title><link>https://aimpact.prandi.net/en/events/2026-06-12-microsoft-build-2026-copilot/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-12-microsoft-build-2026-copilot/</guid><description>Microsoft announces Copilot++ as a native agentic layer in Windows, GitHub Copilot Workspace reaches GA, Azure AI Foundry 2.0 launches, and Phi-4.5 is released as open source.</description><pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate><category>Enterprise AI</category><category>Microsoft Build</category><category>Copilot</category><category>GitHub Copilot</category><category>Azure AI</category></item><item><title>EU AI Act GPAI Compliance Deadline: OpenAI, Google, Anthropic and Meta File Transparency Reports</title><link>https://aimpact.prandi.net/en/events/2026-06-11-eu-ai-act-gpai-compliance/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-11-eu-ai-act-gpai-compliance/</guid><description>June 11, 2026 marks the first real enforcement milestone of the EU AI Act for GPAI model providers: major AI companies must register on the EU AI database and publish transparency reports, with fines up to 3% of global annual turnover for non-compliance.</description><pubDate>Sun, 14 Jun 2026 00:00:00 GMT</pubDate><category>AI Security</category><category>EU AI Act</category><category>GPAI</category><category>Compliance</category><category>Regulation</category></item><item><title>Meta releases Llama 4.1: Scout, Maverick, and Behemoth MoE models under Apache 2.0</title><link>https://aimpact.prandi.net/en/events/2026-06-10-meta-llama-4-1/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-10-meta-llama-4-1/</guid><description>Meta launches Llama 4.1 in three MoE variants — Scout (edge), Maverick (mid-tier with 10M context), and Behemoth (frontier) — all natively multimodal and freely available under Apache 2.0.</description><pubDate>Sat, 13 Jun 2026 00:00:00 GMT</pubDate><category>Open Source Models</category><category>Llama</category><category>Meta</category><category>MoE</category><category>Multimodal</category><category>Open Source</category></item><item><title>OpenAI Codex 2.0: dedicated autonomous coding agent in ChatGPT and API</title><link>https://aimpact.prandi.net/en/events/2026-06-09-openai-codex-v2/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-09-openai-codex-v2/</guid><description>OpenAI releases Codex 2.0 as a fully autonomous coding agent inside ChatGPT and via API, capable of completing entire repository tasks — reading files, running tests, and opening pull requests — inside isolated cloud sandboxes.</description><pubDate>Fri, 12 Jun 2026 00:00:00 GMT</pubDate><category>AI Coding</category><category>Codex</category><category>Coding Agent</category><category>Autonomous Coding</category><category>Sandboxed VM</category></item><item><title>Apple WWDC 2026: Apple Intelligence 2.0 with 4B parameter on-device models</title><link>https://aimpact.prandi.net/en/events/2026-06-06-apple-wwdc-2026-intelligence/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-06-apple-wwdc-2026-intelligence/</guid><description>Apple unveils Apple Intelligence 2.0 at WWDC 2026: on-device models upgraded to 4B parameters on A18 Pro, an autonomous multi-step Siri, Visual Intelligence 2 with real-time scene understanding, and on-device image generation — all with expanded privacy guarantees.</description><pubDate>Tue, 09 Jun 2026 00:00:00 GMT</pubDate><category>Enterprise AI</category><category>Apple Intelligence</category><category>On-Device AI</category><category>Siri</category><category>Visual Intelligence</category></item><item><title>Google I/O 2026: Gemini Ultra 3, Project Astra goes live on Pixel, 2M context with real-time grounding, Veo 3.2, Imagen 4</title><link>https://aimpact.prandi.net/en/events/2026-06-05-google-io-2026-gemini/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-05-google-io-2026-gemini/</guid><description>At Google I/O 2026, Google DeepMind unveiled Gemini Ultra 3 with a 2M-token context window and real-time web grounding, Project Astra now live on Pixel devices, Veo 3.2 and Imagen 4 for creative generation, and a broader rollout of NotebookLM Plus and Android AI mode.</description><pubDate>Tue, 09 Jun 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>Gemini</category><category>Google</category><category>Multimodal</category><category>Video Generation</category></item><item><title>Anthropic releases Claude Fable 5: a new model family beyond the 4.x line</title><link>https://aimpact.prandi.net/en/events/2026-06-04-claude-fable-5/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-04-claude-fable-5/</guid><description>Anthropic introduces Claude Fable 5 (claude-fable-5), a landmark new model family that breaks from the previous naming convention and signals a significant architectural leap, positioned between Sonnet and Opus in capability and speed.</description><pubDate>Sun, 07 Jun 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>Anthropic</category><category>Claude</category><category>Fable</category><category>Foundation Model</category></item><item><title>Google releases Veo 3.2: 4K video generation at 60fps with native lip-sync and integrated audio</title><link>https://aimpact.prandi.net/en/events/2026-06-03-google-veo-3-2-video/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-03-google-veo-3-2-video/</guid><description>Google DeepMind launches Veo 3.2, advancing AI video generation to 4K 60fps with native lip-sync, multi-character scene consistency, and integrated audio generation, available via Gemini API and Vertex AI.</description><pubDate>Sun, 07 Jun 2026 00:00:00 GMT</pubDate><category>Image &amp; Video Gen</category><category>Video Generation</category><category>Google</category><category>Gemini</category><category>Generative AI</category></item><item><title>Anthropic releases Claude Opus 4.8: the most powerful Claude model to date</title><link>https://aimpact.prandi.net/en/events/2026-06-02-claude-opus-4-8/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-06-02-claude-opus-4-8/</guid><description>Anthropic launches Claude Opus 4.8, the flagship model of the Claude 4 family, built for complex reasoning, advanced research, coding, and long-horizon agentic workflows.</description><pubDate>Fri, 05 Jun 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>Claude</category><category>Anthropic</category><category>Reasoning</category><category>Agentic AI</category></item><item><title>OpenAI releases Sora 2: 1080p 60fps, synchronized audio-video, and API-first for creative professionals</title><link>https://aimpact.prandi.net/en/events/2026-05-21-openai-sora-2-video/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-05-21-openai-sora-2-video/</guid><description>OpenAI launches Sora 2 with 1080p 60fps output, 2-minute clips, native audio-video synchronized generation, and integrated inpainting/outpainting. Initially API-only, the model is repositioned as a creative professional tool following the shutdown of the consumer Sora app in April 2026.</description><pubDate>Sun, 24 May 2026 00:00:00 GMT</pubDate><category>Image &amp; Video Gen</category><category>Sora 2</category><category>Video Generation</category><category>OpenAI</category><category>Text-to-Video</category></item><item><title>Realtime voice AI: sub-second latency and multilingual become the norm</title><link>https://aimpact.prandi.net/en/events/2026-05-18-voice-realtime-multilingual/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-05-18-voice-realtime-multilingual/</guid><description>Realtime voice APIs from OpenAI, Google and ElevenLabs converge on &lt; 500ms latency, fluent multilingual, natural prosody. Phone as an agentic channel becomes practical.</description><pubDate>Mon, 18 May 2026 00:00:00 GMT</pubDate><category>Voice &amp; Audio</category><category>Voice</category><category>Realtime</category><category>Speech</category><category>Multilingual</category><category>Latency</category></item><item><title>Boston Dynamics Atlas Electric: manipulation foundation model trained on 10 million robot hours</title><link>https://aimpact.prandi.net/en/events/2026-05-15-boston-dynamics-atlas-electric-ai/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-05-15-boston-dynamics-atlas-electric-ai/</guid><description>Boston Dynamics releases a new manipulation foundation model for Atlas Electric, trained on 10 million hours of robot experience, enabling zero-shot grasping of novel objects in unstructured environments. Now deployed at Hyundai factories.</description><pubDate>Mon, 18 May 2026 00:00:00 GMT</pubDate><category>Robotics</category><category>Boston Dynamics</category><category>Atlas Electric</category><category>Foundation Model</category><category>Manipulation</category></item><item><title>Mistral releases Devstral Small: 7B coding model for agentic tasks on consumer GPU</title><link>https://aimpact.prandi.net/en/events/2026-05-13-mistral-devstral-small/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-05-13-mistral-devstral-small/</guid><description>Mistral releases Devstral Small, a 7-billion-parameter model fine-tuned for agentic coding that outperforms GPT-4o-mini on SWE-bench and runs on just 8GB of VRAM.</description><pubDate>Sat, 16 May 2026 00:00:00 GMT</pubDate><category>Local AI</category><category>Coding Model</category><category>Local LLM</category><category>Agentic AI</category><category>Open Weights</category></item><item><title>MCP at 18 months: the server ecosystem hits critical mass</title><link>https://aimpact.prandi.net/en/events/2026-05-12-mcp-ecosystem-milestone/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-05-12-mcp-ecosystem-milestone/</guid><description>Eighteen months after launch (November 2024), Model Context Protocol consolidates: thousands of public servers, confirmed cross-vendor adoption, first stable official registry.</description><pubDate>Fri, 15 May 2026 00:00:00 GMT</pubDate><category>Agents</category><category>MCP</category><category>Model Context Protocol</category><category>Anthropic</category><category>Standards</category><category>Tooling</category></item><item><title>Google releases Gemini 3.1 Pro with native video understanding</title><link>https://aimpact.prandi.net/en/events/2026-05-09-gemini-multimodal-video-understanding/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-05-09-gemini-multimodal-video-understanding/</guid><description>Gemini 3.1 Pro analyzes videos up to one hour long frame-by-frame, extracts events, and answers questions about video content. It powers YouTube AI summaries and Google Search video clips, with a 2M token context window that natively includes video frames.</description><pubDate>Wed, 13 May 2026 00:00:00 GMT</pubDate><category>Multimodal AI</category><category>Video Understanding</category><category>Gemini</category><category>Long Context</category><category>YouTube AI</category></item><item><title>ServiceNow Now AI Agents GA: autonomous IT service management at scale</title><link>https://aimpact.prandi.net/en/events/2026-05-03-servicenow-now-ai-agents/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-05-03-servicenow-now-ai-agents/</guid><description>ServiceNow releases Now AI Agents to general availability, enabling autonomous end-to-end handling of L1/L2 tickets without human routing, claiming 65% ticket deflection across enterprise deployments.</description><pubDate>Thu, 07 May 2026 00:00:00 GMT</pubDate><category>Enterprise AI</category><category>ITSM</category><category>Autonomous Agents</category><category>Ticket Deflection</category><category>ServiceNow</category></item><item><title>Usable 2-bit quantization: frontier reasoning models drop below 32GB RAM</title><link>https://aimpact.prandi.net/en/events/2026-04-30-local-ai-quantization-breakthrough/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-04-30-local-ai-quantization-breakthrough/</guid><description>New quantization techniques (high-quality 2-bit / 3-bit extensions) let frontier-sized reasoning models run on workstations with 32-64GB unified RAM.</description><pubDate>Mon, 04 May 2026 00:00:00 GMT</pubDate><category>Local AI</category><category>Local AI</category><category>Quantization</category><category>Ollama</category><category>llama.cpp</category><category>On-device</category></item><item><title>OpenAI shuts down the Sora app: consumer AI video can&apos;t sustain the math</title><link>https://aimpact.prandi.net/en/events/2026-04-26-openai-sora-app-shutdown/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-04-26-openai-sora-app-shutdown/</guid><description>OpenAI shuts down the Sora app on April 26, 2026; the Sora 2 API will be turned off September 24. Operating costs estimated around $1M/day, compute shifting to ChatGPT/GPT-5.5 and core enterprise.</description><pubDate>Sat, 02 May 2026 00:00:00 GMT</pubDate><category>Image &amp; Video Gen</category><category>OpenAI</category><category>Sora</category><category>Video Generation</category><category>Shutdown</category><category>Cost</category></item><item><title>DeepSeek V4 Preview: 1.6T parameters, 1M context, open weight in two sizes</title><link>https://aimpact.prandi.net/en/events/2026-04-24-deepseek-v4-preview/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-04-24-deepseek-v4-preview/</guid><description>DeepSeek releases V4 Preview as open source: V4-Pro (1.6T total, 49B active) and V4-Flash (284B total, 13B active). Native 1M-token context, hybrid CSA+HCA attention cutting KV cache by 90%.</description><pubDate>Thu, 30 Apr 2026 00:00:00 GMT</pubDate><category>Open Source Models</category><category>DeepSeek</category><category>Open Source</category><category>MoE</category><category>China</category><category>Long Context</category></item><item><title>GPT-5.5: OpenAI shifts ChatGPT toward an &quot;agent runtime&quot; paradigm</title><link>https://aimpact.prandi.net/en/events/2026-04-23-openai-gpt-5-5/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-04-23-openai-gpt-5-5/</guid><description>OpenAI releases GPT-5.5, GPT-5.5 Thinking, and GPT-5.5 Pro: designed as an &quot;agent runtime&quot; for persistent multi-step workflows. 23% more factually correct vs GPT-5.4. File Library, side-by-side shopping, improved image gen.</description><pubDate>Thu, 30 Apr 2026 00:00:00 GMT</pubDate><category>Foundation Models</category><category>OpenAI</category><category>GPT-5.5</category><category>ChatGPT</category><category>Agents</category><category>Codex</category></item><item><title>EU AI Act: 100-day countdown to the high-risk system rules</title><link>https://aimpact.prandi.net/en/events/2026-04-22-eu-ai-act-high-risk-prep/</link><guid isPermaLink="true">https://aimpact.prandi.net/en/events/2026-04-22-eu-ai-act-high-risk-prep/</guid><description>Around 100 days before high-risk AI system obligations take effect (August 2026), the European Commission publishes operational guidelines and the AI Office activates.</description><pubDate>Sun, 26 Apr 2026 00:00:00 GMT</pubDate><category>AI Security</category><category>EU AI Act</category><category>Regulation</category><category>Compliance</category><category>Europe</category><category>High-risk</category></item></channel></rss>