Open Source AI — August 16, 2026 — Lucien Engelen
Home About Blog Books Booking Lucien
Blog · 16 August 2026

Open Source AI — August 16, 2026

DeepSeek quadruples API prices and OpenAI expands ads in Europe, while Alibaba, Meta, and Nvidia all ship free, downloadable models good enough to run on your own hardware instead of the cloud.

In Plain English: Today's stories cut both ways. DeepSeek — a Chinese AI company known for dirt-cheap prices — just made its API up to 4x more expensive, and OpenAI is adding ads to ChatGPT's free tier in Europe: both signs that running AI in the cloud costs more than the sticker price suggests. At the same time, Alibaba, Meta, and Nvidia all released free, downloadable AI models good enough to run on your own hardware instead of paying for cloud access, and a Nigerian state government proved you can automate real government work with one of these free models instead of signing an expensive vendor contract. Net effect: the case for eventually running AI locally instead of renting it from the cloud got a bit stronger today.


DeepSeek Quadruples API Prices as Peak-Hour Surcharges Begin

DeepSeek is raising V4 Pro and V4 Flash output-token prices roughly fourfold starting today, introducing peak/off-peak dynamic pricing and abandoning its earlier pledge to keep promotional rates permanent. Even one of the cheapest frontier-adjacent labs is now raising per-token prices — direct evidence that inference economics are tightening industry-wide, adding to the case for local inference that sidesteps metered API costs entirely.

Continue Reading →

Engadget · Infrastructure & Economics · Aug 14, 2026


Bond Traders Flag $70 Billion in "Shadow" AI Credit Backstops

Bloomberg reports that Nvidia, Broadcom, and Meta have quietly guaranteed tens of billions in residual value on AI infrastructure financing deals — including Nvidia backstopping up to 25% of a $500B financing vehicle with BlackRock, Goldman Sachs, and others — contingent liabilities that stay off balance sheets until losses look "probable." Analysts warn the structure is pro-cyclical: if AI demand disappoints, chipmakers absorb losses at the worst possible moment, exactly the capex-timing risk underlying skepticism about current infrastructure spending.

Continue Reading →

Bloomberg, via Yahoo Finance · Infrastructure & Economics · Aug 15, 2026


OpenAI Extends ChatGPT Ads to Europe as Free-Tier Monetization Push Continues

OpenAI notified European users this week that ads are coming to ChatGPT's Free and Go tiers later this month, following earlier rollouts in the US, UK, Japan, and South Korea. The expansion shows even the best-funded AI lab reaching for new revenue levers beyond subscriptions and API fees to offset its infrastructure commitments — a sign that current paid usage alone isn't covering the buildout.

Continue Reading →

PPC Land · Infrastructure & Economics · Aug 15, 2026


Alibaba Releases Qwen3.8-27B Under Apache 2.0, Built for Local and Agentic Use

Alibaba's Qwen team open-sourced the Qwen3.8 family, led by a 27B-parameter multimodal model with native 262K context (extensible to 1M) and a toggleable "thinking mode," claiming it beats its own larger Qwen3.7-Plus on coding and office tasks. The release is explicitly aimed at developers building local and agent-based applications — another concrete example of open-weight models closing the capability gap while optimizing for self-hosted use over API dependency.

Continue Reading →

The Decoder · Open-Weight Models · Aug 14, 2026


Meta and Nvidia "Plant a Firm Flag" in Open-Weight Race as Chinese Labs Set the Pace

Within days of each other, Meta shipped Muse Glimmer (30B parameters, Apache 2.0, compressed under 20GB for consumer-hardware and on-device use) and Nvidia released Nemotron 3.5 Lightning plus NeMo Switchyard, an open-source router for directing tasks across models. Coverage frames both moves as a direct response to DeepSeek, Moonshot, and Alibaba's open-weight momentum — with competition increasingly framed around control, hosting location, and cost rather than raw benchmarks, exactly the dynamic driving the shift toward local deployment.

Continue Reading →

CNBC · Open-Weight Models · Aug 15, 2026


Nigerian State Government Automates Operations With Adapted Open-Source AI, Skipping Vendor Contracts

Anambra State's ICT Agency built an internal automation system by customizing an existing open-source model rather than contracting a foundation-model vendor or building from scratch, calling the approach "meaningfully cheaper and more replicable" for resource-constrained governments. It's a small but concrete example of the on-prem/open-weight substitution thesis playing out outside Silicon Valley — an organization with a tight budget bypassing commercial API relationships entirely.

Continue Reading →

TechBuild Africa · On-Prem/On-Device Shift · Aug 13, 2026


The Apple Angle (Standing Perspective)

A standing perspective worth tracking alongside the daily news: some commentators argue Apple — not Nvidia — is quietly best positioned for the on-prem/on-device shift. The case: Apple's unified memory design lets a single ~$9,500 Mac Studio hold up to 512GB of memory, enough to run trillion-parameter open-weight models that would otherwise require five to six Nvidia RTX Pro 6000 cards ($60,000–$75,000, roughly 10x the power draw) to match. As frontier-capable open-weight models keep shipping every few weeks, the argument goes, the bottleneck shifts from "who has the best model" to "who has affordable hardware to run it at home" — a race Apple's Mac Studio line may currently be winning almost by default, largely unnoticed.

Continue Reading →

Limited Edition Jonathan (Substack) · Perspective


Why This Matters

Today's items line up on both sides of the thesis at once: DeepSeek's 4x price hike and OpenAI's expanding ad business are early signs that the commercial API layer is under real margin pressure — and that the $70B in off-balance-sheet financing guarantees now propping up the infrastructure build-out could turn pro-cyclical if that pressure doesn't ease. Meanwhile, Qwen3.8, Meta's Muse Glimmer, and Nvidia's Nemotron 3.5 Lightning are all explicitly positioned for local and on-device deployment rather than API access, and a Nigerian state government just demonstrated, in miniature, what happens when an organization chooses an adapted open-weight model over a vendor contract. None of this settles the debate, but it's a coherent day for the "rising token prices + closing capability gap = shift to on-prem" argument.

Book Lucien for your next event

Keynotes, masterclasses, panels and board-room sessions.

Get in touch