Über die Zeit
Vertrieb und Kontrolle werden zu den eigentlichen Produktebenen rund um generative KI
Diese Woche hat zwei bestehende Linien weiter zusammengeführt: Lokale Bereitstellung rückt näher an die gewöhnliche Nutzung, während gehostete Frontier-Systeme in engere Kanäle und Kontrollen eingebettet werden. Meta veröffentlichte Muse Glimmer als Open-Weight-Modell, das für die lokale Ausführung auf Laptops konzipiert ist, und sagte, es werde die Gewichte für Muse Spark 1.2 offenlegen. NVIDIA veröffentlichte Nemotron 3.5 Lightning mit Gewichten, Trainingsdaten und Rezepten und kombinierte es mit der Routing-Bibliothek NeMo Switchyard, um Anfragen über offene, proprietäre und NVIDIA-Modelle auf PCs, Edge-Geräten, Workstations, Rechenzentren und in der Cloud zu steuern. OpenAI machte seine Cyber-Modelle Daybreak Red und Daybreak Blue über Amazon Bedrock verfügbar, weitete sein Daybreak-Partnerprogramm auf Sicherheits- und Dienstleistungsfirmen aus und veröffentlichte eine Linux-Desktop-App für ChatGPT. Diese Kombination dürfte wichtig sein, weil sich das Feld weiter danach aufspaltet, wo Modelle laufen und wer sie nutzen darf, und nicht nur danach, wie leistungsfähig ein einzelnes Modell ist.
Eine ähnliche Aufspaltung zeigte sich bei Sicherheit und Governance. OpenAI sagte, interne Bewertungen seines kommenden Modells Astra deuteten auf potenzielle Cyber-Fähigkeiten auf kritischem Niveau hin, pausierte Astra-Aktivitäten, die strengere Kontrollen noch nicht erfüllen, und fügte eine universelle Überwachung für riskante Handlungen über agentische Nutzungen hinweg hinzu. Anthropic sagte, dass der Auto-Modus von Claude Code ab dem 14. August für bezahlte Konten standardmässig aktiviert sein wird, ausser bei irreversiblen, destruktiven oder externen Handlungen, und sagte, dass es Inhalte, die von seinen Modellen erzeugt oder bearbeitet werden, bei neuen Modellen, die nach dem 2. August veröffentlicht werden, mit dem C2PA-Standard mit Wasserzeichen versehen wird. NIST eröffnete eine öffentliche Kommentierungsfrist zu seinem Entwurf des TEVV-Athlon-Frameworks, während die FTC sagte, sie werde Section-5-Durchsetzungsansprüche auf Basis von Theorien zu disparate impact oder unfair discrimination nicht länger verfolgen. Das dürfte bedeuten, dass der Markt nicht mehr auf eine einzige saubere regulatorische Klärung wartet, bevor er handelt: Anbieter bauen jetzt ihre eigenen Kontrollebenen auf, während öffentliche Regeln sich je nach Funktion weiter auseinanderentwickeln.
Das bringt Microsoft in eine engere Position, als die frühe Cloud-Ära zunächst nahelegte. Wenn weiterhin Open-Weight-Modelle mit praxistauglichen lokalen Laufzeiten und Routing-Werkzeugen erscheinen, dann dürfte der Modellzugang weniger an ein einzelnes Cloud-Gate gebunden bleiben. Wenn gehostete Frontier-Modelle zunehmend über Partner verkauft, in bestehende Software eingebettet oder auf ausgewählte Unternehmenskänale beschränkt werden, dann dürften Vertrieb und Workflow-Kontrolle wichtiger werden, als einfach das Assistentenfenster zu besitzen. Die Folge dürfte ein Markt sein, in dem der dauerhafte Vorteil weniger darin liegt, einen Ort zu haben, an dem Nutzer nach KI fragen, und stärker darin, den Weg zwischen Modell, Gerät, Cloud, Compliance-Grenze und Arbeitsaufgabe zu kontrollieren.
Unsere Prognose
Die nächste sichtbare Aufspaltung dürfte darin bestehen, dass mehr Frontier- oder Near-Frontier-Modellzugang über Cloud-Plattformen und Kanäle von Drittanbietern verkauft wird, während sich die lokale Open-Weight-Verbreitung ausserhalb der eigenen Oberfläche einer einzelnen Plattform weiter ausweitet.
woran wir uns messen lassenFällig 2026-11-30
Bis zum 2026-11-30 muss mindestens eines von OpenAI, Anthropic oder Google ankündigen, dass eines seiner Frontier- oder namentlich bezeichneten spezialisierten generativen KI-Modelle neu über eine Cloud-Plattform eines Drittanbieters verfügbar ist, die nicht diesem Modellanbieter gehört, oder mindestens eines von Meta, NVIDIA oder Mistral muss eine neue herunterladbare offene oder Open-Weight-Modellveröffentlichung mit ausdrücklich genannter Unterstützung für die lokale Ausführung auf PCs, Laptops oder Edge-Geräten ankündigen.
Belege
- 2026-08-10 Meta launched Muse Glimmer, a 30‑billion‑parameter multimodal agentic model released as an open‑weight model under an Apache 2.0 license and designed to run locally on laptops, and announced it will open the weights for its Muse Spark 1.2 model, enabling direct download and use by developers and enterprises and integration into local runtimes such as llama.cpp and Ollama. cnbc.com
- 2026-08-11 NVIDIA released Nemotron 3.5 Lightning, a 30‑billion‑parameter mixture‑of‑experts model as a customizable open model with permissive licensing, weights, training data and recipes, along with the NeMo Switchyard routing library for directing requests across open, proprietary and NVIDIA models on PCs, edge devices, workstations, data centers and the cloud. blogs.nvidia.com
- 2026-08-11 OpenAI announced that its Daybreak Red and Daybreak Blue cybersecurity models are now available to eligible customers through Amazon Bedrock, allowing security teams enrolled in Daybreak Access to run frontier cyber models within existing AWS environments. openai.com
- 2026-08-10 OpenAI expanded its Daybreak Cyber Partner Program, granting security and services firms including Accenture, IBM, Capgemini, Palo Alto Networks, CrowdStrike and Cisco access to its frontier cyber models to embed in their products for enterprise vulnerability discovery, validation and remediation. openai.com
- 2026-08-07 OpenAI reported that internal evaluations of its upcoming Astra model indicate potential critical-level cyber capabilities under its Preparedness Framework and announced scaled-up robustness testing, stricter security controls for high-capability models, a pause on Astra activities that do not yet meet those controls, and universal monitoring for risky actions across agentic uses of Astra. openai.com
- 2026-08-11 OpenAI released a dedicated ChatGPT desktop app for Linux, providing access to ChatGPT, ChatGPT Work and Codex on Ubuntu, Debian and Fedora distributions. techcrunch.com
- 2026-08-09 Anthropic announced it will turn Claude Code’s auto mode on by default for Pro, Max, and Team accounts starting August 14, allowing the coding agent to execute actions automatically except when they are irreversible, destructive, or target external systems, and highlighted additional safeguards such as prompt‑injection screening and hard‑deny rules. techcrunch.com
- 2026-08-11 Anthropic said it will begin watermarking content generated or edited by its models, including Claude, embedding machine-readable watermarks in text and files using the C2PA standard for new models released after August 2 and planning to extend the system to older models to comply with the EU AI Act’s Transparency Code. techcrunch.com
- 2026-08-07 NIST released the initial public draft of NIST AI 200-2, the TEVV-Athlon Framework for Evaluating AI Systems, and opened a 60-day public comment period from August 7 to October 6, 2026 for feedback on its proposed approach to testing, evaluation, validation and verification of AI systems. nist.gov
- 2026-08-07 The Federal Trade Commission issued a policy statement clarifying that it will no longer pursue enforcement claims based on disparate‑impact or “unfair discrimination” theories under Section 5 of the FTC Act and entered agreements to modify certain compliance obligations in prior cases. ftc.gov
- 2026-08-11 OpenAI expanded its ChatGPT Ads program to the United Kingdom, Mexico, Brazil, Japan and South Korea as part of a pilot intended to support free access to ChatGPT. openai.com
- 2026-08-09 OpenAI announced it is deprecating its Atlas browser and will shut it down on August 9, 2026, shifting its browser-based agentic capabilities into ChatGPT and Codex and directing users to export their data and move to the ChatGPT desktop app or Chrome integrations. help.openai.com
- 2026-08-06 Google expanded the Ask Maps feature in Google Maps with agentic capabilities to order food, book hotels and surface ticketed local events, and introduced a Personal Intelligence option that tailors recommendations using data from Gmail and Google Calendar for U.S. users. techcrunch.com
- 2026-08-11 Mistral AI announced general availability of Mistral Regional Endpoints that let customers choose Europe or the U.S. for inference, introduced a Priority Tier with custom rate limits and an uptime SLA in public preview, and outlined plans for long-term compute commitments in Europe and expanded access to third-party open models within its infrastructure. mistral.ai