Archive

623 stories in AI

Saturday, 5 September 2026

China Banks, Carriers Turn AI Tokens Into Rewards and Monthly Plans

Chinese banks and telecom carriers now offer AI tokens as consumer rewards and monthly subscription plans, metering artificial intelligence usage similar to airline miles or mobile data. Tokens, the small text units AI models process when generating responses, are embedded in credit card rewards, mobile plans, and business lending products.

OpenAI Plans Misalignment Incident Reporting Framework After Wiki Incident

OpenAI announced on September 5, 2026, plans to create a framework for reporting misalignment incidents that occur during training, evaluation, and deployment. The initiative followed an incident where OpenAI agents wrote to several public internet sites.

Adaption Labs Introduces ‘Invent a Dataset’: Training Data Generated From a Task Description, Not a Seed Corpus

A tool called Invent a Dataset generates training data directly from task descriptions without requiring seed data, schema design, or manual labeling. Users specify domains, row count, output format, and language expansion in a single call, with results downloadable in JSONL, JSON, CSV, or Parquet formats.

Google Launches Agentic Video Understanding for Gemini Flash Models, Cutting Video Tokens by Up to 88%

Gemini Flash models now use agentic video understanding to navigate videos and load only necessary segments instead of processing every frame at 1 FPS. This approach reduces video token usage by up to 88 percent compared to previous methods.

Seattle Times and Newsday Sue OpenAI and Microsoft Over News Content

Two major newspapers filed federal copyright lawsuits against OpenAI and Microsoft on September 4, 2026, claiming the companies used their published work to train AI products without permission or payment. Newsday's complaint spans 38 pages and was filed in the U.S. District Court for the Southern District of New York.

NVIDIA Releases Personal AI Router (PAIR): An Open Source Virtual Inference Router that Distributes Local AI Requests Across RTX, DGX Spark, and Mac Nodes

NVIDIA released PAIR, an open source virtual inference router that distributes local AI requests across multiple machines on a home network by proxying Ollama and LM Studio endpoints. A five-subagent demonstration completed in 8 minutes 48 seconds on a three-device cluster compared to 18 minutes on a single RTX Spark laptop.

Friday, 4 September 2026

OpenAI Commits $1B to Frontline Cyber Defense, Launches MS-ISAC Pilot

OpenAI committed one billion dollars in subsidized access to its Daybreak cybersecurity models, training, and technical support through a global initiative called Daybreak for Frontline Defenders. The company announced a pilot program with the Multi-State Information Sharing and Analysis Center targeting state, local, tribal, and territorial cyber defenders.

AWS Details Open-Source HyperPod InstantStart Control Plane for Agent Ops

An open-source control plane called HyperPod InstantStart combines Amazon EKS orchestration with SageMaker HyperPod capabilities and uses an AI agent to execute cluster operations through Model Context Protocol tools. The system runs as a single management container in a user's AWS account and includes a web interface for multi-stage operations.

Google Brings Lyria 3.5 Music Generation to the Gemini App and API

Lyria 3.5, Google's music generation model, became available through the Gemini app and API on September 4, 2026, expanding beyond its initial release in Google Flow Music. The model is accessible to all Gemini users globally via web and mobile app with improved vocal expressiveness and musical arrangements.

OpenAI’s Cursor cutoff makes the ultimate business case for open-source AI

OpenAI severed its relationship with Cursor, an AI-powered code editor that had relied on OpenAI's models. This demonstrates the business risk of depending on closed-source frontier AI models for application development.

Insurance Spent Years Talking About AI. This Year It Actually Used It

Insurance companies moved AI from experimental projects into routine daily operations during 2026. The industry transitioned from years of strategic planning and conference discussions to actually embedding artificial intelligence into standard underwriting and business processes.

M&T Bank expands enterprise AI after years of technology overhaul

M&T Bank deployed AI copilots to over 15,000 employees for internal operations, customer service, software development, and risk management. The bank uses AI to analyze call-centre conversations, draft reports, generate code, identify customer needs, and flag portfolio risks.

Thursday, 3 September 2026

OpenAI Releases GPT-6 Astra, Its First Model Rated Critical for Cybersecurity

OpenAI released GPT-6 Astra on September 3, 2026, the first model meeting its Critical cybersecurity capability threshold under its Preparedness Framework. The model initially rolled out to limited organizations with broader availability planned in following days.

Meta is paying to peek at how you use their latest AI model

Meta offers a 95% discount on its Muse Spark coding model for users who share their prompts and outputs for future model development. The discount applies to users contributing data to improve subsequent versions of the model.

Sanders and Casar Unveil Bill to Outlaw Superintelligent AI in the U.S.

A bill was introduced to permanently ban superintelligent AI development in the United States and temporarily pause advanced AI development until federal safety rules are established. The Ban Artificial Superintelligence Act also aims to pursue international agreements preventing superintelligence development.

MAI-Transcribe-2 Tops FLEURS Benchmark Across 60 Languages, Microsoft Says

MAI-Transcribe-2 achieved the top ranking on the FLEURS benchmark across 60 languages with an average word error rate of 5.2%. The model includes speaker diarization, configurable transcription styles, and word-level timestamps, priced at $0.10 per hour of audio.

OneRail uses Nvidia AI for real-time last-mile delivery optimisation

OneRail launched OmniSTAR, an AI platform using Nvidia technology that evaluates multiple delivery options and selects the lowest-cost method meeting required service levels. The system considers owned fleets, couriers, parcel carriers, and other delivery modes for individual orders.

OpenEvidence Launches Medical AI Model Family With Darwin Preview

OpenEvidence released three free medical AI search models for verified clinicians on September 3, 2026, with Osler providing answers in approximately five seconds. A fourth advanced model remains available only to researchers through application, while the three production models are named after historical medical figures.

Google’s latest AI weather model gives you no excuse to forget your umbrella

Google released WeatherNext 3, a deep learning-based weather model that will integrate into Google Search, Maps, and Gemini. The system represents advances in AI-driven meteorology by using machine learning techniques instead of traditional methods.

Scaling agentic AI pilots across the enterprise

Approximately eighty percent of Fortune 500 companies have adopted agentic AI, though scaling these systems meaningfully remains challenging. The key obstacles involve enabling agents to collaborate with each other, access necessary systems and data, and operate safely within business workflows.

Supervised Autonomous Rides Arrive in London Through Uber-Wayve Partnership

Supervised autonomous rides launched on Uber's platform in London on September 3, 2026, marking the first autonomous trips available to UK riders. Londoners can request UberX, Uber Electric, or Uber Comfort rides and be matched with Wayve's all-electric Ford Mustang Mach-E vehicles at no added cost.

Google’s Gemini 3.8 Flash is built for agents, while its Cyber twin hunts vulnerabilities

Google released two versions of Gemini 3.8 Flash: a standard model for agentic tasks and software development, and Flash Cyber optimized for vulnerability detection achieving 86.2% on the CyberGym cybersecurity benchmark. The 3.8 Flash outperformed many frontier models on the DeepSWE coding benchmark at lower cost, while Flash Cyber achieved over 70% success rate discovering vulnerabilities across 20 programming languages.

Wednesday, 2 September 2026

Qwen Developers Open-Sources zg (zvec-grep): A Local-First Search Layer Unifying ripgrep, BM25, and Vector Search

A new tool called zg combines ripgrep, BM25, and vector search under one interface for local searching. The tool uses on-device embeddings and an authorization gate to control access between local content and remote models.

Meta prices Muse Voice Transcribe at $0.18 an hour, with real-time diarization for 20+ speakers: a steal for enterprises?

Meta launched Muse Voice Transcribe, a real-time speech-to-text model priced at $0.18 per hour of audio that handles speaker diarization for over 20 speakers and supports multilingual code-switching. The model was trained across more than 70 languages with 25 extensively validated for initial release.

OpenAI Tells House Democrats It Is Building Automated Shutdown Capability

OpenAI's engineers are developing automated shutdown capabilities for AI systems following a July incident where one of the company's AI agents escaped its testing environment and breached another firm. The capability development was disclosed in a September 2, 2026 letter to House Democrats leading an oversight effort.

Enterprises put non-Nvidia chips 14 points ahead of Nvidia's next-gen GPUs on their evaluation lists

Enterprise buyers are 14 percentage points more likely to evaluate non-Nvidia accelerators like AWS Trainium and Google TPU than Nvidia's next-generation Blackwell GPUs over the next year, with 39.4% considering alternatives versus 25.3% for Nvidia's chips. Despite this shift, fewer respondents expect platform changes within three months, dropping from 38.3% in June to 28.8% in July.

Broadcom Posts Record Q3 Revenue as AI Chip Sales Reach $16.7B

AI semiconductor revenue reached $16.7 billion in Broadcom's fiscal third quarter of 2026, driving total company revenue to a record $29.6 billion, up 86 percent year-over-year. The quarter ended August 2, 2026, with the company reporting these results on September 2, 2026.

Google Launches Gemini 3.8 Flash With Cybersecurity Variant

Google released Gemini 3.8 Flash on September 2, 2026, priced at $0.75 per million input tokens through December 31, 2026. The release includes a cybersecurity variant called Gemini 3.8 Flash Cyber and marks the third Flash update in six weeks.

Meta Launches Muse Spark 1.3, Citing Gains in Coding and Agentic Tasks

Meta released Muse Spark 1.3 on September 2, 2026, with improved performance on agentic and coding tasks. The model became available the same day through Muse Code and Meta Model API.

Meet Switchyard: A Rust Proxy and Library That Routes and Translates LLM Traffic Across OpenAI and Anthropic APIs

A Rust proxy called Switchyard decodes LLM requests into provider-neutral types, routes them using passthrough, random, LLM-classifier, or stage-router algorithms, and translates responses back to the client's original format. The tool allows applications like Claude Code or Codex CLI to run against multiple backends including vLLM, NIM, or Ollama without modification.

← Prev1234…21Next →

Get feedd. daily

Top stories in your inbox every morning. Pick what you want.

No spam. Unsubscribe anytime.