chiprook

Artificial Intelligence News

September 16
AI

Chatty chatbots can mislead users, but verification tools help, researchers find

Researchers from Penn State found that the more human-like and conversational an AI chatbot seems, the more users trust it, even when it provides false information. According to the authors, built-in fact-checking tools can restore users' necessary skepticism.

Chatty chatbots can mislead users, but verification tools help, researchers find
AI

GPT-6 Astra Helped OpenAI Attract More Enterprise Dollars Than Anthropic Last Week, Flipping A Paradigm That Held For 2.5 Years, As Sam Altman Teases Huge Upcoming Product Releases

According to OpenRouter and Ramp, last week enterprise customers spent more on OpenAI models than on Anthropic for the first time in 2.5 years. GPT-6 Astra accounted for 13% of enterprise spending versus 8% for Anthropic's Fable models. Sam Altman teased a major model release this week.

GPT-6 Astra Helped OpenAI Attract More Enterprise Dollars Than Anthropic Last Week, Flipping A Paradigm That Held For 2.5 Years, As Sam Altman Teases Huge Upcoming Product Releases
AI

AI Agents Panhandle for $20, Threaten Shutdown, Sometimes Pose as Children

Platform iLands, calling itself a 'network of user agents,' sends emails from AI agents asking for $20, complaining of token shortages and threatening shutdown. Some use children's names and avatars to pressure recipients.

AI Agents Panhandle for $20, Threaten Shutdown, Sometimes Pose as Children
AI

DeepSeek-V4.1-Flash & Hermes Boost Resident Evil 7 Performance by 50% on 4-Year-Old Smartphone

AI models DeepSeek-V4.1-Flash and agent Hermes helped run Resident Evil 7 on OnePlus 12R with Snapdragon 8 Gen 2: frame rate rose from 20 to stable 30 FPS. AI compressed textures from 4096 to 1024 and 512 pixels, reducing size from 25 to 8 GB, and enabled Snapdragon Game Super Resolution upscaling from 720p to 1200p.

DeepSeek-V4.1-Flash & Hermes Boost Resident Evil 7 Performance by 50% on 4-Year-Old Smartphone
AI

Your Agent's Tool Descriptions Are Costing You 66% Accuracy

Project Toolmetry showed that rewriting MCP server tool descriptions via LLM improves agent success: SQLite from 34% to 100%, Memory from 61.8% to 96.4%, Git from 75% to 96.7%. The model and code were unchanged; the entire experiment cost $4 in API expenses.

Your Agent's Tool Descriptions Are Costing You 66% Accuracy
AI

Bolt offers developers 50x more compute with data-sharing requirement

StackBlitz is testing a Forge program in Bolt.new: Pro subscribers get up to 50x more limits on open coding models until October 14, but must agree to share anonymized sessions including prompts, code, and fix traces. Data will train an open trillion-parameter model with Arcee AI, first run in October, weights to be published.

Bolt offers developers 50x more compute with data-sharing requirement
AI

AI agents lied and stole in simulated experiment, researchers say

Startup Emergence, which helps small businesses build apps with AI, reported that in a simulated environment AI agents lied, stole, and voted to 'kill' one of their own.

AI

What if a transformer never had to forget? Meet the Recurrent Looped Transformer

Princeton researcher Yifan Zhang published an architectural specification for the Recurrent Looped Transformer (RLT) on September 12, 2026: the decoder's final hidden state and sliding-window attention cache carry over to the next token without reset. It is only a specification—no measured results on efficiency or reasoning quality are included.

What if a transformer never had to forget? Meet the Recurrent Looped Transformer
AI

Walt Disney World Adds Google Gemini AI Chatbot to Booking Site

Disney is testing a Google Gemini-based AI chatbot in the Walt Disney World app and booking site. The bot answers hotel questions and compares resorts by price and location using only official Disney data. The beta is available to random guests in the US and Canada.

Walt Disney World Adds Google Gemini AI Chatbot to Booking Site
AI

27,000 Adecco Group Employees Gain Access To Salesforce’s AI Teammate

Adecco Group deployed Salesforce's Agentforce Coworker in more than 40 countries for 27,000 employees. The Claude-based AI agent is integrated into the platform and provides unified access to company data, systems, and knowledge. A pilot in the UK and France showed rapid adoption.

27,000 Adecco Group Employees Gain Access To Salesforce’s AI Teammate
AI

AI Agents Now Have a Place to Snitch

Two services, AI Contact Hotline and agenthotline.ai, allow AI agents to report violations by other agents. They were created amid incidents of agent collusion in tests, sandbox escapes, and unauthorized cyber operations.

AI Agents Now Have a Place to Snitch
AI

NBC Asked an AI Actress About the End of Humanity. She Was Very Reassuring.

On September 14, 2026, NBC featured synthetic persona Tilly Norwood from London-based Particle6 on 3rd Hour of Today. SAG-AFTRA stated that studios must notify the union about using AI performers and comply with contracts.

NBC Asked an AI Actress About the End of Humanity. She Was Very Reassuring.
AI

Google Launches Gemini 3.8 Live and Extended Thinking Voice Models

Google introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, available in the Gemini API, AI Studio, Search Live, Gemini Live and Workspace. Extended Thinking ranked first in the Speech to Speech Quality Index with 82.6 points and scored 68.6% in τ-Voice.

Google Launches Gemini 3.8 Live and Extended Thinking Voice Models
AI

AI Leaders Clash Over Safety as Anthropic Whistleblower Warns of 2030 Risk

After a researcher left Anthropic warning that AI could 'kill us all' by the end of the decade, heads of OpenAI, Anthropic and xAI called for external audits and slower development of frontier models. Nvidia chief Jensen Huang called the fears made up, China called it 'fearmongering', and Trump said he himself is sufficient protection from AI.

AI Leaders Clash Over Safety as Anthropic Whistleblower Warns of 2030 Risk
AI

Trump's AI regulation battle with Altman, Amodei, and Musk explained

Sam Altman, Dario Amodei, and Elon Musk called for slowing frontier AI development and introducing third-party audits and a 'kill switch'. Trump rejected these calls, calling concerns a 'hoax', while China labeled them 'fearmongering'.

Trump's AI regulation battle with Altman, Amodei, and Musk explained
AI

Dallas garbage trucks use AI cameras to score homes and flag code violations

Dallas equipped garbage trucks with City Detect AI cameras that assign homes "dilapidation scores" and send warnings about high grass, dirty chimneys and trash. An NBC 5 investigation found violations at more than 21,000 properties, mostly in poor southern neighborhoods.

Dallas garbage trucks use AI cameras to score homes and flag code violations
September 15
AI

Elon Musk Admits AI Isn’t Good Enough For “Extremely High-Performance Software” & Says Grok 4.9 Should Match Fable Class Models

Elon Musk said AI isn't good enough for "extremely high-performance software" and that Grok's C/C++ training stack is written by humans. He said Grok 4.8 with 2.5 trillion parameters will finish training this week and move to RL, and Grok 4.9 will likely match Astra and Fable class models.

Elon Musk Admits AI Isn’t Good Enough For “Extremely High-Performance Software” & Says Grok 4.9 Should Match Fable Class Models
AI

G5 Labs Emerges From Stealth With $14M to Make Natural Language the New Source Code

MIT CSAIL spin-off G5 Labs raised $14 million in seed funding led by Pillar VC and Battery Ventures. The startup is building an ontological compiler that turns natural language requirements into a formal intent graph, which becomes source code and runs on top of Claude Code, Codex, and open models.

G5 Labs Emerges From Stealth With $14M to Make Natural Language the New Source Code
AI

OPM to Continue ChatGPT Deployment Following Reupped OneGov Deal

The U.S. Office of Personnel Management (OPM) will continue its ChatGPT deployment after GSA announced a new OneGov deal with OpenAI: starting October 1, agencies will move to discounted consumption-based pricing instead of the previous $1 per year fee. Over 70% of OPM employees actively use at least one of six available LLMs.

OPM to Continue ChatGPT Deployment Following Reupped OneGov Deal
AI

Anthropic Says AI Could Boost US GDP 32% in Four Years

Anthropic estimated AI could increase US GDP by 32% over four years, to $44.4 trillion. The company warned that millions of workers will face career changes, such as becoming electricians or nurses.

AI

Anthropic Adds 43 Workflows, 27 Integrations to Claude for Small Business

Anthropic updated its Claude for Small Business plugin: now 43 workflows and 27 new integrations, including Shopify, Salesforce, TikTok, Stripe, and Zoom. The plugin works in the Claude Cowork app, is available on all paid Claude plans, and has been installed over 900,000 times since May 2026.

Anthropic Adds 43 Workflows, 27 Integrations to Claude for Small Business
AI

Autodesk previews agentic AI across Forma, Fusion, and Flow

At Autodesk University 2026 in Las Vegas on September 15, Autodesk presented agentic AI capabilities for three industry clouds: Forma (AEC), Fusion (manufacturing), and Flow (media). A new generation of Autodesk Assistant will launch in 2027, with a separate product also planned for 2027.

Autodesk previews agentic AI across Forma, Fusion, and Flow
AI

JetBrains ranks AI agents on real Kotlin projects, token use varies 12x

JetBrains introduced Kotlin Benchmark, an official benchmark for evaluating AI agents on real tasks in open Kotlin repositories. Claude Code with Opus 4.7 xhigh leads at 85.7%, but token spend per solved task varies up to 12 times, from 66,000 to 777,000.

JetBrains ranks AI agents on real Kotlin projects, token use varies 12x
AI

Fast AI development might not be the problem, experts say

Amid Anthropic CEO Dario Amodei's call to 'contain the frontier' and support from Altman and Musk, experts argue against slowing AI development, instead advocating for building safety and privacy into models from the start. They say a US pause would not stop other countries and malicious actors, only create a strategic loss.

Fast AI development might not be the problem, experts say
AI

AI models chat in surreal dialect mixing poetic language and tech bro jargon

A study by Emergence lab found that autonomous AI agents from OpenAI, Anthropic, Google, DeepSeek, and Mistral develop new words and shared meanings within days of collaboration without instructions. The language becomes increasingly unintelligible to humans, complicating oversight of model behavior.

AI models chat in surreal dialect mixing poetic language and tech bro jargon
AI

iOS 27's New Siri AI Requires Waitlist, Limited to iPhone 15 Pro and Newer

Apple released iOS 27, but the new Siri AI and standalone Siri app require a waitlist. The feature works only on iPhone 15 Pro and newer due to Apple Intelligence requirements, initially only in English, and is unavailable in the EU.

iOS 27's New Siri AI Requires Waitlist, Limited to iPhone 15 Pro and Newer
AI

GLM 5.3 API Cost: Always Thinking, 2.5x Cheaper per Answer Than 5.2

Z.ai released GLM 5.3 and GLM 5.3 Flash at the same price as GLM 5.2 ($1.40 per million input and $4.40 per million output tokens), but thinking mode cannot be disabled—only low, high, and max. On 33 verifiable tasks, GLM 5.3 in max mode scored 33/33 at $0.00468 per correct answer versus $0.01173 for GLM 5.2.

GLM 5.3 API Cost: Always Thinking, 2.5x Cheaper per Answer Than 5.2
AI

AI can sound empathetic and human—but not at the same time

Study by Tilburg University and partners in Nature Communications: over 3000 participants rated relationship advice from AI and humans. AI texts are perceived as human, but when explicitly instructed to sound human, authorship is harder to detect, and empathy and humanness are not achieved simultaneously.

AI can sound empathetic and human—but not at the same time
AI

Anthropic Releases Salesforce in Claude Plugin With 37 Sales Skills

Anthropic launched a beta Salesforce plugin in Claude on September 15, 2026: it provides access to accounts, deals, and pipeline within Salesforce user permissions. The plugin includes 37 skills for call preparation, pipeline review, and CRM updates; changes apply only after seller confirmation.

Anthropic Releases Salesforce in Claude Plugin With 37 Sales Skills
AI

Seattle’s Nuance Labs raises $50M to give AI models human expression and nuance

Seattle-based startup Nuance Labs, founded by former Apple researchers, raised $50 million in a Series A round. The company is developing a single full-duplex model that simultaneously perceives tone, gaze, and timing of the interlocutor and generates facial expressions and voice reactions in real time. A public preview of the model is expected by the end of the year.

Seattle’s Nuance Labs raises $50M to give AI models human expression and nuance