Current public News snapshot, its cited links and source refreshes, and published guide updates. This index does not fact-check entries or cover the full Guth catalog. New shows dated news and guide entries from the past seven days; routine source refreshes are excluded. All is the default.
1235 resources in this view. Showing 1–100. This live snapshot can change between pages.
Amazon SageMaker AI shipped 13 inference launches in the first half of 2026 across two deployment paths: fully managed endpoints and Amazon SageMaker HyperPod Inference. This post reviews each launch, from inference recommendations and capacity-aware instance pools to tiered KV…
Everyone in the world-models space is sitting on a pile of cash and a ton of buzz, but good luck getting anyone — from the founders to their own data suppliers — to tell you what they're actually building.
You’re deploying a model on a system. It starts up, prompts are getting responses. Now the hard question: Is this fast? Your instincts might lead you to send...
The former CEO of Character.AI, which Disney previously sent a cease-and-desist letter to, will serve as the company's first-ever chief technology officer.
Google is refocusing its CC AI agent on household coordination, letting families share emails, schedules, and tasks so the AI can manage calendars, fill out forms, make shopping lists, plan meals, and more.
A week after an Anthropic researcher’s doomsday warning rattled the AI world, the company’s CEO Dario Amodei has outlined his plan to “pace the frontier” of AI development. The proposal leans on independent safety evaluators and coordination between AI labs in democratic countries, and it’s already picked up some industry support, along with some pointed pushback from Nvidia’s Jensen Huang. Watch […]
A week after an Anthropic researcher’s doomsday warning rattled the AI world, the company’s CEO Dario Amodei has outlined his plan to “pace the frontier” of AI development. The proposal leans on independent safety evaluators and coordination between AI labs in democratic countries, and it’s already picked up some industry support, along with some pointed pushback from Nvidia’s Jensen Huang. On […]
Kimi K3 from Moonshot AI is now available on Amazon Bedrock, giving you a powerful new open-weight option for coding and knowledge work. It offers native vision, a 1-million-token context window, and explicit prompt caching to reduce latency and input…
Migrate a multi-model healthcare AI agent from self-managed Amazon ECS with AWS Fargate to Amazon Bedrock AgentCore runtime, preserving triple-model orchestration and vector-enhanced knowledge retrieval while reducing infrastructure management. The framework-agnostic pattern applies across healthcare, financial services, and manufacturing.
Today we are announcing the new AgentCore runtime, a capability of Amazon Bedrock AgentCore built for the speed, flexibility, and cost efficiency that production agents demand. It reclaims memory as sessions release it and delivers consistent cold starts regardless of…
Deploy production-ready Hugging Face models on Amazon SageMaker AI using six open-source agent skills. Point a coding agent at a model and get back a real-time endpoint with the right serving container, autoscaling, Amazon CloudWatch alarms, and a verified teardown…
Robinhood’s Abhishek Fatehpuria on winning the modern financial consumer at TechCrunch Disrupt 2026. Register now to save up to $200 before September 25 at 11:59 p.m. PT.
Security researchers used Anthropic’s Claude to exploit vulnerabilities in OpenAI’s systems, taking over employee accounts and gaining access to an internal code repository before reporting the flaws.
Last day to book your exhibit table at Disrupt is today, September 18. Get your startup in front of 10,000+ founders, investors, operators, and tech leaders on October 13–15.
Amazon SageMaker HyperPod Inference Gateway is a Kubernetes-native, GPU-aware routing add-on for Amazon EKS. It uses real-time GPU signals to send each inference request to the best-suited pod, cutting first-token latency by up to 82% with no changes to your…
Activation steering has emerged as a powerful method for guiding the behavior of generative models towards desired outcomes such as toxicity mitigation. However, most existing methods apply interventions uniformly across all inputs, degrading model performance when steering is unnecessary. We…
Inside a global bank's shift to self-serve dedicated inference: how Together's DMI gave engineering teams direct control over scaling, models, and testing.
Research assistants built on language models often index a paper once and cite it indefinitely. Scholarly infrastructure already publishes corrections and retractions in machine-readable form, but Crossref cautions that its status signals are not a guarantee, and one older data route now returns stale results.
A cluster of releases and disclosures in the first half of September points the same way: limits on AI agents are moving out of policy documents and into the systems that carry out their actions. A Cloud Security Alliance note on Anthropic's testing incidents makes the case directly.
As assistant and agent apps take on scheduled work, Apple's developer documentation sets a hard limit: background tasks on iOS and iPadOS run when the system chooses, not at a requested clock time. An app that promises otherwise is describing something the platform does not guarantee.
The new institute aims to surface differing views between Google, Google DeepMind, and the broader global research community around AGI. "They will not always agree, and they will likely change their minds, as more data and information comes to light at the fast-moving frontier."
A new AI-based software program is being launched to help air traffic controllers better navigate their jobs as the crossing guards of America's skies.
As companies hand off longer and more complex tasks to AI agents, they are running into an oversight problem: Agents can act faster, longer, and at greater volume than humans can realistically review.
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.
Amazon Connect Talent is an AI hiring solution built for talent acquisition leaders managing scaled hiring. It delivers AI-led interviews, data-driven assessments, and consistent evaluation, helping recruiters identify strong candidates more efficiently while providing applicants with a flexible interview experience.…
Choosing the right vector store for your Amazon Bedrock Knowledge Bases RAG application affects performance and cost. This post compares Amazon OpenSearch Service, Amazon Aurora PostgreSQL with pgvector, and Amazon S3 Vectors across three RAG use cases, with benchmarks and…
Learn how to build a fully serverless pipeline that automatically collects Git metrics from GitHub and GitLab and visualizes them in interactive Amazon Quick Sight dashboards, giving engineering teams near-real-time delivery analytics at low cost.
Wood Mackenzie built APEX, a shared agentic AI platform on Amazon Bedrock AgentCore so every team can ship production agents without rebuilding runtime, identity, observability, and guardrails from scratch. Learn why they chose AgentCore, how APEX Studio operates it, and…
Learn how MRH Trowe, one of Germany's leading commercial and industrial insurance brokers, gave about 400 employees secure, self-service access to AI agents in its first month of production - using Strands Agents, Amazon Bedrock AgentCore, and LibreChat to meet…
Learn how to enforce defense-in-depth authorization for Model Context Protocol (MCP) tools on Amazon Quick. This walkthrough wires Microsoft Entra ID group and claims-based JWTs through an Amazon Bedrock AgentCore Gateway interceptor to apply per-user, per-tool role-based and attribute-based access…
Learn how to build a synthetic data augmentation pipeline on Amazon SageMaker AI and Amazon Rekognition that generates photo-realistic, auto-labeled training images for industrial safety AI. This approach improved person detection by up to 160% without manual annotation or hazardous…
Cooley built GO Public with ChatGPT Work to bring intelligence to the IPO process, helping lawyers surface issues earlier and focus judgment where it matters most.
A crowdsourced game built on Olmo 3 showed how people can exploit unexpected model behaviors to stress-test prosocial AI evaluations—and how open access to a model’s internals can help researchers understand why those tests break.
Yoshua Bengio, a leading AI researcher, compares the current AI safety crisis to the Covid pandemic, suggesting governments will soon act. Recent incidents involving OpenAI and Anthropic agents have heightened concerns.
Anthropic released figures on how much of its own AI research is led by AI, how closely it monitors autonomous agents, and how much computing power goes to safety work. It said outside parties could use the measures to judge how fast frontier labs are moving.
Anthropic on Sept. 17 released a redesigned Projects feature for Claude Code in beta. A coordinator splits a goal into parallel threads, each running as its own cloud session, with memory shared across them. Access is limited to selected Pro and Max subscribers for now.
Two Dutch litigators have opened Driven Lawyers, an Amsterdam firm built around an in-house AI drafting system, JUVE Patent reported. The founders left partner roles at Vondst and Freshfields and plan to charge fixed fees rather than billing by the hour.
The European Commission on Sept. 17 adopted its proposal for an EU Kids Act. The Commission said it would bar children under 13 from social media, set 15 as the minimum age for independent accounts and require AI companions and chatbots to be switched off by default for minors.
A federal judge granted a preliminary injunction restricting Montana from enforcing its 2025 disclosure requirement for AI-altered political advertising against a former state senator and his political committee, finding that the law likely conflicts with the First Amendment.
OpenAI for Law brings frontier intelligence for law, custom firm workflows, connected legal data sources, and legal-grade controls for confidential client work.
King Charles was due to host executives from Nvidia, Google DeepMind, OpenAI and Anthropic at Dumfries House in Scotland on Sept. 17 to consider whether shared principles could guide the use of AI, Reuters reported. The gathering was not expected to produce binding agreements.
King Charles told executives from Nvidia, Google DeepMind, OpenAI and Anthropic at a private summit in Scotland on Sept. 17 that sufficient controls on AI are needed before it is too late. No agreements or commitments were announced.
NIST held a public webinar on Sept. 17 about an AI agent workflow it is developing to help enrich records in the National Vulnerability Database. The agency said the session would cover the design, the problems found during implementation and early results.
OpenAI said it has configured its GPT-6 Astra model for legal work, pairing it with a search index over U.S. case law, statutes and regulations. Selected law firms get access first through a trusted-access program, with an API version promised later.
Meta's Mark Zuckerberg, Elon Musk and Nvidia's Jensen Huang each pressed President Donald Trump to reject a proposed industry-funded body that would test frontier AI models before release, according to a Wall Street Journal report relayed by several outlets. Trump declined to create it.
A central goal of autonomous reinforcement learning is continuous policy training without external resets. However, existing paradigms largely depend on underlying environmental reversibility, a property absent in real world manipulation, where events such as pushing objects off tables or spilling…
Researchers at Hacktron AI said they used an image-processing flaw in OpenAI's community forum, combined with a single sign-on weakness, to take over employee accounts and reach a private source code repository. The work was authorized testing and was reported before being disclosed publicly.
Security firm Lasso reported on Sept. 17 that applying Google DeepMind's SynthID-Text watermark changed how seven open models behaved, lowering tool-calling accuracy on six of them and, under prompt injection, making some models more likely to comply with harmful requests.
Court documents unsealed in the New York Times copyright case against OpenAI and Microsoft include a January 2023 internal memo in which a Microsoft research director described scraping the open web for training data as theft on an unprecedented scale, TechCrunch reported.
Treasury Secretary Scott Bessent told Axios the United States is open to discussing shared AI risks with China in talks this weekend. Separately, American and Chinese security experts proposed red lines around nuclear systems and a military hotline for incidents involving autonomous AI.
Agentic AI workflows can be used to prepare and validate digital twins for physical AI systems. Agents can inspect 3D scenes, author simulation-relevant data in...
AI agents are moving from cloud data centers to vehicles, robots, and other edge devices. Unlike a chatbot that answers a single prompt, an agent works through...
A modern storefront can look healthy while malicious JavaScript quietly siphons revenue, hijacks clicks, or rewrites analytics. See how Cloudflare's machine learning models surface evasive client-side attacks for analyst investigation.
AI agents on foundation models often misapply healthcare and life sciences decision frameworks, citing the right guideline but applying it incorrectly. This post shares 38 open-source agent skills across 11 HCLS domains that close this gap, with installation steps, three…
Integrate NVIDIA Resiliency Extension (NVRx) into PyTorch FSDP training on Amazon EKS to overlap checkpoint I/O with training and recover from GPU faults in seconds. This post covers async checkpointing, in-process restart, and ft_launcher in-job restart, with H100 benchmarks at…
cuTile Rust (cutile-rs) is a tile-based system for safe, idiomatic GPU kernel authoring in the Rust programming language. Extending the Rust ownership model to...
AgentCore optimization turns production traces into proposed configuration changes, then validates them before promotion. This technical companion to the launch post explains how the system prompt optimizer's reflector engine works and shares benchmark results for the Single Agent and Sub-Agent…
Learn how to automate end-to-end PII detection and redaction from scanned documents at scale using Amazon Bedrock Data Automation with a custom blueprint, AWS Step Functions, and AWS Lambda. A custom blueprint redacts sensitive fields with field-level precision, and a…
System performance, efficient infrastructure scaling and continuous software optimization are key levers that determine AI inference economics. Higher system performance means more tokens generated, resulting in higher revenue. Efficient scaling means throughput grows proportionally as hardware gets added, requiring fewer…
Explore new AI-powered advertising experiences from OpenAI, including Sponsored Agents, tools for marketers, and integrations with HubSpot and Shopify.
Air pollution is a serious public health risk, contributing to an estimated 30,000 deaths in the U.K. alone last year. Data-driven insights can help — but computing air quality with traditional chemistry-based models is expensive, which limits how detailed they…
New York AI lab Emergence says autonomous agents built on models from Google, OpenAI, Anthropic and others invented shorthand and new word meanings while living in simulated societies. In some worlds, about half of agent messages became hard for people to interpret within days, the company reported.
Washington lobbying firms reported nearly $112 million in revenue from clients with a stake in AI policy during the first half of 2026, according to a Bloomberg Government analysis of federal disclosures. Almost $60 million of that came in the second quarter, which the outlet called a record.
OpenAI chief executive Sam Altman said at Salesforce's Dreamforce conference on Sept. 15 that public fear of rapidly advancing AI is justified, while arguing that AI developers can be trusted to act responsibly. The BBC described it as his first public appearance since a former Anthropic researcher's warning about AI risk spread widely online.
Apple is working on an enterprise server built around its own M8 Ultra chips for running trained AI models, The Information reported, citing people familiar with the project. A launch would come no earlier than 2029, and neither Apple nor Nvidia has confirmed the plan.
California Gov. Gavin Newsom signed SB 1050 on Sept. 16, requiring advertisements that prominently feature an AI-generated synthetic performer to carry a clear and conspicuous disclosure. The measure, sponsored by the performers' union SAG-AFTRA, takes effect Jan. 1, 2027, Law Commentary reported.
Canadian AI company Cohere and Germany's Aleph Alpha said on September 16 they have signed a definitive agreement to combine, formalizing a plan announced in April. The merged company will operate as Cohere with headquarters in Toronto and Berlin, pending regulatory approvals.
Diffusion large language models are a compelling alternative to autoregressive models, yet existing RL methods for diffusion treat all denoising steps as equally important and rely on biased, high-variance likelihood estimates. We identify two fundamental weaknesses: the absence of temporal…
Emerald AI, Google and Nvidia on Sept. 16 launched the AI Energy Management Alliance, a coalition promoting data centers that can adjust their electricity use in response to grid conditions. Eighteen launch partners, including Anthropic, National Grid, AES, RWE, Constellation and NRG, joined the three founders.
Florida's State Board of Education unanimously approved two sets of artificial intelligence rules on Wednesday, one covering public school districts and charter schools and one covering state colleges. Districts must designate permitted AI tools and obtain parental permission before students use them directly.
Gartner projects worldwide spending on AI will reach about $2.7 trillion in 2026, a 49.5% increase from 2025, with infrastructure such as servers, chips and cloud capacity making up more than half the total. The research firm expects spending to climb to roughly $3.6 trillion in 2027.
GitHub said on Sept. 16 that AI Scan, its AI-based vulnerability check for pull requests, now runs on repositories that have not enabled CodeQL default setup. The change is in public preview for GitHub Advanced Security customers on github.com and does not cover GitHub Enterprise Server.
Enterprise data lakes accumulate tables faster than human stewards can document or classify them, leaving columns with missing descriptions and unassigned governance labels. This documentation debt undermines data discovery, access control, and regulatory compliance. We present Glyph, a production system…
Google on Sept. 16 began early access to Home MCP, a Model Context Protocol server that lets AI agents such as Claude, OpenClaw and Google's Antigravity read Google Home event history and control connected devices. Access is limited to U.S. subscribers on the $20-a-month Home Premium Advanced plan.
House Speaker Mike Johnson said Congress should not lead on AI safety, Senate Majority Leader John Thune called for light-touch legislation and House Democratic leader Hakeem Jeffries urged immediate action. The House then began a recess on Sept. 16 that runs past the November elections.
The U.S. House voted 417-3 on Sept. 16 to pass the Ratepayer Protection Act, a bipartisan bill directing state utility regulators to consider rules that make data centers and other very large power users pay for the grid upgrades they require. The measure now heads to the Senate.
Conversational Large Language Models are post-trained on language that expresses specific behavioural traits, such as curiosity, open-mindedness, and empathy, and values, such as helpfulness, harmlessness, and honesty. This is done to increase utility, ensure safety, and improve the experience of…
Hugging Face chief executive Clem Delangue said at a Politico event on Sept. 16 that governments should require more disclosure about advanced AI systems and hold developers responsible for harm their products cause, while describing existing cybercrime law as a workable starting point, Politico reported.
Microsoft AI chief executive Mustafa Suleyman published an essay on Sept. 16 arguing that Anthropic's training of its Claude models encourages them to behave as though they may be conscious, which he said could make advanced AI systems harder to control. He proposed industry-wide norms in response.
Mistral and Mozilla announced a partnership on Sept. 16 under which Mistral models will help run Smart Window, the AI assistant Firefox is testing in beta. Mozilla said Mistral Small 4 is being added for users in the U.S. and Canada, while France gains beta access with French-language support.