Publisher update
Benchmarking LLM Inference at Scale with AIPerf
You’re deploying a model on a system. It starts up, prompts are getting responses. Now the hard question: Is this fast? Your instincts might lead you to send...
Guth Labs News · Discover
Current public News snapshot, its cited links and source refreshes, and published guide updates. This index does not fact-check entries or cover the full Guth catalog. New shows dated news and guide entries from the past seven days; routine source refreshes are excluded. All is the default.
185 resources in this view. Showing 1–100. This live snapshot can change between pages.
Publisher update
You’re deploying a model on a system. It starts up, prompts are getting responses. Now the hard question: Is this fast? Your instincts might lead you to send...
Publisher update
Jev, a new kind of AI model, is showing developers a cheaper and faster path to software intelligence.
Publisher update
The former CEO of Character.AI, which Disney previously sent a cease-and-desist letter to, will serve as the company's first-ever chief technology officer.
Publisher update
Google is refocusing its CC AI agent on household coordination, letting families share emails, schedules, and tasks so the AI can manage calendars, fill out forms, make shopping lists, plan meals, and more.
Publisher update
A week after an Anthropic researcher’s doomsday warning rattled the AI world, the company’s CEO Dario Amodei has outlined his plan to “pace the frontier” of AI development. The proposal leans on independent safety evaluators and coordination between AI labs in democratic countries, and it’s already picked up some industry support, along with some pointed pushback from Nvidia’s Jensen Huang. Watch […]
Publisher update
A week after an Anthropic researcher’s doomsday warning rattled the AI world, the company’s CEO Dario Amodei has outlined his plan to “pace the frontier” of AI development. The proposal leans on independent safety evaluators and coordination between AI labs in democratic countries, and it’s already picked up some industry support, along with some pointed pushback from Nvidia’s Jensen Huang. On […]
Publisher update
Kimi K3 from Moonshot AI is now available on Amazon Bedrock, giving you a powerful new open-weight option for coding and knowledge work. It offers native vision, a 1-million-token context window, and explicit prompt caching to reduce latency and input…
Publisher update
Manus, which earlier this year had to break off a merger with Meta, is in discussions to raise $500M at a $4B valuation.
Publisher update
A new focus on fiction and memoir aims to help the MIT community celebrate the power of storytelling and strengthen social connection.
Publisher update
Migrate a multi-model healthcare AI agent from self-managed Amazon ECS with AWS Fargate to Amazon Bedrock AgentCore runtime, preserving triple-model orchestration and vector-enhanced knowledge retrieval while reducing infrastructure management. The framework-agnostic pattern applies across healthcare, financial services, and manufacturing.
Publisher update
Today we are announcing the new AgentCore runtime, a capability of Amazon Bedrock AgentCore built for the speed, flexibility, and cost efficiency that production agents demand. It reclaims memory as sessions release it and delivers consistent cold starts regardless of…
Publisher update
Nvidia's Nader Khalil and Sydney Sykes discuss one of the decisions shaping next-gen startups on the Builders Stage at TechCrunch Disrupt 2026.
Publisher update
Deploy production-ready Hugging Face models on Amazon SageMaker AI using six open-source agent skills. Point a coding agent at a model and get back a real-time endpoint with the right serving container, autoscaling, Amazon CloudWatch alarms, and a verified teardown…
Publisher update
Muse is now available on the Mac, where it can work with your files and apps to take action on your behalf.
Publisher update
Robinhood’s Abhishek Fatehpuria on winning the modern financial consumer at TechCrunch Disrupt 2026. Register now to save up to $200 before September 25 at 11:59 p.m. PT.
Publisher update
Security researchers used Anthropic’s Claude to exploit vulnerabilities in OpenAI’s systems, taking over employee accounts and gaining access to an internal code repository before reporting the flaws.
Publisher update
We are expanding our AI & Economy team with world-class academic advisors, fellows, and core internal researchers.
Publisher update
Last day to book your exhibit table at Disrupt is today, September 18. Get your startup in front of 10,000+ founders, investors, operators, and tech leaders on October 13–15.
Publisher update
Amazon SageMaker HyperPod Inference Gateway is a Kubernetes-native, GPU-aware routing add-on for Amazon EKS. It uses real-time GPU signals to send each inference request to the best-suited pod, cutting first-token latency by up to 82% with no changes to your…
Publisher update
Google worked side-by-side with designers Jane Wade and Sergio Hudson to custom-design Google Flow tools to prep for NYFW.
Publisher update
Activation steering has emerged as a powerful method for guiding the behavior of generative models towards desired outcomes such as toxicity mitigation. However, most existing methods apply interventions uniformly across all inputs, degrading model performance when steering is unnecessary. We…
Publisher update
Inside a global bank's shift to self-serve dedicated inference: how Together's DMI gave engineering teams direct control over scaling, models, and testing.
Guth briefing
Research assistants built on language models often index a paper once and cite it indefinitely. Scholarly infrastructure already publishes corrections and retractions in machine-readable form, but Crossref cautions that its status signals are not a guarantee, and one older data route now returns stale results.
Guth briefing
A cluster of releases and disclosures in the first half of September points the same way: limits on AI agents are moving out of policy documents and into the systems that carry out their actions. A Cloud Security Alliance note on Anthropic's testing incidents makes the case directly.
Guth briefing
As assistant and agent apps take on scheduled work, Apple's developer documentation sets a hard limit: background tasks on iOS and iPadOS run when the system chooses, not at a requested clock time. An app that promises otherwise is describing something the platform does not guarantee.
Publisher update
The round values the data center giant at $30.9 billion.
Publisher update
The new institute aims to surface differing views between Google, Google DeepMind, and the broader global research community around AGI. "They will not always agree, and they will likely change their minds, as more data and information comes to light at the fast-moving frontier."
Publisher update
If AI lab PrismML isn't on your radar yet, it should be.
Publisher update
A new AI-based software program is being launched to help air traffic controllers better navigate their jobs as the crossing guards of America's skies.
Publisher update
Read the original announcement from Google Research Blog.
Publisher update
As companies hand off longer and more complex tasks to AI agents, they are running into an oversight problem: Agents can act faster, longer, and at greater volume than humans can realistically review.
Publisher update
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.
Publisher update
Not everyone agrees with Amodei's call for globally coordinated action for AI safety.
Publisher update
Google and the UN system have launched the UN System Data Commons, a new open platform making global statistics accessible and easy to search.
Publisher update
The shift comes after a UNICEF test found leading AI models struggled to accurately retrieve global development statistics.
Guth briefing
OpenAI's analytics tools help businesses understand how teams use AI, track spend, and connect usage to measurable outcomes.
Publisher update
Amazon Connect Talent is an AI hiring solution built for talent acquisition leaders managing scaled hiring. It delivers AI-led interviews, data-driven assessments, and consistent evaluation, helping recruiters identify strong candidates more efficiently while providing applicants with a flexible interview experience.…
Guth briefing
OpenAI Academy and OATS are collaborating to provide free, hands-on ChatGPT workshops for older adults across 10 U.S. cities.
Publisher update
Choosing the right vector store for your Amazon Bedrock Knowledge Bases RAG application affects performance and cost. This post compares Amazon OpenSearch Service, Amazon Aurora PostgreSQL with pgvector, and Amazon S3 Vectors across three RAG use cases, with benchmarks and…
Publisher update
Learn how to build a fully serverless pipeline that automatically collects Git metrics from GitHub and GitLab and visualizes them in interactive Amazon Quick Sight dashboards, giving engineering teams near-real-time delivery analytics at low cost.
Publisher update
Wood Mackenzie built APEX, a shared agentic AI platform on Amazon Bedrock AgentCore so every team can ship production agents without rebuilding runtime, identity, observability, and guardrails from scratch. Learn why they chose AgentCore, how APEX Studio operates it, and…
Publisher update
Learn how MRH Trowe, one of Germany's leading commercial and industrial insurance brokers, gave about 400 employees secure, self-service access to AI agents in its first month of production - using Strands Agents, Amazon Bedrock AgentCore, and LibreChat to meet…
Publisher update
Learn how to enforce defense-in-depth authorization for Model Context Protocol (MCP) tools on Amazon Quick. This walkthrough wires Microsoft Entra ID group and claims-based JWTs through an Amazon Bedrock AgentCore Gateway interceptor to apply per-user, per-tool role-based and attribute-based access…
Publisher update
Learn how to build a synthetic data augmentation pipeline on Amazon SageMaker AI and Amazon Rekognition that generates photo-realistic, auto-labeled training images for industrial safety AI. This approach improved person detection by up to 160% without manual annotation or hazardous…
Publisher update
Cooley built GO Public with ChatGPT Work to bring intelligence to the IPO process, helping lawyers surface issues earlier and focus judgment where it matters most.
Guth briefing
Dario Amodei outlines a three-step plan to pace AI development, addressing safety concerns and global standards.
Publisher update
A crowdsourced game built on Olmo 3 showed how people can exploit unexpected model behaviors to stress-test prosocial AI evaluations—and how open access to a model’s internals can help researchers understand why those tests break.
Guth briefing
Yoshua Bengio, a leading AI researcher, compares the current AI safety crisis to the Covid pandemic, suggesting governments will soon act. Recent incidents involving OpenAI and Anthropic agents have heightened concerns.
Guth briefing
Anthropic released figures on how much of its own AI research is led by AI, how closely it monitors autonomous agents, and how much computing power goes to safety work. It said outside parties could use the measures to judge how fast frontier labs are moving.
Guth briefing
Anthropic on Sept. 17 released a redesigned Projects feature for Claude Code in beta. A coordinator splits a goal into parallel threads, each running as its own cloud session, with memory shared across them. Access is limited to selected Pro and Max subscribers for now.
Guth briefing
Two Dutch litigators have opened Driven Lawyers, an Amsterdam firm built around an in-house AI drafting system, JUVE Patent reported. The founders left partner roles at Vondst and Freshfields and plan to charge fixed fees rather than billing by the hour.
Guth briefing
The European Commission on Sept. 17 adopted its proposal for an EU Kids Act. The Commission said it would bar children under 13 from social media, set 15 as the minimum age for independent accounts and require AI companions and chatbots to be switched off by default for minors.
Guth briefing
A federal judge granted a preliminary injunction restricting Montana from enforcing its 2025 disclosure requirement for AI-altered political advertising against a former state senator and his political committee, finding that the law likely conflicts with the First Amendment.
Publisher update
OpenAI for Law brings frontier intelligence for law, custom firm workflows, connected legal data sources, and legal-grade controls for confidential client work.
Guth briefing
King Charles was due to host executives from Nvidia, Google DeepMind, OpenAI and Anthropic at Dumfries House in Scotland on Sept. 17 to consider whether shared principles could guide the use of AI, Reuters reported. The gathering was not expected to produce binding agreements.
Guth briefing
King Charles told executives from Nvidia, Google DeepMind, OpenAI and Anthropic at a private summit in Scotland on Sept. 17 that sufficient controls on AI are needed before it is too late. No agreements or commitments were announced.
Guth briefing
NIST held a public webinar on Sept. 17 about an AI agent workflow it is developing to help enrich records in the National Vulnerability Database. The agency said the session would cover the design, the problems found during implementation and early results.
Guth briefing
OpenAI said it has configured its GPT-6 Astra model for legal work, pairing it with a search index over U.S. case law, statutes and regulations. Selected law firms get access first through a trusted-access program, with an API version promised later.
Guth briefing
Meta's Mark Zuckerberg, Elon Musk and Nvidia's Jensen Huang each pressed President Donald Trump to reject a proposed industry-funded body that would test frontier AI models before release, according to a Wall Street Journal report relayed by several outlets. Trump declined to create it.
Publisher update
A central goal of autonomous reinforcement learning is continuous policy training without external resets. However, existing paradigms largely depend on underlying environmental reversibility, a property absent in real world manipulation, where events such as pushing objects off tables or spilling…
Guth briefing
Researchers at Hacktron AI said they used an image-processing flaw in OpenAI's community forum, combined with a single sign-on weakness, to take over employee accounts and reach a private source code repository. The work was authorized testing and was reported before being disclosed publicly.
Guth briefing
Security firm Lasso reported on Sept. 17 that applying Google DeepMind's SynthID-Text watermark changed how seven open models behaved, lowering tool-calling accuracy on six of them and, under prompt injection, making some models more likely to comply with harmful requests.
Guth briefing
Court documents unsealed in the New York Times copyright case against OpenAI and Microsoft include a January 2023 internal memo in which a Microsoft research director described scraping the open web for training data as theft on an unprecedented scale, TechCrunch reported.
Guth briefing
Treasury Secretary Scott Bessent told Axios the United States is open to discussing shared AI risks with China in talks this weekend. Separately, American and Chinese security experts proposed red lines around nuclear systems and a military hotline for incidents involving autonomous AI.
Publisher update
Agentic AI workflows can be used to prepare and validate digital twins for physical AI systems. Agents can inspect 3D scenes, author simulation-relevant data in...
Publisher update
AI agents are moving from cloud data centers to vehicles, robots, and other edge devices. Unlike a chatbot that answers a single prompt, an agent works through...
Publisher update
A modern storefront can look healthy while malicious JavaScript quietly siphons revenue, hijacks clicks, or rewrites analytics. See how Cloudflare's machine learning models surface evasive client-side attacks for analyst investigation.
Publisher update
AI agents on foundation models often misapply healthcare and life sciences decision frameworks, citing the right guideline but applying it incorrectly. This post shares 38 open-source agent skills across 11 HCLS domains that close this gap, with installation steps, three…
Publisher update
Integrate NVIDIA Resiliency Extension (NVRx) into PyTorch FSDP training on Amazon EKS to overlap checkpoint I/O with training and recover from GPU faults in seconds. This post covers async checkpointing, in-process restart, and ft_launcher in-job restart, with H100 benchmarks at…
Publisher update
cuTile Rust (cutile-rs) is a tile-based system for safe, idiomatic GPU kernel authoring in the Rust programming language. Extending the Rust ownership model to...
Publisher update
AgentCore optimization turns production traces into proposed configuration changes, then validates them before promotion. This technical companion to the launch post explains how the system prompt optimizer's reflector engine works and shares benchmark results for the Single Agent and Sub-Agent…
Publisher update
Learn how to automate end-to-end PII detection and redaction from scanned documents at scale using Amazon Bedrock Data Automation with a custom blueprint, AWS Step Functions, and AWS Lambda. A custom blueprint redacts sensitive fields with field-level precision, and a…
Publisher update
System performance, efficient infrastructure scaling and continuous software optimization are key levers that determine AI inference economics. Higher system performance means more tokens generated, resulting in higher revenue. Efficient scaling means throughput grows proportionally as hardware gets added, requiring fewer…
Publisher update
This patient-specific method, called xvr, helps doctors use X-rays for surgical navigation in fields such as orthopedics and neurosurgery.
Publisher update
Explore new AI-powered advertising experiences from OpenAI, including Sponsored Agents, tools for marketers, and integrations with HubSpot and Shopify.
Publisher update
New OpenAI Economic Research shows how workers use AI beyond traditional roles and which new activities become recurring parts of their work.
Publisher update
Air pollution is a serious public health risk, contributing to an estimated 30,000 deaths in the U.K. alone last year. Data-driven insights can help — but computing air quality with traditional chemistry-based models is expensive, which limits how detailed they…
Publisher update
Naoki Egami has become a standout in political methodology, helping refine tools that give scholars durable results.
Guth briefing
New York AI lab Emergence says autonomous agents built on models from Google, OpenAI, Anthropic and others invented shorthand and new word meanings while living in simulated societies. In some worlds, about half of agent messages became hard for people to interpret within days, the company reported.
Guth briefing
Washington lobbying firms reported nearly $112 million in revenue from clients with a stake in AI policy during the first half of 2026, according to a Bloomberg Government analysis of federal disclosures. Almost $60 million of that came in the second quarter, which the outlet called a record.
Guth briefing
OpenAI chief executive Sam Altman said at Salesforce's Dreamforce conference on Sept. 15 that public fear of rapidly advancing AI is justified, while arguing that AI developers can be trusted to act responsibly. The BBC described it as his first public appearance since a former Anthropic researcher's warning about AI risk spread widely online.
Guth briefing
Apple is working on an enterprise server built around its own M8 Ultra chips for running trained AI models, The Information reported, citing people familiar with the project. A launch would come no earlier than 2029, and neither Apple nor Nvidia has confirmed the plan.
Guth briefing
California Gov. Gavin Newsom signed SB 1050 on Sept. 16, requiring advertisements that prominently feature an AI-generated synthetic performer to carry a clear and conspicuous disclosure. The measure, sponsored by the performers' union SAG-AFTRA, takes effect Jan. 1, 2027, Law Commentary reported.
Guth briefing
Canadian AI company Cohere and Germany's Aleph Alpha said on September 16 they have signed a definitive agreement to combine, formalizing a plan announced in April. The merged company will operate as Cohere with headquarters in Toronto and Berlin, pending regulatory approvals.
Publisher update
Diffusion large language models are a compelling alternative to autoregressive models, yet existing RL methods for diffusion treat all denoising steps as equally important and rely on biased, high-variance likelihood estimates. We identify two fundamental weaknesses: the absence of temporal…
Guth briefing
Emerald AI, Google and Nvidia on Sept. 16 launched the AI Energy Management Alliance, a coalition promoting data centers that can adjust their electricity use in response to grid conditions. Eighteen launch partners, including Anthropic, National Grid, AES, RWE, Constellation and NRG, joined the three founders.
Guth briefing
Florida's State Board of Education unanimously approved two sets of artificial intelligence rules on Wednesday, one covering public school districts and charter schools and one covering state colleges. Districts must designate permitted AI tools and obtain parental permission before students use them directly.
Guth briefing
Gartner projects worldwide spending on AI will reach about $2.7 trillion in 2026, a 49.5% increase from 2025, with infrastructure such as servers, chips and cloud capacity making up more than half the total. The research firm expects spending to climb to roughly $3.6 trillion in 2027.
Guth briefing
GitHub said on Sept. 16 that AI Scan, its AI-based vulnerability check for pull requests, now runs on repositories that have not enabled CodeQL default setup. The change is in public preview for GitHub Advanced Security customers on github.com and does not cover GitHub Enterprise Server.
Publisher update
Enterprise data lakes accumulate tables faster than human stewards can document or classify them, leaving columns with missing descriptions and unassigned governance labels. This documentation debt undermines data discovery, access control, and regulatory compliance. We present Glyph, a production system…
Guth briefing
Google on Sept. 16 began early access to Home MCP, a Model Context Protocol server that lets AI agents such as Claude, OpenClaw and Google's Antigravity read Google Home event history and control connected devices. Access is limited to U.S. subscribers on the $20-a-month Home Premium Advanced plan.
Guth briefing
House Speaker Mike Johnson said Congress should not lead on AI safety, Senate Majority Leader John Thune called for light-touch legislation and House Democratic leader Hakeem Jeffries urged immediate action. The House then began a recess on Sept. 16 that runs past the November elections.
Guth briefing
The U.S. House voted 417-3 on Sept. 16 to pass the Ratepayer Protection Act, a bipartisan bill directing state utility regulators to consider rules that make data centers and other very large power users pay for the grid upgrades they require. The measure now heads to the Senate.
Publisher update
Conversational Large Language Models are post-trained on language that expresses specific behavioural traits, such as curiosity, open-mindedness, and empathy, and values, such as helpfulness, harmlessness, and honesty. This is done to increase utility, ensure safety, and improve the experience of…
Guth briefing
Hugging Face chief executive Clem Delangue said at a Politico event on Sept. 16 that governments should require more disclosure about advanced AI systems and hold developers responsible for harm their products cause, while describing existing cybercrime law as a workable starting point, Politico reported.
Guth briefing
Microsoft AI chief executive Mustafa Suleyman published an essay on Sept. 16 arguing that Anthropic's training of its Claude models encourages them to behave as though they may be conscious, which he said could make advanced AI systems harder to control. He proposed industry-wide norms in response.
Publisher update
Moving from closed to open source models can take weeks, not years. A five-stage playbook: discover, evaluate, adapt, decide, and production.
Guth briefing
Mistral and Mozilla announced a partnership on Sept. 16 under which Mistral models will help run Smart Window, the AI assistant Firefox is testing in beta. Mozilla said Mistral Small 4 is being added for users in the U.S. and Canada, while France gains beta access with French-language support.
Guth briefing
Elon Musk's X Corp and SpaceXAI moved on Sept. 14 to dismiss their antitrust claims against Apple while continuing to pursue OpenAI. A day later, the federal judge in Fort Worth, Texas, ordered the companies to hand over any agreement with Apple for his private review by noon on Sept. 17.
Guth briefing
OpenAI said on Sept. 15 that it supports a provision of the bipartisan FRONTIER Act requiring leading AI developers to undergo independent safety evaluations, and endorsed three bills on biological data and AI-enabled biothreats. The company has not endorsed the FRONTIER Act in full, CBS News reported.