Curated AI news, research, and engineering updates
AI Info brings together AI news, research posts, engineering write-ups, and product announcements from major labs, companies, and communities in one crawlable hub.
How we contain Claude across products
As agents grow more capable, so does their potential blast radius. The engineering question is how to cap it. Here’s what we’ve learned building containment for claude.ai, Claude Code, and Cowork.
2026-05-25An update on recent Claude Code quality reports
We traced recent reports of Claude Code quality issues to three separate changes. Here's what happened and what we're changing.
2026-04-23Scaling Managed Agents: Decoupling the brain from the hands
Harnesses encode assumptions that go stale as models improve. Managed Agents—our hosted service for long-horizon agent work—is built around interfaces that stay stable as harnesses change.
2026-04-08How we built Claude Code auto mode: a safer way to skip permissions
Claude Code users approve 93% of permission prompts. We built classifiers to automate some decisions, increasing safety while reducing approval fatigue. Here's what it catches, and what it misses.
2026-03-25Harness design for long-running application development
Harness design is key to performance at the frontier of agentic coding. Here's how we pushed Claude further in frontend design and long-running autonomous software engineering.
2026-03-24permissionlesstech / bitchat
bluetooth mesh chat, IRC vibes
Trending todayamnezia-vpn / amnezia-client
Amnezia VPN Client (Desktop+Mobile)
Trending todaymoeru-ai / airi
💖🧸 Self hosted, you-owned Grok Companion, a container of souls of waifu, cyber livings to bring them into our worlds, wishing to achieve Neuro-sama's altitude. Capable of realtime voice chat, Minecraft, Factorio playing. Web / macOS / Windows supported.
Trending todayopengeos / GeoLibre
A lightweight, cloud-native GIS platform for visualizing, exploring, and analyzing geospatial data. It runs in the web browser, on the desktop, on mobile, and inside Jupyter notebooks.
Trending todayyorukot / superfile
Pretty fancy and modern terminal file manager
Trending todayProject Pilot: Can AI control a drone?
Working with Andon Labs, we’ve developed a new series of evaluations that assess AI models’ ability to use a flying drone, culminating in a new benchmark: Drone-Bench.
2026-07-24T17:00:00.000ZHow Canada uses Claude: Findings from the Anthropic Economic Index
2026-07-14T13:00:00.000ZClaude’s values across models and languages
2026-07-13T17:08:00.000ZClaude plays robotics
In project Fetch, we examined how humans can use models to get robots to perform complex tasks. Now, we investigate many models on a large variety of different robotics tasks in simulation, to see how good models are at controlling robots themselves.
2026-07-09T18:00:00.000ZAn off switch for dual-use knowledge in AI models
2026-07-08T21:46:00.000ZAccelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission
Google commits $40M in AI tokens and credits for the Genesis Mission
Wed, 22 Jul 2026 13:38:54 +0000Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
We’re introducing new Gemini models, including Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber.
Tue, 21 Jul 2026 15:16:30 +0000Introducing Gemini 3.5 Flash Cyber
Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.
Fri, 17 Jul 2026 15:00:11 +0000Our approach to bioresilience
Google DeepMind and Isomorphic Labs are sharing our joint approach to bioresilience and AI models.
Thu, 16 Jul 2026 09:30:42 +0000Empowering India’s next generation of innovators with ATL Saathi
Google and AIM launched ATL Saathi, a Gemini-powered AI tool empowering Indian educators in robotics labs.
Mon, 13 Jul 2026 12:37:28 +0000Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization
Hierarchical Interest Representation is a research area for Meta Ads. We’re exploring an upstream representation layer over the universe of Ads entities – users, advertisers, products, services – learning unified embeddings that connect users’ inferred interests with the breadth of what advertisers offer in their deep funnel ads. The innovations in Hierarchical Interest Representation are [...] Read More... The post Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization appeared first on Engineering at Meta.
Wed, 15 Jul 2026 17:00:52 +000010 Years of Meta’s Commitment to Python
This year marks Meta’s 10th consecutive year as a sponsor of the Python Software Foundation (PSF), the charitable organization dedicated to advancing, supporting, and protecting the open-source Python programming language and the community that sustains it. Python is one of the world’s most influential programming languages, and we use it across our engineering stack, from [...] Read More... The post 10 Years of Meta’s Commitment to Python appeared first on Engineering at Meta.
Tue, 30 Jun 2026 16:00:46 +0000RCCLX: Innovating GPU Communications on AMD Platforms
We are open-sourcing the initial version of RCCLX – an enhanced version of RCCL that we developed and tested on Meta’s internal workloads. RCCLX is fully integrated with Torchcomms and aims to empower researchers and developers to accelerate innovation, regardless of their chosen backend. Communication patterns for AI models are constantly evolving, as are hardware [...] Read More... The post RCCLX: Innovating GPU Communications on AMD Platforms appeared first on Engineering at Meta.
Tue, 24 Feb 2026 21:30:54 +0000Scaling LLM Inference: Innovations in Tensor Parallelism, Context Parallelism, and Expert Parallelism
At Meta, we are constantly pushing the boundaries of LLM inference systems to power applications such as the Meta AI App. We’re sharing how we developed and implemented advanced parallelism techniques to optimize key performance metrics related to resource efficiency, throughput, and latency. The rapid evolution of large language models (LLMs) has ushered in a [...] Read More... The post Scaling LLM Inference: Innovations in Tensor Parallelism, Context Parallelism, and Expert Parallelism appeared first on Engineering at Meta.
Fri, 17 Oct 2025 16:00:50 +0000LLMs Are the Key to Mutation Testing and Better Compliance
Following our keynote presentations at FSE 2025 and Eurostar 2025, we’re delving further into the development of Meta’s Automated Compliance Hardening (ACH) tool, an LLM-based tool for software testing that is automating aspects of compliance adherence at Meta, while accelerating developer and product velocity. By leveraging LLMs we’ve been able to overcome the barriers that [...] Read More... The post LLMs Are the Key to Mutation Testing and Better Compliance appeared first on Engineering at Meta.
Tue, 30 Sep 2025 16:00:08 +0000How AI is expanding what people do at work
New OpenAI research shows how AI is expanding what workers do, with ChatGPT users taking on tasks across roles and reshaping job boundaries.
Mon, 27 Jul 2026 03:30:00 GMTLaunching Health in ChatGPT
Health in ChatGPT now lets eligible U.S. users securely connect medical records and Apple Health to get more personalized insights and better understand their health.
Thu, 23 Jul 2026 00:00:00 GMTBuilding AI infrastructure with the Effingham County community
OpenAI announces Project Camellia in Effingham County, Georgia, with commitments to responsible energy, community investment, jobs, and access to Codex.
Wed, 22 Jul 2026 13:00:00 GMTHow news organizations are using AI to advance their vital missions
News organizations are using AI to strengthen reporting, grow audiences, and improve business operations, with OpenAI tools supporting journalists and publishers worldwide.
Wed, 22 Jul 2026 13:00:00 GMTAdvancing the next era of national science
OpenAI outlines its commitment to advancing American science working with the U.S. Department of Energy and national labs to use frontier AI to accelerate discovery.
Wed, 22 Jul 2026 12:00:00 GMTBeyond RAG: Task-aware knowledge compression for enterprise AI on AWS
Traditional RAG hits a ceiling on analytical tasks that span hundreds of documents. This post shows how to use task-aware knowledge compression (TAKC) on AWS to pre-compress entire knowledge bases into task-specific representations, cache them at multiple fidelity tiers, and route each query to the right tier, with an open-source implementation you can deploy.
Mon, 27 Jul 2026 16:11:32 +0000Deepgram enhances Amazon SageMaker AI support with AWS IAM Temporary Delegation
In this post, we cover why Deepgram built on IAM temporary delegation, how the integration works end-to-end, and what it unlocks for customers running Deepgram speech models on SageMaker AI. With this integration, Deepgram has reduced the time for initial investigation on a SageMaker AI support ticket from days to minutes.
Mon, 27 Jul 2026 16:07:44 +0000How Guardoc transforms medical document processing with Amazon Nova models
In this post, we explore how Guardoc Health uses the Amazon Nova family of models, available through Amazon Bedrock, to transform clinical documentation in long-term care.
Mon, 27 Jul 2026 16:05:00 +0000Introducing Claude Opus 5 on AWS: Anthropic’s most capable Opus model
This post covers Opus 5’s improvements and practical guidance for AI engineers integrating the model into agentic systems and production inference workloads on Amazon Bedrock. See the documentation for Claude Platform on AWS.
Fri, 24 Jul 2026 17:59:03 +0000Build an explainable next-best-product recommendation system for banking on AWS
Learn the architecture and design decisions behind an explainable next-best-product recommendation system for banking, built with Amazon SageMaker AI and PyTorch. A multi-tower neural network with learned attention delivers accurate, per-customer recommendations while providing the explainability that banking regulators require.
Fri, 24 Jul 2026 15:42:11 +0000SymptomAI: Towards a conversational AI agent for everyday symptom assessment
General Science
Wed, 22 Jul 2026 21:32:00 +0000Towards a quantum computer that learns from its errors
Machine Intelligence
Wed, 22 Jul 2026 18:40:21 +0000Towards demystifying the creativity of diffusion models
Algorithms & Theory
Wed, 15 Jul 2026 18:06:00 +0000SensorFM: Towards a general intelligence and interface for wearable health data
Generative AI
Thu, 09 Jul 2026 09:56:00 +0000The power of collaboration: How we can reduce traffic congestion
Algorithms & Theory
Tue, 07 Jul 2026 16:42:08 +0000Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization
Hierarchical Interest Representation is a research area for Meta Ads. We’re exploring an upstream representation layer over the universe of Ads entities – users, advertisers, products, services – learning unified embeddings that connect users’ inferred interests with the breadth of what advertisers offer in their deep funnel ads. The innovations in Hierarchical Interest Representation are [...] Read More... The post Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization appeared first on Engineering at Meta.
Wed, 15 Jul 2026 17:00:52 +0000Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler
TL; DR At Meta’s scale, a few milliseconds of latency degradation can have a significant negative impact on ads performance. When a Linux kernel upgrade risked regressing latency across Meta’s ad serving fleet, we turned to sched_ext — the upstream, BPF-based extensible scheduling framework — to build a scheduling policy customized to the Ads delivery [...] Read More... The post Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler appeared first on Engineering at Meta.
Mon, 13 Jul 2026 16:00:50 +0000Meta’s AI Storage Blueprint at Scale
Over the past several years, model capabilities and training dataset sizes have experienced exponential growth. During the past year or so, the time between new-frontier-model releases has gone down from months to weeks. Reliable and fast access to storage is important to both the speed and computational cost of this AI innovation. If AI is [...] Read More... The post Meta’s AI Storage Blueprint at Scale appeared first on Engineering at Meta.
Wed, 01 Jul 2026 16:00:36 +000010 Years of Meta’s Commitment to Python
This year marks Meta’s 10th consecutive year as a sponsor of the Python Software Foundation (PSF), the charitable organization dedicated to advancing, supporting, and protecting the open-source Python programming language and the community that sustains it. Python is one of the world’s most influential programming languages, and we use it across our engineering stack, from [...] Read More... The post 10 Years of Meta’s Commitment to Python appeared first on Engineering at Meta.
Tue, 30 Jun 2026 16:00:46 +0000Privacy-Aware Infrastructure in the AI-Native Era: An Asset Classification Case Study
Privacy controls — systems that enforce retention, access, allowed-purpose, downstream-sharing, or anonymization policies — require a reliable understanding of data to function. Before such a control can operate effectively, it must know exactly what it is looking at. This can be complex, as demonstrated by a field simply named “age“: In one context, it [...] Read More... The post Privacy-Aware Infrastructure in the AI-Native Era: An Asset Classification Case Study appeared first on Engineering at Meta.
Thu, 25 Jun 2026 22:30:51 +0000Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission
Google commits $40M in AI tokens and credits for the Genesis Mission
Wed, 22 Jul 2026 13:38:54 +0000Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
We’re introducing new Gemini models, including Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber.
Tue, 21 Jul 2026 15:16:30 +0000Introducing Gemini 3.5 Flash Cyber
Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.
Fri, 17 Jul 2026 15:00:11 +0000Our approach to bioresilience
Google DeepMind and Isomorphic Labs are sharing our joint approach to bioresilience and AI models.
Thu, 16 Jul 2026 09:30:42 +0000Empowering India’s next generation of innovators with ATL Saathi
Google and AIM launched ATL Saathi, a Gemini-powered AI tool empowering Indian educators in robotics labs.
Mon, 13 Jul 2026 12:37:28 +0000Our position on open-weights models
2026-07-27T18:36:00.000ZCognizant and Anthropic expand their partnership to bring Claude to enterprise clients
2026-07-27T15:32:00.000ZIntroducing Claude Opus 5
Opus 5 is a step change improvement for the Opus tier powering long-running agents while delivering improvements in coding and professional work.
2026-07-24T17:00:00.000ZA research agenda for the Economic Futures Research Fund
We’re sharing the research agenda for the Anthropic Economic Futures Research Fund.
2026-07-22T17:00:00.000ZAsk Claude about the Anthropic Economic Index
We're launching the Anthropic Economic Index connector for Claude, which lets anyone explore real data about AI and work.
2026-07-22T17:00:00.000Z3 Google updates from Galaxy Unpacked 2026
We shared how Samsung users can boost productivity and get time back on new foldables, watches, and glasses coming soon.
Wed, 22 Jul 2026 13:00:00 +0000Connect more of your apps to Search
You’ll be able to securely link and interact with your go-to services directly in AI Mode.
Thu, 16 Jul 2026 16:00:00 +0000Create, edit and star in videos with two Google Vids updates
Gemini Omni and personal avatars in Google Vids make video creation easier than ever.
Thu, 16 Jul 2026 16:00:00 +0000Celebrating 25 years of visual search innovation
Google Images is turning 25. Here’s a look back at some major milestones — and new ways to explore and create visual content.
Tue, 14 Jul 2026 16:00:00 +0000Expanding Managed Agents in Gemini API: background tasks, remote MCP and more
We’re announcing new capabilities in Managed Agents in Gemini API so developers can build reliable, production-ready agents.
Tue, 07 Jul 2026 08:54:00 +0000#498 – Anthony Kaldellis: Roman Empire, Byzantine Empire, Rise & Fall of Empires
Anthony Kaldellis is a historian of the Roman Empire and author of “The New Roman Empire”, a comprehensive history of the Byzantine Empire (Eastern Roman Empire). https://lexfridman.com/sponsors/ep498-sc Transcript: https://lexfridman.com/anthony-kaldellis-transcript CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://amzn.to/49AX7Q1 https://kaldellispublications.weebly.com https://classics.uchicago.edu/people/anthony-kaldellis https://amzn.to/3PTFTqk https://amzn.to/4fgRMRq https://byzantiumandfriends.podbean.com/ https://thehistoryofbyzantium.com/ SPONSORS: Upwork: Platform for hiring freelancers. https://upwork.com/lex Fin: AI agent for customer service. https://fin.ai/lex BetterHelp: Online therapy and counseling. https://betterhelp.com/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Shopify: Sell stuff online. https://shopify.com/lex Perplexity: AI-powered answer engine. https://perplexity.ai/ OUTLINE: PODCAST LINKS: https://lexfridman.com/podcast https://apple.co/2lwqZIr https://spoti.fi/2nEwCF8 https://lexfridman.com/feed/podcast/ https://www.youtube.com/playlist?list=PLrAXtmErZgOdP_8GztsuKi9nrraNbKKp4 https://www.youtube.com/lexclips
Tue, 30 Jun 2026 21:33:40 +0000#497 – Biggest Mysteries in Physics: Antimatter, Dark Energy & ToE – Don Lincoln
Don Lincoln is a particle physicist at Fermilab who has spent decades working at the frontiers of high energy physics. https://lexfridman.com/sponsors/ep497-sc CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://facebook.com/Dr.Don.Lincoln/ https://drdonlincoln.com/ https://bit.ly/4nHeNiF https://bit.ly/3PCIW67 https://x.com/DrDonLincoln https://amzn.to/4uYbkOZ https://shop.thegreatcourses.com/don-lincoln https://adbl.co/4wGioRV https://www.youtube.com/fermilab https://www.fnal.gov/ https://x.com/fermilab SPONSORS: Upwork: Platform for hiring freelancers. https://upwork.com/lex Larridin: Measure AI adoption in your business. https://larridin.com Fin: AI agent for customer service. https://fin.ai/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Shopify: Sell stuff online. https://shopify.com/lex Perplexity: AI-powered answer engine. https://perplexity.ai/ OUTLINE: PODCAST LINKS: https://lexfridman.com/podcast https://apple.co/2lwqZIr https://spoti.fi/2nEwCF8 https://lexfridman.com/feed/podcast/ https://www.youtube.com/playlist?list=PLrAXtmErZgOdP_8GztsuKi9nrraNbKKp4 https://www.youtube.com/lexclips
Fri, 29 May 2026 16:22:02 +0000#496 – FFmpeg: The Incredible Technology Behind Video on the Internet
Jean-Baptiste Kempf is lead developer of VLC and president of VideoLAN. Kieran Kunhya is a longtime FFmpeg contributor, codec engineer, and the person behind the now-infamous FFmpeg account on X. https://lexfridman.com/sponsors/ep496-sc Transcript: https://lexfridman.com/ffmpeg-transcript CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://x.com/FFmpeg https://ffmpeg.org/ https://www.videolan.org/ https://x.com/videolan https://jbkempf.com/ https://www.linkedin.com/in/jbkempf/ https://github.com/jbkempf https://x.com/kierank_ https://bit.ly/3OORhmC https://github.com/kierank SPONSORS: Larridin: Measure AI adoption in your business. https://larridin.com Blitzy: AI agent for large enterprise codebases. https://blitzy.com/lex BetterHelp: Online therapy and counseling. https://betterhelp.com/lex Fin: AI agent for customer service. https://fin.ai/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Perplexity: AI-powered answer engine. https://perplexity.ai/ OUTLINE: (00:00) – Introduction (03:00) – Sponsors, Comments, and Reflections (10:48) – Weirdest things VLC opens (15:12) – How video playback works (24:33) – Video codecs and containers (35:20) – FFmpeg explained (56:20) – Linus Torvalds (1:00:59) – Turning down millions to keep VLC ad-free (1:15:17) – FFmpeg & Google drama (1:34:31) – FFmpeg developers (1:41:08) – VLC and FFmpeg (1:45:42) – History of FFmpeg (1:48:59) – Reverse engineering codecs (2:02:14) – FFmpeg testing (2:06:21) – Assembly code (handwritten) (2:30:39) – Rust programming language (2:39:55) – FFmpeg and Libav fork (2:48:17) – Open source burnout (2:56:04) – x264 and internet video (3:09:20) – Video compression basics (3:16:17) – CIA and fake VLC (3:26:52) – Ultra low latency streaming (3:44:20) – AV2 codec and video patents (3:54:12) – VLC backdoors (4:04:27) – Video archiving (4:11:04) – Future of FFmpeg and VLC
Wed, 06 May 2026 22:06:47 +0000#495 – Vikings, Ragnar, Berserkers, Valhalla & the Warriors of the Viking Age
Lars Brownworth is a historian, teacher, podcaster, and author specializing in Viking history, medieval Europe, and the Byzantine Empire. https://lexfridman.com/sponsors/ep495-sc Transcript: https://lexfridman.com/lars-brownworth-transcript CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://larsbrownworth.com/ https://www.amazon.com/Sea-Wolves-History-Vikings/dp/1909979120 https://amzn.to/4sHY0xw https://12byzantinerulers.com/ https://apple.co/4sgSxNi SPONSORS: Larridin: Measure AI adoption in your business. https://larridin.com BetterHelp: Online therapy and counseling. https://betterhelp.com/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Fin: AI agent for customer service. https://fin.ai/lex Shopify: Sell stuff online. https://shopify.com/lex Perplexity: AI-powered answer engine. https://perplexity.ai/ OUTLINE: PODCAST LINKS: https://lexfridman.com/podcast https://apple.co/2lwqZIr https://spoti.fi/2nEwCF8 https://lexfridman.com/feed/podcast/ https://www.youtube.com/playlist?list=PLrAXtmErZgOdP_8GztsuKi9nrraNbKKp4 https://www.youtube.com/lexclips
Thu, 09 Apr 2026 17:43:17 +0000#494 – Jensen Huang: NVIDIA – The $4 Trillion Company & the AI Revolution
Jensen Huang is the co-founder and CEO of NVIDIA, the world’s most valuable company and the engine powering the AI computing revolution. https://lexfridman.com/sponsors/ep494-sc Transcript: https://lexfridman.com/jensen-huang-transcript CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://nvidia.com https://x.com/nvidia https://x.com/NVIDIAAI https://youtube.com/@nvidia https://www.instagram.com/nvidia/ https://www.linkedin.com/company/nvidia/ https://www.facebook.com/NVIDIA/ https://github.com/NVIDIA https://developer.nvidia.com/nemotron SPONSORS: Perplexity: AI-powered answer engine. https://perplexity.ai/ Shopify: Sell stuff online. https://shopify.com/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Fin: AI agent for customer service. https://fin.ai/lex Quo: Phone system (calls, texts, contacts) for businesses. https://quo.com/lex OUTLINE: PODCAST LINKS: https://lexfridman.com/podcast https://apple.co/2lwqZIr https://spoti.fi/2nEwCF8 https://lexfridman.com/feed/podcast/ https://www.youtube.com/playlist?list=PLrAXtmErZgOdP_8GztsuKi9nrraNbKKp4 https://www.youtube.com/lexclips
Mon, 23 Mar 2026 16:28:42 +0000