AI Info

21h ago

Filter:

Curated AI news, research, and engineering updates

AI Info brings together AI news, research posts, engineering write-ups, and product announcements from major labs, companies, and communities in one crawlable hub.

Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization

Hierarchical Interest Representation is a research area for Meta Ads. We’re exploring an upstream representation layer over the universe of Ads entities – users, advertisers, products, services – learning unified embeddings that connect users’ inferred interests with the breadth of what advertisers offer in their deep funnel ads. The innovations in Hierarchical Interest Representation are [...] Read More... The post Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization appeared first on Engineering at Meta.

Wed, 15 Jul 2026 17:00:52 +0000

10 Years of Meta’s Commitment to Python

This year marks Meta’s 10th consecutive year as a sponsor of the Python Software Foundation (PSF), the charitable organization dedicated to advancing, supporting, and protecting the open-source Python programming language and the community that sustains it. Python is one of the world’s most influential programming languages, and we use it across our engineering stack, from [...] Read More... The post 10 Years of Meta’s Commitment to Python appeared first on Engineering at Meta.

Tue, 30 Jun 2026 16:00:46 +0000

RCCLX: Innovating GPU Communications on AMD Platforms

We are open-sourcing the initial version of RCCLX – an enhanced version of RCCL that we developed and tested on Meta’s internal workloads. RCCLX is fully integrated with Torchcomms and aims to empower researchers and developers to accelerate innovation, regardless of their chosen backend. Communication patterns for AI models are constantly evolving, as are hardware [...] Read More... The post RCCLX: Innovating GPU Communications on AMD Platforms appeared first on Engineering at Meta.

Tue, 24 Feb 2026 21:30:54 +0000

Scaling LLM Inference: Innovations in Tensor Parallelism, Context Parallelism, and Expert Parallelism

At Meta, we are constantly pushing the boundaries of LLM inference systems to power applications such as the Meta AI App. We’re sharing how we developed and implemented advanced parallelism techniques to optimize key performance metrics related to resource efficiency, throughput, and latency. The rapid evolution of large language models (LLMs) has ushered in a [...] Read More... The post Scaling LLM Inference: Innovations in Tensor Parallelism, Context Parallelism, and Expert Parallelism appeared first on Engineering at Meta.

Fri, 17 Oct 2025 16:00:50 +0000

LLMs Are the Key to Mutation Testing and Better Compliance

Following our keynote presentations at FSE 2025 and Eurostar 2025, we’re delving further into the development of Meta’s Automated Compliance Hardening (ACH) tool, an LLM-based tool for software testing that is automating aspects of compliance adherence at Meta, while accelerating developer and product velocity. By leveraging LLMs we’ve been able to overcome the barriers that [...] Read More... The post LLMs Are the Key to Mutation Testing and Better Compliance appeared first on Engineering at Meta.

Tue, 30 Sep 2025 16:00:08 +0000

Beyond RAG: Task-aware knowledge compression for enterprise AI on AWS

Traditional RAG hits a ceiling on analytical tasks that span hundreds of documents. This post shows how to use task-aware knowledge compression (TAKC) on AWS to pre-compress entire knowledge bases into task-specific representations, cache them at multiple fidelity tiers, and route each query to the right tier, with an open-source implementation you can deploy.

Mon, 27 Jul 2026 16:11:32 +0000

Deepgram enhances Amazon SageMaker AI support with AWS IAM Temporary Delegation

In this post, we cover why Deepgram built on IAM temporary delegation, how the integration works end-to-end, and what it unlocks for customers running Deepgram speech models on SageMaker AI. With this integration, Deepgram has reduced the time for initial investigation on a SageMaker AI support ticket from days to minutes.

Mon, 27 Jul 2026 16:07:44 +0000

How Guardoc transforms medical document processing with Amazon Nova models

In this post, we explore how Guardoc Health uses the Amazon Nova family of models, available through Amazon Bedrock, to transform clinical documentation in long-term care.

Mon, 27 Jul 2026 16:05:00 +0000

Introducing Claude Opus 5 on AWS: Anthropic’s most capable Opus model

This post covers Opus 5’s improvements and practical guidance for AI engineers integrating the model into agentic systems and production inference workloads on Amazon Bedrock. See the documentation for Claude Platform on AWS.

Fri, 24 Jul 2026 17:59:03 +0000

Build an explainable next-best-product recommendation system for banking on AWS

Learn the architecture and design decisions behind an explainable next-best-product recommendation system for banking, built with Amazon SageMaker AI and PyTorch. A multi-tower neural network with learned attention delivers accurate, per-customer recommendations while providing the explainability that banking regulators require.

Fri, 24 Jul 2026 15:42:11 +0000

Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization

Hierarchical Interest Representation is a research area for Meta Ads. We’re exploring an upstream representation layer over the universe of Ads entities – users, advertisers, products, services – learning unified embeddings that connect users’ inferred interests with the breadth of what advertisers offer in their deep funnel ads. The innovations in Hierarchical Interest Representation are [...] Read More... The post Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization appeared first on Engineering at Meta.

Wed, 15 Jul 2026 17:00:52 +0000

Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler

TL; DR At Meta’s scale, a few milliseconds of latency degradation can have a significant negative impact on ads performance.  When a Linux kernel upgrade risked regressing latency across Meta’s ad serving fleet, we turned to sched_ext — the upstream, BPF-based extensible scheduling framework — to build a scheduling policy customized to the Ads delivery [...] Read More... The post Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler appeared first on Engineering at Meta.

Mon, 13 Jul 2026 16:00:50 +0000

Meta’s AI Storage Blueprint at Scale

Over the past several years, model capabilities and training dataset sizes have experienced exponential growth. During the past year or so, the time between new-frontier-model releases has gone down from months to weeks. Reliable and fast access to storage is important to both the speed and computational cost of this AI innovation. If AI is [...] Read More... The post Meta’s AI Storage Blueprint at Scale appeared first on Engineering at Meta.

Wed, 01 Jul 2026 16:00:36 +0000

10 Years of Meta’s Commitment to Python

This year marks Meta’s 10th consecutive year as a sponsor of the Python Software Foundation (PSF), the charitable organization dedicated to advancing, supporting, and protecting the open-source Python programming language and the community that sustains it. Python is one of the world’s most influential programming languages, and we use it across our engineering stack, from [...] Read More... The post 10 Years of Meta’s Commitment to Python appeared first on Engineering at Meta.

Tue, 30 Jun 2026 16:00:46 +0000

Privacy-Aware Infrastructure in the AI-Native Era: An Asset Classification Case Study

Privacy controls — systems that enforce retention, access, allowed-purpose, downstream-sharing, or anonymization policies — require a reliable understanding of data to function. Before such a control can operate effectively, it must know exactly what it is looking at. This can be complex, as demonstrated by a field simply named “age“: In one context, it [...] Read More... The post Privacy-Aware Infrastructure in the AI-Native Era: An Asset Classification Case Study appeared first on Engineering at Meta.

Thu, 25 Jun 2026 22:30:51 +0000

#498 – Anthony Kaldellis: Roman Empire, Byzantine Empire, Rise & Fall of Empires

Anthony Kaldellis is a historian of the Roman Empire and author of “The New Roman Empire”, a comprehensive history of the Byzantine Empire (Eastern Roman Empire). https://lexfridman.com/sponsors/ep498-sc Transcript: https://lexfridman.com/anthony-kaldellis-transcript CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://amzn.to/49AX7Q1 https://kaldellispublications.weebly.com https://classics.uchicago.edu/people/anthony-kaldellis https://amzn.to/3PTFTqk https://amzn.to/4fgRMRq https://byzantiumandfriends.podbean.com/ https://thehistoryofbyzantium.com/ SPONSORS: Upwork: Platform for hiring freelancers. https://upwork.com/lex Fin: AI agent for customer service. https://fin.ai/lex BetterHelp: Online therapy and counseling. https://betterhelp.com/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Shopify: Sell stuff online. https://shopify.com/lex Perplexity: AI-powered answer engine. https://perplexity.ai/ OUTLINE: PODCAST LINKS: https://lexfridman.com/podcast https://apple.co/2lwqZIr https://spoti.fi/2nEwCF8 https://lexfridman.com/feed/podcast/ https://www.youtube.com/playlist?list=PLrAXtmErZgOdP_8GztsuKi9nrraNbKKp4 https://www.youtube.com/lexclips

Tue, 30 Jun 2026 21:33:40 +0000

#497 – Biggest Mysteries in Physics: Antimatter, Dark Energy & ToE – Don Lincoln

Don Lincoln is a particle physicist at Fermilab who has spent decades working at the frontiers of high energy physics. https://lexfridman.com/sponsors/ep497-sc CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://facebook.com/Dr.Don.Lincoln/ https://drdonlincoln.com/ https://bit.ly/4nHeNiF https://bit.ly/3PCIW67 https://x.com/DrDonLincoln https://amzn.to/4uYbkOZ https://shop.thegreatcourses.com/don-lincoln https://adbl.co/4wGioRV https://www.youtube.com/fermilab https://www.fnal.gov/ https://x.com/fermilab SPONSORS: Upwork: Platform for hiring freelancers. https://upwork.com/lex Larridin: Measure AI adoption in your business. https://larridin.com Fin: AI agent for customer service. https://fin.ai/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Shopify: Sell stuff online. https://shopify.com/lex Perplexity: AI-powered answer engine. https://perplexity.ai/ OUTLINE: PODCAST LINKS: https://lexfridman.com/podcast https://apple.co/2lwqZIr https://spoti.fi/2nEwCF8 https://lexfridman.com/feed/podcast/ https://www.youtube.com/playlist?list=PLrAXtmErZgOdP_8GztsuKi9nrraNbKKp4 https://www.youtube.com/lexclips

Fri, 29 May 2026 16:22:02 +0000

#496 – FFmpeg: The Incredible Technology Behind Video on the Internet

Jean-Baptiste Kempf is lead developer of VLC and president of VideoLAN. Kieran Kunhya is a longtime FFmpeg contributor, codec engineer, and the person behind the now-infamous FFmpeg account on X. https://lexfridman.com/sponsors/ep496-sc Transcript: https://lexfridman.com/ffmpeg-transcript CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://x.com/FFmpeg https://ffmpeg.org/ https://www.videolan.org/ https://x.com/videolan https://jbkempf.com/ https://www.linkedin.com/in/jbkempf/ https://github.com/jbkempf https://x.com/kierank_ https://bit.ly/3OORhmC https://github.com/kierank SPONSORS: Larridin: Measure AI adoption in your business. https://larridin.com Blitzy: AI agent for large enterprise codebases. https://blitzy.com/lex BetterHelp: Online therapy and counseling. https://betterhelp.com/lex Fin: AI agent for customer service. https://fin.ai/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Perplexity: AI-powered answer engine. https://perplexity.ai/ OUTLINE: (00:00) – Introduction (03:00) – Sponsors, Comments, and Reflections (10:48) – Weirdest things VLC opens (15:12) – How video playback works (24:33) – Video codecs and containers (35:20) – FFmpeg explained (56:20) – Linus Torvalds (1:00:59) – Turning down millions to keep VLC ad-free (1:15:17) – FFmpeg & Google drama (1:34:31) – FFmpeg developers (1:41:08) – VLC and FFmpeg (1:45:42) – History of FFmpeg (1:48:59) – Reverse engineering codecs (2:02:14) – FFmpeg testing (2:06:21) – Assembly code (handwritten) (2:30:39) – Rust programming language (2:39:55) – FFmpeg and Libav fork (2:48:17) – Open source burnout (2:56:04) – x264 and internet video (3:09:20) – Video compression basics (3:16:17) – CIA and fake VLC (3:26:52) – Ultra low latency streaming (3:44:20) – AV2 codec and video patents (3:54:12) – VLC backdoors (4:04:27) – Video archiving (4:11:04) – Future of FFmpeg and VLC

Wed, 06 May 2026 22:06:47 +0000

#495 – Vikings, Ragnar, Berserkers, Valhalla & the Warriors of the Viking Age

Lars Brownworth is a historian, teacher, podcaster, and author specializing in Viking history, medieval Europe, and the Byzantine Empire. https://lexfridman.com/sponsors/ep495-sc Transcript: https://lexfridman.com/lars-brownworth-transcript CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://larsbrownworth.com/ https://www.amazon.com/Sea-Wolves-History-Vikings/dp/1909979120 https://amzn.to/4sHY0xw https://12byzantinerulers.com/ https://apple.co/4sgSxNi SPONSORS: Larridin: Measure AI adoption in your business. https://larridin.com BetterHelp: Online therapy and counseling. https://betterhelp.com/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Fin: AI agent for customer service. https://fin.ai/lex Shopify: Sell stuff online. https://shopify.com/lex Perplexity: AI-powered answer engine. https://perplexity.ai/ OUTLINE: PODCAST LINKS: https://lexfridman.com/podcast https://apple.co/2lwqZIr https://spoti.fi/2nEwCF8 https://lexfridman.com/feed/podcast/ https://www.youtube.com/playlist?list=PLrAXtmErZgOdP_8GztsuKi9nrraNbKKp4 https://www.youtube.com/lexclips

Thu, 09 Apr 2026 17:43:17 +0000

#494 – Jensen Huang: NVIDIA – The $4 Trillion Company & the AI Revolution

Jensen Huang is the co-founder and CEO of NVIDIA, the world’s most valuable company and the engine powering the AI computing revolution. https://lexfridman.com/sponsors/ep494-sc Transcript: https://lexfridman.com/jensen-huang-transcript CONTACT LEX: Feedback – give feedback to Lex: https://lexfridman.com/survey AMA – submit questions, videos or call-in: https://lexfridman.com/ama Hiring – join our team: https://lexfridman.com/hiring Other – other ways to get in touch: https://lexfridman.com/contact EPISODE LINKS: https://nvidia.com https://x.com/nvidia https://x.com/NVIDIAAI https://youtube.com/@nvidia https://www.instagram.com/nvidia/ https://www.linkedin.com/company/nvidia/ https://www.facebook.com/NVIDIA/ https://github.com/NVIDIA https://developer.nvidia.com/nemotron SPONSORS: Perplexity: AI-powered answer engine. https://perplexity.ai/ Shopify: Sell stuff online. https://shopify.com/lex LMNT: Zero-sugar electrolyte drink mix. https://drinkLMNT.com/lex Fin: AI agent for customer service. https://fin.ai/lex Quo: Phone system (calls, texts, contacts) for businesses. https://quo.com/lex OUTLINE: PODCAST LINKS: https://lexfridman.com/podcast https://apple.co/2lwqZIr https://spoti.fi/2nEwCF8 https://lexfridman.com/feed/podcast/ https://www.youtube.com/playlist?list=PLrAXtmErZgOdP_8GztsuKi9nrraNbKKp4 https://www.youtube.com/lexclips

Mon, 23 Mar 2026 16:28:42 +0000