a16z Podcast
Summary & Insights
What if the secret to spatial intelligence isn’t just generating a pretty picture, but predicting exactly what the world looks like from a new vantage point? This is the core hypothesis behind Atlas, a new world model from World Labs that seeks to do for physical space what next-token prediction did for language. By unifying 3D reconstruction and generative AI, Atlas allows users to turn a handful of sparse images—as few as three—into a fully navigable 3D environment. This effectively eliminates the need for the “dense” captures (hundreds of photos) or expensive green screens traditionally required for high-end visual effects, such as the famous “bullet time” sequences from The Matrix.
The technical breakthrough lies in moving beyond simple next-frame prediction. While most video models “guess” what happens next in a sequence, Atlas uses a “spatial context” that anchors every frame to a precise 3D camera pose. This allows the model to handle both reconstruction (exactly replicating a real room) and generation (imagining the parts of the room the camera never saw) simultaneously. The team describes this as a “scaling law” for spatial intelligence: as they increased the model size and compute, the model’s ability to maintain 3D consistency and geometric accuracy improved predictably.
Looking forward, the implications extend far beyond cinema and digital art into the realm of robotics and industrial design. By creating high-fidelity “real-to-sim” pipelines, Atlas can help train robots in virtual environments that are indistinguishable from the real world, allowing them to experience thousands of failure scenarios before ever touching a physical object. The ultimate goal is a “closed-loop” system where the AI doesn’t just see and render a world, but understands the physics and dynamics of that world well enough to plan complex actions within it.
Surprising Insights
- The 100x Efficiency Gain: Traditional 3D reconstruction requires “dense” data—hundreds of images to avoid holes in the model. Atlas can achieve similar results with “sparse” data (3–10 images), representing a massive reduction in capture effort.
- Generation as a Gap-Filler: In 3D reconstruction, any area not captured by a camera is normally a “hole.” Atlas uses generative AI not just for creativity, but as a functional tool to logically fill in those missing pixels based on spatial priors.
- “AI Completeness” of Viewpoints: The founders argue that “new view prediction” is a fundamental primitive of intelligence. Just as predicting the next word can solve complex logic puzzles, predicting the next viewpoint is the core mechanism biological evolution used to develop vision and movement.
- Dynamics through Exposure: Counterintuitively, the best way to create a perfectly static 3D reconstruction is to train the model on as much dynamic, moving data as possible, allowing the model to learn how to factor out motion.
Practical Takeaways
- Leverage Sparse Captures: For creators and architects, move away from the tedious process of taking hundreds of photos of a space. Experiment with a few high-quality, diverse angles and use world models to extrapolate the rest.
- Bridge Real-to-Sim for Robotics: If training physical agents, use generative world models to create “randomized” versions of a real environment. This allows you to test how a robot handles different object colors, sizes, or positions without needing to manually build those variations in the real world.
- Adopt “Stateful” Workflows: Shift from an “ephemeral” generative mindset (prompt $\rightarrow$ image $\rightarrow$ discard) to a “stateful” one. Build a persistent 3D scene as a base and use AI to modify specific elements or camera angles while keeping the rest of the environment consistent.
🛍️ Products & Resources Mentioned
- ⚡ DeviceiPhone — Mentioned as a tool to capture video from multiple angles to be used as input for the Atlas model to create bullet-time effects.View on Amazon →
Điều gì sẽ xảy ra nếu bí mật của trí tuệ không gian không chỉ đơn thuần là tạo ra một bức ảnh đẹp, mà là dự đoán chính xác thế giới trông như thế nào từ một góc nhìn mới? Đây chính là giả thuyết cốt lõi đằng sau Atlas, một mô hình thế giới mới từ World Labs, với mục tiêu thực hiện cho không gian vật lý những điều mà việc “dự đoán token tiếp theo” đã làm được cho ngôn ngữ. Bằng cách hợp nhất tái tạo 3D và AI tạo sinh, Atlas cho phép người dùng biến một vài hình ảnh thưa thớt — thậm chí chỉ cần ba tấm — thành một môi trường 3D có thể điều hướng hoàn toàn. Điều này loại bỏ hiệu quả nhu cầu về các bản chụp “dày đặc” (hàng trăm bức ảnh) hoặc những phông xanh đắt đỏ vốn thường được yêu cầu cho các hiệu ứng hình ảnh cao cấp, chẳng hạn như những phân cảnh “bullet time” nổi tiếng trong phim The Matrix.
Đột phá kỹ thuật nằm ở việc vượt xa khả năng dự đoán khung hình tiếp theo đơn thuần. Trong khi hầu hết các mô hình video “đoán” điều gì xảy ra tiếp theo trong một chuỗi, Atlas sử dụng một “ngữ cảnh không gian” giúp neo giữ mọi khung hình vào một vị trí camera 3D chính xác. Điều này cho phép mô hình xử lý đồng thời cả việc tái tạo (sao chép chính xác một căn phòng thực) và tạo sinh (tưởng tượng ra những phần của căn phòng mà camera chưa bao giờ nhìn thấy). Nhóm phát triển mô tả đây là một “quy luật tỷ lệ” (scaling law) cho trí tuệ không gian: khi họ tăng kích thước mô hình và năng lực tính toán, khả năng duy trì tính nhất quán 3D và độ chính xác hình học của mô hình được cải thiện một cách có hệ thống.
Nhìn về tương lai, những tác động của công nghệ này mở rộng xa hơn cả điện ảnh và nghệ thuật kỹ thuật số, tiến vào lĩnh vực robot và thiết kế công nghiệp. Bằng cách tạo ra các quy trình “real-to-sim” (từ thực tế sang mô phỏng) với độ trung thực cao, Atlas có thể giúp huấn luyện robot trong các môi trường ảo không thể phân biệt được với thế giới thực, cho phép chúng trải nghiệm hàng ngàn kịch bản thất bại trước khi thực sự chạm vào một vật thể vật lý. Mục tiêu cuối cùng là một hệ thống “vòng lặp kín”, nơi AI không chỉ nhìn và kết xuất (render) một thế giới, mà còn hiểu rõ vật lý và động lực học của thế giới đó để lập kế hoạch cho các hành động phức tạp bên trong nó.
Những hiểu biết bất ngờ
- Hiệu quả tăng gấp 100 lần: Tái tạo 3D truyền thống đòi hỏi dữ liệu “dày đặc” — hàng trăm hình ảnh để tránh các “lỗ hổng” trong mô hình. Atlas có thể đạt được kết quả tương tự với dữ liệu “thưa” (3–10 hình ảnh), giúp giảm thiểu đáng kể công sức thu thập dữ liệu.
- Tạo sinh đóng vai trò là công cụ lấp đầy: Trong tái tạo 3D, bất kỳ khu vực nào không được camera chụp lại thường sẽ là một “lỗ hổng”. Atlas sử dụng AI tạo sinh không chỉ để sáng tạo, mà như một công cụ chức năng để lấp đầy các pixel còn thiếu đó một cách logic dựa trên các tiền đề không gian.
- “Tính hoàn thiện AI” của các góc nhìn: Những người sáng lập lập luận rằng “dự đoán góc nhìn mới” là một thành phần cơ bản của trí thông minh. Giống như việc dự đoán từ tiếp theo có thể giải quyết các câu đố logic phức tạp, việc dự đoán góc nhìn tiếp theo chính là cơ chế cốt lõi mà tiến hóa sinh học đã sử dụng để phát triển thị giác và vận động.
- Động lực học thông qua sự tiếp xúc: Ngược với suy nghĩ thông thường, cách tốt nhất để tạo ra một bản tái tạo 3D tĩnh hoàn hảo là huấn luyện mô hình trên càng nhiều dữ liệu động, di chuyển càng tốt, giúp mô hình học được cách tách biệt các yếu tố chuyển động ra khỏi cấu trúc tĩnh.
Bài học thực tiễn
- Tận dụng thu thập dữ liệu thưa: Đối với các nhà sáng tạo và kiến trúc sư, hãy từ bỏ quy trình tẻ nhạt là chụp hàng trăm bức ảnh của một không gian. Hãy thử nghiệm với một vài góc chụp chất lượng cao, đa dạng và sử dụng các mô hình thế giới để suy diễn phần còn lại.
- Kết nối Thực-sang-Mô phỏng cho Robot: Nếu huấn luyện các tác tử vật lý, hãy sử dụng các mô hình thế giới tạo sinh để tạo ra các phiên bản “ngẫu nhiên hóa” của một môi trường thực. Điều này cho phép bạn kiểm tra cách robot xử lý các màu sắc, kích thước hoặc vị trí vật thể khác nhau mà không cần phải xây dựng thủ công những biến thể đó trong thế giới thực.
- Áp dụng quy trình làm việc “có trạng thái” (Stateful): Chuyển từ tư duy tạo sinh “tạm thời” (nhập prompt $\rightarrow$ tạo ảnh $\rightarrow$ xóa bỏ) sang tư duy “có trạng thái”. Hãy xây dựng một cảnh 3D bền vững làm cơ sở và sử dụng AI để sửa đổi các yếu tố cụ thể hoặc góc camera trong khi vẫn giữ cho toàn bộ môi trường nhất quán.
如果空間智能(spatial intelligence)的祕訣不在於僅僅生成一張漂亮的圖片,而是在於能精確預測從一個新視角看過去的世界模樣會是如何?這正是 World Labs 推出的新世界模型 Atlas 的核心假設。Atlas 旨在將「下一標記預測」(next-token prediction)在語言領域所帶來的革命,應用於物理空間。透過將 3D 重建與生成式 AI 統一,Atlas 讓使用者僅需少數幾張稀疏的圖像(最少三張),即可將其轉化為一個完全可導航的 3D 環境。這有效地消除了對「密集」採集(數百張照片)或昂貴綠幕的需求,而這些傳統上是高端視覺特效(例如《駭客任務》中著名的「子彈時間」序列)所必需的。
這項技術突破在於超越了簡單的「下一幀預測」。大多數視訊模型是在「猜測」序列中接下來會發生什麼,而 Atlas 則使用一種「空間上下文」(spatial context),將每一幀錨定在精確的 3D 攝影機姿勢(camera pose)上。這使得模型能同時處理「重建」(精確複製一個真實房間)與「生成」(想像攝影機未曾拍攝到的房間部分)。團隊將此描述為空間智能的「縮放定律」(scaling law):隨著模型規模和算力的增加,模型維持 3D 一致性和幾何準確性的能力會隨之可預見地提升。
展望未來,其影響將遠超電影和數位藝術,延伸至機器人技術和工業設計領域。透過建立高保真的「實實轉模擬」(real-to-sim)管線,Atlas 可以協助在與現實世界無異的虛擬環境中訓練機器人,讓它們在接觸物理實體之前,就能體驗數千種失敗情境。最終目標是建立一個「閉環」系統,使 AI 不僅能看到並渲染世界,還能充分理解該世界的物理特性與動力學,從而在其中規劃複雜的行動。
驚人洞察
- 100 倍的效率提升: 傳統的 3D 重建需要「密集」數據——數百張圖像以避免模型出現漏洞。Atlas 僅需「稀疏」數據(3-10 張圖像)即可達到類似效果,大幅減少了採集工作量。
- 將「生成」作為填補空白的工具: 在 3D 重建中,攝影機未拍攝到的區域通常會形成「漏洞」。Atlas 使用生成式 AI 不僅是為了創意,而是將其作為一種功能性工具,根據空間先驗(spatial priors)邏輯地填補缺失的像素。
- 視角的「AI 完備性」: 創始人認為「新視角預測」是智能的一項基本原語(primitive)。正如預測下一個詞可以解決複雜的邏輯謎題,預測下一個視角則是生物演化用以發展視覺和運動的核心機制。
- 透過暴露學習動力學: 與直覺相反,創建完美靜態 3D 重建的最佳方法,是讓模型在盡可能多且具動態的移動數據上進行訓練,使模型學習如何將運動因素分離出來。
實務啟發
- 利用稀疏採集: 對於創作者和建築師而言,可以擺脫對一個空間拍攝數百張照片的繁瑣過程。嘗試拍攝少數幾張高品質且角度多樣的照片,並利用世界模型來推演其餘部分。
- 為機器人搭建「實轉模」橋樑: 在訓練物理智能體時,利用生成式世界模型創建現實環境的「隨機化」版本。這讓您無需在現實世界中手動構建變體,即可測試機器人如何處理不同的物體顏色、尺寸或位置。
- 採納「有狀態」的工作流: 將思維從「瞬時」的生成模式(提示詞 $\rightarrow$ 圖像 $\rightarrow$ 捨棄)轉向「有狀態」(stateful)模式。建立一個持久的 3D 場景作為基底,並利用 AI 在保持環境一致性的同時,修改特定元素或攝影機角度。
Et si le secret de l’intelligence spatiale ne consistait pas seulement à générer une jolie image, mais à prédire exactement à quoi ressemble le monde depuis un nouveau point de vue ? C’est l’hypothèse centrale d’Atlas, un nouveau modèle de monde conçu par World Labs, qui cherche à faire pour l’espace physique ce que la prédiction du prochain jeton (*next-token prediction*) a fait pour le langage. En unifiant la reconstruction 3D et l’IA générative, Atlas permet aux utilisateurs de transformer une poignée d’images éparses — seulement trois, par exemple — en un environnement 3D entièrement navigable. Cela élimine concrètement le besoin de captures « denses » (des centaines de photos) ou de fonds verts coûteux, traditionnellement requis pour les effets visuels de haut niveau, comme les célèbres séquences en « bullet time » de Matrix.
La percée technique réside dans le dépassement de la simple prédiction de l’image suivante. Alors que la plupart des modèles vidéo « devinent » ce qui arrive ensuite dans une séquence, Atlas utilise un « contexte spatial » qui ancre chaque image à une pose de caméra 3D précise. Cela permet au modèle de gérer simultanément la reconstruction (répliquer exactement une pièce réelle) et la génération (imaginer les parties de la pièce que la caméra n’a jamais vues). L’équipe décrit cela comme une « loi d’échelle » (*scaling law*) pour l’intelligence spatiale : à mesure qu’ils augmentaient la taille du modèle et la puissance de calcul, la capacité du modèle à maintenir la cohérence 3D et la précision géométrique s’améliorait de manière prévisible.
À l’avenir, les implications s’étendent bien au-delà du cinéma et de l’art numérique pour toucher la robotique et le design industriel. En créant des pipelines « real-to-sim » (du réel vers la simulation) de haute fidélité, Atlas peut aider à entraîner des robots dans des environnements virtuels indiscernables du monde réel, leur permettant de vivre des milliers de scénarios d’échec avant même de toucher un objet physique. L’objectif ultime est un système en « boucle fermée » où l’IA ne se contente pas de voir et de rendre un monde, mais en comprend la physique et la dynamique suffisamment bien pour planifier des actions complexes à l’intérieur.
Perspectives Surprenantes
- Un gain d’efficacité de 100x : La reconstruction 3D traditionnelle nécessite des données « denses » — des centaines d’images pour éviter les trous dans le modèle. Atlas peut obtenir des résultats similaires avec des données « éparses » (3 à 10 images), ce qui représente une réduction massive de l’effort de capture.
- La génération comme outil de comblement : Dans la reconstruction 3D, toute zone non capturée par une caméra est normalement un « trou ». Atlas utilise l’IA générative non pas seulement pour la créativité, mais comme un outil fonctionnel pour combler logiquement ces pixels manquants en s’appuyant sur des a priori spatiaux.
- La « complétude IA » des points de vue : Les fondateurs soutiennent que la « prédiction de nouvelle vue » est une primitive fondamentale de l’intelligence. Tout comme la prédiction du mot suivant peut résoudre des puzzles logiques complexes, la prédiction du prochain point de vue est le mécanisme central que l’évolution biologique a utilisé pour développer la vision et le mouvement.
- La dynamique par l’exposition : De manière contre-intuitive, la meilleure façon de créer une reconstruction 3D parfaitement statique est d’entraîner le modèle sur autant de données dynamiques et mouvantes que possible, permettant ainsi au modèle d’apprendre à isoler et éliminer le mouvement.
Conseils Pratiques
- Exploitez les captures éparses : Pour les créateurs et les architectes, délaissez le processus fastidieux consistant à prendre des centaines de photos d’un espace. Expérimentez avec quelques angles diversifiés et de haute qualité, puis utilisez des modèles de monde pour extrapoler le reste.
- Faites le pont Real-to-Sim pour la robotique : Si vous entraînez des agents physiques, utilisez des modèles de monde génératifs pour créer des versions « randomisées » d’un environnement réel. Cela vous permet de tester comment un robot gère différentes couleurs, tailles ou positions d’objets sans avoir à construire manuellement ces variations dans le monde réel.
- Adoptez des flux de travail « stateful » (à état) : Passez d’un état d’esprit génératif « éphémère » (prompt $\rightarrow$ image $\rightarrow$ suppression) à un état d’esprit « stateful ». Construisez une scène 3D persistante comme base et utilisez l’IA pour modifier des éléments spécifiques ou des angles de caméra tout en gardant le reste de l’environnement cohérent.
Was wäre, wenn das Geheimnis räumlicher Intelligenz nicht darin bestünde, einfach nur ein hübsches Bild zu erzeugen, sondern exakt vorherzusagen, wie die Welt von einem neuen Standpunkt aus aussieht? Dies ist die Kernhypothese hinter Atlas, einem neuen Weltmodell von World Labs, das für den physischen Raum das erreichen will, was die Vorhersage des nächsten Tokens (Next-Token Prediction) für die Sprache geleistet hat. Durch die Vereinigung von 3D-Rekonstruktion und generativer KI ermöglicht Atlas es Nutzern, eine Handvoll spärlicher Bilder – bereits ab drei Aufnahmen – in eine vollständig navigierbare 3D-Umgebung zu verwandeln. Dies macht „dichte“ Aufnahmen (Hunderte von Fotos) oder teure Greenscreens, die traditionell für High-End-Visuelleffekte wie die berühmten „Bullet-Time“-Sequenzen aus The Matrix benötigt wurden, effektiv überflüssig.
Der technische Durchbruch liegt darin, über die einfache Vorhersage des nächsten Frames hinauszugehen. Während die meisten Videomodelle „raten“, was als Nächstes in einer Sequenz passiert, nutzt Atlas einen „räumlichen Kontext“, der jeden Frame an eine präzise 3D-Kameraposition bindet. Dies erlaubt es dem Modell, gleichzeitig sowohl die Rekonstruktion (die exakte Nachbildung eines realen Raums) als auch die Generierung (das Imaginieren von Raumteilen, die die Kamera nie gesehen hat) zu bewältigen. Das Team beschreibt dies als ein „Scaling Law“ (Skalierungsgesetz) für räumliche Intelligenz: Mit zunehmender Modellgröße und Rechenleistung verbesserte sich die Fähigkeit des Modells, 3D-Konsistenz und geometrische Genauigkeit beizubehalten, in vorhersehbarer Weise.
Mit Blick auf die Zukunft reichen die Auswirkungen weit über das Kino und die digitale Kunst hinaus bis in den Bereich der Robotik und des Industriedesigns. Durch die Schaffung hochpräziser „Real-to-Sim“-Pipelines kann Atlas dabei helfen, Roboter in virtuellen Umgebungen zu trainieren, die von der realen Welt nicht zu unterscheiden sind. So können sie Tausende von Fehlerszenarien durchspielen, bevor sie jemals ein physisches Objekt berühren. Das ultimative Ziel ist ein „Closed-Loop“-System, in dem die KI eine Welt nicht nur sieht und rendert, sondern die Physik und Dynamik dieser Welt gut genug versteht, um komplexe Aktionen darin zu planen.
Überraschende Erkenntnisse
- 100-facher Effizienzgewinn: Traditionelle 3D-Rekonstruktionen benötigen „dichte“ Daten – Hunderte von Bildern, um Lücken im Modell zu vermeiden. Atlas kann ähnliche Ergebnisse mit „spärlichen“ Daten (3–10 Bilder) erzielen, was eine massive Reduzierung des Aufwands bei der Aufnahme bedeutet.
- Generierung als Lückenfüller: Bei der 3D-Rekonstruktion ist jeder Bereich, der nicht von einer Kamera erfasst wurde, normalerweise ein „Loch“. Atlas nutzt generative KI nicht nur für die Kreativität, sondern als funktionales Werkzeug, um diese fehlenden Pixel basierend auf räumlichen Vorwissen (Spatial Priors) logisch aufzufüllen.
- „KI-Vollständigkeit“ von Blickwinkeln: Die Gründer argumentieren, dass die „Vorhersage neuer Ansichten“ ein grundlegendes Primitiv der Intelligenz ist. So wie die Vorhersage des nächsten Wortes komplexe Logikrätsel lösen kann, ist die Vorhersage des nächsten Blickwinkels der Kernmechanismus, den die biologische Evolution genutzt hat, um Sehen und Bewegung zu entwickeln.
- Dynamik durch Exposition: Paradoxerweise ist der beste Weg, eine perfekt statische 3D-Rekonstruktion zu erstellen, das Modell mit so vielen dynamischen, bewegten Daten wie möglich zu trainieren, damit das Modell lernt, Bewegungen herauszufiltern.
Praktische Ableitungen
- Spärliche Aufnahmen nutzen: Kreative und Architekten sollten sich vom mühsamen Prozess, Hunderte von Fotos eines Raumes zu machen, lösen. Experimentieren Sie mit wenigen, qualitativ hochwertigen und vielfältigen Winkeln und nutzen Sie Weltmodelle, um den Rest zu extrapolieren.
- Real-to-Sim für Robotik überbrücken: Wenn Sie physische Agenten trainieren, nutzen Sie generative Weltmodelle, um „randomisierte“ Versionen einer realen Umgebung zu erstellen. Dies ermöglicht es Ihnen zu testen, wie ein Roboter mit verschiedenen Objektfarben, Größen oder Positionen umgeht, ohne diese Variationen manuell in der realen Welt aufbauen zu müssen.
- „Stateful“ Workflows einführen: Wechseln Sie von einer „ephemeren“ generativen Denkweise (Prompt $\rightarrow$ Bild $\rightarrow$ Verwerfen) zu einer „stateful“ (zustandsbehafteten) Herangehensweise. Erstellen Sie eine persistente 3D-Szene als Basis und nutzen Sie die KI, um spezifische Elemente oder Kamerawinkel zu ändern, während die restliche Umgebung konsistent bleibt.
World Labs co-founders Fei-Fei Li, Justin Johnson, and Ben Mildenhall join a16z General Partner Martin Casado to discuss Atlas, their latest world model, and what it reveals about the pursuit of spatial intelligence.
At the center of Atlas is what the team calls “new view prediction”: given images or views of a scene, the model predicts what that environment should look like from a different position in space and time. This brings generation and 3D reconstruction into the same model, and raises a broader question about whether predicting views could become a useful primitive for understanding the physical world.
They discuss the technical bets behind the model, what it can and can’t yet capture, and the importance of dynamics, editability, and simulation as world models develop. The conversation also explores applications in creative work, architecture, and robotics, where Fei-Fei argues that one of today’s biggest constraints is access to real-world training data.
Resources:
Follow Fei-Fei Li on X:https://x.com/drfeifei
Follow Justin Johnson on X: https://x.com/jcjohnss
Follow Ben Mildenhall on X: https://x.com/BenMildenhall
Follow Martin Casado on X:https://x.com/martin_casado
Learn more about Atlas:https://www.worldlabs.ai/blog/atlas
Stay Updated:
Find a16z on YouTube: YouTube
Find a16z on X
Find a16z on LinkedIn
Listen to the a16z Show on Spotify
Listen to the a16z Show on Apple Podcasts
Follow our host: https://twitter.com/eriktorenberg
Please note that the content here is for informational purposes only; should NOT be taken as legal, business, tax, or investment advice or be used to evaluate any investment or security; and is not directed at any investors or potential investors in any a16z fund. a16z and its affiliates may maintain investments in the companies discussed. For more details please see a16z.com/disclosures.
Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.
-
The Deepfake Dilemma: The Technology, Policy, and Economy
Deepfakes—AI-generated fake videos and voices—have become a widespread concern across politics, social media, and more. As they become easier to create, the threat grows. But so do the tools to detect them. In this episode,…
-
A Big Week in Tech: NotebookLM, OpenAI’s Speech API, & Custom Audio
Last week was another big week in technology. Google’s NotebookLM introduced its Audio Overview feature, enabling users to create customizable podcasts in over 35 languages. OpenAI followed with their real-time speech-to-speech API, making voice integration…
-
From Swipe to Scale: How Tinder Became #1
In 1995, just 2% of couples met online. Today, that number has surged to over 50%, making online dating the top way couples connect. In this episode, a16z General Partner Andrew Chen chats with Tinder…
-
Human Data is Key to AI: Alex Wang from Scale AI
What if the key to unlocking AI’s full potential lies not just in algorithms or compute, but in data? In this episode, a16z General Partner David George sits down with Alex Wang, founder and CEO…
-
The Frontier of Spatial Intelligence with Fei-Fei Li
Fei-Fei Li and Justin Johnson are pioneers in AI. While the world has only recently witnessed a surge in consumer AI, our guests have long been laying the groundwork for innovations that are transforming industries…
-
Apple’s Big Reveals, OpenAI’s Multi-Step Models, and Firefly Does Video
This week in consumer tech: Apple’s big reveals, OpenAI’s multi-step reasoning, and Adobe Firefly’s video model. Olivia Moore and Justine Moore, Partners on the a16z Consumer team, break down the latest announcements and how these…
-
Grand Challenges in Healthcare AI
Vijay Pande, founding general partner, and Julie Yoo, general partner at a16z Bio + Health, come together to discuss the grand challenges facing healthcare AI today. The talk through the implications of AI integration in…
-
Governing democracy, the internet, and boardrooms
with @NoahRFeldman, @ahall_research, @rhhackett Welcome to web3 with a16z. I’m Robert Hackett and today we have a special episode about governance in many forms — from nation states to corporate boards to internet services and…
-
It’s Time to Build in Healthcare
Half of prescribed medications are never taken, and 88% of Americans are metabolically unhealthy. Despite spending 20% of our GDP on healthcare—twice that of any other developed nation—our outcomes still lag behind. In this episode,…
-
Latin America: A Tech Powerhouse?
Latin America is emerging as a tech powerhouse, but it’s not a one-size-fits-all market. In this episode, we explore why what works in Argentina won’t necessarily fly in Brazil or Mexico, and how companies are…
