OpenAI 自研芯片 Jalapeño 投产
OpenAI 与博通合作推出首款自研 AI 芯片 Jalapeño,专为 LLM 推理负载设计。这标志着 OpenAI 从应用层向底层基础设施垂直整合,将直接影响其 API 成本和产能扩张节奏。
OpenAI 选择博通而非台积电直签,说明其走的是 Google TPU 的 fabless 路线而非全栈自研,这降低了风险但也意味着产能仍受制于人。Jalapeño 明确标注为 inference-optimized,暗示训练端仍依赖 NVIDIA,短期内不会动摇 H100 的统治地位。
Karpathy 同时点评 Claude 的 inline 集成模式,两个信号叠加来看:头部公司正在从模型竞赛转向基础设施和交互范式的系统性卡位。做 AI 中间件和开发者工具的创业者需要警惕,平台层的缝隙正在快速收窄。
We’ve designed and built our first AI chip: Jalapeño. Designed from the ground up by OpenAI and brought to production with @Broadcom, Jalapeño is purpose-built for the LLM workloads powering ChatGPT, Codex, the API, and future agentic products. Chips are foundational to the AI economy. Building our own expands our full-stack platform from products to models to infrastructure, and will help us scale intelligence, serve more people, and expand access to AI.
查看原文 →This is a new paradigm for interacting with Claude that is significantly more "inline" with all the other human activity org-wide. Once you do all of the under the hood engineering work to make this "just work" (e.g. across tools, integrations, compute environments, memory, security, etc.), Claude basically joins the team in a seamless way - you can talk to it as you would talk to a person and it can help with a very large variety of workloads. Imo this is the 3rd major redesign of LLM UIUX. The first paradigm was that the LLM is a website you go to, the second was that it is an app you download to your computer. This third one is that it is a self-contained, persistent, asynchronous entity with org-wide tools and context, working alongside teams of humans. It really takes a while to wrap your head around it, but it works and it is awesome.
查看原文 →We have a new version of GPT-5.5 Instant for you, and it's much more fun to talk to. Our most-used model is now better at understanding the intent behind a question and adapting its response accordingly. It also handles complex constraints more reliably and makes shopping and local recommendations more useful and cohesive. Rolling out today to paid users, tomorrow to free users.
查看原文 →RT @PabloTorre: Watching Kyra and her family fight brain cancer — glioblastoma — over the last year has been inspiring, and heartrending, b…
查看原文 →Work at OpenAI is being transformed by agents, in every department. Across our entire company, people are using Codex to do work that is more complex, longer-running, and increasingly cross-functional. Our internal usage offers an early look at how agentic tools may reshape work as they become more capable and broadly available.
查看原文 →Are we honestly calling a Slack bot “a new paradigm”? Have we lost all decency?
查看原文 →RT @CleGuardians: Kahlil Watson's first Major League homer is a game-tying knock to right field! #GuardsBall | #VoteGuards https://t.co/d0…
查看原文 →RT @ihtesham2005: Elon Musk built one of the largest AI compute clusters on earth. Yann LeCun just explained why xAI now rents it out to ri…
查看原文 →RT @DeryaTR_: I’m very excited about this article from @OpenAI on my attempt to use GPT-5 Pro to understand the results of an experiment we…
查看原文 →RT @barriere_dr: C’est elle qui a jeté Fauci en pâture relayé ensuite par RFK et Musk… Je vous le dis Le RN au pouvoir ça sera ça. On peut…
查看原文 →RT @EricTopol: "I do not believe that AI is likely to cure cancer anytime soon"—@2plus2makes5 https://t.co/iU4ysK2xzl In fact, AI has alrea…
查看原文 →RT @DegenerateTBone: Giant baby. Can't answer a few questions from the media about rainbow-colored hats, or take any accountability for his…
查看原文 →RT @jparkerholder: Huge congratulations to the Project Genie team on taking home the Cannes Lions Grand Prix for AI Craft! 🎉 Thank you @Can…
查看原文 →If you are a software engineer still using an IDE or a CLI to build software, you ain’t gonna make it. If you are writing prompts, or using skills, you ain’t gonna make it. Slack is the paradigm. That’s the only way forward. /s
查看原文 →What happens to society’s collective intelligence after two or three years of everyone outsourcing their thinking to a chatbot? How many orders of magnitude dumber are we going to get? Can we start talking about that yet, or nah?
查看原文 →We should build a church for people who open-source their code so everyone can learn from it. Here is the complete source code of a RAG assistant to navigate airline policies. You get the complete source code and video from @lenadroid, walking you through everything she did (I'm linking to the video in the first comment below). The fact that you can watch every engineering decision that Lena made when building this app is pure gold. A few things you'll pick up from this: • It uses LangChain for the retrieval pipeline • It uses LangGraph for conversation state • It stores embeddings in Postgres with pgvector • It indexes documents to ground answers in the source text • It uses Terraform to stand up the infrastructure I'm linking to the video walkthrough and the source code below.
查看原文 →The architecture of this new world model is one of the most interesting things I've seen lately: Let me first explain how most world models work: They predict and render one frame at a time. If you are navigating in one of these worlds, and you look left, the model draws whatever looks right in the moment. Every time you change your viewpoint, the model has to imagine what should be there again, so it's very common for these models to "forget" what's in the world. For example, if you put a toy on the table, look away, then look back, the toy might not be there anymore. Tripo AI is releasing its Project Eden model, which works very differently: The model builds the world first, and then renders it based on that map. That map holds the real state of the world: the geometry, every object, where things are, what's already happened. The picture you see on screen gets generated from the map. This architecture flips the whole thing. Now, you get the following: 1. The world stops forgetting. Leave, come back, and the toy is still on the table because it lives in the map, not in the last frame you saw. 2. You can edit the world, and those changes persist for anyone who enters later. 3. Multiple people and AI agents can coexist in the world and see it from different perspectives. This is early research, but it's looking really promising. They just raised nearly $200M across two rounds to build it out. Tripo will be at SIGGRAPH 2026 (July 19–23, Los Angeles Convention Center). If you work in 3D, embodied AI, simulation, or anything spatial, go connect with them there.
查看原文 →RT @ZcohenCNN: CNN reported in April that US intelligence assessed that roughly half of Iran’s missile launchers had survived US strikes.…
查看原文 →RT @randall_balestr: It's a bird, it's a plane, it's a JEPA! Congrats on that great work that brought SIGReg and JEPAs to the sky--in the r…
查看原文 →RT @PratyakshRao5: What should a world model for agile quadrotor control actually provide? 📄 Arxiv: https://t.co/fMzmm7GODZ 🌐 Project: ht…
查看原文 →RT @KenRoth: More than 4.7 million people in the US have lost their Supplemental Nutrition Assistance Program benefits, also known as food…
查看原文 →