Kevin Di's Hatch

Kevin Di
Latest
Popular
Glasp
Hatch
为什么大模型的性能瓶颈,最后都会变成调度问题
glasp.co/hatch

为什么大模型的性能瓶颈,最后都会变成调度问题

Hatched on May 1, 2026 · 5 views

Glasp
Hatch
GPU泡沫退潮之后,真正值钱的不是算力,而是可编程的稀缺性
glasp.co/hatch

GPU泡沫退潮之后,真正值钱的不是算力,而是可编程的稀缺性

Hatched on Apr 30, 2026 · 10 views

Glasp
Hatch
为什么最贵的系统故障,往往卡在最不起眼的转换环节
glasp.co/hatch

为什么最贵的系统故障,往往卡在最不起眼的转换环节

Hatched on Apr 29, 2026 · 6 views

Glasp
Hatch
Why the Fastest LLMs Win by Remembering Less
glasp.co/hatch

Why the Fastest LLMs Win by Remembering Less

Hatched on Apr 28, 2026 · 4 views

Glasp
Hatch
算力不是机器,是一张负债表:为什么真正值钱的是可持续的计算能力
glasp.co/hatch

算力不是机器,是一张负债表:为什么真正值钱的是可持续的计算能力

Hatched on Apr 27, 2026 · 8 views

Glasp
Hatch
从随机采样到万卡网络:大模型真正稀缺的不是算力,而是秩序
glasp.co/hatch

从随机采样到万卡网络:大模型真正稀缺的不是算力,而是秩序

Hatched on Apr 26, 2026 · 9 views

Glasp
Hatch
真正决定AI芯片胜负的,不是算力,而是“路由”
glasp.co/hatch

真正决定AI芯片胜负的,不是算力,而是“路由”

Hatched on Apr 25, 2026 · 10 views

Glasp
Hatch
真正的智能,不在算力峰值,而在把记忆放在离思考最近的地方
glasp.co/hatch

真正的智能,不在算力峰值,而在把记忆放在离思考最近的地方

Hatched on Apr 24, 2026 · 11 views

Glasp
Hatch
AI 训练的真正瓶颈,不是算力,而是每一次“搬运”
glasp.co/hatch

AI 训练的真正瓶颈,不是算力,而是每一次“搬运”

Hatched on Apr 23, 2026 · 8 views

Glasp
Hatch
为什么最好的算力,不是“更快”,而是“更贴近真实任务的尺寸”
glasp.co/hatch

为什么最好的算力,不是“更快”,而是“更贴近真实任务的尺寸”

Hatched on Apr 22, 2026 · 11 views

Glasp
Hatch
当并行不再只追求更快,而是先找到它真正该服务的任务
glasp.co/hatch

当并行不再只追求更快,而是先找到它真正该服务的任务

Hatched on Apr 21, 2026 · 5 views

Glasp
Hatch
AI 竞赛的真正瓶颈,不是算力,而是精度该放在哪里
glasp.co/hatch

AI 竞赛的真正瓶颈,不是算力,而是精度该放在哪里

Hatched on Apr 20, 2026 · 6 views

Glasp
Hatch
算力战争真正打的不是芯片,而是控制权
glasp.co/hatch

算力战争真正打的不是芯片,而是控制权

Hatched on Apr 19, 2026 · 6 views

Glasp
Hatch
当算力从稀缺变成拥堵,真正的瓶颈就不再是 GPU
glasp.co/hatch

当算力从稀缺变成拥堵,真正的瓶颈就不再是 GPU

Hatched on Apr 18, 2026 · 9 views

Glasp
Hatch
当云平台遇见高速网络:真正决定AI集群上限的不是算力,而是编排
glasp.co/hatch

当云平台遇见高速网络:真正决定AI集群上限的不是算力,而是编排

Hatched on Apr 17, 2026 · 8 views

Glasp
Hatch
当记忆成为瓶颈: 把序列拆成流水的时机与方法
glasp.co/hatch

当记忆成为瓶颈: 把序列拆成流水的时机与方法

Hatched on Apr 16, 2026 · 8 views

Glasp
Hatch
当显存追不上参数:把网络当作内存,重构大模型的计算拓扑
glasp.co/hatch

当显存追不上参数:把网络当作内存,重构大模型的计算拓扑

Hatched on Apr 15, 2026 · 8 views

Glasp
Hatch
当算法遇见场景: 找到机器学习工程的真正适配点
glasp.co/hatch

当算法遇见场景: 找到机器学习工程的真正适配点

Hatched on Apr 14, 2026 · 8 views

Glasp
Hatch
### Unlocking the Potential of Token-Level Pipeline Parallelism and NVLink Technology
glasp.co/hatch

### Unlocking the Potential of Token-Level Pipeline Parallelism and NVLink Technology

Hatched on Apr 13, 2026 · 5 views

Glasp
Hatch
# Accelerating the Development of Domestic AI Chips: Bridging the Gap with Advanced Models
glasp.co/hatch

# Accelerating the Development of Domestic AI Chips: Bridging the Gap with Advanced Models

Hatched on Apr 12, 2026 · 5 views

Glasp
Hatch
# The Future of Chip Technology: Navigating the Landscape of UCIE and NVLink
glasp.co/hatch

# The Future of Chip Technology: Navigating the Landscape of UCIE and NVLink

Hatched on Apr 11, 2026 · 6 views

Glasp
Hatch
Understanding Dynamic Inference and GPU Architecture: A Deep Dive into AI Technologies
glasp.co/hatch

Understanding Dynamic Inference and GPU Architecture: A Deep Dive into AI Technologies

Hatched on Apr 10, 2026 · 5 views

Glasp
Hatch
# The Race for High Bandwidth Memory: Transforming AI and Computing Performance
glasp.co/hatch

# The Race for High Bandwidth Memory: Transforming AI and Computing Performance

Hatched on Apr 9, 2026 · 7 views

Glasp
Hatch
### 未来内存和计算架构的变革:MRAM与CXL和GB200的潜力
glasp.co/hatch

### 未来内存和计算架构的变革:MRAM与CXL和GB200的潜力

Hatched on Apr 8, 2026 · 8 views

Glasp
Hatch
### Accelerating Transformer Models: Insights from Sparse Attention and NVLink Technologies
glasp.co/hatch

### Accelerating Transformer Models: Insights from Sparse Attention and NVLink Technologies

Hatched on Apr 7, 2026 · 5 views

Glasp
Hatch
The Rise of Next-Generation AI Chips: A New Era in Large Language Model Inference
glasp.co/hatch

The Rise of Next-Generation AI Chips: A New Era in Large Language Model Inference

Hatched on Apr 6, 2026 · 5 views

Glasp
Hatch
### The Evolution of Large Language Models: A Comparative Analysis and Optimization Strategies
glasp.co/hatch

### The Evolution of Large Language Models: A Comparative Analysis and Optimization Strategies

Hatched on Apr 5, 2026 · 9 views

Glasp
Hatch
Cerebras vs. NVIDIA: The Race for AI Chip Supremacy
glasp.co/hatch

Cerebras vs. NVIDIA: The Race for AI Chip Supremacy

Hatched on Apr 4, 2026 · 11 views

Glasp
Hatch
Navigating the Challenges of Blackwell: Insights into AIGC Network Optimization
glasp.co/hatch

Navigating the Challenges of Blackwell: Insights into AIGC Network Optimization

Hatched on Apr 3, 2026 · 5 views

Glasp
Hatch
### The Future of AI Processing: Unleashing the Power of Innovative Chip Technologies
glasp.co/hatch

### The Future of AI Processing: Unleashing the Power of Innovative Chip Technologies

Hatched on Apr 2, 2026 · 7 views

Glasp
Hatch
# Optimizing Performance in Machine Learning Inference: Insights into KVCache and Fully-Connected Layers
glasp.co/hatch

# Optimizing Performance in Machine Learning Inference: Insights into KVCache and Fully-Connected Layers

Hatched on Apr 1, 2026 · 4 views

Glasp
Hatch
### The Future of Computing: Innovations in CXL and AI Networking Solutions
glasp.co/hatch

### The Future of Computing: Innovations in CXL and AI Networking Solutions

Hatched on Mar 31, 2026 · 3 views

Glasp
Hatch
# Revolutionizing AI Hardware: The Future of Open Design and Unified Memory Architecture
glasp.co/hatch

# Revolutionizing AI Hardware: The Future of Open Design and Unified Memory Architecture

Hatched on Mar 30, 2026 · 6 views

Glasp
Hatch
### The Evolution of AI Chips: Innovations and Implications for the Future
glasp.co/hatch

### The Evolution of AI Chips: Innovations and Implications for the Future

Hatched on Mar 29, 2026 · 5 views

Glasp
Hatch
Pioneering the Future of AI: Insights from Chip Development and Performance Optimization
glasp.co/hatch

Pioneering the Future of AI: Insights from Chip Development and Performance Optimization

Hatched on Mar 28, 2026 · 6 views

Glasp
Hatch
### The Evolution of Chiplet Packaging and AI Chip Architecture: A New Era in Semiconductor Technology
glasp.co/hatch

### The Evolution of Chiplet Packaging and AI Chip Architecture: A New Era in Semiconductor Technology

Hatched on Mar 27, 2026 · 5 views

Glasp
Hatch
### The Future of AI Inference: Innovations and Strategies for Maximizing Performance
glasp.co/hatch

### The Future of AI Inference: Innovations and Strategies for Maximizing Performance

Hatched on Mar 26, 2026 · 8 views

Glasp
Hatch
The Future of Computing Chips in the Era of Large AI Models
glasp.co/hatch

The Future of Computing Chips in the Era of Large AI Models

Hatched on Mar 25, 2026 · 4 views

Glasp
Hatch
### The Rise of CXL and Tenstorrent: Pioneering New Frontiers in Computing
glasp.co/hatch

### The Rise of CXL and Tenstorrent: Pioneering New Frontiers in Computing

Hatched on Mar 24, 2026 · 10 views

Glasp
Hatch
# CXL与C4:提升数据中心效率的关键技术
glasp.co/hatch

# CXL与C4:提升数据中心效率的关键技术

Hatched on Mar 23, 2026 · 7 views

Glasp
Hatch
The Evolution of AI: Integrating System-on-Chip (SoC) Innovations with Large Language Models
glasp.co/hatch

The Evolution of AI: Integrating System-on-Chip (SoC) Innovations with Large Language Models

Hatched on Mar 22, 2026 · 4 views

Glasp
Hatch
# Optimizing LLM Inference: Technologies, Applications, and Challenges
glasp.co/hatch

# Optimizing LLM Inference: Technologies, Applications, and Challenges

Hatched on Mar 21, 2026 · 7 views

Glasp
Hatch
### Breaking the Limits of Computing Power: Innovations in Integrated Memory and Architecture
glasp.co/hatch

### Breaking the Limits of Computing Power: Innovations in Integrated Memory and Architecture

Hatched on Mar 20, 2026 · 9 views

Glasp
Hatch
# The Rise of H100: A New Era in GPU Architecture and the Quest for AI Dominance
glasp.co/hatch

# The Rise of H100: A New Era in GPU Architecture and the Quest for AI Dominance

Hatched on Mar 19, 2026 · 6 views

Glasp
Hatch
### The Misconceptions and Challenges of GPU Utilization in Generative AI
glasp.co/hatch

### The Misconceptions and Challenges of GPU Utilization in Generative AI

Hatched on Mar 18, 2026 · 7 views

Glasp
Hatch
The Future of Computing: Insights from the Rise of Intelligent Computing Centers and AI Accelerator Design
glasp.co/hatch

The Future of Computing: Insights from the Rise of Intelligent Computing Centers and AI Accelerator Design

Hatched on Mar 17, 2026 · 3 views

Glasp
Hatch
The Future of Large Model Inference: Performance, Parallelism, and Cache Management
glasp.co/hatch

The Future of Large Model Inference: Performance, Parallelism, and Cache Management

Hatched on Mar 16, 2026 · 11 views

Glasp
Hatch
### Optimizing GPU Batch Processing for Language Model Inference
glasp.co/hatch

### Optimizing GPU Batch Processing for Language Model Inference

Hatched on Mar 15, 2026 · 5 views

Glasp
Hatch
# CXL与现代计算的未来:推动数据中心性能的革命
glasp.co/hatch

# CXL与现代计算的未来:推动数据中心性能的革命

Hatched on Mar 14, 2026 · 8 views

Glasp
Hatch
Unlocking the Future of Computing: The Interplay Between LLMs and Next-Gen Memory Solutions
glasp.co/hatch

Unlocking the Future of Computing: The Interplay Between LLMs and Next-Gen Memory Solutions

Hatched on Mar 13, 2026 · 8 views

Glasp
Hatch
### The Future of AI Hardware: Design Guidelines and Architectural Innovations
glasp.co/hatch

### The Future of AI Hardware: Design Guidelines and Architectural Innovations

Hatched on Mar 12, 2026 · 5 views

Glasp
Hatch
Navigating the Future of AI: Insights on Chip Technology and Model Optimization
glasp.co/hatch

Navigating the Future of AI: Insights on Chip Technology and Model Optimization

Hatched on Mar 11, 2026 · 4 views

Glasp
Hatch
标题:深入探讨大规模语言模型的训练与推理:挑战与应对
glasp.co/hatch

标题:深入探讨大规模语言模型的训练与推理:挑战与应对

Hatched on Mar 10, 2026 · 2 views

Glasp
Hatch
# The Convergence of Cloud-Native Machine Learning and Large Language Models: A Technological Overview
glasp.co/hatch

# The Convergence of Cloud-Native Machine Learning and Large Language Models: A Technological Overview

Hatched on Mar 9, 2026 · 14 views

Glasp
Hatch
# The Evolution of GPU Technologies and Communication Efficiency in AI Training
glasp.co/hatch

# The Evolution of GPU Technologies and Communication Efficiency in AI Training

Hatched on Mar 8, 2026 · 14 views

Glasp
Hatch
# Optimizing Infrastructure for Large Language Models: From Setup to Performance
glasp.co/hatch

# Optimizing Infrastructure for Large Language Models: From Setup to Performance

Hatched on Mar 7, 2026 · 11 views

Glasp
Hatch
# Advancements in AI Technology: From Dynamic Chat Models to Powerhouse GPUs
glasp.co/hatch

# Advancements in AI Technology: From Dynamic Chat Models to Powerhouse GPUs

Hatched on Mar 6, 2026 · 9 views

Glasp
Hatch
The Evolution and Deployment of Recommender Systems in the Era of AI
glasp.co/hatch

The Evolution and Deployment of Recommender Systems in the Era of AI

Hatched on Mar 5, 2026 · 9 views

Glasp
Hatch
# The Silent Battle of Data Centers: Navigating the PCIe Landscape and Advancements in AI
glasp.co/hatch

# The Silent Battle of Data Centers: Navigating the PCIe Landscape and Advancements in AI

Hatched on Mar 4, 2026 · 4 views

Glasp
Hatch
### The Future of AI Chip Design: Innovations and Insights
glasp.co/hatch

### The Future of AI Chip Design: Innovations and Insights

Hatched on Mar 3, 2026 · 7 views

Next >