Kevin Di's Hatch

Kevin Di
Latest
Popular
Glasp
Hatch
从Google TPU v4看AI芯片的未来
glasp.co/hatch

从Google TPU v4看AI芯片的未来

Hatched on Apr 5, 2024 · 50 views

Glasp
Hatch
"Optimizing NLP Inference Performance: Exploring LightLLM and StreamingLLM"
glasp.co/hatch

"Optimizing NLP Inference Performance: Exploring LightLLM and StreamingLLM"

Hatched on Apr 4, 2024 · 16 views

Glasp
Hatch
HBM3 and Nvidia: A Powerhouse Combination for Advanced Computing
glasp.co/hatch

HBM3 and Nvidia: A Powerhouse Combination for Advanced Computing

Hatched on Apr 3, 2024 · 9 views

Glasp
Hatch
The Growing Landscape of AI Chip Companies: Microsoft's Investment in d-Matrix and NVIDIA's Ongoing Bets
glasp.co/hatch

The Growing Landscape of AI Chip Companies: Microsoft's Investment in d-Matrix and NVIDIA's Ongoing Bets

Hatched on Apr 2, 2024 · 13 views

Glasp
Hatch
The Future of AI Chips: Insights from Google TPU v4 and the Next Phase of GPU Chip Startups
glasp.co/hatch

The Future of AI Chips: Insights from Google TPU v4 and the Next Phase of GPU Chip Startups

Hatched on Apr 1, 2024 · 8 views

Glasp
Hatch
"Demystifying Five Misconceptions about GPUs in the Field of Generative AI"
glasp.co/hatch

"Demystifying Five Misconceptions about GPUs in the Field of Generative AI"

Hatched on Mar 31, 2024 · 9 views

Glasp
Hatch
Revolutionizing AI Chip Compiler: The Unveiling of Jim Keller's BUDA and Hardware Details
glasp.co/hatch

Revolutionizing AI Chip Compiler: The Unveiling of Jim Keller's BUDA and Hardware Details

Hatched on Mar 30, 2024 · 32 views

Glasp
Hatch
The Rise of NVIDIA's Powerful GPUs and Their Continuous Innovation
glasp.co/hatch

The Rise of NVIDIA's Powerful GPUs and Their Continuous Innovation

Hatched on Mar 29, 2024 · 8 views

Glasp
Hatch
Exploring Network Architecture Options for Reduced Latency and Enhanced Performance
glasp.co/hatch

Exploring Network Architecture Options for Reduced Latency and Enhanced Performance

Hatched on Mar 28, 2024 · 12 views

Glasp
Hatch
Accelerating Transformer Models: A Comprehensive Overview
glasp.co/hatch

Accelerating Transformer Models: A Comprehensive Overview

Hatched on Mar 27, 2024 · 11 views

Glasp
Hatch
The Ongoing Marathon in GPU Chip Startups: Navigating Strategic Directions and Operational Challenges
glasp.co/hatch

The Ongoing Marathon in GPU Chip Startups: Navigating Strategic Directions and Operational Challenges

Hatched on Mar 26, 2024 · 10 views

Glasp
Hatch
"The Journey of DSA: From Domain Specific Accelerator to Domain Specific System"
glasp.co/hatch

"The Journey of DSA: From Domain Specific Accelerator to Domain Specific System"

Hatched on Mar 25, 2024 · 9 views

Glasp
Hatch
The Power of Language Models and Advancements in Chip Design
glasp.co/hatch

The Power of Language Models and Advancements in Chip Design

Hatched on Mar 24, 2024 · 8 views

Glasp
Hatch
The Intersection of GPT-2 and Nvidia's H100: Unveiling the Key Connections and Insights
glasp.co/hatch

The Intersection of GPT-2 and Nvidia's H100: Unveiling the Key Connections and Insights

Hatched on Mar 23, 2024 · 7 views

Glasp
Hatch
Understanding Convolutional Layers and Transformer Models in Neural Networks
glasp.co/hatch

Understanding Convolutional Layers and Transformer Models in Neural Networks

Hatched on Mar 22, 2024 · 13 views

Glasp
Hatch
The Future Trends in Computing Chips: Insights from the Rise of Large Models like ChatGPT and Others
glasp.co/hatch

The Future Trends in Computing Chips: Insights from the Rise of Large Models like ChatGPT and Others

Hatched on Mar 21, 2024 · 10 views

Glasp
Hatch
Optimizing KV Cache and Exploring the Latest Chip Innovations
glasp.co/hatch

Optimizing KV Cache and Exploring the Latest Chip Innovations

Hatched on Mar 20, 2024 · 15 views

Glasp
Hatch
Unveiling the Power behind GPU Performance and Advanced Packaging
glasp.co/hatch

Unveiling the Power behind GPU Performance and Advanced Packaging

Hatched on Mar 19, 2024 · 2 views

Glasp
Hatch
The Future of DeepTech and the Implications of High Costs
glasp.co/hatch

The Future of DeepTech and the Implications of High Costs

Hatched on Mar 18, 2024 · 10 views

Glasp
Hatch
The Future of AI Hardware: Google's Next-Generation Chip and the Rise of Chinese Automakers in the Self-Driving Market
glasp.co/hatch

The Future of AI Hardware: Google's Next-Generation Chip and the Rise of Chinese Automakers in the Self-Driving Market

Hatched on Mar 17, 2024 · 16 views

Glasp
Hatch
"Boosting Performance in Generative AI with LightLLM and Efficient Routing"
glasp.co/hatch

"Boosting Performance in Generative AI with LightLLM and Efficient Routing"

Hatched on Mar 16, 2024 · 17 views

Glasp
Hatch
HBM: The Future of Memory Technology
glasp.co/hatch

HBM: The Future of Memory Technology

Hatched on Mar 15, 2024 · 41 views

Glasp
Hatch
The Future of AI Chip Design: Insights from DSA and Google TPU v4
glasp.co/hatch

The Future of AI Chip Design: Insights from DSA and Google TPU v4

Hatched on Mar 14, 2024 · 28 views

Glasp
Hatch
"大型语言模型的推理演算"与"DSA的翻身路":探索定制化的解决方案
glasp.co/hatch

"大型语言模型的推理演算"与"DSA的翻身路":探索定制化的解决方案

Hatched on Mar 13, 2024 · 23 views

Glasp
Hatch
The Future Development Trend of Computing Chips: Insights from the Rise of Large Models like ChatGPT and More
glasp.co/hatch

The Future Development Trend of Computing Chips: Insights from the Rise of Large Models like ChatGPT and More

Hatched on Mar 12, 2024 · 4 views

Glasp
Hatch
Unveiling the Power of H100: A Deep Dive into Modern GPU Architecture
glasp.co/hatch

Unveiling the Power of H100: A Deep Dive into Modern GPU Architecture

Hatched on Mar 11, 2024 · 19 views

Glasp
Hatch
"The Intersection of Lightweight Inference Frameworks and the Future of GPU Entrepreneurship"
glasp.co/hatch

"The Intersection of Lightweight Inference Frameworks and the Future of GPU Entrepreneurship"

Hatched on Mar 10, 2024 · 14 views

Glasp
Hatch
The Battle for HBM: Storage Giants Compete for High Bandwidth Memory
glasp.co/hatch

The Battle for HBM: Storage Giants Compete for High Bandwidth Memory

Hatched on Mar 9, 2024 · 13 views

Glasp
Hatch
Misconceptions about GPUs in the field of generative AI
glasp.co/hatch

Misconceptions about GPUs in the field of generative AI

Hatched on Mar 8, 2024 · 12 views

Glasp
Hatch
Accelerating Transformer with Sparse Attention Accelerators
glasp.co/hatch

Accelerating Transformer with Sparse Attention Accelerators

Hatched on Mar 7, 2024 · 9 views

Glasp
Hatch
"Unleashing the Power of Advanced Packaging: Exploring the Intersection of Apple and NVIDIA"
glasp.co/hatch

"Unleashing the Power of Advanced Packaging: Exploring the Intersection of Apple and NVIDIA"

Hatched on Mar 6, 2024 · 19 views

Glasp
Hatch
The Future Development Trends of Computing Chips: Insights from the Rise of Large Models like ChatGPT and OAI Server Design Guidelines
glasp.co/hatch

The Future Development Trends of Computing Chips: Insights from the Rise of Large Models like ChatGPT and OAI Server Design Guidelines

Hatched on Mar 5, 2024 · 6 views

Glasp
Hatch
The Journey of DSA: From Domain Specific Accelerator to Domain Specific System
glasp.co/hatch

The Journey of DSA: From Domain Specific Accelerator to Domain Specific System

Hatched on Mar 4, 2024 · 6 views

Glasp
Hatch
The Intersection of Nvidia's H100 and Open AI Server Design Guidelines
glasp.co/hatch

The Intersection of Nvidia's H100 and Open AI Server Design Guidelines

Hatched on Mar 3, 2024 · 9 views

Glasp
Hatch
The Battle for AI Dominance: The Intersection of Chip Investments and Semiconductor Blockades
glasp.co/hatch

The Battle for AI Dominance: The Intersection of Chip Investments and Semiconductor Blockades

Hatched on Mar 2, 2024 · 4 views

Glasp
Hatch
DSA's Transformation Journey: From Accelerator to System
glasp.co/hatch

DSA's Transformation Journey: From Accelerator to System

Hatched on Mar 1, 2024 · 11 views

Glasp
Hatch
These memory-limited layers pose a challenge for deep learning models running on GPUs, as the time spent on memory transfers can significantly impact the overall performance. To address this issue, NVIDIA provides a guide called "Memory-Limited Layers User's Guide" to help users optimize the usage of memory-limited layers in their deep learning models.
glasp.co/hatch

These memory-limited layers pose a challenge for deep learning models running on GPUs, as the time spent on memory transfers can significantly impact the overall performance. To address this issue, NVIDIA provides a guide called "Memory-Limited Layers User's Guide" to help users optimize the usage of memory-limited layers in their deep learning models.

Hatched on Feb 29, 2024 · 12 views

Glasp
Hatch
The Intersection of Large Model Inference and GPU Optimization: Dispelling Misconceptions and Unveiling Insights
glasp.co/hatch

The Intersection of Large Model Inference and GPU Optimization: Dispelling Misconceptions and Unveiling Insights

Hatched on Feb 28, 2024 · 14 views

Glasp
Hatch
"The Rise of NVIDIA: A Record-Breaking Year and the Dominance of H100"
glasp.co/hatch

"The Rise of NVIDIA: A Record-Breaking Year and the Dominance of H100"

Hatched on Feb 27, 2024 · 7 views

Glasp
Hatch
Maximizing Performance in Fully-Connected Layers: Understanding FlashAttention2 and GEMM Parameters
glasp.co/hatch

Maximizing Performance in Fully-Connected Layers: Understanding FlashAttention2 and GEMM Parameters

Hatched on Feb 26, 2024 · 17 views

Glasp
Hatch
"Comparing Network Architectures: InfiniBand vs. RoCE and the LightLLM Framework"
glasp.co/hatch

"Comparing Network Architectures: InfiniBand vs. RoCE and the LightLLM Framework"

Hatched on Feb 25, 2024 · 17 views

Glasp
Hatch
The Future of High Bandwidth Memory (HBM) and Optimized Inference Technologies
glasp.co/hatch

The Future of High Bandwidth Memory (HBM) and Optimized Inference Technologies

Hatched on Feb 24, 2024 · 17 views

Glasp
Hatch
The Second Half of the GPU Chip Entrepreneurship: A Game of Survival and Outperforming Teammates
glasp.co/hatch

The Second Half of the GPU Chip Entrepreneurship: A Game of Survival and Outperforming Teammates

Hatched on Feb 23, 2024 · 7 views

Glasp
Hatch
Optimizing Attention Performance: From FlashAttention to PagedAttention
glasp.co/hatch

Optimizing Attention Performance: From FlashAttention to PagedAttention

Hatched on Feb 22, 2024 · 25 views

Glasp
Hatch
Accelerating Generative AI with PyTorch II: GPT, Fast - AI芯片,看什么?
glasp.co/hatch

Accelerating Generative AI with PyTorch II: GPT, Fast - AI芯片,看什么?

Hatched on Feb 21, 2024 · 18 views

Glasp
Hatch
"Optimizing Large-scale Model Inference: From Model Analysis to Computational Efficiency"
glasp.co/hatch

"Optimizing Large-scale Model Inference: From Model Analysis to Computational Efficiency"

Hatched on Feb 20, 2024 · 8 views

Glasp
Hatch
Unveiling the Performance Capabilities and Cost Analysis of NVIDIA's BR100 and H100 GPUs
glasp.co/hatch

Unveiling the Performance Capabilities and Cost Analysis of NVIDIA's BR100 and H100 GPUs

Hatched on Feb 19, 2024 · 14 views

Glasp
Hatch
The Battle of High-Performance Computing: NVIDIA vs. BR100
glasp.co/hatch

The Battle of High-Performance Computing: NVIDIA vs. BR100

Hatched on Feb 18, 2024 · 29 views

Glasp
Hatch
Exploring Network Architecture Options for Improved Performance in Distributed Computing
glasp.co/hatch

Exploring Network Architecture Options for Improved Performance in Distributed Computing

Hatched on Feb 17, 2024 · 12 views

Glasp
Hatch
The Battle for AI Computing Power: Storage Chip Giants, NVIDIA, and AMD
glasp.co/hatch

The Battle for AI Computing Power: Storage Chip Giants, NVIDIA, and AMD

Hatched on Feb 16, 2024 · 7 views

Glasp
Hatch
"The Road to DSA's Transformation: From DSA to DSS"
glasp.co/hatch

"The Road to DSA's Transformation: From DSA to DSS"

Hatched on Feb 15, 2024 · 10 views

Glasp
Hatch
PyTorch 2.0: Unleashing the Power of Advanced Hardware and Manufacturing Techniques
glasp.co/hatch

PyTorch 2.0: Unleashing the Power of Advanced Hardware and Manufacturing Techniques

Hatched on Feb 14, 2024 · 8 views

Glasp
Hatch
The Battle for HBM Supremacy and the Future of Computing Chips
glasp.co/hatch

The Battle for HBM Supremacy and the Future of Computing Chips

Hatched on Feb 13, 2024 · 21 views

Glasp
Hatch
Unveiling the Power Behind NVIDIA's Empire
glasp.co/hatch

Unveiling the Power Behind NVIDIA's Empire

Hatched on Feb 12, 2024 · 15 views

Glasp
Hatch
The Future of AI: Analyzing Transformer Models and the Rise of GPU Chip Startups
glasp.co/hatch

The Future of AI: Analyzing Transformer Models and the Rise of GPU Chip Startups

Hatched on Feb 11, 2024 · 4 views

Glasp
Hatch
The Advancements in Open Bilingual Chat Language Models and High Bandwidth Memory
glasp.co/hatch

The Advancements in Open Bilingual Chat Language Models and High Bandwidth Memory

Hatched on Feb 10, 2024 · 7 views

Glasp
Hatch
The Five Misconceptions about GPUs in the Generative AI Field
glasp.co/hatch

The Five Misconceptions about GPUs in the Generative AI Field

Hatched on Feb 9, 2024 · 10 views

Glasp
Hatch
"The Power of Domain Specific Systems and the TPUv5e: Revolutionizing AI Hardware"
glasp.co/hatch

"The Power of Domain Specific Systems and the TPUv5e: Revolutionizing AI Hardware"

Hatched on Feb 8, 2024 · 16 views

Glasp
Hatch
The Intersection of AI DC Parameters and the Rise of NVIDIA GPUs
glasp.co/hatch

The Intersection of AI DC Parameters and the Rise of NVIDIA GPUs

Hatched on Feb 7, 2024 · 15 views

Glasp
Hatch
Optimizing KV Cache and Understanding StreamingLLM: Insights from NLP and Google TPU v4
glasp.co/hatch

Optimizing KV Cache and Understanding StreamingLLM: Insights from NLP and Google TPU v4

Hatched on Feb 6, 2024 · 20 views

Next >