Kevin Di's Hatch

Kevin Di
Latest
Popular
Glasp
Hatch
Unveiling the Power of H100: A Deep Dive into Modern GPU Architecture
glasp.co/hatch

Unveiling the Power of H100: A Deep Dive into Modern GPU Architecture

Hatched on Mar 11, 2024 · 19 views

Glasp
Hatch
"The Intersection of Lightweight Inference Frameworks and the Future of GPU Entrepreneurship"
glasp.co/hatch

"The Intersection of Lightweight Inference Frameworks and the Future of GPU Entrepreneurship"

Hatched on Mar 10, 2024 · 14 views

Glasp
Hatch
The Battle for HBM: Storage Giants Compete for High Bandwidth Memory
glasp.co/hatch

The Battle for HBM: Storage Giants Compete for High Bandwidth Memory

Hatched on Mar 9, 2024 · 14 views

Glasp
Hatch
Misconceptions about GPUs in the field of generative AI
glasp.co/hatch

Misconceptions about GPUs in the field of generative AI

Hatched on Mar 8, 2024 · 13 views

Glasp
Hatch
Accelerating Transformer with Sparse Attention Accelerators
glasp.co/hatch

Accelerating Transformer with Sparse Attention Accelerators

Hatched on Mar 7, 2024 · 9 views

Glasp
Hatch
"Unleashing the Power of Advanced Packaging: Exploring the Intersection of Apple and NVIDIA"
glasp.co/hatch

"Unleashing the Power of Advanced Packaging: Exploring the Intersection of Apple and NVIDIA"

Hatched on Mar 6, 2024 · 20 views

Glasp
Hatch
The Future Development Trends of Computing Chips: Insights from the Rise of Large Models like ChatGPT and OAI Server Design Guidelines
glasp.co/hatch

The Future Development Trends of Computing Chips: Insights from the Rise of Large Models like ChatGPT and OAI Server Design Guidelines

Hatched on Mar 5, 2024 · 6 views

Glasp
Hatch
The Journey of DSA: From Domain Specific Accelerator to Domain Specific System
glasp.co/hatch

The Journey of DSA: From Domain Specific Accelerator to Domain Specific System

Hatched on Mar 4, 2024 · 7 views

Glasp
Hatch
The Intersection of Nvidia's H100 and Open AI Server Design Guidelines
glasp.co/hatch

The Intersection of Nvidia's H100 and Open AI Server Design Guidelines

Hatched on Mar 3, 2024 · 12 views

Glasp
Hatch
The Battle for AI Dominance: The Intersection of Chip Investments and Semiconductor Blockades
glasp.co/hatch

The Battle for AI Dominance: The Intersection of Chip Investments and Semiconductor Blockades

Hatched on Mar 2, 2024 · 5 views

Glasp
Hatch
DSA's Transformation Journey: From Accelerator to System
glasp.co/hatch

DSA's Transformation Journey: From Accelerator to System

Hatched on Mar 1, 2024 · 11 views

Glasp
Hatch
These memory-limited layers pose a challenge for deep learning models running on GPUs, as the time spent on memory transfers can significantly impact the overall performance. To address this issue, NVIDIA provides a guide called "Memory-Limited Layers User's Guide" to help users optimize the usage of memory-limited layers in their deep learning models.
glasp.co/hatch

These memory-limited layers pose a challenge for deep learning models running on GPUs, as the time spent on memory transfers can significantly impact the overall performance. To address this issue, NVIDIA provides a guide called "Memory-Limited Layers User's Guide" to help users optimize the usage of memory-limited layers in their deep learning models.

Hatched on Feb 29, 2024 · 14 views

Glasp
Hatch
The Intersection of Large Model Inference and GPU Optimization: Dispelling Misconceptions and Unveiling Insights
glasp.co/hatch

The Intersection of Large Model Inference and GPU Optimization: Dispelling Misconceptions and Unveiling Insights

Hatched on Feb 28, 2024 · 14 views

Glasp
Hatch
"The Rise of NVIDIA: A Record-Breaking Year and the Dominance of H100"
glasp.co/hatch

"The Rise of NVIDIA: A Record-Breaking Year and the Dominance of H100"

Hatched on Feb 27, 2024 · 7 views

Glasp
Hatch
Maximizing Performance in Fully-Connected Layers: Understanding FlashAttention2 and GEMM Parameters
glasp.co/hatch

Maximizing Performance in Fully-Connected Layers: Understanding FlashAttention2 and GEMM Parameters

Hatched on Feb 26, 2024 · 18 views

Glasp
Hatch
"Comparing Network Architectures: InfiniBand vs. RoCE and the LightLLM Framework"
glasp.co/hatch

"Comparing Network Architectures: InfiniBand vs. RoCE and the LightLLM Framework"

Hatched on Feb 25, 2024 · 17 views

Glasp
Hatch
The Future of High Bandwidth Memory (HBM) and Optimized Inference Technologies
glasp.co/hatch

The Future of High Bandwidth Memory (HBM) and Optimized Inference Technologies

Hatched on Feb 24, 2024 · 21 views

Glasp
Hatch
The Second Half of the GPU Chip Entrepreneurship: A Game of Survival and Outperforming Teammates
glasp.co/hatch

The Second Half of the GPU Chip Entrepreneurship: A Game of Survival and Outperforming Teammates

Hatched on Feb 23, 2024 · 8 views

Glasp
Hatch
Optimizing Attention Performance: From FlashAttention to PagedAttention
glasp.co/hatch

Optimizing Attention Performance: From FlashAttention to PagedAttention

Hatched on Feb 22, 2024 · 27 views

Glasp
Hatch
Accelerating Generative AI with PyTorch II: GPT, Fast - AI芯片,看什么?
glasp.co/hatch

Accelerating Generative AI with PyTorch II: GPT, Fast - AI芯片,看什么?

Hatched on Feb 21, 2024 · 18 views

Glasp
Hatch
"Optimizing Large-scale Model Inference: From Model Analysis to Computational Efficiency"
glasp.co/hatch

"Optimizing Large-scale Model Inference: From Model Analysis to Computational Efficiency"

Hatched on Feb 20, 2024 · 9 views

Glasp
Hatch
Unveiling the Performance Capabilities and Cost Analysis of NVIDIA's BR100 and H100 GPUs
glasp.co/hatch

Unveiling the Performance Capabilities and Cost Analysis of NVIDIA's BR100 and H100 GPUs

Hatched on Feb 19, 2024 · 14 views

Glasp
Hatch
The Battle of High-Performance Computing: NVIDIA vs. BR100
glasp.co/hatch

The Battle of High-Performance Computing: NVIDIA vs. BR100

Hatched on Feb 18, 2024 · 29 views

Glasp
Hatch
Exploring Network Architecture Options for Improved Performance in Distributed Computing
glasp.co/hatch

Exploring Network Architecture Options for Improved Performance in Distributed Computing

Hatched on Feb 17, 2024 · 12 views

Glasp
Hatch
The Battle for AI Computing Power: Storage Chip Giants, NVIDIA, and AMD
glasp.co/hatch

The Battle for AI Computing Power: Storage Chip Giants, NVIDIA, and AMD

Hatched on Feb 16, 2024 · 7 views

Glasp
Hatch
"The Road to DSA's Transformation: From DSA to DSS"
glasp.co/hatch

"The Road to DSA's Transformation: From DSA to DSS"

Hatched on Feb 15, 2024 · 12 views

Glasp
Hatch
PyTorch 2.0: Unleashing the Power of Advanced Hardware and Manufacturing Techniques
glasp.co/hatch

PyTorch 2.0: Unleashing the Power of Advanced Hardware and Manufacturing Techniques

Hatched on Feb 14, 2024 · 8 views

Glasp
Hatch
The Battle for HBM Supremacy and the Future of Computing Chips
glasp.co/hatch

The Battle for HBM Supremacy and the Future of Computing Chips

Hatched on Feb 13, 2024 · 21 views

Glasp
Hatch
Unveiling the Power Behind NVIDIA's Empire
glasp.co/hatch

Unveiling the Power Behind NVIDIA's Empire

Hatched on Feb 12, 2024 · 15 views

Glasp
Hatch
The Future of AI: Analyzing Transformer Models and the Rise of GPU Chip Startups
glasp.co/hatch

The Future of AI: Analyzing Transformer Models and the Rise of GPU Chip Startups

Hatched on Feb 11, 2024 · 4 views

Glasp
Hatch
The Advancements in Open Bilingual Chat Language Models and High Bandwidth Memory
glasp.co/hatch

The Advancements in Open Bilingual Chat Language Models and High Bandwidth Memory

Hatched on Feb 10, 2024 · 7 views

Glasp
Hatch
The Five Misconceptions about GPUs in the Generative AI Field
glasp.co/hatch

The Five Misconceptions about GPUs in the Generative AI Field

Hatched on Feb 9, 2024 · 10 views

Glasp
Hatch
"The Power of Domain Specific Systems and the TPUv5e: Revolutionizing AI Hardware"
glasp.co/hatch

"The Power of Domain Specific Systems and the TPUv5e: Revolutionizing AI Hardware"

Hatched on Feb 8, 2024 · 16 views

Glasp
Hatch
The Intersection of AI DC Parameters and the Rise of NVIDIA GPUs
glasp.co/hatch

The Intersection of AI DC Parameters and the Rise of NVIDIA GPUs

Hatched on Feb 7, 2024 · 15 views

Glasp
Hatch
Optimizing KV Cache and Understanding StreamingLLM: Insights from NLP and Google TPU v4
glasp.co/hatch

Optimizing KV Cache and Understanding StreamingLLM: Insights from NLP and Google TPU v4

Hatched on Feb 6, 2024 · 20 views

Glasp
Hatch
The Evolving Landscape of Machine Learning: From Nvidia's CUDA Monopoly to LightLLM's Python-based LLM Framework
glasp.co/hatch

The Evolving Landscape of Machine Learning: From Nvidia's CUDA Monopoly to LightLLM's Python-based LLM Framework

Hatched on Feb 5, 2024 · 31 views

Glasp
Hatch
解析 Transformer 模型 | Way to AGI
glasp.co/hatch

解析 Transformer 模型 | Way to AGI

Hatched on Feb 4, 2024 · 41 views

Glasp
Hatch
"Demystifying the FP32 Performance of BR100 and Debunking Misconceptions about GPUs in Generative AI"
glasp.co/hatch

"Demystifying the FP32 Performance of BR100 and Debunking Misconceptions about GPUs in Generative AI"

Hatched on Feb 3, 2024 · 5 views

Glasp
Hatch
The Intersection of AI and Storage: Optimizing Data Center Parameters and HBM Advancements
glasp.co/hatch

The Intersection of AI and Storage: Optimizing Data Center Parameters and HBM Advancements

Hatched on Feb 2, 2024 · 9 views

Glasp
Hatch
The Rise of NVIDIA GPU: A Battle That Never Stops and No Permanent Friends
glasp.co/hatch

The Rise of NVIDIA GPU: A Battle That Never Stops and No Permanent Friends

Hatched on Feb 1, 2024 · 12 views

Glasp
Hatch
Zig: The Promising Replacement for C in System Programming
glasp.co/hatch

Zig: The Promising Replacement for C in System Programming

Hatched on Jan 31, 2024 · 7 views

Glasp
Hatch
The Magic Behind Breakthroughs in Domestic Mobile Phone 7nm Chips and the Optimization of Inference Calculations
glasp.co/hatch

The Magic Behind Breakthroughs in Domestic Mobile Phone 7nm Chips and the Optimization of Inference Calculations

Hatched on Jan 30, 2024 · 8 views

Glasp
Hatch
Google Gemini Eats The World – Gemini Smashes GPT-4 By 5X, The GPU-Poors
glasp.co/hatch

Google Gemini Eats The World – Gemini Smashes GPT-4 By 5X, The GPU-Poors

Hatched on Jan 29, 2024 · 14 views

Glasp
Hatch
The Rise of Smart Computing Centers and Advanced GPU Servers: A Look into the Future
glasp.co/hatch

The Rise of Smart Computing Centers and Advanced GPU Servers: A Look into the Future

Hatched on Jan 28, 2024 · 9 views

Glasp
Hatch
Zig: The Future Replacement for C in System Programming
glasp.co/hatch

Zig: The Future Replacement for C in System Programming

Hatched on Jan 27, 2024 · 9 views

Glasp
Hatch
The Future of Open AI Server Design: Zig Language as a Game Changer
glasp.co/hatch

The Future of Open AI Server Design: Zig Language as a Game Changer

Hatched on Jan 26, 2024 · 11 views

Glasp
Hatch
The Rise of AI Data Centers: Unveiling the Secrets Behind the "智算中心"
glasp.co/hatch

The Rise of AI Data Centers: Unveiling the Secrets Behind the "智算中心"

Hatched on Jan 25, 2024 · 16 views

Glasp
Hatch
Debunking 5 Misconceptions about GPU in the Field of Generative AI
glasp.co/hatch

Debunking 5 Misconceptions about GPU in the Field of Generative AI

Hatched on Jan 24, 2024 · 8 views

Glasp
Hatch
Boosting Language Model Inference Performance: Insights and Techniques
glasp.co/hatch

Boosting Language Model Inference Performance: Insights and Techniques

Hatched on Jan 23, 2024 · 8 views

Glasp
Hatch
The Future of AI Chips: Insights from Google TPU v4 and FlashAttention
glasp.co/hatch

The Future of AI Chips: Insights from Google TPU v4 and FlashAttention

Hatched on Jan 22, 2024 · 17 views

Glasp
Hatch
"Optimizing Large-scale Model Inference: From Analysis to Computational Efficiency"
glasp.co/hatch

"Optimizing Large-scale Model Inference: From Analysis to Computational Efficiency"

Hatched on Jan 21, 2024 · 11 views

Glasp
Hatch
The Battle for High-Bandwidth Memory (HBM): A Look at the Latest Developments in AI Chips and Storage Technology
glasp.co/hatch

The Battle for High-Bandwidth Memory (HBM): A Look at the Latest Developments in AI Chips and Storage Technology

Hatched on Jan 20, 2024 · 17 views

Glasp
Hatch
The Future of AI Chips: Breakthroughs in Chip Manufacturing and Design
glasp.co/hatch

The Future of AI Chips: Breakthroughs in Chip Manufacturing and Design

Hatched on Jan 19, 2024 · 15 views

Glasp
Hatch
"The Intersection of Model Analysis, Computational Optimization, and Nvidia's CUDA Monopoly: Breaking OpenAI Triton and PyTorch 2.0"
glasp.co/hatch

"The Intersection of Model Analysis, Computational Optimization, and Nvidia's CUDA Monopoly: Breaking OpenAI Triton and PyTorch 2.0"

Hatched on Jan 18, 2024 · 11 views

Glasp
Hatch
Decoding the US Chip Blockade: A Form of Warfare
glasp.co/hatch

Decoding the US Chip Blockade: A Form of Warfare

Hatched on Jan 17, 2024 · 9 views

Glasp
Hatch
TPUv5e: The New Benchmark in Cost-Efficient Inference and Training for <200B Parameter Models
glasp.co/hatch

TPUv5e: The New Benchmark in Cost-Efficient Inference and Training for <200B Parameter Models

Hatched on Jan 16, 2024 · 17 views

Glasp
Hatch
GPU在生成式AI领域的五大误解
glasp.co/hatch

GPU在生成式AI领域的五大误解

Hatched on Jan 15, 2024 · 11 views

Glasp
Hatch
The Future of GPU Chip Startups: A Battle for Survival in the Second Half
glasp.co/hatch

The Future of GPU Chip Startups: A Battle for Survival in the Second Half

Hatched on Jan 14, 2024 · 12 views

Glasp
Hatch
Best Practices for Building and Deploying Recommender Systems and the Magic Behind Breakthroughs in Domestic 7nm Chipsets
glasp.co/hatch

Best Practices for Building and Deploying Recommender Systems and the Magic Behind Breakthroughs in Domestic 7nm Chipsets

Hatched on Jan 13, 2024 · 12 views

Glasp
Hatch
The Art of Language Modeling Inference
glasp.co/hatch

The Art of Language Modeling Inference

Hatched on Jan 12, 2024 · 7 views

Next >