View Detailed Profile
LLM Compression Explained: Build Faster, Efficient AI Models

LLM Compression Explained: Build Faster, Efficient AI Models

Ready to become a certified watsonx

Model Compression & Optimization: Making AI Models Faster | #GirlsWhoML

Model Compression & Optimization: Making AI Models Faster | #GirlsWhoML

How do you take a state-of-the-art

Optimize Your AI - Quantization Explained

Optimize Your AI - Quantization Explained

Run massive

Model Compression: Optimize VLM Inference with These Techniques

Model Compression: Optimize VLM Inference with These Techniques

Model compression

Model Compression Explained: Making AI Smaller & Faster ๐Ÿš€

Model Compression Explained: Making AI Smaller & Faster ๐Ÿš€

Ever wonder how powerful

Quantization vs Pruning vs Distillation: Optimizing NNs for Inference

Quantization vs Pruning vs Distillation: Optimizing NNs for Inference

Try Voice Writer - speak your thoughts and let

How LLMs survive in low precision | Quantization Fundamentals

How LLMs survive in low precision | Quantization Fundamentals

In this video, we discuss the fundamentals of

Model Compression and Optimization: Making AIs Faster and Smaller

Model Compression and Optimization: Making AIs Faster and Smaller

Part of the '

Optimize Your AI Models

Optimize Your AI Models

Dive deep into the world of Large Language

Run AI on ANY Device: Model Compression & Quantization Explained!

Run AI on ANY Device: Model Compression & Quantization Explained!

Discover the secrets of

Model Quantization Explained | GPTQ, AWQ, SmoothQuant & AI Model Compression

Model Quantization Explained | GPTQ, AWQ, SmoothQuant & AI Model Compression

Deploying modern

AI Compression is 300x Better (but we don't use it)

AI Compression is 300x Better (but we don't use it)

It's crazy

ModelShrink Demo | Using GPT-OSS to Automate PyTorch Model Compression

ModelShrink Demo | Using GPT-OSS to Automate PyTorch Model Compression

Our official submission for the OpenAI & Hugging Face Open