Efficient LLM Inference and Generation with SGLang — WalkSelf
⏱ 2 h 36 min 📚 26 aulas 🎧 Versão em áudio

Efficient LLM Inference and Generation with SGLang

Learn how to speed up text and image generation using SGLang to implement advanced caching, RadixAttention, and structured outputs for cost-effective AI applications.

  • 💬 Instrutor de IA
    Pergunte sobre qualquer aula e receba uma resposta clara na hora, quando quiser.
  • 🕐 Comece quando quiser
    Sem horários nem prazos: aprenda no seu ritmo, quando quiser.
  • 🌐 Em português
    Aulas, tarefas e certificado: tudo totalmente no seu idioma.

Sobre este curso

Deploying large language models can be slow and expensive, making inference optimization a critical skill for modern developers. Understanding the mechanics of model serving allows you to build responsive applications without skyrocketing compute costs. This course guides you through the foundational concepts of LLM serving and optimization using the open-source SGLang framework. What you'll learn: - Understand the core mechanics of LLM inference, including token generation phases and latency bottlenecks. - Implement efficient caching strategies using KV cache and RadixAttention to accelerate repetitive prompts. - Configure SGLang to handle concurrent text and image generation tasks efficiently. - Apply structured output techniques to guarantee precise JSON responses from your models. - Optimize hardware resource utilization to make model deployment cheaper and more scalable. You will start with key terminology and the foundational concepts of model serving before exploring SGLang configuration, caching mechanics, and structured routing. Through clear written explanations and detailed code snippets, you will build a practical understanding of modern inference optimization. This course is designed for software developers and AI enthusiasts who want to learn the basics of efficient model serving, with no prior SGLang experience required. Start reading today to make your AI generation pipelines faster and more cost-effective.

O que você vai receber

  • 📜 Certificado de conclusão
    Adicione ao seu perfil do LinkedIn
  • 💬 Tutor AI pessoal
    Travou em uma aula? Pergunte ao seu tutor integrado qualquer coisa, a qualquer hora.
  • 🎧 Versão em áudio incluída
    Estude em qualquer lugar, sem tela
  • ♾️ Acesso vitalício
    Volte quando quiser, sem expirar
  • 📱 Celular ou computador
    Funciona em qualquer dispositivo
  • 💸 Reembolso em 14 dias
    Sem perguntas
  • Curto e focado
    2 h 36 min de conteúdo prático

Avaliações

Ainda não há avaliações — seja o primeiro a compartilhar sua experiência.

Escrever uma avaliação

Pediremos para fazer login após enviar — o rascunho fica salvo.

Outros também fizeram

Perguntas frequentes

O que preciso para fazer este curso? +

Só um celular ou computador com internet. Sem instalações nem hardware especial.

Como faço para pagar? +

Com cartão via Stripe. Não guardamos dados do cartão — o Stripe processa com segurança.

Posso pedir reembolso? +

Sim — reembolso integral em 14 dias, sem perguntas.

Por quanto tempo terei acesso? +

Para sempre. Uma vez comprado, o curso é seu para revisar quando quiser.

Vou receber um certificado? +

Sim. Ao concluir, você recebe um certificado que pode adicionar ao seu perfil do LinkedIn.

Feito para profissionais em
Tecnologia Design Finanças Marketing Saúde Educação Hotelaria Indústria