Planning Multi-Object Images with LLMs and Progressive Diffusion
Learn how to decompose complex text-to-image prompts into structured layouts using language models and generate accurate multi-object scenes step-by-step.
-
💬
مدرب ذكاء اصطناعي
اسأل عن أي درس واحصل على إجابة واضحة فورًا، في أي وقت. -
🕐
ابدأ في أي وقت
بلا جداول أو مواعيد نهائية — تعلّم بوتيرتك، وقتما يناسبك. -
🌐
بالعربية
الدروس والمهام والشهادة — كل ذلك بلغتك بالكامل.
حول هذه الدورة
Generating images with multiple overlapping objects often leads to chaotic results, where AI models struggle to place items exactly where you want them. By introducing a structured planning phase before rendering, you can guide diffusion models to generate complex scenes with high spatial accuracy. This text-only course introduces you to the foundational concepts of progressive multi-object generation, exploring how Large Language Models (LLMs) act as layout planners to decompose a single prompt into step-by-step instructions that progressive diffusion models execute seamlessly.
What you'll learn:
- Understand the core limitations of standard text-to-image models when handling multiple distinct objects
- Learn how LLMs generate spatial layouts and coordinate plans from natural language descriptions
- Explore the mechanics of progressive diffusion and how images are built up layer by layer
- Configure structured layout coordinates to control object placement, scale, and relationships
- Master prompt decomposition techniques to separate background elements from foreground subjects
- Analyze modern regional guidance and attention-masking methods that keep objects visually distinct
You will start with the basic terminology of spatial planning in generative AI before moving on to practical workflows for structuring prompts and layout coordinates. The course guides you through the process of conceptualizing, planning, and refining complex multi-object scenes through clear, written explanations. Designed for beginners interested in the cutting edge of AI image generation, this course requires no prior coding or machine learning background. Start learning how to orchestrate complex AI-generated scenes with precision today.
ما الذي ستحصل عليه
-
📜
شهادة إتمام
أضفها إلى ملفك على LinkedIn -
💬
مدرّس AI شخصي
عالق في دورة؟ اسأل مدرّسك المدمج أي شيء، في أي وقت. -
🎧
النسخة الصوتية مضمَّنة
تعلَّم أثناء تنقُّلك — دون شاشة -
♾️
وصول مدى الحياة
عُد متى شئت، بلا انتهاء -
📱
الهاتف أو الكمبيوتر
يعمل في أي مكان وعلى أي جهاز -
💸
استرداد خلال 14 يومًا
دون أسئلة -
⚡
قصير ومركَّز
2 ساعة 36 دقيقة من المحتوى التطبيقي
المراجعات
لا توجد مراجعات بعد — كن أول من يشارك تجربته.
المتعلمون أخذوا أيضًا
🎓 بشهادة
الذكاء الاصطناعي الخاص مع برامج الماجستير في القانون مفتوحة المصدر: النشر المحلي، وRAG، والوكلاء
شهادة
تطبيق عملي
QR 50.00
→
💼 جاهز لسوق العمل
🎓 بشهادة
ضبط نماذج OpenAI: تخصيص نماذج اللغة الكبيرة ببياناتك الخاصة
شهادة
تطبيق عملي
QR 50.00
→
🏆 الأكثر شعبية
🎓 بشهادة
تطوير أنظمة RAG باستخدام Azure OpenAI و Azure AI Search
شهادة
تطبيق عملي
QR 50.00
→
💼 جاهز لسوق العمل
🎓 بشهادة
تطوير تطبيقات الذكاء الاصطناعي باستخدام LangChain
شهادة
تطبيق عملي
QR 50.00
→
الأسئلة الشائعة
ما الذي أحتاجه لأخذ هذه الدورة؟ +
يكفي هاتف أو كمبيوتر متصل بالإنترنت. بدون تثبيتات أو أجهزة خاصة.
كيف يمكنني الدفع؟ +
بالبطاقة عبر Stripe. لا نخزن بيانات البطاقة — يتولى Stripe ذلك بأمان.
هل يمكنني استرداد المال؟ +
نعم — استرداد كامل خلال 14 يومًا، دون أسئلة.
إلى متى يستمر وصولي؟ +
إلى الأبد. بمجرد الشراء، الدورة لك تعود إليها متى شئت.
هل سأحصل على شهادة؟ +
نعم. عند الإتمام ستحصل على شهادة يمكنك إضافتها إلى ملفك في LinkedIn.
مصمَّم للعاملين في
التقنية
التصميم
المالية
التسويق
الرعاية الصحية
التعليم
الضيافة
التصنيع