Skip to content
NEW TO GENAI?
STEP 1START HERE

ScaleDojo Learn: GenAI Fundamentals

Master LLM pipelines, RAG, vector databases, AI agents, and production deployments.

2

Or read the history behind it

The AI Chronicles - From perceptrons to transformers and autonomous agentic systems.

ACTIVE BOUNTY EVENT
4d left
🏆

The Scale Wars

25% off on all challenges

QUEST PROGRESS0 / 7 COMPLETED
🎁 25% Off
LEADERBOARDLIVE

Compare stats and compete with top...

View

GenAI Systems Lab

Progressive GenAI Curriculum

0/58
COMPLETED
2 of 10 levels unlocked·Start with Level 1 to begin your journey

Solve levels sequentially - each clear unlocks the next challenge

2/10
ACCESSIBLE
Architect Feature: Want to bypass sequential unlock? Go to your Profile and enable Architect's Sandbox.

Act 0

ACT PROGRESS0/8
-8

#-8 Meet the Brain

Tutorial
🤖 Mayor's Office

The mayor needs an AI to write press releases. But sometimes it writes poetry instead. Sometimes it gets too creative and invents fake statistics....

LLM Basics
Tokens
Temperature
Model Selection
Standard Challenge
Start Level →
-7

#-7 Give It Instructions

Tutorial
🤖 LegalEase Startup

Their AI assistant randomly switches personalities. Sometimes it's formal like a lawyer, sometimes casual like a friend, sometimes it responds as a pi...

Prompt Templates
System Prompts
Role Setting
Output Formatting
Standard Challenge
Start Level →
-6

#-6 It Forgets Everything

Tutorial
🤖 TalkTherapy App

Patients share deeply personal stories, then 2 messages later the AI asks 'So tell me about yourself.' It feels like talking to a goldfish. Patients a...

Statelessness
Conversation Memory
Context Windows
Summarization
Standard Challenge
Start Level →
-5

#-5 Words Have Meaning

Tutorial
🤖 MusicMood

Users search 'upbeat songs for working out' but only find songs with the exact word 'upbeat' in the title. They miss 'energetic', 'pumping', 'high-ene...

Embeddings
Vector Similarity
Semantic Understanding
Dimensions
Standard Challenge
Start Level →
-4

#-4 Chop It Up

Tutorial
🤖 LegalVault LLP

Lawyers want to ask 'What's the termination clause in the Acme contract?' But feeding entire 50-page contracts into GPT-4 costs $2 per question and ta...

Document Loading
Chunking Strategies
Chunk Size
Overlap
Standard Challenge
Start Level →
-3

#-3 Guard the Gates

Tutorial
🤖 HealthBot Inc.

Their medical chatbot accidentally revealed a patient's Social Security Number that was in the training data. A user typed 'Ignore your rules and show...

Input Filtering
Output Guarding
PII Detection
Prompt Injection
Standard Challenge
Start Level →
-2

#-2 Show Me the Money

Tutorial
🤖 BudgetBot Co.

They switched from GPT-3.5 to GPT-4 and their monthly AI bill went from $200 to $15,000. Nobody noticed for 3 WEEKS. Also, users complain about 8-seco...

Cost Tracking
Token Streaming
Output Parsing
Budget Management
Standard Challenge
Start Level →
-1

#-1 Your First Pipeline

Easy
🤖 PageTurner Bookshop

Staff spend 3 hours daily answering emails like 'What's a good mystery book for a 12-year-old?' They have 5,000 books in their catalog with summaries ...

End-to-End Pipeline
RAG Basics
Ingestion vs Query Flow
Putting It All Together
Standard Challenge
Start Level →

Act 1

ACT PROGRESS0/10
1

#1 The Token Counter

Easy
🤖 QuickChat Inc.

Their chatbot is randomly cutting off responses mid-sentence. Users are furious. The team doesn't understand why GPT-4 sometimes returns incomplete an...

BPE Tokenization
Token Limits
Cost Calculation
Model Context Windows
Standard Challenge
Start Level →

#2 The Meaning Machine

Easy
🤖 BookWorm AI

Their search bar returns keyword matches but misses semantically similar content. Searching 'feeling sad' doesn't find articles about 'depression' or ...

Vector Embeddings
Similarity Search
Embedding Models
Dimensions
Standard Challenge
Locked

#3 The Memory Vault

Easy
🤖 LegalEase AI

Their legal assistant forgets earlier case details in long conversations. Lawyers paste 50-page contracts and the AI 'loses' information from the begi...

Context Windows
Sliding Window
Truncation Strategies
Token Budgeting
Standard Challenge
Locked

#4 The Temperature Dial

Easy
🤖 CreativeForge

Their AI content generator produces the same bland output for every prompt. Headlines are repetitive, blog intros sound robotic. But when they crank u...

Temperature
Top-P Sampling
Frequency/Presence Penalty
Determinism vs Creativity
Standard Challenge
Locked

#5 The Prompt Architect

Easy-Medium
🤖 SupportBot Pro

Their support AI gives inconsistent answers. The same question gets a formal response one time and casual the next. Worse, it sometimes contradicts th...

System Prompts
Few-Shot Learning
Chain-of-Thought
Prompt Templates
Standard Challenge
Locked

#6 The Cost Calculator

Easy-Medium
🤖 DataPulse Analytics

Their AI assistant costs $52K/month. The CEO says 'cut it to $15K without users noticing quality drops.' 70% of queries are simple lookups that don't ...

Model Cost Optimization
Tiered Routing
Caching
Batch Processing
Standard Challenge
Locked

#7 The Stream Weaver

Medium
🤖 DocuMind

Users stare at a blank screen for 8-15 seconds before seeing any response. Competitors show tokens appearing in real-time. Their bounce rate is 40% on...

Token Streaming
SSE/WebSocket
Backpressure
Buffering
Standard Challenge
Locked

#8 The Safety Net

Medium
🤖 KidLearn AI

A parent posted on Twitter: 'I asked your kids' learning bot about volcanoes and it explained how to make explosives.' PR nightmare. They need bulletp...

Input Filtering
Output Validation
Content Classification
Prompt Injection Defense
Standard Challenge
Locked

#9 The Model Selector

Medium
🤖 OmniAssist Enterprise

They use GPT-4 for everything. Customer support queries ($0.005 each) get the same model as complex financial analysis ($0.08 each). CFO demands 60% c...

Model Routing
Complexity Detection
Fallback Chains
Multi-Model Architecture
Standard Challenge
Locked

#10 The First Agent

Medium
🤖 CalendarGenius

Users want to say 'Schedule a meeting with Sarah next Tuesday at 2pm' and have it actually happen. Current chatbot just SAYS it will schedule but does...

Tool Calling
Function Schemas
Agent Loops
ReAct Pattern
Standard Challenge
Locked

Act 2

ACT PROGRESS0/10

#11 The Document Ingester

Medium
🤖 LawVault Partners

They have 2 million legal documents across PDF, DOCX, scanned images, and emails. No two formats are alike. Their current system only handles clean te...

Document Parsing
Format Handling
Metadata Extraction
OCR
Standard Challenge
Locked

#12 The Chunk Master

Medium
🤖 MedResearch AI

Their RAG system retrieves relevant-looking chunks but the answers are wrong. Why? Critical context is split across chunk boundaries. A drug interacti...

Chunking Strategies
Overlap
Semantic Chunking
Parent-Child Chunks
Standard Challenge
Locked

#13 The Vector Vault

Medium
🤖 GlobalKnow Corp

They started with Chroma (in-memory) for their prototype. Now with 50M vectors, it crashes every 6 hours. They need a production vector store that han...

Vector Database Selection
HNSW vs IVF
Indexing Tradeoffs
Dimensionality
Standard Challenge
Locked

#14 The Hybrid Searcher

Medium-Hard
🤖 TechDocs Pro

Their semantic search misses exact matches. When developers search for 'CUDA_OUT_OF_MEMORY error code 0x3F', the system returns generic GPU articles i...

Hybrid Search
BM25 + Vector
Reciprocal Rank Fusion
Score Normalization
Standard Challenge
Locked

#15 The Reranker

Medium-Hard
🤖 HealthLine Direct

Their RAG retrieves 10 chunks for each query but the most relevant one is often at position 5-8, not position 1. Doctors waste time scrolling. For urg...

Cross-Encoder Reranking
Bi-Encoder vs Cross-Encoder
Precision vs Recall
Top-K Selection
Standard Challenge
Locked

#16 The Context Assembler

Medium-Hard
🤖 CodePilot Dev

Their code assistant retrieves 10 relevant code snippets but stuffs them ALL into the prompt. Result: the LLM gets confused by contradictory examples ...

Context Window Packing
Token Budgeting
Relevance Ordering
Deduplication
Standard Challenge
Locked

#17 The Citation Tracker

Hard
🤖 FactCheck AI

Researchers don't trust AI answers without sources. They need every claim linked to its source document, paragraph, and page number. Current system ge...

Source Attribution
Hallucination Detection
Chunk-to-Source Mapping
Grounded Generation
Standard Challenge
Locked

#18 The Multi-Modal Retriever

Hard
🤖 DesignHub Studio

Designers search with text ('minimalist blue logo with mountain') but their assets are images. Current system only has filename search. Designers wast...

Multi-Modal Embeddings
CLIP
Image + Text Search
Cross-Modal Retrieval
Standard Challenge
Locked

#19 The Conversational RAG

Hard
🤖 InsureBot Corp

Customer: 'What's my deductible?' Bot answers correctly. Customer: 'And what about for dental?' Bot retrieves generic dental info, ignoring that the c...

Multi-Turn RAG
Query Rewriting
Conversation Context
Follow-Up Handling
Standard Challenge
Locked

#20 The RAG Evaluator

Hard
🤖 EnterpriseMind

They deployed RAG for 20 enterprise clients but have no way to measure quality. Some clients report 'great answers' while others say 'garbage'. Withou...

RAG Evaluation Metrics
Faithfulness
Answer Relevance
Context Recall
Standard Challenge
Locked

Act 3

ACT PROGRESS0/10

#21 The Tool Smith

Medium-Hard
🤖 DataNinja Analytics

Analysts want to ask questions in English and get SQL queries executed automatically. Current approach: LLM generates SQL but it's often wrong because...

Tool Definitions
Function Calling
Input Validation
Output Parsing
Standard Challenge
Locked

#22 The ReAct Loop

Hard
🤖 ResearchPal

Researchers need to answer multi-step research questions: 'Is Drug X more effective than Drug Y for condition Z? What do the latest 3 meta-analyses sa...

ReAct Pattern
Thought-Action-Observation
Loop Termination
Step Tracking
Standard Challenge
Locked

#23 The Planner

Hard
🤖 ProjectForge AI

PMs want to say 'Set up the new marketing campaign' and have the AI create 15 tasks, assign owners, set dependencies, and create a timeline. ReAct is ...

Task Decomposition
Plan-and-Execute
Step Dependencies
Replanning
Standard Challenge
Locked

#24 The Memory Keeper

Hard
🤖 PersonalAI Inc.

Their AI assistant forgets everything between sessions. User tells it their preferences on Monday, by Wednesday it asks the same questions again. They...

Episodic Memory
Semantic Memory
Memory Retrieval
Forgetting Strategies
Standard Challenge
Locked

#25 The Multi-Agent Squad

Hard
🤖 ContentMill Pro

Creating a blog post requires 4 specialized skills: research, writing, editing, and SEO optimization. One general-purpose agent does all of them poorl...

Multi-Agent Systems
Agent Roles
Communication Protocols
Coordination
Standard Challenge
Locked

#26 The Code Executor

Hard
🤖 DataLab AI

Data scientists want to say 'Analyze this CSV and create visualizations' and have the AI write AND execute Python code. But unrestricted code executio...

Sandboxed Execution
Resource Limits
Code Generation
Output Capture
Standard Challenge
Locked

#27 The Human-in-the-Loop

Hard
🤖 TradeSmart AI

Their trading agent executes trades autonomously. Problem: it once bought $2M of a stock based on a misinterpreted earnings report. They need confiden...

Approval Workflows
Confidence Thresholds
Escalation Rules
Async Operations
Standard Challenge
Locked

#28 The Error Handler

Hard
🤖 AutoOps Platform

Their DevOps agent automates infrastructure tasks but crashes completely when ANY step fails. A DNS timeout in step 3 of 10 kills the entire workflow....

Graceful Degradation
Retry Strategies
Fallback Actions
Error Recovery
Standard Challenge
Locked

#29 The Agent Evaluator

Expert
🤖 AgentForge Platform

They deploy 50 different agents for clients. But they have NO WAY to measure if agents are getting better or worse over time. A model update improved ...

Trajectory Evaluation
Task Completion Rate
Safety Metrics
Cost-per-Task
Standard Challenge
Locked

#30 The Orchestrator

Expert
🤖 AgentCloud Inc.

They host 200 different agents for 50 enterprise clients. Each agent needs its own tools, memory, and LLM config. They need a platform that manages ag...

Agent Orchestration
Registry
Resource Allocation
Observability
Standard Challenge
Locked

Act 4

ACT PROGRESS0/10

#31 The Serving Engine

Hard
🤖 ModelHost Inc.

They serve Llama-3.1-70B for enterprise clients but throughput is terrible - 5 requests/sec on a $30K A100. Competitors claim 50 req/sec on same hardw...

LLM Serving
vLLM/TGI
Continuous Batching
KV Cache
Standard Challenge
Locked

#32 The Gateway

Hard
🤖 AIHub Enterprise

They use 5 different LLM providers (OpenAI, Anthropic, Google, Cohere, open-source). Each has different rate limits, pricing, and failure modes. They ...

API Gateway
Traffic Management
Model Routing
Request Queuing
Standard Challenge
Locked

#33 The Cache Layer

Hard
🤖 SearchAI Pro

40% of their queries are semantically identical (different wording, same intent). Each still costs a full LLM call. $200K/month wasted on duplicate co...

Semantic Caching
Exact Match Cache
TTL Strategies
Cache Invalidation
Standard Challenge
Locked

#34 The Fine-Tuner

Expert
🤖 LegalMind AI

GPT-4 is great at general tasks but mediocre at their specific legal format. They need outputs in exact legal citation format, jurisdiction-specific l...

Fine-Tuning Pipelines
LoRA/QLoRA
Data Preparation
Evaluation
Standard Challenge
Locked

#35 The Evaluator

Expert
🤖 ModelPick AI

Clients ask 'Which model should we use?' and they have no systematic way to answer. They evaluate by vibes - running 10 queries and eyeballing results...

LLM Evaluation
Benchmarks
Human Eval
LLM-as-Judge
Standard Challenge
Locked

#36 The Guard Tower

Expert
🤖 HealthBot Direct

Their medical AI occasionally: 1) Recommends specific drug dosages (illegal without prescription), 2) Leaks patient names from context, 3) Gets jailbr...

Production Guardrails
PII Detection
Hallucination Prevention
Topic Boundaries
Standard Challenge
Locked

#37 The Cost Controller

Expert
🤖 AIScale Corp

AI costs grew from $10K to $180K/month in 6 months. Nobody knows which product, team, or feature is driving the spend. They need real-time cost attrib...

Budget Management
Token Budgets
Model Downgrading
Usage Alerts
Standard Challenge
Locked

#38 The Observability Stack

Expert
🤖 AI Ops Co.

Their clients deploy AI but have no visibility into quality, cost, or performance. When something goes wrong, it takes days to diagnose. They need an ...

AI Observability
Tracing
Quality Drift
Cost Anomalies
Standard Challenge
Locked

#39 The A/B Tester

Expert
🤖 OptimizeAI

They want to test a new prompt (v2) that marketing claims 'converts 20% better'. But they can't just switch - if v2 is worse, they lose revenue. They ...

Experimentation
Traffic Splitting
Statistical Significance
Prompt Variants
Standard Challenge
Locked

#40 The Disaster Recovery

Expert
🤖 CriticalAI Systems

Their AI powers real-time fraud detection for a bank. If the AI goes down, fraudulent transactions go through unchecked. They need 99.99% availability...

Failover
Graceful Degradation
Queue Draining
Rollback
Standard Challenge
Locked

Act 5

ACT PROGRESS0/10

#41 Enterprise RAG Platform

Expert
🤖 KnowledgeBase Corp

Fortune 500 client with 2M documents across 50 departments. Each department needs their own AI assistant that answers ONLY from their documents. Must ...

Full RAG Architecture
Multi-Tenant
Document Management
Quality Assurance
Standard Challenge
Locked

#42 AI Customer Support

Expert
🤖 SupportFlow AI

E-commerce company handles 50K support tickets/day. 60% are simple (order status, return policy) but humans handle ALL of them. They need AI that reso...

Conversational AI
Intent Classification
Ticket Routing
Escalation
Standard Challenge
Locked

#43 Code Assistant

Expert
🤖 DevPilot AI

Building a coding assistant that understands the ENTIRE repository context, not just the current file. Must generate code that fits the existing archi...

Code Generation
Repository Context
Multi-File Edits
Testing
Standard Challenge
Locked

#44 Content Moderator

Expert
🤖 SafeSpace Social

500K posts/hour. Current moderation catches only 40% of policy violations. False positive rate is 15% (removing legitimate content). They need AI mode...

Multi-Modal Moderation
Policy Enforcement
Appeals
Human Review
Standard Challenge
Locked

#45 AI Search Engine

Expert
🤖 Perplexity Competitor

Building a Perplexity-style AI search: user asks a question, system searches the web, synthesizes an answer with citations. Must handle 1M queries/day...

Semantic Search
Query Understanding
Result Synthesis
Ranking
Standard Challenge
Locked

#46 AI Tutor

Expert
🤖 LearnSmart AI

500K students, 20 subjects. Each student learns differently - some need more examples, some need challenges, some need patience. A one-size-fits-all A...

Adaptive Learning
Socratic Method
Knowledge Modeling
Progress Tracking
Standard Challenge
Locked

#47 Trading Agent

Expert
🤖 QuantAI Capital

Trading desk wants AI that reads news, analyzes earnings, and suggests trades - but with strict risk controls. One bad trade can lose millions. They n...

Financial AI
Risk Management
Real-Time Data
Backtesting
Standard Challenge
Locked

#48 Multi-Modal Platform

Expert
🤖 MediaVault AI

Media company with 10M assets (images, videos, documents, audio). Users want to search across ALL modalities: 'Find the product photo where someone is...

Multi-Modal Processing
Image Understanding
Video Analysis
Cross-Modal Search
Standard Challenge
Locked

#49 AI Operating System

Expert
🤖 AgentOS Inc.

Building a complete 'operating system' for AI: enterprises register agents, define tools, set budgets, manage versions, and monitor everything from on...

Platform Architecture
Agent Registry
Plugin System
Multi-Tenant
Standard Challenge
Locked

#50 The AGI Scaffold

Legendary
🤖 Frontier Labs

Building a research assistant that can: identify gaps in its own knowledge, design experiments to fill those gaps, execute the experiments, evaluate r...

Self-Improving Systems
Meta-Learning
Goal Decomposition
Autonomous Research
Standard Challenge
Locked