Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Apache Hudi Flink Batch Incremental Read
- Workflow:Sgl project Sglang Structured Output Generation
- Workflow:Explodinggradients Ragas Prompt Evaluation And Iteration
- Workflow:Bigscience workshop Petals Distributed Text Generation
- Workflow:Eventual Inc Daft Data Lakehouse ETL
- Workflow:Wandb Weave LLM Integration Tracing
- Workflow:DataExpert io Data engineer handbook Flink Kafka Streaming Pipeline
- Workflow:Danijar Dreamerv3 Distributed Parallel Training
- Workflow:Evidentlyai Evidently LLM Evaluation Monitoring
- Workflow:Speechbrain Speechbrain Speaker Embedding Training
Principles
- Principle:FMInference FlexLLMGen Offloaded Model Loading
- Principle:BerriAI Litellm Database Setup
- Principle:Cleanlab Cleanlab Clean Model Inference
- Principle:Turboderp org Exllamav2 Sampling Configuration
- Principle:Ray project Ray Actor Termination
- Principle:Ggml org Llama cpp Server Startup
- Principle:Unstructured IO Unstructured Embedding Provider Interface
- Principle:Ggml org Llama cpp Vocabulary System
- Principle:Haosulab ManiSkill LeRobot Format Export
- Principle:Apache Beam Twister2 Execution and Result Collection
Implementations
- Implementation:Recommenders team Recommenders SARSingleNode Init
- Implementation:Microsoft DeepSpeedExamples Create HF Model
- Implementation:Triton inference server Server GenQaRaggedModels
- Implementation:Haifengl Smile SVD EVD Results
- Implementation:Huggingface Alignment handbook TrlParser Parse Args And Config
- Implementation:Shiyu coder Kronos WebUI App
- Implementation:Elevenlabs Elevenlabs python ConversationConfigClientOverrideConfigInput
- Implementation:EvolvingLMMs Lab Lmms eval LongVT Utils
- Implementation:Risingwavelabs Risingwave PostgresDialect
- Implementation:Interpretml Interpret PartitionRandomBoosting
Heuristics
- Heuristic:LLMBook zh LLMBook zh github io IGNORE INDEX Loss Masking
- Heuristic:Microsoft LoRA Selective LoRA QV Only
- Heuristic:Ggml org Ggml Sampling Parameter Defaults
- Heuristic:PrefectHQ Prefect Task Timeout Thread Limitation
- Heuristic:ContextualAI HALOs TF32 Matmul Acceleration
- Heuristic:Kubeflow Kubeflow Sequential Infrastructure Deployment
- Heuristic:Huggingface Peft Gradient Checkpointing With Quantization
- Heuristic:Ggml org Llama cpp Context Size Alignment
- Heuristic:Apache Shardingsphere Shadow Routing Hint First Fallback
- Heuristic:ArroyoSystems Arroyo Worker Heartbeat Timeout
Environments
- Environment:Pytorch Serve vLLM Engine Environment
- Environment:Evidentlyai Evidently SQL Storage Environment
- Environment:Isaac sim IsaacGymEnvs Pip Dependencies
- Environment:ArroyoSystems Arroyo Webui Runtime
- Environment:Liu00222 Open Prompt Injection Python Dependencies
- Environment:Openai Openai agents python MCP Dependencies
- Environment:Langfuse Langfuse S3 Compatible Storage
- Environment:Marker Inc Korea AutoRAG Python 3 10 Runtime
- Environment:Deepseek ai Janus JanusFlow Diffusers Environment
- Environment:Nightwatchjs Nightwatch Android Mobile Testing