Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Spcl Graph of thoughts GoT Keyword Counting Pipeline
- Workflow:Neuml Txtai Model Training
- Workflow:Lucidrains X transformers DPO Preference Alignment
- Workflow:Microsoft DeepSpeedExamples CIFAR10 Getting Started
- Workflow:Huggingface Alignment handbook SFT DPO Alignment Pipeline
- Workflow:Astronomer Astronomer cosmos Local dbt DAG rendering
- Workflow:NVIDIA NeMo Aligner REINFORCE Training
- Workflow:MaterializeInc Materialize Upgrade Testing
- Workflow:Scikit learn Scikit learn Data Preprocessing Pipeline
- Workflow:Explodinggradients Ragas Agent Evaluation
Principles
- Principle:Zai org CogVideo Video Encoding
- Principle:ARISE Initiative Robosuite Action Vector Construction
- Principle:Dagster io Dagster Time Based Partitioning
- Principle:Apache Flink Cross Source Checkpointing
- Principle:Scikit learn Scikit learn Baseline Models
- Principle:Microsoft DeepSpeedExamples Model Evaluation And Export
- Principle:Hpcaitech ColossalAI SFT Training Execution
- Principle:ARISE Initiative Robosuite Object Grouping
- Principle:Facebookresearch Habitat lab Task Dataset Selection
- Principle:TobikoData Sqlmesh Web UI
Implementations
- Implementation:SeldonIO Seldon core Open Inference Protocol V2 OpenAPI
- Implementation:Apache Hudi SchemaChangeUtils IsTypeUpdateAllow
- Implementation:Microsoft Agent framework Run All Samples Script
- Implementation:Google deepmind Mujoco Engine Setconst
- Implementation:NVIDIA NeMo Curator DataDesignerStage
- Implementation:Lm sys FastChat OpenAI API Server
- Implementation:Openai CLIP Zeroshot Classifier
- Implementation:PeterL1n BackgroundMattingV2 Camera
- Implementation:Mit han lab Llm awq NVILAQwen2
- Implementation:Microsoft DeepSpeedExamples Net DeepSpeed
Heuristics
- Heuristic:Kubeflow Kubeflow Stale Issue Lifecycle Management
- Heuristic:Speechbrain Speechbrain Nonfinite Loss Handling
- Heuristic:PeterL1n BackgroundMattingV2 Data Augmentation Strategy
- Heuristic:Deepspeedai DeepSpeed ZeRO Pipeline Incompatibility
- Heuristic:Obss Sahi Confidence Threshold Setting
- Heuristic:Treeverse LakeFS Action Cache Wait Tip
- Heuristic:Fede1024 Rust rdkafka Commit Mode Sync Vs Async
- Heuristic:Facebookresearch Audiocraft Generation Parameter Defaults
- Heuristic:AUTOMATIC1111 Stable diffusion webui GTX 16 Series FP16 Workaround
- Heuristic:FlagOpen FlagEmbedding Dynamic Batch Size Reduction
Environments
- Environment:Protectai Llm guard ONNX Runtime Acceleration
- Environment:FlagOpen FlagEmbedding Finetuning Environment
- Environment:Duckdb Duckdb Code Generation Tools
- Environment:Infiniflow Ragflow GPU CUDA Environment
- Environment:Mit han lab Llm awq CUDA Build Environment
- Environment:Tensorflow Serving Kubernetes Deployment Environment
- Environment:Pytorch Serve vLLM Engine Environment
- Environment:Online ml River Build Toolchain
- Environment:Pyro ppl Pyro Distributed Training
- Environment:Google research Deduplicate text datasets Python HuggingFace Environment