Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Ggml org Llama cpp Multimodal Inference
- Workflow:Mlc ai Mlc llm Disaggregated Serving
- Workflow:LLMBook zh LLMBook zh github io Data Preprocessing Pipeline
- Workflow:AUTOMATIC1111 Stable diffusion webui Checkpoint merging
- Workflow:Axolotl ai cloud Axolotl Full Finetuning Distributed
- Workflow:Run llama Llama index OpenAI LLM Finetuning
- Workflow:Microsoft Semantic kernel Kernel Setup And Chat Completion
- Workflow:PeterL1n BackgroundMattingV2 Image matting inference
- Workflow:Huggingface Trl PPO RLHF Training
- Workflow:SeldonIO Seldon core AB Testing Experiment
Principles
- Principle:Sktime Pytorch forecasting MLP Decoder
- Principle:Online ml River Online ML Utilities
- Principle:Princeton nlp SimPO Response Post Processing
- Principle:Huggingface Diffusers Prior Preservation
- Principle:Mlflow Mlflow Prompt Template Design
- Principle:Volcengine Verl Multimodal Data Preparation
- Principle:Teamcapybara Capybara RSpec Integration
- Principle:Huggingface Transformers Commit Bisection Debugging
- Principle:Isaac sim IsaacGymEnvs Automatic Domain Randomization
- Principle:Danijar Dreamerv3 Distributed Environment Execution
Implementations
- Implementation:Datahub project Datahub StreamingDataSourceV2RelationVisitor
- Implementation:Allenai Open instruct Mason Main
- Implementation:Infiniflow Ragflow Request Client
- Implementation:Apache Paimon RESTCatalogOptions
- Implementation:NVIDIA TransformerEngine JAX XLA GEMM
- Implementation:InternLM Lmdeploy Gemm TiledMma
- Implementation:Norrrrrrr lyn WAInjectBench load detector image
- Implementation:Speechbrain Speechbrain Train AISHELL1 Transformer
- Implementation:Microsoft Playwright DefaultFontFamilies
- Implementation:Predibase Lorax Response Format Type
Heuristics
- Heuristic:Intel Ipex llm LoRA Target All Linear Layers
- Heuristic:Openclaw Openclaw Cache TTL Asymmetric Strategy
- Heuristic:Apache Beam Lock Contention Batching
- Heuristic:Run llama Llama index Batch Eval Retry Strategy
- Heuristic:LMCache LMCache CacheGen Quantization Strategy
- Heuristic:EvolvingLMMs Lab Lmms eval Distributed Padding Strategy
- Heuristic:Fastai Fastbook Mixup Data Augmentation
- Heuristic:Googleapis Python genai API Retry Backoff Strategy
- Heuristic:Google deepmind Mujoco MJX Feature Compatibility
- Heuristic:Huggingface Datatrove Gopher Quality Thresholds
Environments
- Environment:Microsoft BIPIA DeepSpeed Finetuning Environment
- Environment:Dotnet Machinelearning TorchSharp Environment
- Environment:MaterializeInc Materialize Dbt Materialize Runtime
- Environment:Trailofbits Fickling Python Runtime
- Environment:Snorkel team Snorkel SpaCy NLP
- Environment:Huggingface Datasets Audio Video Dependencies
- Environment:Huggingface Open r1 vLLM Server
- Environment:NVIDIA TransformerEngine CUDA Toolkit Requirements
- Environment:Mlc ai Mlc llm CUDA GPU Environment
- Environment:Treeverse LakeFS Spark GC Environment