Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Kubeflow Pipelines XGBoost Training Pipeline
- Workflow:OpenRLHF OpenRLHF DPO Training
- Workflow:Mlfoundations Open flamingo Data Preparation
- Workflow:Protectai Modelscan Programmatic Model Scanning
- Workflow:ChenghaoMou Text dedup SimHash Deduplication
- Workflow:NVIDIA TransformerEngine FSDP Distributed Training
- Workflow:Danijar Dreamerv3 Distributed Parallel Training
- Workflow:Avdvg InjectGuard Vector Similarity Detection Pipeline
- Workflow:Mlc ai Web llm Function Calling
- Workflow:Iamhankai Forest of Thought CGDM Post Processing
Principles
- Principle:Mit han lab Llm awq Multimodal Vision Language Inference
- Principle:Dotnet Machinelearning Trial Result Extraction
- Principle:Bitsandbytes foundation Bitsandbytes Paged Optimizer
- Principle:Huggingface Datasets DatasetDict Hub Upload
- Principle:Lm sys FastChat ShareGPT HTML Cleaning
- Principle:NVIDIA TransformerEngine FP8 Delayed Scaling
- Principle:Allenai Open instruct Supervised Finetuning
- Principle:Huggingface Datasets Semantic Versioning
- Principle:Nautechsystems Nautilus trader Exchange Adapter Configuration
- Principle:Mit han lab Llm awq Quantized Linear Module
Implementations
- Implementation:Recommenders team Recommenders BaseModel Run Fast Eval
- Implementation:SeleniumHQ Selenium ChromeDriverService CreateDefaultService
- Implementation:Scikit learn contrib Imbalanced learn CondensedNearestNeighbour
- Implementation:Google deepmind Dm control Suite Humanoid CMU
- Implementation:Datajuicer Data juicer TsvFormatter
- Implementation:OpenBMB UltraFeedback Inference Environment Setup
- Implementation:NVIDIA NeMo Curator DomainClassifier
- Implementation:Puppeteer Puppeteer CDPSession
- Implementation:TobikoData Sqlmesh LoadingContainer
- Implementation:Mlc ai Mlc llm Template Registry
Heuristics
- Heuristic:AnswerDotAI RAGatouille Auto Batch Size For Long Documents
- Heuristic:Explodinggradients Ragas Concurrency And Rate Limiting
- Heuristic:FMInference FlexLLMGen OOM Memory Management
- Heuristic:Elevenlabs Elevenlabs python TTS Model Selection
- Heuristic:Pyro ppl Pyro MCMC Warmup Adaptation
- Heuristic:Princeton nlp Tree of thought llm Functools Partial Model Binding
- Heuristic:Scikit learn Scikit learn Random State Management
- Heuristic:Predibase Lorax Warning Deprecated BitsAndBytes 8bit
- Heuristic:ClickHouse ClickHouse Jemalloc Production Requirement
- Heuristic:Tensorflow Serving GPU Memory And CPU Optimization
Environments
- Environment:Haosulab ManiSkill GPU CUDA Simulation
- Environment:Shiyu coder Kronos HuggingFace Hub Access
- Environment:Deepspeedai DeepSpeed CPU Environment
- Environment:Zai org CogVideo Diffusers Inference Environment
- Environment:OWASP Www project top 10 for large language model applications Pydantic Invoice Agent Runtime
- Environment:InternLM Lmdeploy CUDA GPU Runtime
- Environment:Mlc ai Mlc llm WebGPU Browser Environment
- Environment:LLMBook zh LLMBook zh github io PyTorch CUDA GPU Environment
- Environment:Langfuse Langfuse PostgreSQL 17
- Environment:Cleanlab Cleanlab Image Quality Dependencies