Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Risingwavelabs Risingwave Docker Deployment
- Workflow:Apache Paimon Lance Format Analytics
- Workflow:Predibase Lorax Server Deployment
- Workflow:Recommenders team Recommenders SAR Collaborative Filtering
- Workflow:Sgl project Sglang Online Serving With OpenAI API
- Workflow:Huggingface Optimum FX Graph Optimization
- Workflow:MaterializeInc Materialize dbt Integration
- Workflow:Sktime Pytorch forecasting TFT Hyperparameter Optimization
- Workflow:Triton inference server Server Model Performance Tuning
- Workflow:Google research Deduplicate text datasets Suffix array querying
Principles
- Principle:TA Lib Ta lib python Streaming Data Buffering
- Principle:Apache Shardingsphere Persistence Facade Construction
- Principle:Heibaiying BigData Notes Kafka Rebalancing and Shutdown
- Principle:ClickHouse ClickHouse Test Result Analysis
- Principle:Neuml Txtai Sparse Retrieval
- Principle:Huggingface Optimum Pipeline Inference Execution
- Principle:Trailofbits Fickling Severity Classification
- Principle:Tensorflow Serving Canary Deployment
- Principle:Getgauge Taiko Text Field Selection
- Principle:Apache Shardingsphere Runtime Rule Object Initialization
Implementations
- Implementation:Datajuicer Data juicer PartitionSizeOptimizer Calculate
- Implementation:LMCache LMCache Standalone Starter
- Implementation:Iterative Dvc Utils Core Helpers
- Implementation:Haosulab ManiSkill StereoDepthCamera
- Implementation:Arize ai Phoenix TracerProvider Instrumentation
- Implementation:Anthropics Anthropic sdk python BetaRawMessageDeltaEvent
- Implementation:Zai org CogVideo DiagonalGaussianDistribution
- Implementation:Lakeraai Pint benchmark Pint Benchmark Function
- Implementation:Alibaba ROLL HomeContent Component
- Implementation:Langchain ai Langchain AIMessageChunk Accumulation
Heuristics
- Heuristic:Roboflow Rf detr Small Dataset Oversampling
- Heuristic:Microsoft LoRA Warning Deprecated Legacy Examples
- Heuristic:Allenai Open instruct Gradient Clipping Norm
- Heuristic:Apache Spark Build Fallback Strategies
- Heuristic:PeterL1n BackgroundMattingV2 Mixed Precision Training
- Heuristic:Junyanz Pytorch CycleGAN and pix2pix Instance Norm for Multi GPU
- Heuristic:DistrictDataLabs Yellowbrick Elbow Knee Detection Sensitivity
- Heuristic:Facebookresearch Habitat lab Force Single Threaded PyTorch
- Heuristic:Google deepmind Dm control MJCF Model Composition Gotchas
- Heuristic:Haotian liu LLaVA Flash Attention GPU Requirement
Environments
- Environment:Vllm project Vllm CUDA
- Environment:Alibaba ROLL CUDA GPU Environment
- Environment:Mlflow Mlflow GPU System Metrics Environment
- Environment:LLMBook zh LLMBook zh github io PyTorch CUDA GPU Environment
- Environment:Microsoft Autogen Studio Server Environment
- Environment:CarperAI Trlx NeMo Megatron
- Environment:Trailofbits Fickling PyTorch
- Environment:Google deepmind Dm control GLFW Desktop Rendering
- Environment:Ucbepic Docetl Docker Deployment
- Environment:Ggml org Llama cpp CUDA GPU Environment