Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Zai org CogVideo Video Editing DDIM Inversion
- Workflow:Openai Whisper Language Detection And Decoding
- Workflow:Explodinggradients Ragas Test Data Generation
- Workflow:Apache Dolphinscheduler Workflow Failover Recovery
- Workflow:Getgauge Taiko Interactive Test Recording
- Workflow:VainF Torch Pruning Vision Transformer Pruning
- Workflow:Neuml Txtai Pipeline Workflow Chaining
- Workflow:Alibaba ROLL Agentic RL Training Pipeline
- Workflow:Confident ai Deepeval AI Agent Evaluation
- Workflow:NVIDIA NeMo Curator Semantic Deduplication
Principles
- Principle:Scikit learn contrib Imbalanced learn Sampler Aware Pipeline
- Principle:OpenRLHF OpenRLHF Sequence Regression Model Loading
- Principle:Tencent Ncnn Vulkan Build Configuration
- Principle:Ggml org Ggml Vectorized Math Operations
- Principle:Iterative Dvc Data Structure Utilities
- Principle:Heibaiying BigData Notes HBase Table Creation
- Principle:LLMBook zh LLMBook zh github io SFT Model Loading
- Principle:Dagster io Dagster Rate Limiting Strategy
- Principle:Gretelai Gretel synthetics WGAN GP Training
- Principle:Interpretml Interpret Local Explanation Generation
Implementations
- Implementation:Apache Paimon DlfProvider
- Implementation:Mlc ai Web llm Embedding Model Config
- Implementation:Ollama Ollama MLXRunner Cache
- Implementation:Langfuse Langfuse NormalizeInputOutput
- Implementation:Tencent Ncnn RFCN Example
- Implementation:Rapidsai Cuml Genetic Node
- Implementation:Open compass VLMEvalKit CCOCR OCR Evaluator
- Implementation:Haifengl Smile Streaming Prediction API
- Implementation:Pyro ppl Pyro IndepMessenger
- Implementation:Recommenders team Recommenders A2SVD Model
Heuristics
- Heuristic:Nautechsystems Nautilus trader Inflight Order Check Threshold
- Heuristic:Deepspeedai DeepSpeed Shared Memory Sizing
- Heuristic:Norrrrrrr lyn WAInjectBench Zero Vector Fallback Failed Embeddings
- Heuristic:ARISE Initiative Robomimic BatchNorm To GroupNorm For EMA
- Heuristic:Openai Openai node Stream Usage Interruption
- Heuristic:Allenai Open instruct BFloat16 Training
- Heuristic:HKUDS AI Trader Market Type Auto Detection
- Heuristic:Danijar Dreamerv3 Percentile Return Normalization
- Heuristic:Openai Whisper SDPA Disabling For Attention Extraction
- Heuristic:OpenRLHF OpenRLHF Off Policy IS Correction Tip
Environments
- Environment:Alibaba MNN GPU OpenCL Environment
- Environment:Spcl Graph of thoughts Local LLaMA GPU Inference
- Environment:Ggml org Ggml Vulkan GPU Environment
- Environment:Microsoft Autogen Studio Server Environment
- Environment:Sktime Pytorch forecasting Matplotlib Plotting Dependencies
- Environment:Treeverse LakeFS LakeFS Server Environment
- Environment:Pytorch Serve vLLM Engine Environment
- Environment:Sgl project Sglang Kubernetes
- Environment:Fastai Fastbook Sklearn Environment
- Environment:ThreeSR Awesome Inference Time Scaling GitHub Account Environment