Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Trailofbits Fickling PyTorch Payload Injection
- Workflow:Hpcaitech ColossalAI RAG Application
- Workflow:Sgl project Sglang Offline Batch Inference
- Workflow:Huggingface Transformers Pipeline Inference
- Workflow:Pyro ppl Pyro VAE Training
- Workflow:Treeverse LakeFS Write Audit Publish With Hooks
- Workflow:NVIDIA DALI Image Classification Training PyTorch
- Workflow:Google research Deduplicate text datasets Single file deduplication
- Workflow:Protectai Modelscan CLI Model Scanning
- Workflow:Confident ai Deepeval End to End LLM Evaluation
Principles
- Principle:Openai Whisper Audio Padding And Trimming
- Principle:Ggml org Ggml CPU Tensor Operations
- Principle:Interpretml Interpret Regression Performance
- Principle:Farama Foundation Gymnasium Vectorized Environment Creation
- Principle:Microsoft Semantic kernel Orchestration Runtime
- Principle:Triton inference server Server API Access Restriction
- Principle:Google deepmind Mujoco CPU Model Loading
- Principle:Microsoft Semantic kernel Native Plugin Definition
- Principle:Pola rs Polars Lazy Query Collection
- Principle:Guardrails ai Guardrails LLM Validation Wrapping
Implementations
- Implementation:OpenRLHF OpenRLHF RewardModelTrainer
- Implementation:Mage ai Mage ai Source Discover
- Implementation:SeleniumHQ Selenium Closure UI SelectionModel
- Implementation:Recommenders team Recommenders LSTUR Model
- Implementation:Microsoft Onnxruntime CUDA SoftmaxGrad
- Implementation:Arize ai Phoenix Client Executors
- Implementation:LLMBook zh LLMBook zh github io Apply Rotary Pos Emb
- Implementation:Astronomer Astronomer cosmos Package Exports
- Implementation:Scikit learn Scikit learn PrecisionRecallDisplay
- Implementation:BerriAI Litellm Least Busy Strategy
Heuristics
- Heuristic:ContextualAI HALOs FSDP Sampling Workaround
- Heuristic:Alibaba ROLL Dynamic Batching Token Limits
- Heuristic:Axolotl ai cloud Axolotl Gradient Checkpointing Reentrant Rules
- Heuristic:Microsoft Onnxruntime Flash Attention Optimization
- Heuristic:Eventual Inc Daft Delta Lake S3 Locking
- Heuristic:Datajuicer Data juicer Checkpoint Resumption Strategy
- Heuristic:Guardrails ai Guardrails Async Vs Sync Validation Mode
- Heuristic:Fede1024 Rust rdkafka Transaction Error Recovery
- Heuristic:ClickHouse ClickHouse Test Writing Conventions
- Heuristic:Fastai Fastbook Mixup Data Augmentation
Environments
- Environment:Neuml Txtai Python Core Environment
- Environment:FMInference FlexLLMGen CUDA GPU
- Environment:Apache Spark Release Build Environment
- Environment:Deepset ai Haystack GPU Device Environment
- Environment:Iamhankai Forest of Thought Python CUDA Runtime
- Environment:Huggingface Alignment handbook Evaluation Tools
- Environment:Huggingface Diffusers PyTorch CUDA Runtime
- Environment:Openai Openai node Node 20 Runtime
- Environment:NVIDIA NeMo Curator Python Linux Base
- Environment:OpenBMB UltraFeedback OpenAI API Environment