Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Microsoft Semantic kernel Agent Conversation And Orchestration
- Workflow:Volcengine Verl GRPO Training Pipeline
- Workflow:Explodinggradients Ragas Test Data Generation
- Workflow:Google deepmind Mujoco Model compilation and conversion
- Workflow:NVIDIA NeMo Aligner REINFORCE Training
- Workflow:Apache Shardingsphere Dynamic Rule Configuration Change
- Workflow:Sail sg LongSpec GLIDE Draft Model Training
- Workflow:Scikit learn Scikit learn Cross Validation Evaluation
- Workflow:Promptfoo Promptfoo LLM Evaluation
- Workflow:PrefectHQ Prefect Asset Based Data Pipeline
Principles
- Principle:Duckdb Duckdb Fast Number Parsing
- Principle:Cleanlab Cleanlab Multilabel Label Issue Filtering
- Principle:Mlc ai Web llm Chat Inference
- Principle:Bentoml BentoML Bento Artifact Management
- Principle:Dotnet Machinelearning Experiment Configuration
- Principle:Groq Groq python Batch Results Retrieval
- Principle:CARLA simulator Carla Simulation Configuration
- Principle:Sail sg LongSpec VLLM Inference Client
- Principle:Scikit learn contrib Imbalanced learn Borderline Oversampling
- Principle:Apache Beam Job Submission Twister2
Implementations
- Implementation:NVIDIA TransformerEngine NVFP4Tensor
- Implementation:Online ml River Drift KSWIN
- Implementation:Openclaw Openclaw UpdateCommand
- Implementation:PrefectHQ Prefect AI Cleanup Agent
- Implementation:CarperAI Trlx Logging
- Implementation:FMInference FlexLLMGen DeepSpeed BF16 Optimizer
- Implementation:Hpcaitech ColossalAI Prepare Dataset SFT
- Implementation:Microsoft Playwright Roll Browser
- Implementation:Openai Openai python Module Level Client
- Implementation:Apache Paimon AuthProviderFactory
Heuristics
- Heuristic:Bigscience workshop Petals Randomized Rebalancing Intervals
- Heuristic:Kornia Kornia Morphology Engine Selection
- Heuristic:Facebookresearch Audiocraft Generation Parameter Defaults
- Heuristic:Google deepmind Dm control Prop Settling Physics Tuning
- Heuristic:TA Lib Ta lib python Thread Safety With Abstract API
- Heuristic:Huggingface Datasets Batch Size Optimization
- Heuristic:Neuml Txtai Model Quantization Defaults
- Heuristic:Huggingface Alignment handbook DPO Beta Selection
- Heuristic:ChenghaoMou Text dedup SimHash Optimization Ceiling
- Heuristic:Mbzuai oryx Awesome LLM Post training Checkpoint Every 3 Papers
Environments
- Environment:Alibaba ROLL Diffusion Video Environment
- Environment:Arize ai Phoenix Frontend Node 22
- Environment:Astronomer Astronomer cosmos Kubernetes Provider
- Environment:Bitsandbytes foundation Bitsandbytes CUDA GPU Runtime
- Environment:OpenRLHF OpenRLHF Ray Distributed Environment
- Environment:NVIDIA NeMo Curator NVIDIA DALI
- Environment:Nautechsystems Nautilus trader Binance API Credentials
- Environment:Pola rs Polars Cloud Storage Environment
- Environment:Triton inference server Server Docker Container Build
- Environment:VainF Torch Pruning CUDA GPU Benchmarking