Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Onnx Onnx Model Validation
- Workflow:Mbzuai oryx Awesome LLM Post training Deep Paper Collection
- Workflow:NVIDIA NeMo Curator Image Curation Pipeline
- Workflow:Deepspeedai DeepSpeed ZeRO Distributed Training
- Workflow:Dotnet Machinelearning Binary Classification Pipeline
- Workflow:Datahub project Datahub Protobuf Schema Ingestion
- Workflow:FlagOpen FlagEmbedding Reranker Finetuning
- Workflow:Tencent Ncnn Vulkan GPU Accelerated Inference
- Workflow:InternLM Lmdeploy VLM Inference Pipeline
- Workflow:Pytorch Serve Model Deployment
Principles
- Principle:Puppeteer Puppeteer Browser Version Resolution
- Principle:Pytorch Serve IPEX Quantized Inference
- Principle:Bentoml BentoML Container Image Building
- Principle:Google deepmind Mujoco Parallel Constraint Solving
- Principle:Kubeflow Pipelines XGBoost Model Training
- Principle:OpenRLHF OpenRLHF Agent Based Rollout Collection
- Principle:Spotify Luigi Spark Job Definition
- Principle:Apache Flink Fatal Exception Classification
- Principle:Openclaw Openclaw Routing Verification
- Principle:Openai Evals Custom Eval Implementation
Implementations
- Implementation:Mistralai Client python Function Dispatch Pattern
- Implementation:Astronomer Astronomer cosmos DbtVirtualenvBaseOperator
- Implementation:Open compass VLMEvalKit VGRPBench Aquarium
- Implementation:NVIDIA TransformerEngine Float8CurrentScaling Recipe
- Implementation:Infiniflow Ragflow Next Request
- Implementation:Openai Openai python Response Reasoning Summary Text Done
- Implementation:Ggml org Ggml Ggml backend sched graph compute
- Implementation:Ggml org Ggml Cpu quantization
- Implementation:Datahub project Datahub EntityClient Upsert
- Implementation:Speechbrain Speechbrain Train TimersAndSuch Wav2Vec
Heuristics
- Heuristic:PeterL1n BackgroundMattingV2 Data Augmentation Strategy
- Heuristic:Intel Ipex llm LoRA Target All Linear Layers
- Heuristic:Cleanlab Cleanlab Confident Threshold Heuristic
- Heuristic:Kornia Kornia Numerical Stability Patterns
- Heuristic:Google deepmind Dm control Prop Settling Physics Tuning
- Heuristic:Tencent Ncnn FP16 Precision Selection
- Heuristic:Axolotl ai cloud Axolotl Memory Optimization Tips
- Heuristic:Mlc ai Web llm Grammar Matcher Reuse
- Heuristic:Mlfoundations Open flamingo Deterministic Shard Shuffling
- Heuristic:Explodinggradients Ragas LLM Temperature Defaults
Environments
- Environment:SeldonIO Seldon core Kubernetes Cluster Environment
- Environment:Guardrails ai Guardrails Docker Server Runtime
- Environment:Triton inference server Server Docker Container Build
- Environment:Pytorch Serve CUDA GPU Environment
- Environment:LLMBook zh LLMBook zh github io PyTorch CUDA GPU Environment
- Environment:TobikoData Sqlmesh Snowflake Connection
- Environment:Microsoft DeepSpeedExamples RLHF Training Environment
- Environment:Isaac sim IsaacGymEnvs Pip Dependencies
- Environment:Lucidrains X transformers PyTorch CUDA
- Environment:Mistralai Client python Python SDK Environment