Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Apache Airflow Scheduler Operation and Task Execution
- Workflow:Haosulab ManiSkill Sim2Real Deployment
- Workflow:AUTOMATIC1111 Stable diffusion webui Text to image generation
- Workflow:Roboflow Rf detr Object Detection Inference
- Workflow:SeleniumHQ Selenium Chrome DevTools Protocol Integration
- Workflow:Ray project Ray Actor Lifecycle Management
- Workflow:VainF Torch Pruning Object Detection Pruning
- Workflow:Astronomer Astronomer cosmos Local dbt DAG rendering
- Workflow:Volcengine Verl GRPO Training Pipeline
- Workflow:Deepset ai Haystack Document Preprocessing Pipeline
Principles
- Principle:Huggingface Trl Reward Evaluation and Saving
- Principle:Apache Dolphinscheduler Plugin Module Structure
- Principle:EvolvingLMMs Lab Lmms eval MCP Tool Integration
- Principle:Apache Hudi Write Operation Configuration
- Principle:Huggingface Datatrove Pipeline Failure Diagnosis
- Principle:Langfuse Langfuse LLM Execution for Experiments
- Principle:Triton inference server Server Classification Postprocessing
- Principle:Openai Evals Solver Configuration Patterns
- Principle:Vespa engine Vespa ABI Compatibility Specification
- Principle:Gretelai Gretel synthetics Batch Data Export
Implementations
- Implementation:Cohere ai Cohere python RequestOptions
- Implementation:Microsoft Onnxruntime ORTModule Training Execution
- Implementation:Spotify Luigi Luigi Build Run
- Implementation:Mlc ai Mlc llm FP8 Quantization
- Implementation:Microsoft Onnxruntime CUDA TrainingKernels
- Implementation:Treeverse LakeFS Java SDK Model StsAuthRequest
- Implementation:NVIDIA TransformerEngine Activation C API
- Implementation:Lance format Lance StructEncoding
- Implementation:DevExpress Testcafe TestRunTracker
- Implementation:Lucidrains X transformers Paired Sequence Generator Pattern
Heuristics
- Heuristic:OpenBMB UltraFeedback Score 10 Anomaly Correction
- Heuristic:Fede1024 Rust rdkafka Manual Offset Store Pattern
- Heuristic:Huggingface Alignment handbook Gradient Checkpointing Use Cache
- Heuristic:Romsto Speculative Decoding Shared Tokenizer Requirement
- Heuristic:Elevenlabs Elevenlabs python Text Chunking Splitter Characters
- Heuristic:Openai Whisper Compression Ratio Threshold
- Heuristic:Shiyu coder Kronos Sampling Temperature Tuning
- Heuristic:NVIDIA NeMo Curator GPU Memory Resource Allocation
- Heuristic:Microsoft DeepSpeedExamples LoRA Learning Rate Scaling
- Heuristic:Dotnet Machinelearning Numerical Stability Guards
Environments
- Environment:Google deepmind Dm control GLFW Desktop Rendering
- Environment:Openai Openai node OpenAI API Credentials
- Environment:Deepset ai Haystack HuggingFace Model Environment
- Environment:Roboflow Rf detr Roboflow Deployment Credentials
- Environment:Ray project Ray CI Build Matrix Environment
- Environment:Pyro ppl Pyro Distributed Training
- Environment:Marker Inc Korea AutoRAG Python 3 10 Runtime
- Environment:Apache Dolphinscheduler Java Runtime
- Environment:LLMBook zh LLMBook zh github io PyTorch CUDA GPU Environment
- Environment:Google research Deduplicate text datasets Python HuggingFace Environment