Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Vespa engine Vespa Config subscription lifecycle
- Workflow:Heibaiying BigData Notes Flink Kafka Streaming Pipeline
- Workflow:ARISE Initiative Robomimic Trained Policy Evaluation
- Workflow:Google deepmind Dm control Control Suite RL Training
- Workflow:OpenRLHF OpenRLHF PPO Ray Training
- Workflow:Bigscience workshop Petals Prompt Tuning Chatbot
- Workflow:Farama Foundation Gymnasium Custom Environment Creation
- Workflow:Getgauge Taiko Gauge Integration Testing
- Workflow:Alibaba ROLL Reward Flow Diffusion Pipeline
- Workflow:Mage ai Mage ai Destination Data Loading
Principles
- Principle:Triton inference server Server Resource Management Testing
- Principle:Microsoft Semantic kernel Event Routing
- Principle:Vllm project Vllm LoRA Engine Configuration
- Principle:Norrrrrrr lyn WAInjectBench Validation Checkpoint Selection
- Principle:Scikit learn Scikit learn Stacking Ensemble
- Principle:MaterializeInc Materialize Container Registry Publishing
- Principle:Openai Evals Batch Eval Execution
- Principle:Langgenius Dify Embedding Indexing Configuration
- Principle:Apache Beam Model Enforcement
- Principle:Nightwatchjs Nightwatch Multi Page Composition
Implementations
- Implementation:ArroyoSystems Arroyo Session Window
- Implementation:SeleniumHQ Selenium Closure Uri
- Implementation:Openai Openai python Pydantic Function Tool
- Implementation:Online ml River Optim Nadam
- Implementation:Ggml org Llama cpp Common Header
- Implementation:Pola rs Polars Scan for Streaming
- Implementation:OpenHands OpenHands SetAuthCookieMiddleware
- Implementation:Treeverse LakeFS Java SDK Model RangeMetadata
- Implementation:Ucbepic Docetl CodeOperations
- Implementation:Microsoft Onnxruntime CPU LSTM Grad
Heuristics
- Heuristic:Microsoft DeepSpeedExamples Gradient Checkpointing Tradeoff
- Heuristic:EvolvingLMMs Lab Lmms eval Request Caching Strategy
- Heuristic:Huggingface Optimum Version Conditional Behavior
- Heuristic:Sail sg LongSpec NCCL Distributed Settings
- Heuristic:LLMBook zh LLMBook zh github io BF16 Mixed Precision Default
- Heuristic:Alibaba MNN Weight Quantization Strategy
- Heuristic:Eventual Inc Daft Execution Config Tuning
- Heuristic:Ollama Ollama Quantization Layer Selection
- Heuristic:Openai Openai python Fine Tuning Data Preparation Tips
- Heuristic:Fede1024 Rust rdkafka Manual Offset Store Pattern
Environments
- Environment:Microsoft Semantic kernel OpenAI API Environment
- Environment:Intel Ipex llm RAG LlamaIndex Environment
- Environment:Huggingface Transformers 3D Parallel Multi GPU
- Environment:Evidentlyai Evidently SQL Storage Environment
- Environment:HKUDS AI Trader Python LangChain Runtime
- Environment:FlowiseAI Flowise Node Runtime Environment
- Environment:Kubeflow Kubeflow Git GitHub Environment
- Environment:NVIDIA NeMo Curator Python Linux Base
- Environment:Sgl project Sglang Distributed
- Environment:NVIDIA DALI TensorFlow Environment