Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Mistralai Client python GCP Chat Completion
- Workflow:Intel Ipex llm QLoRA Finetuning
- Workflow:Microsoft Autogen Studio Team Deployment
- Workflow:Ggml org Llama cpp Multimodal Inference
- Workflow:Pola rs Polars DataFrame Aggregation and Grouping
- Workflow:Volcengine Verl Supervised Fine Tuning
- Workflow:Langfuse Langfuse Evaluation pipeline
- Workflow:Kserve Kserve Canary Rollout Deployment
- Workflow:Dagster io Dagster Modal Serverless Pipeline
- Workflow:Openclaw Openclaw Gateway Operations And Diagnostics
Principles
- Principle:Pola rs Polars Lazy Data Scanning
- Principle:Ggml org Llama cpp Terminal IO
- Principle:Mlc ai Web llm Chat Request Configuration
- Principle:LLMBook zh LLMBook zh github io Causal LM Loss Computation
- Principle:Vllm project Vllm Multimodal Output Processing
- Principle:Alibaba ROLL Distillation Dataset Preparation
- Principle:Huggingface Optimum Dummy Input Generation
- Principle:Facebookresearch Audiocraft Pretrained Compression Export
- Principle:CARLA simulator Carla Pedestrian Navigation Mesh
- Principle:OpenGVLab InternVL Segmentation Dataset Configuration
Implementations
- Implementation:Evidentlyai Evidently Custom Descriptors
- Implementation:Scikit learn Scikit learn ColumnTransformer Init
- Implementation:Alibaba ROLL DPOTrainer
- Implementation:Langfuse Langfuse ClickHouse Query Fragments
- Implementation:Huggingface Trl TrlParser DPOConfig
- Implementation:OpenHands OpenHands GithubManager Send Message
- Implementation:Scikit learn Scikit learn IterativeImputer
- Implementation:Diagram of thought Diagram of thought Summarizer Completeness Check
- Implementation:Recommenders team Recommenders DKN Model
- Implementation:TobikoData Sqlmesh Story ModelLineage
Heuristics
- Heuristic:Kubeflow Pipelines Component URL Commit SHA Pinning
- Heuristic:Farama Foundation Gymnasium Seeding Determinism Best Practices
- Heuristic:Huggingface Open r1 vLLM GPU Allocation
- Heuristic:Lakeraai Pint benchmark Optimal Configuration Selection
- Heuristic:Mlfoundations Open flamingo KV Cache Classification Optimization
- Heuristic:Huggingface Datasets Parquet Shard Sizing
- Heuristic:Huggingface Alignment handbook EOS Token Alignment
- Heuristic:Sgl project Sglang Memory Fraction Tuning
- Heuristic:Iamhankai Forest of Thought Tree Iteration Scaling
- Heuristic:Puppeteer Puppeteer Firefox Single Process Workaround
Environments
- Environment:Tensorflow Serving Build Environment
- Environment:Online ml River Build Toolchain
- Environment:Alibaba MNN GPU OpenCL Environment
- Environment:Intel Ipex llm XPU Serving Environment
- Environment:Intel Ipex llm Build Environment
- Environment:Openclaw Openclaw Docker Deployment Environment
- Environment:Vllm project Vllm Python Dependencies
- Environment:Spcl Graph of thoughts Local LLaMA GPU Inference
- Environment:Snorkel team Snorkel SpaCy NLP
- Environment:Langfuse Langfuse Docker Infrastructure